Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
Fuente:
arXiv
Guardado en:
| Autores principales: | Meng, Huiling, Chen, Ningyuan, Gao, Xuefeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reinforcement Learning for Jump-Diffusions, with Financial Applications
por: Gao, Xuefeng, et al.
Publicado: (2024)
por: Gao, Xuefeng, et al.
Publicado: (2024)
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
por: Wang, Tianyu, et al.
Publicado: (2024)
por: Wang, Tianyu, et al.
Publicado: (2024)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
por: Li, Lucky
Publicado: (2024)
por: Li, Lucky
Publicado: (2024)
Multi-Armed Bandits with Minimum Aggregated Revenue Constraints
por: Yahmed, Ahmed Ben, et al.
Publicado: (2025)
por: Yahmed, Ahmed Ben, et al.
Publicado: (2025)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2021)
por: Zeng, Sihan, et al.
Publicado: (2021)
Blind Network Revenue Management and Bandits with Knapsacks under Limited Switches
por: Simchi-Levi, David, et al.
Publicado: (2019)
por: Simchi-Levi, David, et al.
Publicado: (2019)
Reward-Directed Score-Based Diffusion Models via q-Learning
por: Gao, Xuefeng, et al.
Publicado: (2024)
por: Gao, Xuefeng, et al.
Publicado: (2024)
Constrained Pricing in Choice-based Revenue Management
por: Shao, Qian, et al.
Publicado: (2025)
por: Shao, Qian, et al.
Publicado: (2025)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
por: Gao, Xuefeng, et al.
Publicado: (2022)
por: Gao, Xuefeng, et al.
Publicado: (2022)
An accelerated first-order regularized momentum descent ascent algorithm for stochastic nonconvex-concave minimax problems
por: Zhang, Huiling, et al.
Publicado: (2023)
por: Zhang, Huiling, et al.
Publicado: (2023)
Structure-Informed Deep Reinforcement Learning for Inventory Management
por: Maggiar, Alvaro, et al.
Publicado: (2025)
por: Maggiar, Alvaro, et al.
Publicado: (2025)
Regret Bounds for Episodic Risk-Sensitive Linear Quadratic Regulator
por: Xu, Wenhao, et al.
Publicado: (2024)
por: Xu, Wenhao, et al.
Publicado: (2024)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks (YANNs)
por: Braniff, Austin, et al.
Publicado: (2025)
por: Braniff, Austin, et al.
Publicado: (2025)
Extensions of Robbins-Siegmund Theorem with Applications in Reinforcement Learning
por: Liu, Xinyu, et al.
Publicado: (2025)
por: Liu, Xinyu, et al.
Publicado: (2025)
Reinforcement Learning and Regret Bounds for Admission Control
por: Weber, Lucas, et al.
Publicado: (2024)
por: Weber, Lucas, et al.
Publicado: (2024)
Mitigating Covariate Shift in Misspecified Regression with Applications to Reinforcement Learning
por: Amortila, Philip, et al.
Publicado: (2024)
por: Amortila, Philip, et al.
Publicado: (2024)
Convex Chance-Constrained Stochastic Control under Uncertain Specifications with Application to Learning-Based Hybrid Powertrain Control
por: Kato, Teruki, et al.
Publicado: (2026)
por: Kato, Teruki, et al.
Publicado: (2026)
On Sinkhorn's Algorithm and Choice Modeling
por: Qu, Zhaonan, et al.
Publicado: (2023)
por: Qu, Zhaonan, et al.
Publicado: (2023)
A Nonlinear Separation Principle via Contraction Theory: Applications to Neural Networks, Control, and Learning
por: Gokhale, Anand, et al.
Publicado: (2026)
por: Gokhale, Anand, et al.
Publicado: (2026)
Zeroth-Order primal-dual Alternating Projection Gradient Algorithms for Nonconvex Minimax Problems with Coupled linear Constraints
por: Zhang, Huiling, et al.
Publicado: (2024)
por: Zhang, Huiling, et al.
Publicado: (2024)
Completely Parameter-Free Single-Loop Algorithms for Nonconvex-Concave Minimax Problems
por: Yang, Junnan, et al.
Publicado: (2024)
por: Yang, Junnan, et al.
Publicado: (2024)
Semi-on-Demand Transit Feeders with Shared Autonomous Vehicles and Reinforcement-Learning-Based Zonal Dispatching Control
por: Ng, Max T. M., et al.
Publicado: (2025)
por: Ng, Max T. M., et al.
Publicado: (2025)
Grower-in-the-Loop Interactive Reinforcement Learning for Greenhouse Climate Control
por: Xiao, Maxiu, et al.
Publicado: (2025)
por: Xiao, Maxiu, et al.
Publicado: (2025)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes
por: Braniff, Austin, et al.
Publicado: (2026)
por: Braniff, Austin, et al.
Publicado: (2026)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Iterative Minimax Games with Coupled Linear Constraints
por: Zhang, Huiling, et al.
Publicado: (2022)
por: Zhang, Huiling, et al.
Publicado: (2022)
Achieve Performatively Optimal Policy for Performative Reinforcement Learning
por: Chen, Ziyi, et al.
Publicado: (2025)
por: Chen, Ziyi, et al.
Publicado: (2025)
Continuous Q-Score Matching: Diffusion Guided Reinforcement Learning for Continuous-Time Control
por: Hua, Chengxiu, et al.
Publicado: (2025)
por: Hua, Chengxiu, et al.
Publicado: (2025)
Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning
por: Zhan, Donglin, et al.
Publicado: (2025)
por: Zhan, Donglin, et al.
Publicado: (2025)
Solving the Paint Shop Problem with Flexible Management of Multi-Lane Buffers Using Reinforcement Learning and Action Masking
por: Stappert, Mirko, et al.
Publicado: (2025)
por: Stappert, Mirko, et al.
Publicado: (2025)
Hybrid Reinforcement Learning Framework for Mixed-Variable Problems
por: Zhai, Haoyan, et al.
Publicado: (2024)
por: Zhai, Haoyan, et al.
Publicado: (2024)
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
por: Lamperski, Andrew, et al.
Publicado: (2024)
por: Lamperski, Andrew, et al.
Publicado: (2024)
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning
por: Ma, Chaolun, et al.
Publicado: (2022)
por: Ma, Chaolun, et al.
Publicado: (2022)
Fill-and-Spill: Deep Reinforcement Learning Policy Gradient Methods for Reservoir Operation Decision and Control
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
Provably-Stable Neural Network-Based Control of Nonlinear Systems
por: Li, Anran, et al.
Publicado: (2025)
por: Li, Anran, et al.
Publicado: (2025)
A Reinforcement-Learning-Based Multiple-Column Selection Strategy for Column Generation
por: Yuan, Haofeng, et al.
Publicado: (2023)
por: Yuan, Haofeng, et al.
Publicado: (2023)
Principled Penalty-based Methods for Bilevel Reinforcement Learning and RLHF
por: Shen, Han, et al.
Publicado: (2024)
por: Shen, Han, et al.
Publicado: (2024)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
por: Chen, Zijun, et al.
Publicado: (2025)
por: Chen, Zijun, et al.
Publicado: (2025)
Neural Network-assisted Interval Reachability for Systems with Control Barrier Function-Based Safe Controllers
por: Ajeyemi, Damola, et al.
Publicado: (2025)
por: Ajeyemi, Damola, et al.
Publicado: (2025)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
por: Zeng, Sihan, et al.
Publicado: (2026)
por: Zeng, Sihan, et al.
Publicado: (2026)
Ejemplares similares
-
Reinforcement Learning for Jump-Diffusions, with Financial Applications
por: Gao, Xuefeng, et al.
Publicado: (2024) -
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
por: Wang, Tianyu, et al.
Publicado: (2024) -
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
por: Li, Lucky
Publicado: (2024) -
Multi-Armed Bandits with Minimum Aggregated Revenue Constraints
por: Yahmed, Ahmed Ben, et al.
Publicado: (2025) -
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2021)