Constrained Average-Reward Intermittently Observable MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Avrachenkov, Konstantin, Dhiman, Madhu, Kavitha, Veeraruna |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal Control with $L^{\infty}$ cost: incorporating peak minimization
by: Dhiman, Madhu, et al.
Published: (2024)
by: Dhiman, Madhu, et al.
Published: (2024)
Games with Rational and Herding Players
by: Vyas, Raghupati, et al.
Published: (2026)
by: Vyas, Raghupati, et al.
Published: (2026)
Stability of Polling Systems for a Large Class of Markovian Switching Policies
by: Avrachenkov, Konstantin, et al.
Published: (2025)
by: Avrachenkov, Konstantin, et al.
Published: (2025)
Punitive policies to combat misreporting in dynamic supply chains
by: Dhiman, Madhu, et al.
Published: (2025)
by: Dhiman, Madhu, et al.
Published: (2025)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Queue or lounge: strategic design for strategic customer
by: Sultana, Riya, et al.
Published: (2025)
by: Sultana, Riya, et al.
Published: (2025)
Multi-type random game dynamics: limits at discontinuities and cyclic limits
by: Vyas, Raghupati, et al.
Published: (2026)
by: Vyas, Raghupati, et al.
Published: (2026)
Policy Gradient Algorithms in Average-Reward Multichain MDPs
by: Lee, Jongmin, et al.
Published: (2026)
by: Lee, Jongmin, et al.
Published: (2026)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
Balancing Morality and Economics: Population Games with Herding and Inertia
by: Vyas, Raghupati, et al.
Published: (2026)
by: Vyas, Raghupati, et al.
Published: (2026)
On the interplay between pricing, competition and QoS in ride-hailing
by: Walunj, Tushar Shankar, et al.
Published: (2023)
by: Walunj, Tushar Shankar, et al.
Published: (2023)
Computing Stationary Distribution via Dirichlet-Energy Minimization by Coordinate Descent
by: Avrachenkov, Konstantin, et al.
Published: (2026)
by: Avrachenkov, Konstantin, et al.
Published: (2026)
Derivative Estimation from Coarse, Irregular, Noisy Samples: An MLE-Spline Approach
by: Avrachenkov, Konstantin E., et al.
Published: (2025)
by: Avrachenkov, Konstantin E., et al.
Published: (2025)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
by: Chae, Woojin, et al.
Published: (2024)
by: Chae, Woojin, et al.
Published: (2024)
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
by: Wang, Shengbo, et al.
Published: (2026)
by: Wang, Shengbo, et al.
Published: (2026)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Span-Based Optimal Sample Complexity for Average Reward MDPs
by: Zurek, Matthew, et al.
Published: (2023)
by: Zurek, Matthew, et al.
Published: (2023)
Linking PageRank, Time Reversal, and Policy Evaluation
by: Avrachenkov, Konstantin, et al.
Published: (2026)
by: Avrachenkov, Konstantin, et al.
Published: (2026)
The Plug-in Approach for Average-Reward and Discounted MDPs: Optimal Sample Complexity Analysis
by: Zurek, Matthew, et al.
Published: (2024)
by: Zurek, Matthew, et al.
Published: (2024)
Epsilon-Optimal Policies for Average-Cost Separable MDPs with Perturbations
by: Kantawala, Dhairya
Published: (2025)
by: Kantawala, Dhairya
Published: (2025)
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
by: Zurek, Matthew, et al.
Published: (2024)
by: Zurek, Matthew, et al.
Published: (2024)
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)
by: Wang, Weikai, et al.
Published: (2025)
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
by: Goldsztajn, Diego, et al.
Published: (2024)
by: Goldsztajn, Diego, et al.
Published: (2024)
Queues with inspection cost: To see or not to see?
by: Clarkson, Jake, et al.
Published: (2025)
by: Clarkson, Jake, et al.
Published: (2025)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
0/1 Constrained Optimization Solving Sample Average Approximation for Chance Constrained Programming
by: Zhou, Shenglong, et al.
Published: (2022)
by: Zhou, Shenglong, et al.
Published: (2022)
Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Approximate Solution Methods for the Average Reward Criterion in Optimal Tracking Control of Linear Systems
by: Nguyen, Duc Cuong
Published: (2025)
by: Nguyen, Duc Cuong
Published: (2025)
Optimal Non-Asymptotic Rates of Value Iteration for Average-Reward Markov Decision Processes
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
A Proximal DC Algorithm for Sample Average Approximation of Chance Constrained Programming
by: Wang, Peng, et al.
Published: (2023)
by: Wang, Peng, et al.
Published: (2023)
Balancing rationality and social influence: Alpha-rational Nash equilibrium in games with herding
by: Agarwal, Khushboo, et al.
Published: (2024)
by: Agarwal, Khushboo, et al.
Published: (2024)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
by: Chen, Zijun, et al.
Published: (2025)
by: Chen, Zijun, et al.
Published: (2025)
Optimal Sample Complexity for Average Reward Markov Decision Processes
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
by: Zhang, Runyu, et al.
Published: (2023)
by: Zhang, Runyu, et al.
Published: (2023)
Shape-Constrained Distributional Optimization via Importance-Weighted Sample Average Approximation
by: Lam, Henry, et al.
Published: (2024)
by: Lam, Henry, et al.
Published: (2024)
Convergence of actor-critic for entropy regularised MDPs in general action spaces
by: Zorba, Denis, et al.
Published: (2025)
by: Zorba, Denis, et al.
Published: (2025)
Sequential Decision-Making under Uncertainty: A Robust MDPs review
by: Ou, Wenfan, et al.
Published: (2024)
by: Ou, Wenfan, et al.
Published: (2024)
Revisiting Subgradient Dominance in Robust MDPs: Counterexamples, Hardness, and Sufficient Conditions
by: Kitamura, Toshinori, et al.
Published: (2026)
by: Kitamura, Toshinori, et al.
Published: (2026)
Learning Weakly Communicating Average-Reward CMDPs: Strong Duality and Improved Regret
by: Yu, Kihyun, et al.
Published: (2026)
by: Yu, Kihyun, et al.
Published: (2026)
Similar Items
-
Optimal Control with $L^{\infty}$ cost: incorporating peak minimization
by: Dhiman, Madhu, et al.
Published: (2024) -
Games with Rational and Herding Players
by: Vyas, Raghupati, et al.
Published: (2026) -
Stability of Polling Systems for a Large Class of Markovian Switching Policies
by: Avrachenkov, Konstantin, et al.
Published: (2025) -
Punitive policies to combat misreporting in dynamic supply chains
by: Dhiman, Madhu, et al.
Published: (2025) -
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)