Epsilon-Optimal Policies for Average-Cost Separable MDPs with Perturbations
Fuente:
arXiv
Saved in:
| Main Author: | Kantawala, Dhairya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
by: Wang, Shengbo, et al.
Published: (2026)
by: Wang, Shengbo, et al.
Published: (2026)
Policy Gradient Algorithms in Average-Reward Multichain MDPs
by: Lee, Jongmin, et al.
Published: (2026)
by: Lee, Jongmin, et al.
Published: (2026)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
Risk-Aware Finite-Horizon Social Optimal Control of Mean-Field Coupled Linear-Quadratic Subsystems
by: Patel, Dhairya, et al.
Published: (2024)
by: Patel, Dhairya, et al.
Published: (2024)
Constrained Average-Reward Intermittently Observable MDPs
by: Avrachenkov, Konstantin, et al.
Published: (2025)
by: Avrachenkov, Konstantin, et al.
Published: (2025)
Span-Based Optimal Sample Complexity for Average Reward MDPs
by: Zurek, Matthew, et al.
Published: (2023)
by: Zurek, Matthew, et al.
Published: (2023)
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)
by: Wang, Weikai, et al.
Published: (2025)
Linear Dynamics meets Linear MDPs: Closed-Form Optimal Policies via Reinforcement Learning
by: Makdah, Abed AlRahman Al, et al.
Published: (2025)
by: Makdah, Abed AlRahman Al, et al.
Published: (2025)
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
by: Feinberg, Eugene A., et al.
Published: (2024)
by: Feinberg, Eugene A., et al.
Published: (2024)
Infinite Horizon Average Cost Optimality Criteria for Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2023)
by: Bayraktar, Erhan, et al.
Published: (2023)
The Plug-in Approach for Average-Reward and Discounted MDPs: Optimal Sample Complexity Analysis
by: Zurek, Matthew, et al.
Published: (2024)
by: Zurek, Matthew, et al.
Published: (2024)
Efficient Duple Perturbation Robustness in Low-rank MDPs
by: Hu, Yang, et al.
Published: (2024)
by: Hu, Yang, et al.
Published: (2024)
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
by: Zurek, Matthew, et al.
Published: (2024)
by: Zurek, Matthew, et al.
Published: (2024)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
by: Zhang, Runyu, et al.
Published: (2023)
by: Zhang, Runyu, et al.
Published: (2023)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
by: Chae, Woojin, et al.
Published: (2024)
by: Chae, Woojin, et al.
Published: (2024)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Near-Optimal Sample Complexity for MDPs via Anchoring
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
Landscape of Policy Optimization for Finite Horizon MDPs with General State and Action
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Epsilon-Greedy Thompson Sampling to Bayesian Optimization
by: Do, Bach, et al.
Published: (2024)
by: Do, Bach, et al.
Published: (2024)
Policy Learning for Perturbance-wise Linear Quadratic Control Problem
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
Optimal Variance-Dependent Regret Bounds for Infinite-Horizon MDPs
by: Zamir, Guy, et al.
Published: (2026)
by: Zamir, Guy, et al.
Published: (2026)
Policy Synthesis for Interval MDPs via Polyhedral Lyapunov Functions
by: Monir, Negar, et al.
Published: (2026)
by: Monir, Negar, et al.
Published: (2026)
Data-Driven Non-Parametric Model Learning and Adaptive Control of MDPs with Borel spaces: Identifiability and Near Optimal Design
by: Mrani-Zentar, Omar, et al.
Published: (2025)
by: Mrani-Zentar, Omar, et al.
Published: (2025)
Retrospective Cost-based Extremum Seeking Control with Vanishing Perturbation for Online Output Minimization
by: Paredes, Juan A., et al.
Published: (2024)
by: Paredes, Juan A., et al.
Published: (2024)
Approximate Solution Methods for the Average Reward Criterion in Optimal Tracking Control of Linear Systems
by: Nguyen, Duc Cuong
Published: (2025)
by: Nguyen, Duc Cuong
Published: (2025)
Convergence of actor-critic for entropy regularised MDPs in general action spaces
by: Zorba, Denis, et al.
Published: (2025)
by: Zorba, Denis, et al.
Published: (2025)
Sequential Decision-Making under Uncertainty: A Robust MDPs review
by: Ou, Wenfan, et al.
Published: (2024)
by: Ou, Wenfan, et al.
Published: (2024)
Revisiting Subgradient Dominance in Robust MDPs: Counterexamples, Hardness, and Sufficient Conditions
by: Kitamura, Toshinori, et al.
Published: (2026)
by: Kitamura, Toshinori, et al.
Published: (2026)
Optimal Non-Asymptotic Rates of Value Iteration for Average-Reward Markov Decision Processes
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Constraint-Generation Policy Optimization (CGPO): Nonlinear Programming for Policy Optimization in Mixed Discrete-Continuous MDPs
by: Gimelfarb, Michael, et al.
Published: (2024)
by: Gimelfarb, Michael, et al.
Published: (2024)
Thinking Beyond Visibility: A Near-Optimal Policy Framework for Locally Interdependent Multi-Agent MDPs
by: DeWeese, Alex, et al.
Published: (2025)
by: DeWeese, Alex, et al.
Published: (2025)
Performance of NPG in Countable State-Space Average-Cost RL
by: Murthy, Yashaswini, et al.
Published: (2024)
by: Murthy, Yashaswini, et al.
Published: (2024)
Entropic Optimal Transport Problem with Convex Functional Cost
by: Kazeykina, Anna, et al.
Published: (2025)
by: Kazeykina, Anna, et al.
Published: (2025)
Generative Transfer for Entropic Optimal Transport with Unknown Costs
by: Debouchage, Antoine, et al.
Published: (2026)
by: Debouchage, Antoine, et al.
Published: (2026)
Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL
by: Zurek, Matthew, et al.
Published: (2025)
by: Zurek, Matthew, et al.
Published: (2025)
Model approximation in MDPs with unbounded per-step cost
by: Bozkurt, Berk, et al.
Published: (2024)
by: Bozkurt, Berk, et al.
Published: (2024)
Scalable Two-Stage Stochastic Optimal Power Flow via Separable Approximation
by: Lamichhane, Shishir, et al.
Published: (2025)
by: Lamichhane, Shishir, et al.
Published: (2025)
Similar Items
-
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
by: Wang, Shengbo, et al.
Published: (2026) -
Policy Gradient Algorithms in Average-Reward Multichain MDPs
by: Lee, Jongmin, et al.
Published: (2026) -
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024) -
Risk-Aware Finite-Horizon Social Optimal Control of Mean-Field Coupled Linear-Quadratic Subsystems
by: Patel, Dhairya, et al.
Published: (2024) -
Constrained Average-Reward Intermittently Observable MDPs
by: Avrachenkov, Konstantin, et al.
Published: (2025)