Optimal Non-Asymptotic Rates of Value Iteration for Average-Reward Markov Decision Processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Jongmin, Ryu, Ernest K. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Policy Gradient Algorithms in Average-Reward Multichain MDPs
von: Lee, Jongmin, et al.
Veröffentlicht: (2026)
von: Lee, Jongmin, et al.
Veröffentlicht: (2026)
Deflated Dynamics Value Iteration
von: Lee, Jongmin, et al.
Veröffentlicht: (2024)
von: Lee, Jongmin, et al.
Veröffentlicht: (2024)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Convergence Analyses of Davis-Yin Splitting via Scaled Relative Graphs
von: Lee, Jongmin, et al.
Veröffentlicht: (2022)
von: Lee, Jongmin, et al.
Veröffentlicht: (2022)
On the Convergence Rate of MCTS for the Optimal Value Estimation in Markov Decision Processes
von: Chang, Hyeong Soo
Veröffentlicht: (2024)
von: Chang, Hyeong Soo
Veröffentlicht: (2024)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
On Convergence of Average-Reward Q-Learning in Weakly Communicating Markov Decision Processes
von: Wan, Yi, et al.
Veröffentlicht: (2024)
von: Wan, Yi, et al.
Veröffentlicht: (2024)
Optimal First-Order Algorithms as a Function of Inequalities
von: Park, Chanwoo, et al.
Veröffentlicht: (2021)
von: Park, Chanwoo, et al.
Veröffentlicht: (2021)
Markov Decision Processes with Value-at-Risk Criterion
von: Xia, Li, et al.
Veröffentlicht: (2025)
von: Xia, Li, et al.
Veröffentlicht: (2025)
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
Beyond Average Return in Markov Decision Processes
von: Marthe, Alexandre, et al.
Veröffentlicht: (2023)
von: Marthe, Alexandre, et al.
Veröffentlicht: (2023)
Finite-Time Bounds for Average-Reward Fitted Q-Iteration
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
Inexact Policy Iteration Methods for Large-Scale Markov Decision Processes
von: Gargiani, Matilde, et al.
Veröffentlicht: (2024)
von: Gargiani, Matilde, et al.
Veröffentlicht: (2024)
Nesterov Flow May Travel Infinitely Long to Converge to a Minimizer
von: Ryu, Ernest K.
Veröffentlicht: (2026)
von: Ryu, Ernest K.
Veröffentlicht: (2026)
Uniqueness of DRS as the 2 Operator Resolvent-Splitting and Impossibility of 3 Operator Resolvent-Splitting
von: Ryu, Ernest K.
Veröffentlicht: (2018)
von: Ryu, Ernest K.
Veröffentlicht: (2018)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
von: Misra, Rahul, et al.
Veröffentlicht: (2025)
von: Misra, Rahul, et al.
Veröffentlicht: (2025)
Nesterov Acceleration with Operator Decomposition
von: Lee, Jaewook, et al.
Veröffentlicht: (2026)
von: Lee, Jaewook, et al.
Veröffentlicht: (2026)
Absorbing Markov Decision Processes
von: Dufour, François, et al.
Veröffentlicht: (2023)
von: Dufour, François, et al.
Veröffentlicht: (2023)
Point Convergence of Nesterov's Accelerated Gradient Method: An AI-Assisted Proof
von: Jang, Uijeong, et al.
Veröffentlicht: (2025)
von: Jang, Uijeong, et al.
Veröffentlicht: (2025)
Optimal Acceleration for Proximal Minimization of the Sum of Convex and Strongly Convex Functions
von: Chari, Govind M., et al.
Veröffentlicht: (2026)
von: Chari, Govind M., et al.
Veröffentlicht: (2026)
Robust Reward Design for Markov Decision Processes
von: Wu, Shuo, et al.
Veröffentlicht: (2024)
von: Wu, Shuo, et al.
Veröffentlicht: (2024)
A Bayesian Composite Risk Approach for Stochastic Optimal Control and Markov Decision Processes
von: Ma, Wentao, et al.
Veröffentlicht: (2024)
von: Ma, Wentao, et al.
Veröffentlicht: (2024)
Sample Complexity for Markov Decision Processes and Stochastic Optimal Control with Static Risk Measures
von: Chávez, Cristian, et al.
Veröffentlicht: (2026)
von: Chávez, Cristian, et al.
Veröffentlicht: (2026)
Optimal Acceleration for Minimax and Fixed-Point Problems is Not Unique
von: Yoon, TaeHo, et al.
Veröffentlicht: (2024)
von: Yoon, TaeHo, et al.
Veröffentlicht: (2024)
Tractable Robust Markov Decision Processes
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2024)
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2024)
Mean Field Markov Decision Processes
von: Bäuerle, Nicole
Veröffentlicht: (2021)
von: Bäuerle, Nicole
Veröffentlicht: (2021)
Bounding the Difference between the Values of Robust and Non-Robust Markov Decision Problems
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
Accelerated Minimax Algorithms Flock Together
von: Yoon, TaeHo, et al.
Veröffentlicht: (2022)
von: Yoon, TaeHo, et al.
Veröffentlicht: (2022)
Quantum Markov Decision Processes: Dynamic and Semi-Definite Programs for Optimal Solutions
von: Saldi, Naci, et al.
Veröffentlicht: (2024)
von: Saldi, Naci, et al.
Veröffentlicht: (2024)
Dynamic Capital Requirements for Markov Decision Processes
von: Haskell, William B., et al.
Veröffentlicht: (2024)
von: Haskell, William B., et al.
Veröffentlicht: (2024)
Risk-averse formulations of Stochastic Optimal Control and Markov Decision Processes
von: Shapiro, Alexander, et al.
Veröffentlicht: (2025)
von: Shapiro, Alexander, et al.
Veröffentlicht: (2025)
Convex Approximations of Random Constrained Markov Decision Processes
von: Varagapriya, V, et al.
Veröffentlicht: (2025)
von: Varagapriya, V, et al.
Veröffentlicht: (2025)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
von: Boone, Victor, et al.
Veröffentlicht: (2024)
von: Boone, Victor, et al.
Veröffentlicht: (2024)
Operator Splitting for Convex Constrained Markov Decision Processes
von: Grontas, Panagiotis D., et al.
Veröffentlicht: (2024)
von: Grontas, Panagiotis D., et al.
Veröffentlicht: (2024)
Markov Decision Process Design: A Framework for Integrating Strategic and Operational Decisions
von: Brown, Seth, et al.
Veröffentlicht: (2023)
von: Brown, Seth, et al.
Veröffentlicht: (2023)
Stochastic Mirror Descent under Iterate-Dependent Markov Noise: Analysis in the Asymptotic and Finite Time Regimes
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
Approximate Solution Methods for the Average Reward Criterion in Optimal Tracking Control of Linear Systems
von: Nguyen, Duc Cuong
Veröffentlicht: (2025)
von: Nguyen, Duc Cuong
Veröffentlicht: (2025)
Optimal Sensor and Actuator Selection for Factored Markov Decision Processes: Complexity, Approximability and Algorithms
von: Bhargav, Jayanth, et al.
Veröffentlicht: (2024)
von: Bhargav, Jayanth, et al.
Veröffentlicht: (2024)
Asymptotic Optimality in Data-Driven Decision Making
von: Salač, Radek, et al.
Veröffentlicht: (2025)
von: Salač, Radek, et al.
Veröffentlicht: (2025)
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
von: Goldsztajn, Diego, et al.
Veröffentlicht: (2024)
von: Goldsztajn, Diego, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Policy Gradient Algorithms in Average-Reward Multichain MDPs
von: Lee, Jongmin, et al.
Veröffentlicht: (2026) -
Deflated Dynamics Value Iteration
von: Lee, Jongmin, et al.
Veröffentlicht: (2024) -
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
Convergence Analyses of Davis-Yin Splitting via Scaled Relative Graphs
von: Lee, Jongmin, et al.
Veröffentlicht: (2022) -
On the Convergence Rate of MCTS for the Optimal Value Estimation in Markov Decision Processes
von: Chang, Hyeong Soo
Veröffentlicht: (2024)