A Single-Loop Robust Policy Gradient Method for Robust Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Zhenwei, Xue, Chenyu, Deng, Qi, Ye, Yinyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Achieving Instance-dependent Sample Complexity for Constrained Markov Decision Process
by: Jiang, Jiashuo, et al.
Published: (2024)
by: Jiang, Jiashuo, et al.
Published: (2024)
A Homogenization Approach for Gradient-Dominated Stochastic Optimization
by: Tan, Jiyuan, et al.
Published: (2023)
by: Tan, Jiyuan, et al.
Published: (2023)
Tractable Robust Markov Decision Processes
by: Grand-Clément, Julien, et al.
Published: (2024)
by: Grand-Clément, Julien, et al.
Published: (2024)
Decoupling Learning and Decision-Making: Breaking the $\mathcal{O}(\sqrt{T})$ Barrier in Online Resource Allocation with First-Order Methods
by: Gao, Wenzhi, et al.
Published: (2024)
by: Gao, Wenzhi, et al.
Published: (2024)
A Practical GPU-Enhanced Matrix-Free Primal-Dual Method for Large-Scale Conic Programs
by: Lin, Zhenwei, et al.
Published: (2025)
by: Lin, Zhenwei, et al.
Published: (2025)
A Technical Note on the Implementation and Use of PDCS
by: Lin, Zhenwei, et al.
Published: (2026)
by: Lin, Zhenwei, et al.
Published: (2026)
Robust Deterministic Policies for Markov Decision Processes under Budgeted Uncertainty
by: Wu, Fei, et al.
Published: (2024)
by: Wu, Fei, et al.
Published: (2024)
Robust Markov Decision Processes on Continuous State Spaces
by: Li, Mengmeng, et al.
Published: (2026)
by: Li, Mengmeng, et al.
Published: (2026)
Beyond $\mathcal{O}(\sqrt{T})$ Regret: Decoupling Learning and Decision-making in Online Linear Programming
by: Gao, Wenzhi, et al.
Published: (2025)
by: Gao, Wenzhi, et al.
Published: (2025)
Computational Hardness of Static Distributionally Robust Markov Decision Processes
by: Li, Yan
Published: (2025)
by: Li, Yan
Published: (2025)
Accelerating Low-Rank Factorization-Based Semidefinite Programming Algorithms on GPU
by: Han, Qiushi, et al.
Published: (2024)
by: Han, Qiushi, et al.
Published: (2024)
Inexact Policy Iteration Methods for Large-Scale Markov Decision Processes
by: Gargiani, Matilde, et al.
Published: (2024)
by: Gargiani, Matilde, et al.
Published: (2024)
A Homogeneous Second-Order Descent Method for Nonconvex Optimization
by: Zhang, Chuwen, et al.
Published: (2022)
by: Zhang, Chuwen, et al.
Published: (2022)
Robust Markov Decision Processes: A Place Where AI and Formal Methods Meet
by: Suilen, Marnix, et al.
Published: (2024)
by: Suilen, Marnix, et al.
Published: (2024)
A Low-Rank ADMM Splitting Approach for Semidefinite Programming
by: Han, Qiushi, et al.
Published: (2024)
by: Han, Qiushi, et al.
Published: (2024)
Uniformly Optimal and Parameter-free First-order Methods for Convex and Function-constrained Optimization
by: Deng, Qi, et al.
Published: (2024)
by: Deng, Qi, et al.
Published: (2024)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025)
by: Misra, Rahul, et al.
Published: (2025)
An MILP-Based Solution Scheme for Factored and Robust Factored Markov Decision Processes
by: Liu, Huikang, et al.
Published: (2024)
by: Liu, Huikang, et al.
Published: (2024)
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
by: Kitamura, Toshinori, et al.
Published: (2024)
by: Kitamura, Toshinori, et al.
Published: (2024)
Gradient Methods with Online Scaling
by: Gao, Wenzhi, et al.
Published: (2024)
by: Gao, Wenzhi, et al.
Published: (2024)
An Enhanced ADMM-based Interior Point Method for Linear and Conic Optimization
by: Deng, Qi, et al.
Published: (2022)
by: Deng, Qi, et al.
Published: (2022)
Decentralized Gradient-Free Methods for Stochastic Non-Smooth Non-Convex Optimization
by: Lin, Zhenwei, et al.
Published: (2023)
by: Lin, Zhenwei, et al.
Published: (2023)
Efficient Algorithms for Robust Markov Decision Processes with $s$-Rectangular Ambiguity Sets
by: Ho, Chin Pang, et al.
Published: (2026)
by: Ho, Chin Pang, et al.
Published: (2026)
Trust Region Methods For Nonconvex Stochastic Optimization Beyond Lipschitz Smoothness
by: Xie, Chenghan, et al.
Published: (2023)
by: Xie, Chenghan, et al.
Published: (2023)
Gradient Methods with Online Scaling Part I. Theoretical Foundations
by: Gao, Wenzhi, et al.
Published: (2025)
by: Gao, Wenzhi, et al.
Published: (2025)
Gradient Methods with Online Scaling Part II. Practical Aspects
by: Chu, Ya-Chi, et al.
Published: (2025)
by: Chu, Ya-Chi, et al.
Published: (2025)
Robust Reward Design for Markov Decision Processes
by: Wu, Shuo, et al.
Published: (2024)
by: Wu, Shuo, et al.
Published: (2024)
Bounding the Difference between the Values of Robust and Non-Robust Markov Decision Problems
by: Neufeld, Ariel, et al.
Published: (2023)
by: Neufeld, Ariel, et al.
Published: (2023)
Restarted Primal-Dual Hybrid Conjugate Gradient Method for Large-Scale Quadratic Programming
by: Huang, Yicheng, et al.
Published: (2024)
by: Huang, Yicheng, et al.
Published: (2024)
Semismooth Newton Methods for Risk-Averse Markov Decision Processes
by: Gargiani, Matilde, et al.
Published: (2025)
by: Gargiani, Matilde, et al.
Published: (2025)
Adaptively Robust LLM Inference Optimization under Prediction Uncertainty
by: Chen, Zixi, et al.
Published: (2025)
by: Chen, Zixi, et al.
Published: (2025)
Revisiting Randomized Smoothing: Nonsmooth Nonconvex Optimization Beyond Global Lipschitz Continuity
by: Xia, Jingfan, et al.
Published: (2025)
by: Xia, Jingfan, et al.
Published: (2025)
Transition Uncertainties in Constrained Markov Decision Models: A Robust Optimization Approach
by: Varagapriya, V
Published: (2025)
by: Varagapriya, V
Published: (2025)
Absorbing Markov Decision Processes
by: Dufour, François, et al.
Published: (2023)
by: Dufour, François, et al.
Published: (2023)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
by: Choi, Jimin, et al.
Published: (2025)
by: Choi, Jimin, et al.
Published: (2025)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
by: Wang, Shengbo, et al.
Published: (2025)
by: Wang, Shengbo, et al.
Published: (2025)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
by: Xu, Mingyuan, et al.
Published: (2026)
by: Xu, Mingyuan, et al.
Published: (2026)
Absorbing Markov Decision Processes: Geometric Properties and Sufficiency of Finite Mixtures of Deterministic Policies
by: Dufour, Francois, et al.
Published: (2025)
by: Dufour, Francois, et al.
Published: (2025)
Mean Field Markov Decision Processes
by: Bäuerle, Nicole
Published: (2021)
by: Bäuerle, Nicole
Published: (2021)
Quantum Markov Decision Processes: General Theory, Approximations, and Classes of Policies
by: Saldi, Naci, et al.
Published: (2024)
by: Saldi, Naci, et al.
Published: (2024)
Similar Items
-
Achieving Instance-dependent Sample Complexity for Constrained Markov Decision Process
by: Jiang, Jiashuo, et al.
Published: (2024) -
A Homogenization Approach for Gradient-Dominated Stochastic Optimization
by: Tan, Jiyuan, et al.
Published: (2023) -
Tractable Robust Markov Decision Processes
by: Grand-Clément, Julien, et al.
Published: (2024) -
Decoupling Learning and Decision-Making: Breaking the $\mathcal{O}(\sqrt{T})$ Barrier in Online Resource Allocation with First-Order Methods
by: Gao, Wenzhi, et al.
Published: (2024) -
A Practical GPU-Enhanced Matrix-Free Primal-Dual Method for Large-Scale Conic Programs
by: Lin, Zhenwei, et al.
Published: (2025)