Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
Fuente:
arXiv
Guardado en:
| Autores principales: | Rozada, Sergio, Ding, Dongsheng, Marques, Antonio G., Ribeiro, Alejandro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
por: Cheng, Ziheng, et al.
Publicado: (2025)
por: Cheng, Ziheng, et al.
Publicado: (2025)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
por: Chen, Weiqin, et al.
Publicado: (2024)
por: Chen, Weiqin, et al.
Publicado: (2024)
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2022)
por: Ding, Dongsheng, et al.
Publicado: (2022)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
por: Ozaslan, Ibrahim K., et al.
Publicado: (2024)
por: Ozaslan, Ibrahim K., et al.
Publicado: (2024)
Matrix Low-Rank Approximation For Policy Gradient Methods
por: Rozada, Sergio, et al.
Publicado: (2024)
por: Rozada, Sergio, et al.
Publicado: (2024)
Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients
por: Alvo, Matias, et al.
Publicado: (2026)
por: Alvo, Matias, et al.
Publicado: (2026)
Primal-Dual Spectral Representation for Off-policy Evaluation
por: Hu, Yang, et al.
Publicado: (2024)
por: Hu, Yang, et al.
Publicado: (2024)
WARP: A Benchmark for Primal-Dual Warm-Starting of Interior-Point Solvers
por: Suri, Dhruv, et al.
Publicado: (2026)
por: Suri, Dhruv, et al.
Publicado: (2026)
Resilient Constrained Reinforcement Learning
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Constrained Diffusion Models via Dual Training
por: Khalafi, Shervin, et al.
Publicado: (2024)
por: Khalafi, Shervin, et al.
Publicado: (2024)
Self-Certifying Primal-Dual Optimization Proxies for Large-Scale Batch Economic Dispatch
por: Klamkin, Michael, et al.
Publicado: (2025)
por: Klamkin, Michael, et al.
Publicado: (2025)
Primal-Dual Bundle Methods for Linear Equality-Constrained Problems
por: Zheng, Zhuoqing, et al.
Publicado: (2025)
por: Zheng, Zhuoqing, et al.
Publicado: (2025)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
por: Ganesh, Swetha, et al.
Publicado: (2024)
por: Ganesh, Swetha, et al.
Publicado: (2024)
On the convergence of doubly stochastic Primal-Dual Hybrid Gradient Method
por: Xiao, Yiheng, et al.
Publicado: (2026)
por: Xiao, Yiheng, et al.
Publicado: (2026)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
por: Xiao, Minheng, et al.
Publicado: (2024)
por: Xiao, Minheng, et al.
Publicado: (2024)
Delightful Policy Gradient
por: Osband, Ian
Publicado: (2026)
por: Osband, Ian
Publicado: (2026)
Delightful Distributed Policy Gradient
por: Osband, Ian
Publicado: (2026)
por: Osband, Ian
Publicado: (2026)
Policy-based Primal-Dual Methods for Concave CMDP with Variance Reduction
por: Ying, Donghao, et al.
Publicado: (2022)
por: Ying, Donghao, et al.
Publicado: (2022)
Quasi-Quadratic Gradient: A New Direction for Accelerating the BFGS Method in Quasi-Newton Optimization
por: Chiang, John
Publicado: (2026)
por: Chiang, John
Publicado: (2026)
An Overview and Comparison of Spectral Bundle Methods for Primal and Dual Semidefinite Programs
por: Liao, Feng-Yi, et al.
Publicado: (2023)
por: Liao, Feng-Yi, et al.
Publicado: (2023)
Restarted Primal-Dual Hybrid Conjugate Gradient Method for Large-Scale Quadratic Programming
por: Huang, Yicheng, et al.
Publicado: (2024)
por: Huang, Yicheng, et al.
Publicado: (2024)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
por: Basu, Debabrota, et al.
Publicado: (2025)
por: Basu, Debabrota, et al.
Publicado: (2025)
MUSIC: Accelerated Convergence for Distributed Optimization With Inexact and Exact Methods
por: Wu, Mou, et al.
Publicado: (2024)
por: Wu, Mou, et al.
Publicado: (2024)
Thinking Beyond Visibility: A Near-Optimal Policy Framework for Locally Interdependent Multi-Agent MDPs
por: DeWeese, Alex, et al.
Publicado: (2025)
por: DeWeese, Alex, et al.
Publicado: (2025)
Provable Offline Reinforcement Learning for Structured Cyclic MDPs
por: Lee, Kyungbok, et al.
Publicado: (2026)
por: Lee, Kyungbok, et al.
Publicado: (2026)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
An Accelerated Primal Dual Algorithm with Backtracking for Decentralized Constrained Optimization
por: Xu, Qiushui, et al.
Publicado: (2025)
por: Xu, Qiushui, et al.
Publicado: (2025)
Boosting Gradient Ascent for Continuous DR-submodular Maximization
por: Zhang, Qixin, et al.
Publicado: (2024)
por: Zhang, Qixin, et al.
Publicado: (2024)
Power-Constrained Policy Gradient Methods for LQR
por: Verma, Ashwin, et al.
Publicado: (2025)
por: Verma, Ashwin, et al.
Publicado: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
por: Zhang, Xiangyuan, et al.
Publicado: (2023)
On the Differentiability of the Primal-Dual Interior-Point Method
por: Tracy, Kevin, et al.
Publicado: (2024)
por: Tracy, Kevin, et al.
Publicado: (2024)
Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model
por: Mortensen, Oliver, et al.
Publicado: (2025)
por: Mortensen, Oliver, et al.
Publicado: (2025)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Double Momentum Method for Lower-Level Constrained Bilevel Optimization
por: Shi, Wanli, et al.
Publicado: (2024)
por: Shi, Wanli, et al.
Publicado: (2024)
WANCO: Weak Adversarial Networks for Constrained Optimization problems
por: Bao, Gang, et al.
Publicado: (2024)
por: Bao, Gang, et al.
Publicado: (2024)
Sampled-Data Primal-Dual Gradient Dynamics in Model Predictive Control
por: Moriyasu, Ryuta, et al.
Publicado: (2024)
por: Moriyasu, Ryuta, et al.
Publicado: (2024)
Exponential Stability of Primal-Dual Gradient Dynamics with Non-Strong Convexity
por: Chen, Xin, et al.
Publicado: (2019)
por: Chen, Xin, et al.
Publicado: (2019)
A Relaxed Primal-Dual Hybrid Gradient Method with Line Search
por: McManus, Alex, et al.
Publicado: (2025)
por: McManus, Alex, et al.
Publicado: (2025)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
por: Zhang, Runyu, et al.
Publicado: (2023)
por: Zhang, Runyu, et al.
Publicado: (2023)
Ejemplares similares
-
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023) -
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
por: Cheng, Ziheng, et al.
Publicado: (2025) -
Adaptive Primal-Dual Method for Safe Reinforcement Learning
por: Chen, Weiqin, et al.
Publicado: (2024) -
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2022) -
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
por: Ozaslan, Ibrahim K., et al.
Publicado: (2024)