Achieving Zero Constraint Violation for Constrained Reinforcement Learning via Conservative Natural Policy Gradient Primal-Dual Algorithm
Fuente:
arXiv
Saved in:
| Main Authors: | Bai, Qinbo, Bedi, Amrit Singh, Aggarwal, Vaneet |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm
by: Bai, Qinbo, et al.
Published: (2021)
by: Bai, Qinbo, et al.
Published: (2021)
Learning General Parameterized Policies for Infinite Horizon Average Reward Constrained MDPs via Primal-Dual Policy Gradient Algorithm
by: Bai, Qinbo, et al.
Published: (2024)
by: Bai, Qinbo, et al.
Published: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Global Convergence for Average Reward Constrained MDPs with Primal-Dual Actor Critic Algorithm
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms
by: Aggarwal, Vaneet, et al.
Published: (2024)
by: Aggarwal, Vaneet, et al.
Published: (2024)
Regret Analysis of Policy Gradient Algorithm for Infinite Horizon Average Reward Markov Decision Processes
by: Bai, Qinbo, et al.
Published: (2023)
by: Bai, Qinbo, et al.
Published: (2023)
A Distributed Primal-Dual Method for Constrained Multi-agent Reinforcement Learning with General Parameterization
by: Kahe, Ali, et al.
Published: (2024)
by: Kahe, Ali, et al.
Published: (2024)
Sampled-Data Primal-Dual Gradient Dynamics in Model Predictive Control
by: Moriyasu, Ryuta, et al.
Published: (2024)
by: Moriyasu, Ryuta, et al.
Published: (2024)
Exponential Stability of Primal-Dual Gradient Dynamics with Non-Strong Convexity
by: Chen, Xin, et al.
Published: (2019)
by: Chen, Xin, et al.
Published: (2019)
Pricing Short-Circuit Current via a Primal-Dual Formulation for Preserving Integrality Constraints
by: Wang, Peng, et al.
Published: (2025)
by: Wang, Peng, et al.
Published: (2025)
Closing the Gap: Achieving Global Convergence (Last Iterate) of Actor-Critic under Markovian Sampling with Neural Network Parametrization
by: Gaur, Mudit, et al.
Published: (2024)
by: Gaur, Mudit, et al.
Published: (2024)
A Mixing-Accelerated Primal-Dual Proximal Algorithm for Distributed Nonconvex Optimization
by: Ou, Zichong, et al.
Published: (2023)
by: Ou, Zichong, et al.
Published: (2023)
On The Sample Complexity Bounds In Bilevel Reinforcement Learning
by: Gaur, Mudit, et al.
Published: (2025)
by: Gaur, Mudit, et al.
Published: (2025)
DiffOP: Reinforcement Learning of Optimization-Based Control Policies via Implicit Policy Gradients
by: Bian, Yuexin, et al.
Published: (2024)
by: Bian, Yuexin, et al.
Published: (2024)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Learning-Augmented Primal-Dual Control Design for Secondary Frequency Regulation
by: Yu, Yixuan, et al.
Published: (2026)
by: Yu, Yixuan, et al.
Published: (2026)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
A Common Lyapunov Matrix Approach to the Exponential Stability of Augmented Primal-Dual Gradient Flow as LPV Systems
by: Li, Mengmou, et al.
Published: (2026)
by: Li, Mengmou, et al.
Published: (2026)
BAGEL: Projection-Free Algorithm for Adversarially Constrained Online Convex Optimization
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Accelerating Quantum Reinforcement Learning with a Quantum Natural Policy Gradient Based Approach
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Safe Exploration in Reinforcement Learning: Training Backup Control Barrier Functions with Zero Training Time Safety Violations
by: Rabiee, Pedram, et al.
Published: (2023)
by: Rabiee, Pedram, et al.
Published: (2023)
Delay-Robust Primal-Dual Dynamics for Distributed Optimization
by: Şen, Gökçen Devlet, et al.
Published: (2026)
by: Şen, Gökçen Devlet, et al.
Published: (2026)
Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Off-Policy Reinforcement Learning with Anytime Safety Guarantees via Robust Safe Gradient Flow
by: Mestres, Pol, et al.
Published: (2025)
by: Mestres, Pol, et al.
Published: (2025)
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning
by: Baheri, Ali
Published: (2025)
by: Baheri, Ali
Published: (2025)
An Overview and Comparison of Spectral Bundle Methods for Primal and Dual Semidefinite Programs
by: Liao, Feng-Yi, et al.
Published: (2023)
by: Liao, Feng-Yi, et al.
Published: (2023)
A Dual Perspective of Reinforcement Learning for Imposing Policy Constraints
by: De Cooman, Bram, et al.
Published: (2024)
by: De Cooman, Bram, et al.
Published: (2024)
Every Call is Precious: Global Optimization of Black-Box Functions with Unknown Lipschitz Constants
by: Fourati, Fares, et al.
Published: (2025)
by: Fourati, Fares, et al.
Published: (2025)
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
by: Yifru, Lunet, et al.
Published: (2024)
by: Yifru, Lunet, et al.
Published: (2024)
Consensus Planning with Primal, Dual, and Proximal Agents
by: Maggiar, Alvaro, et al.
Published: (2024)
by: Maggiar, Alvaro, et al.
Published: (2024)
A Distributed Gradient-based Algorithm for Optimization Problems with Coupled Equality Constraints
by: Qiu, Chenyang, et al.
Published: (2025)
by: Qiu, Chenyang, et al.
Published: (2025)
Mean-Field Approximation of Cooperative Constrained Multi-Agent Reinforcement Learning (CMARL)
by: Mondal, Washim Uddin, et al.
Published: (2022)
by: Mondal, Washim Uddin, et al.
Published: (2022)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Adaptive Event-Triggered Policy Gradient for Multi-Agent Reinforcement Learning
by: Siddique, Umer, et al.
Published: (2025)
by: Siddique, Umer, et al.
Published: (2025)
Sample-Optimal Zero-Violation Safety For Continuous Control
by: Ray, Ritabrata, et al.
Published: (2024)
by: Ray, Ritabrata, et al.
Published: (2024)
Stability-Certified On-Policy Data-Driven LQR via Recursive Learning and Policy Gradient
by: Sforni, Lorenzo, et al.
Published: (2024)
by: Sforni, Lorenzo, et al.
Published: (2024)
A Bundle-based Augmented Lagrangian Framework: Algorithm, Convergence, and Primal-dual Principles
by: Liao, Feng-Yi, et al.
Published: (2025)
by: Liao, Feng-Yi, et al.
Published: (2025)
Distributed Constraint-coupled Resource Allocation: Anytime Feasibility and Violation Robustness
by: Wu, Wenwen, et al.
Published: (2025)
by: Wu, Wenwen, et al.
Published: (2025)
On The Global Convergence Of Online RLHF With Neural Parametrization
by: Gaur, Mudit, et al.
Published: (2024)
by: Gaur, Mudit, et al.
Published: (2024)
Similar Items
-
Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm
by: Bai, Qinbo, et al.
Published: (2021) -
Learning General Parameterized Policies for Infinite Horizon Average Reward Constrained MDPs via Primal-Dual Policy Gradient Algorithm
by: Bai, Qinbo, et al.
Published: (2024) -
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023) -
Global Convergence for Average Reward Constrained MDPs with Primal-Dual Actor Critic Algorithm
by: Xu, Yang, et al.
Published: (2025) -
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms
by: Aggarwal, Vaneet, et al.
Published: (2024)