Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Dongsheng, Zhang, Kaiqing, Duan, Jiali, Başar, Tamer, Jovanović, Mihailo R. |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Variational Contraction Conditions for Iterative Algorithms in Multi-Population Discrete-Time Regularized Mean-Field Games
by: Aydın, Uğur, et al.
Published: (2026)
by: Aydın, Uğur, et al.
Published: (2026)
Automated algorithm design via Nevanlinna-Pick interpolation
by: Ozaslan, Ibrahim K., et al.
Published: (2025)
by: Ozaslan, Ibrahim K., et al.
Published: (2025)
Best Response Strategies for Asymmetric Sensing in Linear-Quadratic Differential Games
by: Aggarwal, Shubham, et al.
Published: (2024)
by: Aggarwal, Shubham, et al.
Published: (2024)
Stability properties of gradient flow dynamics for the symmetric low-rank matrix factorization problem
by: Mohammadi, Hesameddin, et al.
Published: (2024)
by: Mohammadi, Hesameddin, et al.
Published: (2024)
Accelerated forward-backward and Douglas-Rachford splitting dynamics
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
From exponential to finite/fixed-time stability: Applications to optimization
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
Extremum and Nash Equilibrium Seeking with Delays and PDEs: Designs & Applications
by: Oliveira, Tiago Roux, et al.
Published: (2024)
by: Oliveira, Tiago Roux, et al.
Published: (2024)
Tradeoffs between convergence rate and noise amplification for momentum-based accelerated optimization algorithms
by: Mohammadi, Hesameddin, et al.
Published: (2022)
by: Mohammadi, Hesameddin, et al.
Published: (2022)
The Silence that Speaks: Neural Estimation via Communication Gaps
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Disentangling Resilience from Robustness: Contextual Dualism, Interactionism, and Game-Theoretic Paradigms
by: Zhu, Quanyan, et al.
Published: (2024)
by: Zhu, Quanyan, et al.
Published: (2024)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Near Optimal Convergence to Coarse Correlated Equilibrium in General-Sum Markov Games
by: Yorulmaz, Asrin Efe, et al.
Published: (2025)
by: Yorulmaz, Asrin Efe, et al.
Published: (2025)
Convergence analysis of primal-dual augmented Lagrangian methods and duality theory
by: Dolgopolik, M. V.
Published: (2024)
by: Dolgopolik, M. V.
Published: (2024)
Automated algorithm design for convex optimization problems with linear equality constraints
by: Ozaslan, Ibrahim K., et al.
Published: (2025)
by: Ozaslan, Ibrahim K., et al.
Published: (2025)
Nash Equilibrium Seeking for Noncooperative Duopoly Games via Event-Triggered Control
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2024)
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2024)
Lie-Bracket Nash Equilibrium Seeking with Bounded Update Rates for Noncooperative Games
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2025)
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2025)
Distributed Event-Triggered Nash Equilibrium Seeking for Noncooperative Games
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2025)
by: Rodrigues, Victor Hugo Pereira, et al.
Published: (2025)
Stochastic-Robust Planning of Networked Hydrogen-Electrical Microgrids: A Study on Induced Refueling Demand
by: Sun, Xunhang, et al.
Published: (2024)
by: Sun, Xunhang, et al.
Published: (2024)
Data-Driven Two-Stage Distributionally Robust Dispatch of Multi-Energy Microgrid
by: Sun, Xunhang, et al.
Published: (2025)
by: Sun, Xunhang, et al.
Published: (2025)
Policy Optimization for PDE Control with a Warm Start
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
Communication and Control Co-design in Non-cooperative Games
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Alignment of large language models with constrained learning
by: Zhang, Botong, et al.
Published: (2025)
by: Zhang, Botong, et al.
Published: (2025)
Harnessing Information in Incentive Design
by: Velicheti, Raj Kiriti, et al.
Published: (2025)
by: Velicheti, Raj Kiriti, et al.
Published: (2025)
Semantic Communication in Multi-team Dynamic Games: A Mean Field Perspective
by: Aggarwal, Shubham, et al.
Published: (2024)
by: Aggarwal, Shubham, et al.
Published: (2024)
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
by: Lan, Guanghui, et al.
Published: (2024)
by: Lan, Guanghui, et al.
Published: (2024)
An agent-based decentralized threshold policy finding the constrained shortest paths
by: Rosset, Francesca, et al.
Published: (2023)
by: Rosset, Francesca, et al.
Published: (2023)
Robust Constrained Consensus and Inequality-constrained Distributed Optimization with Guaranteed Differential Privacy and Accurate Convergence
by: Wang, Yongqiang, et al.
Published: (2024)
by: Wang, Yongqiang, et al.
Published: (2024)
Distributed Offloading in Multi-Access Edge Computing Systems: A Mean-Field Perspective
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
A Bundle-based Augmented Lagrangian Framework: Algorithm, Convergence, and Primal-dual Principles
by: Liao, Feng-Yi, et al.
Published: (2025)
by: Liao, Feng-Yi, et al.
Published: (2025)
An efficient primal dual semismooth Newton method for semidefinite programming
by: Deng, Zhanwang, et al.
Published: (2025)
by: Deng, Zhanwang, et al.
Published: (2025)
A Stackelberg Game Framework with Drainability Guardrails for Pricing and Scaling in Multi-Tenant GPU Cloud Platforms
by: Yan, Junji, et al.
Published: (2026)
by: Yan, Junji, et al.
Published: (2026)
On the convergence analysis of the decentralized projected gradient descent method
by: Choi, Woocheol, et al.
Published: (2023)
by: Choi, Woocheol, et al.
Published: (2023)
Convergence result for the gradient-push algorithm and its application to boost up the Push-DIging algorithm
by: Choi, Hyogi, et al.
Published: (2024)
by: Choi, Hyogi, et al.
Published: (2024)
Integral control of the proximal gradient method for unbiased sparse optimization
by: Cerone, V., et al.
Published: (2025)
by: Cerone, V., et al.
Published: (2025)
A primal-dual interior point trust region method for second-order stationary points of Riemannian inequality-constrained optimization problems
by: Obara, Mitsuaki, et al.
Published: (2025)
by: Obara, Mitsuaki, et al.
Published: (2025)
Nested Extremum Seeking Converges to Stackelberg Equilibrium
by: Ratto, Brad, et al.
Published: (2026)
by: Ratto, Brad, et al.
Published: (2026)
Newton and interior-point methods for (constrained) nonconvex-nonconcave minmax optimization with stability and instability guarantees
by: Chinchilla, Raphael, et al.
Published: (2022)
by: Chinchilla, Raphael, et al.
Published: (2022)
Similar Items
-
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023) -
Variational Contraction Conditions for Iterative Algorithms in Multi-Population Discrete-Time Regularized Mean-Field Games
by: Aydın, Uğur, et al.
Published: (2026) -
Automated algorithm design via Nevanlinna-Pick interpolation
by: Ozaslan, Ibrahim K., et al.
Published: (2025) -
Best Response Strategies for Asymmetric Sensing in Linear-Quadratic Differential Games
by: Aggarwal, Shubham, et al.
Published: (2024) -
Stability properties of gradient flow dynamics for the symmetric low-rank matrix factorization problem
by: Mohammadi, Hesameddin, et al.
Published: (2024)