Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Feiran, You, Keyou |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Linear Convergence of Data-Enabled Policy Optimization for Linear Quadratic Tracking
by: Kang, Shubo, et al.
Published: (2024)
by: Kang, Shubo, et al.
Published: (2024)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
Asynchronous Parallel Policy Gradient Methods for the Linear Quadratic Regulator
by: Sha, Xingyu, et al.
Published: (2024)
by: Sha, Xingyu, et al.
Published: (2024)
A Bayesian Perspective on the Data-Driven LQR
by: Schwaller, Thierry, et al.
Published: (2026)
by: Schwaller, Thierry, et al.
Published: (2026)
Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Regularization for Covariance Parameterization of Direct Data-Driven LQR Control
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
On the (almost) Global Exponential Convergence of the Overparameterized Policy Optimization for the LQR Problem
by: Wafi, Moh Kamalul, et al.
Published: (2025)
by: Wafi, Moh Kamalul, et al.
Published: (2025)
Policy Gradient Bounds in Multitask LQR
by: Stamouli, Charis, et al.
Published: (2025)
by: Stamouli, Charis, et al.
Published: (2025)
Direct Data-Driven Linear Quadratic Tracking via Policy Optimization
by: Kang, Shubo, et al.
Published: (2026)
by: Kang, Shubo, et al.
Published: (2026)
On the Effect of Quadratic Regularization in Direct Data-Driven LQR
by: Klädtke, Manuel, et al.
Published: (2026)
by: Klädtke, Manuel, et al.
Published: (2026)
Power-Constrained Policy Gradient Methods for LQR
by: Verma, Ashwin, et al.
Published: (2025)
by: Verma, Ashwin, et al.
Published: (2025)
A Globally Convergent Policy Gradient Method for Linear Quadratic Gaussian (LQG) Control
by: Sadamoto, Tomonori, et al.
Published: (2023)
by: Sadamoto, Tomonori, et al.
Published: (2023)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
Reliably Learn to Trim Multiparametric Quadratic Programs via Constraint Removal
by: Hou, Zhinan, et al.
Published: (2024)
by: Hou, Zhinan, et al.
Published: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
A Strong Duality Result for Constrained POMDPs with Multiple Cooperative Agents
by: Khan, Nouman, et al.
Published: (2023)
by: Khan, Nouman, et al.
Published: (2023)
Beyond Quadratic Costs in LQR: Bregman Divergence Control
by: Hassibi, Babak, et al.
Published: (2025)
by: Hassibi, Babak, et al.
Published: (2025)
ReLU Networks for Model Predictive Control: Network Complexity and Performance Guarantees
by: Li, Xingchen, et al.
Published: (2026)
by: Li, Xingchen, et al.
Published: (2026)
Small-Disturbance Input-to-State Stability of Perturbed Gradient Flows: Applications to LQR Problem
by: Cui, Leilei, et al.
Published: (2023)
by: Cui, Leilei, et al.
Published: (2023)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Policy Gradient Method for LQG Control via Input-Output-History Representation: Convergence to $O(ε)$-Stationary Points
by: Sadamoto, Tomonori, et al.
Published: (2025)
by: Sadamoto, Tomonori, et al.
Published: (2025)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
by: Carnevale, Guido, et al.
Published: (2024)
by: Carnevale, Guido, et al.
Published: (2024)
Strong Duality in Risk-Constrained Nonconvex Functional Programming
by: Kalogerias, Dionysis, et al.
Published: (2022)
by: Kalogerias, Dionysis, et al.
Published: (2022)
Stochastic LQR Design With Disturbance Preview
by: Liu, Jietian, et al.
Published: (2024)
by: Liu, Jietian, et al.
Published: (2024)
The Distributionally Robust Infinite-Horizon LQR
by: Hajar, Joudi, et al.
Published: (2024)
by: Hajar, Joudi, et al.
Published: (2024)
Mixed Regular and Impulsive Sampled-data LQR
by: Daafouz, Jamal, et al.
Published: (2024)
by: Daafouz, Jamal, et al.
Published: (2024)
An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control
by: Persson, Niklas, et al.
Published: (2025)
by: Persson, Niklas, et al.
Published: (2025)
Almost Sure Convergence of Networked Policy Gradient over Time-Varying Networks in Markov Potential Games
by: Aydin, Sarper, et al.
Published: (2024)
by: Aydin, Sarper, et al.
Published: (2024)
An optimistic planning algorithm for switched discrete-time LQR
by: Granzotto, Mathieu, et al.
Published: (2025)
by: Granzotto, Mathieu, et al.
Published: (2025)
Harnessing Data for Accelerating Model Predictive Control by Constraint Removal
by: Hou, Zhinan, et al.
Published: (2024)
by: Hou, Zhinan, et al.
Published: (2024)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR
by: Zeng, Xiong, et al.
Published: (2024)
by: Zeng, Xiong, et al.
Published: (2024)
A System Level Approach to LQR Control of the Diffusion Equation
by: McCurdy, Addie, et al.
Published: (2025)
by: McCurdy, Addie, et al.
Published: (2025)
LQR based $ω-$stabilization of a heat equation with memory
by: Sistla, Bhargav Pavan Kumar, et al.
Published: (2025)
by: Sistla, Bhargav Pavan Kumar, et al.
Published: (2025)
Adaptive Control of Unknown Linear Switched Systems via Policy Gradient Methods
by: Laurent, Felix, et al.
Published: (2026)
by: Laurent, Felix, et al.
Published: (2026)
Analyzing the Impact of Computation in Adaptive Dynamic Programming for Stochastic LQR Problem
by: Cao, Wenhan, et al.
Published: (2024)
by: Cao, Wenhan, et al.
Published: (2024)
Adapt and Stabilize, Then Learn and Optimize: A New Approach to Adaptive LQR
by: Fisher, Peter A., et al.
Published: (2025)
by: Fisher, Peter A., et al.
Published: (2025)
Distributionally Robust Regret Optimal LQR with Common Stage-Law Ambiguity
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2026)
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2026)
On Convergence of the Iteratively Preconditioned Gradient-Descent (IPG) Observer
by: Chakrabarti, Kushal, et al.
Published: (2024)
by: Chakrabarti, Kushal, et al.
Published: (2024)
Similar Items
-
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024) -
Linear Convergence of Data-Enabled Policy Optimization for Linear Quadratic Tracking
by: Kang, Shubo, et al.
Published: (2024) -
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025) -
Asynchronous Parallel Policy Gradient Methods for the Linear Quadratic Regulator
by: Sha, Xingyu, et al.
Published: (2024) -
A Bayesian Perspective on the Data-Driven LQR
by: Schwaller, Thierry, et al.
Published: (2026)