Stability-Certified On-Policy Data-Driven LQR via Recursive Learning and Policy Gradient
Fuente:
arXiv
Saved in:
| Main Authors: | Sforni, Lorenzo, Carnevale, Guido, Notarnicola, Ivano, Notarstefano, Giuseppe |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
by: Carnevale, Guido, et al.
Published: (2024)
by: Carnevale, Guido, et al.
Published: (2024)
Data-Driven Distributed Optimization via Aggregative Tracking and Deep-Learning
by: Brumali, Riccardo, et al.
Published: (2025)
by: Brumali, Riccardo, et al.
Published: (2025)
MR-ARL: Model Reference Adaptive Reinforcement Learning for Robustly Stable On-Policy Data-Driven LQR
by: Borghesi, Marco, et al.
Published: (2024)
by: Borghesi, Marco, et al.
Published: (2024)
DATA-DRIVEN PRONTO: a Model-free Solution for Numerical Optimal Control
by: Borghesi, Marco, et al.
Published: (2025)
by: Borghesi, Marco, et al.
Published: (2025)
Safe Control of Feedback-Interconnected Systems via Singular Perturbations
by: Di Gregorio, Stefano, et al.
Published: (2026)
by: Di Gregorio, Stefano, et al.
Published: (2026)
Nonconvex Distributed Feedback Optimization for Aggregative Cooperative Robotics
by: Carnevale, Guido, et al.
Published: (2023)
by: Carnevale, Guido, et al.
Published: (2023)
Nonlinear MPC for Feedback-Interconnected Systems: a Suboptimal and Reduced-Order Model Approach
by: Di Gregorio, Stefano, et al.
Published: (2025)
by: Di Gregorio, Stefano, et al.
Published: (2025)
Extremum Seeking Tracking for Derivative-free Distributed Optimization
by: Mimmo, Nicola, et al.
Published: (2021)
by: Mimmo, Nicola, et al.
Published: (2021)
Multi-Robot Target Monitoring and Encirclement via Triggered Distributed Feedback Optimization
by: Pichierri, Lorenzo, et al.
Published: (2024)
by: Pichierri, Lorenzo, et al.
Published: (2024)
Accelerated ADMM: Automated Parameter Tuning and Improved Linear Convergence
by: Tavakoli, Meisam, et al.
Published: (2025)
by: Tavakoli, Meisam, et al.
Published: (2025)
Policy Gradient Bounds in Multitask LQR
by: Stamouli, Charis, et al.
Published: (2025)
by: Stamouli, Charis, et al.
Published: (2025)
On Sufficient Richness for Linear Time-Invariant Systems
by: Borghesi, Marco, et al.
Published: (2025)
by: Borghesi, Marco, et al.
Published: (2025)
Tracking-based distributed equilibrium seeking for aggregative games
by: Carnevale, Guido, et al.
Published: (2022)
by: Carnevale, Guido, et al.
Published: (2022)
Distributed equilibrium seeking in aggregative games: linear convergence under singular perturbations lens
by: Carnevale, Guido, et al.
Published: (2025)
by: Carnevale, Guido, et al.
Published: (2025)
Policy Gradient for LQR with Domain Randomization
by: Fujinami, Tesshu, et al.
Published: (2025)
by: Fujinami, Tesshu, et al.
Published: (2025)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
ADMM-Tracking Gradient for Distributed Optimization over Asynchronous and Unreliable Networks
by: Carnevale, Guido, et al.
Published: (2023)
by: Carnevale, Guido, et al.
Published: (2023)
Convergence Guarantees of Model-free Policy Gradient Methods for LQR with Stochastic Data
by: Song, Bowen, et al.
Published: (2025)
by: Song, Bowen, et al.
Published: (2025)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Data-Driven Stabilization of Continuous-Time LTI Systems from Noisy Input-Output Data
by: Bosso, Alessandro, et al.
Published: (2025)
by: Bosso, Alessandro, et al.
Published: (2025)
A Stochastic Gradient Descent Approach to Design Policy Gradient Methods for LQR
by: Song, Bowen, et al.
Published: (2026)
by: Song, Bowen, et al.
Published: (2026)
A Unifying System Theory Framework for Distributed Optimization and Games
by: Carnevale, Guido, et al.
Published: (2024)
by: Carnevale, Guido, et al.
Published: (2024)
A Tutorial on Distributed Optimization for Cooperative Robotics: from Setups and Algorithms to Toolboxes and Research Directions
by: Testa, Andrea, et al.
Published: (2023)
by: Testa, Andrea, et al.
Published: (2023)
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
On Globally Optimal Stochastic Policy Gradient Methods for Domain Randomized LQR Synthesis
by: Nguyen-Le, Alex, et al.
Published: (2026)
by: Nguyen-Le, Alex, et al.
Published: (2026)
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
by: Song, Bowen, et al.
Published: (2025)
by: Song, Bowen, et al.
Published: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Data-Driven Control of Continuous-Time LTI Systems via Non-Minimal Realizations
by: Bosso, Alessandro, et al.
Published: (2025)
by: Bosso, Alessandro, et al.
Published: (2025)
On Reward-Balancing Methods for Reinforcement Learning
by: Baroncini, Simone, et al.
Published: (2026)
by: Baroncini, Simone, et al.
Published: (2026)
Derivative-Free Data-Driven Control of Continuous-Time Linear Time-Invariant Systems
by: Bosso, Alessandro, et al.
Published: (2024)
by: Bosso, Alessandro, et al.
Published: (2024)
A Distributed Bilevel Framework for the Macroscopic Optimization of Multi-Agent Systems
by: Brumali, Riccardo, et al.
Published: (2026)
by: Brumali, Riccardo, et al.
Published: (2026)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
by: Long, Kehan, et al.
Published: (2025)
by: Long, Kehan, et al.
Published: (2025)
Modular Distributed Nonconvex Learning with Error Feedback
by: Carnevale, Guido, et al.
Published: (2025)
by: Carnevale, Guido, et al.
Published: (2025)
A Bayesian Perspective on the Data-Driven LQR
by: Schwaller, Thierry, et al.
Published: (2026)
by: Schwaller, Thierry, et al.
Published: (2026)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Affine-coupled Distributed Optimization via Distributed Proximal Jacobian ADMM with Quantized Communication
by: Du, Xu, et al.
Published: (2026)
by: Du, Xu, et al.
Published: (2026)
On the (almost) Global Exponential Convergence of the Overparameterized Policy Optimization for the LQR Problem
by: Wafi, Moh Kamalul, et al.
Published: (2025)
by: Wafi, Moh Kamalul, et al.
Published: (2025)
On the Effect of Quadratic Regularization in Direct Data-Driven LQR
by: Klädtke, Manuel, et al.
Published: (2026)
by: Klädtke, Manuel, et al.
Published: (2026)
Small-Disturbance Input-to-State Stability of Perturbed Gradient Flows: Applications to LQR Problem
by: Cui, Leilei, et al.
Published: (2023)
by: Cui, Leilei, et al.
Published: (2023)
Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR
by: Zeng, Xiong, et al.
Published: (2024)
by: Zeng, Xiong, et al.
Published: (2024)
Similar Items
-
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
by: Carnevale, Guido, et al.
Published: (2024) -
Data-Driven Distributed Optimization via Aggregative Tracking and Deep-Learning
by: Brumali, Riccardo, et al.
Published: (2025) -
MR-ARL: Model Reference Adaptive Reinforcement Learning for Robustly Stable On-Policy Data-Driven LQR
by: Borghesi, Marco, et al.
Published: (2024) -
DATA-DRIVEN PRONTO: a Model-free Solution for Numerical Optimal Control
by: Borghesi, Marco, et al.
Published: (2025) -
Safe Control of Feedback-Interconnected Systems via Singular Perturbations
by: Di Gregorio, Stefano, et al.
Published: (2026)