On the (almost) Global Exponential Convergence of the Overparameterized Policy Optimization for the LQR Problem
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wafi, Moh Kamalul, de Oliveira, Arthur Castello B., Sontag, Eduardo D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Boundedness of solutions in feedback systems with antithetic controllers
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
von: de Oliveira, Arthur Castello Branco, et al.
Veröffentlicht: (2025)
von: de Oliveira, Arthur Castello Branco, et al.
Veröffentlicht: (2025)
Convergence Analysis of Gradient Flow for Overparameterized LQR Formulations
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2024)
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2024)
When is cumulative dose response monotonic? Analysis of incoherent feedforward motifs
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
System Identification on the Families of Auto-Regressive with Least-Square-Batch Algorithm
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
Estimation and Fault Detection on Hydraulic System with Adaptive-Scaling Kalman and Consensus Filtering
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
Filtering Module on Satellite Tracking
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)
Small-Disturbance Input-to-State Stability of Perturbed Gradient Flows: Applications to LQR Problem
von: Cui, Leilei, et al.
Veröffentlicht: (2023)
von: Cui, Leilei, et al.
Veröffentlicht: (2023)
Remarks on the Polyak-Lojasiewicz inequality and the convergence of gradient systems
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2025)
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2025)
A Passivity-Agnostic Framework for Distributed Adaptive Synchronization under Unknown Leader Dynamics
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
Fault-Tolerant Control Design in Scrubber Plant with Fault on Sensor Sensitivity
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2023)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2023)
Distributed Adaptive Control of Disturbed Interconnected Systems with High-Order Tuners
von: Wafi, Moh. Kamalul, et al.
Veröffentlicht: (2024)
von: Wafi, Moh. Kamalul, et al.
Veröffentlicht: (2024)
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
Discrete-Time State-Feedback Controller with Canonical Form on Inverted Pendulum (on a cart)
von: Widjiantoro, Bambang L., et al.
Veröffentlicht: (2023)
von: Widjiantoro, Bambang L., et al.
Veröffentlicht: (2023)
Distributed Estimation with Decentralized Control for Quadruple-Tank Process
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2023)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2023)
On incremental and semi-global exponential stability of gradient flows satisfying generalized Łojasiewicz inequalities
von: Oliveira, Andreas, et al.
Veröffentlicht: (2026)
von: Oliveira, Andreas, et al.
Veröffentlicht: (2026)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
Model Reference Adaptive Control of Networked Systems with State and Input Delays
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
Adaptive Kalman Filtering with Exact Linearization and Decoupling Control on Three-Tank Process
von: Widjiantoro, Bambang L., et al.
Veröffentlicht: (2023)
von: Widjiantoro, Bambang L., et al.
Veröffentlicht: (2023)
Policy Gradient Bounds in Multitask LQR
von: Stamouli, Charis, et al.
Veröffentlicht: (2025)
von: Stamouli, Charis, et al.
Veröffentlicht: (2025)
Safe-by-Design: Approximate Nonlinear Model Predictive Control with Real Time Feasibility
von: Olucak, Jan, et al.
Veröffentlicht: (2025)
von: Olucak, Jan, et al.
Veröffentlicht: (2025)
Distributed Adaptive Estimation with ISS Guarantees for Sensor Networks with Partially Unknown Source Dynamics
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
von: Zhao, Feiran, et al.
Veröffentlicht: (2025)
von: Zhao, Feiran, et al.
Veröffentlicht: (2025)
Analyzing the Impact of Computation in Adaptive Dynamic Programming for Stochastic LQR Problem
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
Distributed Nonlinear Control of Networked Two-Wheeled Robots under Adversarial Interactions
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
Adapt and Stabilize, Then Learn and Optimize: A New Approach to Adaptive LQR
von: Fisher, Peter A., et al.
Veröffentlicht: (2025)
von: Fisher, Peter A., et al.
Veröffentlicht: (2025)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
von: Carnevale, Guido, et al.
Veröffentlicht: (2024)
von: Carnevale, Guido, et al.
Veröffentlicht: (2024)
Cooperative $\mathcal{H}_\infty$ Fault-Tolerant Tracking with ISS Guarantees for Networked Systems with Sensor Faults
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026)
Stochastic LQR Design With Disturbance Preview
von: Liu, Jietian, et al.
Veröffentlicht: (2024)
von: Liu, Jietian, et al.
Veröffentlicht: (2024)
The Distributionally Robust Infinite-Horizon LQR
von: Hajar, Joudi, et al.
Veröffentlicht: (2024)
von: Hajar, Joudi, et al.
Veröffentlicht: (2024)
Perturbed Gradient Descent Algorithms are Small-Disturbance Input-to-State Stable
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
Policy Optimization with Differentiable MPC: Convergence Analysis under Uncertainty
von: Zuliani, Riccardo, et al.
Veröffentlicht: (2026)
von: Zuliani, Riccardo, et al.
Veröffentlicht: (2026)
A Bayesian Perspective on the Data-Driven LQR
von: Schwaller, Thierry, et al.
Veröffentlicht: (2026)
von: Schwaller, Thierry, et al.
Veröffentlicht: (2026)
Mixed Regular and Impulsive Sampled-data LQR
von: Daafouz, Jamal, et al.
Veröffentlicht: (2024)
von: Daafouz, Jamal, et al.
Veröffentlicht: (2024)
Distributed Adaptive Time-Varying Optimization with Global Asymptotic Convergence
von: Jiang, Liangze, et al.
Veröffentlicht: (2024)
von: Jiang, Liangze, et al.
Veröffentlicht: (2024)
Linear Convergence of Data-Enabled Policy Optimization for Linear Quadratic Tracking
von: Kang, Shubo, et al.
Veröffentlicht: (2024)
von: Kang, Shubo, et al.
Veröffentlicht: (2024)
A Globally Convergent Policy Gradient Method for Linear Quadratic Gaussian (LQG) Control
von: Sadamoto, Tomonori, et al.
Veröffentlicht: (2023)
von: Sadamoto, Tomonori, et al.
Veröffentlicht: (2023)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
von: Rodriguez-Gil, Jhojan A., et al.
Veröffentlicht: (2026)
von: Rodriguez-Gil, Jhojan A., et al.
Veröffentlicht: (2026)
Beyond Quadratic Costs in LQR: Bregman Divergence Control
von: Hassibi, Babak, et al.
Veröffentlicht: (2025)
von: Hassibi, Babak, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Boundedness of solutions in feedback systems with antithetic controllers
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026) -
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
von: de Oliveira, Arthur Castello Branco, et al.
Veröffentlicht: (2025) -
Convergence Analysis of Gradient Flow for Overparameterized LQR Formulations
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2024) -
When is cumulative dose response monotonic? Analysis of incoherent feedforward motifs
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2026) -
System Identification on the Families of Auto-Regressive with Least-Square-Batch Algorithm
von: Wafi, Moh Kamalul
Veröffentlicht: (2023)