On Globally Optimal Stochastic Policy Gradient Methods for Domain Randomized LQR Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen-Le, Alex, Matni, Nikolai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Policy Gradient for LQR with Domain Randomization
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025)
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025)
A Stochastic Gradient Descent Approach to Design Policy Gradient Methods for LQR
von: Song, Bowen, et al.
Veröffentlicht: (2026)
von: Song, Bowen, et al.
Veröffentlicht: (2026)
Convergence Guarantees of Model-free Policy Gradient Methods for LQR with Stochastic Data
von: Song, Bowen, et al.
Veröffentlicht: (2025)
von: Song, Bowen, et al.
Veröffentlicht: (2025)
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
Domain Randomization is Sample Efficient for Linear Quadratic Control
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025)
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025)
Policy Gradient Bounds in Multitask LQR
von: Stamouli, Charis, et al.
Veröffentlicht: (2025)
von: Stamouli, Charis, et al.
Veröffentlicht: (2025)
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
von: Song, Bowen, et al.
Veröffentlicht: (2025)
von: Song, Bowen, et al.
Veröffentlicht: (2025)
Coordinating Planning and Tracking in Layered Control Policies via Actor-Critic Learning
von: Yang, Fengjun, et al.
Veröffentlicht: (2024)
von: Yang, Fengjun, et al.
Veröffentlicht: (2024)
A Quantitative Framework for Navigating Controller Design Tradeoffs under Computational Constraints
von: Verhoek, Chris, et al.
Veröffentlicht: (2026)
von: Verhoek, Chris, et al.
Veröffentlicht: (2026)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
von: Zhao, Feiran, et al.
Veröffentlicht: (2025)
von: Zhao, Feiran, et al.
Veröffentlicht: (2025)
Single Trajectory Conformal Prediction
von: Lee, Brian, et al.
Veröffentlicht: (2024)
von: Lee, Brian, et al.
Veröffentlicht: (2024)
LQR for Systems with Probabilistic Parametric Uncertainties: A Gradient Method
von: Cui, Leilei, et al.
Veröffentlicht: (2026)
von: Cui, Leilei, et al.
Veröffentlicht: (2026)
Stability-Certified On-Policy Data-Driven LQR via Recursive Learning and Policy Gradient
von: Sforni, Lorenzo, et al.
Veröffentlicht: (2024)
von: Sforni, Lorenzo, et al.
Veröffentlicht: (2024)
Safe Planning in Interactive Environments via Iterative Policy Updates and Adversarially Robust Conformal Prediction
von: Mirzaeedodangeh, Omid, et al.
Veröffentlicht: (2025)
von: Mirzaeedodangeh, Omid, et al.
Veröffentlicht: (2025)
On the (almost) Global Exponential Convergence of the Overparameterized Policy Optimization for the LQR Problem
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
von: Wafi, Moh Kamalul, et al.
Veröffentlicht: (2025)
Stochastic LQR Design With Disturbance Preview
von: Liu, Jietian, et al.
Veröffentlicht: (2024)
von: Liu, Jietian, et al.
Veröffentlicht: (2024)
Learning Flatness-Preserving Residuals for Pure-Feedback Systems
von: Yang, Fengjun, et al.
Veröffentlicht: (2025)
von: Yang, Fengjun, et al.
Veröffentlicht: (2025)
Scalable Distributed Nonlinear Control Under Flatness-Preserving Coupling
von: Yang, Fengjun, et al.
Veröffentlicht: (2025)
von: Yang, Fengjun, et al.
Veröffentlicht: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
von: Zhang, Xiangyuan, et al.
Veröffentlicht: (2023)
von: Zhang, Xiangyuan, et al.
Veröffentlicht: (2023)
Closed-loop Analysis of ADMM-based Suboptimal Linear Model Predictive Control
von: Srikanthan, Anusha, et al.
Veröffentlicht: (2024)
von: Srikanthan, Anusha, et al.
Veröffentlicht: (2024)
Convergence Analysis of Gradient Flow for Overparameterized LQR Formulations
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2024)
von: de Oliveira, Arthur Castello B., et al.
Veröffentlicht: (2024)
Separation is Optimal for LQR under Intermittent Feedback
von: Etcibasi, Abdullah Y., et al.
Veröffentlicht: (2026)
von: Etcibasi, Abdullah Y., et al.
Veröffentlicht: (2026)
Towards a Theory of Control Architecture: A quantitative framework for layered multi-rate control
von: Matni, Nikolai, et al.
Veröffentlicht: (2024)
von: Matni, Nikolai, et al.
Veröffentlicht: (2024)
Analyzing the Impact of Computation in Adaptive Dynamic Programming for Stochastic LQR Problem
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
von: Zhao, Feiran, et al.
Veröffentlicht: (2024)
Nonasymptotic Regret Analysis of Adaptive Linear Quadratic Control with Model Misspecification
von: Lee, Bruce D., et al.
Veröffentlicht: (2023)
von: Lee, Bruce D., et al.
Veröffentlicht: (2023)
State space models, emergence, and ergodicity: How many parameters are needed for stable predictions?
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2024)
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2024)
Learning Complex Motion Plans using Neural ODEs with Safety and Stability Guarantees
von: Nawaz, Farhad, et al.
Veröffentlicht: (2023)
von: Nawaz, Farhad, et al.
Veröffentlicht: (2023)
Distributionally Robust Regret Optimal LQR with Common Stage-Law Ambiguity
von: Fiechtner, Lukas-Benedikt, et al.
Veröffentlicht: (2026)
von: Fiechtner, Lukas-Benedikt, et al.
Veröffentlicht: (2026)
Power-Constrained Policy Gradient Methods for LQR
von: Verma, Ashwin, et al.
Veröffentlicht: (2025)
von: Verma, Ashwin, et al.
Veröffentlicht: (2025)
Small-Disturbance Input-to-State Stability of Perturbed Gradient Flows: Applications to LQR Problem
von: Cui, Leilei, et al.
Veröffentlicht: (2023)
von: Cui, Leilei, et al.
Veröffentlicht: (2023)
Statistical Efficiency of Single- and Multi-step Models for Forecasting and Control
von: Somalwar, Anne, et al.
Veröffentlicht: (2026)
von: Somalwar, Anne, et al.
Veröffentlicht: (2026)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
von: Rodriguez-Gil, Jhojan A., et al.
Veröffentlicht: (2026)
von: Rodriguez-Gil, Jhojan A., et al.
Veröffentlicht: (2026)
A Globally Convergent Policy Gradient Method for Linear Quadratic Gaussian (LQG) Control
von: Sadamoto, Tomonori, et al.
Veröffentlicht: (2023)
von: Sadamoto, Tomonori, et al.
Veröffentlicht: (2023)
Towards Optimal Passive Feedback Control of LTI Systems under LQR Performance
von: Gießler, Armin, et al.
Veröffentlicht: (2026)
von: Gießler, Armin, et al.
Veröffentlicht: (2026)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
von: Carnevale, Guido, et al.
Veröffentlicht: (2024)
von: Carnevale, Guido, et al.
Veröffentlicht: (2024)
Bridging Continuous-time LQR and Reinforcement Learning via Gradient Flow of the Bellman Error
von: Gießler, Armin, et al.
Veröffentlicht: (2025)
von: Gießler, Armin, et al.
Veröffentlicht: (2025)
The Fragility of Learning LQG Controllers
von: Lee, Bruce D., et al.
Veröffentlicht: (2026)
von: Lee, Bruce D., et al.
Veröffentlicht: (2026)
Stochastic Trajectory Influence Functions for LQR: Joint Sensitivity Through Dynamics and Noise Covariance
von: Li, Jiachen, et al.
Veröffentlicht: (2026)
von: Li, Jiachen, et al.
Veröffentlicht: (2026)
Quad-LCD: Layered Control Decomposition Enables Actuator-Feasible Quadrotor Trajectory Planning
von: Srikanthan, Anusha, et al.
Veröffentlicht: (2025)
von: Srikanthan, Anusha, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Policy Gradient for LQR with Domain Randomization
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025) -
A Stochastic Gradient Descent Approach to Design Policy Gradient Methods for LQR
von: Song, Bowen, et al.
Veröffentlicht: (2026) -
Convergence Guarantees of Model-free Policy Gradient Methods for LQR with Stochastic Data
von: Song, Bowen, et al.
Veröffentlicht: (2025) -
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
von: Zhao, Feiran, et al.
Veröffentlicht: (2024) -
Domain Randomization is Sample Efficient for Linear Quadratic Control
von: Fujinami, Tesshu, et al.
Veröffentlicht: (2025)