Policy Gradient for LQR with Domain Randomization
Fuente:
arXiv
Saved in:
| Main Authors: | Fujinami, Tesshu, Lee, Bruce D., Matni, Nikolai, Pappas, George J. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain Randomization is Sample Efficient for Linear Quadratic Control
by: Fujinami, Tesshu, et al.
Published: (2025)
by: Fujinami, Tesshu, et al.
Published: (2025)
On Globally Optimal Stochastic Policy Gradient Methods for Domain Randomized LQR Synthesis
by: Nguyen-Le, Alex, et al.
Published: (2026)
by: Nguyen-Le, Alex, et al.
Published: (2026)
Learning with Imperfect Models: When Multi-step Prediction Mitigates Compounding Error
by: Somalwar, Anne, et al.
Published: (2025)
by: Somalwar, Anne, et al.
Published: (2025)
Active Learning for Control-Oriented Identification of Nonlinear Systems
by: Lee, Bruce D., et al.
Published: (2024)
by: Lee, Bruce D., et al.
Published: (2024)
State space models, emergence, and ergodicity: How many parameters are needed for stable predictions?
by: Ziemann, Ingvar, et al.
Published: (2024)
by: Ziemann, Ingvar, et al.
Published: (2024)
Guarantees for Nonlinear Representation Learning: Non-identical Covariates, Dependent Data, Fewer Samples
by: Zhang, Thomas T., et al.
Published: (2024)
by: Zhang, Thomas T., et al.
Published: (2024)
A Tutorial on the Non-Asymptotic Theory of System Identification
by: Ziemann, Ingvar, et al.
Published: (2023)
by: Ziemann, Ingvar, et al.
Published: (2023)
Nonasymptotic Regret Analysis of Adaptive Linear Quadratic Control with Model Misspecification
by: Lee, Bruce D., et al.
Published: (2023)
by: Lee, Bruce D., et al.
Published: (2023)
Coordinating Planning and Tracking in Layered Control Policies via Actor-Critic Learning
by: Yang, Fengjun, et al.
Published: (2024)
by: Yang, Fengjun, et al.
Published: (2024)
Single Trajectory Conformal Prediction
by: Lee, Brian, et al.
Published: (2024)
by: Lee, Brian, et al.
Published: (2024)
Policy Gradient Bounds in Multitask LQR
by: Stamouli, Charis, et al.
Published: (2025)
by: Stamouli, Charis, et al.
Published: (2025)
Statistical Efficiency of Single- and Multi-step Models for Forecasting and Control
by: Somalwar, Anne, et al.
Published: (2026)
by: Somalwar, Anne, et al.
Published: (2026)
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
by: Lee, Bruce D., et al.
Published: (2024)
by: Lee, Bruce D., et al.
Published: (2024)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Action Chunking and Exploratory Data Collection Yield Exponential Improvements in Behavior Cloning for Continuous Control
by: Zhang, Thomas T., et al.
Published: (2025)
by: Zhang, Thomas T., et al.
Published: (2025)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Sample-Efficient Linear Representation Learning from Non-IID Non-Isotropic Data
by: Zhang, Thomas T. C. K., et al.
Published: (2023)
by: Zhang, Thomas T. C. K., et al.
Published: (2023)
Why Change Your Controller When You Can Change Your Planner: Drag-Aware Trajectory Generation for Quadrotor Systems
by: Zhang, Hanli, et al.
Published: (2024)
by: Zhang, Hanli, et al.
Published: (2024)
Learning Robust Output Control Barrier Functions from Safe Expert Demonstrations
by: Lindemann, Lars, et al.
Published: (2021)
by: Lindemann, Lars, et al.
Published: (2021)
Conformal Prediction Regions for Time Series using Linear Complementarity Programming
by: Cleaveland, Matthew, et al.
Published: (2023)
by: Cleaveland, Matthew, et al.
Published: (2023)
Adversarial Robustness of Deep State Space Models for Forecasting
by: Anand, Sribalaji C., et al.
Published: (2026)
by: Anand, Sribalaji C., et al.
Published: (2026)
Recursively Feasible Shrinking-Horizon MPC in Dynamic Environments with Conformal Prediction Guarantees
by: Stamouli, Charis, et al.
Published: (2024)
by: Stamouli, Charis, et al.
Published: (2024)
Almost Surely $\sqrt{T}$ Regret for Adaptive LQR
by: Lu, Yiwen, et al.
Published: (2023)
by: Lu, Yiwen, et al.
Published: (2023)
Distributionally Robust Imitation Learning: Layered Control Architecture for Certifiable Autonomy
by: Gahlawat, Aditya, et al.
Published: (2025)
by: Gahlawat, Aditya, et al.
Published: (2025)
eCP: Equivariant Conformal Prediction with pre-trained models
by: Bousias, Nikolaos, et al.
Published: (2026)
by: Bousias, Nikolaos, et al.
Published: (2026)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
On the System Theoretic Offline Learning of Continuous-Time LQR with Exogenous Disturbances
by: Mukherjee, Sayak, et al.
Published: (2025)
by: Mukherjee, Sayak, et al.
Published: (2025)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
by: Jing, Gangshan, et al.
Published: (2021)
by: Jing, Gangshan, et al.
Published: (2021)
Multi-Modal Conformal Prediction Regions with Simple Structures by Optimizing Convex Shape Templates
by: Tumu, Renukanandan, et al.
Published: (2023)
by: Tumu, Renukanandan, et al.
Published: (2023)
Rate-Optimal Non-Asymptotics for the Quadratic Prediction Error Method
by: Stamouli, Charis, et al.
Published: (2024)
by: Stamouli, Charis, et al.
Published: (2024)
Safe Reinforcement Learning-Based Vibration Control: Overcoming Training Risks with LQR Guidance
by: Thorat, Rohan Vitthal, et al.
Published: (2025)
by: Thorat, Rohan Vitthal, et al.
Published: (2025)
Anytime Acceleration of Gradient Descent
by: Zhang, Zihan, et al.
Published: (2024)
by: Zhang, Zihan, et al.
Published: (2024)
Algorithm-Relative Trajectory Valuation in Policy Gradient Control
by: Li, Shihao, et al.
Published: (2025)
by: Li, Shihao, et al.
Published: (2025)
Learning Local Control Barrier Functions for Hybrid Systems
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Flying Quadrotors in Tight Formations using Learning-based Model Predictive Control
by: Chee, Kong Yao, et al.
Published: (2024)
by: Chee, Kong Yao, et al.
Published: (2024)
Solving Reach-Avoid-Stay Problems Using Deep Deterministic Policy Gradients
by: Chenevert, Gabriel, et al.
Published: (2024)
by: Chenevert, Gabriel, et al.
Published: (2024)
A Quantitative Framework for Navigating Controller Design Tradeoffs under Computational Constraints
by: Verhoek, Chris, et al.
Published: (2026)
by: Verhoek, Chris, et al.
Published: (2026)
The Fragility of Learning LQG Controllers
by: Lee, Bruce D., et al.
Published: (2026)
by: Lee, Bruce D., et al.
Published: (2026)
Similar Items
-
Domain Randomization is Sample Efficient for Linear Quadratic Control
by: Fujinami, Tesshu, et al.
Published: (2025) -
On Globally Optimal Stochastic Policy Gradient Methods for Domain Randomized LQR Synthesis
by: Nguyen-Le, Alex, et al.
Published: (2026) -
Learning with Imperfect Models: When Multi-step Prediction Mitigates Compounding Error
by: Somalwar, Anne, et al.
Published: (2025) -
Active Learning for Control-Oriented Identification of Nonlinear Systems
by: Lee, Bruce D., et al.
Published: (2024) -
State space models, emergence, and ergodicity: How many parameters are needed for stable predictions?
by: Ziemann, Ingvar, et al.
Published: (2024)