Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Kanakeri, Vinay, Bajaj, Shivam, Verma, Ashwin, Gupta, Vijay, Mitra, Aritra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2024)
by: Kanakeri, Vinay, et al.
Published: (2024)
Power-Constrained Policy Gradient Methods for LQR
by: Verma, Ashwin, et al.
Published: (2025)
by: Verma, Ashwin, et al.
Published: (2025)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
by: Ye, Lintao, et al.
Published: (2024)
by: Ye, Lintao, et al.
Published: (2024)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
by: Mitra, Aritra
Published: (2024)
by: Mitra, Aritra
Published: (2024)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
by: Zhu, Feng, et al.
Published: (2024)
by: Zhu, Feng, et al.
Published: (2024)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
by: Zhu, Feng, et al.
Published: (2026)
by: Zhu, Feng, et al.
Published: (2026)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
Almost Surely $\sqrt{T}$ Regret for Adaptive LQR
by: Lu, Yiwen, et al.
Published: (2023)
by: Lu, Yiwen, et al.
Published: (2023)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
by: Zhu, Feng, et al.
Published: (2025)
by: Zhu, Feng, et al.
Published: (2025)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
by: Ye, Lintao, et al.
Published: (2022)
by: Ye, Lintao, et al.
Published: (2022)
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
On the (almost) Global Exponential Convergence of the Overparameterized Policy Optimization for the LQR Problem
by: Wafi, Moh Kamalul, et al.
Published: (2025)
by: Wafi, Moh Kamalul, et al.
Published: (2025)
Policy Gradient Bounds in Multitask LQR
by: Stamouli, Charis, et al.
Published: (2025)
by: Stamouli, Charis, et al.
Published: (2025)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
by: Zhang, Xiaole, et al.
Published: (2025)
by: Zhang, Xiaole, et al.
Published: (2025)
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
by: Chang, Ting-Jui, et al.
Published: (2024)
by: Chang, Ting-Jui, et al.
Published: (2024)
On Policy Stochasticity in Mutual Information Optimal Control of Linear Systems
by: Enami, Shoju, et al.
Published: (2025)
by: Enami, Shoju, et al.
Published: (2025)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
by: Jing, Gangshan, et al.
Published: (2021)
by: Jing, Gangshan, et al.
Published: (2021)
LQR based $ω-$stabilization of a heat equation with memory
by: Sistla, Bhargav Pavan Kumar, et al.
Published: (2025)
by: Sistla, Bhargav Pavan Kumar, et al.
Published: (2025)
A Moreau Envelope Approach for LQR Meta-Policy Estimation
by: Aravind, Ashwin, et al.
Published: (2024)
by: Aravind, Ashwin, et al.
Published: (2024)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
by: Cao, John, et al.
Published: (2025)
by: Cao, John, et al.
Published: (2025)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
by: Carnevale, Guido, et al.
Published: (2024)
by: Carnevale, Guido, et al.
Published: (2024)
Hybrid Energy-Aware Reward Shaping: A Unified Lightweight Physics-Guided Methodology for Policy Optimization
by: Liao, Qijun, et al.
Published: (2026)
by: Liao, Qijun, et al.
Published: (2026)
Leveraging Offline Data from Similar Systems for Online Linear Quadratic Control
by: Bajaj, Shivam, et al.
Published: (2025)
by: Bajaj, Shivam, et al.
Published: (2025)
On Constraints in First-Order Optimization: A View from Non-Smooth Dynamical Systems
by: Muehlebach, Michael, et al.
Published: (2021)
by: Muehlebach, Michael, et al.
Published: (2021)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
A Bayesian Perspective on the Data-Driven LQR
by: Schwaller, Thierry, et al.
Published: (2026)
by: Schwaller, Thierry, et al.
Published: (2026)
FORESEE: Prediction with Expansion-Compression Unscented Transform for Online Policy Optimization
by: Parwana, Hardik, et al.
Published: (2022)
by: Parwana, Hardik, et al.
Published: (2022)
Data-Driven Adversarial Online Control for Unknown Linear Systems
by: Liu, Zishun, et al.
Published: (2023)
by: Liu, Zishun, et al.
Published: (2023)
Interpolation Conditions for Data Consistency and Prediction in Noisy Linear Systems
by: Vanelli, Martina, et al.
Published: (2025)
by: Vanelli, Martina, et al.
Published: (2025)
Learning of Linear Dynamical Systems as a Non-Commutative Polynomial Optimization Problem
by: Zhou, Quan, et al.
Published: (2020)
by: Zhou, Quan, et al.
Published: (2020)
Symplectic Inductive Bias for Data-Driven Target Reachability in Hamiltonian Systems
by: Ouyang, Zhuo, et al.
Published: (2026)
by: Ouyang, Zhuo, et al.
Published: (2026)
Similar Items
-
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2025) -
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2024) -
Power-Constrained Policy Gradient Methods for LQR
by: Verma, Ashwin, et al.
Published: (2025) -
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
by: Ye, Lintao, et al.
Published: (2024) -
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
by: Mitra, Aritra
Published: (2024)