Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
Fuente:
arXiv
Guardado en:
| Autores principales: | Ye, Lintao, Mitra, Aritra, Gupta, Vijay |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
por: Ye, Lintao, et al.
Publicado: (2022)
por: Ye, Lintao, et al.
Publicado: (2022)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
por: Maity, Sreejeet, et al.
Publicado: (2025)
por: Maity, Sreejeet, et al.
Publicado: (2025)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
por: Mitra, Aritra
Publicado: (2024)
por: Mitra, Aritra
Publicado: (2024)
Online Actuator Selection and Controller Design for Linear Quadratic Regulation with Unknown System Model
por: Ye, Lintao, et al.
Publicado: (2022)
por: Ye, Lintao, et al.
Publicado: (2022)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
por: Maity, Sreejeet, et al.
Publicado: (2025)
por: Maity, Sreejeet, et al.
Publicado: (2025)
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
por: Kanakeri, Vinay, et al.
Publicado: (2024)
por: Kanakeri, Vinay, et al.
Publicado: (2024)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
por: Kanakeri, Vinay, et al.
Publicado: (2025)
por: Kanakeri, Vinay, et al.
Publicado: (2025)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
por: Zhu, Feng, et al.
Publicado: (2024)
por: Zhu, Feng, et al.
Publicado: (2024)
Robust Q-Learning under Corrupted Rewards
por: Maity, Sreejeet, et al.
Publicado: (2024)
por: Maity, Sreejeet, et al.
Publicado: (2024)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026)
por: Zhang, Ankang, et al.
Publicado: (2026)
Learning to Sparsify Stochastic Linear Bandits
por: Wang, Zhengmiao, et al.
Publicado: (2026)
por: Wang, Zhengmiao, et al.
Publicado: (2026)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2024)
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2024)
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
por: Kanakeri, Vinay, et al.
Publicado: (2025)
por: Kanakeri, Vinay, et al.
Publicado: (2025)
Tsallis Entropy Regularization for Linearly Solvable MDP and Linear Quadratic Regulator
por: Hashizume, Yota, et al.
Publicado: (2024)
por: Hashizume, Yota, et al.
Publicado: (2024)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2025)
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
por: Zhu, Feng, et al.
Publicado: (2025)
por: Zhu, Feng, et al.
Publicado: (2025)
Near-Optimal Distributed Linear-Quadratic Regulator for Networked Systems
por: Shin, Sungho, et al.
Publicado: (2022)
por: Shin, Sungho, et al.
Publicado: (2022)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
por: Zhu, Feng, et al.
Publicado: (2026)
por: Zhu, Feng, et al.
Publicado: (2026)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
por: Mitra, Aritra, et al.
Publicado: (2023)
por: Mitra, Aritra, et al.
Publicado: (2023)
Power-Constrained Policy Gradient Methods for LQR
por: Verma, Ashwin, et al.
Publicado: (2025)
por: Verma, Ashwin, et al.
Publicado: (2025)
Online Learning of Kalman Filtering: From Output to State Estimation
por: Ye, Lintao, et al.
Publicado: (2026)
por: Ye, Lintao, et al.
Publicado: (2026)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
por: Tian, Yi, et al.
Publicado: (2022)
por: Tian, Yi, et al.
Publicado: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
por: Tian, Yi, et al.
Publicado: (2026)
por: Tian, Yi, et al.
Publicado: (2026)
Fixed Horizon Linear Quadratic Covariance Steering in Continuous Time with Hilbert-Schmidt Terminal Cost
por: Sial, Tushar, et al.
Publicado: (2025)
por: Sial, Tushar, et al.
Publicado: (2025)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
por: Hutchinson, Spencer, et al.
Publicado: (2026)
por: Hutchinson, Spencer, et al.
Publicado: (2026)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
por: Zhang, Xiaole, et al.
Publicado: (2025)
por: Zhang, Xiaole, et al.
Publicado: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
por: Adibi, Arman, et al.
Publicado: (2024)
por: Adibi, Arman, et al.
Publicado: (2024)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
por: Srikant, R.
Publicado: (2024)
por: Srikant, R.
Publicado: (2024)
Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2024)
por: Huang, Yilie, et al.
Publicado: (2024)
Expressivity of Quadratic Neural ODEs
por: Hanson, Joshua, et al.
Publicado: (2025)
por: Hanson, Joshua, et al.
Publicado: (2025)
Robustness of Iteratively Pre-Conditioned Gradient-Descent Method: The Case of Distributed Linear Regression Problem
por: Chakrabarti, Kushal, et al.
Publicado: (2021)
por: Chakrabarti, Kushal, et al.
Publicado: (2021)
Iterative Pre-Conditioning for Expediting the Gradient-Descent Method: The Distributed Linear Least-Squares Problem
por: Chakrabarti, Kushal, et al.
Publicado: (2020)
por: Chakrabarti, Kushal, et al.
Publicado: (2020)
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
por: Chang, Ting-Jui, et al.
Publicado: (2024)
por: Chang, Ting-Jui, et al.
Publicado: (2024)
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2025)
por: Huang, Yilie, et al.
Publicado: (2025)
Learning Linear Dynamics from Bilinear Observations
por: Sattar, Yahya, et al.
Publicado: (2024)
por: Sattar, Yahya, et al.
Publicado: (2024)
A PAC-Bayes Approach for Controlling Unknown Linear Discrete-time Systems
por: Luo, Yujia, et al.
Publicado: (2026)
por: Luo, Yujia, et al.
Publicado: (2026)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
por: Wang, Han, et al.
Publicado: (2023)
por: Wang, Han, et al.
Publicado: (2023)
A Complete Set of Quadratic Constraints for Repeated ReLU and Generalizations
por: Noori, Sahel Vahedi, et al.
Publicado: (2024)
por: Noori, Sahel Vahedi, et al.
Publicado: (2024)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
por: Sattar, Yahya, et al.
Publicado: (2025)
por: Sattar, Yahya, et al.
Publicado: (2025)
Ejemplares similares
-
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
por: Ye, Lintao, et al.
Publicado: (2022) -
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
por: Maity, Sreejeet, et al.
Publicado: (2025) -
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
por: Mitra, Aritra
Publicado: (2024) -
Online Actuator Selection and Controller Design for Linear Quadratic Regulation with Unknown System Model
por: Ye, Lintao, et al.
Publicado: (2022) -
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
por: Maity, Sreejeet, et al.
Publicado: (2025)