Learning Stabilizing Policies via an Unstable Subspace Representation
Fuente:
arXiv
Guardado en:
| Autores principales: | Toso, Leonardo F., Ye, Lintao, Anderson, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR
por: Toso, Leonardo F., et al.
Publicado: (2024)
por: Toso, Leonardo F., et al.
Publicado: (2024)
Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning
por: Zhan, Donglin, et al.
Publicado: (2025)
por: Zhan, Donglin, et al.
Publicado: (2025)
On the Gradient Domination of the LQG Problem
por: Fallah, Kasra, et al.
Publicado: (2025)
por: Fallah, Kasra, et al.
Publicado: (2025)
Adversarially Robust Multitask Adaptive Control
por: Fallah, Kasra, et al.
Publicado: (2025)
por: Fallah, Kasra, et al.
Publicado: (2025)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026)
por: Zhang, Ankang, et al.
Publicado: (2026)
Learning Invariant Visual Representations for Planning with Joint-Embedding Predictive World Models
por: Toso, Leonardo F., et al.
Publicado: (2026)
por: Toso, Leonardo F., et al.
Publicado: (2026)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
por: Ye, Lintao, et al.
Publicado: (2024)
por: Ye, Lintao, et al.
Publicado: (2024)
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
por: Ye, Lintao, et al.
Publicado: (2022)
por: Ye, Lintao, et al.
Publicado: (2022)
Online Learning of Kalman Filtering: From Output to State Estimation
por: Ye, Lintao, et al.
Publicado: (2026)
por: Ye, Lintao, et al.
Publicado: (2026)
Non-Asymptotic Bounds for Closed-Loop Identification of Unstable Nonlinear Stochastic Systems
por: Siriya, Seth, et al.
Publicado: (2024)
por: Siriya, Seth, et al.
Publicado: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Learning to Sparsify Stochastic Linear Bandits
por: Wang, Zhengmiao, et al.
Publicado: (2026)
por: Wang, Zhengmiao, et al.
Publicado: (2026)
Asynchronous Heterogeneous Linear Quadratic Regulator Design
por: Toso, Leonardo F., et al.
Publicado: (2024)
por: Toso, Leonardo F., et al.
Publicado: (2024)
Policy Gradient Bounds in Multitask LQR
por: Stamouli, Charis, et al.
Publicado: (2025)
por: Stamouli, Charis, et al.
Publicado: (2025)
Data-Efficient and Robust Task Selection for Meta-Learning
por: Zhan, Donglin, et al.
Publicado: (2024)
por: Zhan, Donglin, et al.
Publicado: (2024)
Subspace Optimization for Efficient Federated Learning under Heterogeneous Data
por: Zhu, Shuchen, et al.
Publicado: (2026)
por: Zhu, Shuchen, et al.
Publicado: (2026)
Online Convex Optimization with Memory and Limited Predictions
por: Wang, Zhengmiao, et al.
Publicado: (2024)
por: Wang, Zhengmiao, et al.
Publicado: (2024)
Stochastic Subspace Descent Accelerated via Bi-fidelity Line Search
por: Cheng, Nuojin, et al.
Publicado: (2025)
por: Cheng, Nuojin, et al.
Publicado: (2025)
Masked Subspace Clustering Methods
por: Song, Jiebo, et al.
Publicado: (2025)
por: Song, Jiebo, et al.
Publicado: (2025)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
por: Long, Kehan, et al.
Publicado: (2025)
por: Long, Kehan, et al.
Publicado: (2025)
Sample-Efficient Linear Representation Learning from Non-IID Non-Isotropic Data
por: Zhang, Thomas T. C. K., et al.
Publicado: (2023)
por: Zhang, Thomas T. C. K., et al.
Publicado: (2023)
Collaborative Bayesian Optimization via Wasserstein Barycenters
por: Zhan, Donglin, et al.
Publicado: (2025)
por: Zhan, Donglin, et al.
Publicado: (2025)
Subspace Optimization for Large Language Models with Convergence Guarantees
por: He, Yutong, et al.
Publicado: (2024)
por: He, Yutong, et al.
Publicado: (2024)
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
por: Lee, Bruce D., et al.
Publicado: (2024)
por: Lee, Bruce D., et al.
Publicado: (2024)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
por: Cao, John, et al.
Publicado: (2025)
por: Cao, John, et al.
Publicado: (2025)
Global Convergence of Iteratively Reweighted Least Squares for Robust Subspace Recovery
por: Lerman, Gilad, et al.
Publicado: (2025)
por: Lerman, Gilad, et al.
Publicado: (2025)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
por: Wang, Han, et al.
Publicado: (2023)
por: Wang, Han, et al.
Publicado: (2023)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
por: Shen, Kaichen, et al.
Publicado: (2026)
por: Shen, Kaichen, et al.
Publicado: (2026)
Stability Evaluation via Distributional Perturbation Analysis
por: Blanchet, Jose, et al.
Publicado: (2024)
por: Blanchet, Jose, et al.
Publicado: (2024)
Mathematical Foundations of Deep Learning
por: Ye, Xiaojing
Publicado: (2026)
por: Ye, Xiaojing
Publicado: (2026)
A Minibatch-SGD-Based Learning Meta-Policy for Inventory Systems with Myopic Optimal Policy
por: Lyu, Jiameng, et al.
Publicado: (2024)
por: Lyu, Jiameng, et al.
Publicado: (2024)
Krylov Cubic Regularized Newton: A Subspace Second-Order Method with Dimension-Free Convergence Rate
por: Jiang, Ruichen, et al.
Publicado: (2024)
por: Jiang, Ruichen, et al.
Publicado: (2024)
Achieve Performatively Optimal Policy for Performative Reinforcement Learning
por: Chen, Ziyi, et al.
Publicado: (2025)
por: Chen, Ziyi, et al.
Publicado: (2025)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
por: Lin, Junan, et al.
Publicado: (2026)
por: Lin, Junan, et al.
Publicado: (2026)
Deceptive Sequential Decision-Making via Regularized Policy Optimization
por: Kim, Yerin, et al.
Publicado: (2025)
por: Kim, Yerin, et al.
Publicado: (2025)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
por: Schmidt, Carolin, et al.
Publicado: (2024)
por: Schmidt, Carolin, et al.
Publicado: (2024)
On Building Myopic MPC Policies using Supervised Learning
por: Orrico, Christopher A., et al.
Publicado: (2024)
por: Orrico, Christopher A., et al.
Publicado: (2024)
Learning Aligned Stability in Neural ODEs Reconciling Accuracy with Robustness
por: Luo, Chaoyang, et al.
Publicado: (2025)
por: Luo, Chaoyang, et al.
Publicado: (2025)
Why Smooth Stability Assumptions Fail for ReLU Learning
por: Katende, Ronald
Publicado: (2025)
por: Katende, Ronald
Publicado: (2025)
Generalization Guarantees for Learning Branch-and-Cut Policies in Integer Programming
por: Cheng, Hongyu, et al.
Publicado: (2025)
por: Cheng, Hongyu, et al.
Publicado: (2025)
Ejemplares similares
-
Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR
por: Toso, Leonardo F., et al.
Publicado: (2024) -
Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning
por: Zhan, Donglin, et al.
Publicado: (2025) -
On the Gradient Domination of the LQG Problem
por: Fallah, Kasra, et al.
Publicado: (2025) -
Adversarially Robust Multitask Adaptive Control
por: Fallah, Kasra, et al.
Publicado: (2025) -
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026)