Model-Free Output Feedback Stabilization via Policy Gradient Methods
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Ankang, Chi, Ming, Wang, Xiaoling, Ye, Lintao |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Online Learning of Kalman Filtering: From Output to State Estimation
par: Ye, Lintao, et autres
Publié: (2026)
par: Ye, Lintao, et autres
Publié: (2026)
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
par: Ye, Lintao, et autres
Publié: (2022)
par: Ye, Lintao, et autres
Publié: (2022)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
par: Ye, Lintao, et autres
Publié: (2024)
par: Ye, Lintao, et autres
Publié: (2024)
Learning to Sparsify Stochastic Linear Bandits
par: Wang, Zhengmiao, et autres
Publié: (2026)
par: Wang, Zhengmiao, et autres
Publié: (2026)
Learning Stabilizing Policies via an Unstable Subspace Representation
par: Toso, Leonardo F., et autres
Publié: (2025)
par: Toso, Leonardo F., et autres
Publié: (2025)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
par: Ding, Dongsheng, et autres
Publié: (2023)
par: Ding, Dongsheng, et autres
Publié: (2023)
Decentralized Riemannian Conjugate Gradient Method on the Stiefel Manifold
par: Chen, Jun, et autres
Publié: (2023)
par: Chen, Jun, et autres
Publié: (2023)
Online Convex Optimization with Memory and Limited Predictions
par: Wang, Zhengmiao, et autres
Publié: (2024)
par: Wang, Zhengmiao, et autres
Publié: (2024)
Predictor-Based Output-Feedback Control of Linear Systems with Time-Varying Input and Measurement Delays via Neural-Approximated Prediction Horizons
par: Bhan, Luke, et autres
Publié: (2026)
par: Bhan, Luke, et autres
Publié: (2026)
Sample-Free Safety Assessment of Neural Network Controllers via Taylor Methods
par: Evans, Adam, et autres
Publié: (2026)
par: Evans, Adam, et autres
Publié: (2026)
Online Actuator Selection and Controller Design for Linear Quadratic Regulation with Unknown System Model
par: Ye, Lintao, et autres
Publié: (2022)
par: Ye, Lintao, et autres
Publié: (2022)
A Concise Lyapunov Analysis of Nesterov's Accelerated Gradient Method
par: Liu, Jun
Publié: (2025)
par: Liu, Jun
Publié: (2025)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
par: Cao, John, et autres
Publié: (2025)
par: Cao, John, et autres
Publié: (2025)
Lyapunov-stable Neural Control for State and Output Feedback: A Novel Formulation
par: Yang, Lujie, et autres
Publié: (2024)
par: Yang, Lujie, et autres
Publié: (2024)
Robustness of Iteratively Pre-Conditioned Gradient-Descent Method: The Case of Distributed Linear Regression Problem
par: Chakrabarti, Kushal, et autres
Publié: (2021)
par: Chakrabarti, Kushal, et autres
Publié: (2021)
Iterative Pre-Conditioning for Expediting the Gradient-Descent Method: The Distributed Linear Least-Squares Problem
par: Chakrabarti, Kushal, et autres
Publié: (2020)
par: Chakrabarti, Kushal, et autres
Publié: (2020)
Anytime Acceleration of Gradient Descent
par: Zhang, Zihan, et autres
Publié: (2024)
par: Zhang, Zihan, et autres
Publié: (2024)
Adaptive Output Feedback MPC with Guaranteed Stability and Robustness
par: Dey, Anchita, et autres
Publié: (2025)
par: Dey, Anchita, et autres
Publié: (2025)
Gradient Estimation and Variance Reduction in Stochastic and Deterministic Models
par: Keane, Ronan
Publié: (2024)
par: Keane, Ronan
Publié: (2024)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
par: Zhang, Xiangyuan, et autres
Publié: (2023)
par: Zhang, Xiangyuan, et autres
Publié: (2023)
Policy Gradient Method for LQG Control via Input-Output-History Representation: Convergence to $O(ε)$-Stationary Points
par: Sadamoto, Tomonori, et autres
Publié: (2025)
par: Sadamoto, Tomonori, et autres
Publié: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
par: Zhang, Chenyu, et autres
Publié: (2024)
par: Zhang, Chenyu, et autres
Publié: (2024)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
par: Long, Kehan, et autres
Publié: (2025)
par: Long, Kehan, et autres
Publié: (2025)
Modular Distributed Nonconvex Learning with Error Feedback
par: Carnevale, Guido, et autres
Publié: (2025)
par: Carnevale, Guido, et autres
Publié: (2025)
On the Gradient Domination of the LQG Problem
par: Fallah, Kasra, et autres
Publié: (2025)
par: Fallah, Kasra, et autres
Publié: (2025)
A Variance-Reduced Stochastic Gradient Tracking Algorithm for Decentralized Optimization with Orthogonality Constraints
par: Wang, Lei, et autres
Publié: (2022)
par: Wang, Lei, et autres
Publié: (2022)
Gradient-Informed Monte Carlo Fine-Tuning of Diffusion Models for Low-Thrust Trajectory Design
par: Graebner, Jannik, et autres
Publié: (2025)
par: Graebner, Jannik, et autres
Publié: (2025)
Safe Gradient Flow for Bilevel Optimization
par: Sharifi, Sina, et autres
Publié: (2025)
par: Sharifi, Sina, et autres
Publié: (2025)
Gradient Methods with Online Scaling
par: Gao, Wenzhi, et autres
Publié: (2024)
par: Gao, Wenzhi, et autres
Publié: (2024)
Efficient Interaction-Aware Interval Analysis of Neural Network Feedback Loops
par: Jafarpour, Saber, et autres
Publié: (2023)
par: Jafarpour, Saber, et autres
Publié: (2023)
Causal Optimal Coupling for Gaussian Input-Output Distributional Data
par: Xu, Daran, et autres
Publié: (2026)
par: Xu, Daran, et autres
Publié: (2026)
Negative Imaginary Neural ODEs: Learning to Control Mechanical Systems with Stability Guarantees
par: Shi, Kanghong, et autres
Publié: (2025)
par: Shi, Kanghong, et autres
Publié: (2025)
On Stability in Optimistic Bilevel Optimization
par: Royset, Johannes O.
Publié: (2024)
par: Royset, Johannes O.
Publié: (2024)
Distributionally Robust System Level Synthesis With Output Feedback Affine Control Policy
par: Li, Yun, et autres
Publié: (2025)
par: Li, Yun, et autres
Publié: (2025)
Decentralized Optimization on Compact Submanifolds by Quantized Riemannian Gradient Tracking
par: Chen, Jun, et autres
Publié: (2025)
par: Chen, Jun, et autres
Publié: (2025)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
par: Chen, Xiaokai, et autres
Publié: (2024)
par: Chen, Xiaokai, et autres
Publié: (2024)
DoWG Unleashed: An Efficient Universal Parameter-Free Gradient Descent Method
par: Khaled, Ahmed, et autres
Publié: (2023)
par: Khaled, Ahmed, et autres
Publié: (2023)
A Control Theoretic Framework for Adaptive Gradient Optimizers in Machine Learning
par: Chakrabarti, Kushal, et autres
Publié: (2022)
par: Chakrabarti, Kushal, et autres
Publié: (2022)
Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization
par: Zhao, Feiran, et autres
Publié: (2024)
par: Zhao, Feiran, et autres
Publié: (2024)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
par: Ozaslan, Ibrahim K., et autres
Publié: (2024)
par: Ozaslan, Ibrahim K., et autres
Publié: (2024)
Documents similaires
-
Online Learning of Kalman Filtering: From Output to State Estimation
par: Ye, Lintao, et autres
Publié: (2026) -
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
par: Ye, Lintao, et autres
Publié: (2022) -
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
par: Ye, Lintao, et autres
Publié: (2024) -
Learning to Sparsify Stochastic Linear Bandits
par: Wang, Zhengmiao, et autres
Publié: (2026) -
Learning Stabilizing Policies via an Unstable Subspace Representation
par: Toso, Leonardo F., et autres
Publié: (2025)