Robustness of Online Identification-based Policy Iteration to Noisy Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Song, Bowen, Iannelli, Andrea |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Role of Identification in Data-driven Policy Iteration: A System Theoretic Study
por: Song, Bowen, et al.
Publicado: (2024)
por: Song, Bowen, et al.
Publicado: (2024)
Convergence and Robustness of Value and Policy Iteration for the Linear Quadratic Regulator
por: Song, Bowen, et al.
Publicado: (2024)
por: Song, Bowen, et al.
Publicado: (2024)
Convergence Guarantees of Model-free Policy Gradient Methods for LQR with Stochastic Data
por: Song, Bowen, et al.
Publicado: (2025)
por: Song, Bowen, et al.
Publicado: (2025)
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
por: Song, Bowen, et al.
Publicado: (2025)
por: Song, Bowen, et al.
Publicado: (2025)
Data-Driven Stabilization of Continuous-Time LTI Systems from Noisy Input-Output Data
por: Bosso, Alessandro, et al.
Publicado: (2025)
por: Bosso, Alessandro, et al.
Publicado: (2025)
Hidden Convexity in Active Learning: A Convexified Online Input Design for ARX Systems
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2025)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2025)
A Stochastic Gradient Descent Approach to Design Policy Gradient Methods for LQR
por: Song, Bowen, et al.
Publicado: (2026)
por: Song, Bowen, et al.
Publicado: (2026)
Online Convex Optimization and Integral Quadratic Constraints: An automated approach to regret analysis
por: Jakob, Fabian, et al.
Publicado: (2025)
por: Jakob, Fabian, et al.
Publicado: (2025)
Adaptive control mechanisms in gradient descent algorithms
por: Iannelli, Andrea
Publicado: (2025)
por: Iannelli, Andrea
Publicado: (2025)
Sample Complexity Bounds for Linear System Identification from a Finite Set
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
A hybrid systems framework for data-based adaptive control of linear time-varying systems
por: Iannelli, Andrea, et al.
Publicado: (2024)
por: Iannelli, Andrea, et al.
Publicado: (2024)
Closed-Loop Finite-Time Analysis of Suboptimal Online Control
por: Karapetyan, Aren, et al.
Publicado: (2023)
por: Karapetyan, Aren, et al.
Publicado: (2023)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
por: Manenti, Massimiliano, et al.
Publicado: (2025)
por: Manenti, Massimiliano, et al.
Publicado: (2025)
A Unified Bayesian Framework for Data-Driven Smoothing, Prediction, and Control
por: Yin, Mingzhou, et al.
Publicado: (2025)
por: Yin, Mingzhou, et al.
Publicado: (2025)
The role of identification in data‐driven policy iteration: A system theoretic study
por: Bowen Song, et al.
Publicado: (2024)
por: Bowen Song, et al.
Publicado: (2024)
Derivative-Free Data-Driven Control of Continuous-Time Linear Time-Invariant Systems
por: Bosso, Alessandro, et al.
Publicado: (2024)
por: Bosso, Alessandro, et al.
Publicado: (2024)
Analysis and Synthesis of Switched Optimization Algorithms
por: Miller, Jared, et al.
Publicado: (2025)
por: Miller, Jared, et al.
Publicado: (2025)
A polynomial-based QCQP solver for encrypted optimization
por: Schlor, Sebastian, et al.
Publicado: (2025)
por: Schlor, Sebastian, et al.
Publicado: (2025)
Data-Driven Control of Continuous-Time LTI Systems via Non-Minimal Realizations
por: Bosso, Alessandro, et al.
Publicado: (2025)
por: Bosso, Alessandro, et al.
Publicado: (2025)
A harmonic framework for the identification of linear time-periodic systems
por: Vernerey, Flora, et al.
Publicado: (2023)
por: Vernerey, Flora, et al.
Publicado: (2023)
Beyond Bounded Noise: Stochastic Set-Membership Estimation for Nonlinear Systems
por: Brändle, Felix, et al.
Publicado: (2026)
por: Brändle, Felix, et al.
Publicado: (2026)
High Effort, Low Gain: Fundamental Limits of Active Learning for Linear Dynamical Systems
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2025)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2025)
Data-Driven Synthesis of Robust Positively Invariant Sets from Noisy Data
por: Wang, Chi, et al.
Publicado: (2026)
por: Wang, Chi, et al.
Publicado: (2026)
End-to-end guarantees for indirect data-driven control of bilinear systems with finite stochastic data
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
Accelerated ADMM: Automated Parameter Tuning and Improved Linear Convergence
por: Tavakoli, Meisam, et al.
Publicado: (2025)
por: Tavakoli, Meisam, et al.
Publicado: (2025)
A distributed framework for linear adaptive MPC
por: Parsi, Anilkumar, et al.
Publicado: (2021)
por: Parsi, Anilkumar, et al.
Publicado: (2021)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Controller Synthesis from Noisy-Input Noisy-Output Data
por: Li, Lidong, et al.
Publicado: (2024)
por: Li, Lidong, et al.
Publicado: (2024)
Learning Soft Constrained MPC Value Functions: Efficient MPC Design and Implementation providing Stability and Safety Guarantees
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
On-Line Policy Iteration with Trajectory-Driven Policy Generation
por: Li, Yuchao, et al.
Publicado: (2026)
por: Li, Yuchao, et al.
Publicado: (2026)
Physics-Informed Neural Network Policy Iteration: Algorithms, Convergence, and Verification
por: Meng, Yiming, et al.
Publicado: (2024)
por: Meng, Yiming, et al.
Publicado: (2024)
Adaptive Time-Domain Harmonic Control for Noise-Vibration-Harshness Reduction of Electric Drives
por: Herburger, Klaus, et al.
Publicado: (2025)
por: Herburger, Klaus, et al.
Publicado: (2025)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
por: Carnevale, Guido, et al.
Publicado: (2024)
por: Carnevale, Guido, et al.
Publicado: (2024)
A quantitative and constructive proof of Willems' Fundamental Lemma and its implications
por: Berberich, Julian, et al.
Publicado: (2022)
por: Berberich, Julian, et al.
Publicado: (2022)
On the Regret of Recursive Methods for Discrete-Time Adaptive Control with Matched Uncertainty
por: Karapetyan, Aren, et al.
Publicado: (2024)
por: Karapetyan, Aren, et al.
Publicado: (2024)
Data-Enabled Policy and Value Iteration for Continuous-Time Linear Quadratic Output Feedback Control
por: Xie, Jun, et al.
Publicado: (2026)
por: Xie, Jun, et al.
Publicado: (2026)
Dual-Loop Robust Control of Biased Koopman Operator Model by Noisy Data of Nonlinear Systems
por: He, Tianyi, et al.
Publicado: (2024)
por: He, Tianyi, et al.
Publicado: (2024)
Learning Robust Regions of Attraction Using Rollout-Enhanced Physics-Informed Neural Networks with Policy Iteration
por: Wang, Junkai, et al.
Publicado: (2025)
por: Wang, Junkai, et al.
Publicado: (2025)
Online Optimization with Unknown Time-Varying Parameters from Noisy Gradient Measurements
por: Tripathi, Shivanshu, et al.
Publicado: (2026)
por: Tripathi, Shivanshu, et al.
Publicado: (2026)
A Linear Parameter-Varying Framework for the Analysis of Time-Varying Optimization Algorithms
por: Jakob, Fabian, et al.
Publicado: (2025)
por: Jakob, Fabian, et al.
Publicado: (2025)
Ejemplares similares
-
The Role of Identification in Data-driven Policy Iteration: A System Theoretic Study
por: Song, Bowen, et al.
Publicado: (2024) -
Convergence and Robustness of Value and Policy Iteration for the Linear Quadratic Regulator
por: Song, Bowen, et al.
Publicado: (2024) -
Convergence Guarantees of Model-free Policy Gradient Methods for LQR with Stochastic Data
por: Song, Bowen, et al.
Publicado: (2025) -
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
por: Song, Bowen, et al.
Publicado: (2025) -
Data-Driven Stabilization of Continuous-Time LTI Systems from Noisy Input-Output Data
por: Bosso, Alessandro, et al.
Publicado: (2025)