Optimistic Online LQR via Intrinsic Rewards
Fuente:
arXiv
Saved in:
| Main Authors: | Bartos, Marcell, Lee, Bruce D., Treven, Lenart, Krause, Andreas, Dörfler, Florian, Zeilinger, Melanie N. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters
by: Bartos, Marcell, et al.
Published: (2025)
by: Bartos, Marcell, et al.
Published: (2025)
Safe and Near-Optimal Control with Online Dynamics Learning
by: Prajapat, Manish, et al.
Published: (2025)
by: Prajapat, Manish, et al.
Published: (2025)
Stochastic MPC with Online-optimized Policies and Closed-loop Guarantees
by: Bartos, Marcell, et al.
Published: (2025)
by: Bartos, Marcell, et al.
Published: (2025)
Finite-Sample-Based Reachability for Safe Control with Gaussian Process Dynamics
by: Prajapat, Manish, et al.
Published: (2025)
by: Prajapat, Manish, et al.
Published: (2025)
Safe Guaranteed Exploration for Non-linear Systems
by: Prajapat, Manish, et al.
Published: (2024)
by: Prajapat, Manish, et al.
Published: (2024)
Towards safe and tractable Gaussian process-based MPC: Efficient sampling within a sequential quadratic programming framework
by: Prajapat, Manish, et al.
Published: (2024)
by: Prajapat, Manish, et al.
Published: (2024)
A Bayesian Perspective on the Data-Driven LQR
by: Schwaller, Thierry, et al.
Published: (2026)
by: Schwaller, Thierry, et al.
Published: (2026)
A robust and adaptive MPC formulation for Gaussian process models
by: Dubied, Mathieu, et al.
Published: (2025)
by: Dubied, Mathieu, et al.
Published: (2025)
Approximate non-linear model predictive control with safety-augmented neural networks
by: Hose, Henrik, et al.
Published: (2023)
by: Hose, Henrik, et al.
Published: (2023)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
Regularization for Covariance Parameterization of Direct Data-Driven LQR Control
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
Automatic nonlinear MPC approximation with closed-loop guarantees
by: Tokmak, Abdullah, et al.
Published: (2023)
by: Tokmak, Abdullah, et al.
Published: (2023)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
On the Effect of Quadratic Regularization in Direct Data-Driven LQR
by: Klädtke, Manuel, et al.
Published: (2026)
by: Klädtke, Manuel, et al.
Published: (2026)
On Stability in Optimistic Bilevel Optimization
by: Royset, Johannes O.
Published: (2024)
by: Royset, Johannes O.
Published: (2024)
Physics-Informed Graph Neural Network for Dynamic Reconfiguration of Power Systems
by: Authier, Jules, et al.
Published: (2023)
by: Authier, Jules, et al.
Published: (2023)
Decision-Dependent Stochastic Optimization: The Role of Distribution Dynamics
by: He, Zhiyu, et al.
Published: (2025)
by: He, Zhiyu, et al.
Published: (2025)
Almost Surely $\sqrt{T}$ Regret for Adaptive LQR
by: Lu, Yiwen, et al.
Published: (2023)
by: Lu, Yiwen, et al.
Published: (2023)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Towards a Systems Theory of Algorithms
by: Dörfler, Florian, et al.
Published: (2024)
by: Dörfler, Florian, et al.
Published: (2024)
Online Feedback Optimization for Monotone Systems without Timescale Separation
by: Bianchi, Mattia, et al.
Published: (2025)
by: Bianchi, Mattia, et al.
Published: (2025)
Approximate predictive control barrier function for discrete-time systems
by: Didier, Alexandre, et al.
Published: (2024)
by: Didier, Alexandre, et al.
Published: (2024)
Predictive control for nonlinear stochastic systems: Closed-loop guarantees with unbounded noise
by: Köhler, Johannes, et al.
Published: (2024)
by: Köhler, Johannes, et al.
Published: (2024)
A multiobjective approach to robust predictive control barrier functions for discrete-time systems
by: Didier, Alexandre, et al.
Published: (2025)
by: Didier, Alexandre, et al.
Published: (2025)
A model predictive control framework with robust stability guarantees under unbounded disturbances
by: Köhler, Johannes, et al.
Published: (2022)
by: Köhler, Johannes, et al.
Published: (2022)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
A Stability Condition for Online Feedback Optimization without Timescale Separation
by: Bianchi, Mattia, et al.
Published: (2024)
by: Bianchi, Mattia, et al.
Published: (2024)
VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis
by: Leeman, Antoine P., et al.
Published: (2026)
by: Leeman, Antoine P., et al.
Published: (2026)
Network-aware Recommender System via Online Feedback Optimization
by: Chandrasekaran, Sanjay, et al.
Published: (2024)
by: Chandrasekaran, Sanjay, et al.
Published: (2024)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
by: Jing, Gangshan, et al.
Published: (2021)
by: Jing, Gangshan, et al.
Published: (2021)
Optimistic Online Non-stochastic Control via FTRL
by: Mhaisen, Naram, et al.
Published: (2024)
by: Mhaisen, Naram, et al.
Published: (2024)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Online Nonstochastic Prediction: Logarithmic Regret via Predictive Online Least Squares
by: Pai, Chih-Fan, et al.
Published: (2026)
by: Pai, Chih-Fan, et al.
Published: (2026)
Guaranteed Robust Nonlinear MPC via Disturbance Feedback
by: Leeman, Antoine P., et al.
Published: (2025)
by: Leeman, Antoine P., et al.
Published: (2025)
GREAT: Grassmannian REcursive Algorithm for Tracking & Online System Identification
by: Sasfi, András, et al.
Published: (2024)
by: Sasfi, András, et al.
Published: (2024)
Zero-Order Optimization for Gaussian Process-based Model Predictive Control
by: Lahr, Amon, et al.
Published: (2022)
by: Lahr, Amon, et al.
Published: (2022)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Robust reduced-order model predictive control using peak-to-peak analysis of filtered signals
by: Köhler, Johannes, et al.
Published: (2025)
by: Köhler, Johannes, et al.
Published: (2025)
Dynamic Programming in Probability Spaces via Optimal Transport
by: Terpin, Antonio, et al.
Published: (2023)
by: Terpin, Antonio, et al.
Published: (2023)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
Similar Items
-
Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters
by: Bartos, Marcell, et al.
Published: (2025) -
Safe and Near-Optimal Control with Online Dynamics Learning
by: Prajapat, Manish, et al.
Published: (2025) -
Stochastic MPC with Online-optimized Policies and Closed-loop Guarantees
by: Bartos, Marcell, et al.
Published: (2025) -
Finite-Sample-Based Reachability for Safe Control with Gaussian Process Dynamics
by: Prajapat, Manish, et al.
Published: (2025) -
Safe Guaranteed Exploration for Non-linear Systems
by: Prajapat, Manish, et al.
Published: (2024)