Regret Bounds for Episodic Risk-Sensitive Linear Quadratic Regulator
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Wenhao, Gao, Xuefeng, He, Xuedong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
von: Ye, Lintao, et al.
Veröffentlicht: (2022)
von: Ye, Lintao, et al.
Veröffentlicht: (2022)
Regret Lower Bounds for Learning Linear Quadratic Gaussian Systems
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2022)
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2022)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
Accelerated Optimization Landscape of Linear-Quadratic Regulator
von: Feng, Lechen, et al.
Veröffentlicht: (2023)
von: Feng, Lechen, et al.
Veröffentlicht: (2023)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
von: Han, Yinbin, et al.
Veröffentlicht: (2023)
von: Han, Yinbin, et al.
Veröffentlicht: (2023)
Tsallis Entropy Regularization for Linearly Solvable MDP and Linear Quadratic Regulator
von: Hashizume, Yota, et al.
Veröffentlicht: (2024)
von: Hashizume, Yota, et al.
Veröffentlicht: (2024)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2025)
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2024)
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2024)
Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
von: Huang, Yilie, et al.
Veröffentlicht: (2024)
von: Huang, Yilie, et al.
Veröffentlicht: (2024)
Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters
von: Li, Deyue
Veröffentlicht: (2023)
von: Li, Deyue
Veröffentlicht: (2023)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
Reinforcement Learning and Regret Bounds for Admission Control
von: Weber, Lucas, et al.
Veröffentlicht: (2024)
von: Weber, Lucas, et al.
Veröffentlicht: (2024)
Small Gradient Norm Regret for Online Convex Optimization
von: Gao, Wenzhi, et al.
Veröffentlicht: (2026)
von: Gao, Wenzhi, et al.
Veröffentlicht: (2026)
Near-Optimal Distributed Linear-Quadratic Regulator for Networked Systems
von: Shin, Sungho, et al.
Veröffentlicht: (2022)
von: Shin, Sungho, et al.
Veröffentlicht: (2022)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
von: Guo, Xin, et al.
Veröffentlicht: (2023)
von: Guo, Xin, et al.
Veröffentlicht: (2023)
Variance-Aware Regret Bounds for Stochastic Contextual Dueling Bandits
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR
von: Toso, Leonardo F., et al.
Veröffentlicht: (2024)
von: Toso, Leonardo F., et al.
Veröffentlicht: (2024)
Beyond $\mathcal{O}(\sqrt{T})$ Regret: Decoupling Learning and Decision-making in Online Linear Programming
von: Gao, Wenzhi, et al.
Veröffentlicht: (2025)
von: Gao, Wenzhi, et al.
Veröffentlicht: (2025)
Regret Bounds for Expected Improvement Algorithms in Gaussian Process Bandit Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
Follow The Approximate Sparse Leader for No-Regret Online Sparse Linear Approximation
von: Mukhopadhyay, Samrat, et al.
Veröffentlicht: (2025)
von: Mukhopadhyay, Samrat, et al.
Veröffentlicht: (2025)
Achieving Better Local Regret Bound for Online Non-Convex Bilevel Optimization
von: Jia, Tingkai, et al.
Veröffentlicht: (2026)
von: Jia, Tingkai, et al.
Veröffentlicht: (2026)
Identification and Adaptive Control of Markov Jump Systems: Sample Complexity and Regret Bounds
von: Sattar, Yahya, et al.
Veröffentlicht: (2021)
von: Sattar, Yahya, et al.
Veröffentlicht: (2021)
Online Optimization on Hadamard Manifolds: Curvature Independent Regret Bounds on Horospherically Convex Objectives
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
Modified Meta-Thompson Sampling for Linear Bandits and Its Bayes Regret Analysis
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
von: Chang, Ting-Jui, et al.
Veröffentlicht: (2024)
von: Chang, Ting-Jui, et al.
Veröffentlicht: (2024)
Improved Regret Bound for Safe Reinforcement Learning via Tighter Cost Pessimism and Reward Optimism
von: Yu, Kihyun, et al.
Veröffentlicht: (2024)
von: Yu, Kihyun, et al.
Veröffentlicht: (2024)
Accelerated Gradient Methods with Biased Gradient Estimates: Risk Sensitivity, High-Probability Guarantees, and Large Deviation Bounds
von: Gürbüzbalaban, Mert, et al.
Veröffentlicht: (2025)
von: Gürbüzbalaban, Mert, et al.
Veröffentlicht: (2025)
Tight Regret Bounds for Bayesian Optimization in One Dimension
von: Scarlett, Jonathan
Veröffentlicht: (2018)
von: Scarlett, Jonathan
Veröffentlicht: (2018)
Learning An Interpretable Risk Scoring System for Maximizing Decision Net Benefit
von: Chi, Wenhao, et al.
Veröffentlicht: (2026)
von: Chi, Wenhao, et al.
Veröffentlicht: (2026)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
von: Gao, Xuefeng, et al.
Veröffentlicht: (2022)
von: Gao, Xuefeng, et al.
Veröffentlicht: (2022)
Two-Timescale Optimization Framework for Sparse-Feedback Linear-Quadratic Optimal Control
von: Feng, Lechen, et al.
Veröffentlicht: (2024)
von: Feng, Lechen, et al.
Veröffentlicht: (2024)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
von: Li, Lucky
Veröffentlicht: (2024)
von: Li, Lucky
Veröffentlicht: (2024)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
von: Carmona, René, et al.
Veröffentlicht: (2019)
von: Carmona, René, et al.
Veröffentlicht: (2019)
Optimal Variance-Dependent Regret Bounds for Infinite-Horizon MDPs
von: Zamir, Guy, et al.
Veröffentlicht: (2026)
von: Zamir, Guy, et al.
Veröffentlicht: (2026)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
von: Tian, Yi, et al.
Veröffentlicht: (2022)
von: Tian, Yi, et al.
Veröffentlicht: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
von: Tian, Yi, et al.
Veröffentlicht: (2026)
von: Tian, Yi, et al.
Veröffentlicht: (2026)
Robust Risk-Sensitive Reinforcement Learning with Conditional Value-at-Risk
von: Ni, Xinyi, et al.
Veröffentlicht: (2024)
von: Ni, Xinyi, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
von: Meng, Huiling, et al.
Veröffentlicht: (2024)
von: Meng, Huiling, et al.
Veröffentlicht: (2024)
Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Penalty Approach
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
A Neural Network Framework for Discovering Closed-form Solutions to Quadratic Programs with Linear Constraints
von: Beylunioglu, Fuat Can, et al.
Veröffentlicht: (2025)
von: Beylunioglu, Fuat Can, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
von: Ye, Lintao, et al.
Veröffentlicht: (2022) -
Regret Lower Bounds for Learning Linear Quadratic Gaussian Systems
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2022) -
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026) -
Accelerated Optimization Landscape of Linear-Quadratic Regulator
von: Feng, Lechen, et al.
Veröffentlicht: (2023) -
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
von: Han, Yinbin, et al.
Veröffentlicht: (2023)