The Confusing Instance Principle for Online Linear Quadratic Control
Fuente:
arXiv
Saved in:
| Main Authors: | Radji, Waris, Maillard, Odalric-Ambrym |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How Hard is it to Confuse a World Model?
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
by: Kobanda, Anthony, et al.
Published: (2025)
by: Kobanda, Anthony, et al.
Published: (2025)
Leveraging priors on distribution functions for multi-arm bandits
by: Vashishtha, Sumit, et al.
Published: (2025)
by: Vashishtha, Sumit, et al.
Published: (2025)
The regret lower bound for communicating Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
by: Maillard, Odalric-Ambrym, et al.
Published: (2024)
by: Maillard, Odalric-Ambrym, et al.
Published: (2024)
A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks
by: Kobanda, Anthony, et al.
Published: (2025)
by: Kobanda, Anthony, et al.
Published: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
by: Prevost, Adrien, et al.
Published: (2025)
by: Prevost, Adrien, et al.
Published: (2025)
Pliable rejection sampling
by: Erraqabi, Akram, et al.
Published: (2026)
by: Erraqabi, Akram, et al.
Published: (2026)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
by: Kobanda, Anthony, et al.
Published: (2024)
by: Kobanda, Anthony, et al.
Published: (2024)
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
by: Kobanda, Anthony, et al.
Published: (2026)
by: Kobanda, Anthony, et al.
Published: (2026)
Provably Efficient Exploration in Reward Machines with Low Regret
by: Bourel, Hippolyte, et al.
Published: (2024)
by: Bourel, Hippolyte, et al.
Published: (2024)
Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
Interpolation pour l'augmentation de donnees : Application à la gestion des adventices de la canne a sucre a la Reunion
by: Ferber, Frederick Fabre, et al.
Published: (2025)
by: Ferber, Frederick Fabre, et al.
Published: (2025)
AdaStop: adaptive statistical testing for sound comparisons of Deep RL agents
by: Mathieu, Timothée, et al.
Published: (2023)
by: Mathieu, Timothée, et al.
Published: (2023)
Power Mean Estimation in Stochastic Monte-Carlo Tree_Search
by: Dam, Tuan, et al.
Published: (2024)
by: Dam, Tuan, et al.
Published: (2024)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
by: Kohler, Hector, et al.
Published: (2025)
by: Kohler, Hector, et al.
Published: (2025)
Latent Linear Quadratic Regulator for Robotic Control Tasks
by: Zhang, Yuan, et al.
Published: (2024)
by: Zhang, Yuan, et al.
Published: (2024)
Instance-Adaptive Online Multicalibration
by: Huang, Zhiming, et al.
Published: (2026)
by: Huang, Zhiming, et al.
Published: (2026)
Foundations of Safe Online Reinforcement Learning in the Linear Quadratic Regulator: $\sqrt{T}$-Regret
by: Schiffer, Benjamin, et al.
Published: (2025)
by: Schiffer, Benjamin, et al.
Published: (2025)
Foundations of Safe Online Reinforcement Learning in the Linear Quadratic Regulator: Generalized Baselines
by: Schiffer, Benjamin, et al.
Published: (2024)
by: Schiffer, Benjamin, et al.
Published: (2024)
Kriging and Gaussian Process Interpolation for Georeferenced Data Augmentation
by: Ferber, Frédérick Fabre, et al.
Published: (2025)
by: Ferber, Frédérick Fabre, et al.
Published: (2025)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
by: Guo, Xin, et al.
Published: (2023)
by: Guo, Xin, et al.
Published: (2023)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
Nonasymptotic Regret Analysis of Adaptive Linear Quadratic Control with Model Misspecification
by: Lee, Bruce D., et al.
Published: (2023)
by: Lee, Bruce D., et al.
Published: (2023)
Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems
by: Chekan, Jafar Abbaszadeh, et al.
Published: (2024)
by: Chekan, Jafar Abbaszadeh, et al.
Published: (2024)
PHONOS: PHOnetic Neutralization for Online Streaming Applications
by: Quamer, Waris, et al.
Published: (2026)
by: Quamer, Waris, et al.
Published: (2026)
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
by: Lee, Bruce D., et al.
Published: (2024)
by: Lee, Bruce D., et al.
Published: (2024)
Two-Timescale Optimization Framework for Sparse-Feedback Linear-Quadratic Optimal Control
by: Feng, Lechen, et al.
Published: (2024)
by: Feng, Lechen, et al.
Published: (2024)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
by: Li, Lucky
Published: (2024)
by: Li, Lucky
Published: (2024)
Improving the Linearized Laplace Approximation via Quadratic Approximations
by: Jiménez, Pedro, et al.
Published: (2026)
by: Jiménez, Pedro, et al.
Published: (2026)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
by: Hutchinson, Spencer, et al.
Published: (2026)
by: Hutchinson, Spencer, et al.
Published: (2026)
Stacked Confusion Reject Plots (SCORE)
by: Hasler, Stephan, et al.
Published: (2024)
by: Hasler, Stephan, et al.
Published: (2024)
Instance-Dependent Regret Bounds for Nonstochastic Linear Partial Monitoring
by: Di Gennaro, Federico, et al.
Published: (2025)
by: Di Gennaro, Federico, et al.
Published: (2025)
A Noise Sensitivity Exponent Controls Large Statistical-to-Computational Gaps in Single- and Multi-Index Models
by: Defilippis, Leonardo, et al.
Published: (2026)
by: Defilippis, Leonardo, et al.
Published: (2026)
Principled Data Augmentation for Learning to Solve Quadratic Programming Problems
by: Qian, Chendi, et al.
Published: (2025)
by: Qian, Chendi, et al.
Published: (2025)
Online Algorithm for Aggregating Experts' Predictions with Unbounded Quadratic Loss
by: Korotin, Alexander, et al.
Published: (2025)
by: Korotin, Alexander, et al.
Published: (2025)
Blending Complementary Memory Systems in Hybrid Quadratic-Linear Transformers
by: Irie, Kazuki, et al.
Published: (2025)
by: Irie, Kazuki, et al.
Published: (2025)
Generalizing Linear Autoencoder Recommenders with Decoupled Expected Quadratic Loss
by: Guo, Ruixin, et al.
Published: (2026)
by: Guo, Ruixin, et al.
Published: (2026)
Scalar Federated Learning for Linear Quadratic Regulator
by: Rostami, Mohammadreza, et al.
Published: (2026)
by: Rostami, Mohammadreza, et al.
Published: (2026)
Accelerated Optimization Landscape of Linear-Quadratic Regulator
by: Feng, Lechen, et al.
Published: (2023)
by: Feng, Lechen, et al.
Published: (2023)
Similar Items
-
How Hard is it to Confuse a World Model?
by: Radji, Waris, et al.
Published: (2025) -
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
by: Kobanda, Anthony, et al.
Published: (2025) -
Leveraging priors on distribution functions for multi-arm bandits
by: Vashishtha, Sumit, et al.
Published: (2025) -
The regret lower bound for communicating Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025) -
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
by: Maillard, Odalric-Ambrym, et al.
Published: (2024)