Tight Rates for Bandit Control Beyond Quadratics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Y. Jennifer, Lu, Zhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Population Dynamics Control with Partial Observations
von: Lu, Zhou, et al.
Veröffentlicht: (2025)
von: Lu, Zhou, et al.
Veröffentlicht: (2025)
Online Control in Population Dynamics
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
Bandit Allocational Instability
von: Chen, Yilun, et al.
Veröffentlicht: (2026)
von: Chen, Yilun, et al.
Veröffentlicht: (2026)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
von: Guo, Xin, et al.
Veröffentlicht: (2023)
von: Guo, Xin, et al.
Veröffentlicht: (2023)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
Bandit Convex Optimisation
von: Lattimore, Tor
Veröffentlicht: (2024)
von: Lattimore, Tor
Veröffentlicht: (2024)
Model Predictive Control is Almost Optimal for Restless Bandit
von: Gast, Nicolas, et al.
Veröffentlicht: (2024)
von: Gast, Nicolas, et al.
Veröffentlicht: (2024)
Two-Timescale Optimization Framework for Sparse-Feedback Linear-Quadratic Optimal Control
von: Feng, Lechen, et al.
Veröffentlicht: (2024)
von: Feng, Lechen, et al.
Veröffentlicht: (2024)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
von: Li, Lucky
Veröffentlicht: (2024)
von: Li, Lucky
Veröffentlicht: (2024)
Neural Network-Based Bandit: A Medium Access Control for the IIoT Alarm Scenario
von: Raghuwanshi, Prasoon, et al.
Veröffentlicht: (2024)
von: Raghuwanshi, Prasoon, et al.
Veröffentlicht: (2024)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
von: Tian, Yi, et al.
Veröffentlicht: (2022)
von: Tian, Yi, et al.
Veröffentlicht: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
von: Tian, Yi, et al.
Veröffentlicht: (2026)
von: Tian, Yi, et al.
Veröffentlicht: (2026)
Tight Generalization Bounds for Noiseless Inverse Optimization
von: Fatemi, Pouria, et al.
Veröffentlicht: (2026)
von: Fatemi, Pouria, et al.
Veröffentlicht: (2026)
Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Penalty Approach
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
Tight Bounds for Online Convex Optimization with Adversarial Constraints
von: Sinha, Abhishek, et al.
Veröffentlicht: (2024)
von: Sinha, Abhishek, et al.
Veröffentlicht: (2024)
Tight Mixed-Integer Optimization Formulations for Prescriptive Trees
von: Biggs, Max, et al.
Veröffentlicht: (2023)
von: Biggs, Max, et al.
Veröffentlicht: (2023)
Contextual Bandits with Budgeted Information Reveal
von: Gan, Kyra, et al.
Veröffentlicht: (2023)
von: Gan, Kyra, et al.
Veröffentlicht: (2023)
Decentralized Contextual Bandits with Network Adaptivity
von: Deng, Chuyun, et al.
Veröffentlicht: (2025)
von: Deng, Chuyun, et al.
Veröffentlicht: (2025)
The Safety-Privacy Tradeoff in Linear Bandits
von: Zibaie, Arghavan, et al.
Veröffentlicht: (2025)
von: Zibaie, Arghavan, et al.
Veröffentlicht: (2025)
Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Non-Penalty Approach
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
von: Feng, Lechen, et al.
Veröffentlicht: (2025)
Expressivity of Quadratic Neural ODEs
von: Hanson, Joshua, et al.
Veröffentlicht: (2025)
von: Hanson, Joshua, et al.
Veröffentlicht: (2025)
An Efficient Unsupervised Framework for Convex Quadratic Programs via Deep Unrolling
von: Yang, Linxin, et al.
Veröffentlicht: (2024)
von: Yang, Linxin, et al.
Veröffentlicht: (2024)
Logarithmic Regret for Unconstrained Submodular Maximization Stochastic Bandit
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
Learning to Sparsify Stochastic Linear Bandits
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
A Tight Theory of Error Feedback Algorithms in Distributed Optimization
von: Thomsen, Daniel Berg, et al.
Veröffentlicht: (2026)
von: Thomsen, Daniel Berg, et al.
Veröffentlicht: (2026)
Representative Action Selection for Large Action Space: From Bandits to MDPs
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
Representative Action Selection for Large Action Space Bandit Families
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
von: Zhou, Quan, et al.
Veröffentlicht: (2025)
Limits of Convergence-Rate Control for Open-Weight Safety
von: Rosati, Domenic, et al.
Veröffentlicht: (2026)
von: Rosati, Domenic, et al.
Veröffentlicht: (2026)
Insights on Muon from Simple Quadratics
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
Online Newton Method for Bandit Convex Optimisation
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
Sharper Guarantees for Misspecified Kernelized Bandit Optimization
von: Maran, Davide, et al.
Veröffentlicht: (2026)
von: Maran, Davide, et al.
Veröffentlicht: (2026)
Combinatorial Causal Bandits without Graph Skeleton
von: Feng, Shi, et al.
Veröffentlicht: (2023)
von: Feng, Shi, et al.
Veröffentlicht: (2023)
Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
A Sequential Quadratic Programming Method with High Probability Complexity Bounds for Nonlinear Equality Constrained Stochastic Optimization
von: Berahas, Albert S., et al.
Veröffentlicht: (2023)
von: Berahas, Albert S., et al.
Veröffentlicht: (2023)
Tight Convergence Rate Bounds for Optimization Under Power Law Spectral Conditions
von: Velikanov, Maksim, et al.
Veröffentlicht: (2022)
von: Velikanov, Maksim, et al.
Veröffentlicht: (2022)
Tight Robustness Certificates and Wasserstein Distributional Attacks for Deep Neural Networks
von: Le, Bach C., et al.
Veröffentlicht: (2025)
von: Le, Bach C., et al.
Veröffentlicht: (2025)
On the Effectiveness of the z-Transform Method in Quadratic Optimization
von: Bach, Francis
Veröffentlicht: (2025)
von: Bach, Francis
Veröffentlicht: (2025)
Accelerated Optimization Landscape of Linear-Quadratic Regulator
von: Feng, Lechen, et al.
Veröffentlicht: (2023)
von: Feng, Lechen, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Population Dynamics Control with Partial Observations
von: Lu, Zhou, et al.
Veröffentlicht: (2025) -
Online Control in Population Dynamics
von: Golowich, Noah, et al.
Veröffentlicht: (2024) -
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026) -
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024) -
Bandit Allocational Instability
von: Chen, Yilun, et al.
Veröffentlicht: (2026)