Linear bandits with polylogarithmic minimax regret
Fuente:
arXiv
Guardado en:
| Autores principales: | Lumbreras, Josep, Tomamichel, Marco |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning pure quantum states (almost) without regret
por: Lumbreras, Josep, et al.
Publicado: (2024)
por: Lumbreras, Josep, et al.
Publicado: (2024)
Quantum contextual bandits and recommender systems for quantum data
por: Brahmachari, Shrigyan, et al.
Publicado: (2023)
por: Brahmachari, Shrigyan, et al.
Publicado: (2023)
Quantum state-agnostic work extraction (almost) without dissipation
por: Lumbreras, Josep, et al.
Publicado: (2025)
por: Lumbreras, Josep, et al.
Publicado: (2025)
Learning Pure Quantum States in Any Dimension (Almost) Without Regret
por: Lumbreras, Josep, et al.
Publicado: (2026)
por: Lumbreras, Josep, et al.
Publicado: (2026)
Bandits roaming Hilbert space
por: Lumbreras, Josep
Publicado: (2025)
por: Lumbreras, Josep
Publicado: (2025)
Spectral bandits
por: Kocák, Tomáš, et al.
Publicado: (2026)
por: Kocák, Tomáš, et al.
Publicado: (2026)
minimax: Efficient Baselines for Autocurricula in JAX
por: Jiang, Minqi, et al.
Publicado: (2023)
por: Jiang, Minqi, et al.
Publicado: (2023)
Reinforcement learning for quantum processes with memory
por: Lumbreras, Josep, et al.
Publicado: (2026)
por: Lumbreras, Josep, et al.
Publicado: (2026)
Adversarial bandit optimization for approximately linear functions
por: Cheng, Zhuoyu, et al.
Publicado: (2025)
por: Cheng, Zhuoyu, et al.
Publicado: (2025)
On the optimal regret of collaborative personalized linear bandits
por: Huang, Bruce, et al.
Publicado: (2025)
por: Huang, Bruce, et al.
Publicado: (2025)
Reinforcement learning with combinatorial actions for coupled restless bandits
por: Xu, Lily, et al.
Publicado: (2025)
por: Xu, Lily, et al.
Publicado: (2025)
Functional multi-armed bandit and the best function identification problems
por: Dorn, Yuriy, et al.
Publicado: (2025)
por: Dorn, Yuriy, et al.
Publicado: (2025)
Prior-informed optimization of treatment recommendation via bandit algorithms trained on large language model-processed historical records
por: Nessari, Saman, et al.
Publicado: (2025)
por: Nessari, Saman, et al.
Publicado: (2025)
A conversion theorem and minimax optimality for continuum contextual bandits
por: Akhavan, Arya, et al.
Publicado: (2024)
por: Akhavan, Arya, et al.
Publicado: (2024)
Softmax gradient policy for variance minimization and risk-averse multi armed bandits
por: Turinici, Gabriel
Publicado: (2026)
por: Turinici, Gabriel
Publicado: (2026)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
por: Lin, Xiaoqiang, et al.
Publicado: (2023)
por: Lin, Xiaoqiang, et al.
Publicado: (2023)
UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
por: Han, Qiyang, et al.
Publicado: (2024)
por: Han, Qiyang, et al.
Publicado: (2024)
CATS-Linear: Classification Auxiliary Linear Model for Time Series Forecasting
por: Jibao, Zipo, et al.
Publicado: (2025)
por: Jibao, Zipo, et al.
Publicado: (2025)
Perturbation-Induced Linearization: Constructing Unlearnable Data with Solely Linear Classifiers
por: Liu, Jinlin, et al.
Publicado: (2026)
por: Liu, Jinlin, et al.
Publicado: (2026)
Exact Linear Attention
por: Ou, Weinuo
Publicado: (2026)
por: Ou, Weinuo
Publicado: (2026)
Kaczmarz Linear Attention
por: Zou, Jiaxuan, et al.
Publicado: (2026)
por: Zou, Jiaxuan, et al.
Publicado: (2026)
vLinear: A Powerful Linear Model for Multivariate Time Series Forecasting
por: Yue, Wenzhen, et al.
Publicado: (2026)
por: Yue, Wenzhen, et al.
Publicado: (2026)
Distilling Linearized Behavior into Non-Linear Fine-Tuning for Effective Task Arithmetic
por: Sommariva, Thomas, et al.
Publicado: (2026)
por: Sommariva, Thomas, et al.
Publicado: (2026)
Local Linear Attention: An Optimal Interpolation of Linear and Softmax Attention For Test-Time Regression
por: Zuo, Yifei, et al.
Publicado: (2025)
por: Zuo, Yifei, et al.
Publicado: (2025)
Transolver is a Linear Transformer: Revisiting Physics-Attention through the Lens of Linear Attention
por: Hu, Wenjie, et al.
Publicado: (2025)
por: Hu, Wenjie, et al.
Publicado: (2025)
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
por: Beck, Maximilian, et al.
Publicado: (2025)
por: Beck, Maximilian, et al.
Publicado: (2025)
Adaptive Locally Linear Embedding
por: Goli, Ali, et al.
Publicado: (2025)
por: Goli, Ali, et al.
Publicado: (2025)
LinearizeLLM: An Agent-Based Framework for LLM-Driven Exact Linear Reformulation of Nonlinear Optimization Problems
por: Kandora, Paul-Niklas Ken, et al.
Publicado: (2025)
por: Kandora, Paul-Niklas Ken, et al.
Publicado: (2025)
A Differentiable Integer Linear Programming Solver for Explanation-Based Natural Language Inference
por: Thayaparan, Mokanarangan, et al.
Publicado: (2024)
por: Thayaparan, Mokanarangan, et al.
Publicado: (2024)
Enhancing Linear Attention with Residual Learning
por: Lai, Xunhao, et al.
Publicado: (2025)
por: Lai, Xunhao, et al.
Publicado: (2025)
Representation Alignment Rests on Linear Structure
por: Bangachev, Kiril, et al.
Publicado: (2026)
por: Bangachev, Kiril, et al.
Publicado: (2026)
The Laplacian Keyboard: Beyond the Linear Span
por: Chandrasekar, Siddarth, et al.
Publicado: (2026)
por: Chandrasekar, Siddarth, et al.
Publicado: (2026)
Linearity-based neural network compression
por: Dobler, Silas, et al.
Publicado: (2025)
por: Dobler, Silas, et al.
Publicado: (2025)
Linear Causal Discovery with Interventional Constraints
por: Guo, Zhigao, et al.
Publicado: (2025)
por: Guo, Zhigao, et al.
Publicado: (2025)
Convergent Linear Representations of Emergent Misalignment
por: Soligo, Anna, et al.
Publicado: (2025)
por: Soligo, Anna, et al.
Publicado: (2025)
Grokking in Linear Models for Logistic Regression
por: Das, Nataraj, et al.
Publicado: (2026)
por: Das, Nataraj, et al.
Publicado: (2026)
DP2Unlearning: An Efficient and Guaranteed Unlearning Framework for LLMs
por: Mahmud, Tamim Al, et al.
Publicado: (2025)
por: Mahmud, Tamim Al, et al.
Publicado: (2025)
A Generalization Bound for Nearly-Linear Networks
por: Golikov, Eugene
Publicado: (2024)
por: Golikov, Eugene
Publicado: (2024)
Federated Linear Contextual Bandits with Heterogeneous Clients
por: Blaser, Ethan, et al.
Publicado: (2024)
por: Blaser, Ethan, et al.
Publicado: (2024)
On the Relation Between Linear Diffusion and Power Iteration
por: Weitzner, Dana, et al.
Publicado: (2024)
por: Weitzner, Dana, et al.
Publicado: (2024)
Ejemplares similares
-
Learning pure quantum states (almost) without regret
por: Lumbreras, Josep, et al.
Publicado: (2024) -
Quantum contextual bandits and recommender systems for quantum data
por: Brahmachari, Shrigyan, et al.
Publicado: (2023) -
Quantum state-agnostic work extraction (almost) without dissipation
por: Lumbreras, Josep, et al.
Publicado: (2025) -
Learning Pure Quantum States in Any Dimension (Almost) Without Regret
por: Lumbreras, Josep, et al.
Publicado: (2026) -
Bandits roaming Hilbert space
por: Lumbreras, Josep
Publicado: (2025)