A Perturbation Approach to Unconstrained Linear Bandits
Fuente:
arXiv
Salvato in:
| Autori principali: | Jacobsen, Andrew, Baudry, Dorian, Ito, Shinji, Cesa-Bianchi, Nicolò |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
di: Rumi, Alberto, et al.
Pubblicazione: (2026)
di: Rumi, Alberto, et al.
Pubblicazione: (2026)
Gradient-Variation Regret Bounds for Unconstrained Online Learning
di: Zhao, Yuheng, et al.
Pubblicazione: (2026)
di: Zhao, Yuheng, et al.
Pubblicazione: (2026)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
di: Jin, Tianyuan, et al.
Pubblicazione: (2024)
di: Jin, Tianyuan, et al.
Pubblicazione: (2024)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
di: Qiu, Hao, et al.
Pubblicazione: (2026)
di: Qiu, Hao, et al.
Pubblicazione: (2026)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
di: Kuroki, Yuko, et al.
Pubblicazione: (2023)
di: Kuroki, Yuko, et al.
Pubblicazione: (2023)
Dynamic Regret Reduces to Kernelized Static Regret
di: Jacobsen, Andrew, et al.
Pubblicazione: (2025)
di: Jacobsen, Andrew, et al.
Pubblicazione: (2025)
Improved Regret Bounds for Bandits with Expert Advice
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2024)
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2024)
Beyond Bandit Feedback in Online Multiclass Classification
di: van der Hoeven, Dirk, et al.
Pubblicazione: (2021)
di: van der Hoeven, Dirk, et al.
Pubblicazione: (2021)
Online Linear Regression with Paid Stochastic Features
di: Merlis, Nadav, et al.
Pubblicazione: (2025)
di: Merlis, Nadav, et al.
Pubblicazione: (2025)
A General Recipe for the Analysis of Randomized Multi-Armed Bandit Algorithms
di: Baudry, Dorian, et al.
Pubblicazione: (2023)
di: Baudry, Dorian, et al.
Pubblicazione: (2023)
Information Capacity Regret Bounds for Bandits with Mediator Feedback
di: Eldowa, Khaled, et al.
Pubblicazione: (2024)
di: Eldowa, Khaled, et al.
Pubblicazione: (2024)
Instance-Dependent Regret Bounds for Nonstochastic Linear Partial Monitoring
di: Di Gennaro, Federico, et al.
Pubblicazione: (2025)
di: Di Gennaro, Federico, et al.
Pubblicazione: (2025)
Online Budget Allocation with Censored Semi-Bandit Feedback
di: Bachoc, François, et al.
Pubblicazione: (2025)
di: Bachoc, François, et al.
Pubblicazione: (2025)
Revisiting Follow-the-Perturbed-Leader with Unbounded Perturbations in Bandit Problems
di: Lee, Jongyeong, et al.
Pubblicazione: (2025)
di: Lee, Jongyeong, et al.
Pubblicazione: (2025)
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
di: Kato, Masahiro, et al.
Pubblicazione: (2024)
di: Kato, Masahiro, et al.
Pubblicazione: (2024)
Non-stationary Bandit Convex Optimization: A Comprehensive Study
di: Liu, Xiaoqi, et al.
Pubblicazione: (2025)
di: Liu, Xiaoqi, et al.
Pubblicazione: (2025)
Lookahead identification in adversarial bandits: accuracy and memory bounds
di: Brukhim, Nataly, et al.
Pubblicazione: (2026)
di: Brukhim, Nataly, et al.
Pubblicazione: (2026)
Distributed Online Optimization with Stochastic Agent Availability
di: Achddou, Juliette, et al.
Pubblicazione: (2024)
di: Achddou, Juliette, et al.
Pubblicazione: (2024)
Multitask Online Learning: Listen to the Neighborhood Buzz
di: Achddou, Juliette, et al.
Pubblicazione: (2023)
di: Achddou, Juliette, et al.
Pubblicazione: (2023)
Heavy-tailed Linear Bandits: Adversarial Robustness, Best-of-both-worlds, and Beyond
di: Zhao, Canzhe, et al.
Pubblicazione: (2025)
di: Zhao, Canzhe, et al.
Pubblicazione: (2025)
Adaptive maximization of social welfare
di: Cesa-Bianchi, Nicolo, et al.
Pubblicazione: (2023)
di: Cesa-Bianchi, Nicolo, et al.
Pubblicazione: (2023)
Follow-the-Perturbed-Leader with Fréchet-type Tail Distributions: Optimality in Adversarial Bandits and Best-of-Both-Worlds
di: Lee, Jongyeong, et al.
Pubblicazione: (2024)
di: Lee, Jongyeong, et al.
Pubblicazione: (2024)
Cooperative Online Learning with Feedback Graphs
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2021)
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2021)
Influential Bandits: Pulling an Arm May Change the Environment
di: Sato, Ryoma, et al.
Pubblicazione: (2025)
di: Sato, Ryoma, et al.
Pubblicazione: (2025)
Bandit Max-Min Fair Allocation
di: Harada, Tsubasa, et al.
Pubblicazione: (2025)
di: Harada, Tsubasa, et al.
Pubblicazione: (2025)
Best-of-Both Worlds for linear contextual bandits with paid observations
di: Boyer, Nathan, et al.
Pubblicazione: (2025)
di: Boyer, Nathan, et al.
Pubblicazione: (2025)
The Value of Reward Lookahead in Reinforcement Learning
di: Merlis, Nadav, et al.
Pubblicazione: (2024)
di: Merlis, Nadav, et al.
Pubblicazione: (2024)
A Tight Lower Bound for Non-stochastic Multi-armed Bandits with Expert Advice
di: Chase, Zachary, et al.
Pubblicazione: (2025)
di: Chase, Zachary, et al.
Pubblicazione: (2025)
Self-Concordant Perturbations for Linear Bandits
di: Lévy, Lucas, et al.
Pubblicazione: (2025)
di: Lévy, Lucas, et al.
Pubblicazione: (2025)
Fair Online Bilateral Trade
di: Bachoc, François, et al.
Pubblicazione: (2024)
di: Bachoc, François, et al.
Pubblicazione: (2024)
A Regret Analysis of Bilateral Trade
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2021)
di: Cesa-Bianchi, Nicolò, et al.
Pubblicazione: (2021)
Combinatorial Allocation Bandits with Nonlinear Arm Utility
di: Shibukawa, Yuki, et al.
Pubblicazione: (2026)
di: Shibukawa, Yuki, et al.
Pubblicazione: (2026)
Replicability is Asymptotically Free in Multi-armed Bandits
di: Komiyama, Junpei, et al.
Pubblicazione: (2024)
di: Komiyama, Junpei, et al.
Pubblicazione: (2024)
New Classes of the Greedy-Applicable Arm Feature Distributions in the Sparse Linear Bandit Problem
di: Ichikawa, Koji, et al.
Pubblicazione: (2023)
di: Ichikawa, Koji, et al.
Pubblicazione: (2023)
Learning on the Edge: Online Learning with Stochastic Feedback Graphs
di: Esposito, Emmanuel, et al.
Pubblicazione: (2022)
di: Esposito, Emmanuel, et al.
Pubblicazione: (2022)
A Theory of Interpretable Approximations
di: Bressan, Marco, et al.
Pubblicazione: (2024)
di: Bressan, Marco, et al.
Pubblicazione: (2024)
A Probabilistic Approach to Learning the Degree of Equivariance in Steerable CNNs
di: Veefkind, Lars, et al.
Pubblicazione: (2024)
di: Veefkind, Lars, et al.
Pubblicazione: (2024)
Adversarial Bandit Optimization with Globally Bounded Perturbations to Linear Losses
di: Cheng, Zhuoyu, et al.
Pubblicazione: (2026)
di: Cheng, Zhuoyu, et al.
Pubblicazione: (2026)
Of Dice and Games: A Theory of Generalized Boosting
di: Bressan, Marco, et al.
Pubblicazione: (2024)
di: Bressan, Marco, et al.
Pubblicazione: (2024)
Adapting to Stochastic and Adversarial Losses in Episodic MDPs with Aggregate Bandit Feedback
di: Ito, Shinji, et al.
Pubblicazione: (2025)
di: Ito, Shinji, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
di: Rumi, Alberto, et al.
Pubblicazione: (2026) -
Gradient-Variation Regret Bounds for Unconstrained Online Learning
di: Zhao, Yuheng, et al.
Pubblicazione: (2026) -
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
di: Jin, Tianyuan, et al.
Pubblicazione: (2024) -
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
di: Qiu, Hao, et al.
Pubblicazione: (2026) -
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
di: Kuroki, Yuko, et al.
Pubblicazione: (2023)