Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kuroki, Yuko, Rumi, Alberto, Tsuchiya, Taira, Vitale, Fabio, Cesa-Bianchi, Nicolò |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
von: Rumi, Alberto, et al.
Veröffentlicht: (2026)
von: Rumi, Alberto, et al.
Veröffentlicht: (2026)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
Efficient Best-of-Both-Worlds Algorithms for Contextual Combinatorial Semi-Bandits
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
A Perturbation Approach to Unconstrained Linear Bandits
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2026)
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2026)
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
Adaptive Learning Rate for Follow-the-Regularized-Leader: Competitive Analysis and Best-of-Both-Worlds
von: Ito, Shinji, et al.
Veröffentlicht: (2024)
von: Ito, Shinji, et al.
Veröffentlicht: (2024)
A Simple and Adaptive Learning Rate for FTRL in Online Learning with Minimax Regret of $Θ(T^{2/3})$ and its Application to Best-of-Both-Worlds
von: Tsuchiya, Taira, et al.
Veröffentlicht: (2024)
von: Tsuchiya, Taira, et al.
Veröffentlicht: (2024)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
Information Capacity Regret Bounds for Bandits with Mediator Feedback
von: Eldowa, Khaled, et al.
Veröffentlicht: (2024)
von: Eldowa, Khaled, et al.
Veröffentlicht: (2024)
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
Beyond Bandit Feedback in Online Multiclass Classification
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2021)
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2021)
Online Linear Regression with Paid Stochastic Features
von: Merlis, Nadav, et al.
Veröffentlicht: (2025)
von: Merlis, Nadav, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds Policy Optimization for CMDPs with Bandit Feedback
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
Instance-Dependent Regret Bounds for Nonstochastic Linear Partial Monitoring
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2025)
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2025)
Online Budget Allocation with Censored Semi-Bandit Feedback
von: Bachoc, François, et al.
Veröffentlicht: (2025)
von: Bachoc, François, et al.
Veröffentlicht: (2025)
Multi-Play Combinatorial Semi-Bandit Problem
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2025)
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2025)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Bandits with Abstention under Expert Advice
von: Pasteris, Stephen, et al.
Veröffentlicht: (2024)
von: Pasteris, Stephen, et al.
Veröffentlicht: (2024)
Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
A Further Efficient Algorithm with Best-of-Both-Worlds Guarantees for $m$-Set Semi-Bandit Problem
von: Chen, Botao, et al.
Veröffentlicht: (2026)
von: Chen, Botao, et al.
Veröffentlicht: (2026)
Multitask Online Learning: Listen to the Neighborhood Buzz
von: Achddou, Juliette, et al.
Veröffentlicht: (2023)
von: Achddou, Juliette, et al.
Veröffentlicht: (2023)
Lookahead identification in adversarial bandits: accuracy and memory bounds
von: Brukhim, Nataly, et al.
Veröffentlicht: (2026)
von: Brukhim, Nataly, et al.
Veröffentlicht: (2026)
Distributed Online Optimization with Stochastic Agent Availability
von: Achddou, Juliette, et al.
Veröffentlicht: (2024)
von: Achddou, Juliette, et al.
Veröffentlicht: (2024)
Adaptive maximization of social welfare
von: Cesa-Bianchi, Nicolo, et al.
Veröffentlicht: (2023)
von: Cesa-Bianchi, Nicolo, et al.
Veröffentlicht: (2023)
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
von: Akash, S, et al.
Veröffentlicht: (2026)
von: Akash, S, et al.
Veröffentlicht: (2026)
A Best-of-Both-Worlds Algorithm for Constrained MDPs with Long-Term Constraints
von: Germano, Jacopo, et al.
Veröffentlicht: (2023)
von: Germano, Jacopo, et al.
Veröffentlicht: (2023)
Bandit and Delayed Feedback in Online Structured Prediction
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
A Reduction Algorithm for Markovian Contextual Linear Bandits
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
Follow-the-Perturbed-Leader Approaches Best-of-Both-Worlds for the m-Set Semi-Bandit Problems
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
Online Inverse Linear Optimization: Efficient Logarithmic-Regret Algorithm, Robustness to Suboptimality, and Lower Bound
von: Sakaue, Shinsaku, et al.
Veröffentlicht: (2025)
von: Sakaue, Shinsaku, et al.
Veröffentlicht: (2025)
Dynamic Structure Estimation from Bandit Feedback using Nonvanishing Exponential Sums
von: Ohnishi, Motoya, et al.
Veröffentlicht: (2022)
von: Ohnishi, Motoya, et al.
Veröffentlicht: (2022)
Fast Best-in-Class Regret for Contextual Bandits
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
Online Minimization of Polarization and Disagreement via Low-Rank Matrix Bandits
von: Cinus, Federico, et al.
Veröffentlicht: (2025)
von: Cinus, Federico, et al.
Veröffentlicht: (2025)
Gradient-Variation Regret Bounds for Unconstrained Online Learning
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
Cooperative Online Learning with Feedback Graphs
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2021)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2021)
Dynamic Regret Reduces to Kernelized Static Regret
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2025)
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2025)
An Improved Algorithm for Adversarial Linear Contextual Bandits via Reduction
von: van Erven, Tim, et al.
Veröffentlicht: (2025)
von: van Erven, Tim, et al.
Veröffentlicht: (2025)
Online Control of Linear Systems under Unbounded Noise
von: Ito, Kaito, et al.
Veröffentlicht: (2024)
von: Ito, Kaito, et al.
Veröffentlicht: (2024)
Follow-the-Perturbed-Leader with Fréchet-type Tail Distributions: Optimality in Adversarial Bandits and Best-of-Both-Worlds
von: Lee, Jongyeong, et al.
Veröffentlicht: (2024)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2024)
Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates
von: Yu, Sanghoon, et al.
Veröffentlicht: (2026)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
von: Rumi, Alberto, et al.
Veröffentlicht: (2026) -
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024) -
Efficient Best-of-Both-Worlds Algorithms for Contextual Combinatorial Semi-Bandits
von: Li, Mengmeng, et al.
Veröffentlicht: (2025) -
A Perturbation Approach to Unconstrained Linear Bandits
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2026) -
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)