Follow-the-Perturbed-Leader Approaches Best-of-Both-Worlds for the m-Set Semi-Bandit Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhan, Jingxin, Xin, Yuchen, Sun, Chenjie, Zhang, Zhihua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
Follow-the-Perturbed-Leader with Fréchet-type Tail Distributions: Optimality in Adversarial Bandits and Best-of-Both-Worlds
von: Lee, Jongyeong, et al.
Veröffentlicht: (2024)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2024)
A Further Efficient Algorithm with Best-of-Both-Worlds Guarantees for $m$-Set Semi-Bandit Problem
von: Chen, Botao, et al.
Veröffentlicht: (2026)
von: Chen, Botao, et al.
Veröffentlicht: (2026)
Note on Follow-the-Perturbed-Leader in Combinatorial Semi-Bandit Problems
von: Chen, Botao, et al.
Veröffentlicht: (2025)
von: Chen, Botao, et al.
Veröffentlicht: (2025)
A Regularized Online Newton Method for Stochastic Convex Bandits with Linear Vanishing Noise
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
Revisiting Follow-the-Perturbed-Leader with Unbounded Perturbations in Bandit Problems
von: Lee, Jongyeong, et al.
Veröffentlicht: (2025)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2025)
Last-Iterate Analyses of FTRL with the 1/2-Tsallis Entropy in Stochastic Bandits
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)
Efficient Best-of-Both-Worlds Algorithms for Contextual Combinatorial Semi-Bandits
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
Adaptive Learning Rate for Follow-the-Regularized-Leader: Competitive Analysis and Best-of-Both-Worlds
von: Ito, Shinji, et al.
Veröffentlicht: (2024)
von: Ito, Shinji, et al.
Veröffentlicht: (2024)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds Policy Optimization for CMDPs with Bandit Feedback
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
The Best of Both Worlds: On the Dilemma of Out-of-distribution Detection
von: Zhang, Qingyang, et al.
Veröffentlicht: (2024)
von: Zhang, Qingyang, et al.
Veröffentlicht: (2024)
Data-dependent Bounds with $T$-Optimal Best-of-Both-Worlds Guarantees in Multi-Armed Bandits using Stability-Penalty Matching
von: Nguyen, Quan, et al.
Veröffentlicht: (2025)
von: Nguyen, Quan, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
von: Akash, S, et al.
Veröffentlicht: (2026)
von: Akash, S, et al.
Veröffentlicht: (2026)
Learning to Bid in FCR Markets: A Best-of-Both-Worlds Approach
von: Potfer, Marius, et al.
Veröffentlicht: (2026)
von: Potfer, Marius, et al.
Veröffentlicht: (2026)
Best-of-Both Worlds for linear contextual bandits with paid observations
von: Boyer, Nathan, et al.
Veröffentlicht: (2025)
von: Boyer, Nathan, et al.
Veröffentlicht: (2025)
Best of Both Worlds: Regret Minimization versus Minimax Play
von: Müller, Adrian, et al.
Veröffentlicht: (2025)
von: Müller, Adrian, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds for Heavy-Tailed Markov Decision Processes
von: Chen, Yu, et al.
Veröffentlicht: (2026)
von: Chen, Yu, et al.
Veröffentlicht: (2026)
Multi-Play Combinatorial Semi-Bandit Problem
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2025)
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2025)
On the Problem of Best Arm Retention
von: Chen, Houshuang, et al.
Veröffentlicht: (2025)
von: Chen, Houshuang, et al.
Veröffentlicht: (2025)
A Perturbation Approach to Unconstrained Linear Bandits
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2026)
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2026)
A Best-of-Both-Worlds Algorithm for Constrained MDPs with Long-Term Constraints
von: Germano, Jacopo, et al.
Veröffentlicht: (2023)
von: Germano, Jacopo, et al.
Veröffentlicht: (2023)
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models
von: Behrouz, Ali, et al.
Veröffentlicht: (2024)
von: Behrouz, Ali, et al.
Veröffentlicht: (2024)
uniINF: Best-of-Both-Worlds Algorithm for Parameter-Free Heavy-Tailed MABs
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
Best of Both Worlds: Practical and Theoretically Optimal Submodular Maximization in Parallel
von: Chen, Yixin, et al.
Veröffentlicht: (2021)
von: Chen, Yixin, et al.
Veröffentlicht: (2021)
A Best-of-Both-Worlds Proof for Tsallis-INF without Fenchel Conjugates
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
Best of Both Worlds Guarantees for Smoothed Online Quadratic Optimization
von: Bhuyan, Neelkamal, et al.
Veröffentlicht: (2023)
von: Bhuyan, Neelkamal, et al.
Veröffentlicht: (2023)
Self-Concordant Perturbations for Linear Bandits
von: Lévy, Lucas, et al.
Veröffentlicht: (2025)
von: Lévy, Lucas, et al.
Veröffentlicht: (2025)
Fast Best-in-Class Regret for Contextual Bandits
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
Best Group Identification in Multi-Objective Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Constrained Best Arm Identification in Grouped Bandits
von: Dharod, Sahil, et al.
Veröffentlicht: (2024)
von: Dharod, Sahil, et al.
Veröffentlicht: (2024)
Best Arm Identification for Stochastic Rising Bandits
von: Mussi, Marco, et al.
Veröffentlicht: (2023)
von: Mussi, Marco, et al.
Veröffentlicht: (2023)
Best of Many in Both Worlds: Online Resource Allocation with Predictions under Unknown Arrival Model
von: An, Lin, et al.
Veröffentlicht: (2024)
von: An, Lin, et al.
Veröffentlicht: (2024)
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
Best-Arm Identification in Unimodal Bandits
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
The Best of Both Worlds: Hybridizing Neural Operators and Solvers for Stable Long-Horizon Inference
von: Roy, Rajyasri, et al.
Veröffentlicht: (2025)
von: Roy, Rajyasri, et al.
Veröffentlicht: (2025)
Exploration via Feature Perturbation in Contextual Bandits
von: Yi, Seouh-won, et al.
Veröffentlicht: (2025)
von: Yi, Seouh-won, et al.
Veröffentlicht: (2025)
Oracle-Efficient Combinatorial Semi-Bandits
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025) -
Follow-the-Perturbed-Leader with Fréchet-type Tail Distributions: Optimality in Adversarial Bandits and Best-of-Both-Worlds
von: Lee, Jongyeong, et al.
Veröffentlicht: (2024) -
A Further Efficient Algorithm with Best-of-Both-Worlds Guarantees for $m$-Set Semi-Bandit Problem
von: Chen, Botao, et al.
Veröffentlicht: (2026) -
Note on Follow-the-Perturbed-Leader in Combinatorial Semi-Bandit Problems
von: Chen, Botao, et al.
Veröffentlicht: (2025) -
A Regularized Online Newton Method for Stochastic Convex Bandits with Linear Vanishing Noise
von: Zhan, Jingxin, et al.
Veröffentlicht: (2025)