Algorithm Design and Stronger Guarantees for the Improving Multi-Armed Bandits Problem
Fuente:
arXiv
Saved in:
| Main Authors: | Blum, Avrim, Garicano, Marten, Ravichandran, Kavya, Sharma, Dravyansh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Nearly-tight Approximation Guarantees for the Improving Multi-Armed Bandits Problem
by: Blum, Avrim, et al.
Published: (2024)
by: Blum, Avrim, et al.
Published: (2024)
A Model for Combinatorial Dictionary Learning and Inference
by: Blum, Avrim, et al.
Published: (2024)
by: Blum, Avrim, et al.
Published: (2024)
On Learning Verifiers and Implications to Chain-of-Thought Reasoning
by: Balcan, Maria-Florina, et al.
Published: (2025)
by: Balcan, Maria-Florina, et al.
Published: (2025)
Online Learnability of Chain-of-Thought Verifiers: Soundness and Completeness Trade-offs
by: Balcan, Maria-Florina, et al.
Published: (2026)
by: Balcan, Maria-Florina, et al.
Published: (2026)
Tuning Algorithmic and Architectural Hyperparameters in Graph-Based Semi-Supervised Learning with Provable Guarantees
by: Du, Ally Yalei, et al.
Published: (2025)
by: Du, Ally Yalei, et al.
Published: (2025)
PAC Learning with Improvements
by: Attias, Idan, et al.
Published: (2025)
by: Attias, Idan, et al.
Published: (2025)
Non-Stationary Restless Multi-Armed Bandits with Provable Guarantee
by: Hung, Yu-Heng, et al.
Published: (2025)
by: Hung, Yu-Heng, et al.
Published: (2025)
Achieving PAC Guarantees in Mechanism Design through Multi-Armed Bandits
by: Osogami, Takayuki, et al.
Published: (2024)
by: Osogami, Takayuki, et al.
Published: (2024)
Gradient Descent with Provably Tuned Learning-rate Schedules
by: Sharma, Dravyansh
Published: (2025)
by: Sharma, Dravyansh
Published: (2025)
Optimal Streaming Algorithms for Multi-Armed Bandits
by: Jin, Tianyuan, et al.
Published: (2024)
by: Jin, Tianyuan, et al.
Published: (2024)
Recovering from Biased Data: Can Fairness Constraints Improve Accuracy?
by: Blum, Avrim, et al.
Published: (2019)
by: Blum, Avrim, et al.
Published: (2019)
Distributed Algorithms for Multi-Agent Multi-Armed Bandits with Collision
by: Zhou, Daoyuan, et al.
Published: (2025)
by: Zhou, Daoyuan, et al.
Published: (2025)
Autonomous Drug Design with Multi-Armed Bandits
by: Svensson, Hampus Gummesson, et al.
Published: (2022)
by: Svensson, Hampus Gummesson, et al.
Published: (2022)
Multi-Armed Bandits with Interference
by: Jia, Su, et al.
Published: (2024)
by: Jia, Su, et al.
Published: (2024)
Imprecise Multi-Armed Bandits
by: Kosoy, Vanessa
Published: (2024)
by: Kosoy, Vanessa
Published: (2024)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
by: Xu, Tianyi, et al.
Published: (2025)
by: Xu, Tianyi, et al.
Published: (2025)
Pessimism Traps and Algorithmic Interventions
by: Blum, Avrim, et al.
Published: (2024)
by: Blum, Avrim, et al.
Published: (2024)
A General Recipe for the Analysis of Randomized Multi-Armed Bandit Algorithms
by: Baudry, Dorian, et al.
Published: (2023)
by: Baudry, Dorian, et al.
Published: (2023)
The Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandit with Many Arms
by: Bayati, Mohsen, et al.
Published: (2020)
by: Bayati, Mohsen, et al.
Published: (2020)
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
Meet Me at the Arm: The Cooperative Multi-Armed Bandits Problem with Shareable Arms
by: Hu, Xinyi, et al.
Published: (2025)
by: Hu, Xinyi, et al.
Published: (2025)
Open Problem: Tight Bounds for Kernelized Multi-Armed Bandits with Bernoulli Rewards
by: Mussi, Marco, et al.
Published: (2024)
by: Mussi, Marco, et al.
Published: (2024)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
by: Davoodi, Mansoor, et al.
Published: (2025)
by: Davoodi, Mansoor, et al.
Published: (2025)
Multi-Label Learning with Stronger Consistency Guarantees
by: Mao, Anqi, et al.
Published: (2024)
by: Mao, Anqi, et al.
Published: (2024)
Multi-Armed Bandits with Network Interference
by: Agarwal, Abhineet, et al.
Published: (2024)
by: Agarwal, Abhineet, et al.
Published: (2024)
An Experimental Design for Anytime-Valid Causal Inference on Multi-Armed Bandits
by: Liang, Biyonka, et al.
Published: (2023)
by: Liang, Biyonka, et al.
Published: (2023)
On the Regularity and Fairness of Combinatorial Multi-Armed Bandit
by: Wu, Xiaoyi, et al.
Published: (2025)
by: Wu, Xiaoyi, et al.
Published: (2025)
Rising Multi-Armed Bandits with Known Horizons
by: Song, Seockbean, et al.
Published: (2026)
by: Song, Seockbean, et al.
Published: (2026)
Collaborating in Multi-Armed Bandits with Strategic Agents
by: Barnea, Idan, et al.
Published: (2026)
by: Barnea, Idan, et al.
Published: (2026)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
by: Nakamura, Shintaro, et al.
Published: (2023)
by: Nakamura, Shintaro, et al.
Published: (2023)
Algorithm Configuration for Structured Pfaffian Settings
by: Balcan, Maria-Florina, et al.
Published: (2024)
by: Balcan, Maria-Florina, et al.
Published: (2024)
Multi-Agent Combinatorial-Multi-Armed-Bandit framework for the Submodular Welfare Problem under Bandit Feedback
by: Pokhriyal, Subham, et al.
Published: (2026)
by: Pokhriyal, Subham, et al.
Published: (2026)
Offline-to-online hyperparameter transfer for stochastic bandits
by: Sharma, Dravyansh, et al.
Published: (2025)
by: Sharma, Dravyansh, et al.
Published: (2025)
Introduction to Multi-Armed Bandits
by: Slivkins, Aleksandrs
Published: (2019)
by: Slivkins, Aleksandrs
Published: (2019)
Learning accurate and interpretable tree-based models
by: Balcan, Maria-Florina, et al.
Published: (2024)
by: Balcan, Maria-Florina, et al.
Published: (2024)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Data-dependent Bounds with $T$-Optimal Best-of-Both-Worlds Guarantees in Multi-Armed Bandits using Stability-Penalty Matching
by: Nguyen, Quan, et al.
Published: (2025)
by: Nguyen, Quan, et al.
Published: (2025)
Improving Reward-Conditioned Policies for Multi-Armed Bandits using Normalized Weight Functions
by: Xu, Kai, et al.
Published: (2024)
by: Xu, Kai, et al.
Published: (2024)
Competitive strategies to use "warm start" algorithms with predictions
by: Srinivas, Vaidehi, et al.
Published: (2024)
by: Srinivas, Vaidehi, et al.
Published: (2024)
Strategic Multi-Armed Bandit Problems Under Debt-Free Reporting
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
Similar Items
-
Nearly-tight Approximation Guarantees for the Improving Multi-Armed Bandits Problem
by: Blum, Avrim, et al.
Published: (2024) -
A Model for Combinatorial Dictionary Learning and Inference
by: Blum, Avrim, et al.
Published: (2024) -
On Learning Verifiers and Implications to Chain-of-Thought Reasoning
by: Balcan, Maria-Florina, et al.
Published: (2025) -
Online Learnability of Chain-of-Thought Verifiers: Soundness and Completeness Trade-offs
by: Balcan, Maria-Florina, et al.
Published: (2026) -
Tuning Algorithmic and Architectural Hyperparameters in Graph-Based Semi-Supervised Learning with Provable Guarantees
by: Du, Ally Yalei, et al.
Published: (2025)