Introduction to Multi-Armed Bandits
Fuente:
arXiv
Saved in:
| Main Author: | Slivkins, Aleksandrs |
|---|---|
| Format: | Preprint |
| Published: |
2019
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Greedy Algorithm for Structured Bandits: A Sharp Characterization of Asymptotic Success / Failure
by: Slivkins, Aleksandrs, et al.
Published: (2025)
by: Slivkins, Aleksandrs, et al.
Published: (2025)
Adaptive Discretization against an Adversary: Lipschitz bandits, Dynamic Pricing, and Auction Tuning
by: Podimata, Chara, et al.
Published: (2020)
by: Podimata, Chara, et al.
Published: (2020)
Bandit Social Learning: Exploration under Myopic Behavior
by: Banihashem, Kiarash, et al.
Published: (2023)
by: Banihashem, Kiarash, et al.
Published: (2023)
Convergence of a L2 regularized Policy Gradient Algorithm for the Multi Armed Bandit
by: Anita, Stefana, et al.
Published: (2024)
by: Anita, Stefana, et al.
Published: (2024)
Adversarial Attacks on Combinatorial Multi-Armed Bandits
by: Balasubramanian, Rishab, et al.
Published: (2023)
by: Balasubramanian, Rishab, et al.
Published: (2023)
Unlearning Offline Stochastic Multi-Armed Bandits
by: Ye, Zichun, et al.
Published: (2026)
by: Ye, Zichun, et al.
Published: (2026)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
by: Davoodi, Mansoor, et al.
Published: (2025)
by: Davoodi, Mansoor, et al.
Published: (2025)
Stochastic Submodular Bandits with Delayed Composite Anonymous Bandit Feedback
by: Pedramfar, Mohammad, et al.
Published: (2023)
by: Pedramfar, Mohammad, et al.
Published: (2023)
Nearly-tight Approximation Guarantees for the Improving Multi-Armed Bandits Problem
by: Blum, Avrim, et al.
Published: (2024)
by: Blum, Avrim, et al.
Published: (2024)
Understanding Memory-Regret Trade-Off for Streaming Stochastic Multi-Armed Bandits
by: He, Yuchen, et al.
Published: (2024)
by: He, Yuchen, et al.
Published: (2024)
Incentivizing Exploration with Selective Data Disclosure
by: Immorlica, Nicole, et al.
Published: (2018)
by: Immorlica, Nicole, et al.
Published: (2018)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
by: Ye, Zichun, et al.
Published: (2025)
by: Ye, Zichun, et al.
Published: (2025)
Pareto Optimal Algorithmic Recourse in Multi-cost Function
by: Chen, Wen-Ling, et al.
Published: (2025)
by: Chen, Wen-Ling, et al.
Published: (2025)
AlgoSelect: Universal Algorithm Selection via the Comb Operator
by: Yao, Jasper
Published: (2025)
by: Yao, Jasper
Published: (2025)
An Algorithm for Learning Smaller Representations of Models With Scarce Data
by: de Wynter, Adrian
Published: (2020)
by: de Wynter, Adrian
Published: (2020)
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
by: Li, Dongyue, et al.
Published: (2025)
by: Li, Dongyue, et al.
Published: (2025)
Optimizing Text Search: A Novel Pattern Matching Algorithm Based on Ukkonen's Approach
by: Guan, Xinyu, et al.
Published: (2025)
by: Guan, Xinyu, et al.
Published: (2025)
Streaming Attention Approximation via Discrepancy Theory
by: Kochetkova, Ekaterina, et al.
Published: (2025)
by: Kochetkova, Ekaterina, et al.
Published: (2025)
Discovering Data Structures: Nearest Neighbor Search and Beyond
by: Salemohamed, Omar, et al.
Published: (2024)
by: Salemohamed, Omar, et al.
Published: (2024)
Optimal Survival Trees: A Dynamic Programming Approach
by: Huisman, Tim, et al.
Published: (2024)
by: Huisman, Tim, et al.
Published: (2024)
OpenTensor: Reproducing Faster Matrix Multiplication Discovering Algorithms
by: Sun, Yiwen, et al.
Published: (2024)
by: Sun, Yiwen, et al.
Published: (2024)
Mini-Batch Kernel $k$-means
by: Jourdan, Ben, et al.
Published: (2024)
by: Jourdan, Ben, et al.
Published: (2024)
Model Stealing for Any Low-Rank Language Model
by: Liu, Allen, et al.
Published: (2024)
by: Liu, Allen, et al.
Published: (2024)
Block-Diagonal Guided DBSCAN Clustering
by: Zhao, Weibing
Published: (2024)
by: Zhao, Weibing
Published: (2024)
Contract Scheduling with Distributional and Multiple Advice
by: Angelopoulos, Spyros, et al.
Published: (2024)
by: Angelopoulos, Spyros, et al.
Published: (2024)
SubGen: Token Generation in Sublinear Time and Memory
by: Zandieh, Amir, et al.
Published: (2024)
by: Zandieh, Amir, et al.
Published: (2024)
Anytime-Constrained Reinforcement Learning
by: McMahan, Jeremy, et al.
Published: (2023)
by: McMahan, Jeremy, et al.
Published: (2023)
A Fixed-Parameter Tractable Algorithm for Counting Markov Equivalence Classes with the same Skeleton
by: Sharma, Vidya Sagar
Published: (2023)
by: Sharma, Vidya Sagar
Published: (2023)
The sample complexity of multi-distribution learning
by: Peng, Binghui
Published: (2023)
by: Peng, Binghui
Published: (2023)
Approximate Lifted Model Construction
by: Luttermann, Malte, et al.
Published: (2025)
by: Luttermann, Malte, et al.
Published: (2025)
Online Learning with Probing for Sequential User-Centric Selection
by: Xu, Tianyi, et al.
Published: (2025)
by: Xu, Tianyi, et al.
Published: (2025)
Online bipartite matching with imperfect advice
by: Choo, Davin, et al.
Published: (2024)
by: Choo, Davin, et al.
Published: (2024)
Lower Bound on the Greedy Approximation Ratio for Adaptive Submodular Cover
by: Harris, Blake, et al.
Published: (2024)
by: Harris, Blake, et al.
Published: (2024)
Estimating Causal Effects in Partially Directed Parametric Causal Factor Graphs
by: Luttermann, Malte, et al.
Published: (2024)
by: Luttermann, Malte, et al.
Published: (2024)
Learning-Augmented Priority Queues
by: Benomar, Ziyad, et al.
Published: (2024)
by: Benomar, Ziyad, et al.
Published: (2024)
Demand Selection for VRP with Emission Quota
by: Najar, Farid, et al.
Published: (2025)
by: Najar, Farid, et al.
Published: (2025)
Constructing Decision Trees from Data Streams
by: Pham, Huy, et al.
Published: (2024)
by: Pham, Huy, et al.
Published: (2024)
Provably Learning from Modern Language Models via Low Logit Rank
by: Golowich, Noah, et al.
Published: (2025)
by: Golowich, Noah, et al.
Published: (2025)
Fast EXP3 Algorithms
by: Sato, Ryoma, et al.
Published: (2025)
by: Sato, Ryoma, et al.
Published: (2025)
Learning-augmented smooth integer programs with PAC-learnable oracles
by: He, Hao-Yuan, et al.
Published: (2026)
by: He, Hao-Yuan, et al.
Published: (2026)
Similar Items
-
Greedy Algorithm for Structured Bandits: A Sharp Characterization of Asymptotic Success / Failure
by: Slivkins, Aleksandrs, et al.
Published: (2025) -
Adaptive Discretization against an Adversary: Lipschitz bandits, Dynamic Pricing, and Auction Tuning
by: Podimata, Chara, et al.
Published: (2020) -
Bandit Social Learning: Exploration under Myopic Behavior
by: Banihashem, Kiarash, et al.
Published: (2023) -
Convergence of a L2 regularized Policy Gradient Algorithm for the Multi Armed Bandit
by: Anita, Stefana, et al.
Published: (2024) -
Adversarial Attacks on Combinatorial Multi-Armed Bandits
by: Balasubramanian, Rishab, et al.
Published: (2023)