Gespeichert in:
| Hauptverfasser: | Blanchard, Moïse, Goyal, Vineet |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.10313 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Individual Regret in Cooperative Stochastic Multi-Armed Bandits
von: Barnea, Idan, et al.
Veröffentlicht: (2024)
von: Barnea, Idan, et al.
Veröffentlicht: (2024)
Graph-Dependent Regret Bounds in Multi-Armed Bandits with Interference
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025)
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025)
Bandit Max-Min Fair Allocation
von: Harada, Tsubasa, et al.
Veröffentlicht: (2025)
von: Harada, Tsubasa, et al.
Veröffentlicht: (2025)
Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning
von: Lee, Harin, et al.
Veröffentlicht: (2026)
von: Lee, Harin, et al.
Veröffentlicht: (2026)
Collaborating in Multi-Armed Bandits with Strategic Agents
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
Asymptotically and Minimax Optimal Regret Bounds for Multi-Armed Bandits with Abstention
von: Yang, Junwen, et al.
Veröffentlicht: (2024)
von: Yang, Junwen, et al.
Veröffentlicht: (2024)
Materials Discovery using Max K-Armed Bandit
von: Kikkawa, Nobuaki, et al.
Veröffentlicht: (2022)
von: Kikkawa, Nobuaki, et al.
Veröffentlicht: (2022)
Understanding Memory-Regret Trade-Off for Streaming Stochastic Multi-Armed Bandits
von: He, Yuchen, et al.
Veröffentlicht: (2024)
von: He, Yuchen, et al.
Veröffentlicht: (2024)
Agnostic Smoothed Online Learning without Knowledge of the Base Measure
von: Blanchard, Moïse
Veröffentlicht: (2024)
von: Blanchard, Moïse
Veröffentlicht: (2024)
Multi-Armed Bandits with Interference
von: Jia, Su, et al.
Veröffentlicht: (2024)
von: Jia, Su, et al.
Veröffentlicht: (2024)
Imprecise Multi-Armed Bandits
von: Kosoy, Vanessa
Veröffentlicht: (2024)
von: Kosoy, Vanessa
Veröffentlicht: (2024)
MNL-Bandit with Knapsacks: a near-optimal algorithm
von: Aznag, Abdellah, et al.
Veröffentlicht: (2021)
von: Aznag, Abdellah, et al.
Veröffentlicht: (2021)
Collaborative Multi-Agent Heterogeneous Multi-Armed Bandits
von: Chawla, Ronshee, et al.
Veröffentlicht: (2023)
von: Chawla, Ronshee, et al.
Veröffentlicht: (2023)
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Online Min-Max Optimization: From Individual Regrets to Cumulative Saddle Points
von: Vyas, Abhijeet, et al.
Veröffentlicht: (2026)
von: Vyas, Abhijeet, et al.
Veröffentlicht: (2026)
Multi-Armed Bandits with Network Interference
von: Agarwal, Abhineet, et al.
Veröffentlicht: (2024)
von: Agarwal, Abhineet, et al.
Veröffentlicht: (2024)
Introduction to Multi-Armed Bandits
von: Slivkins, Aleksandrs
Veröffentlicht: (2019)
von: Slivkins, Aleksandrs
Veröffentlicht: (2019)
Gradient Descent is Pareto-Optimal in the Oracle Complexity and Memory Tradeoff for Feasibility Problems
von: Blanchard, Moise
Veröffentlicht: (2024)
von: Blanchard, Moise
Veröffentlicht: (2024)
Near-Optimal Privacy-Preserving Learning for Max-Min Fair Multi-Agent Bandits
von: Leshem, Amir
Veröffentlicht: (2023)
von: Leshem, Amir
Veröffentlicht: (2023)
On the Regularity and Fairness of Combinatorial Multi-Armed Bandit
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
Rising Multi-Armed Bandits with Known Horizons
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
Optimal Streaming Algorithms for Multi-Armed Bandits
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
MinMaxMin $Q$-learning
von: Soffair, Nitsan, et al.
Veröffentlicht: (2024)
von: Soffair, Nitsan, et al.
Veröffentlicht: (2024)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
Autonomous Drug Design with Multi-Armed Bandits
von: Svensson, Hampus Gummesson, et al.
Veröffentlicht: (2022)
von: Svensson, Hampus Gummesson, et al.
Veröffentlicht: (2022)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
Byzantine-Resilient Decentralized Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2023)
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2023)
Multi-Armed Bandits With Best-Action Queries
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2026)
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2026)
Stochastic Multi-Armed Bandits with Limited Control Variates
von: Verma, Arun, et al.
Veröffentlicht: (2026)
von: Verma, Arun, et al.
Veröffentlicht: (2026)
Optimism in the Face of Ambiguity Principle for Multi-Armed Bandits
von: Li, Mengmeng, et al.
Veröffentlicht: (2024)
von: Li, Mengmeng, et al.
Veröffentlicht: (2024)
Federated Multi-Armed Bandits Under Byzantine Attacks
von: Saday, Artun, et al.
Veröffentlicht: (2022)
von: Saday, Artun, et al.
Veröffentlicht: (2022)
Distributed Algorithms for Multi-Agent Multi-Armed Bandits with Collision
von: Zhou, Daoyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Daoyuan, et al.
Veröffentlicht: (2025)
Distributionally-Constrained Adversaries in Online Learning
von: Blanchard, Moïse, et al.
Veröffentlicht: (2025)
von: Blanchard, Moïse, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Combinatorial Multi-Armed Bandits
von: Balasubramanian, Rishab, et al.
Veröffentlicht: (2023)
von: Balasubramanian, Rishab, et al.
Veröffentlicht: (2023)
Unlearning Offline Stochastic Multi-Armed Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2026)
von: Ye, Zichun, et al.
Veröffentlicht: (2026)
Distribution-Free Sequential Prediction with Abstentions
von: Yu, Jialin, et al.
Veröffentlicht: (2026)
von: Yu, Jialin, et al.
Veröffentlicht: (2026)
Global Rewards in Restless Multi-Armed Bandits
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Individual Regret in Cooperative Stochastic Multi-Armed Bandits
von: Barnea, Idan, et al.
Veröffentlicht: (2024) -
Graph-Dependent Regret Bounds in Multi-Armed Bandits with Interference
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025) -
Bandit Max-Min Fair Allocation
von: Harada, Tsubasa, et al.
Veröffentlicht: (2025) -
Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning
von: Lee, Harin, et al.
Veröffentlicht: (2026) -
Collaborating in Multi-Armed Bandits with Strategic Agents
von: Barnea, Idan, et al.
Veröffentlicht: (2026)