One Good Source is All You Need: Near-Optimal Regret for Bandits under Heterogeneous Noise
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhat, Amith, Luo, Haipeng, Saha, Aadirupa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Near-Optimal Regret in Linear MDPs with Aggregate Bandit Feedback
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
Stop Relying on No-Choice and Do not Repeat the Moves: Optimal, Efficient and Practical Algorithms for Assortment Optimization
von: Saha, Aadirupa, et al.
Veröffentlicht: (2024)
von: Saha, Aadirupa, et al.
Veröffentlicht: (2024)
Revisiting Social Welfare in Bandits: UCB is (Nearly) All You Need
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2025)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2025)
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
Nearly Minimax Optimal Regret for Multinomial Logistic Bandit
von: Lee, Joongkyu, et al.
Veröffentlicht: (2024)
von: Lee, Joongkyu, et al.
Veröffentlicht: (2024)
Strategic Linear Contextual Bandits
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
Gaussian Process Upper Confidence Bound Achieves Nearly-Optimal Regret in Noise-Free Gaussian Process Bandits
von: Iwazaki, Shogo
Veröffentlicht: (2025)
von: Iwazaki, Shogo
Veröffentlicht: (2025)
Near-Optimal Last-Iterate Convergence for Zero-Sum Games with Bandit Feedback and Opponent Actions
von: Hait, Soumita, et al.
Veröffentlicht: (2026)
von: Hait, Soumita, et al.
Veröffentlicht: (2026)
DP-Dueling: Learning from Preference Feedback without Compromising User Privacy
von: Saha, Aadirupa, et al.
Veröffentlicht: (2024)
von: Saha, Aadirupa, et al.
Veröffentlicht: (2024)
Near-Optimal Regret for Efficient Stochastic Combinatorial Semi-Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
Generalized Linear Bandits: Almost Optimal Regret with One-Pass Update
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Optimal Regret for Single Index Bandits
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
Data-Driven Upper Confidence Bounds with Near-Optimal Regret for Heavy-Tailed Bandits
von: Tamás, Ambrus, et al.
Veröffentlicht: (2024)
von: Tamás, Ambrus, et al.
Veröffentlicht: (2024)
LLM-as-Judge on a Budget
von: Saha, Aadirupa, et al.
Veröffentlicht: (2026)
von: Saha, Aadirupa, et al.
Veröffentlicht: (2026)
Tracking the Best Expert Privately
von: Saha, Aadirupa, et al.
Veröffentlicht: (2025)
von: Saha, Aadirupa, et al.
Veröffentlicht: (2025)
On the Vulnerability of Fairness Constrained Learning to Malicious Noise
von: Blum, Avrim, et al.
Veröffentlicht: (2023)
von: Blum, Avrim, et al.
Veröffentlicht: (2023)
Optimal Regret for Policy Optimization in Contextual Bandits
von: Levy, Orin, et al.
Veröffentlicht: (2026)
von: Levy, Orin, et al.
Veröffentlicht: (2026)
Near-optimal Per-Action Regret Bounds for Sleeping Bandits
von: Nguyen, Quan, et al.
Veröffentlicht: (2024)
von: Nguyen, Quan, et al.
Veröffentlicht: (2024)
Regret Bounds for Noise-Free Cascaded Kernelized Bandits
von: Li, Zihan, et al.
Veröffentlicht: (2022)
von: Li, Zihan, et al.
Veröffentlicht: (2022)
Uni-LoRA: One Vector is All You Need
von: Li, Kaiyang, et al.
Veröffentlicht: (2025)
von: Li, Kaiyang, et al.
Veröffentlicht: (2025)
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
von: Azize, Achraf, et al.
Veröffentlicht: (2025)
von: Azize, Achraf, et al.
Veröffentlicht: (2025)
Contextual Multinomial Logit Bandits with General Value Functions
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
No-Regret Linear Bandits under Gap-Adjusted Misspecification
von: Liu, Chong, et al.
Veröffentlicht: (2025)
von: Liu, Chong, et al.
Veröffentlicht: (2025)
No-Regret Learning Under Adversarial Resource Constraints: A Spending Plan Is All You Need!
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2025)
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2025)
Noise is All You Need: Private Second-Order Convergence of Noisy SGD
von: Avdiukhin, Dmitrii, et al.
Veröffentlicht: (2024)
von: Avdiukhin, Dmitrii, et al.
Veröffentlicht: (2024)
Preference-centric Bandits: Optimality of Mixtures and Regret-efficient Algorithms
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
Regretful Decisions under Label Noise
von: Nagaraj, Sujay, et al.
Veröffentlicht: (2025)
von: Nagaraj, Sujay, et al.
Veröffentlicht: (2025)
Alternating Regret for Online Convex Optimization
von: Hait, Soumita, et al.
Veröffentlicht: (2025)
von: Hait, Soumita, et al.
Veröffentlicht: (2025)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
On the Optimal Regret of Locally Private Linear Contextual Bandit
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Accuracy is Not All You Need
von: Dutta, Abhinav, et al.
Veröffentlicht: (2024)
von: Dutta, Abhinav, et al.
Veröffentlicht: (2024)
Achieving Optimal Static and Dynamic Regret Simultaneously in Bandits with Deterministic Losses
von: Qian, Jian, et al.
Veröffentlicht: (2026)
von: Qian, Jian, et al.
Veröffentlicht: (2026)
Learning What to Recommend: Minimax Optimal Simple Regret in Logistic Bandits
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
Risk-sensitive Bandits: Arm Mixture Optimality and Regret-efficient Algorithms
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards
von: Panda, Subhodip, et al.
Veröffentlicht: (2026)
von: Panda, Subhodip, et al.
Veröffentlicht: (2026)
Near-optimal Regret Using Policy Optimization in Online MDPs with Aggregate Bandit Feedback
von: Lancewicki, Tal, et al.
Veröffentlicht: (2025)
von: Lancewicki, Tal, et al.
Veröffentlicht: (2025)
Optimal Multitask Linear Regression and Contextual Bandits under Sparse Heterogeneity
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Near-Optimal Regret in Linear MDPs with Aggregate Bandit Feedback
von: Cassel, Asaf, et al.
Veröffentlicht: (2024) -
Stop Relying on No-Choice and Do not Repeat the Moves: Optimal, Efficient and Practical Algorithms for Assortment Optimization
von: Saha, Aadirupa, et al.
Veröffentlicht: (2024) -
Revisiting Social Welfare in Bandits: UCB is (Nearly) All You Need
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2025) -
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026) -
Nearly Minimax Optimal Regret for Multinomial Logistic Bandit
von: Lee, Joongkyu, et al.
Veröffentlicht: (2024)