Few Batches or Little Memory, But Not Both: Simultaneous Space and Adaptivity Constraints in Stochastic Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Ruiyuan, Lyu, Zicheng, Zhu, Xiaoyi, Huang, Zengfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Nearly Tight Bounds for Cross-Learning Contextual Bandits with Graphical Feedback
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2025)
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2025)
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
Load--Reserve Wasserstein Propagation for Isotropic Diffusion Samplers
von: Lyu, Zicheng, et al.
Veröffentlicht: (2026)
von: Lyu, Zicheng, et al.
Veröffentlicht: (2026)
Batched Stochastic Bandit for Nondegenerate Functions
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
Sublinear Spectral Clustering Oracle with Little Memory
von: Shen, Ranran, et al.
Veröffentlicht: (2026)
von: Shen, Ranran, et al.
Veröffentlicht: (2026)
Batch Ensemble for Variance Dependent Regret in Stochastic Bandits
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
A Near-optimal, Scalable and Parallelizable Framework for Stochastic Bandits Robust to Adversarial Corruptions and Beyond
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Robust Batched Bandits
von: Guo, Yunwen, et al.
Veröffentlicht: (2025)
von: Guo, Yunwen, et al.
Veröffentlicht: (2025)
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
Increasing Both Batch Size and Learning Rate Accelerates Stochastic Gradient Descent
von: Umeda, Hikaru, et al.
Veröffentlicht: (2024)
von: Umeda, Hikaru, et al.
Veröffentlicht: (2024)
The Batch Complexity of Bandit Pure Exploration
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
Batched Nonparametric Contextual Bandits
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
Beyond Primal-Dual Methods in Bandits with Stochastic and Adversarial Constraints
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
Optimal and Practical Batched Linear Bandit Algorithm
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
The Adaptivity Barrier in Batched Nonparametric Bandits: Sharp Characterization of the Price of Unknown Margin
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
On the Regularity and Fairness of Combinatorial Multi-Armed Bandit
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
Batched Kernelized Bandits: Refinements and Extensions
von: Ma, Chenkai, et al.
Veröffentlicht: (2026)
von: Ma, Chenkai, et al.
Veröffentlicht: (2026)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
von: Akash, S, et al.
Veröffentlicht: (2026)
von: Akash, S, et al.
Veröffentlicht: (2026)
Best-of-Both-Worlds Policy Optimization for CMDPs with Bandit Feedback
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
von: Stradi, Francesco Emanuele, et al.
Veröffentlicht: (2024)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Online Bandits with (Biased) Offline Data: Adaptive Learning under Distribution Mismatch
von: Cheung, Wang Chi, et al.
Veröffentlicht: (2024)
von: Cheung, Wang Chi, et al.
Veröffentlicht: (2024)
No-Regret is not enough! Bandits with General Constraints through Adaptive Regret Minimization
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
Efficient and Adaptive Posterior Sampling Algorithms for Bandits
von: Hu, Bingshan, et al.
Veröffentlicht: (2024)
von: Hu, Bingshan, et al.
Veröffentlicht: (2024)
Your Graph Recommender is Provably a Single-view Graph Contrastive Learning
von: Yang, Wenjie, et al.
Veröffentlicht: (2024)
von: Yang, Wenjie, et al.
Veröffentlicht: (2024)
Continuum-armed Bandit Optimization with Batch Pairwise Comparison Oracles
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
Batched Online Contextual Sparse Bandits with Sequential Inclusion of Features
von: Swiers, Rowan, et al.
Veröffentlicht: (2024)
von: Swiers, Rowan, et al.
Veröffentlicht: (2024)
Space Complexity of Euclidean Clustering
von: Zhu, Xiaoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaoyi, et al.
Veröffentlicht: (2024)
Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
von: Kim, Chaiwon, et al.
Veröffentlicht: (2025)
Efficient Best-of-Both-Worlds Algorithms for Contextual Combinatorial Semi-Bandits
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
An Adaptive Approach for Infinitely Many-armed Bandits under Generalized Rotting Constraints
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
On the Low-Complexity of Fair Learning for Combinatorial Multi-Armed Bandit
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Efficient Clustering in Stochastic Bandits
von: Chandran, G Dhinesh, et al.
Veröffentlicht: (2026)
von: Chandran, G Dhinesh, et al.
Veröffentlicht: (2026)
Stochastic Bandits for Egalitarian Assignment
von: Lim, Eugene, et al.
Veröffentlicht: (2024)
von: Lim, Eugene, et al.
Veröffentlicht: (2024)
Stochastic Gradient Succeeds for Bandits
von: Mei, Jincheng, et al.
Veröffentlicht: (2024)
von: Mei, Jincheng, et al.
Veröffentlicht: (2024)
FreshGNN: Reducing Memory Access via Stable Historical Embeddings for Graph Neural Network Training
von: Huang, Kezhao, et al.
Veröffentlicht: (2023)
von: Huang, Kezhao, et al.
Veröffentlicht: (2023)
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Nearly Tight Bounds for Cross-Learning Contextual Bandits with Graphical Feedback
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2025) -
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024) -
Load--Reserve Wasserstein Propagation for Isotropic Diffusion Samplers
von: Lyu, Zicheng, et al.
Veröffentlicht: (2026) -
Batched Stochastic Bandit for Nondegenerate Functions
von: Liu, Yu, et al.
Veröffentlicht: (2024) -
Sublinear Spectral Clustering Oracle with Little Memory
von: Shen, Ranran, et al.
Veröffentlicht: (2026)