Pure Exploration in Asynchronous Federated Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zichen, Li, Chuanhao, Song, Chenyu, Wang, Lianghui, Gu, Quanquan, Wang, Huazheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information Acquisition
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
Design-Based Bandits Under Network Interference: Trade-Off Between Regret and Statistical Inference
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
Federated Linear Contextual Bandits with Heterogeneous Clients
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
von: Li, Zitian, et al.
Veröffentlicht: (2026)
von: Li, Zitian, et al.
Veröffentlicht: (2026)
Stealthy Adversarial Attacks on Stochastic Multi-Armed Bandits
von: Wang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Wang, Zhiwei, et al.
Veröffentlicht: (2024)
Incentivized Truthful Communication for Federated Bandits
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
Variance-Aware Feel-Good Thompson Sampling for Contextual Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
Pure Exploration in Bandits with Linear Constraints
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
The Batch Complexity of Bandit Pure Exploration
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
von: He, Jiafan, et al.
Veröffentlicht: (2025)
von: He, Jiafan, et al.
Veröffentlicht: (2025)
Near Optimal Pure Exploration in Logistic Bandits
von: Rivera, Eduardo Ochoa, et al.
Veröffentlicht: (2024)
von: Rivera, Eduardo Ochoa, et al.
Veröffentlicht: (2024)
Feel-Good Thompson Sampling for Contextual Dueling Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
Decentralized Asynchronous Multi-player Bandits
von: Fan, Jingqi, et al.
Veröffentlicht: (2025)
von: Fan, Jingqi, et al.
Veröffentlicht: (2025)
Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies
von: Lu, Muyun, et al.
Veröffentlicht: (2026)
von: Lu, Muyun, et al.
Veröffentlicht: (2026)
Nearly Optimal Algorithms for Contextual Dueling Bandits from Adversarial Feedback
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
Federated Contextual Cascading Bandits with Asynchronous Communication and Heterogeneous Users
von: Yang, Hantao, et al.
Veröffentlicht: (2024)
von: Yang, Hantao, et al.
Veröffentlicht: (2024)
Sharp Analysis for KL-Regularized Contextual Bandits and RLHF
von: Zhao, Heyang, et al.
Veröffentlicht: (2024)
von: Zhao, Heyang, et al.
Veröffentlicht: (2024)
Conversational Dueling Bandits in Generalized Linear Models
von: Yang, Shuhua, et al.
Veröffentlicht: (2024)
von: Yang, Shuhua, et al.
Veröffentlicht: (2024)
Tree Search-Based Evolutionary Bandits for Protein Sequence Optimization
von: Qiu, Jiahao, et al.
Veröffentlicht: (2024)
von: Qiu, Jiahao, et al.
Veröffentlicht: (2024)
Adversarial Attacks on Combinatorial Multi-Armed Bandits
von: Balasubramanian, Rishab, et al.
Veröffentlicht: (2023)
von: Balasubramanian, Rishab, et al.
Veröffentlicht: (2023)
Corruption-Robust Algorithms with Uncertainty Weighting for Nonlinear Contextual Bandits and Markov Decision Processes
von: Ye, Chenlu, et al.
Veröffentlicht: (2022)
von: Ye, Chenlu, et al.
Veröffentlicht: (2022)
Momentum Approximation in Asynchronous Private Federated Learning
von: Yu, Tao, et al.
Veröffentlicht: (2024)
von: Yu, Tao, et al.
Veröffentlicht: (2024)
Federated Linear Dueling Bandits
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
Pure Exploration with Feedback Graphs
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
The Best Arm Evades: Near-optimal Multi-pass Streaming Lower Bounds for Pure Exploration in Multi-armed Bandits
von: Assadi, Sepehr, et al.
Veröffentlicht: (2023)
von: Assadi, Sepehr, et al.
Veröffentlicht: (2023)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
von: Zhang, Junkai, et al.
Veröffentlicht: (2023)
von: Zhang, Junkai, et al.
Veröffentlicht: (2023)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2023)
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2023)
Rising Multi-Armed Bandits with Known Horizons
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
FCOM: A Federated Collaborative Online Monitoring Framework via Representation Learning
von: Kosolwattana, Tanapol, et al.
Veröffentlicht: (2024)
von: Kosolwattana, Tanapol, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Variance-Aware Regret Bounds for Stochastic Contextual Dueling Bandits
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Reward Maximization for Pure Exploration: Minimax Optimal Good Arm Identification for Nonparametric Multi-Armed Bandits
von: Cho, Brian, et al.
Veröffentlicht: (2024)
von: Cho, Brian, et al.
Veröffentlicht: (2024)
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
von: Zhao, Qingyue, et al.
Veröffentlicht: (2026)
von: Zhao, Qingyue, et al.
Veröffentlicht: (2026)
Towards a Sharp Analysis of Offline Policy Learning for $f$-Divergence-Regularized Contextual Bandits
von: Zhao, Qingyue, et al.
Veröffentlicht: (2025)
von: Zhao, Qingyue, et al.
Veröffentlicht: (2025)
Pure Exploration with Infinite Answers
von: Poiani, Riccardo, et al.
Veröffentlicht: (2025)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2025)
Preference-based Pure Exploration
von: Shukla, Apurv, et al.
Veröffentlicht: (2024)
von: Shukla, Apurv, et al.
Veröffentlicht: (2024)
Infrequent Exploration in Linear Bandits
von: Lee, Harin, et al.
Veröffentlicht: (2025)
von: Lee, Harin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information Acquisition
von: Wang, Zichen, et al.
Veröffentlicht: (2025) -
Design-Based Bandits Under Network Interference: Trade-Off Between Regret and Statistical Inference
von: Wang, Zichen, et al.
Veröffentlicht: (2025) -
Federated Linear Contextual Bandits with Heterogeneous Clients
von: Blaser, Ethan, et al.
Veröffentlicht: (2024) -
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
von: Li, Zitian, et al.
Veröffentlicht: (2026) -
Stealthy Adversarial Attacks on Stochastic Multi-Armed Bandits
von: Wang, Zhiwei, et al.
Veröffentlicht: (2024)