A/B Testing and Best-arm Identification for Linear Bandits with Robustness to Non-stationarity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiong, Zhihan, Camilleri, Romain, Fazel, Maryam, Jain, Lalit, Jamieson, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
On the Power of Adaptivity for $\varepsilon$-Best Arm Identification in Linear Bandits
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
Nearly Minimax Optimal Submodular Maximization with Bandit Feedback
von: Tajdini, Artin, et al.
Veröffentlicht: (2023)
von: Tajdini, Artin, et al.
Veröffentlicht: (2023)
Learning to Actively Learn: A Robust Approach
von: Zhang, Jifan, et al.
Veröffentlicht: (2020)
von: Zhang, Jifan, et al.
Veröffentlicht: (2020)
Dual Approximation Policy Optimization
von: Xiong, Zhihan, et al.
Veröffentlicht: (2024)
von: Xiong, Zhihan, et al.
Veröffentlicht: (2024)
A Black-box Approach for Non-stationary Multi-agent Reinforcement Learning
von: Jiang, Haozhe, et al.
Veröffentlicht: (2023)
von: Jiang, Haozhe, et al.
Veröffentlicht: (2023)
Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
A Two-armed Bandit Framework for A/B Testing
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
Hybrid Preference Optimization for Alignment: Provably Faster Convergence Rates by Combining Offline Preferences with Online Exploration
von: Bose, Avinandan, et al.
Veröffentlicht: (2024)
von: Bose, Avinandan, et al.
Veröffentlicht: (2024)
Fairness of Exposure in Online Restless Multi-armed Bandits
von: Sood, Archit, et al.
Veröffentlicht: (2024)
von: Sood, Archit, et al.
Veröffentlicht: (2024)
Heavy-tailed Linear Bandits: Adversarial Robustness, Best-of-both-worlds, and Beyond
von: Zhao, Canzhe, et al.
Veröffentlicht: (2025)
von: Zhao, Canzhe, et al.
Veröffentlicht: (2025)
Best Arm Identification for Stochastic Rising Bandits
von: Mussi, Marco, et al.
Veröffentlicht: (2023)
von: Mussi, Marco, et al.
Veröffentlicht: (2023)
Best Group Identification in Multi-Objective Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Constrained Best Arm Identification in Grouped Bandits
von: Dharod, Sahil, et al.
Veröffentlicht: (2024)
von: Dharod, Sahil, et al.
Veröffentlicht: (2024)
Offline congestion games: How feedback type affects data coverage requirement
von: Jiang, Haozhe, et al.
Veröffentlicht: (2022)
von: Jiang, Haozhe, et al.
Veröffentlicht: (2022)
Multi-Agent Best Arm Identification in Stochastic Linear Bandits
von: Agrawal, Sanjana, et al.
Veröffentlicht: (2024)
von: Agrawal, Sanjana, et al.
Veröffentlicht: (2024)
Best-Arm Identification in Unimodal Bandits
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Online SuBmodular + SuPermodular (BP) Maximization with Bandit Feedback
von: Narang, Adhyyan, et al.
Veröffentlicht: (2022)
von: Narang, Adhyyan, et al.
Veröffentlicht: (2022)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
Almost Minimax Optimal Best Arm Identification in Piecewise Stationary Linear Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2024)
von: Hou, Yunlong, et al.
Veröffentlicht: (2024)
Optimal Posterior Sampling for Policy Identification in Tabular Markov Decision Processes
von: Kone, Cyrille, et al.
Veröffentlicht: (2026)
von: Kone, Cyrille, et al.
Veröffentlicht: (2026)
Nearly Optimal Best Arm Identification for Semiparametric Bandits
von: Kim, Seok-Jin
Veröffentlicht: (2026)
von: Kim, Seok-Jin
Veröffentlicht: (2026)
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
von: Mukherjee, Raunak, et al.
Veröffentlicht: (2026)
von: Mukherjee, Raunak, et al.
Veröffentlicht: (2026)
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
Near Optimal Best Arm Identification for Clustered Bandits
von: Yash, et al.
Veröffentlicht: (2025)
von: Yash, et al.
Veröffentlicht: (2025)
Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
von: Kim, HyunGi, et al.
Veröffentlicht: (2025)
von: Kim, HyunGi, et al.
Veröffentlicht: (2025)
Best of Three Worlds: Adaptive Experimentation for Digital Marketing in Practice
von: Fiez, Tanner, et al.
Veröffentlicht: (2024)
von: Fiez, Tanner, et al.
Veröffentlicht: (2024)
LoRe: Personalizing LLMs via Low-Rank Reward Modeling
von: Bose, Avinandan, et al.
Veröffentlicht: (2025)
von: Bose, Avinandan, et al.
Veröffentlicht: (2025)
Multi-armed Bandits with Missing Outcome
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
von: Sharma, Shradha, et al.
Veröffentlicht: (2026)
von: Sharma, Shradha, et al.
Veröffentlicht: (2026)
A Tight Lower Bound for Non-stochastic Multi-armed Bandits with Expert Advice
von: Chase, Zachary, et al.
Veröffentlicht: (2025)
von: Chase, Zachary, et al.
Veröffentlicht: (2025)
Optimal Best Arm Identification with Fixed Confidence in Restless Bandits
von: Karthik, P. N., et al.
Veröffentlicht: (2023)
von: Karthik, P. N., et al.
Veröffentlicht: (2023)
Robust Causal Bandits for Linear Models
von: Yan, Zirui, et al.
Veröffentlicht: (2023)
von: Yan, Zirui, et al.
Veröffentlicht: (2023)
Decentralized Blockchain-based Robust Multi-agent Multi-armed Bandit
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
Unregularized Linear Convergence in Zero-Sum Game from Preference Feedback
von: Chen, Shulun, et al.
Veröffentlicht: (2025)
von: Chen, Shulun, et al.
Veröffentlicht: (2025)
Sparse Linear Bandits with Blocking Constraints
von: Jain, Adit, et al.
Veröffentlicht: (2024)
von: Jain, Adit, et al.
Veröffentlicht: (2024)
Adapting to Stochastic and Adversarial Losses in Episodic MDPs with Aggregate Bandit Feedback
von: Ito, Shinji, et al.
Veröffentlicht: (2025)
von: Ito, Shinji, et al.
Veröffentlicht: (2025)
LC-Tsallis-INF: Generalized Best-of-Both-Worlds Linear Contextual Bandits
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026) -
On the Power of Adaptivity for $\varepsilon$-Best Arm Identification in Linear Bandits
von: Maiti, Arnab, et al.
Veröffentlicht: (2026) -
Nearly Minimax Optimal Submodular Maximization with Bandit Feedback
von: Tajdini, Artin, et al.
Veröffentlicht: (2023) -
Learning to Actively Learn: A Robust Approach
von: Zhang, Jifan, et al.
Veröffentlicht: (2020) -
Dual Approximation Policy Optimization
von: Xiong, Zhihan, et al.
Veröffentlicht: (2024)