Probe-then-Commit Multi-Objective Bandits: Theoretical Benefits of Limited Multi-Arm Feedback
Fuente:
arXiv
Guardado en:
| Autor principal: | Shi, Ming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-thresholding Good Arm Identification with Bandit Feedback
por: Jiang, Xuanke, et al.
Publicado: (2025)
por: Jiang, Xuanke, et al.
Publicado: (2025)
Communication-Corruption Coupling and Verification in Cooperative Multi-Objective Bandits
por: Shi, Ming
Publicado: (2026)
por: Shi, Ming
Publicado: (2026)
Provably Efficient Multi-Objective Bandit Algorithms under Preference-Centric Customization
por: Cao, Linfeng, et al.
Publicado: (2025)
por: Cao, Linfeng, et al.
Publicado: (2025)
Does Feedback Help in Bandits with Arm Erasures?
por: Karakas, Merve, et al.
Publicado: (2025)
por: Karakas, Merve, et al.
Publicado: (2025)
Maximal Objectives in the Multi-armed Bandit with Applications
por: Ozbay, Eren, et al.
Publicado: (2020)
por: Ozbay, Eren, et al.
Publicado: (2020)
Best Group Identification in Multi-Objective Bandits
por: Shahverdikondori, Mohammad, et al.
Publicado: (2025)
por: Shahverdikondori, Mohammad, et al.
Publicado: (2025)
Networked Restless Multi-Arm Bandits with Reinforcement Learning
por: Zhang, Hanmo, et al.
Publicado: (2025)
por: Zhang, Hanmo, et al.
Publicado: (2025)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
por: Park, Somangchan, et al.
Publicado: (2025)
por: Park, Somangchan, et al.
Publicado: (2025)
MultiScale Contextual Bandits for Long Term Objectives
por: Rastogi, Richa, et al.
Publicado: (2025)
por: Rastogi, Richa, et al.
Publicado: (2025)
Meet Me at the Arm: The Cooperative Multi-Armed Bandits Problem with Shareable Arms
por: Hu, Xinyi, et al.
Publicado: (2025)
por: Hu, Xinyi, et al.
Publicado: (2025)
Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization
por: Wang, John, et al.
Publicado: (2026)
por: Wang, John, et al.
Publicado: (2026)
Blessings of Multiple Good Arms in Multi-Objective Linear Bandits
por: Ann, Heesang, et al.
Publicado: (2026)
por: Ann, Heesang, et al.
Publicado: (2026)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
por: Davoodi, Mansoor, et al.
Publicado: (2025)
por: Davoodi, Mansoor, et al.
Publicado: (2025)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
por: Xu, Tianyi, et al.
Publicado: (2025)
por: Xu, Tianyi, et al.
Publicado: (2025)
Learning with Limited Shared Information in Multi-agent Multi-armed Bandit
por: Shao, Junning, et al.
Publicado: (2025)
por: Shao, Junning, et al.
Publicado: (2025)
Adversarial Bandits with Multi-User Delayed Feedback: Theory and Application
por: Li, Yandi, et al.
Publicado: (2023)
por: Li, Yandi, et al.
Publicado: (2023)
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
por: Li, Shaoang, et al.
Publicado: (2025)
por: Li, Shaoang, et al.
Publicado: (2025)
Combinatorial Multi-armed Bandits: Arm Selection via Group Testing
por: Mukherjee, Arpan, et al.
Publicado: (2024)
por: Mukherjee, Arpan, et al.
Publicado: (2024)
Stochastic Multi-Armed Bandits with Limited Control Variates
por: Verma, Arun, et al.
Publicado: (2026)
por: Verma, Arun, et al.
Publicado: (2026)
Multi-Agent Best Arm Identification in Stochastic Linear Bandits
por: Agrawal, Sanjana, et al.
Publicado: (2024)
por: Agrawal, Sanjana, et al.
Publicado: (2024)
Feedback Control for Multi-Objective Graph Self-Supervision
por: Grover, Karish, et al.
Publicado: (2026)
por: Grover, Karish, et al.
Publicado: (2026)
Optimal Multi-Objective Best Arm Identification with Fixed Confidence
por: Chen, Zhirui, et al.
Publicado: (2025)
por: Chen, Zhirui, et al.
Publicado: (2025)
Quantile Multi-Armed Bandits with 1-bit Feedback
por: Lau, Ivan, et al.
Publicado: (2025)
por: Lau, Ivan, et al.
Publicado: (2025)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
por: Hou, Yunlong, et al.
Publicado: (2026)
por: Hou, Yunlong, et al.
Publicado: (2026)
Multi-Agent Combinatorial-Multi-Armed-Bandit framework for the Submodular Welfare Problem under Bandit Feedback
por: Pokhriyal, Subham, et al.
Publicado: (2026)
por: Pokhriyal, Subham, et al.
Publicado: (2026)
Theoretical Study of Conflict-Avoidant Multi-Objective Reinforcement Learning
por: Wang, Yudan, et al.
Publicado: (2024)
por: Wang, Yudan, et al.
Publicado: (2024)
A Hybrid Meta-Learning and Multi-Armed Bandit Approach for Context-Specific Multi-Objective Recommendation Optimization
por: Cunha, Tiago, et al.
Publicado: (2024)
por: Cunha, Tiago, et al.
Publicado: (2024)
Combinatorial Allocation Bandits with Nonlinear Arm Utility
por: Shibukawa, Yuki, et al.
Publicado: (2026)
por: Shibukawa, Yuki, et al.
Publicado: (2026)
EVaR-Optimal Arm Identification in Bandits
por: Ahmadipour, Mehrasa, et al.
Publicado: (2025)
por: Ahmadipour, Mehrasa, et al.
Publicado: (2025)
Lasso Bandit with Compatibility Condition on Optimal Arm
por: Lee, Harin, et al.
Publicado: (2024)
por: Lee, Harin, et al.
Publicado: (2024)
Constrained Best Arm Identification in Grouped Bandits
por: Dharod, Sahil, et al.
Publicado: (2024)
por: Dharod, Sahil, et al.
Publicado: (2024)
Best Arm Identification for Stochastic Rising Bandits
por: Mussi, Marco, et al.
Publicado: (2023)
por: Mussi, Marco, et al.
Publicado: (2023)
Reward Maximization for Pure Exploration: Minimax Optimal Good Arm Identification for Nonparametric Multi-Armed Bandits
por: Cho, Brian, et al.
Publicado: (2024)
por: Cho, Brian, et al.
Publicado: (2024)
Best-Arm Identification in Unimodal Bandits
por: Poiani, Riccardo, et al.
Publicado: (2024)
por: Poiani, Riccardo, et al.
Publicado: (2024)
The Best Arm Evades: Near-optimal Multi-pass Streaming Lower Bounds for Pure Exploration in Multi-armed Bandits
por: Assadi, Sepehr, et al.
Publicado: (2023)
por: Assadi, Sepehr, et al.
Publicado: (2023)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
por: Xiong, Guojun, et al.
Publicado: (2024)
por: Xiong, Guojun, et al.
Publicado: (2024)
Theoretical Benefit and Limitation of Diffusion Language Model
por: Feng, Guhao, et al.
Publicado: (2025)
por: Feng, Guhao, et al.
Publicado: (2025)
Optimal Arm Elimination Algorithms for Combinatorial Bandits
por: Wen, Yuxiao, et al.
Publicado: (2025)
por: Wen, Yuxiao, et al.
Publicado: (2025)
Nearly Optimal Best Arm Identification for Semiparametric Bandits
por: Kim, Seok-Jin
Publicado: (2026)
por: Kim, Seok-Jin
Publicado: (2026)
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
por: Mukherjee, Raunak, et al.
Publicado: (2026)
por: Mukherjee, Raunak, et al.
Publicado: (2026)
Ejemplares similares
-
Multi-thresholding Good Arm Identification with Bandit Feedback
por: Jiang, Xuanke, et al.
Publicado: (2025) -
Communication-Corruption Coupling and Verification in Cooperative Multi-Objective Bandits
por: Shi, Ming
Publicado: (2026) -
Provably Efficient Multi-Objective Bandit Algorithms under Preference-Centric Customization
por: Cao, Linfeng, et al.
Publicado: (2025) -
Does Feedback Help in Bandits with Arm Erasures?
por: Karakas, Merve, et al.
Publicado: (2025) -
Maximal Objectives in the Multi-armed Bandit with Applications
por: Ozbay, Eren, et al.
Publicado: (2020)