Continuum-armed Bandit Optimization with Batch Pairwise Comparison Oracles
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Xiangyu, Chen, Xi, Wang, Yining, Zeng, Zhiyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust Batched Bandits
von: Guo, Yunwen, et al.
Veröffentlicht: (2025)
von: Guo, Yunwen, et al.
Veröffentlicht: (2025)
Offline Learning for Combinatorial Multi-armed Bandits
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
Batched Stochastic Bandit for Nondegenerate Functions
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
von: Chen, Gongpu, et al.
Veröffentlicht: (2024)
von: Chen, Gongpu, et al.
Veröffentlicht: (2024)
Oracle-Efficient Combinatorial Semi-Bandits
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
Multi-armed Bandits with Missing Outcome
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
Online Statistical Inference for Contextual Bandits via Stochastic Gradient Descent
von: Chang, Xiangyu, et al.
Veröffentlicht: (2022)
von: Chang, Xiangyu, et al.
Veröffentlicht: (2022)
Batched Kernelized Bandits: Refinements and Extensions
von: Ma, Chenkai, et al.
Veröffentlicht: (2026)
von: Ma, Chenkai, et al.
Veröffentlicht: (2026)
ComPO: Preference Alignment via Comparison Oracles
von: Chen, Peter, et al.
Veröffentlicht: (2025)
von: Chen, Peter, et al.
Veröffentlicht: (2025)
The Batch Complexity of Bandit Pure Exploration
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
von: Tuynman, Adrienne, et al.
Veröffentlicht: (2025)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Replicability is Asymptotically Free in Multi-armed Bandits
von: Komiyama, Junpei, et al.
Veröffentlicht: (2024)
von: Komiyama, Junpei, et al.
Veröffentlicht: (2024)
Maximal Objectives in the Multi-armed Bandit with Applications
von: Ozbay, Eren, et al.
Veröffentlicht: (2020)
von: Ozbay, Eren, et al.
Veröffentlicht: (2020)
Batched Nonparametric Contextual Bandits
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
Hybrid Combinatorial Multi-armed Bandits with Probabilistically Triggered Arms
von: Zhou, Kongchang, et al.
Veröffentlicht: (2025)
von: Zhou, Kongchang, et al.
Veröffentlicht: (2025)
Deceptive Exploration in Multi-armed Bandits
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
Causally Abstracted Multi-armed Bandits
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
Design Experiments to Compare Multi-armed Bandit Algorithms
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
Optimal and Practical Batched Linear Bandit Algorithm
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
Pairwise Elimination with Instance-Dependent Guarantees for Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2025)
von: Juneja, Ishank, et al.
Veröffentlicht: (2025)
Multi-agent Multi-armed Bandit with Fully Heavy-tailed Dynamics
von: Wang, Xingyu, et al.
Veröffentlicht: (2025)
von: Wang, Xingyu, et al.
Veröffentlicht: (2025)
Feedback Descent: Open-Ended Text Optimization via Pairwise Comparison
von: Lee, Yoonho, et al.
Veröffentlicht: (2025)
von: Lee, Yoonho, et al.
Veröffentlicht: (2025)
Fairness of Exposure in Online Restless Multi-armed Bandits
von: Sood, Archit, et al.
Veröffentlicht: (2024)
von: Sood, Archit, et al.
Veröffentlicht: (2024)
Federated $\mathcal{X}$-armed Bandit with Flexible Personalisation
von: Arabzadeh, Ali, et al.
Veröffentlicht: (2024)
von: Arabzadeh, Ali, et al.
Veröffentlicht: (2024)
Learning with Limited Shared Information in Multi-agent Multi-armed Bandit
von: Shao, Junning, et al.
Veröffentlicht: (2025)
von: Shao, Junning, et al.
Veröffentlicht: (2025)
Efficient and Optimal Policy Gradient Algorithm for Corrupted Multi-armed Bandits
von: Liu, Jiayuan, et al.
Veröffentlicht: (2025)
von: Liu, Jiayuan, et al.
Veröffentlicht: (2025)
A Two-armed Bandit Framework for A/B Testing
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
Heterogeneous Multi-agent Multi-armed Bandits on Stochastic Block Models
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
Transfer Learning for Contextual Multi-armed Bandits
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
Batch Ensemble for Variance Dependent Regret in Stochastic Bandits
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
Locally Private Nonparametric Contextual Multi-armed Bandits
von: Ma, Yuheng, et al.
Veröffentlicht: (2025)
von: Ma, Yuheng, et al.
Veröffentlicht: (2025)
Metric Learning from Limited Pairwise Preference Comparisons
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
Transfer in Sequential Multi-armed Bandits via Reward Samples
von: R, Rahul N, et al.
Veröffentlicht: (2024)
von: R, Rahul N, et al.
Veröffentlicht: (2024)
Falcon: Fair Active Learning using Multi-armed Bandits
von: Tae, Ki Hyun, et al.
Veröffentlicht: (2024)
von: Tae, Ki Hyun, et al.
Veröffentlicht: (2024)
Batch size invariant Adam
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
Pairwise Comparisons without Stochastic Transitivity: Model, Theory and Applications
von: Lee, Sze Ming, et al.
Veröffentlicht: (2025)
von: Lee, Sze Ming, et al.
Veröffentlicht: (2025)
Nearly Tight Bounds for Exploration in Streaming Multi-armed Bandits with Known Optimality Gap
von: Karpov, Nikolai, et al.
Veröffentlicht: (2025)
von: Karpov, Nikolai, et al.
Veröffentlicht: (2025)
Learning from Similarity/Dissimilarity and Pairwise Comparison
von: Tate, Tomoya, et al.
Veröffentlicht: (2026)
von: Tate, Tomoya, et al.
Veröffentlicht: (2026)
Batched Online Contextual Sparse Bandits with Sequential Inclusion of Features
von: Swiers, Rowan, et al.
Veröffentlicht: (2024)
von: Swiers, Rowan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Robust Batched Bandits
von: Guo, Yunwen, et al.
Veröffentlicht: (2025) -
Offline Learning for Combinatorial Multi-armed Bandits
von: Liu, Xutong, et al.
Veröffentlicht: (2025) -
Batched Stochastic Bandit for Nondegenerate Functions
von: Liu, Yu, et al.
Veröffentlicht: (2024) -
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
von: Chen, Gongpu, et al.
Veröffentlicht: (2024) -
Oracle-Efficient Combinatorial Semi-Bandits
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)