Efficient and Optimal Policy Gradient Algorithm for Corrupted Multi-armed Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jiayuan, Wang, Siwei, Fang, Zhixuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning with Limited Shared Information in Multi-agent Multi-armed Bandit
von: Shao, Junning, et al.
Veröffentlicht: (2025)
von: Shao, Junning, et al.
Veröffentlicht: (2025)
Offline Learning for Combinatorial Multi-armed Bandits
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
von: Chen, Gongpu, et al.
Veröffentlicht: (2024)
von: Chen, Gongpu, et al.
Veröffentlicht: (2024)
Extended UCB Policies for Multi-armed Bandit Problems
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
Design Experiments to Compare Multi-armed Bandit Algorithms
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
Hybrid Combinatorial Multi-armed Bandits with Probabilistically Triggered Arms
von: Zhou, Kongchang, et al.
Veröffentlicht: (2025)
von: Zhou, Kongchang, et al.
Veröffentlicht: (2025)
Multi-armed Bandits with Missing Outcome
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
Corruption-Robust Linear Bandits: Minimax Optimality and Gap-Dependent Misspecification
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
Multi-Agent Stochastic Bandits Robust to Adversarial Corruptions
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2024)
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2024)
Heterogeneous Multi-agent Multi-armed Bandits on Stochastic Block Models
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
Optimal Control of Fluid Restless Multi-armed Bandits: A Machine Learning Approach
von: Bertsimas, Dimitris, et al.
Veröffentlicht: (2025)
von: Bertsimas, Dimitris, et al.
Veröffentlicht: (2025)
Replicability is Asymptotically Free in Multi-armed Bandits
von: Komiyama, Junpei, et al.
Veröffentlicht: (2024)
von: Komiyama, Junpei, et al.
Veröffentlicht: (2024)
Maximal Objectives in the Multi-armed Bandit with Applications
von: Ozbay, Eren, et al.
Veröffentlicht: (2020)
von: Ozbay, Eren, et al.
Veröffentlicht: (2020)
Multi-agent Multi-armed Bandit with Fully Heavy-tailed Dynamics
von: Wang, Xingyu, et al.
Veröffentlicht: (2025)
von: Wang, Xingyu, et al.
Veröffentlicht: (2025)
Nearly Tight Bounds for Exploration in Streaming Multi-armed Bandits with Known Optimality Gap
von: Karpov, Nikolai, et al.
Veröffentlicht: (2025)
von: Karpov, Nikolai, et al.
Veröffentlicht: (2025)
Deceptive Exploration in Multi-armed Bandits
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
Causally Abstracted Multi-armed Bandits
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
Optimal Streaming Algorithms for Multi-Armed Bandits
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
Decentralized Asynchronous Multi-player Bandits
von: Fan, Jingqi, et al.
Veröffentlicht: (2025)
von: Fan, Jingqi, et al.
Veröffentlicht: (2025)
Cascading Bandits Robust to Adversarial Corruptions
von: Xie, Jize, et al.
Veröffentlicht: (2025)
von: Xie, Jize, et al.
Veröffentlicht: (2025)
Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
Fairness of Exposure in Online Restless Multi-armed Bandits
von: Sood, Archit, et al.
Veröffentlicht: (2024)
von: Sood, Archit, et al.
Veröffentlicht: (2024)
Constructing Adversarial Examples for Vertical Federated Learning: Optimal Client Corruption through Multi-Armed Bandit
von: Yao, Duanyi, et al.
Veröffentlicht: (2024)
von: Yao, Duanyi, et al.
Veröffentlicht: (2024)
Transfer Learning for Contextual Multi-armed Bandits
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
Communication-Corruption Coupling and Verification in Cooperative Multi-Objective Bandits
von: Shi, Ming
Veröffentlicht: (2026)
von: Shi, Ming
Veröffentlicht: (2026)
Locally Private Nonparametric Contextual Multi-armed Bandits
von: Ma, Yuheng, et al.
Veröffentlicht: (2025)
von: Ma, Yuheng, et al.
Veröffentlicht: (2025)
Transfer in Sequential Multi-armed Bandits via Reward Samples
von: R, Rahul N, et al.
Veröffentlicht: (2024)
von: R, Rahul N, et al.
Veröffentlicht: (2024)
Falcon: Fair Active Learning using Multi-armed Bandits
von: Tae, Ki Hyun, et al.
Veröffentlicht: (2024)
von: Tae, Ki Hyun, et al.
Veröffentlicht: (2024)
Corruption-Robust Algorithms with Uncertainty Weighting for Nonlinear Contextual Bandits and Markov Decision Processes
von: Ye, Chenlu, et al.
Veröffentlicht: (2022)
von: Ye, Chenlu, et al.
Veröffentlicht: (2022)
Kullback-Leibler Maillard Sampling for Multi-armed Bandits with Bounded Rewards
von: Qin, Hao, et al.
Veröffentlicht: (2023)
von: Qin, Hao, et al.
Veröffentlicht: (2023)
Adaptive Endpointing with Deep Contextual Multi-armed Bandits
von: Min, Do June, et al.
Veröffentlicht: (2023)
von: Min, Do June, et al.
Veröffentlicht: (2023)
Rising Multi-Armed Bandits with Known Horizons
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
von: Manupriya, Piyushi, et al.
Veröffentlicht: (2025)
von: Manupriya, Piyushi, et al.
Veröffentlicht: (2025)
Decentralized Blockchain-based Robust Multi-agent Multi-armed Bandit
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
Convergence of a L2 regularized Policy Gradient Algorithm for the Multi Armed Bandit
von: Anita, Stefana, et al.
Veröffentlicht: (2024)
von: Anita, Stefana, et al.
Veröffentlicht: (2024)
Efficient and Interpretable Bandit Algorithms
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
Online Learning to Rank under Corruption: A Robust Cascading Bandits Approach
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2025)
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2025)
Optimal Regret for Policy Optimization in Contextual Bandits
von: Levy, Orin, et al.
Veröffentlicht: (2026)
von: Levy, Orin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning with Limited Shared Information in Multi-agent Multi-armed Bandit
von: Shao, Junning, et al.
Veröffentlicht: (2025) -
Offline Learning for Combinatorial Multi-armed Bandits
von: Liu, Xutong, et al.
Veröffentlicht: (2025) -
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025) -
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026) -
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
von: Chen, Gongpu, et al.
Veröffentlicht: (2024)