Graph-Dependent Regret Bounds in Multi-Armed Bandits with Interference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jamshidi, Fateme, Shahverdikondori, Mohammad, Kiyavash, Negar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Active Context Selection Improves Simple Regret in Contextual Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
Graph Learning Is Suboptimal in Causal Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Best Group Identification in Multi-Objective Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Confounded Budgeted Causal Bandits
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2024)
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2024)
Sample Complexity of Nonparametric Closeness Testing for Continuous Distributions and Its Application to Causal Discovery with Hidden Confounding
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025)
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025)
QWO: Speeding Up Permutation-Based Causal Discovery in LiGAMs
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2024)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2024)
Pure Exploration Beyond Reward Feedback: The Role of Post-Action Context
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Multi-Armed Bandits with Interference
von: Jia, Su, et al.
Veröffentlicht: (2024)
von: Jia, Su, et al.
Veröffentlicht: (2024)
Multi-armed Bandits with Missing Outcome
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2024)
Asymptotically and Minimax Optimal Regret Bounds for Multi-Armed Bandits with Abstention
von: Yang, Junwen, et al.
Veröffentlicht: (2024)
von: Yang, Junwen, et al.
Veröffentlicht: (2024)
Multi-Armed Bandits with Network Interference
von: Agarwal, Abhineet, et al.
Veröffentlicht: (2024)
von: Agarwal, Abhineet, et al.
Veröffentlicht: (2024)
Individual Regret in Cooperative Stochastic Multi-Armed Bandits
von: Barnea, Idan, et al.
Veröffentlicht: (2024)
von: Barnea, Idan, et al.
Veröffentlicht: (2024)
Collaborative Min-Max Regret in Grouped Multi-Armed Bandits
von: Blanchard, Moïse, et al.
Veröffentlicht: (2025)
von: Blanchard, Moïse, et al.
Veröffentlicht: (2025)
Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning
von: Lee, Harin, et al.
Veröffentlicht: (2026)
von: Lee, Harin, et al.
Veröffentlicht: (2026)
Neighborhood-Aware Graph Labeling Problem
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
von: He, Jiafan, et al.
Veröffentlicht: (2025)
von: He, Jiafan, et al.
Veröffentlicht: (2025)
Learning Peer Influence Probabilities with Linear Contextual Bandits
von: Faruk, Ahmed Sayeed, et al.
Veröffentlicht: (2025)
von: Faruk, Ahmed Sayeed, et al.
Veröffentlicht: (2025)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Understanding Memory-Regret Trade-Off for Streaming Stochastic Multi-Armed Bandits
von: He, Yuchen, et al.
Veröffentlicht: (2024)
von: He, Yuchen, et al.
Veröffentlicht: (2024)
Imprecise Multi-Armed Bandits
von: Kosoy, Vanessa
Veröffentlicht: (2024)
von: Kosoy, Vanessa
Veröffentlicht: (2024)
Multi-Domain Causal Discovery in Bijective Causal Models
von: Jalaldoust, Kasra, et al.
Veröffentlicht: (2025)
von: Jalaldoust, Kasra, et al.
Veröffentlicht: (2025)
s-ID: Causal Effect Identification in a Sub-Population
von: Abouei, Amir Mohammad, et al.
Veröffentlicht: (2023)
von: Abouei, Amir Mohammad, et al.
Veröffentlicht: (2023)
Open Problem: Tight Bounds for Kernelized Multi-Armed Bandits with Bernoulli Rewards
von: Mussi, Marco, et al.
Veröffentlicht: (2024)
von: Mussi, Marco, et al.
Veröffentlicht: (2024)
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
Unlearning Offline Stochastic Multi-Armed Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2026)
von: Ye, Zichun, et al.
Veröffentlicht: (2026)
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
Bayesian Regret Minimization in Offline Bandits
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
Queue Length Regret Bounds for Contextual Queueing Bandits
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
Information Capacity Regret Bounds for Bandits with Mediator Feedback
von: Eldowa, Khaled, et al.
Veröffentlicht: (2024)
von: Eldowa, Khaled, et al.
Veröffentlicht: (2024)
Batch Ensemble for Variance Dependent Regret in Stochastic Bandits
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
On the Regularity and Fairness of Combinatorial Multi-Armed Bandit
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
Rising Multi-Armed Bandits with Known Horizons
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
von: Song, Seockbean, et al.
Veröffentlicht: (2026)
Collaborating in Multi-Armed Bandits with Strategic Agents
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
Optimal Streaming Algorithms for Multi-Armed Bandits
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Active Context Selection Improves Simple Regret in Contextual Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026) -
Graph Learning Is Suboptimal in Causal Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025) -
Best Group Identification in Multi-Objective Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025) -
Confounded Budgeted Causal Bandits
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2024) -
Sample Complexity of Nonparametric Closeness Testing for Continuous Distributions and Its Application to Causal Discovery with Hidden Confounding
von: Jamshidi, Fateme, et al.
Veröffentlicht: (2025)