Batch-Size Independent Regret Bounds for Combinatorial Semi-Bandits with Probabilistically Triggered Arms or Independent Arms
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Xutong, Zuo, Jinhang, Wang, Siwei, Joe-Wong, Carlee, Lui, John C. S., Chen, Wei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
di: Liu, Xutong, et al.
Pubblicazione: (2023)
di: Liu, Xutong, et al.
Pubblicazione: (2023)
Offline Learning for Combinatorial Multi-armed Bandits
di: Liu, Xutong, et al.
Pubblicazione: (2025)
di: Liu, Xutong, et al.
Pubblicazione: (2025)
Hybrid Combinatorial Multi-armed Bandits with Probabilistically Triggered Arms
di: Zhou, Kongchang, et al.
Pubblicazione: (2025)
di: Zhou, Kongchang, et al.
Pubblicazione: (2025)
Continuous Semantic Caching for Low-Cost LLM Serving
di: Atalar, Baran, et al.
Pubblicazione: (2026)
di: Atalar, Baran, et al.
Pubblicazione: (2026)
Semantic Caching for Low-Cost LLM Serving: From Offline Learning to Online Adaptation
di: Liu, Xutong, et al.
Pubblicazione: (2025)
di: Liu, Xutong, et al.
Pubblicazione: (2025)
Combinatorial Multivariant Multi-Armed Bandits with Applications to Episodic Reinforcement Learning and Beyond
di: Liu, Xutong, et al.
Pubblicazione: (2024)
di: Liu, Xutong, et al.
Pubblicazione: (2024)
Stochastic Bandits Robust to Adversarial Attacks
di: Wang, Xuchuang, et al.
Pubblicazione: (2024)
di: Wang, Xuchuang, et al.
Pubblicazione: (2024)
Near-Optimal Regret for Efficient Stochastic Combinatorial Semi-Bandits
di: Ye, Zichun, et al.
Pubblicazione: (2025)
di: Ye, Zichun, et al.
Pubblicazione: (2025)
Neural Combinatorial Clustered Bandits for Recommendation Systems
di: Atalar, Baran, et al.
Pubblicazione: (2024)
di: Atalar, Baran, et al.
Pubblicazione: (2024)
Combinatorial Logistic Bandits
di: Liu, Xutong, et al.
Pubblicazione: (2024)
di: Liu, Xutong, et al.
Pubblicazione: (2024)
Offline Clustering of Linear Bandits: The Power of Clusters under Limited Data
di: Liu, Jingyuan, et al.
Pubblicazione: (2025)
di: Liu, Jingyuan, et al.
Pubblicazione: (2025)
Fusing Reward and Dueling Feedback in Stochastic Bandits
di: Wang, Xuchuang, et al.
Pubblicazione: (2025)
di: Wang, Xuchuang, et al.
Pubblicazione: (2025)
Online Multi-LLM Selection via Contextual Bandits under Unstructured Context Evolution
di: Poon, Manhin, et al.
Pubblicazione: (2025)
di: Poon, Manhin, et al.
Pubblicazione: (2025)
Intelligent Communication Planning for Constrained Environmental IoT Sensing with Reinforcement Learning
di: Hu, Yi, et al.
Pubblicazione: (2023)
di: Hu, Yi, et al.
Pubblicazione: (2023)
A Unified Online-Offline Framework for Co-Branding Campaign Recommendations
di: Dai, Xiangxiang, et al.
Pubblicazione: (2025)
di: Dai, Xiangxiang, et al.
Pubblicazione: (2025)
Graph Feedback Bandits with Similar Arms
di: Qi, Han, et al.
Pubblicazione: (2024)
di: Qi, Han, et al.
Pubblicazione: (2024)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
Worst-Case Regret Bounds for Combinatorial Thompson Sampling in Sleeping Semi-Bandits
di: Huang, Zhiming, et al.
Pubblicazione: (2026)
di: Huang, Zhiming, et al.
Pubblicazione: (2026)
Merit-based Fair Combinatorial Semi-Bandit with Unrestricted Feedback Delays
di: Chen, Ziqun, et al.
Pubblicazione: (2024)
di: Chen, Ziqun, et al.
Pubblicazione: (2024)
Tin-Tin: Towards Tiny Learning on Tiny Devices with Integer-based Neural Network Training
di: Hu, Yi, et al.
Pubblicazione: (2025)
di: Hu, Yi, et al.
Pubblicazione: (2025)
CoRAST: Towards Foundation Model-Powered Correlated Data Analysis in Resource-Constrained CPS and IoT
di: Hu, Yi, et al.
Pubblicazione: (2024)
di: Hu, Yi, et al.
Pubblicazione: (2024)
Best Arm Identification in Generalized Linear Bandits via Hybrid Feedback
di: Zeng, Qirun, et al.
Pubblicazione: (2026)
di: Zeng, Qirun, et al.
Pubblicazione: (2026)
DYNAMITE: Dynamic Interplay of Mini-Batch Size and Aggregation Frequency for Federated Learning with Static and Streaming Dataset
di: Liu, Weijie, et al.
Pubblicazione: (2023)
di: Liu, Weijie, et al.
Pubblicazione: (2023)
Blessings of Multiple Good Arms in Multi-Objective Linear Bandits
di: Ann, Heesang, et al.
Pubblicazione: (2026)
di: Ann, Heesang, et al.
Pubblicazione: (2026)
Identifying All ε-Best Arms in (Misspecified) Linear Bandits
di: Li, Zhekai, et al.
Pubblicazione: (2025)
di: Li, Zhekai, et al.
Pubblicazione: (2025)
Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
di: Qi, Han, et al.
Pubblicazione: (2025)
di: Qi, Han, et al.
Pubblicazione: (2025)
On-line Learning in Tree MDPs by Treating Policies as Bandit Arms
di: Shah, Anvay, et al.
Pubblicazione: (2026)
di: Shah, Anvay, et al.
Pubblicazione: (2026)
Arms and the People
Pubblicazione: (2022)
Pubblicazione: (2022)
The Royal Arms
di: artfletch
Pubblicazione: (2021)
di: artfletch
Pubblicazione: (2021)
Ladies in Arms
Pubblicazione: (2024)
Pubblicazione: (2024)
Arms Trafficking
Pubblicazione: (2023)
Pubblicazione: (2023)
Comrades in Arms
di: Smith, Tom
Pubblicazione: (2020)
di: Smith, Tom
Pubblicazione: (2020)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
di: Juneja, Ishank, et al.
Pubblicazione: (2026)
di: Juneja, Ishank, et al.
Pubblicazione: (2026)
Pairwise Elimination with Instance-Dependent Guarantees for Bandits with Cost Subsidy
di: Juneja, Ishank, et al.
Pubblicazione: (2025)
di: Juneja, Ishank, et al.
Pubblicazione: (2025)
Speed Up the Cold-Start Learning in Two-Sided Bandits with Many Arms
di: Bayati, Mohsen, et al.
Pubblicazione: (2022)
di: Bayati, Mohsen, et al.
Pubblicazione: (2022)
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
di: Atalar, Baran, et al.
Pubblicazione: (2025)
di: Atalar, Baran, et al.
Pubblicazione: (2025)
GenTwoArmsTrialSize: An R Statistical Software Package to estimate Generalized Two Arms Randomized Clinical Trial Sample Size
di: Soltanifar, Mohsen, et al.
Pubblicazione: (2024)
di: Soltanifar, Mohsen, et al.
Pubblicazione: (2024)
Combinatorial Rising Bandits
di: Song, Seockbean, et al.
Pubblicazione: (2024)
di: Song, Seockbean, et al.
Pubblicazione: (2024)
The Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandit with Many Arms
di: Bayati, Mohsen, et al.
Pubblicazione: (2020)
di: Bayati, Mohsen, et al.
Pubblicazione: (2020)
Strategic Arms with Side Communication Prevail Over Low-Regret MAB Algorithms
di: Yahmed, Ahmed Ben, et al.
Pubblicazione: (2024)
di: Yahmed, Ahmed Ben, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
di: Liu, Xutong, et al.
Pubblicazione: (2023) -
Offline Learning for Combinatorial Multi-armed Bandits
di: Liu, Xutong, et al.
Pubblicazione: (2025) -
Hybrid Combinatorial Multi-armed Bandits with Probabilistically Triggered Arms
di: Zhou, Kongchang, et al.
Pubblicazione: (2025) -
Continuous Semantic Caching for Low-Cost LLM Serving
di: Atalar, Baran, et al.
Pubblicazione: (2026) -
Semantic Caching for Low-Cost LLM Serving: From Offline Learning to Online Adaptation
di: Liu, Xutong, et al.
Pubblicazione: (2025)