Gespeichert in:
| Hauptverfasser: | Feng, Qing, Ma, Tianyi, Zhu, Ruihao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2406.06802 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Peril of (Even a Little) Nonstationarity in Satisficing Regret Minimization
von: Zhang, Yixuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2026)
$(ε, u)$-Adaptive Regret Minimization in Heavy-Tailed Bandits
von: Genalti, Gianmarco, et al.
Veröffentlicht: (2023)
von: Genalti, Gianmarco, et al.
Veröffentlicht: (2023)
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
Neural Risk-sensitive Satisficing in Contextual Bandits
von: Ito, Shogo, et al.
Veröffentlicht: (2025)
von: Ito, Shogo, et al.
Veröffentlicht: (2025)
Identifying All ε-Best Arms in (Misspecified) Linear Bandits
von: Li, Zhekai, et al.
Veröffentlicht: (2025)
von: Li, Zhekai, et al.
Veröffentlicht: (2025)
Catoni-Style Change Point Detection for Regret Minimization in Non-Stationary Heavy-Tailed Bandits
von: Genalti, Gianmarco, et al.
Veröffentlicht: (2025)
von: Genalti, Gianmarco, et al.
Veröffentlicht: (2025)
Bayesian Regret Minimization in Offline Bandits
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits
von: Özyıldırım, Emre, et al.
Veröffentlicht: (2026)
von: Özyıldırım, Emre, et al.
Veröffentlicht: (2026)
Efficient Swap Regret Minimization in Combinatorial Bandits
von: Kontogiannis, Andreas, et al.
Veröffentlicht: (2026)
von: Kontogiannis, Andreas, et al.
Veröffentlicht: (2026)
No-Regret is not enough! Bandits with General Constraints through Adaptive Regret Minimization
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
Bandits with Stochastic Experts: Constant Regret, Empirical Experts and Episodes
von: Sharma, Nihal, et al.
Veröffentlicht: (2021)
von: Sharma, Nihal, et al.
Veröffentlicht: (2021)
Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards
von: Panda, Subhodip, et al.
Veröffentlicht: (2026)
von: Panda, Subhodip, et al.
Veröffentlicht: (2026)
Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
Efficient and Interpretable Bandit Algorithms
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
Robust Satisficing Gaussian Process Bandits Under Adversarial Attacks
von: Saday, Artun, et al.
Veröffentlicht: (2025)
von: Saday, Artun, et al.
Veröffentlicht: (2025)
Data-Driven Upper Confidence Bounds with Near-Optimal Regret for Heavy-Tailed Bandits
von: Tamás, Ambrus, et al.
Veröffentlicht: (2024)
von: Tamás, Ambrus, et al.
Veröffentlicht: (2024)
Risk-Aware Linear Bandits: Theory and Applications in Smart Order Routing
von: Ji, Jingwei, et al.
Veröffentlicht: (2022)
von: Ji, Jingwei, et al.
Veröffentlicht: (2022)
Tail Distribution of Regret in Optimistic Reinforcement Learning
von: Khodadadian, Sajad, et al.
Veröffentlicht: (2025)
von: Khodadadian, Sajad, et al.
Veröffentlicht: (2025)
Order-Optimal Regret in Distributed Kernel Bandits using Uniform Sampling with Shared Randomness
von: Pavlovic, Nikola, et al.
Veröffentlicht: (2024)
von: Pavlovic, Nikola, et al.
Veröffentlicht: (2024)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
Chained Information-Theoretic bounds and Tight Regret Rate for Linear Bandit Problems
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
Multi-Armed Bandits With Machine Learning-Generated Surrogate Rewards
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning
von: Lee, Harin, et al.
Veröffentlicht: (2026)
von: Lee, Harin, et al.
Veröffentlicht: (2026)
Optimal Regret for Single Index Bandits
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback
von: Ba, Wenjia, et al.
Veröffentlicht: (2021)
von: Ba, Wenjia, et al.
Veröffentlicht: (2021)
Beyond the Lower Bound: Bridging Regret Minimization and Best Arm Identification in Lexicographic Bandits
von: Xue, Bo, et al.
Veröffentlicht: (2025)
von: Xue, Bo, et al.
Veröffentlicht: (2025)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
Statistical Properties of Robust Satisficing
von: Li, Zhiyi, et al.
Veröffentlicht: (2024)
von: Li, Zhiyi, et al.
Veröffentlicht: (2024)
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
Fast Best-in-Class Regret for Contextual Bandits
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
Optimal Regret for Policy Optimization in Contextual Bandits
von: Levy, Orin, et al.
Veröffentlicht: (2026)
von: Levy, Orin, et al.
Veröffentlicht: (2026)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
A Simple and Optimal Policy Design with Safety against Heavy-Tailed Risk for Stochastic Bandits
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
On the KL-Divergence-based Robust Satisficing Model
von: Yan, Haojie, et al.
Veröffentlicht: (2024)
von: Yan, Haojie, et al.
Veröffentlicht: (2024)
Satisficing Exploration for Deep Reinforcement Learning
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
Imitation Learning via Focused Satisficing
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
Distributed Linear Bandits under Communication Constraints
von: Salgia, Sudeep, et al.
Veröffentlicht: (2022)
von: Salgia, Sudeep, et al.
Veröffentlicht: (2022)
Distributed No-Regret Learning for Multi-Stage Systems with End-to-End Bandit Feedback
von: Hou, I-Hong
Veröffentlicht: (2024)
von: Hou, I-Hong
Veröffentlicht: (2024)
Ähnliche Einträge
-
On the Peril of (Even a Little) Nonstationarity in Satisficing Regret Minimization
von: Zhang, Yixuan, et al.
Veröffentlicht: (2026) -
$(ε, u)$-Adaptive Regret Minimization in Heavy-Tailed Bandits
von: Genalti, Gianmarco, et al.
Veröffentlicht: (2023) -
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023) -
Neural Risk-sensitive Satisficing in Contextual Bandits
von: Ito, Shogo, et al.
Veröffentlicht: (2025) -
Identifying All ε-Best Arms in (Misspecified) Linear Bandits
von: Li, Zhekai, et al.
Veröffentlicht: (2025)