A Simple and Optimal Policy Design with Safety against Heavy-Tailed Risk for Stochastic Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Simchi-Levi, David, Zheng, Zeyu, Zhu, Feng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
Sobolev Norm Learning Rates for Conditional Mean Embeddings
von: Talwai, Prem, et al.
Veröffentlicht: (2021)
von: Talwai, Prem, et al.
Veröffentlicht: (2021)
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
von: Lattimore, Tor
Veröffentlicht: (2026)
von: Lattimore, Tor
Veröffentlicht: (2026)
Asymptotic Classification Error for Heavy-Tailed Renewal Processes
von: Rong, Xinhui, et al.
Veröffentlicht: (2024)
von: Rong, Xinhui, et al.
Veröffentlicht: (2024)
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
Beyond Covariance Matrix: The Statistical Complexity of Private Linear Regression
von: Chen, Fan, et al.
Veröffentlicht: (2025)
von: Chen, Fan, et al.
Veröffentlicht: (2025)
Truncated LinUCB for Stochastic Linear Bandits
von: Song, Yanglei, et al.
Veröffentlicht: (2022)
von: Song, Yanglei, et al.
Veröffentlicht: (2022)
Extended UCB Policies for Multi-armed Bandit Problems
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
A Separation in Heavy-Tailed Sampling: Gaussian vs. Stable Oracles for Proximal Samplers
von: He, Ye, et al.
Veröffentlicht: (2024)
von: He, Ye, et al.
Veröffentlicht: (2024)
Breaking the Heavy-Tailed Noise Barrier in Stochastic Optimization Problems
von: Puchkin, Nikita, et al.
Veröffentlicht: (2023)
von: Puchkin, Nikita, et al.
Veröffentlicht: (2023)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
Privacy of SGD under Gaussian or Heavy-Tailed Noise: Guarantees without Gradient Clipping
von: Şimşekli, Umut, et al.
Veröffentlicht: (2024)
von: Şimşekli, Umut, et al.
Veröffentlicht: (2024)
Diffusion Models with Heavy-Tailed Targets: Score Estimation and Sampling Guarantees
von: Yu, Yifeng, et al.
Veröffentlicht: (2026)
von: Yu, Yifeng, et al.
Veröffentlicht: (2026)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
Sparsified-Learning for High-Dimensional Heavy-Tailed Locally Stationary Time Series, Concentration and Oracle Inequalities
von: Wang, Yingjie, et al.
Veröffentlicht: (2025)
von: Wang, Yingjie, et al.
Veröffentlicht: (2025)
Nonparametric Regression in Dirichlet Spaces: A Random Obstacle Approach
von: Talwai, Prem, et al.
Veröffentlicht: (2024)
von: Talwai, Prem, et al.
Veröffentlicht: (2024)
Towards Efficient and Optimal Covariance-Adaptive Algorithms for Combinatorial Semi-Bandits
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
Design Experiments to Compare Multi-armed Bandit Algorithms
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
From Spikes to Heavy Tails: Unveiling the Spectral Evolution of Neural Networks
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2024)
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2024)
Estimation of Stochastic Optimal Transport Maps
von: Nietert, Sloan, et al.
Veröffentlicht: (2025)
von: Nietert, Sloan, et al.
Veröffentlicht: (2025)
Navigating Sparsities in High-Dimensional Linear Contextual Bandits
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
Blind Network Revenue Management and Bandits with Knapsacks under Limited Switches
von: Simchi-Levi, David, et al.
Veröffentlicht: (2019)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2019)
Order-Optimal Sequential 1-Bit Mean Estimation in General Tail Regimes
von: Lau, Ivan, et al.
Veröffentlicht: (2026)
von: Lau, Ivan, et al.
Veröffentlicht: (2026)
Multi-Armed Bandits With Machine Learning-Generated Surrogate Rewards
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
On the Optimal Regret of Locally Private Linear Contextual Bandit
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
von: Azize, Achraf, et al.
Veröffentlicht: (2025)
von: Azize, Achraf, et al.
Veröffentlicht: (2025)
Risk-Controlled Post-Processing of Decision Policies
von: Joshi, Sunay, et al.
Veröffentlicht: (2026)
von: Joshi, Sunay, et al.
Veröffentlicht: (2026)
The Fragility of Optimized Bandit Algorithms
von: Fan, Lin, et al.
Veröffentlicht: (2021)
von: Fan, Lin, et al.
Veröffentlicht: (2021)
Batched Nonparametric Contextual Bandits
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
Multivariate Stochastic Dominance via Optimal Transport and Applications to Models Benchmarking
von: Rioux, Gabriel, et al.
Veröffentlicht: (2024)
von: Rioux, Gabriel, et al.
Veröffentlicht: (2024)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Optimal Excess Risk Bounds for Empirical Risk Minimization on $p$-Norm Linear Regression
von: Hanchi, Ayoub El, et al.
Veröffentlicht: (2023)
von: Hanchi, Ayoub El, et al.
Veröffentlicht: (2023)
Stochastic Optimization in Semi-Discrete Optimal Transport: Convergence Analysis and Minimax Rate
von: Genans, Ferdinand, et al.
Veröffentlicht: (2025)
von: Genans, Ferdinand, et al.
Veröffentlicht: (2025)
Adaptive Smooth Non-Stationary Bandits
von: Suk, Joe
Veröffentlicht: (2024)
von: Suk, Joe
Veröffentlicht: (2024)
Transfer Learning for Contextual Multi-armed Bandits
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023) -
Sobolev Norm Learning Rates for Conditional Mean Embeddings
von: Talwai, Prem, et al.
Veröffentlicht: (2021) -
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
von: Liu, Jingyu, et al.
Veröffentlicht: (2025) -
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
von: Prevost, Adrien, et al.
Veröffentlicht: (2025) -
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
von: Lattimore, Tor
Veröffentlicht: (2026)