UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Qiyang, Khamaru, Koulik, Zhang, Cun-Hui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Thompson sampling: Precise arm-pull dynamics and adaptive inference
by: Han, Qiyang
Published: (2026)
by: Han, Qiyang
Published: (2026)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
by: Praharaj, Samya, et al.
Published: (2025)
by: Praharaj, Samya, et al.
Published: (2025)
Avoiding the Price of Adaptivity: Inference in Linear Contextual Bandits via Stability
by: Praharaj, Samya, et al.
Published: (2025)
by: Praharaj, Samya, et al.
Published: (2025)
Inference with the Upper Confidence Bound Algorithm
by: Khamaru, Koulik, et al.
Published: (2024)
by: Khamaru, Koulik, et al.
Published: (2024)
Semi-parametric inference based on adaptively collected data
by: Lin, Licong, et al.
Published: (2023)
by: Lin, Licong, et al.
Published: (2023)
Gradient descent inference in empirical risk minimization
by: Han, Qiyang, et al.
Published: (2024)
by: Han, Qiyang, et al.
Published: (2024)
Minimax-optimal trust-aware multi-armed bandits
by: Cai, Changxiao, et al.
Published: (2024)
by: Cai, Changxiao, et al.
Published: (2024)
Design Stability in Adaptive Experiments: Implications for Treatment Effect Estimation
by: Sengupta, Saikat, et al.
Published: (2025)
by: Sengupta, Saikat, et al.
Published: (2025)
Long-time dynamics and universality of nonconvex gradient descent
by: Han, Qiyang
Published: (2025)
by: Han, Qiyang
Published: (2025)
Optimal training-conditional regret for online conformal prediction
by: Liang, Jiadong, et al.
Published: (2026)
by: Liang, Jiadong, et al.
Published: (2026)
Precise Asymptotics and Refined Regret of Variance-Aware UCB
by: Fan, Yingying, et al.
Published: (2024)
by: Fan, Yingying, et al.
Published: (2024)
Extended UCB Policies for Multi-armed Bandit Problems
by: Liu, Keqin, et al.
Published: (2011)
by: Liu, Keqin, et al.
Published: (2011)
Efficient Inference after Directionally Stable Adaptive Experiments
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Anytime-valid off-policy inference for contextual bandits
by: Waudby-Smith, Ian, et al.
Published: (2022)
by: Waudby-Smith, Ian, et al.
Published: (2022)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
by: Han, Qiyang, et al.
Published: (2025)
by: Han, Qiyang, et al.
Published: (2025)
Precise Asymptotics for Spectral Methods in Mixed Generalized Linear Models
by: Zhang, Yihan, et al.
Published: (2022)
by: Zhang, Yihan, et al.
Published: (2022)
Entrywise dynamics and universality of general first order methods
by: Han, Qiyang
Published: (2024)
by: Han, Qiyang
Published: (2024)
Theoretical limits of descending $\ell_0$ sparse-regression ML algorithms
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Spectral Estimators for Multi-Index Models: Precise Asymptotics and Optimal Weak Recovery
by: Kovačević, Filip, et al.
Published: (2025)
by: Kovačević, Filip, et al.
Published: (2025)
Precise analysis of ridge interpolators under heavy correlations -- a Random Duality Theory view
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
The distribution of Ridgeless least squares interpolators
by: Han, Qiyang, et al.
Published: (2023)
by: Han, Qiyang, et al.
Published: (2023)
Truncated LinUCB for Stochastic Linear Bandits
by: Song, Yanglei, et al.
Published: (2022)
by: Song, Yanglei, et al.
Published: (2022)
Gaussian random projections of convex cones: approximate kinematic formulae and applications
by: Han, Qiyang, et al.
Published: (2022)
by: Han, Qiyang, et al.
Published: (2022)
Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits
by: Rajaraman, Nived, et al.
Published: (2023)
by: Rajaraman, Nived, et al.
Published: (2023)
Assouad, Fano, and Le Cam with Interaction: A Unifying Lower Bound Framework and Characterization for Bandit Learnability
by: Chen, Fan, et al.
Published: (2024)
by: Chen, Fan, et al.
Published: (2024)
Stochastic contextual bandits with graph feedback: from independence number to MAS number
by: Wen, Yuxiao, et al.
Published: (2024)
by: Wen, Yuxiao, et al.
Published: (2024)
A conversion theorem and minimax optimality for continuum contextual bandits
by: Akhavan, Arya, et al.
Published: (2024)
by: Akhavan, Arya, et al.
Published: (2024)
Statistical Inference under Adaptive Sampling with LinUCB
by: Fan, Wei, et al.
Published: (2025)
by: Fan, Wei, et al.
Published: (2025)
Fundamental limits of community detection from multi-view data: multi-layer, dynamic and partially labeled block models
by: Yang, Xiaodong, et al.
Published: (2024)
by: Yang, Xiaodong, et al.
Published: (2024)
Minimax Optimality of Score-based Diffusion Models: Beyond the Density Lower Bound Assumptions
by: Zhang, Kaihong, et al.
Published: (2024)
by: Zhang, Kaihong, et al.
Published: (2024)
Federated PCA and Estimation for Spiked Covariance Matrices: Optimal Rates and Efficient Algorithm
by: Li, Jingyang, et al.
Published: (2024)
by: Li, Jingyang, et al.
Published: (2024)
A leave-one-out approach to approximate message passing
by: Bao, Zhigang, et al.
Published: (2023)
by: Bao, Zhigang, et al.
Published: (2023)
On the Nonasymptotic Scaling Guarantee of Hyperparameter Estimation in Inhomogeneous, Weakly-Dependent Complex Network Dynamical Systems
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
Online selective conformal inference: adaptive scores, convergence rate and optimality
by: Humbert, Pierre, et al.
Published: (2025)
by: Humbert, Pierre, et al.
Published: (2025)
Information-Geometric Decomposition of Generalization Error in Unsupervised Learning
by: Kim, Gilhan
Published: (2026)
by: Kim, Gilhan
Published: (2026)
Sharp One-Dimensional Sub-Gaussian Comparison in Convex Order
by: Zhang, Yihan
Published: (2026)
by: Zhang, Yihan
Published: (2026)
Fine-Grained Uncertainty Quantification via Collisions
by: Friedbaum, Jesse, et al.
Published: (2024)
by: Friedbaum, Jesse, et al.
Published: (2024)
Efficient Unbiased Sparsification
by: Barnes, Leighton, et al.
Published: (2024)
by: Barnes, Leighton, et al.
Published: (2024)
Top-$K$ ranking with a monotone adversary
by: Yang, Yuepeng, et al.
Published: (2024)
by: Yang, Yuepeng, et al.
Published: (2024)
Characterizing Dependence of Samples along the Langevin Dynamics and Algorithms via Contraction of $Φ$-Mutual Information
by: Liang, Jiaming, et al.
Published: (2024)
by: Liang, Jiaming, et al.
Published: (2024)
Similar Items
-
Thompson sampling: Precise arm-pull dynamics and adaptive inference
by: Han, Qiyang
Published: (2026) -
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
by: Praharaj, Samya, et al.
Published: (2025) -
Avoiding the Price of Adaptivity: Inference in Linear Contextual Bandits via Stability
by: Praharaj, Samya, et al.
Published: (2025) -
Inference with the Upper Confidence Bound Algorithm
by: Khamaru, Koulik, et al.
Published: (2024) -
Semi-parametric inference based on adaptively collected data
by: Lin, Licong, et al.
Published: (2023)