Extended UCB Policies for Multi-armed Bandit Problems
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Keqin, Zheng, Tianshuo, Zhou, Zhi-Hua |
|---|---|
| Format: | Preprint |
| Published: |
2011
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Truncated LinUCB for Stochastic Linear Bandits
by: Song, Yanglei, et al.
Published: (2022)
by: Song, Yanglei, et al.
Published: (2022)
UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
by: Han, Qiyang, et al.
Published: (2024)
by: Han, Qiyang, et al.
Published: (2024)
Transfer Learning for Contextual Multi-armed Bandits
by: Cai, Changxiao, et al.
Published: (2022)
by: Cai, Changxiao, et al.
Published: (2022)
Early Stopping in Contextual Bandits and Inferences
by: Cui, Zihan
Published: (2025)
by: Cui, Zihan
Published: (2025)
Design Experiments to Compare Multi-armed Bandit Algorithms
by: Meng, Huiling, et al.
Published: (2026)
by: Meng, Huiling, et al.
Published: (2026)
Predicting path-dependent processes by deep learning
by: Zheng, Xudong, et al.
Published: (2024)
by: Zheng, Xudong, et al.
Published: (2024)
Tractability from overparametrization: The example of the negative perceptron
by: Montanari, Andrea, et al.
Published: (2021)
by: Montanari, Andrea, et al.
Published: (2021)
Nonlinear spiked covariance matrices and signal propagation in deep neural networks
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
Statistical limits of correlation detection in trees
by: Ganassali, Luca, et al.
Published: (2022)
by: Ganassali, Luca, et al.
Published: (2022)
Spectral Estimators for Multi-Index Models: Precise Asymptotics and Optimal Weak Recovery
by: Kovačević, Filip, et al.
Published: (2025)
by: Kovačević, Filip, et al.
Published: (2025)
Precise Asymptotics and Refined Regret of Variance-Aware UCB
by: Fan, Yingying, et al.
Published: (2024)
by: Fan, Yingying, et al.
Published: (2024)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
by: Prevost, Adrien, et al.
Published: (2025)
by: Prevost, Adrien, et al.
Published: (2025)
On Universality of Non-Separable Approximate Message Passing Algorithms
by: Lovig, Max, et al.
Published: (2025)
by: Lovig, Max, et al.
Published: (2025)
Restricted Spectral Gap Decomposition for Simulated Tempering Targeting Mixture Distributions
by: Garg, Jhanvi, et al.
Published: (2025)
by: Garg, Jhanvi, et al.
Published: (2025)
Unifying AMP Algorithms for Rotationally-Invariant Models
by: Liu, Songbin, et al.
Published: (2024)
by: Liu, Songbin, et al.
Published: (2024)
Optimality of Approximate Message Passing Algorithms for Spiked Matrix Models with Rotationally Invariant Noise
by: Dudeja, Rishabh, et al.
Published: (2024)
by: Dudeja, Rishabh, et al.
Published: (2024)
Learning with Expected Signatures: Theory and Applications
by: Lucchese, Lorenzo, et al.
Published: (2025)
by: Lucchese, Lorenzo, et al.
Published: (2025)
Limit theorems of Chatterjee's rank correlation
by: Lin, Zhexiao, et al.
Published: (2022)
by: Lin, Zhexiao, et al.
Published: (2022)
Nonlinear Bayesian Update via Ensemble Kernel Regression with Clustering and Subsampling
by: Lee, Yoonsang
Published: (2025)
by: Lee, Yoonsang
Published: (2025)
Diffusion Models with Heavy-Tailed Targets: Score Estimation and Sampling Guarantees
by: Yu, Yifeng, et al.
Published: (2026)
by: Yu, Yifeng, et al.
Published: (2026)
Hamiltonian Monte Carlo with Asymmetrical Momentum Distributions
by: Ghosh, Soumyadip, et al.
Published: (2021)
by: Ghosh, Soumyadip, et al.
Published: (2021)
Gradient-flow SDEs have unique transient population dynamics
by: Guan, Vincent, et al.
Published: (2025)
by: Guan, Vincent, et al.
Published: (2025)
Off-the-grid prediction and testing for linear combination of translated features
by: Butucea, Cristina, et al.
Published: (2022)
by: Butucea, Cristina, et al.
Published: (2022)
Simple Relative Deviation Bounds for Covariance and Gram Matrices
by: Barzilai, Daniel, et al.
Published: (2024)
by: Barzilai, Daniel, et al.
Published: (2024)
kTULA: A Langevin sampling algorithm with improved KL bounds under super-linear log-gradients
by: Lytras, Iosif, et al.
Published: (2025)
by: Lytras, Iosif, et al.
Published: (2025)
On Experiments
by: van Rooyen, Brendan
Published: (2025)
by: van Rooyen, Brendan
Published: (2025)
Limit Theorems for Stochastic Gradient Descent in High-Dimensional Single-Layer Networks
by: Rangriz, Parsa
Published: (2025)
by: Rangriz, Parsa
Published: (2025)
A variational approach to dimension-free self-normalized concentration
by: Chugg, Ben, et al.
Published: (2025)
by: Chugg, Ben, et al.
Published: (2025)
A Computational Transition for Detecting Multivariate Shuffled Linear Regression by Low-Degree Polynomials
by: Li, Zhangsong
Published: (2025)
by: Li, Zhangsong
Published: (2025)
Finite-Dimensional Gaussian Approximation for Deep Neural Networks: Universality in Random Weights
by: Balasubramanian, Krishnakumar, et al.
Published: (2025)
by: Balasubramanian, Krishnakumar, et al.
Published: (2025)
Convergence Bounds for Sequential Monte Carlo on Multimodal Distributions using Soft Decomposition
by: Lee, Holden, et al.
Published: (2024)
by: Lee, Holden, et al.
Published: (2024)
Simultaneous off-the-grid learning of mixtures issued from a continuous dictionary
by: Butucea, Cristina, et al.
Published: (2022)
by: Butucea, Cristina, et al.
Published: (2022)
A Fourier representation of kernel Stein discrepancy with application to Goodness-of-Fit tests for measures on infinite dimensional Hilbert spaces
by: Wynne, George, et al.
Published: (2022)
by: Wynne, George, et al.
Published: (2022)
Dimension-Free Bounds for Generalized First-Order Methods via Gaussian Coupling
by: Reeves, Galen
Published: (2025)
by: Reeves, Galen
Published: (2025)
Optimization, Isoperimetric Inequalities, and Sampling via Lyapunov Potentials
by: Chen, August Y., et al.
Published: (2024)
by: Chen, August Y., et al.
Published: (2024)
The feasibility of multi-graph alignment: a Bayesian approach
by: Vassaux, Louis, et al.
Published: (2025)
by: Vassaux, Louis, et al.
Published: (2025)
An extension of McDiarmid's inequality
by: Combes, Richard
Published: (2015)
by: Combes, Richard
Published: (2015)
Multivariate Gaussian Approximation for Random Forest via Region-based Stabilization
by: Shi, Zhaoyang, et al.
Published: (2024)
by: Shi, Zhaoyang, et al.
Published: (2024)
Sampling conditioned diffusions via Pathspace Projected Monte Carlo
by: Grafke, Tobias
Published: (2025)
by: Grafke, Tobias
Published: (2025)
Phase Transition for Stochastic Block Model with more than $\sqrt{n}$ Communities
by: Carpentier, Alexandra, et al.
Published: (2025)
by: Carpentier, Alexandra, et al.
Published: (2025)
Similar Items
-
Truncated LinUCB for Stochastic Linear Bandits
by: Song, Yanglei, et al.
Published: (2022) -
UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
by: Han, Qiyang, et al.
Published: (2024) -
Transfer Learning for Contextual Multi-armed Bandits
by: Cai, Changxiao, et al.
Published: (2022) -
Early Stopping in Contextual Bandits and Inferences
by: Cui, Zihan
Published: (2025) -
Design Experiments to Compare Multi-armed Bandit Algorithms
by: Meng, Huiling, et al.
Published: (2026)