Truncated LinUCB for Stochastic Linear Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Yanglei, zhou, Meng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Statistical Inference under Adaptive Sampling with LinUCB
von: Fan, Wei, et al.
Veröffentlicht: (2025)
von: Fan, Wei, et al.
Veröffentlicht: (2025)
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
Tractable Instances of Bilinear Maximization: Implementing LinUCB on Ellipsoids
von: Zhang, Raymond, et al.
Veröffentlicht: (2025)
von: Zhang, Raymond, et al.
Veröffentlicht: (2025)
Extended UCB Policies for Multi-armed Bandit Problems
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
Precise Asymptotics and Refined Regret of Variance-Aware UCB
von: Fan, Yingying, et al.
Veröffentlicht: (2024)
von: Fan, Yingying, et al.
Veröffentlicht: (2024)
Testing the Feasibility of Linear Programs with Bandit Feedback
von: Gangrade, Aditya, et al.
Veröffentlicht: (2024)
von: Gangrade, Aditya, et al.
Veröffentlicht: (2024)
Scalable LinUCB: Low-Rank Design Matrix Updates for Recommenders with Large Action Spaces
von: Shustova, Evgenia, et al.
Veröffentlicht: (2025)
von: Shustova, Evgenia, et al.
Veröffentlicht: (2025)
UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
von: Han, Qiyang, et al.
Veröffentlicht: (2024)
von: Han, Qiyang, et al.
Veröffentlicht: (2024)
Navigating Sparsities in High-Dimensional Linear Contextual Bandits
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
FLIPHAT: Joint Differential Privacy for High Dimensional Sparse Linear Bandits
von: Chakraborty, Sunrit, et al.
Veröffentlicht: (2024)
von: Chakraborty, Sunrit, et al.
Veröffentlicht: (2024)
Design Experiments to Compare Multi-armed Bandit Algorithms
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
Avoiding the Price of Adaptivity: Inference in Linear Contextual Bandits via Stability
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
A Simple and Optimal Policy Design with Safety against Heavy-Tailed Risk for Stochastic Bandits
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
Linear Regression with Unknown Truncation Beyond Gaussian Features
von: Kouridakis, Alexandros, et al.
Veröffentlicht: (2026)
von: Kouridakis, Alexandros, et al.
Veröffentlicht: (2026)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
von: Lattimore, Tor
Veröffentlicht: (2026)
von: Lattimore, Tor
Veröffentlicht: (2026)
Mean Testing under Truncation beyond Gaussian
von: Wang, Yuhao, et al.
Veröffentlicht: (2026)
von: Wang, Yuhao, et al.
Veröffentlicht: (2026)
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2023)
Training Diagonal Linear Networks with Stochastic Sharpness-Aware Minimization
von: Clara, Gabriel, et al.
Veröffentlicht: (2025)
von: Clara, Gabriel, et al.
Veröffentlicht: (2025)
The Fragility of Optimized Bandit Algorithms
von: Fan, Lin, et al.
Veröffentlicht: (2021)
von: Fan, Lin, et al.
Veröffentlicht: (2021)
Batched Nonparametric Contextual Bandits
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
von: Jiang, Rong, et al.
Veröffentlicht: (2024)
Adaptive Smooth Non-Stationary Bandits
von: Suk, Joe
Veröffentlicht: (2024)
von: Suk, Joe
Veröffentlicht: (2024)
Transfer Learning for Contextual Multi-armed Bandits
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
Multitask Learning and Bandits via Robust Statistics
von: Xu, Kan, et al.
Veröffentlicht: (2021)
von: Xu, Kan, et al.
Veröffentlicht: (2021)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
BELIEF in Dependence: Leveraging Atomic Linearity in Data Bits for Rethinking Generalized Linear Models
von: Brown, Benjamin, et al.
Veröffentlicht: (2022)
von: Brown, Benjamin, et al.
Veröffentlicht: (2022)
Online Clustering of Data Sequences with Bandit Information
von: Chandran, G Dhinesh, et al.
Veröffentlicht: (2025)
von: Chandran, G Dhinesh, et al.
Veröffentlicht: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
Multi-Armed Bandits With Machine Learning-Generated Surrogate Rewards
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
von: Ji, Wenlong, et al.
Veröffentlicht: (2025)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
Towards Efficient and Optimal Covariance-Adaptive Algorithms for Combinatorial Semi-Bandits
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
The Sample Complexity of Multiple Change Point Identification under Bandit Feedback
von: Graf, Maximilian, et al.
Veröffentlicht: (2026)
von: Graf, Maximilian, et al.
Veröffentlicht: (2026)
The Adaptivity Barrier in Batched Nonparametric Bandits: Sharp Characterization of the Price of Unknown Margin
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
Upper Counterfactual Confidence Bounds: a New Optimism Principle for Contextual Bandits
von: Xu, Yunbei, et al.
Veröffentlicht: (2020)
von: Xu, Yunbei, et al.
Veröffentlicht: (2020)
Sequential Multiple Testing: A Second-Order Asymptotic Analysis
von: Liu, Jingyu, et al.
Veröffentlicht: (2026)
von: Liu, Jingyu, et al.
Veröffentlicht: (2026)
Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
PAC Learning with Bandit Feedback: Sharp Sample Complexity in the Realizable Setting
von: Hanneke, Steve, et al.
Veröffentlicht: (2026)
von: Hanneke, Steve, et al.
Veröffentlicht: (2026)
Choosing the Better Bandit Algorithm under Data Sharing: When Do A/B Experiments Work?
von: Li, Shuangning, et al.
Veröffentlicht: (2025)
von: Li, Shuangning, et al.
Veröffentlicht: (2025)
Concentrated Differential Privacy for Bandits
von: Azize, Achraf, et al.
Veröffentlicht: (2023)
von: Azize, Achraf, et al.
Veröffentlicht: (2023)
Batched Nonparametric Bandits via k-Nearest Neighbor UCB
von: Arya, Sakshi
Veröffentlicht: (2025)
von: Arya, Sakshi
Veröffentlicht: (2025)
Ähnliche Einträge
-
Statistical Inference under Adaptive Sampling with LinUCB
von: Fan, Wei, et al.
Veröffentlicht: (2025) -
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
von: Liu, Jingyu, et al.
Veröffentlicht: (2025) -
Tractable Instances of Bilinear Maximization: Implementing LinUCB on Ellipsoids
von: Zhang, Raymond, et al.
Veröffentlicht: (2025) -
Extended UCB Policies for Multi-armed Bandit Problems
von: Liu, Keqin, et al.
Veröffentlicht: (2011) -
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)