Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Yifan, Duchi, John C., Van Roy, Benjamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
von: Rumi, Alberto, et al.
Veröffentlicht: (2026)
von: Rumi, Alberto, et al.
Veröffentlicht: (2026)
No-Regret Linear Bandits under Gap-Adjusted Misspecification
von: Liu, Chong, et al.
Veröffentlicht: (2025)
von: Liu, Chong, et al.
Veröffentlicht: (2025)
Near-Optimal Regret in Linear MDPs with Aggregate Bandit Feedback
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
von: Cassel, Asaf, et al.
Veröffentlicht: (2024)
On the Optimal Regret of Locally Private Linear Contextual Bandit
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
Generalized Linear Bandits: Almost Optimal Regret with One-Pass Update
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
von: Tajdini, Artin, et al.
Veröffentlicht: (2025)
Chained Information-Theoretic bounds and Tight Regret Rate for Linear Bandit Problems
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
Local Anti-Concentration Class: Logarithmic Regret for Greedy Linear Contextual Bandit
von: Kim, Seok-Jin, et al.
Veröffentlicht: (2024)
von: Kim, Seok-Jin, et al.
Veröffentlicht: (2024)
Time-Varying Gaussian Process Bandits with Unknown Prior
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
Non-Stationary Bandit Learning via Predictive Sampling
von: Liu, Yueyang, et al.
Veröffentlicht: (2022)
von: Liu, Yueyang, et al.
Veröffentlicht: (2022)
Regret Bounds for Expected Improvement Algorithms in Gaussian Process Bandit Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
Tighter Regret Lower Bound for Gaussian Process Bandits with Squared Exponential Kernel in Hypersphere
von: Iwazaki, Shogo
Veröffentlicht: (2026)
von: Iwazaki, Shogo
Veröffentlicht: (2026)
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
Some Robustness Properties of Label Cleaning
von: Cheng, Chen, et al.
Veröffentlicht: (2025)
von: Cheng, Chen, et al.
Veröffentlicht: (2025)
Optimal Regret for Single Index Bandits
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
von: Dey, Devdan, et al.
Veröffentlicht: (2026)
Bayesian Regret Minimization in Offline Bandits
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
von: Petrik, Marek, et al.
Veröffentlicht: (2023)
No-Regret is not enough! Bandits with General Constraints through Adaptive Regret Minimization
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
von: Bernasconi, Martino, et al.
Veröffentlicht: (2024)
Gaussian Process Upper Confidence Bound Achieves Nearly-Optimal Regret in Noise-Free Gaussian Process Bandits
von: Iwazaki, Shogo
Veröffentlicht: (2025)
von: Iwazaki, Shogo
Veröffentlicht: (2025)
Adaptive Prior Selection in Gaussian Process Bandits with Thompson Sampling
von: Sandberg, Jack, et al.
Veröffentlicht: (2025)
von: Sandberg, Jack, et al.
Veröffentlicht: (2025)
p-Mean Regret for Stochastic Bandits
von: Krishna, Anand, et al.
Veröffentlicht: (2024)
von: Krishna, Anand, et al.
Veröffentlicht: (2024)
Satisficing Regret Minimization in Bandits: Constant Rate and Light-Tailed Distribution
von: Feng, Qing, et al.
Veröffentlicht: (2024)
von: Feng, Qing, et al.
Veröffentlicht: (2024)
Modified Meta-Thompson Sampling for Linear Bandits and Its Bayes Regret Analysis
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Tangential Randomization in Linear Bandits (TRAiL): Guaranteed Inference and Regret Bounds
von: Güçlü, Arda, et al.
Veröffentlicht: (2024)
von: Güçlü, Arda, et al.
Veröffentlicht: (2024)
Distribution free M-estimation
von: Areces, Felipe, et al.
Veröffentlicht: (2025)
von: Areces, Felipe, et al.
Veröffentlicht: (2025)
Improved Regret Bounds for Online Fair Division with Bandit Learning
von: Schiffer, Benjamin, et al.
Veröffentlicht: (2025)
von: Schiffer, Benjamin, et al.
Veröffentlicht: (2025)
Revisiting Matrix Sketching in Linear Bandits: Achieving Sublinear Regret via Dyadic Block Sketching
von: Wen, Dongxie, et al.
Veröffentlicht: (2024)
von: Wen, Dongxie, et al.
Veröffentlicht: (2024)
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
Optimal Regret for Policy Optimization in Contextual Bandits
von: Levy, Orin, et al.
Veröffentlicht: (2026)
von: Levy, Orin, et al.
Veröffentlicht: (2026)
Efficient Swap Regret Minimization in Combinatorial Bandits
von: Kontogiannis, Andreas, et al.
Veröffentlicht: (2026)
von: Kontogiannis, Andreas, et al.
Veröffentlicht: (2026)
Fast Best-in-Class Regret for Contextual Bandits
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
von: Girard, Samuel, et al.
Veröffentlicht: (2025)
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
Minimum Empirical Divergence for Sub-Gaussian Linear Bandits
von: Balagopalan, Kapilan, et al.
Veröffentlicht: (2024)
von: Balagopalan, Kapilan, et al.
Veröffentlicht: (2024)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
Honor Among Bandits: No-Regret Learning for Online Fair Division
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2024)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2024)
Structured Diffusion Models with Mixture of Gaussians as Prior Distribution
von: Jia, Nanshan, et al.
Veröffentlicht: (2024)
von: Jia, Nanshan, et al.
Veröffentlicht: (2024)
Queue Length Regret Bounds for Contextual Queueing Bandits
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
Finite-Time Regret Analysis of Retry-Aware Bandits
von: Tong, Bingkui, et al.
Veröffentlicht: (2026)
von: Tong, Bingkui, et al.
Veröffentlicht: (2026)
How Does Variance Shape the Regret in Contextual Bandits?
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
Improved Regret for Bandit Convex Optimization with Delayed Feedback
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024) -
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
von: Rumi, Alberto, et al.
Veröffentlicht: (2026) -
No-Regret Linear Bandits under Gap-Adjusted Misspecification
von: Liu, Chong, et al.
Veröffentlicht: (2025) -
Near-Optimal Regret in Linear MDPs with Aggregate Bandit Feedback
von: Cassel, Asaf, et al.
Veröffentlicht: (2024) -
On the Optimal Regret of Locally Private Linear Contextual Bandit
von: Li, Jiachun, et al.
Veröffentlicht: (2024)