Generalized Linear Bandits: Almost Optimal Regret with One-Pass Update
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yu-Jie, Xu, Sheng-An, Zhao, Peng, Sugiyama, Masashi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Heavy-Tailed Linear Bandits: Huber Regression with One-Pass Update
di: Wang, Jing, et al.
Pubblicazione: (2025)
di: Wang, Jing, et al.
Pubblicazione: (2025)
Non-stationary Online Learning for Curved Losses: Improved Dynamic Regret via Mixability
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2025)
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2025)
Near-Optimal Regret in Adversarial Kernel Bandits
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2026)
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2026)
Multi-Player Approaches for Dueling Bandits
di: Raveh, Or, et al.
Pubblicazione: (2024)
di: Raveh, Or, et al.
Pubblicazione: (2024)
The Survival Bandit Problem
di: Riou, Charles, et al.
Pubblicazione: (2022)
di: Riou, Charles, et al.
Pubblicazione: (2022)
Near-Optimal Regret in Linear MDPs with Aggregate Bandit Feedback
di: Cassel, Asaf, et al.
Pubblicazione: (2024)
di: Cassel, Asaf, et al.
Pubblicazione: (2024)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
di: Nakamura, Shintaro, et al.
Pubblicazione: (2023)
di: Nakamura, Shintaro, et al.
Pubblicazione: (2023)
On the Optimal Regret of Locally Private Linear Contextual Bandit
di: Li, Jiachun, et al.
Pubblicazione: (2024)
di: Li, Jiachun, et al.
Pubblicazione: (2024)
Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates
di: Yu, Sanghoon, et al.
Pubblicazione: (2026)
di: Yu, Sanghoon, et al.
Pubblicazione: (2026)
Adapting to Continuous Covariate Shift via Online Density Ratio Estimation
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2023)
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2023)
Optimal Regret for Single Index Bandits
di: Dey, Devdan, et al.
Pubblicazione: (2026)
di: Dey, Devdan, et al.
Pubblicazione: (2026)
No-Regret Linear Bandits under Gap-Adjusted Misspecification
di: Liu, Chong, et al.
Pubblicazione: (2025)
di: Liu, Chong, et al.
Pubblicazione: (2025)
Revisiting Matrix Sketching in Linear Bandits: Achieving Sublinear Regret via Dyadic Block Sketching
di: Wen, Dongxie, et al.
Pubblicazione: (2024)
di: Wen, Dongxie, et al.
Pubblicazione: (2024)
Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
Optimal Regret for Policy Optimization in Contextual Bandits
di: Levy, Orin, et al.
Pubblicazione: (2026)
di: Levy, Orin, et al.
Pubblicazione: (2026)
Near-Optimal Dynamic Regret for Adversarial Linear Mixture MDPs
di: Li, Long-Fei, et al.
Pubblicazione: (2024)
di: Li, Long-Fei, et al.
Pubblicazione: (2024)
Achieving Optimal Static and Dynamic Regret Simultaneously in Bandits with Deterministic Losses
di: Qian, Jian, et al.
Pubblicazione: (2026)
di: Qian, Jian, et al.
Pubblicazione: (2026)
Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards
di: Panda, Subhodip, et al.
Pubblicazione: (2026)
di: Panda, Subhodip, et al.
Pubblicazione: (2026)
Parameter-Free Dynamic Regret for Unconstrained Linear Bandits
di: Rumi, Alberto, et al.
Pubblicazione: (2026)
di: Rumi, Alberto, et al.
Pubblicazione: (2026)
Nearly Minimax Optimal Regret for Multinomial Logistic Bandit
di: Lee, Joongkyu, et al.
Pubblicazione: (2024)
di: Lee, Joongkyu, et al.
Pubblicazione: (2024)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
di: Ye, Zichun, et al.
Pubblicazione: (2025)
di: Ye, Zichun, et al.
Pubblicazione: (2025)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
di: Braun, Guillaume, et al.
Pubblicazione: (2024)
di: Braun, Guillaume, et al.
Pubblicazione: (2024)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
di: Ji, Kaixuan, et al.
Pubblicazione: (2026)
di: Ji, Kaixuan, et al.
Pubblicazione: (2026)
One Good Source is All You Need: Near-Optimal Regret for Bandits under Heterogeneous Noise
di: Bhat, Amith, et al.
Pubblicazione: (2026)
di: Bhat, Amith, et al.
Pubblicazione: (2026)
Almost Minimax Optimal Best Arm Identification in Piecewise Stationary Linear Bandits
di: Hou, Yunlong, et al.
Pubblicazione: (2024)
di: Hou, Yunlong, et al.
Pubblicazione: (2024)
No-Regret is not enough! Bandits with General Constraints through Adaptive Regret Minimization
di: Bernasconi, Martino, et al.
Pubblicazione: (2024)
di: Bernasconi, Martino, et al.
Pubblicazione: (2024)
Optimal Batched Linear Bandits
di: Ren, Xuanfei, et al.
Pubblicazione: (2024)
di: Ren, Xuanfei, et al.
Pubblicazione: (2024)
Preference-centric Bandits: Optimality of Mixtures and Regret-efficient Algorithms
di: Tatlı, Meltem, et al.
Pubblicazione: (2025)
di: Tatlı, Meltem, et al.
Pubblicazione: (2025)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
di: Qiu, Hao, et al.
Pubblicazione: (2026)
di: Qiu, Hao, et al.
Pubblicazione: (2026)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
di: Réveillard, William, et al.
Pubblicazione: (2025)
di: Réveillard, William, et al.
Pubblicazione: (2025)
Optimal and Practical Batched Linear Bandit Algorithm
di: Yu, Sanghoon, et al.
Pubblicazione: (2025)
di: Yu, Sanghoon, et al.
Pubblicazione: (2025)
Federated Q-Learning with Reference-Advantage Decomposition: Almost Optimal Regret and Logarithmic Communication Cost
di: Zheng, Zhong, et al.
Pubblicazione: (2024)
di: Zheng, Zhong, et al.
Pubblicazione: (2024)
Risk-sensitive Bandits: Arm Mixture Optimality and Regret-efficient Algorithms
di: Tatlı, Meltem, et al.
Pubblicazione: (2025)
di: Tatlı, Meltem, et al.
Pubblicazione: (2025)
Learning What to Recommend: Minimax Optimal Simple Regret in Logistic Bandits
di: Liu, Shuai, et al.
Pubblicazione: (2026)
di: Liu, Shuai, et al.
Pubblicazione: (2026)
Optimal Thresholding Linear Bandit
di: Rivera, Eduardo Ochoa, et al.
Pubblicazione: (2024)
di: Rivera, Eduardo Ochoa, et al.
Pubblicazione: (2024)
Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
di: Tajdini, Artin, et al.
Pubblicazione: (2025)
di: Tajdini, Artin, et al.
Pubblicazione: (2025)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
Model Predictive Control is Almost Optimal for Restless Bandit
di: Gast, Nicolas, et al.
Pubblicazione: (2024)
di: Gast, Nicolas, et al.
Pubblicazione: (2024)
Linear Bandits on Ellipsoids: Minimax Optimal Algorithms
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
Chained Information-Theoretic bounds and Tight Regret Rate for Linear Bandit Problems
di: Gouverneur, Amaury, et al.
Pubblicazione: (2024)
di: Gouverneur, Amaury, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Heavy-Tailed Linear Bandits: Huber Regression with One-Pass Update
di: Wang, Jing, et al.
Pubblicazione: (2025) -
Non-stationary Online Learning for Curved Losses: Improved Dynamic Regret via Mixability
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2025) -
Near-Optimal Regret in Adversarial Kernel Bandits
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2026) -
Multi-Player Approaches for Dueling Bandits
di: Raveh, Or, et al.
Pubblicazione: (2024) -
The Survival Bandit Problem
di: Riou, Charles, et al.
Pubblicazione: (2022)