Non-Stationary Bandit Learning via Predictive Sampling
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Yueyang, Kuang, Xu, Van Roy, Benjamin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Non-Stationary Lipschitz Bandits
por: Nguyen, Nicolas, et al.
Publicado: (2025)
por: Nguyen, Nicolas, et al.
Publicado: (2025)
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
por: Li, Shaoang, et al.
Publicado: (2025)
por: Li, Shaoang, et al.
Publicado: (2025)
Non-Stationary Latent Auto-Regressive Bandits
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Adaptive Smooth Non-Stationary Bandits
por: Suk, Joe
Publicado: (2024)
por: Suk, Joe
Publicado: (2024)
Near-Optimal Algorithm for Non-Stationary Kernelized Bandits
por: Iwazaki, Shogo, et al.
Publicado: (2024)
por: Iwazaki, Shogo, et al.
Publicado: (2024)
Smooth Non-Stationary Bandits
por: Jia, Su, et al.
Publicado: (2023)
por: Jia, Su, et al.
Publicado: (2023)
Incentivized Exploration of Non-Stationary Stochastic Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2024)
por: Chakraborty, Sourav, et al.
Publicado: (2024)
Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels
por: Shisher, Md Kamran Chowdhury, et al.
Publicado: (2025)
por: Shisher, Md Kamran Chowdhury, et al.
Publicado: (2025)
Adaptive Requesting in Decentralized Edge Networks via Non-Stationary Bandits
por: Zhuang, Yi, et al.
Publicado: (2026)
por: Zhuang, Yi, et al.
Publicado: (2026)
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
por: Maynard-Zhang, Leo, et al.
Publicado: (2026)
por: Maynard-Zhang, Leo, et al.
Publicado: (2026)
Non-Stationary Restless Multi-Armed Bandits with Provable Guarantee
por: Hung, Yu-Heng, et al.
Publicado: (2025)
por: Hung, Yu-Heng, et al.
Publicado: (2025)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
por: Veness, Joel, et al.
Publicado: (2025)
por: Veness, Joel, et al.
Publicado: (2025)
Inverse Contextual Bandits without Rewards: Learning from a Non-Stationary Learner via Suffix Imitation
por: Kong, Yuqi, et al.
Publicado: (2026)
por: Kong, Yuqi, et al.
Publicado: (2026)
Fooling Algorithms in Non-Stationary Bandits using Belief Inertia
por: Mendelson, Gal, et al.
Publicado: (2025)
por: Mendelson, Gal, et al.
Publicado: (2025)
A Practical Algorithm for Feature-Rich, Non-Stationary Bandit Problems
por: Loh, Wei Min, et al.
Publicado: (2026)
por: Loh, Wei Min, et al.
Publicado: (2026)
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
por: Werge, Nicklas, et al.
Publicado: (2023)
por: Werge, Nicklas, et al.
Publicado: (2023)
Non-Stationary Dueling Bandits Under a Weighted Borda Criterion
por: Suk, Joe, et al.
Publicado: (2024)
por: Suk, Joe, et al.
Publicado: (2024)
Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
por: Zhu, Yifan, et al.
Publicado: (2026)
por: Zhu, Yifan, et al.
Publicado: (2026)
Posterior Sampling for Continuing Environments
por: Xu, Wanqiao, et al.
Publicado: (2022)
por: Xu, Wanqiao, et al.
Publicado: (2022)
Wavelet Predictive Representations for Non-Stationary Reinforcement Learning
por: Wang, Min, et al.
Publicado: (2025)
por: Wang, Min, et al.
Publicado: (2025)
Aligning AI Agents via Information-Directed Sampling
por: Jeon, Hong Jun, et al.
Publicado: (2024)
por: Jeon, Hong Jun, et al.
Publicado: (2024)
DAL: A Practical Prior-Free Black-Box Framework for Non-Stationary Bandits
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
Sliding-Window Thompson Sampling for Non-Stationary Settings
por: Fiandri, Marco, et al.
Publicado: (2024)
por: Fiandri, Marco, et al.
Publicado: (2024)
Hedging Memory Horizons for Non-Stationary Prediction via Online Aggregation
por: Wang, Yutong, et al.
Publicado: (2026)
por: Wang, Yutong, et al.
Publicado: (2026)
Catoni-Style Change Point Detection for Regret Minimization in Non-Stationary Heavy-Tailed Bandits
por: Genalti, Gianmarco, et al.
Publicado: (2025)
por: Genalti, Gianmarco, et al.
Publicado: (2025)
Contextual Bandits with Non-Stationary Correlated Rewards for User Association in MmWave Vehicular Networks
por: He, Xiaoyang, et al.
Publicado: (2024)
por: He, Xiaoyang, et al.
Publicado: (2024)
Adaptive Client Sampling in Federated Learning via Online Learning with Bandit Feedback
por: Zhao, Boxin, et al.
Publicado: (2021)
por: Zhao, Boxin, et al.
Publicado: (2021)
Continual Learning as Computationally Constrained Reinforcement Learning
por: Kumar, Saurabh, et al.
Publicado: (2023)
por: Kumar, Saurabh, et al.
Publicado: (2023)
Offline Contextual Bandit with Counterfactual Sample Identification
por: Gilotte, Alexandre, et al.
Publicado: (2025)
por: Gilotte, Alexandre, et al.
Publicado: (2025)
Non-Stationary Online Structured Prediction with Surrogate Losses
por: Sakaue, Shinsaku, et al.
Publicado: (2025)
por: Sakaue, Shinsaku, et al.
Publicado: (2025)
Improved Regret Analysis in Gaussian Process Bandits: Optimality for Noiseless Reward, RKHS norm, and Non-Stationary Variance
por: Iwazaki, Shogo, et al.
Publicado: (2025)
por: Iwazaki, Shogo, et al.
Publicado: (2025)
A Modularized Framework for Piecewise-Stationary Restless Bandits
por: Li, Kuan-Ta, et al.
Publicado: (2026)
por: Li, Kuan-Ta, et al.
Publicado: (2026)
Adaptive Calibration in Non-Stationary Environments
por: Liu, Junyan, et al.
Publicado: (2026)
por: Liu, Junyan, et al.
Publicado: (2026)
Performative Prediction with Bandit Feedback: Learning through Reparameterization
por: Chen, Yatong, et al.
Publicado: (2023)
por: Chen, Yatong, et al.
Publicado: (2023)
StatioCL: Contrastive Learning for Time Series via Non-Stationary and Temporal Contrast
por: Wu, Yu, et al.
Publicado: (2024)
por: Wu, Yu, et al.
Publicado: (2024)
Q-Learning with Shift-Aware Upper Confidence Bound in Non-Stationary Reinforcement Learning
por: Bui, Ha Manh, et al.
Publicado: (2025)
por: Bui, Ha Manh, et al.
Publicado: (2025)
Maintaining Plasticity in Continual Learning via Regenerative Regularization
por: Kumar, Saurabh, et al.
Publicado: (2023)
por: Kumar, Saurabh, et al.
Publicado: (2023)
Non-Stationary Online Resource Allocation: Learning from a Single Sample
por: Feng, Yiding, et al.
Publicado: (2026)
por: Feng, Yiding, et al.
Publicado: (2026)
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
por: Chai, Jinhang, et al.
Publicado: (2025)
por: Chai, Jinhang, et al.
Publicado: (2025)
In-Context Learning for Non-Stationary MIMO Equalization
por: Jiang, Jiachen, et al.
Publicado: (2025)
por: Jiang, Jiachen, et al.
Publicado: (2025)
Ejemplares similares
-
Non-Stationary Lipschitz Bandits
por: Nguyen, Nicolas, et al.
Publicado: (2025) -
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
por: Li, Shaoang, et al.
Publicado: (2025) -
Non-Stationary Latent Auto-Regressive Bandits
por: Trella, Anna L., et al.
Publicado: (2024) -
Adaptive Smooth Non-Stationary Bandits
por: Suk, Joe
Publicado: (2024) -
Near-Optimal Algorithm for Non-Stationary Kernelized Bandits
por: Iwazaki, Shogo, et al.
Publicado: (2024)