A Practical Algorithm for Feature-Rich, Non-Stationary Bandit Problems
Fuente:
arXiv
Guardado en:
| Autores principales: | Loh, Wei Min, Sinha, Sajib Kumer, Agarwal, Ankur, Poupart, Pascal |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Basis Transformers for Multi-Task Tabular Regression
por: Loh, Wei Min, et al.
Publicado: (2025)
por: Loh, Wei Min, et al.
Publicado: (2025)
Non-Stationary Dueling Bandits Under a Weighted Borda Criterion
por: Suk, Joe, et al.
Publicado: (2024)
por: Suk, Joe, et al.
Publicado: (2024)
Near-Optimal Algorithm for Non-Stationary Kernelized Bandits
por: Iwazaki, Shogo, et al.
Publicado: (2024)
por: Iwazaki, Shogo, et al.
Publicado: (2024)
Non-Stationary Lipschitz Bandits
por: Nguyen, Nicolas, et al.
Publicado: (2025)
por: Nguyen, Nicolas, et al.
Publicado: (2025)
Fooling Algorithms in Non-Stationary Bandits using Belief Inertia
por: Mendelson, Gal, et al.
Publicado: (2025)
por: Mendelson, Gal, et al.
Publicado: (2025)
DAL: A Practical Prior-Free Black-Box Framework for Non-Stationary Bandits
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
por: Goyal, Tanmay, et al.
Publicado: (2025)
por: Goyal, Tanmay, et al.
Publicado: (2025)
Optimal and Practical Batched Linear Bandit Algorithm
por: Yu, Sanghoon, et al.
Publicado: (2025)
por: Yu, Sanghoon, et al.
Publicado: (2025)
Adaptive Smooth Non-Stationary Bandits
por: Suk, Joe
Publicado: (2024)
por: Suk, Joe
Publicado: (2024)
A Minimalist Method for Fine-tuning Text-to-Image Diffusion Models
por: Miao, Yanting, et al.
Publicado: (2025)
por: Miao, Yanting, et al.
Publicado: (2025)
Non-Stationary Bandit Learning via Predictive Sampling
por: Liu, Yueyang, et al.
Publicado: (2022)
por: Liu, Yueyang, et al.
Publicado: (2022)
Smooth Non-Stationary Bandits
por: Jia, Su, et al.
Publicado: (2023)
por: Jia, Su, et al.
Publicado: (2023)
Incentivized Exploration of Non-Stationary Stochastic Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2024)
por: Chakraborty, Sourav, et al.
Publicado: (2024)
Non-Stationary Latent Auto-Regressive Bandits
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
por: Maynard-Zhang, Leo, et al.
Publicado: (2026)
por: Maynard-Zhang, Leo, et al.
Publicado: (2026)
Why Online Reinforcement Learning is Causal
por: Schulte, Oliver, et al.
Publicado: (2024)
por: Schulte, Oliver, et al.
Publicado: (2024)
Non-Stationary Restless Multi-Armed Bandits with Provable Guarantee
por: Hung, Yu-Heng, et al.
Publicado: (2025)
por: Hung, Yu-Heng, et al.
Publicado: (2025)
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
por: Li, Shaoang, et al.
Publicado: (2025)
por: Li, Shaoang, et al.
Publicado: (2025)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
por: Veness, Joel, et al.
Publicado: (2025)
por: Veness, Joel, et al.
Publicado: (2025)
Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates
por: Yu, Sanghoon, et al.
Publicado: (2026)
por: Yu, Sanghoon, et al.
Publicado: (2026)
BanditQ: Fair Bandits with Guaranteed Rewards
por: Sinha, Abhishek
Publicado: (2023)
por: Sinha, Abhishek
Publicado: (2023)
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
por: Werge, Nicklas, et al.
Publicado: (2023)
por: Werge, Nicklas, et al.
Publicado: (2023)
TDHook: A Lightweight Framework for Interpretability
por: Poupart, Yoann
Publicado: (2025)
por: Poupart, Yoann
Publicado: (2025)
Subject-driven Text-to-Image Generation via Preference-based Reinforcement Learning
por: Miao, Yanting, et al.
Publicado: (2024)
por: Miao, Yanting, et al.
Publicado: (2024)
Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels
por: Shisher, Md Kamran Chowdhury, et al.
Publicado: (2025)
por: Shisher, Md Kamran Chowdhury, et al.
Publicado: (2025)
Adaptive Requesting in Decentralized Edge Networks via Non-Stationary Bandits
por: Zhuang, Yi, et al.
Publicado: (2026)
por: Zhuang, Yi, et al.
Publicado: (2026)
A Linear Programming Enhanced Genetic Algorithm for Hyperparameter Tuning in Machine Learning
por: Sinha, Ankur, et al.
Publicado: (2024)
por: Sinha, Ankur, et al.
Publicado: (2024)
FinBloom: Knowledge Grounding Large Language Model with Real-time Financial Data
por: Sinha, Ankur, et al.
Publicado: (2025)
por: Sinha, Ankur, et al.
Publicado: (2025)
FedLog: Personalized Federated Classification with Less Communication and More Flexibility
por: Yu, Haolin, et al.
Publicado: (2024)
por: Yu, Haolin, et al.
Publicado: (2024)
Catoni-Style Change Point Detection for Regret Minimization in Non-Stationary Heavy-Tailed Bandits
por: Genalti, Gianmarco, et al.
Publicado: (2025)
por: Genalti, Gianmarco, et al.
Publicado: (2025)
Contextual Bandits with Non-Stationary Correlated Rewards for User Association in MmWave Vehicular Networks
por: He, Xiaoyang, et al.
Publicado: (2024)
por: He, Xiaoyang, et al.
Publicado: (2024)
Exploration via Feature Perturbation in Contextual Bandits
por: Yi, Seouh-won, et al.
Publicado: (2025)
por: Yi, Seouh-won, et al.
Publicado: (2025)
Constrained Contextual Bandits with Adversarial Contexts
por: Sarkar, Dhruv, et al.
Publicado: (2026)
por: Sarkar, Dhruv, et al.
Publicado: (2026)
A Modularized Framework for Piecewise-Stationary Restless Bandits
por: Li, Kuan-Ta, et al.
Publicado: (2026)
por: Li, Kuan-Ta, et al.
Publicado: (2026)
Measures of Variability for Risk-averse Policy Gradient
por: Luo, Yudong, et al.
Publicado: (2025)
por: Luo, Yudong, et al.
Publicado: (2025)
Time Is Effort: Estimating Human Post-Editing Time for Grammar Error Correction Tool Evaluation
por: Vadehra, Ankit, et al.
Publicado: (2025)
por: Vadehra, Ankit, et al.
Publicado: (2025)
Linear Bandits with Partially Observable Features
por: Kim, Wonyoung, et al.
Publicado: (2025)
por: Kim, Wonyoung, et al.
Publicado: (2025)
Calibrated One Round Federated Learning with Bayesian Inference in the Predictive Space
por: Hasan, Mohsin, et al.
Publicado: (2023)
por: Hasan, Mohsin, et al.
Publicado: (2023)
Preventing Arbitrarily High Confidence on Far-Away Data in Point-Estimated Discriminative Neural Networks
por: Rashid, Ahmad, et al.
Publicado: (2023)
por: Rashid, Ahmad, et al.
Publicado: (2023)
Multi-Armed Bandits with Network Interference
por: Agarwal, Abhineet, et al.
Publicado: (2024)
por: Agarwal, Abhineet, et al.
Publicado: (2024)
Ejemplares similares
-
Basis Transformers for Multi-Task Tabular Regression
por: Loh, Wei Min, et al.
Publicado: (2025) -
Non-Stationary Dueling Bandits Under a Weighted Borda Criterion
por: Suk, Joe, et al.
Publicado: (2024) -
Near-Optimal Algorithm for Non-Stationary Kernelized Bandits
por: Iwazaki, Shogo, et al.
Publicado: (2024) -
Non-Stationary Lipschitz Bandits
por: Nguyen, Nicolas, et al.
Publicado: (2025) -
Fooling Algorithms in Non-Stationary Bandits using Belief Inertia
por: Mendelson, Gal, et al.
Publicado: (2025)