VITS : Variational Inference Thompson Sampling for contextual bandits
Fuente:
arXiv
Guardado en:
| Autores principales: | Clavier, Pierre, Huix, Tom, Durmus, Alain |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
por: Huix, Tom, et al.
Publicado: (2024)
por: Huix, Tom, et al.
Publicado: (2024)
Variance-sensitive Thompson sampling for generalised linear bandits, revisited
por: Perneczky, Tom, et al.
Publicado: (2026)
por: Perneczky, Tom, et al.
Publicado: (2026)
Central Limit Theorem for Bayesian Neural Network trained with Variational Inference
por: Descours, Arnaud, et al.
Publicado: (2024)
por: Descours, Arnaud, et al.
Publicado: (2024)
Best-of-Both Worlds for linear contextual bandits with paid observations
por: Boyer, Nathan, et al.
Publicado: (2025)
por: Boyer, Nathan, et al.
Publicado: (2025)
Optimal cross-learning for contextual bandits with unknown context distributions
por: Schneider, Jon, et al.
Publicado: (2024)
por: Schneider, Jon, et al.
Publicado: (2024)
Model selection for behavioral learning data and applications to contextual bandits
por: Aubert, Julien, et al.
Publicado: (2025)
por: Aubert, Julien, et al.
Publicado: (2025)
Leveraging heterogeneous spillover in maximizing contextual bandit rewards
por: Faruk, Ahmed Sayeed, et al.
Publicado: (2023)
por: Faruk, Ahmed Sayeed, et al.
Publicado: (2023)
Anytime-valid off-policy inference for contextual bandits
por: Waudby-Smith, Ian, et al.
Publicado: (2022)
por: Waudby-Smith, Ian, et al.
Publicado: (2022)
A conversion theorem and minimax optimality for continuum contextual bandits
por: Akhavan, Arya, et al.
Publicado: (2024)
por: Akhavan, Arya, et al.
Publicado: (2024)
Quantum contextual bandits and recommender systems for quantum data
por: Brahmachari, Shrigyan, et al.
Publicado: (2023)
por: Brahmachari, Shrigyan, et al.
Publicado: (2023)
Vector preference-based contextual bandits under distributional shifts
por: Shukla, Apurv, et al.
Publicado: (2025)
por: Shukla, Apurv, et al.
Publicado: (2025)
Variational Diffusion Posterior Sampling with Midpoint Guidance
por: Moufad, Badr, et al.
Publicado: (2024)
por: Moufad, Badr, et al.
Publicado: (2024)
Precision autotuning for linear solvers via contextual bandit-based RL
por: Carson, Erin, et al.
Publicado: (2026)
por: Carson, Erin, et al.
Publicado: (2026)
Implicit Bias in Noisy-SGD: With Applications to Differentially Private Training
por: Sander, Tom, et al.
Publicado: (2024)
por: Sander, Tom, et al.
Publicado: (2024)
Counterfactual Inference under Thompson Sampling
por: Jeunen, Olivier
Publicado: (2025)
por: Jeunen, Olivier
Publicado: (2025)
On Sampling with Approximate Transport Maps
por: Grenioux, Louis, et al.
Publicado: (2023)
por: Grenioux, Louis, et al.
Publicado: (2023)
Stable Thompson Sampling: Valid Inference via Variance Inflation
por: Halder, Budhaditya, et al.
Publicado: (2025)
por: Halder, Budhaditya, et al.
Publicado: (2025)
Sampling from multi-modal distributions on Riemannian manifolds with training-free stochastic interpolants
por: Durmus, Alain, et al.
Publicado: (2026)
por: Durmus, Alain, et al.
Publicado: (2026)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
por: Janati, Yazid, et al.
Publicado: (2025)
por: Janati, Yazid, et al.
Publicado: (2025)
Stochastic contextual bandits with graph feedback: from independence number to MAS number
por: Wen, Yuxiao, et al.
Publicado: (2024)
por: Wen, Yuxiao, et al.
Publicado: (2024)
Extreme bandits
por: Carpentier, Alexandra, et al.
Publicado: (2026)
por: Carpentier, Alexandra, et al.
Publicado: (2026)
Covariance-adapting algorithm for semi-bandits with application to sparse rewards
por: Perrault, Pierre, et al.
Publicado: (2026)
por: Perrault, Pierre, et al.
Publicado: (2026)
Optimism Stabilizes Thompson Sampling for Adaptive Inference
por: Yan, Shunxing, et al.
Publicado: (2026)
por: Yan, Shunxing, et al.
Publicado: (2026)
Stochastic Localization via Iterative Posterior Sampling
por: Grenioux, Louis, et al.
Publicado: (2024)
por: Grenioux, Louis, et al.
Publicado: (2024)
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning
por: Aravindan, Siddharth, et al.
Publicado: (2025)
por: Aravindan, Siddharth, et al.
Publicado: (2025)
A single algorithm for both restless and rested rotting bandits
por: Seznec, Julien, et al.
Publicado: (2026)
por: Seznec, Julien, et al.
Publicado: (2026)
Divide-and-Conquer Posterior Sampling for Denoising Diffusion Priors
por: Janati, Yazid, et al.
Publicado: (2024)
por: Janati, Yazid, et al.
Publicado: (2024)
Learned Reference-based Diffusion Sampling for multi-modal distributions
por: Noble, Maxence, et al.
Publicado: (2024)
por: Noble, Maxence, et al.
Publicado: (2024)
Thompson Sampling for Repeated Newsvendor
por: Chen, Li, et al.
Publicado: (2025)
por: Chen, Li, et al.
Publicado: (2025)
Constrained Linear Thompson Sampling
por: Gangrade, Aditya, et al.
Publicado: (2025)
por: Gangrade, Aditya, et al.
Publicado: (2025)
Towards Minimax Optimality of Model-based Robust Reinforcement Learning
por: Clavier, Pierre, et al.
Publicado: (2023)
por: Clavier, Pierre, et al.
Publicado: (2023)
Watermarking Makes Language Models Radioactive
por: Sander, Tom, et al.
Publicado: (2024)
por: Sander, Tom, et al.
Publicado: (2024)
Spectral bandits
por: Kocák, Tomáš, et al.
Publicado: (2026)
por: Kocák, Tomáš, et al.
Publicado: (2026)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
por: Namkoong, Hongseok, et al.
Publicado: (2020)
por: Namkoong, Hongseok, et al.
Publicado: (2020)
Regenerative Particle Thompson Sampling
por: Zhou, Zeyu, et al.
Publicado: (2022)
por: Zhou, Zeyu, et al.
Publicado: (2022)
Adaptive Data Augmentation for Thompson Sampling
por: Kim, Wonyoung
Publicado: (2025)
por: Kim, Wonyoung
Publicado: (2025)
A Broader View of Thompson Sampling
por: Qu, Yanlin, et al.
Publicado: (2025)
por: Qu, Yanlin, et al.
Publicado: (2025)
Active clustering with bandit feedback
por: Thuot, Victor, et al.
Publicado: (2024)
por: Thuot, Victor, et al.
Publicado: (2024)
Graph Neural Thompson Sampling
por: Wu, Shuang, et al.
Publicado: (2024)
por: Wu, Shuang, et al.
Publicado: (2024)
Heavy-Tailed Diffusion with Denoising Lévy Probabilistic Models
por: Shariatian, Dario, et al.
Publicado: (2024)
por: Shariatian, Dario, et al.
Publicado: (2024)
Ejemplares similares
-
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
por: Huix, Tom, et al.
Publicado: (2024) -
Variance-sensitive Thompson sampling for generalised linear bandits, revisited
por: Perneczky, Tom, et al.
Publicado: (2026) -
Central Limit Theorem for Bayesian Neural Network trained with Variational Inference
por: Descours, Arnaud, et al.
Publicado: (2024) -
Best-of-Both Worlds for linear contextual bandits with paid observations
por: Boyer, Nathan, et al.
Publicado: (2025) -
Optimal cross-learning for contextual bandits with unknown context distributions
por: Schneider, Jon, et al.
Publicado: (2024)