EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Aravindan, Siddharth, Mittal, Dixant, Lee, Wee Sun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Differentiable Tree Search Network
by: Mittal, Dixant, et al.
Published: (2024)
by: Mittal, Dixant, et al.
Published: (2024)
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling
by: Bayrooti, Jasmine, et al.
Published: (2024)
by: Bayrooti, Jasmine, et al.
Published: (2024)
EVaR-Optimal Arm Identification in Bandits
by: Ahmadipour, Mehrasa, et al.
Published: (2025)
by: Ahmadipour, Mehrasa, et al.
Published: (2025)
EUBRL: Epistemic Uncertainty Directed Bayesian Reinforcement Learning
by: Ma, Jianfei, et al.
Published: (2025)
by: Ma, Jianfei, et al.
Published: (2025)
Thompson Sampling-Based Learning and Control for Unknown Dynamic Systems
by: Zheng, Kaikai, et al.
Published: (2025)
by: Zheng, Kaikai, et al.
Published: (2025)
VITS : Variational Inference Thompson Sampling for contextual bandits
by: Clavier, Pierre, et al.
Published: (2023)
by: Clavier, Pierre, et al.
Published: (2023)
Improved Bayesian Regret Bounds for Thompson Sampling in Reinforcement Learning
by: Moradipari, Ahmadreza, et al.
Published: (2023)
by: Moradipari, Ahmadreza, et al.
Published: (2023)
Risk-averse Total-reward MDPs with ERM and EVaR
by: Su, Xihong, et al.
Published: (2024)
by: Su, Xihong, et al.
Published: (2024)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
by: Namkoong, Hongseok, et al.
Published: (2020)
by: Namkoong, Hongseok, et al.
Published: (2020)
Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds
by: Fu, Guoji, et al.
Published: (2026)
by: Fu, Guoji, et al.
Published: (2026)
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
by: Singh, Sagalpreet, et al.
Published: (2025)
by: Singh, Sagalpreet, et al.
Published: (2025)
Continual Reinforcement Learning by Planning with Online World Models
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Sampling-Based Safe Reinforcement Learning
by: Vignola, Luca, et al.
Published: (2026)
by: Vignola, Luca, et al.
Published: (2026)
Online Learning of Decision Trees with Thompson Sampling
by: Chaouki, Ayman, et al.
Published: (2024)
by: Chaouki, Ayman, et al.
Published: (2024)
RASR: Risk-Averse Soft-Robust MDPs with EVaR and Entropic Risk
by: Hau, Jia Lin, et al.
Published: (2022)
by: Hau, Jia Lin, et al.
Published: (2022)
Continual Learning of Numerous Tasks from Long-tail Distributions
by: Kang, Liwei, et al.
Published: (2024)
by: Kang, Liwei, et al.
Published: (2024)
Thompson Sampling for Repeated Newsvendor
by: Chen, Li, et al.
Published: (2025)
by: Chen, Li, et al.
Published: (2025)
Constrained Linear Thompson Sampling
by: Gangrade, Aditya, et al.
Published: (2025)
by: Gangrade, Aditya, et al.
Published: (2025)
Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions
by: Fu, Guoji, et al.
Published: (2025)
by: Fu, Guoji, et al.
Published: (2025)
A Meta Reinforcement Learning Approach to Goals-Based Wealth Management
by: Das, Sanjiv R., et al.
Published: (2026)
by: Das, Sanjiv R., et al.
Published: (2026)
Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift
by: Li, Bochao, et al.
Published: (2026)
by: Li, Bochao, et al.
Published: (2026)
Regenerative Particle Thompson Sampling
by: Zhou, Zeyu, et al.
Published: (2022)
by: Zhou, Zeyu, et al.
Published: (2022)
Adaptive Data Augmentation for Thompson Sampling
by: Kim, Wonyoung
Published: (2025)
by: Kim, Wonyoung
Published: (2025)
A Broader View of Thompson Sampling
by: Qu, Yanlin, et al.
Published: (2025)
by: Qu, Yanlin, et al.
Published: (2025)
Graph Neural Thompson Sampling
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
Distributionally Robust Model-Based Offline Reinforcement Learning with Near-Optimal Sample Complexity
by: Shi, Laixi, et al.
Published: (2022)
by: Shi, Laixi, et al.
Published: (2022)
Fast, Precise Thompson Sampling for Bayesian Optimization
by: Sweet, David
Published: (2024)
by: Sweet, David
Published: (2024)
Thompson Sampling in Partially Observable Contextual Bandits
by: Park, Hongju, et al.
Published: (2024)
by: Park, Hongju, et al.
Published: (2024)
On Regret Bounds of Thompson Sampling for Bayesian Optimization
by: Takeno, Shion, et al.
Published: (2026)
by: Takeno, Shion, et al.
Published: (2026)
Agnostic Learning of Arbitrary ReLU Activation under Gaussian Marginals
by: Guo, Anxin, et al.
Published: (2024)
by: Guo, Anxin, et al.
Published: (2024)
Fast Online Learning with Gaussian Prior-Driven Hierarchical Unimodal Thompson Sampling
by: Zhao, Tianchi, et al.
Published: (2026)
by: Zhao, Tianchi, et al.
Published: (2026)
Thompson Sampling For Combinatorial Bandits: Polynomial Regret and Mismatched Sampling Paradox
by: Zhang, Raymond, et al.
Published: (2024)
by: Zhang, Raymond, et al.
Published: (2024)
Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow
by: Chao, Chen-Hao, et al.
Published: (2024)
by: Chao, Chen-Hao, et al.
Published: (2024)
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation
by: Lee, Donghoon, et al.
Published: (2025)
by: Lee, Donghoon, et al.
Published: (2025)
Locality Sensitive Sparse Encoding for Learning World Models Online
by: Liu, Zichen, et al.
Published: (2024)
by: Liu, Zichen, et al.
Published: (2024)
On Rollouts in Model-Based Reinforcement Learning
by: Frauenknecht, Bernd, et al.
Published: (2025)
by: Frauenknecht, Bernd, et al.
Published: (2025)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
by: Park, Somangchan, et al.
Published: (2025)
by: Park, Somangchan, et al.
Published: (2025)
On Thompson Sampling and Bilateral Uncertainty in Additive Bayesian Optimization
by: Wycoff, Nathan
Published: (2025)
by: Wycoff, Nathan
Published: (2025)
Thompson Sampling in Online RLHF with General Function Approximation
by: Feng, Songtao, et al.
Published: (2025)
by: Feng, Songtao, et al.
Published: (2025)
Thompson Sampling-like Algorithms for Stochastic Rising Bandits
by: Fiandri, Marco, et al.
Published: (2025)
by: Fiandri, Marco, et al.
Published: (2025)
Similar Items
-
Differentiable Tree Search Network
by: Mittal, Dixant, et al.
Published: (2024) -
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling
by: Bayrooti, Jasmine, et al.
Published: (2024) -
EVaR-Optimal Arm Identification in Bandits
by: Ahmadipour, Mehrasa, et al.
Published: (2025) -
EUBRL: Epistemic Uncertainty Directed Bayesian Reinforcement Learning
by: Ma, Jianfei, et al.
Published: (2025) -
Thompson Sampling-Based Learning and Control for Unknown Dynamic Systems
by: Zheng, Kaikai, et al.
Published: (2025)