Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Bui, Ha Manh, Jazbec, Metod, Nalisnick, Eric, Liu, Anqi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Uncertainty in Diffusion Models
by: Jazbec, Metod, et al.
Published: (2025)
by: Jazbec, Metod, et al.
Published: (2025)
Density-Regression: Efficient and Distance-Aware Deep Regressor for Uncertainty Estimation under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2024)
by: Bui, Ha Manh, et al.
Published: (2024)
Monitoring Risks in Test-Time Adaptation
by: Schirmer, Mona, et al.
Published: (2025)
by: Schirmer, Mona, et al.
Published: (2025)
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2023)
by: Bui, Ha Manh, et al.
Published: (2023)
Early-Exit Neural Networks with Nested Prediction Sets
by: Jazbec, Metod, et al.
Published: (2023)
by: Jazbec, Metod, et al.
Published: (2023)
Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting
by: Wynn, Andrea, et al.
Published: (2025)
by: Wynn, Andrea, et al.
Published: (2025)
Q-Learning with Shift-Aware Upper Confidence Bound in Non-Stationary Reinforcement Learning
by: Bui, Ha Manh, et al.
Published: (2025)
by: Bui, Ha Manh, et al.
Published: (2025)
Calibrated Uncertainty Sampling for Active Learning
by: Bui, Ha Manh, et al.
Published: (2025)
by: Bui, Ha Manh, et al.
Published: (2025)
A Tale of Two Temperatures: Simple, Efficient, and Diverse Sampling from Diffusion Language Models
by: Olausson, Theo X., et al.
Published: (2026)
by: Olausson, Theo X., et al.
Published: (2026)
Variance-Aware Linear UCB with Deep Representation for Neural Contextual Bandits
by: Bui, Ha Manh, et al.
Published: (2024)
by: Bui, Ha Manh, et al.
Published: (2024)
Fast yet Safe: Early-Exiting with Risk Control
by: Jazbec, Metod, et al.
Published: (2024)
by: Jazbec, Metod, et al.
Published: (2024)
Learning Unmasking Policies for Diffusion Language Models
by: Jazbec, Metod, et al.
Published: (2025)
by: Jazbec, Metod, et al.
Published: (2025)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
by: Liu, Xu-Hui, et al.
Published: (2024)
by: Liu, Xu-Hui, et al.
Published: (2024)
A Simple Unified Uncertainty-Guided Framework for Offline-to-Online Reinforcement Learning
by: Guo, Siyuan, et al.
Published: (2023)
by: Guo, Siyuan, et al.
Published: (2023)
Towards Robust Offline-to-Online Reinforcement Learning via Uncertainty and Smoothness
by: Wen, Xiaoyu, et al.
Published: (2023)
by: Wen, Xiaoyu, et al.
Published: (2023)
Dynamic Vocabulary Pruning in Early-Exit LLMs
by: Vincenti, Jort, et al.
Published: (2024)
by: Vincenti, Jort, et al.
Published: (2024)
Offline-to-Online Reinforcement Learning with Classifier-Free Diffusion Generation
by: Huang, Xiao, et al.
Published: (2025)
by: Huang, Xiao, et al.
Published: (2025)
Uncertainty Aware Tropical Cyclone Wind Speed Estimation from Satellite Data
by: Lehmann, Nils, et al.
Published: (2024)
by: Lehmann, Nils, et al.
Published: (2024)
Uncertainty-Aware Rank-One MIMO Q Network Framework for Accelerated Offline Reinforcement Learning
by: Nguyen, Thanh, et al.
Published: (2026)
by: Nguyen, Thanh, et al.
Published: (2026)
MOORL: A Framework for Integrating Offline-Online Reinforcement Learning
by: Chaudhary, Gaurav, et al.
Published: (2025)
by: Chaudhary, Gaurav, et al.
Published: (2025)
On Calibration in Multi-Distribution Learning
by: Verma, Rajeev, et al.
Published: (2024)
by: Verma, Rajeev, et al.
Published: (2024)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
by: Guo, Yihong, et al.
Published: (2025)
by: Guo, Yihong, et al.
Published: (2025)
Online Pre-Training for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2025)
by: Shin, Yongjae, et al.
Published: (2025)
Information-Directed Offline-to-Online Reinforcement Learning
by: Chen, Keru
Published: (2026)
by: Chen, Keru
Published: (2026)
Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data
by: Zhou, Zhiyuan, et al.
Published: (2024)
by: Zhou, Zhiyuan, et al.
Published: (2024)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
by: Panaganti, Kishan, et al.
Published: (2024)
by: Panaganti, Kishan, et al.
Published: (2024)
Efficient Online Reinforcement Learning for Diffusion Policy
by: Ma, Haitong, et al.
Published: (2025)
by: Ma, Haitong, et al.
Published: (2025)
Active Advantage-Aligned Online Reinforcement Learning with Offline Data
by: Liu, Xuefeng, et al.
Published: (2025)
by: Liu, Xuefeng, et al.
Published: (2025)
Bayesian Design Principles for Offline-to-Online Reinforcement Learning
by: Hu, Hao, et al.
Published: (2024)
by: Hu, Hao, et al.
Published: (2024)
Online Optimization for Offline Safe Reinforcement Learning
by: Chemingui, Yassine, et al.
Published: (2025)
by: Chemingui, Yassine, et al.
Published: (2025)
The Three Regimes of Offline-to-Online Reinforcement Learning
by: Li, Lu, et al.
Published: (2025)
by: Li, Lu, et al.
Published: (2025)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
by: Wang, Qi, et al.
Published: (2023)
by: Wang, Qi, et al.
Published: (2023)
Lightning UQ Box: A Comprehensive Framework for Uncertainty Quantification in Deep Learning
by: Lehmann, Nils, et al.
Published: (2024)
by: Lehmann, Nils, et al.
Published: (2024)
When to Trust Your Simulator: Dynamics-Aware Hybrid Offline-and-Online Reinforcement Learning
by: Niu, Haoyi, et al.
Published: (2022)
by: Niu, Haoyi, et al.
Published: (2022)
SUMO: Search-Based Uncertainty Estimation for Model-Based Offline Reinforcement Learning
by: Qiao, Zhongjian, et al.
Published: (2024)
by: Qiao, Zhongjian, et al.
Published: (2024)
Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning
by: Gireesh, Nandiraju, et al.
Published: (2026)
by: Gireesh, Nandiraju, et al.
Published: (2026)
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2023)
by: McInroe, Trevor, et al.
Published: (2023)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
by: Han, Zean, et al.
Published: (2026)
by: Han, Zean, et al.
Published: (2026)
Scalable Generative Modeling of Weighted Graphs
by: Williams, Richard, et al.
Published: (2025)
by: Williams, Richard, et al.
Published: (2025)
Adaptive Bounding Box Uncertainties via Two-Step Conformal Prediction
by: Timans, Alexander, et al.
Published: (2024)
by: Timans, Alexander, et al.
Published: (2024)
Similar Items
-
Generative Uncertainty in Diffusion Models
by: Jazbec, Metod, et al.
Published: (2025) -
Density-Regression: Efficient and Distance-Aware Deep Regressor for Uncertainty Estimation under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2024) -
Monitoring Risks in Test-Time Adaptation
by: Schirmer, Mona, et al.
Published: (2025) -
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2023) -
Early-Exit Neural Networks with Nested Prediction Sets
by: Jazbec, Metod, et al.
Published: (2023)