Exploratory Diffusion Model for Unsupervised Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Ying, Chengyang, Chen, Huayu, Zhou, Xinning, Hao, Zhongkai, Su, Hang, Zhu, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
by: Ying, Chengyang, et al.
Published: (2022)
by: Ying, Chengyang, et al.
Published: (2022)
PEAC: Unsupervised Pre-training for Cross-Embodiment Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2024)
by: Ying, Chengyang, et al.
Published: (2024)
On the Reuse Bias in Off-Policy Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2022)
by: Ying, Chengyang, et al.
Published: (2022)
Task Aware Dreamer for Task Generalization in Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2023)
by: Ying, Chengyang, et al.
Published: (2023)
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning
by: Zhou, Xinning, et al.
Published: (2025)
by: Zhou, Xinning, et al.
Published: (2025)
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
Unsupervised Graph Anomaly Detection via Multi-Hypersphere Heterophilic Graph Learning
by: Ni, Hang, et al.
Published: (2025)
by: Ni, Hang, et al.
Published: (2025)
Visual Generation Without Guidance
by: Chen, Huayu, et al.
Published: (2025)
by: Chen, Huayu, et al.
Published: (2025)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
by: Liu, Songming, et al.
Published: (2024)
by: Liu, Songming, et al.
Published: (2024)
Diffusion Models for Reinforcement Learning: A Survey
by: Zhu, Zhengbang, et al.
Published: (2023)
by: Zhu, Zhengbang, et al.
Published: (2023)
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning
by: Wang, Qingjun, et al.
Published: (2026)
by: Wang, Qingjun, et al.
Published: (2026)
Why the Maximum Second Derivative of Activations Matters for Adversarial Robustness
by: Yu, Yunrui, et al.
Published: (2026)
by: Yu, Yunrui, et al.
Published: (2026)
dUltra: Ultra-Fast Diffusion Language Models via Reinforcement Learning
by: Chen, Shirui, et al.
Published: (2025)
by: Chen, Shirui, et al.
Published: (2025)
Diffusion Policies creating a Trust Region for Offline Reinforcement Learning
by: Chen, Tianyu, et al.
Published: (2024)
by: Chen, Tianyu, et al.
Published: (2024)
Transform then Explore: a Simple and Effective Technique for Exploratory Combinatorial Optimization with Reinforcement Learning
by: Pu, Tianle, et al.
Published: (2024)
by: Pu, Tianle, et al.
Published: (2024)
Ranking-aware Reinforcement Learning for Ordinal Ranking
by: Hao, Aiming, et al.
Published: (2026)
by: Hao, Aiming, et al.
Published: (2026)
IPD: Boosting Sequential Policy with Imaginary Planning Distillation in Offline Reinforcement Learning
by: Qin, Yihao, et al.
Published: (2026)
by: Qin, Yihao, et al.
Published: (2026)
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
by: Cui, Ganqu, et al.
Published: (2025)
by: Cui, Ganqu, et al.
Published: (2025)
Guiding Diffusion Models with Reinforcement Learning for Stable Molecule Generation
by: Zhou, Zhijian, et al.
Published: (2025)
by: Zhou, Zhijian, et al.
Published: (2025)
Exploratory Machine Learning with Unknown Unknowns
by: Zhao, Peng, et al.
Published: (2020)
by: Zhao, Peng, et al.
Published: (2020)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
by: Hugessen, Adriana, et al.
Published: (2024)
by: Hugessen, Adriana, et al.
Published: (2024)
DLPO: Diffusion Model Loss-Guided Reinforcement Learning for Fine-Tuning Text-to-Speech Diffusion Models
by: Chen, Jingyi, et al.
Published: (2024)
by: Chen, Jingyi, et al.
Published: (2024)
Subjectivity in Unsupervised Machine Learning Model Selection
by: Chen, Wanyi, et al.
Published: (2023)
by: Chen, Wanyi, et al.
Published: (2023)
Stabilizing Reinforcement Learning for Diffusion Language Models
by: Zhong, Jianyuan, et al.
Published: (2026)
by: Zhong, Jianyuan, et al.
Published: (2026)
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Bridging Large Language Models and Graph Structure Learning Models for Robust Representation Learning
by: Su, Guangxin, et al.
Published: (2024)
by: Su, Guangxin, et al.
Published: (2024)
Learning Design-Score Manifold to Guide Diffusion Models for Offline Optimization
by: Zhou, Tailin, et al.
Published: (2025)
by: Zhou, Tailin, et al.
Published: (2025)
Diffusion Spectral Representation for Reinforcement Learning
by: Shribak, Dmitry, et al.
Published: (2024)
by: Shribak, Dmitry, et al.
Published: (2024)
Overcoming Overfitting in Reinforcement Learning via Gaussian Process Diffusion Policy
by: Horprasert, Amornyos, et al.
Published: (2025)
by: Horprasert, Amornyos, et al.
Published: (2025)
Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
by: Chen, Huayu, et al.
Published: (2024)
by: Chen, Huayu, et al.
Published: (2024)
Diffusion Models for Reinforcement Learning: Foundations, Taxonomy, and Development
by: Xu, Changfu, et al.
Published: (2025)
by: Xu, Changfu, et al.
Published: (2025)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2026)
by: Foffano, Daniele, et al.
Published: (2026)
Adding Conditional Control to Diffusion Models with Reinforcement Learning
by: Zhao, Yulai, et al.
Published: (2024)
by: Zhao, Yulai, et al.
Published: (2024)
Rethinking the Design Space of Reinforcement Learning for Diffusion Models: On the Importance of Likelihood Estimation Beyond Loss Design
by: Choi, Jaemoo, et al.
Published: (2026)
by: Choi, Jaemoo, et al.
Published: (2026)
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving
by: Zhang, Zhihao, et al.
Published: (2025)
by: Zhang, Zhihao, et al.
Published: (2025)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
by: Frans, Kevin, et al.
Published: (2024)
by: Frans, Kevin, et al.
Published: (2024)
Rethinking the Representation in Federated Unsupervised Learning with Non-IID Data
by: Liao, Xinting, et al.
Published: (2024)
by: Liao, Xinting, et al.
Published: (2024)
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
by: Chen, Rufeng, et al.
Published: (2026)
by: Chen, Rufeng, et al.
Published: (2026)
Physics-informed Imitative Reinforcement Learning for Real-world Driving
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Similar Items
-
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
by: Ying, Chengyang, et al.
Published: (2022) -
PEAC: Unsupervised Pre-training for Cross-Embodiment Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2024) -
On the Reuse Bias in Off-Policy Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2022) -
Task Aware Dreamer for Task Generalization in Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2023) -
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning
by: Zhou, Xinning, et al.
Published: (2025)