Efficient Online Reinforcement Learning for Diffusion Policy
Fuente:
arXiv
Guardado en:
| Autores principales: | Ma, Haitong, Chen, Tianyi, Wang, Kai, Li, Na, Dai, Bo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
One-Step Flow Policy Mirror Descent
por: Chen, Tianyi, et al.
Publicado: (2025)
por: Chen, Tianyi, et al.
Publicado: (2025)
Efficient Duple Perturbation Robustness in Low-rank MDPs
por: Hu, Yang, et al.
Publicado: (2024)
por: Hu, Yang, et al.
Publicado: (2024)
Skill Transfer and Discovery for Sim-to-Real Learning: A Representation-Based Viewpoint
por: Ma, Haitong, et al.
Publicado: (2024)
por: Ma, Haitong, et al.
Publicado: (2024)
FlowRL: A Taxonomy and Modular Framework for Reinforcement Learning with Diffusion Policies
por: Gao, Chenxiao, et al.
Publicado: (2026)
por: Gao, Chenxiao, et al.
Publicado: (2026)
Offline Imitation Learning upon Arbitrary Demonstrations by Pre-Training Dynamics Representations
por: Ma, Haitong, et al.
Publicado: (2025)
por: Ma, Haitong, et al.
Publicado: (2025)
Stochastic Nonlinear Control via Finite-dimensional Spectral Dynamic Embedding
por: Ren, Zhaolin, et al.
Publicado: (2023)
por: Ren, Zhaolin, et al.
Publicado: (2023)
Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning
por: Zhang, Chubin, et al.
Publicado: (2025)
por: Zhang, Chubin, et al.
Publicado: (2025)
Primal-Dual Spectral Representation for Off-policy Evaluation
por: Hu, Yang, et al.
Publicado: (2024)
por: Hu, Yang, et al.
Publicado: (2024)
Decentralized Diffusion Policy Learning for Enhanced Exploration in Cooperative Multi-agent Reinforcement Learning
por: Zhang, Yuyang, et al.
Publicado: (2026)
por: Zhang, Yuyang, et al.
Publicado: (2026)
Max-Entropy Reinforcement Learning with Flow Matching and A Case Study on LQR
por: Zhang, Yuyang, et al.
Publicado: (2025)
por: Zhang, Yuyang, et al.
Publicado: (2025)
Diffusion Spectral Representation for Reinforcement Learning
por: Shribak, Dmitry, et al.
Publicado: (2024)
por: Shribak, Dmitry, et al.
Publicado: (2024)
Flow-Based Policy for Online Reinforcement Learning
por: Lv, Lei, et al.
Publicado: (2025)
por: Lv, Lei, et al.
Publicado: (2025)
DiffPoGAN: Diffusion Policies with Generative Adversarial Networks for Offline Reinforcement Learning
por: Hu, Xuemin, et al.
Publicado: (2024)
por: Hu, Xuemin, et al.
Publicado: (2024)
Distributed Thompson sampling under constrained communication
por: Zerefa, Saba, et al.
Publicado: (2024)
por: Zerefa, Saba, et al.
Publicado: (2024)
Spectral Representation-based Reinforcement Learning
por: Gao, Chenxiao, et al.
Publicado: (2025)
por: Gao, Chenxiao, et al.
Publicado: (2025)
Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning
por: Bui, Ha Manh, et al.
Publicado: (2026)
por: Bui, Ha Manh, et al.
Publicado: (2026)
Spectral Ghost in Representation Learning: from Component Analysis to Self-Supervised Learning
por: Dai, Bo, et al.
Publicado: (2026)
por: Dai, Bo, et al.
Publicado: (2026)
Online Estimation and Inference for Robust Policy Evaluation in Reinforcement Learning
por: Liu, Weidong, et al.
Publicado: (2023)
por: Liu, Weidong, et al.
Publicado: (2023)
Reverse Flow Matching: A Unified Framework for Online Reinforcement Learning with Diffusion and Flow Policies
por: Li, Zeyang, et al.
Publicado: (2026)
por: Li, Zeyang, et al.
Publicado: (2026)
Efficient Multi-Policy Evaluation for Reinforcement Learning
por: Liu, Shuze Daniel, et al.
Publicado: (2024)
por: Liu, Shuze Daniel, et al.
Publicado: (2024)
Diffusion Policies for Risk-Averse Behavior Modeling in Offline Reinforcement Learning
por: Chen, Xiaocong, et al.
Publicado: (2024)
por: Chen, Xiaocong, et al.
Publicado: (2024)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
por: Liu, Xu-Hui, et al.
Publicado: (2024)
por: Liu, Xu-Hui, et al.
Publicado: (2024)
Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning
por: Sun, Mingyang, et al.
Publicado: (2025)
por: Sun, Mingyang, et al.
Publicado: (2025)
Efficient Policy Evaluation with Safety Constraint for Reinforcement Learning
por: Chen, Claire, et al.
Publicado: (2024)
por: Chen, Claire, et al.
Publicado: (2024)
Diffusion Policies creating a Trust Region for Offline Reinforcement Learning
por: Chen, Tianyu, et al.
Publicado: (2024)
por: Chen, Tianyu, et al.
Publicado: (2024)
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance
por: He, Jinmin, et al.
Publicado: (2025)
por: He, Jinmin, et al.
Publicado: (2025)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
por: Zhang, Tianle, et al.
Publicado: (2024)
por: Zhang, Tianle, et al.
Publicado: (2024)
Sample-Efficient Policy Constraint Offline Deep Reinforcement Learning based on Sample Filtering
por: Chen, Yuanhao, et al.
Publicado: (2025)
por: Chen, Yuanhao, et al.
Publicado: (2025)
Policy Improvement Reinforcement Learning
por: Wang, Huaiyang, et al.
Publicado: (2026)
por: Wang, Huaiyang, et al.
Publicado: (2026)
GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
por: Ding, Shutong, et al.
Publicado: (2025)
por: Ding, Shutong, et al.
Publicado: (2025)
DiffCPS: Diffusion Model based Constrained Policy Search for Offline Reinforcement Learning
por: He, Longxiang, et al.
Publicado: (2023)
por: He, Longxiang, et al.
Publicado: (2023)
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
por: Chen, Keru, et al.
Publicado: (2024)
por: Chen, Keru, et al.
Publicado: (2024)
RLOMM: An Efficient and Robust Online Map Matching Framework with Reinforcement Learning
por: Chen, Minxiao, et al.
Publicado: (2025)
por: Chen, Minxiao, et al.
Publicado: (2025)
Provable Memory Efficient Self-Play Algorithm for Model-free Reinforcement Learning
por: Li, Na, et al.
Publicado: (2025)
por: Li, Na, et al.
Publicado: (2025)
JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning
por: Lim, Jing Yu, et al.
Publicado: (2026)
por: Lim, Jing Yu, et al.
Publicado: (2026)
Contrastive UCB: Provably Efficient Contrastive Self-Supervised Learning in Online Reinforcement Learning
por: Qiu, Shuang, et al.
Publicado: (2022)
por: Qiu, Shuang, et al.
Publicado: (2022)
DRARL: Disengagement-Reason-Augmented Reinforcement Learning for Efficient Improvement of Autonomous Driving Policy
por: Zhou, Weitao, et al.
Publicado: (2025)
por: Zhou, Weitao, et al.
Publicado: (2025)
Online Matching via Reinforcement Learning: An Expert Policy Orchestration Strategy
por: Mignacco, Chiara, et al.
Publicado: (2025)
por: Mignacco, Chiara, et al.
Publicado: (2025)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
por: Zhang, Tonghe, et al.
Publicado: (2025)
por: Zhang, Tonghe, et al.
Publicado: (2025)
A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
por: Kim, JaeYoon, et al.
Publicado: (2024)
por: Kim, JaeYoon, et al.
Publicado: (2024)
Ejemplares similares
-
One-Step Flow Policy Mirror Descent
por: Chen, Tianyi, et al.
Publicado: (2025) -
Efficient Duple Perturbation Robustness in Low-rank MDPs
por: Hu, Yang, et al.
Publicado: (2024) -
Skill Transfer and Discovery for Sim-to-Real Learning: A Representation-Based Viewpoint
por: Ma, Haitong, et al.
Publicado: (2024) -
FlowRL: A Taxonomy and Modular Framework for Reinforcement Learning with Diffusion Policies
por: Gao, Chenxiao, et al.
Publicado: (2026) -
Offline Imitation Learning upon Arbitrary Demonstrations by Pre-Training Dynamics Representations
por: Ma, Haitong, et al.
Publicado: (2025)