Diffusion Models as Optimizers for Efficient Planning in Offline RL
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Renming, Pei, Yunqiang, Wang, Guoqing, Zhang, Yangming, Yang, Yang, Wang, Peng, Shen, Hengtao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
Focus On What Matters: Separated Models For Visual-Based RL Generalization
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
A Survey on Efficient Vision-Language-Action Models
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient
von: You, Haoxiang, et al.
Veröffentlicht: (2026)
von: You, Haoxiang, et al.
Veröffentlicht: (2026)
DiMSam: Diffusion Models as Samplers for Task and Motion Planning under Partial Observability
von: Fang, Xiaolin, et al.
Veröffentlicht: (2023)
von: Fang, Xiaolin, et al.
Veröffentlicht: (2023)
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation
von: Xing, Youguang, et al.
Veröffentlicht: (2025)
von: Xing, Youguang, et al.
Veröffentlicht: (2025)
Efficient Adaptation of Pre-trained Vision Transformer underpinned by Approximately Orthogonal Fine-Tuning Strategy
von: Yang, Yiting, et al.
Veröffentlicht: (2025)
von: Yang, Yiting, et al.
Veröffentlicht: (2025)
ExoPredicator: Learning Abstract Models of Dynamic Worlds for Robot Planning
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
Where Bits Matter in World Model Planning: A Paired Mixed-Bit Study for Efficient Spatial Reasoning
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion
von: Zhang, Lunjun, et al.
Veröffentlicht: (2023)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2023)
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
von: Feng, Yao, et al.
Veröffentlicht: (2025)
von: Feng, Yao, et al.
Veröffentlicht: (2025)
DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
ChainFlow-VLA: Causal Flow Planning with Vision-Language Models
von: Wang, Xiyang, et al.
Veröffentlicht: (2026)
von: Wang, Xiyang, et al.
Veröffentlicht: (2026)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
von: Liu, Songming, et al.
Veröffentlicht: (2024)
von: Liu, Songming, et al.
Veröffentlicht: (2024)
Aether: Geometric-Aware Unified World Modeling
von: Aether Team, et al.
Veröffentlicht: (2025)
von: Aether Team, et al.
Veröffentlicht: (2025)
SPA: 3D Spatial-Awareness Enables Effective Embodied Representation
von: Zhu, Haoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Haoyi, et al.
Veröffentlicht: (2024)
Learning Human-Humanoid Coordination for Collaborative Object Carrying
von: Du, Yushi, et al.
Veröffentlicht: (2025)
von: Du, Yushi, et al.
Veröffentlicht: (2025)
CHOrD: Generation of Collision-Free, House-Scale, and Organized Digital Twins for 3D Indoor Scenes with Controllable Floor Plans and Optimal Layouts
von: Su, Chong, et al.
Veröffentlicht: (2025)
von: Su, Chong, et al.
Veröffentlicht: (2025)
Fractional Diffusion Bridge Models
von: Nobis, Gabriel, et al.
Veröffentlicht: (2025)
von: Nobis, Gabriel, et al.
Veröffentlicht: (2025)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
GenRL: Multimodal-foundation world models for generalization in embodied agents
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025)
von: Chun, Junha, et al.
Veröffentlicht: (2025)
E0: Enhancing Generalization and Fine-Grained Control in VLA Models via Tweedie Discrete Diffusion
von: Zhan, Zhihao, et al.
Veröffentlicht: (2025)
von: Zhan, Zhihao, et al.
Veröffentlicht: (2025)
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
von: Huang, Yixuan, et al.
Veröffentlicht: (2023)
von: Huang, Yixuan, et al.
Veröffentlicht: (2023)
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
OMEGA: Efficient Occlusion-Aware Navigation for Air-Ground Robot in Dynamic Environments via State Space Model
von: Wang, Junming, et al.
Veröffentlicht: (2024)
von: Wang, Junming, et al.
Veröffentlicht: (2024)
CLOVER: Closed-Loop Value Estimation and Ranking for End-to-End Autonomous Driving Planning
von: Ang, Sining, et al.
Veröffentlicht: (2026)
von: Ang, Sining, et al.
Veröffentlicht: (2026)
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
Efficient Driving Behavior Narration and Reasoning on Edge Device Using Large Language Models
von: Huang, Yizhou, et al.
Veröffentlicht: (2024)
von: Huang, Yizhou, et al.
Veröffentlicht: (2024)
RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset
von: Wang, Yongzhong, et al.
Veröffentlicht: (2026)
von: Wang, Yongzhong, et al.
Veröffentlicht: (2026)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
von: Fang, Huang, et al.
Veröffentlicht: (2025)
von: Fang, Huang, et al.
Veröffentlicht: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
SIMPL: A Simple and Efficient Multi-agent Motion Prediction Baseline for Autonomous Driving
von: Zhang, Lu, et al.
Veröffentlicht: (2024)
von: Zhang, Lu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024) -
Focus On What Matters: Separated Models For Visual-Based RL Generalization
von: Zhang, Di, et al.
Veröffentlicht: (2024) -
Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026) -
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024) -
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)