DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Yang, Wang, Xiaofeng, Shao, Hao, Wang, Letian, Zhao, Guosheng, Shao, Jiangnan, Zhu, Jiagang, Yu, Tingdong, Zhu, Zheng, Huang, Guan, Waslander, Steven L. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
di: Zhao, Guosheng, et al.
Pubblicazione: (2024)
di: Zhao, Guosheng, et al.
Pubblicazione: (2024)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
di: Zhou, Yang, et al.
Pubblicazione: (2026)
di: Zhou, Yang, et al.
Pubblicazione: (2026)
DriveDreamer4D: World Models Are Effective Data Machines for 4D Driving Scene Representation
di: Zhao, Guosheng, et al.
Pubblicazione: (2024)
di: Zhao, Guosheng, et al.
Pubblicazione: (2024)
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
di: Shao, Hao, et al.
Pubblicazione: (2026)
di: Shao, Hao, et al.
Pubblicazione: (2026)
OpenNav: Open-World Navigation with Multimodal Large Language Models
di: Yuan, Mingfeng, et al.
Pubblicazione: (2025)
di: Yuan, Mingfeng, et al.
Pubblicazione: (2025)
UniDriveDreamer: A Single-Stage Multimodal World Model for Autonomous Driving
di: Zhao, Guosheng, et al.
Pubblicazione: (2026)
di: Zhao, Guosheng, et al.
Pubblicazione: (2026)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
di: Ni, Chaojun, et al.
Pubblicazione: (2024)
di: Ni, Chaojun, et al.
Pubblicazione: (2024)
SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction
di: Zhou, Yang, et al.
Pubblicazione: (2024)
di: Zhou, Yang, et al.
Pubblicazione: (2024)
SmartRefine: A Scenario-Adaptive Refinement Framework for Efficient Motion Prediction
di: Zhou, Yang, et al.
Pubblicazione: (2024)
di: Zhou, Yang, et al.
Pubblicazione: (2024)
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
di: Wang, Boyuan, et al.
Pubblicazione: (2025)
di: Wang, Boyuan, et al.
Pubblicazione: (2025)
ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting
di: Papais, Sandro, et al.
Pubblicazione: (2025)
di: Papais, Sandro, et al.
Pubblicazione: (2025)
ChronoDreamer: Action-Conditioned World Model as an Online Simulator for Robotic Planning
di: Zhou, Zhenhao, et al.
Pubblicazione: (2025)
di: Zhou, Zhenhao, et al.
Pubblicazione: (2025)
Grounded World Model for Semantically Generalizable Planning
di: Li, Quanyi, et al.
Pubblicazione: (2026)
di: Li, Quanyi, et al.
Pubblicazione: (2026)
Nightmare Dreamer: Dreaming About Unsafe States And Planning Ahead
di: Oseni, Oluwatosin, et al.
Pubblicazione: (2026)
di: Oseni, Oluwatosin, et al.
Pubblicazione: (2026)
ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
di: Zhao, Guosheng, et al.
Pubblicazione: (2025)
di: Zhao, Guosheng, et al.
Pubblicazione: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
di: Li, Ying, et al.
Pubblicazione: (2025)
di: Li, Ying, et al.
Pubblicazione: (2025)
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis
di: Lang, Xiaolei, et al.
Pubblicazione: (2026)
di: Lang, Xiaolei, et al.
Pubblicazione: (2026)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
di: Li, Haoyun, et al.
Pubblicazione: (2025)
di: Li, Haoyun, et al.
Pubblicazione: (2025)
DreamerAD: Efficient Reinforcement Learning via Latent World Model for Autonomous Driving
di: Yang, Pengxuan, et al.
Pubblicazione: (2026)
di: Yang, Pengxuan, et al.
Pubblicazione: (2026)
GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control
di: Chen, Anthony, et al.
Pubblicazione: (2025)
di: Chen, Anthony, et al.
Pubblicazione: (2025)
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
di: GigaBrain Team, et al.
Pubblicazione: (2025)
di: GigaBrain Team, et al.
Pubblicazione: (2025)
CarDreamer: Open-Source Learning Platform for World Model based Autonomous Driving
di: Gao, Dechen, et al.
Pubblicazione: (2024)
di: Gao, Dechen, et al.
Pubblicazione: (2024)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
di: Li, Yongkang, et al.
Pubblicazione: (2026)
di: Li, Yongkang, et al.
Pubblicazione: (2026)
Uncertainty-Constrained Differential Dynamic Programming in Belief Space for Vision Based Robots
di: Rahman, Shatil, et al.
Pubblicazione: (2020)
di: Rahman, Shatil, et al.
Pubblicazione: (2020)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
di: jia, Feiyang, et al.
Pubblicazione: (2026)
di: jia, Feiyang, et al.
Pubblicazione: (2026)
RoboTransfer: Controllable Geometry-Consistent Video Diffusion for Manipulation Policy Transfer
di: Liu, Liu, et al.
Pubblicazione: (2025)
di: Liu, Liu, et al.
Pubblicazione: (2025)
SWTrack: Multiple Hypothesis Sliding Window 3D Multi-Object Tracking
di: Papais, Sandro, et al.
Pubblicazione: (2024)
di: Papais, Sandro, et al.
Pubblicazione: (2024)
Uncertainty-Aware Prediction and Application in Planning for Autonomous Driving: Definitions, Methods, and Comparison
di: Shao, Wenbo, et al.
Pubblicazione: (2024)
di: Shao, Wenbo, et al.
Pubblicazione: (2024)
Real-Time Generative Policy via Langevin-Guided Flow Matching for Autonomous Driving
di: Zhu, Tianze, et al.
Pubblicazione: (2026)
di: Zhu, Tianze, et al.
Pubblicazione: (2026)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
di: GigaWorld Team, et al.
Pubblicazione: (2025)
di: GigaWorld Team, et al.
Pubblicazione: (2025)
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
di: Guo, Peizheng, et al.
Pubblicazione: (2026)
di: Guo, Peizheng, et al.
Pubblicazione: (2026)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
aUToPath: Unified Planning and Control for Autonomous Vehicles in Urban Environments Using Hybrid Lattice and Free-Space Search
di: Patel, Tanmay P., et al.
Pubblicazione: (2025)
di: Patel, Tanmay P., et al.
Pubblicazione: (2025)
RynnVLA-002: A Unified Vision-Language-Action and World Model
di: Cen, Jun, et al.
Pubblicazione: (2025)
di: Cen, Jun, et al.
Pubblicazione: (2025)
Online Temporal Fusion for Vectorized Map Construction in Mapless Autonomous Driving
di: Chen, Jiagang, et al.
Pubblicazione: (2024)
di: Chen, Jiagang, et al.
Pubblicazione: (2024)
RoboDreamer: Learning Compositional World Models for Robot Imagination
di: Zhou, Siyuan, et al.
Pubblicazione: (2024)
di: Zhou, Siyuan, et al.
Pubblicazione: (2024)
Trends in Motion Prediction Toward Deployable and Generalizable Autonomy: A Revisit and Perspectives
di: Wang, Letian, et al.
Pubblicazione: (2025)
di: Wang, Letian, et al.
Pubblicazione: (2025)
V-Dreamer: Automating Robotic Simulation and Trajectory Synthesis via Video Generation Priors
di: He, Songjia, et al.
Pubblicazione: (2026)
di: He, Songjia, et al.
Pubblicazione: (2026)
UncertaintyTrack: Exploiting Detection and Localization Uncertainty in Multi-Object Tracking
di: Lee, Chang Won, et al.
Pubblicazione: (2024)
di: Lee, Chang Won, et al.
Pubblicazione: (2024)
VLA-R1: Enhancing Reasoning in Vision-Language-Action Models
di: Ye, Angen, et al.
Pubblicazione: (2025)
di: Ye, Angen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
di: Zhao, Guosheng, et al.
Pubblicazione: (2024) -
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
di: Zhou, Yang, et al.
Pubblicazione: (2026) -
DriveDreamer4D: World Models Are Effective Data Machines for 4D Driving Scene Representation
di: Zhao, Guosheng, et al.
Pubblicazione: (2024) -
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
di: Shao, Hao, et al.
Pubblicazione: (2026) -
OpenNav: Open-World Navigation with Multimodal Large Language Models
di: Yuan, Mingfeng, et al.
Pubblicazione: (2025)