DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Chen, Xu, Jinrui, Shi, Shaoshuai, Sheng, Kehua, Zhang, Bo, Jiang, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
von: Shi, Chen, et al.
Veröffentlicht: (2025)
von: Shi, Chen, et al.
Veröffentlicht: (2025)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene Reconstruction
von: Shi, Chen, et al.
Veröffentlicht: (2025)
von: Shi, Chen, et al.
Veröffentlicht: (2025)
Percept-WAM: Perception-Enhanced World-Awareness-Action Model for Robust End-to-End Autonomous Driving
von: Han, Jianhua, et al.
Veröffentlicht: (2025)
von: Han, Jianhua, et al.
Veröffentlicht: (2025)
JiSAM: Alleviate Labeling Burden and Corner Case Problems in Autonomous Driving via Minimal Real-World Data
von: Chen, Runjian, et al.
Veröffentlicht: (2025)
von: Chen, Runjian, et al.
Veröffentlicht: (2025)
PriorFusion: Unified Integration of Priors for Robust Road Perception in Autonomous Driving
von: Tang, Xuewei, et al.
Veröffentlicht: (2025)
von: Tang, Xuewei, et al.
Veröffentlicht: (2025)
OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models
von: Zhang, Yiwei, et al.
Veröffentlicht: (2026)
von: Zhang, Yiwei, et al.
Veröffentlicht: (2026)
SOLVE: Synergy of Language-Vision and End-to-End Networks for Autonomous Driving
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT
von: Hu, Xiaotao, et al.
Veröffentlicht: (2024)
von: Hu, Xiaotao, et al.
Veröffentlicht: (2024)
ColaVLA: Leveraging Cognitive Latent Reasoning for Hierarchical Parallel Trajectory Planning in Autonomous Driving
von: Peng, Qihang, et al.
Veröffentlicht: (2025)
von: Peng, Qihang, et al.
Veröffentlicht: (2025)
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
von: Xiong, Zhexiao, et al.
Veröffentlicht: (2026)
von: Xiong, Zhexiao, et al.
Veröffentlicht: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
WAM-Flow: Parallel Coarse-to-Fine Motion Planning via Discrete Flow Matching for Autonomous Driving
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
WAM-Diff: A Masked Diffusion VLA Framework with MoE and Online Reinforcement Learning for Autonomous Driving
von: Xu, Mingwang, et al.
Veröffentlicht: (2025)
von: Xu, Mingwang, et al.
Veröffentlicht: (2025)
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
von: Ren, Xuanchi, et al.
Veröffentlicht: (2025)
von: Ren, Xuanchi, et al.
Veröffentlicht: (2025)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
ConsisDrive: Identity-Preserving Driving World Models for Video Generation by Instance Mask
von: Yang, Zhuoran, et al.
Veröffentlicht: (2026)
von: Yang, Zhuoran, et al.
Veröffentlicht: (2026)
DriveFuture: Future-Aware Latent World Models for Autonomous Driving
von: Hong, Yufeng, et al.
Veröffentlicht: (2026)
von: Hong, Yufeng, et al.
Veröffentlicht: (2026)
Map-World: Masked Action planning and Path-Integral World Model for Autonomous Driving
von: Hu, Bin, et al.
Veröffentlicht: (2025)
von: Hu, Bin, et al.
Veröffentlicht: (2025)
LIVE: Long-horizon Interactive Video World Modeling
von: Huang, Junchao, et al.
Veröffentlicht: (2026)
von: Huang, Junchao, et al.
Veröffentlicht: (2026)
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving
von: Wei, Julong, et al.
Veröffentlicht: (2024)
von: Wei, Julong, et al.
Veröffentlicht: (2024)
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
Learning Vision-Language-Action World Models for Autonomous Driving
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
InstaDrive: Instance-Aware Driving World Models for Realistic and Consistent Video Generation
von: Yang, Zhuoran, et al.
Veröffentlicht: (2026)
von: Yang, Zhuoran, et al.
Veröffentlicht: (2026)
M3Net: Multimodal Multi-task Learning for 3D Detection, Segmentation, and Occupancy Prediction in Autonomous Driving
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving
von: Sheng, Zihao, et al.
Veröffentlicht: (2026)
von: Sheng, Zihao, et al.
Veröffentlicht: (2026)
DriveWorld: 4D Pre-trained Scene Understanding via World Models for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2024)
von: Min, Chen, et al.
Veröffentlicht: (2024)
DriveCamSim: Generalizable Camera Simulation via Explicit Camera Modeling for Autonomous Driving
von: Sun, Wenchao, et al.
Veröffentlicht: (2025)
von: Sun, Wenchao, et al.
Veröffentlicht: (2025)
E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving
von: Tang, Yihong, et al.
Veröffentlicht: (2025)
von: Tang, Yihong, et al.
Veröffentlicht: (2025)
Are AI-Generated Driving Videos Ready for Autonomous Driving? A Diagnostic Evaluation Framework
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving
von: Wen, Yuqing, et al.
Veröffentlicht: (2024)
von: Wen, Yuqing, et al.
Veröffentlicht: (2024)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
HybridWorldSim: A Scalable and Controllable High-fidelity Simulator for Autonomous Driving
von: Li, Qiang, et al.
Veröffentlicht: (2025)
von: Li, Qiang, et al.
Veröffentlicht: (2025)
MindDrive: An All-in-One Framework Bridging World Models and Vision-Language Model for End-to-End Autonomous Driving
von: Sun, Bin, et al.
Veröffentlicht: (2025)
von: Sun, Bin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
von: Shi, Chen, et al.
Veröffentlicht: (2025) -
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
von: Wang, Linbo, et al.
Veröffentlicht: (2026) -
UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene Reconstruction
von: Shi, Chen, et al.
Veröffentlicht: (2025) -
Percept-WAM: Perception-Enhanced World-Awareness-Action Model for Robust End-to-End Autonomous Driving
von: Han, Jianhua, et al.
Veröffentlicht: (2025) -
JiSAM: Alleviate Labeling Burden and Corner Case Problems in Autonomous Driving via Minimal Real-World Data
von: Chen, Runjian, et al.
Veröffentlicht: (2025)