UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yongkang, Zhou, Lijun, Yan, Sixu, Liao, Bencheng, Yan, Tianyi, Xiong, Kaixin, Chen, Long, Xie, Hongwei, Wang, Bing, Chen, Guang, Ye, Hangjun, Liu, Wenyu, Sun, Haiyang, Wang, Xinggang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World
von: Xia, Tianze, et al.
Veröffentlicht: (2025)
von: Xia, Tianze, et al.
Veröffentlicht: (2025)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning
von: Zou, Jialv, et al.
Veröffentlicht: (2024)
von: Zou, Jialv, et al.
Veröffentlicht: (2024)
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
DriveFine: Refining-Augmented Masked Diffusion VLA for Precise and Robust Driving
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
DiffusionDriveV2: Reinforcement Learning-Constrained Truncated Diffusion Modeling in End-to-End Autonomous Driving
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model
von: Mei, Xiaodong, et al.
Veröffentlicht: (2026)
von: Mei, Xiaodong, et al.
Veröffentlicht: (2026)
LaST-VLA: Thinking in Latent Spatio-Temporal Space for Vision-Language-Action in Autonomous Driving
von: Luo, Yuechen, et al.
Veröffentlicht: (2026)
von: Luo, Yuechen, et al.
Veröffentlicht: (2026)
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
von: Jiang, Bo, et al.
Veröffentlicht: (2025)
von: Jiang, Bo, et al.
Veröffentlicht: (2025)
OmniMamba: Efficient and Unified Multimodal Understanding and Generation via State Space Models
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
MaTVLM: Hybrid Mamba-Transformer for Efficient Vision-Language Modeling
von: Li, Yingyue, et al.
Veröffentlicht: (2025)
von: Li, Yingyue, et al.
Veröffentlicht: (2025)
Rethinking Driving World Model as Synthetic Data Generator for Perception Tasks
von: Zeng, Kai, et al.
Veröffentlicht: (2025)
von: Zeng, Kai, et al.
Veröffentlicht: (2025)
UFO: Unifying Feed-Forward and Optimization-based Methods for Large Driving Scene Modeling
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2026)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
von: Shang, Shuyao, et al.
Veröffentlicht: (2026)
von: Shang, Shuyao, et al.
Veröffentlicht: (2026)
SAMoE-VLA: A Scene Adaptive Mixture-of-Experts Vision-Language-Action Model for Autonomous Driving
von: You, Zihan, et al.
Veröffentlicht: (2026)
von: You, Zihan, et al.
Veröffentlicht: (2026)
MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
WorldSplat: Gaussian-Centric Feed-Forward 4D Scene Generation for Autonomous Driving
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving
von: Huang, Yuzhou, et al.
Veröffentlicht: (2026)
von: Huang, Yuzhou, et al.
Veröffentlicht: (2026)
Mirage: One-Step Video Diffusion for Photorealistic and Coherent Asset Editing in Driving Scenes
von: Wang, Shuyun, et al.
Veröffentlicht: (2025)
von: Wang, Shuyun, et al.
Veröffentlicht: (2025)
UniOcc: A Unified Benchmark for Occupancy Forecasting and Prediction in Autonomous Driving
von: Wang, Yuping, et al.
Veröffentlicht: (2025)
von: Wang, Yuping, et al.
Veröffentlicht: (2025)
Fully Unified Motion Planning for End-to-End Autonomous Driving
von: Liu, Lin, et al.
Veröffentlicht: (2025)
von: Liu, Lin, et al.
Veröffentlicht: (2025)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving
von: Li, Yingyan, et al.
Veröffentlicht: (2025)
von: Li, Yingyan, et al.
Veröffentlicht: (2025)
DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models
von: Zeng, Lunbin, et al.
Veröffentlicht: (2025)
von: Zeng, Lunbin, et al.
Veröffentlicht: (2025)
CoGen: 3D Consistent Video Generation via Adaptive Conditioning for Autonomous Driving
von: Ji, Yishen, et al.
Veröffentlicht: (2025)
von: Ji, Yishen, et al.
Veröffentlicht: (2025)
DriveVA: Video Action Models are Zero-Shot Drivers
von: Liu, Mengmeng, et al.
Veröffentlicht: (2026)
von: Liu, Mengmeng, et al.
Veröffentlicht: (2026)
InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models
von: Tao, Hongyuan, et al.
Veröffentlicht: (2025)
von: Tao, Hongyuan, et al.
Veröffentlicht: (2025)
EvoDriveVLA: Evolving Driving VLA Models via Collaborative Perception-Planning Distillation
von: Cao, Jiajun, et al.
Veröffentlicht: (2026)
von: Cao, Jiajun, et al.
Veröffentlicht: (2026)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
Uni-Gaussians: Unifying Camera and Lidar Simulation with Gaussians for Dynamic Driving Scenarios
von: Yuan, Zikang, et al.
Veröffentlicht: (2025)
von: Yuan, Zikang, et al.
Veröffentlicht: (2025)
StyleVLA: Driving Style-Aware Vision Language Action Model for Autonomous Driving
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
VGGDrive: Empowering Vision-Language Models with Cross-View Geometric Grounding for Autonomous Driving
von: Wang, Jie, et al.
Veröffentlicht: (2026)
von: Wang, Jie, et al.
Veröffentlicht: (2026)
A Unified Perception-Language-Action Framework for Adaptive Autonomous Driving
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025) -
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World
von: Xia, Tianze, et al.
Veröffentlicht: (2025) -
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024) -
MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning
von: Zou, Jialv, et al.
Veröffentlicht: (2024) -
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
von: Jiang, Bo, et al.
Veröffentlicht: (2024)