VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Yi, Hu, Yuxin, Zhang, Zaiwei, Meyer, Gregory P., Mustikovela, Siva Karthik, Srinivasa, Siddhartha, Wolff, Eric M., Huang, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLMine: Long-Tail Data Mining with Vision Language Models
by: Ye, Mao, et al.
Published: (2024)
by: Ye, Mao, et al.
Published: (2024)
WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model
by: Zhang, Songyan, et al.
Published: (2024)
by: Zhang, Songyan, et al.
Published: (2024)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024)
by: Zhang, Zaiwei, et al.
Published: (2024)
GenAD: Generative End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024)
by: Zheng, Wenzhao, et al.
Published: (2024)
AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models
by: Han, Yuxuan, et al.
Published: (2026)
by: Han, Yuxuan, et al.
Published: (2026)
GaussianAD: Gaussian-Centric End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024)
by: Zheng, Wenzhao, et al.
Published: (2024)
E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving
by: Tang, Yihong, et al.
Published: (2025)
by: Tang, Yihong, et al.
Published: (2025)
SOLVE: Synergy of Language-Vision and End-to-End Networks for Autonomous Driving
by: Chen, Xuesong, et al.
Published: (2025)
by: Chen, Xuesong, et al.
Published: (2025)
V2X-VLM: End-to-End V2X Cooperative Autonomous Driving Through Large Vision-Language Models
by: You, Junwei, et al.
Published: (2024)
by: You, Junwei, et al.
Published: (2024)
RS2AD: End-to-End Autonomous Driving Data Generation from Roadside Sensor Observations
by: Xing, Ruidan, et al.
Published: (2025)
by: Xing, Ruidan, et al.
Published: (2025)
FocalAD: Local Motion Planning for End-to-End Autonomous Driving
by: Sun, Bin, et al.
Published: (2025)
by: Sun, Bin, et al.
Published: (2025)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
by: Ding, Kairui, et al.
Published: (2024)
by: Ding, Kairui, et al.
Published: (2024)
SimpleLLM4AD: An End-to-End Vision-Language Model with Graph Visual Question Answering for Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2024)
by: Zheng, Peiru, et al.
Published: (2024)
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025)
by: Qin, Jack, et al.
Published: (2025)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
by: Jiang, Bo, et al.
Published: (2024)
by: Jiang, Bo, et al.
Published: (2024)
BeLLA: End-to-End Birds Eye View Large Language Assistant for Autonomous Driving
by: Mohan, Karthik, et al.
Published: (2025)
by: Mohan, Karthik, et al.
Published: (2025)
VLM-E2E: Enhancing End-to-End Autonomous Driving with Multimodal Driver Attention Fusion
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving
by: Jiang, Hao, et al.
Published: (2025)
by: Jiang, Hao, et al.
Published: (2025)
ResAD: Normalized Residual Trajectory Modeling for End-to-End Autonomous Driving
by: Zheng, Zhiyu, et al.
Published: (2025)
by: Zheng, Zhiyu, et al.
Published: (2025)
ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving
by: Lu, Han, et al.
Published: (2024)
by: Lu, Han, et al.
Published: (2024)
SimpleVSF: VLM-Scoring Fusion for Trajectory Prediction of End-to-End Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2025)
by: Zheng, Peiru, et al.
Published: (2025)
SparseAD: Sparse Query-Centric Paradigm for Efficient End-to-End Autonomous Driving
by: Zhang, Diankun, et al.
Published: (2024)
by: Zhang, Diankun, et al.
Published: (2024)
SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data Integration
by: Kim, Jongsuk, et al.
Published: (2025)
by: Kim, Jongsuk, et al.
Published: (2025)
LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving
by: Song, Nan, et al.
Published: (2025)
by: Song, Nan, et al.
Published: (2025)
ReAL-AD: Towards Human-Like Reasoning in End-to-End Autonomous Driving
by: Lu, Yuhang, et al.
Published: (2025)
by: Lu, Yuhang, et al.
Published: (2025)
SUPER-AD: Semantic Uncertainty-aware Planning for End-to-End Robust Autonomous Driving
by: Ryu, Wonjeong, et al.
Published: (2025)
by: Ryu, Wonjeong, et al.
Published: (2025)
Prioritizing Perception-Guided Self-Supervision: A New Paradigm for Causal Modeling in End-to-End Autonomous Driving
by: Huang, Yi, et al.
Published: (2025)
by: Huang, Yi, et al.
Published: (2025)
GraphAD: Interaction Scene Graph for End-to-end Autonomous Driving
by: Zhang, Yunpeng, et al.
Published: (2024)
by: Zhang, Yunpeng, et al.
Published: (2024)
AD-R1: Closed-Loop Reinforcement Learning for End-to-End Autonomous Driving with Impartial World Models
by: Yan, Tianyi, et al.
Published: (2025)
by: Yan, Tianyi, et al.
Published: (2025)
VECTOR-Drive: Tightly Coupled Vision-Language and Trajectory Expert Routing for End-to-End Autonomous Driving
by: Zhao, Rui, et al.
Published: (2026)
by: Zhao, Rui, et al.
Published: (2026)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
by: Yang, Zhenjie, et al.
Published: (2025)
by: Yang, Zhenjie, et al.
Published: (2025)
SpaRC-AD: A Baseline for Radar-Camera Fusion in End-to-End Autonomous Driving
by: Wolters, Philipp, et al.
Published: (2025)
by: Wolters, Philipp, et al.
Published: (2025)
Flash3D: Super-scaling Point Transformers through Joint Hardware-Geometry Locality
by: Chen, Liyan, et al.
Published: (2024)
by: Chen, Liyan, et al.
Published: (2024)
Towards Collaborative Autonomous Driving: Simulation Platform and End-to-End System
by: Liu, Genjia, et al.
Published: (2024)
by: Liu, Genjia, et al.
Published: (2024)
UniSTPA: A Safety Analysis Framework for End-to-End Autonomous Driving
by: Kou, Hongrui, et al.
Published: (2025)
by: Kou, Hongrui, et al.
Published: (2025)
ComDrive: Comfort-Oriented End-to-End Autonomous Driving
by: Wang, Junming, et al.
Published: (2024)
by: Wang, Junming, et al.
Published: (2024)
Exploring the Causality of End-to-End Autonomous Driving
by: Li, Jiankun, et al.
Published: (2024)
by: Li, Jiankun, et al.
Published: (2024)
Efficient and Explainable End-to-End Autonomous Driving via Masked Vision-Language-Action Diffusion
by: Zhang, Jiaru, et al.
Published: (2026)
by: Zhang, Jiaru, et al.
Published: (2026)
2nd Place Solution for CVPR2024 E2E Challenge: End-to-End Autonomous Driving Using Vision Language Model
by: Guo, Zilong, et al.
Published: (2025)
by: Guo, Zilong, et al.
Published: (2025)
AD$^2$: Analysis and Detection of Adversarial Threats in Visual Perception for End-to-End Autonomous Driving Systems
by: Sahu, Ishan, et al.
Published: (2026)
by: Sahu, Ishan, et al.
Published: (2026)
Similar Items
-
VLMine: Long-Tail Data Mining with Vision Language Models
by: Ye, Mao, et al.
Published: (2024) -
WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model
by: Zhang, Songyan, et al.
Published: (2024) -
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024) -
GenAD: Generative End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024) -
AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models
by: Han, Yuxuan, et al.
Published: (2026)