SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Peizheng, Zhang, Zhenghao, Holtz, David, Yu, Hang, Yang, Yutong, Lai, Yuzhi, Song, Rui, Geiger, Andreas, Zell, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SEER-VAR: Semantic Egocentric Environment Reasoner for Vehicle Augmented Reality
by: Lai, Yuzhi, et al.
Published: (2025)
by: Lai, Yuzhi, et al.
Published: (2025)
Sticky-Glance: Robust Intent Recognition for Human Robot Collaboration via Single-Glance
by: Lai, Yuzhi, et al.
Published: (2026)
by: Lai, Yuzhi, et al.
Published: (2026)
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
by: Tian, Xiaoyu, et al.
Published: (2024)
by: Tian, Xiaoyu, et al.
Published: (2024)
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026)
by: Gerstenecker, Simon, et al.
Published: (2026)
DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
by: Zheng, Weicheng, et al.
Published: (2025)
by: Zheng, Weicheng, et al.
Published: (2025)
LIX: Implicitly Infusing Spatial Geometric Prior Knowledge into Visual Semantic Segmentation for Autonomous Driving
by: Guo, Sicen, et al.
Published: (2024)
by: Guo, Sicen, et al.
Published: (2024)
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)
by: Chen, Li, et al.
Published: (2023)
SeFlow: A Self-Supervised Scene Flow Method in Autonomous Driving
by: Zhang, Qingwen, et al.
Published: (2024)
by: Zhang, Qingwen, et al.
Published: (2024)
ReSim: Reliable World Simulation for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2025)
by: Yang, Jiazhi, et al.
Published: (2025)
PlanT 2.0: Exposing Biases and Structural Flaws in Closed-Loop Driving
by: Gerstenecker, Simon, et al.
Published: (2025)
by: Gerstenecker, Simon, et al.
Published: (2025)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
by: Feng, Bowen, et al.
Published: (2025)
by: Feng, Bowen, et al.
Published: (2025)
CoDa-4DGS: Dynamic Gaussian Splatting with Context and Deformation Awareness for Autonomous Driving
by: Song, Rui, et al.
Published: (2025)
by: Song, Rui, et al.
Published: (2025)
DriveFuture: Future-Aware Latent World Models for Autonomous Driving
by: Hong, Yufeng, et al.
Published: (2026)
by: Hong, Yufeng, et al.
Published: (2026)
Spatial Retrieval Augmented Autonomous Driving
by: Jia, Xiaosong, et al.
Published: (2025)
by: Jia, Xiaosong, et al.
Published: (2025)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
by: Fu, Yongjie, et al.
Published: (2024)
by: Fu, Yongjie, et al.
Published: (2024)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024)
by: Chitta, Kashyap, et al.
Published: (2024)
GenAD: Generalized Predictive Model for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2024)
by: Yang, Jiazhi, et al.
Published: (2024)
VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events
by: Bhat, Mohammad Qazim, et al.
Published: (2026)
by: Bhat, Mohammad Qazim, et al.
Published: (2026)
Comparative Analysis of Patch Attack on VLM-Based Autonomous Driving Architectures
by: Fernandez, David, et al.
Published: (2026)
by: Fernandez, David, et al.
Published: (2026)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
by: Sima, Chonghao, et al.
Published: (2025)
by: Sima, Chonghao, et al.
Published: (2025)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
by: Chen, Zeming, et al.
Published: (2025)
by: Chen, Zeming, et al.
Published: (2025)
Pseudo-Simulation for Autonomous Driving
by: Cao, Wei, et al.
Published: (2025)
by: Cao, Wei, et al.
Published: (2025)
3D Object Detection for Autonomous Driving: A Survey
by: Qian, Rui, et al.
Published: (2021)
by: Qian, Rui, et al.
Published: (2021)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
by: Wu, Yuzhi, et al.
Published: (2024)
by: Wu, Yuzhi, et al.
Published: (2024)
Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving
by: Nie, Ming, et al.
Published: (2023)
by: Nie, Ming, et al.
Published: (2023)
StyleDrive: Towards Driving-Style Aware Benchmarking of End-To-End Autonomous Driving
by: Hao, Ruiyang, et al.
Published: (2025)
by: Hao, Ruiyang, et al.
Published: (2025)
DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving
by: Su, Haisheng, et al.
Published: (2026)
by: Su, Haisheng, et al.
Published: (2026)
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025)
by: Qin, Jack, et al.
Published: (2025)
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024)
by: Zimmerlin, Julian, et al.
Published: (2024)
DriveLM: Driving with Graph Visual Question Answering
by: Sima, Chonghao, et al.
Published: (2023)
by: Sima, Chonghao, et al.
Published: (2023)
HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
LaVida Drive: Vision-Text Interaction VLM for Autonomous Driving with Token Selection, Recovery and Enhancement
by: Jiao, Siwen, et al.
Published: (2024)
by: Jiao, Siwen, et al.
Published: (2024)
DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
by: He, Yuankai, et al.
Published: (2025)
by: He, Yuankai, et al.
Published: (2025)
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
by: Lai, Yuzhi, et al.
Published: (2025)
by: Lai, Yuzhi, et al.
Published: (2025)
Real-time Traffic Object Detection for Autonomous Driving
by: Khan, Abdul Hannan, et al.
Published: (2024)
by: Khan, Abdul Hannan, et al.
Published: (2024)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
by: Dauner, Daniel, et al.
Published: (2026)
by: Dauner, Daniel, et al.
Published: (2026)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
by: Huang, Zilin, et al.
Published: (2026)
by: Huang, Zilin, et al.
Published: (2026)
VDT-Auto: End-to-end Autonomous Driving with VLM-Guided Diffusion Transformers
by: Guo, Ziang, et al.
Published: (2025)
by: Guo, Ziang, et al.
Published: (2025)
Similar Items
-
SEER-VAR: Semantic Egocentric Environment Reasoner for Vehicle Augmented Reality
by: Lai, Yuzhi, et al.
Published: (2025) -
Sticky-Glance: Robust Intent Recognition for Human Robot Collaboration via Single-Glance
by: Lai, Yuzhi, et al.
Published: (2026) -
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
by: Tian, Xiaoyu, et al.
Published: (2024) -
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026) -
DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
by: Zheng, Weicheng, et al.
Published: (2025)