Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
Fuente:
arXiv
Guardado en:
| Autores principales: | Xie, Shaoyuan, Kong, Lingdong, Dong, Yuhao, Sima, Chonghao, Zhang, Wenwei, Chen, Qi Alfred, Liu, Ziwei, Pan, Liang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving
por: Xie, Shaoyuan, et al.
Publicado: (2024)
por: Xie, Shaoyuan, et al.
Publicado: (2024)
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
por: Kong, Lingdong, et al.
Publicado: (2024)
por: Kong, Lingdong, et al.
Publicado: (2024)
LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving
por: Kong, Lingdong, et al.
Publicado: (2025)
por: Kong, Lingdong, et al.
Publicado: (2025)
Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding
por: Kong, Lingdong, et al.
Publicado: (2024)
por: Kong, Lingdong, et al.
Publicado: (2024)
Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining
por: Xu, Xiang, et al.
Publicado: (2025)
por: Xu, Xiang, et al.
Publicado: (2025)
DynamicCity: Large-Scale 4D Occupancy Generation from Dynamic Scenes
por: Bian, Hengwei, et al.
Publicado: (2024)
por: Bian, Hengwei, et al.
Publicado: (2024)
4D Contrastive Superflows are Dense 3D Representation Learners
por: Xu, Xiang, et al.
Publicado: (2024)
por: Xu, Xiang, et al.
Publicado: (2024)
3EED: Ground Everything Everywhere in 3D
por: Li, Rong, et al.
Publicado: (2025)
por: Li, Rong, et al.
Publicado: (2025)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
por: Sima, Chonghao, et al.
Publicado: (2025)
por: Sima, Chonghao, et al.
Publicado: (2025)
Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
por: Hu, Tianshuai, et al.
Publicado: (2025)
por: Hu, Tianshuai, et al.
Publicado: (2025)
LiMoE: Mixture of LiDAR Representation Learners from Automotive Scenes
por: Xu, Xiang, et al.
Publicado: (2025)
por: Xu, Xiang, et al.
Publicado: (2025)
A Comprehensive Study of Bug-Fix Patterns in Autonomous Driving Systems
por: Chen, Yuntianyi, et al.
Publicado: (2025)
por: Chen, Yuntianyi, et al.
Publicado: (2025)
An Empirical Study of Training State-of-the-Art LiDAR Segmentation Models
por: Sun, Jiahao, et al.
Publicado: (2024)
por: Sun, Jiahao, et al.
Publicado: (2024)
BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations
por: Monninger, Thomas, et al.
Publicado: (2026)
por: Monninger, Thomas, et al.
Publicado: (2026)
Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement Learning
por: Li, Yue, et al.
Publicado: (2025)
por: Li, Yue, et al.
Publicado: (2025)
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
por: Jiang, Bo, et al.
Publicado: (2025)
por: Jiang, Bo, et al.
Publicado: (2025)
Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving
por: Tang, Zecong, et al.
Publicado: (2026)
por: Tang, Zecong, et al.
Publicado: (2026)
Open-Source Autonomous Driving Software Platforms: Comparison of Autoware and Apollo
por: Jung, Hee-Yang, et al.
Publicado: (2025)
por: Jung, Hee-Yang, et al.
Publicado: (2025)
Reducing Text Bias in Synthetically Generated MCQAs for VLMs in Autonomous Driving
por: Kulgod, Sutej, et al.
Publicado: (2026)
por: Kulgod, Sutej, et al.
Publicado: (2026)
Scalable Offline Metrics for Autonomous Driving
por: Aich, Animikh, et al.
Publicado: (2025)
por: Aich, Animikh, et al.
Publicado: (2025)
HSImul3R: Physics-in-the-Loop Reconstruction of Simulation-Ready Human-Scene Interactions
por: Cao, Yukang, et al.
Publicado: (2026)
por: Cao, Yukang, et al.
Publicado: (2026)
U4D: Uncertainty-Aware 4D World Modeling from LiDAR Sequences
por: Xu, Xiang, et al.
Publicado: (2025)
por: Xu, Xiang, et al.
Publicado: (2025)
Multi-Space Alignments Towards Universal LiDAR Segmentation
por: Liu, Youquan, et al.
Publicado: (2024)
por: Liu, Youquan, et al.
Publicado: (2024)
PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image
por: Cao, Ziang, et al.
Publicado: (2025)
por: Cao, Ziang, et al.
Publicado: (2025)
VAP: The Vulnerability-Adaptive Protection Paradigm Toward Reliable Autonomous Machines
por: Wan, Zishen, et al.
Publicado: (2024)
por: Wan, Zishen, et al.
Publicado: (2024)
Perspective-Invariant 3D Object Detection
por: Liang, Ao, et al.
Publicado: (2025)
por: Liang, Ao, et al.
Publicado: (2025)
Adaptive Evolution Factor Risk Ellipse Framework for Reliable and Safe Autonomous Driving
por: Yuan, Fujiang, et al.
Publicado: (2025)
por: Yuan, Fujiang, et al.
Publicado: (2025)
Forging Spatial Intelligence: A Roadmap of Multi-Modal Data Pre-Training for Autonomous Systems
por: Wang, Song, et al.
Publicado: (2025)
por: Wang, Song, et al.
Publicado: (2025)
Are We Ready for Real-Time LiDAR Semantic Segmentation in Autonomous Driving?
por: Haidar, Samir Abou, et al.
Publicado: (2024)
por: Haidar, Samir Abou, et al.
Publicado: (2024)
NavThinker: Action-Conditioned World Models for Coupled Prediction and Planning in Social Navigation
por: Hu, Tianshuai, et al.
Publicado: (2026)
por: Hu, Tianshuai, et al.
Publicado: (2026)
Semi-SMD: Semi-Supervised Metric Depth Estimation via Surrounding Cameras for Autonomous Driving
por: Xie, Yusen, et al.
Publicado: (2025)
por: Xie, Yusen, et al.
Publicado: (2025)
ReSim: Reliable World Simulation for Autonomous Driving
por: Yang, Jiazhi, et al.
Publicado: (2025)
por: Yang, Jiazhi, et al.
Publicado: (2025)
Modified-Emergency Index (MEI): A Criticality Metric for Autonomous Driving in Lateral Conflict
por: Cheng, Hao, et al.
Publicado: (2025)
por: Cheng, Hao, et al.
Publicado: (2025)
Kinematics-Aware Latent World Models for Data-Efficient Autonomous Driving
por: Li, Jiazhuo, et al.
Publicado: (2026)
por: Li, Jiazhuo, et al.
Publicado: (2026)
Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis
por: Xu, Xiang, et al.
Publicado: (2026)
por: Xu, Xiang, et al.
Publicado: (2026)
Is Your VLM for Autonomous Driving Safety-Ready? A Comprehensive Benchmark for Evaluating External and In-Cabin Risks
por: Meng, Xianhui, et al.
Publicado: (2025)
por: Meng, Xianhui, et al.
Publicado: (2025)
From Steering to Pedalling: Do Autonomous Driving VLMs Generalize to Cyclist-Assistive Spatial Perception and Planning?
por: Nakka, Krishna Kanth, et al.
Publicado: (2026)
por: Nakka, Krishna Kanth, et al.
Publicado: (2026)
Large Language Model based Interactive Decision-Making for Autonomous Driving
por: Dong, Xinwei, et al.
Publicado: (2026)
por: Dong, Xinwei, et al.
Publicado: (2026)
A Vision-Language-Action Model with Visual Prompt for OFF-Road Autonomous Driving
por: Zhang, Liangdong, et al.
Publicado: (2026)
por: Zhang, Liangdong, et al.
Publicado: (2026)
TeX-NeRF: Neural Radiance Fields from Pseudo-TeX Vision
por: Zhong, Chonghao, et al.
Publicado: (2024)
por: Zhong, Chonghao, et al.
Publicado: (2024)
Ejemplares similares
-
Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving
por: Xie, Shaoyuan, et al.
Publicado: (2024) -
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
por: Kong, Lingdong, et al.
Publicado: (2024) -
LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving
por: Kong, Lingdong, et al.
Publicado: (2025) -
Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding
por: Kong, Lingdong, et al.
Publicado: (2024) -
Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining
por: Xu, Xiang, et al.
Publicado: (2025)