DriveLM: Driving with Graph Visual Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Sima, Chonghao, Renz, Katrin, Chitta, Kashyap, Chen, Li, Zhang, Hanxue, Xie, Chengen, Beißwenger, Jens, Luo, Ping, Geiger, Andreas, Li, Hongyang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024)
by: Zimmerlin, Julian, et al.
Published: (2024)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
by: Sima, Chonghao, et al.
Published: (2025)
by: Sima, Chonghao, et al.
Published: (2025)
FLARE: Learning Future-Aware Latent Representations from Vision-Language Models for Autonomous Driving
by: Xie, Chengen, et al.
Published: (2026)
by: Xie, Chengen, et al.
Published: (2026)
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026)
by: Gerstenecker, Simon, et al.
Published: (2026)
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)
by: Chen, Li, et al.
Published: (2023)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024)
by: Chitta, Kashyap, et al.
Published: (2024)
PlanT 2.0: Exposing Biases and Structural Flaws in Closed-Loop Driving
by: Gerstenecker, Simon, et al.
Published: (2025)
by: Gerstenecker, Simon, et al.
Published: (2025)
Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability
by: Gao, Shenyuan, et al.
Published: (2024)
by: Gao, Shenyuan, et al.
Published: (2024)
ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models
by: Hamdan, Shadi, et al.
Published: (2025)
by: Hamdan, Shadi, et al.
Published: (2025)
ReSim: Reliable World Simulation for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2025)
by: Yang, Jiazhi, et al.
Published: (2025)
GenAD: Generalized Predictive Model for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2024)
by: Yang, Jiazhi, et al.
Published: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
by: Jaeger, Bernhard, et al.
Published: (2025)
by: Jaeger, Bernhard, et al.
Published: (2025)
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
by: Nguyen, Long, et al.
Published: (2025)
by: Nguyen, Long, et al.
Published: (2025)
Pseudo-Simulation for Autonomous Driving
by: Cao, Wei, et al.
Published: (2025)
by: Cao, Wei, et al.
Published: (2025)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
by: Ding, Kairui, et al.
Published: (2024)
by: Ding, Kairui, et al.
Published: (2024)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
by: Tao, Mingzhe, et al.
Published: (2026)
by: Tao, Mingzhe, et al.
Published: (2026)
TinyDrive: Multiscale Visual Question Answering with Selective Token Routing for Autonomous Driving
by: Hassani, Hossein, et al.
Published: (2025)
by: Hassani, Hossein, et al.
Published: (2025)
LingoQA: Visual Question Answering for Autonomous Driving
by: Marcu, Ana-Maria, et al.
Published: (2023)
by: Marcu, Ana-Maria, et al.
Published: (2023)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
by: Dauner, Daniel, et al.
Published: (2026)
by: Dauner, Daniel, et al.
Published: (2026)
VLM-Assisted Continual learning for Visual Question Answering in Self-Driving
by: Lin, Yuxin, et al.
Published: (2025)
by: Lin, Yuxin, et al.
Published: (2025)
Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
by: Xie, Shaoyuan, et al.
Published: (2025)
by: Xie, Shaoyuan, et al.
Published: (2025)
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
by: Yang, Zetong, et al.
Published: (2023)
by: Yang, Zetong, et al.
Published: (2023)
Answering Questions in Stages: Prompt Chaining for Contract QA
by: Roegiest, Adam, et al.
Published: (2024)
by: Roegiest, Adam, et al.
Published: (2024)
SimLingo: Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
by: Renz, Katrin, et al.
Published: (2025)
by: Renz, Katrin, et al.
Published: (2025)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
by: Wang, Tianqi, et al.
Published: (2024)
by: Wang, Tianqi, et al.
Published: (2024)
STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes
by: Ishihara, Keishi, et al.
Published: (2025)
by: Ishihara, Keishi, et al.
Published: (2025)
Efficient Visual Question Answering Pipeline for Autonomous Driving via Scene Region Compression
by: Cai, Yuliang, et al.
Published: (2026)
by: Cai, Yuliang, et al.
Published: (2026)
SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving
by: Li, Peizheng, et al.
Published: (2025)
by: Li, Peizheng, et al.
Published: (2025)
Enhancing Generalization in Medical Visual Question Answering Tasks via Gradient-Guided Model Perturbation
by: Liu, Gang, et al.
Published: (2024)
by: Liu, Gang, et al.
Published: (2024)
SimpleLLM4AD: An End-to-End Vision-Language Model with Graph Visual Question Answering for Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2024)
by: Zheng, Peiru, et al.
Published: (2024)
NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
by: Qian, Tianwen, et al.
Published: (2023)
by: Qian, Tianwen, et al.
Published: (2023)
Optimizing Visual Question Answering Models for Driving: Bridging the Gap Between Human and Machine Attention Patterns
by: Rekanar, Kaavya, et al.
Published: (2024)
by: Rekanar, Kaavya, et al.
Published: (2024)
Latent Chain-of-Thought World Modeling for End-to-End Driving
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
Precise Drive with VLM: First Prize Solution for PRCV 2024 Drive LM challenge
by: Huang, Bin, et al.
Published: (2024)
by: Huang, Bin, et al.
Published: (2024)
WaymoQA: A Multi-View Visual Question Answering Dataset for Safety-Critical Reasoning in Autonomous Driving
by: Yu, Seungjun, et al.
Published: (2025)
by: Yu, Seungjun, et al.
Published: (2025)
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
by: Chahe, Amirhosein, et al.
Published: (2025)
by: Chahe, Amirhosein, et al.
Published: (2025)
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models
by: Mohamud, Safaa Abdullahi Moallim, et al.
Published: (2025)
by: Mohamud, Safaa Abdullahi Moallim, et al.
Published: (2025)
Natural Language Understanding and Inference with MLLM in Visual Question Answering: A Survey
by: Kuang, Jiayi, et al.
Published: (2024)
by: Kuang, Jiayi, et al.
Published: (2024)
Adversarial Training with OCR Modality Perturbation for Scene-Text Visual Question Answering
by: Shen, Zhixuan, et al.
Published: (2024)
by: Shen, Zhixuan, et al.
Published: (2024)
Test-time Correction: An Online 3D Detection System via Visual Prompting
by: Zhang, Hanxue, et al.
Published: (2024)
by: Zhang, Hanxue, et al.
Published: (2024)
Similar Items
-
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024) -
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
by: Sima, Chonghao, et al.
Published: (2025) -
FLARE: Learning Future-Aware Latent Representations from Vision-Language Models for Autonomous Driving
by: Xie, Chengen, et al.
Published: (2026) -
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026) -
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)