VERDI: VLM-Embedded Reasoning for Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Bowen, Mei, Zhiting, Ost, Julian, Ghilotti, Filippo, Li, Baiang, Girgis, Roger, Majumdar, Anirudha, Heide, Felix |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments
by: Rowe, Luke, et al.
Published: (2025)
by: Rowe, Luke, et al.
Published: (2025)
Inverse Neural Rendering for Explainable Multi-Object Tracking
by: Ost, Julian, et al.
Published: (2024)
by: Ost, Julian, et al.
Published: (2024)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
TruckDrive: Long-Range Autonomous Highway Driving Dataset
by: Ghilotti, Filippo, et al.
Published: (2026)
by: Ghilotti, Filippo, et al.
Published: (2026)
ScenarioControl: Vision-Language Controllable Vectorized Latent Scenario Generation
by: Gao, Lili, et al.
Published: (2026)
by: Gao, Lili, et al.
Published: (2026)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
by: Li, Huiqiong, et al.
Published: (2026)
by: Li, Huiqiong, et al.
Published: (2026)
Poutine: Vision-Language-Trajectory Pre-Training and Reinforcement Learning Post-Training Enable Robust End-to-End Autonomous Driving
by: Rowe, Luke, et al.
Published: (2025)
by: Rowe, Luke, et al.
Published: (2025)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
by: Yin, Tenny, et al.
Published: (2025)
by: Yin, Tenny, et al.
Published: (2025)
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
by: Shorinwa, Ola, et al.
Published: (2025)
by: Shorinwa, Ola, et al.
Published: (2025)
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025)
by: Qin, Jack, et al.
Published: (2025)
VDT-Auto: End-to-end Autonomous Driving with VLM-Guided Diffusion Transformers
by: Guo, Ziang, et al.
Published: (2025)
by: Guo, Ziang, et al.
Published: (2025)
Weak-to-Strong Knowledge Distillation Accelerates Visual Learning
by: Li, Baiang, et al.
Published: (2026)
by: Li, Baiang, et al.
Published: (2026)
WorldFlow3D: Flowing Through 3D Distributions for Unbounded World Generation
by: Joshi, Amogh, et al.
Published: (2026)
by: Joshi, Amogh, et al.
Published: (2026)
Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement Learning
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
by: Jiang, Bo, et al.
Published: (2025)
by: Jiang, Bo, et al.
Published: (2025)
UniLiPs: Unified LiDAR Pseudo-Labeling with Geometry-Grounded Dynamic Scene Decomposition
by: Ghilotti, Filippo, et al.
Published: (2026)
by: Ghilotti, Filippo, et al.
Published: (2026)
Self-Supervised Sparse Sensor Fusion for Long Range Perception
by: Palladin, Edoardo, et al.
Published: (2025)
by: Palladin, Edoardo, et al.
Published: (2025)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
by: Huang, Zilin, et al.
Published: (2026)
by: Huang, Zilin, et al.
Published: (2026)
Work Zones challenge VLM Trajectory Planning: Toward Mitigation and Robust Autonomous Driving
by: Liao, Yifan, et al.
Published: (2025)
by: Liao, Yifan, et al.
Published: (2025)
A Survey of World Models for Autonomous Driving
by: Feng, Tuo, et al.
Published: (2025)
by: Feng, Tuo, et al.
Published: (2025)
Learning Direct Control Policies with Flow Matching for Autonomous Driving
by: Ceresini, Marcello, et al.
Published: (2026)
by: Ceresini, Marcello, et al.
Published: (2026)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
by: Shang, Shuyao, et al.
Published: (2026)
by: Shang, Shuyao, et al.
Published: (2026)
AutoDrive-R$^2$: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving
by: Yuan, Zhenlong, et al.
Published: (2025)
by: Yuan, Zhenlong, et al.
Published: (2025)
Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
by: Zhang, Dapeng, et al.
Published: (2025)
by: Zhang, Dapeng, et al.
Published: (2025)
ReconDrive: Fast Feed-Forward 4D Gaussian Splatting for Autonomous Driving Scene Reconstruction
by: Yu, Haibao, et al.
Published: (2026)
by: Yu, Haibao, et al.
Published: (2026)
LHPF: Look back the History and Plan for the Future in Autonomous Driving
by: Wang, Sheng, et al.
Published: (2024)
by: Wang, Sheng, et al.
Published: (2024)
StyleDrive: Towards Driving-Style Aware Benchmarking of End-To-End Autonomous Driving
by: Hao, Ruiyang, et al.
Published: (2025)
by: Hao, Ruiyang, et al.
Published: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
by: Huang, Zilin, et al.
Published: (2024)
by: Huang, Zilin, et al.
Published: (2024)
AnoVox: A Benchmark for Multimodal Anomaly Detection in Autonomous Driving
by: Bogdoll, Daniel, et al.
Published: (2024)
by: Bogdoll, Daniel, et al.
Published: (2024)
SEPT: Standard-Definition Map Enhanced Scene Perception and Topology Reasoning for Autonomous Driving
by: Pei, Muleilan, et al.
Published: (2025)
by: Pei, Muleilan, et al.
Published: (2025)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
by: Huang, Zhijian, et al.
Published: (2024)
by: Huang, Zhijian, et al.
Published: (2024)
RoadRunner -- Learning Traversability Estimation for Autonomous Off-road Driving
by: Frey, Jonas, et al.
Published: (2024)
by: Frey, Jonas, et al.
Published: (2024)
WorldVLM: Combining World Model Forecasting and Vision-Language Reasoning
by: Englmeier, Stefan, et al.
Published: (2026)
by: Englmeier, Stefan, et al.
Published: (2026)
DriveSafer: End-to-End Autonomous Driving with Safety Guidance
by: Sural, Shounak, et al.
Published: (2026)
by: Sural, Shounak, et al.
Published: (2026)
ComDrive: Comfort-Oriented End-to-End Autonomous Driving
by: Wang, Junming, et al.
Published: (2024)
by: Wang, Junming, et al.
Published: (2024)
CurricuVLM: Towards Safe Autonomous Driving via Personalized Safety-Critical Curriculum Learning with Vision-Language Models
by: Sheng, Zihao, et al.
Published: (2025)
by: Sheng, Zihao, et al.
Published: (2025)
Scalable Offline Metrics for Autonomous Driving
by: Aich, Animikh, et al.
Published: (2025)
by: Aich, Animikh, et al.
Published: (2025)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
by: Liao, Bencheng, et al.
Published: (2024)
by: Liao, Bencheng, et al.
Published: (2024)
Similar Items
-
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025) -
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments
by: Rowe, Luke, et al.
Published: (2025) -
Inverse Neural Rendering for Explainable Multi-Object Tracking
by: Ost, Julian, et al.
Published: (2024) -
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
TruckDrive: Long-Range Autonomous Highway Driving Dataset
by: Ghilotti, Filippo, et al.
Published: (2026)