Probing Multimodal LLMs as World Models for Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Sreeram, Shiva, Wang, Tsun-Hsuan, Maalouf, Alaa, Rosman, Guy, Karaman, Sertac, Rus, Daniela |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
See Less, Drive Better: Generalizable End-to-End Autonomous Driving via Foundation Models Stochastic Patch Selection
by: Mallak, Amir, et al.
Published: (2026)
by: Mallak, Amir, et al.
Published: (2026)
Learning autonomous driving from aerial imagery
by: Murali, Varun, et al.
Published: (2024)
by: Murali, Varun, et al.
Published: (2024)
Generating Out-Of-Distribution Scenarios Using Language Models
by: Aasi, Erfan, et al.
Published: (2024)
by: Aasi, Erfan, et al.
Published: (2024)
Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
by: Sreeram, Shiva, et al.
Published: (2025)
by: Sreeram, Shiva, et al.
Published: (2025)
SAFe-Copilot: Unified Shared Autonomy Framework
by: Nguyen, Phat, et al.
Published: (2025)
by: Nguyen, Phat, et al.
Published: (2025)
Robustness Is a Function, Not a Number: A Factorized Comprehensive Study of OOD Robustness in Vision-Based Driving
by: Mallak, Amir, et al.
Published: (2026)
by: Mallak, Amir, et al.
Published: (2026)
Text-to-Drive: Diverse Driving Behavior Synthesis via Large Language Models
by: Nguyen, Phat, et al.
Published: (2024)
by: Nguyen, Phat, et al.
Published: (2024)
Follow Anything: Open-set detection, tracking, and following in real-time
by: Maalouf, Alaa, et al.
Published: (2023)
by: Maalouf, Alaa, et al.
Published: (2023)
ReGen: Generative Robot Simulation via Inverse Design
by: Nguyen, Phat, et al.
Published: (2025)
by: Nguyen, Phat, et al.
Published: (2025)
Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling
by: Qiu, Xiaowen, et al.
Published: (2025)
by: Qiu, Xiaowen, et al.
Published: (2025)
A Survey of World Models for Autonomous Driving
by: Feng, Tuo, et al.
Published: (2025)
by: Feng, Tuo, et al.
Published: (2025)
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
by: Shao, Hao, et al.
Published: (2026)
by: Shao, Hao, et al.
Published: (2026)
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
by: Liu, Xiaolu, et al.
Published: (2026)
by: Liu, Xiaolu, et al.
Published: (2026)
Surgical Foundation Model Leveraging Compression and Entropy Maximization for Image-Guided Surgical Assistance
by: Yin, Lianhao, et al.
Published: (2025)
by: Yin, Lianhao, et al.
Published: (2025)
Prompts to Summaries: Zero-Shot Language-Guided Video Summarization with Large Language and Video Models
by: Barbara, Mario, et al.
Published: (2025)
by: Barbara, Mario, et al.
Published: (2025)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
by: jia, Feiyang, et al.
Published: (2026)
by: jia, Feiyang, et al.
Published: (2026)
A Simulator Dataset to Support the Study of Impaired Driving
by: Gideon, John, et al.
Published: (2025)
by: Gideon, John, et al.
Published: (2025)
Learning to Drive from a World Model
by: Goff, Mitchell, et al.
Published: (2025)
by: Goff, Mitchell, et al.
Published: (2025)
GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control
by: Chen, Anthony, et al.
Published: (2025)
by: Chen, Anthony, et al.
Published: (2025)
Holistic Surgical Phase Recognition with Hierarchical Input Dependent State Space Models
by: Wu, Haoyang, et al.
Published: (2025)
by: Wu, Haoyang, et al.
Published: (2025)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
by: Liu, Qiqi, et al.
Published: (2026)
by: Liu, Qiqi, et al.
Published: (2026)
Decentralized Vision-Based Autonomous Aerial Wildlife Monitoring
by: Chahine, Makram, et al.
Published: (2025)
by: Chahine, Makram, et al.
Published: (2025)
Latent Chain-of-Thought World Modeling for End-to-End Driving
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
Is Your Driving World Model an All-Around Player?
by: Kong, Lingdong, et al.
Published: (2026)
by: Kong, Lingdong, et al.
Published: (2026)
Map-World: Masked Action planning and Path-Integral World Model for Autonomous Driving
by: Hu, Bin, et al.
Published: (2025)
by: Hu, Bin, et al.
Published: (2025)
WorldRFT: Latent World Model Planning with Reinforcement Fine-Tuning for Autonomous Driving
by: Yang, Pengxuan, et al.
Published: (2025)
by: Yang, Pengxuan, et al.
Published: (2025)
Tracking Meets Large Multimodal Models for Driving Scenario Understanding
by: Ishaq, Ayesha, et al.
Published: (2025)
by: Ishaq, Ayesha, et al.
Published: (2025)
Hypergraph-Transformer (HGT) for Interactive Event Prediction in Laparoscopic and Robotic Surgery
by: Yin, Lianhao, et al.
Published: (2024)
by: Yin, Lianhao, et al.
Published: (2024)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
by: Wang, Linbo, et al.
Published: (2026)
by: Wang, Linbo, et al.
Published: (2026)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
by: Huang, Zhijian, et al.
Published: (2024)
by: Huang, Zhijian, et al.
Published: (2024)
V2V-LLM: Vehicle-to-Vehicle Cooperative Autonomous Driving with Multimodal Large Language Models
by: Chiu, Hsu-kuang, et al.
Published: (2025)
by: Chiu, Hsu-kuang, et al.
Published: (2025)
DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding
by: Ishaq, Ayesha, et al.
Published: (2025)
by: Ishaq, Ayesha, et al.
Published: (2025)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
by: Shang, Shuyao, et al.
Published: (2026)
by: Shang, Shuyao, et al.
Published: (2026)
Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving
by: Geng, Jiaheng, et al.
Published: (2025)
by: Geng, Jiaheng, et al.
Published: (2025)
SimScale: Learning to Drive via Real-World Simulation at Scale
by: Tian, Haochen, et al.
Published: (2025)
by: Tian, Haochen, et al.
Published: (2025)
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving
by: Wei, Julong, et al.
Published: (2024)
by: Wei, Julong, et al.
Published: (2024)
HEAT: Heterogeneous End-to-End Autonomous Driving via Trajectory-Guided World Models
by: Cho, Hoonhee, et al.
Published: (2026)
by: Cho, Hoonhee, et al.
Published: (2026)
ReSim: Reliable World Simulation for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2025)
by: Yang, Jiazhi, et al.
Published: (2025)
Action Images: End-to-End Policy Learning via Multiview Video Generation
by: Zhen, Haoyu, et al.
Published: (2026)
by: Zhen, Haoyu, et al.
Published: (2026)
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving
by: Zhang, Lingjun, et al.
Published: (2026)
by: Zhang, Lingjun, et al.
Published: (2026)
Similar Items
-
See Less, Drive Better: Generalizable End-to-End Autonomous Driving via Foundation Models Stochastic Patch Selection
by: Mallak, Amir, et al.
Published: (2026) -
Learning autonomous driving from aerial imagery
by: Murali, Varun, et al.
Published: (2024) -
Generating Out-Of-Distribution Scenarios Using Language Models
by: Aasi, Erfan, et al.
Published: (2024) -
Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
by: Sreeram, Shiva, et al.
Published: (2025) -
SAFe-Copilot: Unified Shared Autonomy Framework
by: Nguyen, Phat, et al.
Published: (2025)