X-Driver: Explainable Autonomous Driving with Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Wei, Zhang, Jiyuan, Zheng, Binxiong, Hu, Yufeng, Lin, Yingzhan, Zeng, Zengfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings
by: Wasi, Azmine Toushik, et al.
Published: (2026)
by: Wasi, Azmine Toushik, et al.
Published: (2026)
Classification of Driver Behaviour Using External Observation Techniques for Autonomous Vehicles
by: Nell, Ian, et al.
Published: (2025)
by: Nell, Ian, et al.
Published: (2025)
Cross-Platform Scaling of Vision-Language-Action Models from Edge to Cloud GPUs
by: Taherin, Amir, et al.
Published: (2025)
by: Taherin, Amir, et al.
Published: (2025)
VLAD: A VLM-Augmented Autonomous Driving Framework with Hierarchical Planning and Interpretable Decision Process
by: Gariboldi, Cristian, et al.
Published: (2025)
by: Gariboldi, Cristian, et al.
Published: (2025)
Future Aspects in Human Action Recognition: Exploring Emerging Techniques and Ethical Influences
by: Gasteratos, Antonios, et al.
Published: (2024)
by: Gasteratos, Antonios, et al.
Published: (2024)
Pedestrian Intention Prediction via Vision-Language Foundation Models
by: Azarmi, Mohsen, et al.
Published: (2025)
by: Azarmi, Mohsen, et al.
Published: (2025)
Hyperspectral Sensors and Autonomous Driving: Technologies, Limitations, and Opportunities
by: Shah, Imad Ali, et al.
Published: (2025)
by: Shah, Imad Ali, et al.
Published: (2025)
Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models
by: Wang, Dongdong, et al.
Published: (2026)
by: Wang, Dongdong, et al.
Published: (2026)
Human Cognition in Machines: A Unified Perspective of World Models
by: Rupprecht, Timothy, et al.
Published: (2026)
by: Rupprecht, Timothy, et al.
Published: (2026)
All-Optical Segmentation via Diffractive Neural Networks for Autonomous Driving
by: Li, Yingjie, et al.
Published: (2026)
by: Li, Yingjie, et al.
Published: (2026)
Fast Quantum Convolutional Neural Networks for Low-Complexity Object Detection in Autonomous Driving Applications
by: Baek, Hankyul, et al.
Published: (2023)
by: Baek, Hankyul, et al.
Published: (2023)
CMAP: Cross-Modal Adaptive Prompting for Multi-Domain Task-Incremental Learning
by: Mandalika, Sriram
Published: (2026)
by: Mandalika, Sriram
Published: (2026)
Simultaneous Estimation of Manipulation Skill and Hand Grasp Force from Forearm Ultrasound Images
by: Bimbraw, Keshav, et al.
Published: (2025)
by: Bimbraw, Keshav, et al.
Published: (2025)
Driver-Net: Multi-Camera Fusion for Assessing Driver Take-Over Readiness in Automated Vehicles
by: Rezaei, Mahdi, et al.
Published: (2025)
by: Rezaei, Mahdi, et al.
Published: (2025)
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm
by: Zhang, Jiangning, et al.
Published: (2022)
by: Zhang, Jiangning, et al.
Published: (2022)
EvGNN: An Event-driven Graph Neural Network Accelerator for Edge Vision
by: Yang, Yufeng, et al.
Published: (2024)
by: Yang, Yufeng, et al.
Published: (2024)
From Review to Design: Ethical Multimodal Driver Monitoring Systems for Risk Mitigation, Incident Response, and Accountability in Automated Vehicles
by: Khana, Bilal, et al.
Published: (2026)
by: Khana, Bilal, et al.
Published: (2026)
From Images to Insights: Explainable Biodiversity Monitoring with Plain Language Habitat Explanations
by: Zhou, Yutong, et al.
Published: (2025)
by: Zhou, Yutong, et al.
Published: (2025)
MolMetaLM: a Physicochemical Knowledge-Guided Molecular Meta Language Model
by: Wu, Yifan, et al.
Published: (2024)
by: Wu, Yifan, et al.
Published: (2024)
TOFFE -- Temporally-binned Object Flow from Events for High-speed and Energy-Efficient Object Detection and Tracking
by: Kosta, Adarsh Kumar, et al.
Published: (2025)
by: Kosta, Adarsh Kumar, et al.
Published: (2025)
FEDORA: Flying Event Dataset fOr Reactive behAvior
by: Joshi, Amogh, et al.
Published: (2023)
by: Joshi, Amogh, et al.
Published: (2023)
Toward Fully Autonomous Driving: AI, Challenges, Opportunities, and Needs
by: Ullrich, Lars, et al.
Published: (2026)
by: Ullrich, Lars, et al.
Published: (2026)
AgentThink: A Unified Framework for Tool-Augmented Chain-of-Thought Reasoning in Vision-Language Models for Autonomous Driving
by: Qian, Kangan, et al.
Published: (2025)
by: Qian, Kangan, et al.
Published: (2025)
Assistive Image Annotation Systems with Deep Learning and Natural Language Capabilities: A Review
by: Mots'oehli, Moseli
Published: (2024)
by: Mots'oehli, Moseli
Published: (2024)
WeatherEdit: Controllable Weather Editing with 4D Gaussian Field
by: Qian, Chenghao, et al.
Published: (2025)
by: Qian, Chenghao, et al.
Published: (2025)
Neural 3D Object Reconstruction with Small-Scale Unmanned Aerial Vehicles
by: Veres-Vitàlyos, Àlmos, et al.
Published: (2025)
by: Veres-Vitàlyos, Àlmos, et al.
Published: (2025)
Designing for Difference: How Human Characteristics Shape Perceptions of Collaborative Robots
by: Livanec, Sabrina, et al.
Published: (2025)
by: Livanec, Sabrina, et al.
Published: (2025)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
by: Goodge, Adam, et al.
Published: (2025)
by: Goodge, Adam, et al.
Published: (2025)
VOLMO: Versatile and Open Large Models for Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2026)
by: Qin, Zhenyue, et al.
Published: (2026)
VIPER Strike: Defeating Visual Reasoning CAPTCHAs via Structured Vision-Language Inference
by: Qi, Minfeng, et al.
Published: (2026)
by: Qi, Minfeng, et al.
Published: (2026)
From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage
by: Ruan, Cihan, et al.
Published: (2026)
by: Ruan, Cihan, et al.
Published: (2026)
Can Foundation Models Revolutionize Mobile AR Sparse Sensing?
by: Zhao, Yiqin, et al.
Published: (2025)
by: Zhao, Yiqin, et al.
Published: (2025)
Efficient Motion Sickness Assessment: Recreation of On-Road Driving on a Compact Test Track
by: Harmankaya, Huseyin, et al.
Published: (2024)
by: Harmankaya, Huseyin, et al.
Published: (2024)
Towards Evaluating Large Language Models for Graph Query Generation
by: Munir, Siraj, et al.
Published: (2024)
by: Munir, Siraj, et al.
Published: (2024)
LaScA: Language-Conditioned Scalable Modelling of Affective Dynamics
by: Pinitas, Kosmas, et al.
Published: (2026)
by: Pinitas, Kosmas, et al.
Published: (2026)
Multi-Faceted Evaluation of Modeling Languages for Augmented Reality Applications -- The Case of ARWFML
by: Muff, Fabian, et al.
Published: (2024)
by: Muff, Fabian, et al.
Published: (2024)
A Superalignment Framework in Autonomous Driving with Large Language Models
by: Kong, Xiangrui, et al.
Published: (2024)
by: Kong, Xiangrui, et al.
Published: (2024)
Simulating Refractive Distortions and Weather-Induced Artifacts for Resource-Constrained Autonomous Perception
by: Mots'oehli, Moseli, et al.
Published: (2025)
by: Mots'oehli, Moseli, et al.
Published: (2025)
SafeDrive: Knowledge- and Data-Driven Risk-Sensitive Decision-Making for Autonomous Vehicles with Large Language Models
by: Zhou, Zhiyuan, et al.
Published: (2024)
by: Zhou, Zhiyuan, et al.
Published: (2024)
UAV-Based Intelligent Traffic Surveillance System: Real-Time Vehicle Detection, Classification, Tracking, and Behavioral Analysis
by: Khanpour, Ali, et al.
Published: (2025)
by: Khanpour, Ali, et al.
Published: (2025)
Similar Items
-
TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings
by: Wasi, Azmine Toushik, et al.
Published: (2026) -
Classification of Driver Behaviour Using External Observation Techniques for Autonomous Vehicles
by: Nell, Ian, et al.
Published: (2025) -
Cross-Platform Scaling of Vision-Language-Action Models from Edge to Cloud GPUs
by: Taherin, Amir, et al.
Published: (2025) -
VLAD: A VLM-Augmented Autonomous Driving Framework with Hierarchical Planning and Interpretable Decision Process
by: Gariboldi, Cristian, et al.
Published: (2025) -
Future Aspects in Human Action Recognition: Exploring Emerging Techniques and Ethical Influences
by: Gasteratos, Antonios, et al.
Published: (2024)