Efficient Driving Behavior Narration and Reasoning on Edge Device Using Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, Yizhou, Cheng, Yihua, Wang, Kezhi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Trajectory Mamba: Efficient Attention-Mamba Forecasting Model Based on Selective SSM
por: Huang, Yizhou, et al.
Publicado: (2025)
por: Huang, Yizhou, et al.
Publicado: (2025)
V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views
por: You, Junwei, et al.
Publicado: (2026)
por: You, Junwei, et al.
Publicado: (2026)
Behavioral Cloning Models Reality Check for Autonomous Driving
por: Yildirim, Mustafa, et al.
Publicado: (2024)
por: Yildirim, Mustafa, et al.
Publicado: (2024)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
por: Zhou, Enshen, et al.
Publicado: (2025)
por: Zhou, Enshen, et al.
Publicado: (2025)
SORT3D: Spatial Object-centric Reasoning Toolbox for Zero-Shot 3D Grounding Using Large Language Models
por: Zantout, Nader, et al.
Publicado: (2025)
por: Zantout, Nader, et al.
Publicado: (2025)
NeuFlow: Real-time, High-accuracy Optical Flow Estimation on Robots Using Edge Devices
por: Zhang, Zhiyong, et al.
Publicado: (2024)
por: Zhang, Zhiyong, et al.
Publicado: (2024)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
por: Huang, Zilin, et al.
Publicado: (2026)
por: Huang, Zilin, et al.
Publicado: (2026)
DriveGPT: Scaling Autoregressive Behavior Models for Driving
por: Huang, Xin, et al.
Publicado: (2024)
por: Huang, Xin, et al.
Publicado: (2024)
A Survey on Vision-Language-Action Models for Autonomous Driving
por: Jiang, Sicong, et al.
Publicado: (2025)
por: Jiang, Sicong, et al.
Publicado: (2025)
NuPlanQA: A Large-Scale Dataset and Benchmark for Multi-View Driving Scene Understanding in Multi-Modal Large Language Models
por: Park, Sung-Yeon, et al.
Publicado: (2025)
por: Park, Sung-Yeon, et al.
Publicado: (2025)
CoT-Drive: Efficient Motion Forecasting for Autonomous Driving with LLMs and Chain-of-Thought Prompting
por: Liao, Haicheng, et al.
Publicado: (2025)
por: Liao, Haicheng, et al.
Publicado: (2025)
TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics
por: Han, Yi, et al.
Publicado: (2025)
por: Han, Yi, et al.
Publicado: (2025)
Integrating Object Detection Modality into Visual Language Model for Enhanced Autonomous Driving Agent
por: He, Linfeng, et al.
Publicado: (2024)
por: He, Linfeng, et al.
Publicado: (2024)
Is VLA Reasoning Faithful? Probing Safety of Chain-of-Causation in Autonomous Driving Models
por: Mayumu, Nicanor, et al.
Publicado: (2026)
por: Mayumu, Nicanor, et al.
Publicado: (2026)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
por: Feng, Bowen, et al.
Publicado: (2025)
por: Feng, Bowen, et al.
Publicado: (2025)
Less is More: Lean yet Powerful Vision-Language Model for Autonomous Driving
por: Yang, Sheng, et al.
Publicado: (2025)
por: Yang, Sheng, et al.
Publicado: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
por: Huang, Zilin, et al.
Publicado: (2024)
por: Huang, Zilin, et al.
Publicado: (2024)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
por: Wu, Yiming, et al.
Publicado: (2025)
por: Wu, Yiming, et al.
Publicado: (2025)
Tag Map: A Text-Based Map for Spatial Reasoning and Navigation with Large Language Models
por: Zhang, Mike, et al.
Publicado: (2024)
por: Zhang, Mike, et al.
Publicado: (2024)
DriveCritic: Towards Context-Aware, Human-Aligned Evaluation for Autonomous Driving with Vision-Language Models
por: Song, Jingyu, et al.
Publicado: (2025)
por: Song, Jingyu, et al.
Publicado: (2025)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
por: Yang, Zhenjie, et al.
Publicado: (2025)
por: Yang, Zhenjie, et al.
Publicado: (2025)
SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling
por: You, Junwei, et al.
Publicado: (2025)
por: You, Junwei, et al.
Publicado: (2025)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
por: Gao, Haoxiang, et al.
Publicado: (2025)
por: Gao, Haoxiang, et al.
Publicado: (2025)
NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models
por: Zhou, Gengze, et al.
Publicado: (2024)
por: Zhou, Gengze, et al.
Publicado: (2024)
CurricuVLM: Towards Safe Autonomous Driving via Personalized Safety-Critical Curriculum Learning with Vision-Language Models
por: Sheng, Zihao, et al.
Publicado: (2025)
por: Sheng, Zihao, et al.
Publicado: (2025)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
por: Zhou, Yang, et al.
Publicado: (2026)
por: Zhou, Yang, et al.
Publicado: (2026)
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
por: Shao, Hao, et al.
Publicado: (2026)
por: Shao, Hao, et al.
Publicado: (2026)
HiST-VLA: A Hierarchical Spatio-Temporal Vision-Language-Action Model for End-to-End Autonomous Driving
por: Wang, Yiru, et al.
Publicado: (2026)
por: Wang, Yiru, et al.
Publicado: (2026)
Vega: Learning to Drive with Natural Language Instructions
por: Zuo, Sicheng, et al.
Publicado: (2026)
por: Zuo, Sicheng, et al.
Publicado: (2026)
A Low-Rank Method for Vision Language Model Hallucination Mitigation in Autonomous Driving
por: Long, Keke, et al.
Publicado: (2025)
por: Long, Keke, et al.
Publicado: (2025)
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
por: Jiang, Anqing, et al.
Publicado: (2025)
por: Jiang, Anqing, et al.
Publicado: (2025)
VECTOR-Drive: Tightly Coupled Vision-Language and Trajectory Expert Routing for End-to-End Autonomous Driving
por: Zhao, Rui, et al.
Publicado: (2026)
por: Zhao, Rui, et al.
Publicado: (2026)
VLA-Thinker: Boosting Vision-Language-Action Models through Thinking-with-Image Reasoning
por: Wang, Chaoyang, et al.
Publicado: (2026)
por: Wang, Chaoyang, et al.
Publicado: (2026)
DeeAD: Dynamic Early Exit of Vision-Language Action for Efficient Autonomous Driving
por: HU, Haibo, et al.
Publicado: (2025)
por: HU, Haibo, et al.
Publicado: (2025)
DRAMA-X: A Fine-grained Intent Prediction and Risk Reasoning Benchmark For Driving
por: Godbole, Mihir, et al.
Publicado: (2025)
por: Godbole, Mihir, et al.
Publicado: (2025)
DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
por: Zhou, Yang, et al.
Publicado: (2026)
por: Zhou, Yang, et al.
Publicado: (2026)
Synthesizing the Kill Chain: A Zero-Shot Framework for Target Verification and Tactical Reasoning on the Edge
por: Barkley, Jesse, et al.
Publicado: (2026)
por: Barkley, Jesse, et al.
Publicado: (2026)
Learning Velocity and Acceleration: Self-Supervised Motion Consistency for Pedestrian Trajectory Prediction
por: Huang, Yizhou, et al.
Publicado: (2025)
por: Huang, Yizhou, et al.
Publicado: (2025)
LLM-RG: Referential Grounding in Outdoor Scenarios using Large Language Models
por: Saxena, Pranav, et al.
Publicado: (2025)
por: Saxena, Pranav, et al.
Publicado: (2025)
A Language Agent for Autonomous Driving
por: Mao, Jiageng, et al.
Publicado: (2023)
por: Mao, Jiageng, et al.
Publicado: (2023)
Ejemplares similares
-
Trajectory Mamba: Efficient Attention-Mamba Forecasting Model Based on Selective SSM
por: Huang, Yizhou, et al.
Publicado: (2025) -
V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views
por: You, Junwei, et al.
Publicado: (2026) -
Behavioral Cloning Models Reality Check for Autonomous Driving
por: Yildirim, Mustafa, et al.
Publicado: (2024) -
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
por: Zhou, Enshen, et al.
Publicado: (2025) -
SORT3D: Spatial Object-centric Reasoning Toolbox for Zero-Shot 3D Grounding Using Large Language Models
por: Zantout, Nader, et al.
Publicado: (2025)