DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Weicheng, Mao, Xiaofei, Ye, Nanfei, Li, Pengxiang, Zhan, Kun, Lang, Xianpeng, Zhao, Hang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2024)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving
von: Hou, Xinmeng, et al.
Veröffentlicht: (2025)
von: Hou, Xinmeng, et al.
Veröffentlicht: (2025)
Generalizing Motion Planners with Mixture of Experts for Autonomous Driving
von: Sun, Qiao, et al.
Veröffentlicht: (2024)
von: Sun, Qiao, et al.
Veröffentlicht: (2024)
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
von: Jiang, Xuefeng, et al.
Veröffentlicht: (2025)
von: Jiang, Xuefeng, et al.
Veröffentlicht: (2025)
DriveCombo: Benchmarking Compositional Traffic Rule Reasoning in Autonomous Driving
von: Ma, Enhui, et al.
Veröffentlicht: (2026)
von: Ma, Enhui, et al.
Veröffentlicht: (2026)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving
von: Li, Peizheng, et al.
Veröffentlicht: (2025)
von: Li, Peizheng, et al.
Veröffentlicht: (2025)
Learning Personalized Driving Styles via Reinforcement Learning from Human Feedback
von: Li, Derun, et al.
Veröffentlicht: (2025)
von: Li, Derun, et al.
Veröffentlicht: (2025)
SimpleVSF: VLM-Scoring Fusion for Trajectory Prediction of End-to-End Autonomous Driving
von: Zheng, Peiru, et al.
Veröffentlicht: (2025)
von: Zheng, Peiru, et al.
Veröffentlicht: (2025)
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
von: Ma, Enhui, et al.
Veröffentlicht: (2024)
von: Ma, Enhui, et al.
Veröffentlicht: (2024)
DriveMA: Driving Vision-Language-Action Models with verifiable Meta-Actions
von: Zheng, Weicheng, et al.
Veröffentlicht: (2026)
von: Zheng, Weicheng, et al.
Veröffentlicht: (2026)
A Language Agent for Autonomous Driving
von: Mao, Jiageng, et al.
Veröffentlicht: (2023)
von: Mao, Jiageng, et al.
Veröffentlicht: (2023)
TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving
von: Li, Jiaxiang, et al.
Veröffentlicht: (2026)
von: Li, Jiaxiang, et al.
Veröffentlicht: (2026)
AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models
von: Han, Yuxuan, et al.
Veröffentlicht: (2026)
von: Han, Yuxuan, et al.
Veröffentlicht: (2026)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
von: Tang, Tao, et al.
Veröffentlicht: (2024)
von: Tang, Tao, et al.
Veröffentlicht: (2024)
OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving
von: Yan, Tianyi, et al.
Veröffentlicht: (2024)
von: Yan, Tianyi, et al.
Veröffentlicht: (2024)
GaussianAD: Gaussian-Centric End-to-End Autonomous Driving
von: Zheng, Wenzhao, et al.
Veröffentlicht: (2024)
von: Zheng, Wenzhao, et al.
Veröffentlicht: (2024)
CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies
von: Chen, Keyu, et al.
Veröffentlicht: (2026)
von: Chen, Keyu, et al.
Veröffentlicht: (2026)
Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving
von: Kim, Donghyun, et al.
Veröffentlicht: (2026)
von: Kim, Donghyun, et al.
Veröffentlicht: (2026)
Data Scaling Laws for Imitation Learning-Based End-to-End Autonomous Driving
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
DAP: A Discrete-token Autoregressive Planner for Autonomous Driving
von: Ye, Bowen, et al.
Veröffentlicht: (2025)
von: Ye, Bowen, et al.
Veröffentlicht: (2025)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
von: Feng, Bowen, et al.
Veröffentlicht: (2025)
von: Feng, Bowen, et al.
Veröffentlicht: (2025)
SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Driving
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
DriVLM: Domain Adaptation of Vision-Language Models in Autonomous Driving
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
AdaThinkDrive: Adaptive Thinking via Reinforcement Learning for Autonomous Driving
von: Luo, Yuechen, et al.
Veröffentlicht: (2025)
von: Luo, Yuechen, et al.
Veröffentlicht: (2025)
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
von: Song, Ruoyu, et al.
Veröffentlicht: (2024)
von: Song, Ruoyu, et al.
Veröffentlicht: (2024)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
DriveAction: A Benchmark for Exploring Human-like Driving Decisions in VLA Models
von: Hao, Yuhan, et al.
Veröffentlicht: (2025)
von: Hao, Yuhan, et al.
Veröffentlicht: (2025)
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
DriveTester: A Unified Platform for Simulation-Based Autonomous Driving Testing
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
von: He, Yuankai, et al.
Veröffentlicht: (2025)
von: He, Yuankai, et al.
Veröffentlicht: (2025)
DynRsl-VLM: Enhancing Autonomous Driving Perception with Dynamic Resolution Vision-Language Models
von: Zhou, Xirui, et al.
Veröffentlicht: (2025)
von: Zhou, Xirui, et al.
Veröffentlicht: (2025)
CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
NaviDriveVLM: Decoupling High-Level Reasoning and Motion Planning for Autonomous Driving
von: Tao, Ximeng, et al.
Veröffentlicht: (2026)
von: Tao, Ximeng, et al.
Veröffentlicht: (2026)
TokenFLEX: Unified VLM Training for Flexible Visual Tokens Inference
von: Hu, Junshan, et al.
Veröffentlicht: (2025)
von: Hu, Junshan, et al.
Veröffentlicht: (2025)
Listen, Look, Drive: Coupling Audio Instructions for User-aware VLA-based Autonomous Driving
von: Guo, Ziang, et al.
Veröffentlicht: (2026)
von: Guo, Ziang, et al.
Veröffentlicht: (2026)
Energy-Efficient Autonomous Driving with Adaptive Perception and Robust Decision
von: Xia, Yuyang, et al.
Veröffentlicht: (2025)
von: Xia, Yuyang, et al.
Veröffentlicht: (2025)
Flow Matching-Based Autonomous Driving Planning with Advanced Interactive Behavior Modeling
von: Tan, Tianyi, et al.
Veröffentlicht: (2025)
von: Tan, Tianyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2024) -
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
von: Li, Pengxiang, et al.
Veröffentlicht: (2025) -
DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving
von: Hou, Xinmeng, et al.
Veröffentlicht: (2025) -
Generalizing Motion Planners with Mixture of Experts for Autonomous Driving
von: Sun, Qiao, et al.
Veröffentlicht: (2024) -
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
von: Jiang, Xuefeng, et al.
Veröffentlicht: (2025)