A Unified Perception-Language-Action Framework for Adaptive Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yi, Haß, Erik Leo, Chao, Kuo-Yi, Petrovic, Nenad, Song, Yinglei, Wu, Chengdong, Knoll, Alois |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PRAM-R: A Perception-Reasoning-Action-Memory Framework with LLM-Guided Modality Routing for Adaptive Autonomous Driving
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
From Code to Road: A Vehicle-in-the-Loop and Digital Twin-Based Framework for Central Car Server Testing in Autonomous Driving
von: Wu, Chengdong, et al.
Veröffentlicht: (2026)
von: Wu, Chengdong, et al.
Veröffentlicht: (2026)
Digital-Twin Losses for Lane-Compliant Trajectory Prediction at Urban Intersections
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026)
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026)
Beyond the Vehicle: Cooperative Localization by Fusing Point Clouds for GPS-Challenged Urban Scenarios
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026)
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
GarchingSim: An Autonomous Driving Simulator with Photorealistic Scenes and Minimalist Workflow
von: Zhou, Liguo, et al.
Veröffentlicht: (2024)
von: Zhou, Liguo, et al.
Veröffentlicht: (2024)
Percept-WAM: Perception-Enhanced World-Awareness-Action Model for Robust End-to-End Autonomous Driving
von: Han, Jianhua, et al.
Veröffentlicht: (2025)
von: Han, Jianhua, et al.
Veröffentlicht: (2025)
AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
A Survey on Occupancy Perception for Autonomous Driving: The Information Fusion Perspective
von: Xu, Huaiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Huaiyuan, et al.
Veröffentlicht: (2024)
ZOPP: A Framework of Zero-shot Offboard Panoptic Perception for Autonomous Driving
von: Ma, Tao, et al.
Veröffentlicht: (2024)
von: Ma, Tao, et al.
Veröffentlicht: (2024)
A Survey of World Models for Autonomous Driving
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving
von: Wei, Julong, et al.
Veröffentlicht: (2024)
von: Wei, Julong, et al.
Veröffentlicht: (2024)
SimLingo: Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
von: Renz, Katrin, et al.
Veröffentlicht: (2025)
von: Renz, Katrin, et al.
Veröffentlicht: (2025)
Collaborative Perception Datasets in Autonomous Driving: A Survey
von: Yazgan, Melih, et al.
Veröffentlicht: (2024)
von: Yazgan, Melih, et al.
Veröffentlicht: (2024)
Joint Perception and Prediction for Autonomous Driving: A Survey
von: Dal'Col, Lucas, et al.
Veröffentlicht: (2024)
von: Dal'Col, Lucas, et al.
Veröffentlicht: (2024)
A Survey on Vision-Language-Action Models for Autonomous Driving
von: Jiang, Sicong, et al.
Veröffentlicht: (2025)
von: Jiang, Sicong, et al.
Veröffentlicht: (2025)
GEMINUS: Dual-aware Global and Scene-Adaptive Mixture-of-Experts for End-to-End Autonomous Driving
von: Wan, Chi, et al.
Veröffentlicht: (2025)
von: Wan, Chi, et al.
Veröffentlicht: (2025)
OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
2nd Place Solution for CVPR2024 E2E Challenge: End-to-End Autonomous Driving Using Vision Language Model
von: Guo, Zilong, et al.
Veröffentlicht: (2025)
von: Guo, Zilong, et al.
Veröffentlicht: (2025)
HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
DeFlow: Decoder of Scene Flow Network in Autonomous Driving
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
von: Shang, Shuyao, et al.
Veröffentlicht: (2026)
von: Shang, Shuyao, et al.
Veröffentlicht: (2026)
Towards Latency-Aware 3D Streaming Perception for Autonomous Driving
von: Peng, Jiaqi, et al.
Veröffentlicht: (2025)
von: Peng, Jiaqi, et al.
Veröffentlicht: (2025)
TrajDiff: End-to-end Autonomous Driving without Perception Annotation
von: Gui, Xingtai, et al.
Veröffentlicht: (2025)
von: Gui, Xingtai, et al.
Veröffentlicht: (2025)
Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving
von: Xie, Shaoyuan, et al.
Veröffentlicht: (2024)
von: Xie, Shaoyuan, et al.
Veröffentlicht: (2024)
Panoptic Perception for Autonomous Driving: A Survey
von: Li, Yunge, et al.
Veröffentlicht: (2024)
von: Li, Yunge, et al.
Veröffentlicht: (2024)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
DiffAD: A Unified Diffusion Modeling Approach for Autonomous Driving
von: Wang, Tao, et al.
Veröffentlicht: (2025)
von: Wang, Tao, et al.
Veröffentlicht: (2025)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
von: Dauner, Daniel, et al.
Veröffentlicht: (2026)
von: Dauner, Daniel, et al.
Veröffentlicht: (2026)
Unified Video Action Model
von: Li, Shuang, et al.
Veröffentlicht: (2025)
von: Li, Shuang, et al.
Veröffentlicht: (2025)
Progressive Bird's Eye View Perception for Safety-Critical Autonomous Driving: A Comprehensive Survey
von: Gong, Yan, et al.
Veröffentlicht: (2025)
von: Gong, Yan, et al.
Veröffentlicht: (2025)
To New Beginnings: A Survey of Unified Perception in Autonomous Vehicle Software
von: Stratil, Loïc, et al.
Veröffentlicht: (2025)
von: Stratil, Loïc, et al.
Veröffentlicht: (2025)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PRAM-R: A Perception-Reasoning-Action-Memory Framework with LLM-Guided Modality Routing for Adaptive Autonomous Driving
von: Zhang, Yi, et al.
Veröffentlicht: (2026) -
From Code to Road: A Vehicle-in-the-Loop and Digital Twin-Based Framework for Central Car Server Testing in Autonomous Driving
von: Wu, Chengdong, et al.
Veröffentlicht: (2026) -
Digital-Twin Losses for Lane-Compliant Trajectory Prediction at Urban Intersections
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026) -
Beyond the Vehicle: Cooperative Localization by Fusing Point Clouds for GPS-Challenged Urban Scenarios
von: Chao, Kuo-Yi, et al.
Veröffentlicht: (2026) -
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2026)