ALN-P3: Unified Language Alignment for Perception, Prediction, and Planning in Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Yunsheng, Yaman, Burhaneddin, Ye, Xin, Yurt, Mahmut, Luo, Jingru, Mallik, Abhirup, Wang, Ziran, Ren, Liu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MTA: Multimodal Task Alignment for BEV Perception and Captioning
by: Ma, Yunsheng, et al.
Published: (2024)
by: Ma, Yunsheng, et al.
Published: (2024)
LTDA-Drive: LLMs-guided Generative Models based Long-tail Data Augmentation for Autonomous Driving
by: Yurt, Mahmut, et al.
Published: (2025)
by: Yurt, Mahmut, et al.
Published: (2025)
VLP: Vision Language Planning for Autonomous Driving
by: Pan, Chenbin, et al.
Published: (2024)
by: Pan, Chenbin, et al.
Published: (2024)
LORD: Large Models based Opposite Reward Design for Autonomous Driving
by: Ye, Xin, et al.
Published: (2024)
by: Ye, Xin, et al.
Published: (2024)
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
by: Ye, Xin, et al.
Published: (2025)
by: Ye, Xin, et al.
Published: (2025)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
by: Xiong, Zhexiao, et al.
Published: (2026)
by: Xiong, Zhexiao, et al.
Published: (2026)
AdaWM: Adaptive World Model based Planning for Autonomous Driving
by: Wang, Hang, et al.
Published: (2025)
by: Wang, Hang, et al.
Published: (2025)
Reconstruction Matters: Learning Geometry-Aligned BEV Representation through 3D Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2026)
by: Lu, Yiren, et al.
Published: (2026)
Video Token Sparsification for Efficient Multimodal LLMs in Autonomous Driving
by: Ma, Yunsheng, et al.
Published: (2024)
by: Ma, Yunsheng, et al.
Published: (2024)
ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving
by: Sheng, Zihao, et al.
Published: (2026)
by: Sheng, Zihao, et al.
Published: (2026)
CLIP-BEVFormer: Enhancing Multi-View Image-Based BEV Detector with Ground Truth Flow
by: Pan, Chenbin, et al.
Published: (2024)
by: Pan, Chenbin, et al.
Published: (2024)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
by: Ma, Yunsheng, et al.
Published: (2022)
by: Ma, Yunsheng, et al.
Published: (2022)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
by: Li, Yongkang, et al.
Published: (2026)
by: Li, Yongkang, et al.
Published: (2026)
LaMPilot: An Open Benchmark Dataset for Autonomous Driving with Language Model Programs
by: Ma, Yunsheng, et al.
Published: (2023)
by: Ma, Yunsheng, et al.
Published: (2023)
DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
PaPr: Training-Free One-Step Patch Pruning with Lightweight ConvNets for Faster Inference
by: Mahmud, Tanvir, et al.
Published: (2024)
by: Mahmud, Tanvir, et al.
Published: (2024)
NuPlanQA: A Large-Scale Dataset and Benchmark for Multi-View Driving Scene Understanding in Multi-Modal Large Language Models
by: Park, Sung-Yeon, et al.
Published: (2025)
by: Park, Sung-Yeon, et al.
Published: (2025)
ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving
by: Li, Jingyu, et al.
Published: (2025)
by: Li, Jingyu, et al.
Published: (2025)
$AutoDrive\text{-}P^3$: Unified Chain of Perception-Prediction-Planning Thought via Reinforcement Fine-Tuning
by: Ye, Yuqi, et al.
Published: (2026)
by: Ye, Yuqi, et al.
Published: (2026)
Perception in Plan: Coupled Perception and Planning for End-to-End Autonomous Driving
by: Zhang, Bozhou, et al.
Published: (2025)
by: Zhang, Bozhou, et al.
Published: (2025)
Quantifying Uncertainty in Motion Prediction with Variational Bayesian Mixture
by: Lu, Juanwu, et al.
Published: (2024)
by: Lu, Juanwu, et al.
Published: (2024)
Fully Unified Motion Planning for End-to-End Autonomous Driving
by: Liu, Lin, et al.
Published: (2025)
by: Liu, Lin, et al.
Published: (2025)
Comparison of Large Language Models for Deployment Requirements
by: Yaman, Alper, et al.
Published: (2025)
by: Yaman, Alper, et al.
Published: (2025)
A Unified Perception-Language-Action Framework for Adaptive Autonomous Driving
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
PriorFusion: Unified Integration of Priors for Robust Road Perception in Autonomous Driving
by: Tang, Xuewei, et al.
Published: (2025)
by: Tang, Xuewei, et al.
Published: (2025)
LLM4AD: Large Language Models for Autonomous Driving -- Concept, Review, Benchmark, Experiments, and Future Trends
by: Cui, Can, et al.
Published: (2024)
by: Cui, Can, et al.
Published: (2024)
PPAD: Iterative Interactions of Prediction and Planning for End-to-end Autonomous Driving
by: Chen, Zhili, et al.
Published: (2023)
by: Chen, Zhili, et al.
Published: (2023)
Joint Perception and Prediction for Autonomous Driving: A Survey
by: Dal'Col, Lucas, et al.
Published: (2024)
by: Dal'Col, Lucas, et al.
Published: (2024)
Efficient and Explainable End-to-End Autonomous Driving via Masked Vision-Language-Action Diffusion
by: Zhang, Jiaru, et al.
Published: (2026)
by: Zhang, Jiaru, et al.
Published: (2026)
Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning
by: Zhang, Bozhou, et al.
Published: (2025)
by: Zhang, Bozhou, et al.
Published: (2025)
ITPNet: Towards Instantaneous Trajectory Prediction for Autonomous Driving
by: Li, Rongqing, et al.
Published: (2024)
by: Li, Rongqing, et al.
Published: (2024)
Hybrid-Prediction Integrated Planning for Autonomous Driving
by: Liu, Haochen, et al.
Published: (2024)
by: Liu, Haochen, et al.
Published: (2024)
A Language Agent for Autonomous Driving
by: Mao, Jiageng, et al.
Published: (2023)
by: Mao, Jiageng, et al.
Published: (2023)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
by: jia, Feiyang, et al.
Published: (2026)
by: jia, Feiyang, et al.
Published: (2026)
3D Object Visibility Prediction in Autonomous Driving
by: Luo, Chuanyu, et al.
Published: (2024)
by: Luo, Chuanyu, et al.
Published: (2024)
Unifying Language-Action Understanding and Generation for Autonomous Driving
by: Wang, Xinyang, et al.
Published: (2026)
by: Wang, Xinyang, et al.
Published: (2026)
Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking
by: Li, Jingru, et al.
Published: (2026)
by: Li, Jingru, et al.
Published: (2026)
On the Foundation Model for Cardiac MRI Reconstruction
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
DriveMLM: Aligning Multi-Modal Large Language Models with Behavioral Planning States for Autonomous Driving
by: Cui, Erfei, et al.
Published: (2023)
by: Cui, Erfei, et al.
Published: (2023)
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World
by: Xia, Tianze, et al.
Published: (2025)
by: Xia, Tianze, et al.
Published: (2025)
Similar Items
-
MTA: Multimodal Task Alignment for BEV Perception and Captioning
by: Ma, Yunsheng, et al.
Published: (2024) -
LTDA-Drive: LLMs-guided Generative Models based Long-tail Data Augmentation for Autonomous Driving
by: Yurt, Mahmut, et al.
Published: (2025) -
VLP: Vision Language Planning for Autonomous Driving
by: Pan, Chenbin, et al.
Published: (2024) -
LORD: Large Models based Opposite Reward Design for Autonomous Driving
by: Ye, Xin, et al.
Published: (2024) -
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
by: Ye, Xin, et al.
Published: (2025)