Distilling Multi-modal Large Language Models for Autonomous Driving
Fuente:
arXiv
Guardado en:
| Autores principales: | Hegde, Deepti, Yasarla, Rajeev, Cai, Hong, Han, Shizhong, Bhattacharyya, Apratim, Mahajan, Shweta, Liu, Litian, Garrepalli, Risheek, Patel, Vishal M., Porikli, Fatih |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2025)
por: Yasarla, Rajeev, et al.
Publicado: (2025)
Generative Scenario Rollouts for End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2026)
por: Yasarla, Rajeev, et al.
Publicado: (2026)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2026)
por: Yasarla, Rajeev, et al.
Publicado: (2026)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
por: Garrepalli, Risheek, et al.
Publicado: (2024)
por: Garrepalli, Risheek, et al.
Publicado: (2024)
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
por: Yasarla, Rajeev, et al.
Publicado: (2023)
por: Yasarla, Rajeev, et al.
Publicado: (2023)
DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos
por: Yasarla, Rajeev, et al.
Publicado: (2025)
por: Yasarla, Rajeev, et al.
Publicado: (2025)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
por: Mahajan, Shweta, et al.
Publicado: (2025)
por: Mahajan, Shweta, et al.
Publicado: (2025)
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
por: Yasarla, Rajeev, et al.
Publicado: (2024)
por: Yasarla, Rajeev, et al.
Publicado: (2024)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
por: Kadambi, Shreya, et al.
Publicado: (2025)
por: Kadambi, Shreya, et al.
Publicado: (2025)
SciFlow: Empowering Lightweight Optical Flow Models with Self-Cleaning Iterations
por: Lin, Jamie Menjay, et al.
Publicado: (2024)
por: Lin, Jamie Menjay, et al.
Publicado: (2024)
OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
por: Jeong, Jisoo, et al.
Publicado: (2024)
por: Jeong, Jisoo, et al.
Publicado: (2024)
Clockwork Diffusion: Efficient Generation With Model-Step Distillation
por: Habibian, Amirhossein, et al.
Publicado: (2023)
por: Habibian, Amirhossein, et al.
Publicado: (2023)
ToSA: Token Selective Attention for Efficient Vision Transformers
por: Singh, Manish Kumar, et al.
Publicado: (2024)
por: Singh, Manish Kumar, et al.
Publicado: (2024)
Spiffy: Multiplying Diffusion LLM Acceleration via Lossless Speculative Decoding
por: Agrawal, Sudhanshu, et al.
Publicado: (2025)
por: Agrawal, Sudhanshu, et al.
Publicado: (2025)
A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs
por: Goel, Raghavv, et al.
Publicado: (2026)
por: Goel, Raghavv, et al.
Publicado: (2026)
HexaGen3D: StableDiffusion is just one step away from Fast and Diverse Text-to-3D Generation
por: Mercier, Antoine, et al.
Publicado: (2024)
por: Mercier, Antoine, et al.
Publicado: (2024)
DuoLoRA : Cycle-consistent and Rank-disentangled Content-Style Personalization
por: Roy, Aniket, et al.
Publicado: (2025)
por: Roy, Aniket, et al.
Publicado: (2025)
Notes-to-Self: Scratchpad Augmented VLAs for Memory Dependent Manipulation Tasks
por: Haresh, Sanjay, et al.
Publicado: (2026)
por: Haresh, Sanjay, et al.
Publicado: (2026)
Learning Safe Autonomous Driving Policies Using Predictive Safety Representations
por: Keswani, Mahesh, et al.
Publicado: (2025)
por: Keswani, Mahesh, et al.
Publicado: (2025)
ClevrSkills: Compositional Language and Visual Reasoning in Robotics
por: Haresh, Sanjay, et al.
Publicado: (2024)
por: Haresh, Sanjay, et al.
Publicado: (2024)
CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers
por: Li, Zhuojin, et al.
Publicado: (2026)
por: Li, Zhuojin, et al.
Publicado: (2026)
DSDrive: Distilling Large Language Model for Lightweight End-to-End Autonomous Driving with Unified Reasoning and Planning
por: Liu, Wenru, et al.
Publicado: (2025)
por: Liu, Wenru, et al.
Publicado: (2025)
From Words to Wheels: Automated Style-Customized Policy Generation for Autonomous Driving
por: Han, Xu, et al.
Publicado: (2024)
por: Han, Xu, et al.
Publicado: (2024)
Explanation for Trajectory Planning using Multi-modal Large Language Model for Autonomous Driving
por: Yamazaki, Shota, et al.
Publicado: (2024)
por: Yamazaki, Shota, et al.
Publicado: (2024)
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
por: Xu, Tianshuo, et al.
Publicado: (2025)
por: Xu, Tianshuo, et al.
Publicado: (2025)
ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution
por: Fu, Chuyao, et al.
Publicado: (2026)
por: Fu, Chuyao, et al.
Publicado: (2026)
CALMM-Drive: Confidence-Aware Autonomous Driving with Large Multimodal Model
por: Yao, Ruoyu, et al.
Publicado: (2024)
por: Yao, Ruoyu, et al.
Publicado: (2024)
DistillDrive: End-to-End Multi-Mode Autonomous Driving Distillation by Isomorphic Hetero-Source Planning Model
por: Yu, Rui, et al.
Publicado: (2025)
por: Yu, Rui, et al.
Publicado: (2025)
LimSim Series: An Autonomous Driving Simulation Platform for Validation and Enhancement
por: Fu, Daocheng, et al.
Publicado: (2025)
por: Fu, Daocheng, et al.
Publicado: (2025)
METDrive: Multi-modal End-to-end Autonomous Driving with Temporal Guidance
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language Models
por: Wen, Licheng, et al.
Publicado: (2023)
por: Wen, Licheng, et al.
Publicado: (2023)
FALO: Fast and Accurate LiDAR 3D Object Detection on Resource-Constrained Devices
por: Han, Shizhong, et al.
Publicado: (2025)
por: Han, Shizhong, et al.
Publicado: (2025)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
por: Borse, Shubhankar, et al.
Publicado: (2025)
por: Borse, Shubhankar, et al.
Publicado: (2025)
HyperNet Fields: Efficiently Training Hypernetworks without Ground Truth by Learning Weight Trajectories
por: Hedlin, Eric, et al.
Publicado: (2024)
por: Hedlin, Eric, et al.
Publicado: (2024)
Large Language Model based Interactive Decision-Making for Autonomous Driving
por: Dong, Xinwei, et al.
Publicado: (2026)
por: Dong, Xinwei, et al.
Publicado: (2026)
Risk-Aware Reinforcement Learning for Autonomous Driving: Improving Safety When Driving through Intersection
por: Leng, Bo, et al.
Publicado: (2025)
por: Leng, Bo, et al.
Publicado: (2025)
Hi-Drive: Hierarchical POMDP Planning for Safe Autonomous Driving in Diverse Urban Environments
por: Jin, Xuanjin, et al.
Publicado: (2024)
por: Jin, Xuanjin, et al.
Publicado: (2024)
Continual Adaptation for Autonomous Driving with the Mixture of Progressive Experts Network
por: Cui, Yixin, et al.
Publicado: (2025)
por: Cui, Yixin, et al.
Publicado: (2025)
Seamless Virtual Reality with Integrated Synchronizer and Synthesizer for Autonomous Driving
por: Li, He, et al.
Publicado: (2024)
por: Li, He, et al.
Publicado: (2024)
BePo: Dual Representation for 3D Occupancy Prediction
por: Shi, Yunxiao, et al.
Publicado: (2025)
por: Shi, Yunxiao, et al.
Publicado: (2025)
Ejemplares similares
-
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2025) -
Generative Scenario Rollouts for End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2026) -
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2026) -
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
por: Garrepalli, Risheek, et al.
Publicado: (2024) -
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
por: Yasarla, Rajeev, et al.
Publicado: (2023)