DriVLM: Domain Adaptation of Vision-Language Models in Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Xuran, Yoo, Chang D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Cross-Modal Domain Adaptation in Brain Disease Diagnosis: Maximum Mean Discrepancy-based Convolutional Neural Networks
by: Zhu, Xuran
Published: (2024)
by: Zhu, Xuran
Published: (2024)
VLM-C4L: Continual Core Dataset Learning with Corner Case Optimization via Vision-Language Models for Autonomous Driving
by: Hu, Haibo, et al.
Published: (2025)
by: Hu, Haibo, et al.
Published: (2025)
DriVLMe: Enhancing LLM-based Autonomous Driving Agents with Embodied and Social Experiences
by: Huang, Yidong, et al.
Published: (2024)
by: Huang, Yidong, et al.
Published: (2024)
V2X-VLM: End-to-End V2X Cooperative Autonomous Driving Through Large Vision-Language Models
by: You, Junwei, et al.
Published: (2024)
by: You, Junwei, et al.
Published: (2024)
An Empirical Study of Sample Selection Strategies for Large Language Model Repair
by: Li, Xuran, et al.
Published: (2025)
by: Li, Xuran, et al.
Published: (2025)
NaviDriveVLM: Decoupling High-Level Reasoning and Motion Planning for Autonomous Driving
by: Tao, Ximeng, et al.
Published: (2026)
by: Tao, Ximeng, et al.
Published: (2026)
Prompt-Driven Domain Adaptation for End-to-End Autonomous Driving via In-Context RL
by: Khurram, Aleesha, et al.
Published: (2025)
by: Khurram, Aleesha, et al.
Published: (2025)
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation
by: Lee, Donghoon, et al.
Published: (2025)
by: Lee, Donghoon, et al.
Published: (2025)
Policy Learning from Large Vision-Language Model Feedback without Reward Modeling
by: Luu, Tung M., et al.
Published: (2025)
by: Luu, Tung M., et al.
Published: (2025)
Universal Camouflage Attack on Vision-Language Models for Autonomous Driving
by: Kong, Dehong, et al.
Published: (2025)
by: Kong, Dehong, et al.
Published: (2025)
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization
by: Pham, Trung X., et al.
Published: (2025)
by: Pham, Trung X., et al.
Published: (2025)
Reward Generation via Large Vision-Language Model in Offline Reinforcement Learning
by: Lee, Younghwan, et al.
Published: (2025)
by: Lee, Younghwan, et al.
Published: (2025)
VLM Q-Learning: Aligning Vision-Language Models for Interactive Decision-Making
by: Grigsby, Jake, et al.
Published: (2025)
by: Grigsby, Jake, et al.
Published: (2025)
AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference
by: Feng, Yilin, et al.
Published: (2026)
by: Feng, Yilin, et al.
Published: (2026)
FastVLM: Efficient Vision Encoding for Vision Language Models
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
Domain Adaptation with a Single Vision-Language Embedding
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving
by: Xing, Shuo, et al.
Published: (2024)
by: Xing, Shuo, et al.
Published: (2024)
FastVLM: Self-Speculative Decoding for Fast Vision-Language Model Inference
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
Inference for Deep Neural Network Estimators in Generalized Nonparametric Models
by: Meng, Xuran, et al.
Published: (2025)
by: Meng, Xuran, et al.
Published: (2025)
Source-Free Domain Adaptation Guided by Vision and Vision-Language Pre-Training
by: Zhang, Wenyu, et al.
Published: (2024)
by: Zhang, Wenyu, et al.
Published: (2024)
Meta-Learning Online Dynamics Model Adaptation in Off-Road Autonomous Driving
by: Levy, Jacob, et al.
Published: (2025)
by: Levy, Jacob, et al.
Published: (2025)
Autonomous Source Knowledge Selection in Multi-Domain Adaptation
by: Li, Keqiuyin, et al.
Published: (2025)
by: Li, Keqiuyin, et al.
Published: (2025)
Noisy Test-Time Adaptation in Vision-Language Models
by: Cao, Chentao, et al.
Published: (2025)
by: Cao, Chentao, et al.
Published: (2025)
GraphVLM: Benchmarking Vision Language Models for Multimodal Graph Learning
by: Liu, Jiajin, et al.
Published: (2026)
by: Liu, Jiajin, et al.
Published: (2026)
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
by: Bao, Chen, et al.
Published: (2024)
by: Bao, Chen, et al.
Published: (2024)
Language-Induced Priors for Domain Adaptation
by: Chen, Qiyuan, et al.
Published: (2026)
by: Chen, Qiyuan, et al.
Published: (2026)
CF-VLM:CounterFactual Vision-Language Fine-tuning
by: Zhang, Jusheng, et al.
Published: (2025)
by: Zhang, Jusheng, et al.
Published: (2025)
In-Context Policy Adaptation via Cross-Domain Skill Diffusion
by: Yoo, Minjong, et al.
Published: (2025)
by: Yoo, Minjong, et al.
Published: (2025)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
by: Gao, Haoxiang, et al.
Published: (2025)
by: Gao, Haoxiang, et al.
Published: (2025)
Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models
by: Luu, Tung Minh, et al.
Published: (2025)
by: Luu, Tung Minh, et al.
Published: (2025)
Learning to Prompt Your Domain for Vision-Language Models
by: Wei, Guoyizhe, et al.
Published: (2023)
by: Wei, Guoyizhe, et al.
Published: (2023)
GeoFlowVLM: Geometry-Aware Joint Uncertainty for Frozen Vision-Language Embedding
by: Nautiyal, Mayank, et al.
Published: (2026)
by: Nautiyal, Mayank, et al.
Published: (2026)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
by: Wang, Yufei, et al.
Published: (2024)
by: Wang, Yufei, et al.
Published: (2024)
Partial Identifiability for Domain Adaptation
by: Kong, Lingjing, et al.
Published: (2023)
by: Kong, Lingjing, et al.
Published: (2023)
Model Adaptation for Time Constrained Embodied Control
by: Song, Jaehyun, et al.
Published: (2024)
by: Song, Jaehyun, et al.
Published: (2024)
Vision-Language Model Selection and Reuse for Downstream Adaptation
by: Tan, Hao-Zhe, et al.
Published: (2025)
by: Tan, Hao-Zhe, et al.
Published: (2025)
BendVLM: Test-Time Debiasing of Vision-Language Embeddings
by: Gerych, Walter, et al.
Published: (2024)
by: Gerych, Walter, et al.
Published: (2024)
SteerVLM: Robust Model Control through Lightweight Activation Steering for Vision Language Models
by: Sivakumar, Anushka, et al.
Published: (2025)
by: Sivakumar, Anushka, et al.
Published: (2025)
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
by: Zhong, Siru, et al.
Published: (2025)
by: Zhong, Siru, et al.
Published: (2025)
Similar Items
-
VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
by: Xu, Yi, et al.
Published: (2024) -
Cross-Modal Domain Adaptation in Brain Disease Diagnosis: Maximum Mean Discrepancy-based Convolutional Neural Networks
by: Zhu, Xuran
Published: (2024) -
VLM-C4L: Continual Core Dataset Learning with Corner Case Optimization via Vision-Language Models for Autonomous Driving
by: Hu, Haibo, et al.
Published: (2025) -
DriVLMe: Enhancing LLM-based Autonomous Driving Agents with Embodied and Social Experiences
by: Huang, Yidong, et al.
Published: (2024) -
V2X-VLM: End-to-End V2X Cooperative Autonomous Driving Through Large Vision-Language Models
by: You, Junwei, et al.
Published: (2024)