Research on the Application of Computer Vision Based on Deep Learning in Autonomous Driving Technology
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jingyu, Cao, Jin, Chang, Jinghao, Li, Xinjin, Liu, Houze, Li, Zhenglin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Examination of Offline-Trained Encoders in Vision-Based Deep Reinforcement Learning for Autonomous Driving
von: Mohammed, Shawan, et al.
Veröffentlicht: (2024)
von: Mohammed, Shawan, et al.
Veröffentlicht: (2024)
DriveCritic: Towards Context-Aware, Human-Aligned Evaluation for Autonomous Driving with Vision-Language Models
von: Song, Jingyu, et al.
Veröffentlicht: (2025)
von: Song, Jingyu, et al.
Veröffentlicht: (2025)
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
Enhancing Breast Cancer Detection with Vision Transformers and Graph Neural Networks
von: Cai, Yeming, et al.
Veröffentlicht: (2025)
von: Cai, Yeming, et al.
Veröffentlicht: (2025)
Vision Language Models in Autonomous Driving: A Survey and Outlook
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2023)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2023)
YOLO-PPA based Efficient Traffic Sign Detection for Cruise Control in Autonomous Driving
von: Zhang, Jingyu, et al.
Veröffentlicht: (2024)
von: Zhang, Jingyu, et al.
Veröffentlicht: (2024)
Dynamic Universal Approximation Theory: The Basic Theory for Deep Learning-Based Computer Vision Models
von: Wang, Wei, et al.
Veröffentlicht: (2024)
von: Wang, Wei, et al.
Veröffentlicht: (2024)
Learning Vision-Language-Action World Models for Autonomous Driving
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
Research on Driving Scenario Technology Based on Multimodal Large Lauguage Model Optimization
von: Mengjie, Wang, et al.
Veröffentlicht: (2025)
von: Mengjie, Wang, et al.
Veröffentlicht: (2025)
Advanced Feature Manipulation for Enhanced Change Detection Leveraging Natural Language Models
von: Li, Zhenglin, et al.
Veröffentlicht: (2024)
von: Li, Zhenglin, et al.
Veröffentlicht: (2024)
Transformer-Based Framework for Motion Capture Denoising and Anomaly Detection in Medical Rehabilitation
von: Cai, Yeming, et al.
Veröffentlicht: (2025)
von: Cai, Yeming, et al.
Veröffentlicht: (2025)
Black-Box Adversarial Attack on Vision Language Models for Autonomous Driving
von: Wang, Lu, et al.
Veröffentlicht: (2025)
von: Wang, Lu, et al.
Veröffentlicht: (2025)
Natural Reflection Backdoor Attack on Vision Language Model for Autonomous Driving
von: Liu, Ming, et al.
Veröffentlicht: (2025)
von: Liu, Ming, et al.
Veröffentlicht: (2025)
Research on Edge Detection of LiDAR Images Based on Artificial Intelligence Technology
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
VLADriver-RAG: Retrieval-Augmented Vision-Language-Action Models for Autonomous Driving
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events
von: Bhat, Mohammad Qazim, et al.
Veröffentlicht: (2026)
von: Bhat, Mohammad Qazim, et al.
Veröffentlicht: (2026)
Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving
von: Lian, Weitong, et al.
Veröffentlicht: (2026)
von: Lian, Weitong, et al.
Veröffentlicht: (2026)
OG-Gaussian: Occupancy Based Street Gaussians for Autonomous Driving
von: Shen, Yedong, et al.
Veröffentlicht: (2025)
von: Shen, Yedong, et al.
Veröffentlicht: (2025)
Research on Intelligent Aided Diagnosis System of Medical Image Based on Computer Deep Learning
von: Yuan, Jiajie, et al.
Veröffentlicht: (2024)
von: Yuan, Jiajie, et al.
Veröffentlicht: (2024)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
Evaluation of Safety Cognition Capability in Vision-Language Models for Autonomous Driving
von: Zhang, Enming, et al.
Veröffentlicht: (2025)
von: Zhang, Enming, et al.
Veröffentlicht: (2025)
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
von: Yang, Yiming, et al.
Veröffentlicht: (2025)
von: Yang, Yiming, et al.
Veröffentlicht: (2025)
LaVida Drive: Vision-Text Interaction VLM for Autonomous Driving with Token Selection, Recovery and Enhancement
von: Jiao, Siwen, et al.
Veröffentlicht: (2024)
von: Jiao, Siwen, et al.
Veröffentlicht: (2024)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
A Low-Rank Method for Vision Language Model Hallucination Mitigation in Autonomous Driving
von: Long, Keke, et al.
Veröffentlicht: (2025)
von: Long, Keke, et al.
Veröffentlicht: (2025)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2026)
von: Huang, Zilin, et al.
Veröffentlicht: (2026)
VECTOR-Drive: Tightly Coupled Vision-Language and Trajectory Expert Routing for End-to-End Autonomous Driving
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
RALAD: Bridging the Real-to-Sim Domain Gap in Autonomous Driving with Retrieval-Augmented Learning
von: Zuo, Jiacheng, et al.
Veröffentlicht: (2025)
von: Zuo, Jiacheng, et al.
Veröffentlicht: (2025)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling
von: You, Junwei, et al.
Veröffentlicht: (2025)
von: You, Junwei, et al.
Veröffentlicht: (2025)
Computer Vision and Deep Learning for 4D Augmented Reality
von: Shivashankar, Karthik
Veröffentlicht: (2025)
von: Shivashankar, Karthik
Veröffentlicht: (2025)
Autonomous Computer Vision Development with Agentic AI
von: Kim, Jin, et al.
Veröffentlicht: (2025)
von: Kim, Jin, et al.
Veröffentlicht: (2025)
RAD: Retrieval-Augmented Decision-Making of Meta-Actions with Vision-Language Models in Autonomous Driving
von: Wang, Yujin, et al.
Veröffentlicht: (2025)
von: Wang, Yujin, et al.
Veröffentlicht: (2025)
RAC3: Retrieval-Augmented Corner Case Comprehension for Autonomous Driving with Vision-Language Models
von: Wang, Yujin, et al.
Veröffentlicht: (2024)
von: Wang, Yujin, et al.
Veröffentlicht: (2024)
VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
AutoSplat: Constrained Gaussian Splatting for Autonomous Driving Scene Reconstruction
von: Khan, Mustafa, et al.
Veröffentlicht: (2024)
von: Khan, Mustafa, et al.
Veröffentlicht: (2024)
A Survey on Vision-Language-Action Models for Autonomous Driving
von: Jiang, Sicong, et al.
Veröffentlicht: (2025)
von: Jiang, Sicong, et al.
Veröffentlicht: (2025)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Examination of Offline-Trained Encoders in Vision-Based Deep Reinforcement Learning for Autonomous Driving
von: Mohammed, Shawan, et al.
Veröffentlicht: (2024) -
DriveCritic: Towards Context-Aware, Human-Aligned Evaluation for Autonomous Driving with Vision-Language Models
von: Song, Jingyu, et al.
Veröffentlicht: (2025) -
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
von: Jiang, Anqing, et al.
Veröffentlicht: (2025) -
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025) -
Enhancing Breast Cancer Detection with Vision Transformers and Graph Neural Networks
von: Cai, Yeming, et al.
Veröffentlicht: (2025)