Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Greer, Ross, Trivedi, Mohan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Perception Without Vision for Trajectory Prediction: Ego Vehicle Dynamics as Scene Representation for Efficient Active Learning in Autonomous Driving
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024)
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024)
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
von: Keskar, Maitrayee, et al.
Veröffentlicht: (2025)
von: Keskar, Maitrayee, et al.
Veröffentlicht: (2025)
Vision and Language: Novel Representations and Artificial intelligence for Driving Scene Safety Assessment and Autonomous Vehicle Planning
von: Greer, Ross, et al.
Veröffentlicht: (2026)
von: Greer, Ross, et al.
Veröffentlicht: (2026)
Learning to Find Missing Video Frames with Synthetic Data Augmentation: A General Framework and Application in Generating Thermal Images Using RGB Cameras
von: Andersen, Mathias Viborg, et al.
Veröffentlicht: (2024)
von: Andersen, Mathias Viborg, et al.
Veröffentlicht: (2024)
ActiveAnno3D -- An Active Learning Framework for Multi-Modal 3D Object Detection
von: Ghita, Ahmed, et al.
Veröffentlicht: (2024)
von: Ghita, Ahmed, et al.
Veröffentlicht: (2024)
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Towards Safe and Reliable Autonomous Driving: Dynamic Occupancy Set Prediction
von: Shao, Wenbo, et al.
Veröffentlicht: (2024)
von: Shao, Wenbo, et al.
Veröffentlicht: (2024)
Active Learning from Scene Embeddings for End-to-End Autonomous Driving
von: Jiang, Wenhao, et al.
Veröffentlicht: (2025)
von: Jiang, Wenhao, et al.
Veröffentlicht: (2025)
Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection
von: Mohamed, Sondos, et al.
Veröffentlicht: (2024)
von: Mohamed, Sondos, et al.
Veröffentlicht: (2024)
Finding the Reflection Point: Unpadding Images to Remove Data Augmentation Artifacts in Large Open Source Image Datasets for Machine Learning
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
von: Roy, Parthib, et al.
Veröffentlicht: (2024)
von: Roy, Parthib, et al.
Veröffentlicht: (2024)
Natural Language Instructions for Scene-Responsive Human-in-the-Loop Motion Planning in Autonomous Driving using Vision-Language-Action Models
von: Martinez-Sanchez, Angel, et al.
Veröffentlicht: (2026)
von: Martinez-Sanchez, Angel, et al.
Veröffentlicht: (2026)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
Safety-Critical Learning for Long-Tail Events: The TUM Traffic Accident Dataset
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
von: Choi, Lucas, et al.
Veröffentlicht: (2024)
Automated Data Curation Using GPS & NLP to Generate Instruction-Action Pairs for Autonomous Vehicle Vision-Language Navigation Datasets
von: Roque, Guillermo, et al.
Veröffentlicht: (2025)
von: Roque, Guillermo, et al.
Veröffentlicht: (2025)
LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving
von: Song, Nan, et al.
Veröffentlicht: (2025)
von: Song, Nan, et al.
Veröffentlicht: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
Agro-Consensus: Semantic Self-Consistency in Vision-Language Models for Crop Disease Management in Developing Countries
von: Gupta, Mihir, et al.
Veröffentlicht: (2025)
von: Gupta, Mihir, et al.
Veröffentlicht: (2025)
Safe2Drive: Evaluating Safe Driving Behaviors of E2E Autonomous Driving Models
von: Sahu, Nishad, et al.
Veröffentlicht: (2026)
von: Sahu, Nishad, et al.
Veröffentlicht: (2026)
MindDrive: An All-in-One Framework Bridging World Models and Vision-Language Model for End-to-End Autonomous Driving
von: Sun, Bin, et al.
Veröffentlicht: (2025)
von: Sun, Bin, et al.
Veröffentlicht: (2025)
SMc2f: Robust Scenario Mining for Robotic Autonomy from Coarse to Fine
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
SafeAlign-VLA: A Negative-Enhanced Safe Alignment Framework for Risk-Aware Autonomous Driving
von: Tian, Kefei, et al.
Veröffentlicht: (2026)
von: Tian, Kefei, et al.
Veröffentlicht: (2026)
Learning Vision-Language-Action World Models for Autonomous Driving
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
Towards Vision Zero: The TUM Traffic Accid3nD Dataset
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
BeLLA: End-to-End Birds Eye View Large Language Assistant for Autonomous Driving
von: Mohan, Karthik, et al.
Veröffentlicht: (2025)
von: Mohan, Karthik, et al.
Veröffentlicht: (2025)
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
von: Bossen, Tonko E. W., et al.
Veröffentlicht: (2025)
von: Bossen, Tonko E. W., et al.
Veröffentlicht: (2025)
SafeMVDrive: Multi-view Safety-Critical Driving Video Synthesis in the Real World Domain
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
CurricuVLM: Towards Safe Autonomous Driving via Personalized Safety-Critical Curriculum Learning with Vision-Language Models
von: Sheng, Zihao, et al.
Veröffentlicht: (2025)
von: Sheng, Zihao, et al.
Veröffentlicht: (2025)
JiSAM: Alleviate Labeling Burden and Corner Case Problems in Autonomous Driving via Minimal Real-World Data
von: Chen, Runjian, et al.
Veröffentlicht: (2025)
von: Chen, Runjian, et al.
Veröffentlicht: (2025)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
von: Shi, Chen, et al.
Veröffentlicht: (2025)
von: Shi, Chen, et al.
Veröffentlicht: (2025)
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT
von: Hu, Xiaotao, et al.
Veröffentlicht: (2024)
von: Hu, Xiaotao, et al.
Veröffentlicht: (2024)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
Efficient and Explainable End-to-End Autonomous Driving via Masked Vision-Language-Action Diffusion
von: Zhang, Jiaru, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaru, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Perception Without Vision for Trajectory Prediction: Ego Vehicle Dynamics as Scene Representation for Efficient Active Learning in Autonomous Driving
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024) -
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
von: Keskar, Maitrayee, et al.
Veröffentlicht: (2025)