Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
Fuente:
arXiv
Saved in:
| Main Authors: | Shriram, Shashank, Perisetla, Srinivasa, Keskar, Aryan, Krishnaswamy, Harsha, Bossen, Tonko Emil Westerhof, Møgelmose, Andreas, Greer, Ross |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
by: Roy, Parthib, et al.
Published: (2024)
by: Roy, Parthib, et al.
Published: (2024)
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
by: Bossen, Tonko E. W., et al.
Published: (2025)
by: Bossen, Tonko E. W., et al.
Published: (2025)
Vision and Language: Novel Representations and Artificial intelligence for Driving Scene Safety Assessment and Autonomous Vehicle Planning
by: Greer, Ross, et al.
Published: (2026)
by: Greer, Ross, et al.
Published: (2026)
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
by: Choi, Lucas, et al.
Published: (2024)
by: Choi, Lucas, et al.
Published: (2024)
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
by: Keskar, Maitrayee, et al.
Published: (2025)
by: Keskar, Maitrayee, et al.
Published: (2025)
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
by: Choi, Lucas, et al.
Published: (2024)
by: Choi, Lucas, et al.
Published: (2024)
Perception Without Vision for Trajectory Prediction: Ego Vehicle Dynamics as Scene Representation for Efficient Active Learning in Autonomous Driving
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
by: Gopalkrishnan, Akshay, et al.
Published: (2024)
by: Gopalkrishnan, Akshay, et al.
Published: (2024)
CRASH: Cognitive Reasoning Agent for Safety Hazards in Autonomous Driving
by: Silva, Erick, et al.
Published: (2026)
by: Silva, Erick, et al.
Published: (2026)
Learning to Find Missing Video Frames with Synthetic Data Augmentation: A General Framework and Application in Generating Thermal Images Using RGB Cameras
by: Andersen, Mathias Viborg, et al.
Published: (2024)
by: Andersen, Mathias Viborg, et al.
Published: (2024)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
by: Kirchner, Sven, et al.
Published: (2025)
by: Kirchner, Sven, et al.
Published: (2025)
Natural Language Instructions for Scene-Responsive Human-in-the-Loop Motion Planning in Autonomous Driving using Vision-Language-Action Models
by: Martinez-Sanchez, Angel, et al.
Published: (2026)
by: Martinez-Sanchez, Angel, et al.
Published: (2026)
Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
by: Choi, Lucas, et al.
Published: (2025)
by: Choi, Lucas, et al.
Published: (2025)
ActiveAnno3D -- An Active Learning Framework for Multi-Modal 3D Object Detection
by: Ghita, Ahmed, et al.
Published: (2024)
by: Ghita, Ahmed, et al.
Published: (2024)
Spatial Optimization of Interconnected Systems in Non-Convex Design Spaces
by: Westerhof, S., et al.
Published: (2026)
by: Westerhof, S., et al.
Published: (2026)
APHP anonymous sample data SPARTE
by: Westerhof, Berend
Published: (2025)
by: Westerhof, Berend
Published: (2025)
A Hybrid Optimization Framework for Spatial Packaging of Interconnected Systems
by: Westerhof, S., et al.
Published: (2026)
by: Westerhof, S., et al.
Published: (2026)
ContextVLM: Zero-Shot and Few-Shot Context Understanding for Autonomous Driving using Vision Language Models
by: Sural, Shounak, et al.
Published: (2024)
by: Sural, Shounak, et al.
Published: (2024)
Zero-shot Hazard Identification in Autonomous Driving: A Case Study on the COOOL Benchmark
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
INSIGHT: Enhancing Autonomous Driving Safety through Vision-Language Models on Context-Aware Hazard Detection and Edge Case Evaluation
by: Chen, Dianwei, et al.
Published: (2025)
by: Chen, Dianwei, et al.
Published: (2025)
RELATOS DE VIAJE KAWÉSQAR
by: José Tonko P.
Published: (2008)
by: José Tonko P.
Published: (2008)
CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation
by: Liang, Xiwen, et al.
Published: (2023)
by: Liang, Xiwen, et al.
Published: (2023)
Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty
by: Aryan, Prakash, et al.
Published: (2026)
by: Aryan, Prakash, et al.
Published: (2026)
Evaluating Driver Perceptions of Integrated Safety Monitoring Systems for Alcohol Impairment and Distraction
by: Patibandla, RoshikNagaSai, et al.
Published: (2025)
by: Patibandla, RoshikNagaSai, et al.
Published: (2025)
MVAdapt: Zero-Shot Multi-Vehicle Adaptation for End-to-End Autonomous Driving
by: Oh, Haesung, et al.
Published: (2026)
by: Oh, Haesung, et al.
Published: (2026)
Enhancing Autonomous Driving Safety Analysis with Generative AI: A Comparative Study on Automated Hazard and Risk Assessment
by: Abbaspour, Alireza, et al.
Published: (2024)
by: Abbaspour, Alireza, et al.
Published: (2024)
Automated Data Curation Using GPS & NLP to Generate Instruction-Action Pairs for Autonomous Vehicle Vision-Language Navigation Datasets
by: Roque, Guillermo, et al.
Published: (2025)
by: Roque, Guillermo, et al.
Published: (2025)
Autonomous Manipulation of Hazardous Chemicals and Delicate Objects in a Self-Driving Laboratory: A Sliding Mode Approach
by: Sulaiman, Shifa, et al.
Published: (2026)
by: Sulaiman, Shifa, et al.
Published: (2026)
Agro-Consensus: Semantic Self-Consistency in Vision-Language Models for Crop Disease Management in Developing Countries
by: Gupta, Mihir, et al.
Published: (2025)
by: Gupta, Mihir, et al.
Published: (2025)
Evaluation of Safety Cognition Capability in Vision-Language Models for Autonomous Driving
by: Zhang, Enming, et al.
Published: (2025)
by: Zhang, Enming, et al.
Published: (2025)
Visuo-Tactile Zero-Shot Object Recognition with Vision-Language Model
by: Ueda, Shiori, et al.
Published: (2024)
by: Ueda, Shiori, et al.
Published: (2024)
Toward Autonomous Laboratory Safety Monitoring with Vision Language Models: Learning to See Hazards Through Scene Structure
by: Chakraborty, Trishna, et al.
Published: (2026)
by: Chakraborty, Trishna, et al.
Published: (2026)
Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making
by: Greer, Ross, et al.
Published: (2026)
by: Greer, Ross, et al.
Published: (2026)
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
by: Ghosh, Anurag, et al.
Published: (2026)
by: Ghosh, Anurag, et al.
Published: (2026)
Type-Error Ablation and AI Coding Agents
by: Krishnamurthi, Shriram, et al.
Published: (2026)
by: Krishnamurthi, Shriram, et al.
Published: (2026)
AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
Similar Items
-
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
by: Roy, Parthib, et al.
Published: (2024) -
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
by: Bossen, Tonko E. W., et al.
Published: (2025) -
Vision and Language: Novel Representations and Artificial intelligence for Driving Scene Safety Assessment and Autonomous Vehicle Planning
by: Greer, Ross, et al.
Published: (2026) -
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
by: Greer, Ross, et al.
Published: (2024) -
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
by: Choi, Lucas, et al.
Published: (2024)