What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Aich, Abhishek, Garg, Sparsh, BG, Vijay Kumar, Kashgari, Turgun Yusuf, Chandraker, Manmohan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations
by: Garg, Sparsh, et al.
Published: (2025)
by: Garg, Sparsh, et al.
Published: (2025)
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
by: Yao, Manyi, et al.
Published: (2025)
by: Yao, Manyi, et al.
Published: (2025)
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
by: Aich, Abhishek, et al.
Published: (2024)
by: Aich, Abhishek, et al.
Published: (2024)
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
by: Khan, Zaid, et al.
Published: (2024)
by: Khan, Zaid, et al.
Published: (2024)
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
by: Lambertenghi, Stefano Carlo, et al.
Published: (2025)
by: Lambertenghi, Stefano Carlo, et al.
Published: (2025)
ROMAN: Reward-Orchestrated Multi-Head Attention Network for Autonomous Driving System Testing
by: Chi, Jianlei, et al.
Published: (2026)
by: Chi, Jianlei, et al.
Published: (2026)
TigAug: Data Augmentation for Testing Traffic Light Detection in Autonomous Driving Systems
by: Lu, You, et al.
Published: (2025)
by: Lu, You, et al.
Published: (2025)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
by: Abbasishahkoo, Amin, et al.
Published: (2026)
by: Abbasishahkoo, Amin, et al.
Published: (2026)
How Far Can VLMs Go for Visual Bug Detection? Studying 19,738 Keyframes from 41 Hours of Gameplay Videos
by: Lu, Wentao, et al.
Published: (2026)
by: Lu, Wentao, et al.
Published: (2026)
CAFD: Concept-Aware DNN Fault Detection using VLMs
by: Abbasishahkoo, Amin, et al.
Published: (2026)
by: Abbasishahkoo, Amin, et al.
Published: (2026)
AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving
by: Liang, Mingfu, et al.
Published: (2024)
by: Liang, Mingfu, et al.
Published: (2024)
HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes
by: Soroco, Mauricio, et al.
Published: (2026)
by: Soroco, Mauricio, et al.
Published: (2026)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
by: Mzoughi, Seif, et al.
Published: (2025)
by: Mzoughi, Seif, et al.
Published: (2025)
Interpretable Gallbladder Ultrasound Diagnosis: A Lightweight Web-Mobile Software Platform with Real-Time XAI
by: Bhoyan, Fuyad Hasan, et al.
Published: (2025)
by: Bhoyan, Fuyad Hasan, et al.
Published: (2025)
Bias Behind the Wheel: Fairness Testing of Autonomous Driving Systems
by: Li, Xinyue, et al.
Published: (2023)
by: Li, Xinyue, et al.
Published: (2023)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
by: Lyu, Yunbo, et al.
Published: (2025)
by: Lyu, Yunbo, et al.
Published: (2025)
Coverage-Guided Road Selection and Prioritization for Efficient Testing in Autonomous Driving Systems
by: Ali, Qurban, et al.
Published: (2026)
by: Ali, Qurban, et al.
Published: (2026)
Efficient Domain Augmentation for Autonomous Driving Testing Using Diffusion Models
by: Baresi, Luciano, et al.
Published: (2024)
by: Baresi, Luciano, et al.
Published: (2024)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
by: Amini, Mohammad Hossein, et al.
Published: (2025)
by: Amini, Mohammad Hossein, et al.
Published: (2025)
Testing Autonomous Driving Systems -- What Really Matters and What Doesn't
by: Li, Changwen, et al.
Published: (2025)
by: Li, Changwen, et al.
Published: (2025)
MVOS_HSI: A Python Library for Preprocessing Agricultural Crop Hyperspectral Data
by: Aggarwal, Rishik, et al.
Published: (2026)
by: Aggarwal, Rishik, et al.
Published: (2026)
DiffGAN: A Test Generation Approach for Differential Testing of Deep Neural Networks for Image Analysis
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
Image-Specific Adaptation of Transformer Encoders for Compute-Efficient Segmentation
by: Yao, Manyi, et al.
Published: (2024)
by: Yao, Manyi, et al.
Published: (2024)
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025)
by: Duran, Matias, et al.
Published: (2025)
Distribution-aware Fairness Test Generation
by: Rajan, Sai Sathiesh, et al.
Published: (2023)
by: Rajan, Sai Sathiesh, et al.
Published: (2023)
Tuned Contrastive Learning
by: Animesh, Chaitanya, et al.
Published: (2023)
by: Animesh, Chaitanya, et al.
Published: (2023)
Towards Assessing Deep Learning Test Input Generators
by: Mzoughi, Seif, et al.
Published: (2025)
by: Mzoughi, Seif, et al.
Published: (2025)
Generalizing Test Cases for Comprehensive Test Scenario Coverage
by: Qi, Binhang, et al.
Published: (2026)
by: Qi, Binhang, et al.
Published: (2026)
Assessing Quality Metrics for Neural Reality Gap Input Mitigation in Autonomous Driving Testing
by: Lambertenghi, Stefano Carlo, et al.
Published: (2024)
by: Lambertenghi, Stefano Carlo, et al.
Published: (2024)
Computer Vision Intelligence Test Modeling and Generation: A Case Study on Smart OCR
by: Shu, Jing, et al.
Published: (2024)
by: Shu, Jing, et al.
Published: (2024)
Generalized Coverage Criteria for Combinatorial Sequence Testing
by: Elyasaf, Achiya, et al.
Published: (2022)
by: Elyasaf, Achiya, et al.
Published: (2022)
Testing Agentic Workflows with Structural Coverage Criteria
by: Kahani, Nafiseh, et al.
Published: (2026)
by: Kahani, Nafiseh, et al.
Published: (2026)
Towards Making Flowchart Images Machine Interpretable
by: Shukla, Shreya, et al.
Published: (2025)
by: Shukla, Shreya, et al.
Published: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
by: Lu, Zimu, et al.
Published: (2026)
by: Lu, Zimu, et al.
Published: (2026)
Testing with AI Agents: An Empirical Study of Test Generation Frequency, Quality, and Coverage
by: Yoshimoto, Suzuka, et al.
Published: (2026)
by: Yoshimoto, Suzuka, et al.
Published: (2026)
SWAN -- Enabling Fast and Mobile Histopathology Image Annotation through Swipeable Interfaces
by: Banerjee, Sweta, et al.
Published: (2025)
by: Banerjee, Sweta, et al.
Published: (2025)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
by: Khadka, Krishna, et al.
Published: (2026)
by: Khadka, Krishna, et al.
Published: (2026)
Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
by: Hong, Danfeng, et al.
Published: (2026)
by: Hong, Danfeng, et al.
Published: (2026)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
by: Chen, Yifei, et al.
Published: (2025)
by: Chen, Yifei, et al.
Published: (2025)
Similar Items
-
Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations
by: Garg, Sparsh, et al.
Published: (2025) -
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
by: Yao, Manyi, et al.
Published: (2025) -
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
by: Aich, Abhishek, et al.
Published: (2024) -
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
by: Khan, Zaid, et al.
Published: (2024) -
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
by: Lambertenghi, Stefano Carlo, et al.
Published: (2025)