Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Mzoughi, Seif, Elshafeia, Mohamed, Khomh, Foutse |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Assessing Deep Learning Test Input Generators
by: Mzoughi, Seif, et al.
Published: (2025)
by: Mzoughi, Seif, et al.
Published: (2025)
Trimming the Risk: Towards Reliable Continuous Training for Deep Learning Inspection Systems
by: Abbassi, Altaf Allah, et al.
Published: (2024)
by: Abbassi, Altaf Allah, et al.
Published: (2024)
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025)
by: Duran, Matias, et al.
Published: (2025)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024)
by: Abukhalaf, Seif, et al.
Published: (2024)
Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
by: Lyu, Yunbo, et al.
Published: (2025)
by: Lyu, Yunbo, et al.
Published: (2025)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
CUARewardBench: A Benchmark for Evaluating Reward Models on Computer-using Agent
by: Lin, Haojia, et al.
Published: (2025)
by: Lin, Haojia, et al.
Published: (2025)
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
by: Lambertenghi, Stefano Carlo, et al.
Published: (2025)
by: Lambertenghi, Stefano Carlo, et al.
Published: (2025)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
by: Aich, Abhishek, et al.
Published: (2026)
by: Aich, Abhishek, et al.
Published: (2026)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
TigAug: Data Augmentation for Testing Traffic Light Detection in Autonomous Driving Systems
by: Lu, You, et al.
Published: (2025)
by: Lu, You, et al.
Published: (2025)
ROMAN: Reward-Orchestrated Multi-Head Attention Network for Autonomous Driving System Testing
by: Chi, Jianlei, et al.
Published: (2026)
by: Chi, Jianlei, et al.
Published: (2026)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
by: Amini, Mohammad Hossein, et al.
Published: (2025)
by: Amini, Mohammad Hossein, et al.
Published: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
by: Lu, Zimu, et al.
Published: (2026)
by: Lu, Zimu, et al.
Published: (2026)
Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
by: Hong, Danfeng, et al.
Published: (2026)
by: Hong, Danfeng, et al.
Published: (2026)
Investigating Traffic Accident Detection Using Multimodal Large Language Models
by: Skender, Ilhan, et al.
Published: (2025)
by: Skender, Ilhan, et al.
Published: (2025)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
by: Khadka, Krishna, et al.
Published: (2026)
by: Khadka, Krishna, et al.
Published: (2026)
GUing: A Mobile GUI Search Engine using a Vision-Language Model
by: Wei, Jialiang, et al.
Published: (2024)
by: Wei, Jialiang, et al.
Published: (2024)
Can Vision-Language Models Handle Long-Context Code? An Empirical Study on Visual Compression
by: Zhong, Jianping, et al.
Published: (2026)
by: Zhong, Jianping, et al.
Published: (2026)
Computer Vision Intelligence Test Modeling and Generation: A Case Study on Smart OCR
by: Shu, Jing, et al.
Published: (2024)
by: Shu, Jing, et al.
Published: (2024)
DiffGAN: A Test Generation Approach for Differential Testing of Deep Neural Networks for Image Analysis
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
Mock Deep Testing: Toward Separate Development of Data and Models for Deep Learning
by: Manke, Ruchira, et al.
Published: (2025)
by: Manke, Ruchira, et al.
Published: (2025)
Distribution-aware Fairness Test Generation
by: Rajan, Sai Sathiesh, et al.
Published: (2023)
by: Rajan, Sai Sathiesh, et al.
Published: (2023)
ChartMimic: Evaluating LMM's Cross-Modal Reasoning Capability via Chart-to-Code Generation
by: Yang, Cheng, et al.
Published: (2024)
by: Yang, Cheng, et al.
Published: (2024)
SWAN -- Enabling Fast and Mobile Histopathology Image Annotation through Swipeable Interfaces
by: Banerjee, Sweta, et al.
Published: (2025)
by: Banerjee, Sweta, et al.
Published: (2025)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
by: Chen, Yifei, et al.
Published: (2025)
by: Chen, Yifei, et al.
Published: (2025)
Cross-Breed Pig Identification Using Auricular Vein Pattern Recognition: A Machine Learning Approach for Small-Scale Farming Applications
by: Nsengiyumvaa, Emmanuel, et al.
Published: (2025)
by: Nsengiyumvaa, Emmanuel, et al.
Published: (2025)
Interpretable Gallbladder Ultrasound Diagnosis: A Lightweight Web-Mobile Software Platform with Real-Time XAI
by: Bhoyan, Fuyad Hasan, et al.
Published: (2025)
by: Bhoyan, Fuyad Hasan, et al.
Published: (2025)
ARI3D: A Software for Interactive Quantification of Regions in X-Ray CT 3D Images
by: Albrecht, Jan Phillipp, et al.
Published: (2025)
by: Albrecht, Jan Phillipp, et al.
Published: (2025)
Open-source automatic pipeline for efficient conversion of large-scale point clouds to IFC format
by: Zbirovský, Slávek, et al.
Published: (2025)
by: Zbirovský, Slávek, et al.
Published: (2025)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
by: Khalid, Waleed, et al.
Published: (2025)
by: Khalid, Waleed, et al.
Published: (2025)
How Far Can VLMs Go for Visual Bug Detection? Studying 19,738 Keyframes from 41 Hours of Gameplay Videos
by: Lu, Wentao, et al.
Published: (2026)
by: Lu, Wentao, et al.
Published: (2026)
Natural Adversaries: Fuzzing Autonomous Vehicles with Realistic Roadside Object Placements
by: Sun, Yang, et al.
Published: (2024)
by: Sun, Yang, et al.
Published: (2024)
Earth Embeddings as Products: Taxonomy, Ecosystem, and Standardized Access
by: Fang, Heng, et al.
Published: (2026)
by: Fang, Heng, et al.
Published: (2026)
Ear-Keeper: A Cross-Platform AI System for Rapid and Accurate Ear Disease Diagnosis
by: Lu, Feiyan, et al.
Published: (2023)
by: Lu, Feiyan, et al.
Published: (2023)
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
by: Meng, Fanqing, et al.
Published: (2026)
by: Meng, Fanqing, et al.
Published: (2026)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
by: Abbasishahkoo, Amin, et al.
Published: (2026)
by: Abbasishahkoo, Amin, et al.
Published: (2026)
A Plausibility Study of Using Augmented Reality in the Ventriculoperitoneal Shunt Operations
by: Dorji, Tandin, et al.
Published: (2024)
by: Dorji, Tandin, et al.
Published: (2024)
VideoGameBunny: Towards vision assistants for video games
by: Taesiri, Mohammad Reza, et al.
Published: (2024)
by: Taesiri, Mohammad Reza, et al.
Published: (2024)
DOne: Decoupling Structure and Rendering for High-Fidelity Design-to-Code Generation
by: Huang, Xinhao, et al.
Published: (2026)
by: Huang, Xinhao, et al.
Published: (2026)
Similar Items
-
Towards Assessing Deep Learning Test Input Generators
by: Mzoughi, Seif, et al.
Published: (2025) -
Trimming the Risk: Towards Reliable Continuous Training for Deep Learning Inspection Systems
by: Abbassi, Altaf Allah, et al.
Published: (2024) -
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025) -
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024) -
Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
by: Lyu, Yunbo, et al.
Published: (2025)