MuFF: Stable and Sensitive Post-training Mutation Testing for Deep Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jinhan, Humbatova, Nargiz, Jahangirova, Gunel, Yoo, Shin, Tonella, Paolo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
Revisiting "Revisiting Neuron Coverage for DNN Testing: A Layer-Wise and Distribution-Aware Criterion": A Critical Review and Implications on DNN Coverage Testing
by: Kim, Jinhan, et al.
Published: (2026)
by: Kim, Jinhan, et al.
Published: (2026)
An Empirical Study of Fault Localisation Techniques for Deep Learning
by: Humbatova, Nargiz, et al.
Published: (2024)
by: Humbatova, Nargiz, et al.
Published: (2024)
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
by: Jahangirova, Gunel, et al.
Published: (2024)
by: Jahangirova, Gunel, et al.
Published: (2024)
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
muPRL: A Mutation Testing Pipeline for Deep Reinforcement Learning based on Real Faults
by: Thomas, Deepak-George, et al.
Published: (2024)
by: Thomas, Deepak-George, et al.
Published: (2024)
TopoMap: A Feature-based Semantic Discriminator of the Topographical Regions in the Test Input Space
by: De Vita, Gianmarco, et al.
Published: (2025)
by: De Vita, Gianmarco, et al.
Published: (2025)
GenMorph: Automatically Generating Metamorphic Relations via Genetic Programming
by: Ayerdi, Jon, et al.
Published: (2023)
by: Ayerdi, Jon, et al.
Published: (2023)
PCLA: A Framework for Testing Autonomous Agents in the CARLA Simulator
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2025)
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2025)
Understanding LLM-Driven Test Oracle Generation
by: Bodicoat, Adam, et al.
Published: (2026)
by: Bodicoat, Adam, et al.
Published: (2026)
Detecting Trojaned DNNs via Spectral Regression Analysis
by: Pasini, Samuele, et al.
Published: (2026)
by: Pasini, Samuele, et al.
Published: (2026)
A Taxonomy of Real Faults in Hybrid Quantum-Classical Architectures
by: Bensoussan, Avner, et al.
Published: (2025)
by: Bensoussan, Avner, et al.
Published: (2025)
How Does Chunking Affect Retrieval-Augmented Code Completion? A Controlled Empirical Study
by: Wu, Xinjian, et al.
Published: (2026)
by: Wu, Xinjian, et al.
Published: (2026)
A Taxonomy of System-Level Attacks on Deep Learning Models in Autonomous Vehicles
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2024)
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2024)
DevMuT: Testing Deep Learning Framework via Developer Expertise-Based Mutation
by: Mu, Yanzhou, et al.
Published: (2025)
by: Mu, Yanzhou, et al.
Published: (2025)
MuMuTestUp: Mutation-based Multi-Agent Test Case Update
by: Tian, Dawei, et al.
Published: (2026)
by: Tian, Dawei, et al.
Published: (2026)
Cross-site scripting adversarial attacks based on deep reinforcement learning: Evaluation and extension study
by: Pasini, Samuele, et al.
Published: (2025)
by: Pasini, Samuele, et al.
Published: (2025)
Testing of Deep Reinforcement Learning Agents with Surrogate Models
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
Adaptive Random Testing with Q-grams: The Illusion Comes True
by: Biagiola, Matteo, et al.
Published: (2024)
by: Biagiola, Matteo, et al.
Published: (2024)
Improving the Readability of Automatically Generated Tests using Large Language Models
by: Biagiola, Matteo, et al.
Published: (2024)
by: Biagiola, Matteo, et al.
Published: (2024)
MuSe: a Mutation Testing Plugin for the Remix IDE
by: Iuliano, Gerardo, et al.
Published: (2026)
by: Iuliano, Gerardo, et al.
Published: (2026)
Comparative Analysis of Carbon Footprint in Manual vs. LLM-Assisted Code Development
by: Cheung, Kuen Sum, et al.
Published: (2025)
by: Cheung, Kuen Sum, et al.
Published: (2025)
Boundary State Generation for Testing and Improvement of Autonomous Driving Systems
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
Neural Embeddings for Web Testing
by: Kanaththage, Kasun, et al.
Published: (2023)
by: Kanaththage, Kasun, et al.
Published: (2023)
Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs
by: Pasini, Samuele, et al.
Published: (2024)
by: Pasini, Samuele, et al.
Published: (2024)
DANDI: Diffusion as Normative Distribution for Deep Neural Network Input
by: Kim, Somin, et al.
Published: (2025)
by: Kim, Somin, et al.
Published: (2025)
Deep Learning Framework Testing via Model Mutation: How Far Are We?
by: Mu, Yanzhou, et al.
Published: (2025)
by: Mu, Yanzhou, et al.
Published: (2025)
Mutation-Based Deep Learning Framework Testing Method in JavaScript Environment
by: Zou, Yinglong, et al.
Published: (2024)
by: Zou, Yinglong, et al.
Published: (2024)
Bridging Research and Practice in Simulation-based Testing of Industrial Robot Navigation Systems
by: Khatiri, Sajad, et al.
Published: (2025)
by: Khatiri, Sajad, et al.
Published: (2025)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
by: Ravi, Ravin, et al.
Published: (2026)
by: Ravi, Ravin, et al.
Published: (2026)
Identifying Inaccurate Descriptions in LLM-generated Code Comments via Test Execution
by: Kang, Sungmin, et al.
Published: (2024)
by: Kang, Sungmin, et al.
Published: (2024)
Predicting Safety Misbehaviours in Autonomous Driving Systems using Uncertainty Quantification
by: Grewal, Ruben, et al.
Published: (2024)
by: Grewal, Ruben, et al.
Published: (2024)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
Benchmarking and Evaluating VLMs for Software Architecture Diagram Understanding
by: Ouyang, Shuyin, et al.
Published: (2026)
by: Ouyang, Shuyin, et al.
Published: (2026)
Are Benchmark Tests Strong Enough? Mutation-Guided Diagnosis and Augmentation of Regression Suites
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
by: Li, Ziyu, et al.
Published: (2024)
by: Li, Ziyu, et al.
Published: (2024)
When Uncertainty Leads to Unsafety: Empirical Insights into the Role of Uncertainty in Unmanned Aerial Vehicle Safety
by: Khatiri, Sajad, et al.
Published: (2025)
by: Khatiri, Sajad, et al.
Published: (2025)
Clotho: Measuring Task-Specific Pre-Generation Test Adequacy for LLM Inputs
by: Yoon, Juyeon, et al.
Published: (2025)
by: Yoon, Juyeon, et al.
Published: (2025)
Two is Better Than One: Digital Siblings to Improve Autonomous Driving Testing
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
Atropos: Improving Cost-Benefit Trade-off of LLM-based Agents under Self-Consistency with Early Termination and Model Hotswap
by: Kim, Naryeong, et al.
Published: (2026)
by: Kim, Naryeong, et al.
Published: (2026)
Similar Items
-
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
by: Kim, Jinhan, et al.
Published: (2025) -
Revisiting "Revisiting Neuron Coverage for DNN Testing: A Layer-Wise and Distribution-Aware Criterion": A Critical Review and Implications on DNN Coverage Testing
by: Kim, Jinhan, et al.
Published: (2026) -
An Empirical Study of Fault Localisation Techniques for Deep Learning
by: Humbatova, Nargiz, et al.
Published: (2024) -
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
by: Jahangirova, Gunel, et al.
Published: (2024) -
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
by: Kim, Jinhan, et al.
Published: (2025)