An Empirical Study of Fault Localisation Techniques for Deep Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Humbatova, Nargiz, Kim, Jinhan, Jahangirova, Gunel, Yoo, Shin, Tonella, Paolo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
by: Jahangirova, Gunel, et al.
Published: (2024)
by: Jahangirova, Gunel, et al.
Published: (2024)
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
MuFF: Stable and Sensitive Post-training Mutation Testing for Deep Learning
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
Revisiting "Revisiting Neuron Coverage for DNN Testing: A Layer-Wise and Distribution-Aware Criterion": A Critical Review and Implications on DNN Coverage Testing
by: Kim, Jinhan, et al.
Published: (2026)
by: Kim, Jinhan, et al.
Published: (2026)
muPRL: A Mutation Testing Pipeline for Deep Reinforcement Learning based on Real Faults
by: Thomas, Deepak-George, et al.
Published: (2024)
by: Thomas, Deepak-George, et al.
Published: (2024)
TopoMap: A Feature-based Semantic Discriminator of the Topographical Regions in the Test Input Space
by: De Vita, Gianmarco, et al.
Published: (2025)
by: De Vita, Gianmarco, et al.
Published: (2025)
Understanding LLM-Driven Test Oracle Generation
by: Bodicoat, Adam, et al.
Published: (2026)
by: Bodicoat, Adam, et al.
Published: (2026)
Testing of Deep Reinforcement Learning Agents with Surrogate Models
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
Detecting Trojaned DNNs via Spectral Regression Analysis
by: Pasini, Samuele, et al.
Published: (2026)
by: Pasini, Samuele, et al.
Published: (2026)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
COSMosFL: Ensemble of Small Language Models for Fault Localisation
by: Cho, Hyunjoon, et al.
Published: (2025)
by: Cho, Hyunjoon, et al.
Published: (2025)
Cross-site scripting adversarial attacks based on deep reinforcement learning: Evaluation and extension study
by: Pasini, Samuele, et al.
Published: (2025)
by: Pasini, Samuele, et al.
Published: (2025)
Order Matters! An Empirical Study on Large Language Models' Input Order Bias in Software Fault Localization
by: Rafi, Md Nakhla, et al.
Published: (2024)
by: Rafi, Md Nakhla, et al.
Published: (2024)
GenMorph: Automatically Generating Metamorphic Relations via Genetic Programming
by: Ayerdi, Jon, et al.
Published: (2023)
by: Ayerdi, Jon, et al.
Published: (2023)
A Taxonomy of Real Faults in Hybrid Quantum-Classical Architectures
by: Bensoussan, Avner, et al.
Published: (2025)
by: Bensoussan, Avner, et al.
Published: (2025)
How Does Chunking Affect Retrieval-Augmented Code Completion? A Controlled Empirical Study
by: Wu, Xinjian, et al.
Published: (2026)
by: Wu, Xinjian, et al.
Published: (2026)
Boundary State Generation for Testing and Improvement of Autonomous Driving Systems
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
by: Vitale, Antonio, et al.
Published: (2026)
by: Vitale, Antonio, et al.
Published: (2026)
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
by: Zhao, Zhimin, et al.
Published: (2026)
by: Zhao, Zhimin, et al.
Published: (2026)
Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs
by: Pasini, Samuele, et al.
Published: (2024)
by: Pasini, Samuele, et al.
Published: (2024)
Efficient Domain Augmentation for Autonomous Driving Testing Using Diffusion Models
by: Baresi, Luciano, et al.
Published: (2024)
by: Baresi, Luciano, et al.
Published: (2024)
DANDI: Diffusion as Normative Distribution for Deep Neural Network Input
by: Kim, Somin, et al.
Published: (2025)
by: Kim, Somin, et al.
Published: (2025)
More with Less: An Empirical Study of Turn-Control Strategies for Efficient Coding Agents
by: Gao, Pengfei, et al.
Published: (2025)
by: Gao, Pengfei, et al.
Published: (2025)
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
by: Latendresse, Jasmine, et al.
Published: (2025)
by: Latendresse, Jasmine, et al.
Published: (2025)
Parameter-Efficient Fine-Tuning of Large Language Models for Unit Test Generation: An Empirical Study
by: Storhaug, André, et al.
Published: (2024)
by: Storhaug, André, et al.
Published: (2024)
DeepKnowledge: Generalisation-Driven Deep Learning Testing
by: Missaoui, Sondess, et al.
Published: (2024)
by: Missaoui, Sondess, et al.
Published: (2024)
SETA: Statistical Fault Attribution for Compound AI Systems
by: Chowdhury, Sayak, et al.
Published: (2026)
by: Chowdhury, Sayak, et al.
Published: (2026)
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation
by: Daghighfarsoodeh, Alireza, et al.
Published: (2025)
by: Daghighfarsoodeh, Alireza, et al.
Published: (2025)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
by: Khatoonabadi, SayedHassan, et al.
Published: (2023)
by: Khatoonabadi, SayedHassan, et al.
Published: (2023)
Assessing the Impact of Code Changes on the Fault Localizability of Large Language Models
by: Haroon, Sabaat, et al.
Published: (2025)
by: Haroon, Sabaat, et al.
Published: (2025)
On the Effectiveness of Machine Learning-based Call Graph Pruning: An Empirical Study
by: Mir, Amir M., et al.
Published: (2024)
by: Mir, Amir M., et al.
Published: (2024)
On the Replicability and Reproducibility of Deep Learning in Software Engineering
by: Liu, Chao, et al.
Published: (2020)
by: Liu, Chao, et al.
Published: (2020)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
by: Bouchoucha, Rached, et al.
Published: (2024)
by: Bouchoucha, Rached, et al.
Published: (2024)
Deep Configuration Performance Learning: A Systematic Survey and Taxonomy
by: Gong, Jingzhi, et al.
Published: (2024)
by: Gong, Jingzhi, et al.
Published: (2024)
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study
by: Dhar, Rudra, et al.
Published: (2024)
by: Dhar, Rudra, et al.
Published: (2024)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
by: Vulićević, Jelena Ilić
Published: (2026)
by: Vulićević, Jelena Ilić
Published: (2026)
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt
by: Sutoyo, Edi, et al.
Published: (2024)
by: Sutoyo, Edi, et al.
Published: (2024)
Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
by: Shah, Mehil B, et al.
Published: (2025)
by: Shah, Mehil B, et al.
Published: (2025)
SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents
by: Zolfagharian, Amirhossein, et al.
Published: (2023)
by: Zolfagharian, Amirhossein, et al.
Published: (2023)
Similar Items
-
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
by: Jahangirova, Gunel, et al.
Published: (2024) -
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
by: Kim, Jinhan, et al.
Published: (2025) -
MuFF: Stable and Sensitive Post-training Mutation Testing for Deep Learning
by: Kim, Jinhan, et al.
Published: (2025) -
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
by: Kim, Jinhan, et al.
Published: (2025) -
Revisiting "Revisiting Neuron Coverage for DNN Testing: A Layer-Wise and Distribution-Aware Criterion": A Critical Review and Implications on DNN Coverage Testing
by: Kim, Jinhan, et al.
Published: (2026)