Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
Fuente:
arXiv
Saved in:
| Main Authors: | Shah, Mehil B, Rahman, Mohammad Masudur, Khomh, Foutse |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Enhancing the Reproducibility of Deep Learning Bugs: An Empirical Study
by: Shah, Mehil B., et al.
Published: (2024)
by: Shah, Mehil B., et al.
Published: (2024)
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
by: Shah, Mehil B, et al.
Published: (2024)
by: Shah, Mehil B, et al.
Published: (2024)
Improved Bug Localization with AI Agents Leveraging Hypothesis and Dynamic Cognition
by: Samir, Asif Mohammed, et al.
Published: (2026)
by: Samir, Asif Mohammed, et al.
Published: (2026)
Towards Understanding the Challenges of Bug Localization in Deep Learning Systems
by: Jahan, Sigma, et al.
Published: (2024)
by: Jahan, Sigma, et al.
Published: (2024)
Characterizing Faults in Agentic AI: A Taxonomy of Types, Symptoms, and Root Causes
by: Shah, Mehil B, et al.
Published: (2026)
by: Shah, Mehil B, et al.
Published: (2026)
Machine Learning Robustness: A Primer
by: Braiek, Houssem Ben, et al.
Published: (2024)
by: Braiek, Houssem Ben, et al.
Published: (2024)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
by: Majdinasab, Vahid, et al.
Published: (2025)
by: Majdinasab, Vahid, et al.
Published: (2025)
What Information Contributes to Log-based Anomaly Detection? Insights from a Configurable Transformer-Based Approach
by: Wu, Xingfang, et al.
Published: (2024)
by: Wu, Xingfang, et al.
Published: (2024)
Be a Partner, not a Bystander in Software Engineering Practice: Bridging the Gaps between Academia and Industry
by: Rahman, Mohammad Masudur, et al.
Published: (2026)
by: Rahman, Mohammad Masudur, et al.
Published: (2026)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
by: Bouchoucha, Rached, et al.
Published: (2024)
by: Bouchoucha, Rached, et al.
Published: (2024)
Leveraging Data Characteristics for Bug Localization in Deep Learning Programs
by: Manke, Ruchira, et al.
Published: (2024)
by: Manke, Ruchira, et al.
Published: (2024)
Fault Localization in Deep Learning-based Software: A System-level Approach
by: Morovati, Mohammad Mehdi, et al.
Published: (2024)
by: Morovati, Mohammad Mehdi, et al.
Published: (2024)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
A Survey of Bugs in AI-Generated Code
by: Gao, Ruofan, et al.
Published: (2025)
by: Gao, Ruofan, et al.
Published: (2025)
GIST: Generated Inputs Sets Transferability in Deep Learning
by: Tambon, Florian, et al.
Published: (2023)
by: Tambon, Florian, et al.
Published: (2023)
Improved Detection and Diagnosis of Faults in Deep Neural Networks Using Hierarchical and Explainable Classification
by: Jahan, Sigma, et al.
Published: (2025)
by: Jahan, Sigma, et al.
Published: (2025)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024)
by: Abukhalaf, Seif, et al.
Published: (2024)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
by: Oueslati, Khouloud, et al.
Published: (2025)
by: Oueslati, Khouloud, et al.
Published: (2025)
SDLog: A Deep Learning Framework for Detecting Sensitive Information in Software Logs
by: Aghili, Roozbeh, et al.
Published: (2025)
by: Aghili, Roozbeh, et al.
Published: (2025)
Improved IR-based Bug Localization with Intelligent Relevance Feedback
by: Samir, Asif Mohammed, et al.
Published: (2025)
by: Samir, Asif Mohammed, et al.
Published: (2025)
On the Replicability and Reproducibility of Deep Learning in Software Engineering
by: Liu, Chao, et al.
Published: (2020)
by: Liu, Chao, et al.
Published: (2020)
Bugs in Large Language Models Generated Code: An Empirical Study
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
FairFLRep: Fairness aware fault localization and repair of Deep Neural Networks
by: Openja, Moses, et al.
Published: (2025)
by: Openja, Moses, et al.
Published: (2025)
Can Hessian-Based Insights Support Fault Diagnosis in Attention-based Models?
by: Jahan, Sigma, et al.
Published: (2025)
by: Jahan, Sigma, et al.
Published: (2025)
SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents
by: Mündler, Niels, et al.
Published: (2024)
by: Mündler, Niels, et al.
Published: (2024)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
by: Wu, Xingfang, et al.
Published: (2023)
by: Wu, Xingfang, et al.
Published: (2023)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
by: Da Silva, Leuson, et al.
Published: (2024)
by: Da Silva, Leuson, et al.
Published: (2024)
Explaining Software Bugs Leveraging Code Structures in Neural Machine Translation
by: Mahbub, Parvez, et al.
Published: (2022)
by: Mahbub, Parvez, et al.
Published: (2022)
Are Large Language Models Memorizing Bug Benchmarks?
by: Ramos, Daniel, et al.
Published: (2024)
by: Ramos, Daniel, et al.
Published: (2024)
Are Sparse Autoencoders Useful for Java Function Bug Detection?
by: Melo, Rui, et al.
Published: (2025)
by: Melo, Rui, et al.
Published: (2025)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
by: Taraghi, Mina, et al.
Published: (2024)
by: Taraghi, Mina, et al.
Published: (2024)
BugMentor: Generating Answers to Follow-up Questions from Software Bug Reports using Structured Information Retrieval and Neural Text Generation
by: Mukherjee, Usmi, et al.
Published: (2023)
by: Mukherjee, Usmi, et al.
Published: (2023)
Mining Action Rules for Defect Reduction Planning
by: Oueslati, Khouloud, et al.
Published: (2024)
by: Oueslati, Khouloud, et al.
Published: (2024)
An Efficient Model Maintenance Approach for MLOps
by: Majidi, Forough, et al.
Published: (2024)
by: Majidi, Forough, et al.
Published: (2024)
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation
by: Daghighfarsoodeh, Alireza, et al.
Published: (2025)
by: Daghighfarsoodeh, Alireza, et al.
Published: (2025)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
by: Vitale, Antonio, et al.
Published: (2026)
by: Vitale, Antonio, et al.
Published: (2026)
SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents
by: Zolfagharian, Amirhossein, et al.
Published: (2023)
by: Zolfagharian, Amirhossein, et al.
Published: (2023)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
by: Vulićević, Jelena Ilić
Published: (2026)
by: Vulićević, Jelena Ilić
Published: (2026)
VibeTensor: System Software for Deep Learning, Fully Generated by AI Agents
by: Xu, Bing, et al.
Published: (2026)
by: Xu, Bing, et al.
Published: (2026)
Similar Items
-
Towards Enhancing the Reproducibility of Deep Learning Bugs: An Empirical Study
by: Shah, Mehil B., et al.
Published: (2024) -
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
by: Shah, Mehil B, et al.
Published: (2024) -
Improved Bug Localization with AI Agents Leveraging Hypothesis and Dynamic Cognition
by: Samir, Asif Mohammed, et al.
Published: (2026) -
Towards Understanding the Challenges of Bug Localization in Deep Learning Systems
by: Jahan, Sigma, et al.
Published: (2024) -
Characterizing Faults in Agentic AI: A Taxonomy of Types, Symptoms, and Root Causes
by: Shah, Mehil B, et al.
Published: (2026)