Metamorphic Testing of Deep Code Models: A Systematic Literature Review
Fuente:
arXiv
Saved in:
| Main Authors: | Asgari, Ali, de Koning, Milan, Derakhshanfar, Pouria, Panichella, Annibale |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
by: De Koning, Milan, et al.
Published: (2026)
by: De Koning, Milan, et al.
Published: (2026)
What Challenges Do Developers Face in AI Agent Systems? An Empirical Study on Stack Overflow & GitHub Issues
by: Asgari, Ali, et al.
Published: (2025)
by: Asgari, Ali, et al.
Published: (2025)
Test Wars: A Comparative Study of SBST, Symbolic Execution, and LLM-Based Approaches to Unit Test Generation
by: Abdullin, Azat, et al.
Published: (2025)
by: Abdullin, Azat, et al.
Published: (2025)
TestSpark: IntelliJ IDEA's Ultimate Test Generation Companion
by: Sapozhnikov, Arkadii, et al.
Published: (2024)
by: Sapozhnikov, Arkadii, et al.
Published: (2024)
Evolutionary Generative Fuzzing for Differential Testing of the Kotlin Compiler
by: Georgescu, Calin, et al.
Published: (2024)
by: Georgescu, Calin, et al.
Published: (2024)
Automated Test-Case Generation for REST APIs Using Model Inference Search Heuristic
by: Cao, Clinton, et al.
Published: (2024)
by: Cao, Clinton, et al.
Published: (2024)
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
by: Bartlett, Antony, et al.
Published: (2025)
by: Bartlett, Antony, et al.
Published: (2025)
SBFT Tool Competition 2024 -- Python Test Case Generation Track
by: Erni, Nicolas, et al.
Published: (2024)
by: Erni, Nicolas, et al.
Published: (2024)
Automated Unit Test Case Generation: A Systematic Literature Review
by: Wang, Jason, et al.
Published: (2025)
by: Wang, Jason, et al.
Published: (2025)
Metamorphic Testing of Large Language Models for Natural Language Processing
by: Cho, Steven, et al.
Published: (2025)
by: Cho, Steven, et al.
Published: (2025)
Evaluating Human Trajectory Prediction with Metamorphic Testing
by: Spieker, Helge, et al.
Published: (2024)
by: Spieker, Helge, et al.
Published: (2024)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
by: Gao, Xuanqi, et al.
Published: (2025)
by: Gao, Xuanqi, et al.
Published: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
by: Wang, Xiaoyin, et al.
Published: (2024)
by: Wang, Xiaoyin, et al.
Published: (2024)
Large Language Models for Software Engineering: A Systematic Literature Review
by: Hou, Xinyi, et al.
Published: (2023)
by: Hou, Xinyi, et al.
Published: (2023)
A Systematic Literature Review on Explainability for Machine/Deep Learning-based Software Engineering Research
by: Cao, Sicong, et al.
Published: (2024)
by: Cao, Sicong, et al.
Published: (2024)
Search-based Selection of Metamorphic Relations for Optimized Robustness Testing of Large Language Models
by: Hyun, Sangwon, et al.
Published: (2025)
by: Hyun, Sangwon, et al.
Published: (2025)
Maintainability Challenges in ML: A Systematic Literature Review
by: Shivashankar, Karthik, et al.
Published: (2024)
by: Shivashankar, Karthik, et al.
Published: (2024)
A Systematic Literature Review on Detecting Software Vulnerabilities with Large Language Models
by: Kaniewski, Sabrina, et al.
Published: (2025)
by: Kaniewski, Sabrina, et al.
Published: (2025)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
by: Xu, Congying, et al.
Published: (2026)
by: Xu, Congying, et al.
Published: (2026)
Generative AI for Requirements Engineering: A Systematic Literature Review
by: Cheng, Haowei, et al.
Published: (2024)
by: Cheng, Haowei, et al.
Published: (2024)
Turbulence: Systematically and Automatically Testing Instruction-Tuned Large Language Models for Code
by: Honarvar, Shahin, et al.
Published: (2023)
by: Honarvar, Shahin, et al.
Published: (2023)
Novice Developers' Perspectives on Adopting LLMs for Software Development: A Systematic Literature Review
by: Ferino, Samuel, et al.
Published: (2025)
by: Ferino, Samuel, et al.
Published: (2025)
Breaking the Silence: the Threats of Using LLMs in Software Engineering
by: Sallou, June, et al.
Published: (2023)
by: Sallou, June, et al.
Published: (2023)
Metamorphic Testing for Audio Content Moderation Software
by: Wang, Wenxuan, et al.
Published: (2025)
by: Wang, Wenxuan, et al.
Published: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
by: Giramata, Suavis, et al.
Published: (2025)
by: Giramata, Suavis, et al.
Published: (2025)
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code
by: He, Kaifeng, et al.
Published: (2026)
by: He, Kaifeng, et al.
Published: (2026)
Impact and Implications of Generative AI for Enterprise Architects in Agile Environments: A Systematic Literature Review
by: Kooy, Stefan Julian, et al.
Published: (2025)
by: Kooy, Stefan Julian, et al.
Published: (2025)
Effective Black Box Testing of Sentiment Analysis Classification Networks
by: Karbasizadeh, Parsa, et al.
Published: (2024)
by: Karbasizadeh, Parsa, et al.
Published: (2024)
Are LLMs Reliable Code Reviewers? Systematic Overcorrection in Requirement Conformance Judgement
by: Jin, Haolin, et al.
Published: (2026)
by: Jin, Haolin, et al.
Published: (2026)
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
by: Zhou, Zenghui, et al.
Published: (2026)
by: Zhou, Zenghui, et al.
Published: (2026)
Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations
by: Kulshreshtha, Ashir, et al.
Published: (2026)
by: Kulshreshtha, Ashir, et al.
Published: (2026)
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025)
by: Duran, Matias, et al.
Published: (2025)
RefExpo: Unveiling Software Project Structures through Advanced Dependency Graph Extraction
by: Haratian, Vahid, et al.
Published: (2024)
by: Haratian, Vahid, et al.
Published: (2024)
Predicting the Understandability of Computational Notebooks through Code Metrics Analysis
by: Ghahfarokhi, Mojtaba Mostafavi, et al.
Published: (2024)
by: Ghahfarokhi, Mojtaba Mostafavi, et al.
Published: (2024)
Advances in Artificial Intelligence forDiabetes Prediction: Insights from a Systematic Literature Review
by: Khokhar, Pir Bakhsh, et al.
Published: (2024)
by: Khokhar, Pir Bakhsh, et al.
Published: (2024)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
by: Fu, Jia, et al.
Published: (2025)
by: Fu, Jia, et al.
Published: (2025)
Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review
by: Kamalı, Hüseyin Özgür, et al.
Published: (2026)
by: Kamalı, Hüseyin Özgür, et al.
Published: (2026)
Evaluating Large Language Models for Code Review
by: Cihan, Umut, et al.
Published: (2025)
by: Cihan, Umut, et al.
Published: (2025)
DeepCode: Open Agentic Coding
by: Li, Zongwei, et al.
Published: (2025)
by: Li, Zongwei, et al.
Published: (2025)
On Assessing the Relevance of Code Reviews Authored by Generative Models
by: Heumüller, Robert, et al.
Published: (2025)
by: Heumüller, Robert, et al.
Published: (2025)
Similar Items
-
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
by: De Koning, Milan, et al.
Published: (2026) -
What Challenges Do Developers Face in AI Agent Systems? An Empirical Study on Stack Overflow & GitHub Issues
by: Asgari, Ali, et al.
Published: (2025) -
Test Wars: A Comparative Study of SBST, Symbolic Execution, and LLM-Based Approaches to Unit Test Generation
by: Abdullin, Azat, et al.
Published: (2025) -
TestSpark: IntelliJ IDEA's Ultimate Test Generation Companion
by: Sapozhnikov, Arkadii, et al.
Published: (2024) -
Evolutionary Generative Fuzzing for Differential Testing of the Kotlin Compiler
by: Georgescu, Calin, et al.
Published: (2024)