Validating LLM-Generated Programs with Metamorphic Prompt Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xiaoyin, Zhu, Dakai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
by: De Koning, Milan, et al.
Published: (2026)
by: De Koning, Milan, et al.
Published: (2026)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
by: Xu, Congying, et al.
Published: (2026)
by: Xu, Congying, et al.
Published: (2026)
Evaluating Human Trajectory Prediction with Metamorphic Testing
by: Spieker, Helge, et al.
Published: (2024)
by: Spieker, Helge, et al.
Published: (2024)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
by: Gao, Xuanqi, et al.
Published: (2025)
by: Gao, Xuanqi, et al.
Published: (2025)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
by: Cui, Yi
Published: (2025)
by: Cui, Yi
Published: (2025)
Metamorphic Testing of Large Language Models for Natural Language Processing
by: Cho, Steven, et al.
Published: (2025)
by: Cho, Steven, et al.
Published: (2025)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
by: Asgari, Ali, et al.
Published: (2025)
by: Asgari, Ali, et al.
Published: (2025)
LLM Test Generation via Iterative Hybrid Program Analysis
by: Gu, Sijia, et al.
Published: (2025)
by: Gu, Sijia, et al.
Published: (2025)
Static Program Analysis Guided LLM Based Unit Test Generation
by: Roychowdhury, Sujoy, et al.
Published: (2025)
by: Roychowdhury, Sujoy, et al.
Published: (2025)
The Prompt Alchemist: Automated LLM-Tailored Prompt Optimization for Test Case Generation
by: Gao, Shuzheng, et al.
Published: (2025)
by: Gao, Shuzheng, et al.
Published: (2025)
Metamorphic Testing for Audio Content Moderation Software
by: Wang, Wenxuan, et al.
Published: (2025)
by: Wang, Wenxuan, et al.
Published: (2025)
PromptPex: Automatic Test Generation for Language Model Prompts
by: Sharma, Reshabh K, et al.
Published: (2025)
by: Sharma, Reshabh K, et al.
Published: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
by: Baqar, Mohammad, et al.
Published: (2024)
by: Baqar, Mohammad, et al.
Published: (2024)
TENET: Leveraging Tests Beyond Validation for Code Generation
by: Hu, Yiran, et al.
Published: (2025)
by: Hu, Yiran, et al.
Published: (2025)
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
by: Zhou, Zenghui, et al.
Published: (2026)
by: Zhou, Zenghui, et al.
Published: (2026)
Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations
by: Kulshreshtha, Ashir, et al.
Published: (2026)
by: Kulshreshtha, Ashir, et al.
Published: (2026)
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025)
by: Duran, Matias, et al.
Published: (2025)
Integrating Artificial Intelligence with Human Expertise: An In-depth Analysis of ChatGPT's Capabilities in Generating Metamorphic Relations
by: Zhang, Yifan, et al.
Published: (2025)
by: Zhang, Yifan, et al.
Published: (2025)
Call-Chain-Aware LLM-Based Test Generation for Java Projects
by: Wang, Guancheng, et al.
Published: (2026)
by: Wang, Guancheng, et al.
Published: (2026)
The Readability Spectrum: Patterns, Issues, and Prompt Effects in LLM-Generated Code
by: Ye, Hengzhi, et al.
Published: (2026)
by: Ye, Hengzhi, et al.
Published: (2026)
STELP: Secure Transpilation and Execution of LLM-Generated Programs
by: Shinde, Swapnil, et al.
Published: (2026)
by: Shinde, Swapnil, et al.
Published: (2026)
VerilogReader: LLM-Aided Hardware Test Generation
by: Ma, Ruiyang, et al.
Published: (2024)
by: Ma, Ruiyang, et al.
Published: (2024)
Can LLM Generate Regression Tests for Software Commits?
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Bias Testing and Mitigation in LLM-based Code Generation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
Online Prompt Selection for Program Synthesis
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
by: Giramata, Suavis, et al.
Published: (2025)
by: Giramata, Suavis, et al.
Published: (2025)
Evaluating LLM-Based Test Generation Under Software Evolution
by: Haroon, Sabaat, et al.
Published: (2026)
by: Haroon, Sabaat, et al.
Published: (2026)
HoarePrompt: Structural Reasoning About Program Correctness in Natural Language
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
by: Nandal, Deeksha, et al.
Published: (2026)
by: Nandal, Deeksha, et al.
Published: (2026)
Impact of Code Context and Prompting Strategies on Automated Unit Test Generation with Modern General-Purpose Large Language Models
by: Walczak, Jakub, et al.
Published: (2025)
by: Walczak, Jakub, et al.
Published: (2025)
Prompt2DAG: A Modular Methodology for LLM-Based Data Enrichment Pipeline Generation
by: Alidu, Abubakari, et al.
Published: (2025)
by: Alidu, Abubakari, et al.
Published: (2025)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
by: Gan, Tiantian, et al.
Published: (2025)
by: Gan, Tiantian, et al.
Published: (2025)
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming
by: Sung, Sicheol, et al.
Published: (2025)
by: Sung, Sicheol, et al.
Published: (2025)
CoCoEvo: Co-Evolution of Programs and Test Cases to Enhance Code Generation
by: Li, Kefan, et al.
Published: (2025)
by: Li, Kefan, et al.
Published: (2025)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
by: Chen, Zhi, et al.
Published: (2026)
by: Chen, Zhi, et al.
Published: (2026)
Investigating The Smells of LLM Generated Code
by: Paul, Debalina Ghosh, et al.
Published: (2025)
by: Paul, Debalina Ghosh, et al.
Published: (2025)
LLM-Based Test Case Generation in DBMS through Monte Carlo Tree Search
by: Chen, Yujia, et al.
Published: (2026)
by: Chen, Yujia, et al.
Published: (2026)
Search-based Selection of Metamorphic Relations for Optimized Robustness Testing of Large Language Models
by: Hyun, Sangwon, et al.
Published: (2025)
by: Hyun, Sangwon, et al.
Published: (2025)
Task-oriented Prompt Enhancement via Script Generation
by: Wang, Chung-Yu, et al.
Published: (2024)
by: Wang, Chung-Yu, et al.
Published: (2024)
Automated Validation of LLM-based Evaluators for Software Engineering Artifacts
by: Fandina, Ora Nova, et al.
Published: (2025)
by: Fandina, Ora Nova, et al.
Published: (2025)
Similar Items
-
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
by: De Koning, Milan, et al.
Published: (2026) -
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
by: Xu, Congying, et al.
Published: (2026) -
Evaluating Human Trajectory Prediction with Metamorphic Testing
by: Spieker, Helge, et al.
Published: (2024) -
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
by: Gao, Xuanqi, et al.
Published: (2025) -
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
by: Cui, Yi
Published: (2025)