Metamorphic Testing of Large Language Models for Natural Language Processing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cho, Steven, Ruberto, Stefano, Terragni, Valerio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLMORPH: Automated Metamorphic Testing of Large Language Models
von: Cho, Steven, et al.
Veröffentlicht: (2026)
von: Cho, Steven, et al.
Veröffentlicht: (2026)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
von: Terragni, Valerio
Veröffentlicht: (2026)
von: Terragni, Valerio
Veröffentlicht: (2026)
Automated Trustworthiness Testing for Machine Learning Classifiers
von: Cho, Steven, et al.
Veröffentlicht: (2024)
von: Cho, Steven, et al.
Veröffentlicht: (2024)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
von: Xu, Congying, et al.
Veröffentlicht: (2026)
von: Xu, Congying, et al.
Veröffentlicht: (2026)
Understanding LLM-Driven Test Oracle Generation
von: Bodicoat, Adam, et al.
Veröffentlicht: (2026)
von: Bodicoat, Adam, et al.
Veröffentlicht: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
von: Ravi, Ravin, et al.
Veröffentlicht: (2026)
von: Ravi, Ravin, et al.
Veröffentlicht: (2026)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
Search-based Selection of Metamorphic Relations for Optimized Robustness Testing of Large Language Models
von: Hyun, Sangwon, et al.
Veröffentlicht: (2025)
von: Hyun, Sangwon, et al.
Veröffentlicht: (2025)
MR-Scout: Automated Synthesis of Metamorphic Relations from Existing Test Cases
von: Xu, Congying, et al.
Veröffentlicht: (2023)
von: Xu, Congying, et al.
Veröffentlicht: (2023)
Evaluating Human Trajectory Prediction with Metamorphic Testing
von: Spieker, Helge, et al.
Veröffentlicht: (2024)
von: Spieker, Helge, et al.
Veröffentlicht: (2024)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
von: Asgari, Ali, et al.
Veröffentlicht: (2025)
von: Asgari, Ali, et al.
Veröffentlicht: (2025)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
von: Gao, Xuanqi, et al.
Veröffentlicht: (2025)
von: Gao, Xuanqi, et al.
Veröffentlicht: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
MR-Adopt: Automatic Deduction of Input Transformation Function for Metamorphic Testing
von: Xu, Congying, et al.
Veröffentlicht: (2024)
von: Xu, Congying, et al.
Veröffentlicht: (2024)
GenMorph: Automatically Generating Metamorphic Relations via Genetic Programming
von: Ayerdi, Jon, et al.
Veröffentlicht: (2023)
von: Ayerdi, Jon, et al.
Veröffentlicht: (2023)
Assessing the Business Process Modeling Competences of Large Language Models
von: Lauer, Chantale, et al.
Veröffentlicht: (2026)
von: Lauer, Chantale, et al.
Veröffentlicht: (2026)
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
von: De Koning, Milan, et al.
Veröffentlicht: (2026)
von: De Koning, Milan, et al.
Veröffentlicht: (2026)
Large Language Models for Software Testing: A Research Roadmap
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
Exploring the Integration of Large Language Models in Industrial Test Maintenance Processes
von: Liu, Jingxiong, et al.
Veröffentlicht: (2024)
von: Liu, Jingxiong, et al.
Veröffentlicht: (2024)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
Large Language Models as Test Case Generators: Performance Evaluation and Enhancement
von: Li, Kefan, et al.
Veröffentlicht: (2024)
von: Li, Kefan, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
exLong: Generating Exceptional Behavior Tests with Large Language Models
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
Navigating the Labyrinth: Path-Sensitive Unit Test Generation with Large Language Models
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
LangBiTe: A Platform for Testing Bias in Large Language Models
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
Turbulence: Systematically and Automatically Testing Instruction-Tuned Large Language Models for Code
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
A System for Automated Unit Test Generation Using Large Language Models and Assessment of Generated Test Suites
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
HFuzzer: Testing Large Language Models for Package Hallucinations via Phrase-based Fuzzing
von: Zhao, Yukai, et al.
Veröffentlicht: (2025)
von: Zhao, Yukai, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for the Generation of Unit Tests with Equivalence Partitions and Boundary Values
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
Automated User Story Generation with Test Case Specification Using Large Language Model
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
Metamorphic Testing for Audio Content Moderation Software
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents
von: Lin, Feng, et al.
Veröffentlicht: (2024)
von: Lin, Feng, et al.
Veröffentlicht: (2024)
Comparison of Static Application Security Testing Tools and Large Language Models for Repo-level Vulnerability Detection
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
LlamaRestTest: Effective REST API Testing with Small Language Models
von: Kim, Myeongsoo, et al.
Veröffentlicht: (2025)
von: Kim, Myeongsoo, et al.
Veröffentlicht: (2025)
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
von: Zhou, Zenghui, et al.
Veröffentlicht: (2026)
von: Zhou, Zenghui, et al.
Veröffentlicht: (2026)
Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations
von: Kulshreshtha, Ashir, et al.
Veröffentlicht: (2026)
von: Kulshreshtha, Ashir, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LLMORPH: Automated Metamorphic Testing of Large Language Models
von: Cho, Steven, et al.
Veröffentlicht: (2026) -
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
von: Terragni, Valerio
Veröffentlicht: (2026) -
Automated Trustworthiness Testing for Machine Learning Classifiers
von: Cho, Steven, et al.
Veröffentlicht: (2024) -
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
von: Xu, Congying, et al.
Veröffentlicht: (2026) -
Understanding LLM-Driven Test Oracle Generation
von: Bodicoat, Adam, et al.
Veröffentlicht: (2026)