Large Language Models as Test Case Generators: Performance Evaluation and Enhancement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Kefan, Yuan, Yuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoCoEvo: Co-Evolution of Programs and Test Cases to Enhance Code Generation
von: Li, Kefan, et al.
Veröffentlicht: (2025)
von: Li, Kefan, et al.
Veröffentlicht: (2025)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
Automated User Story Generation with Test Case Specification Using Large Language Model
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for the Generation of Unit Tests with Equivalence Partitions and Boundary Values
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
Task Abstention for Large Language Models in Code Generation
von: Zhou, Yanke, et al.
Veröffentlicht: (2026)
von: Zhou, Yanke, et al.
Veröffentlicht: (2026)
Output Format Biases in the Evaluation of Large Language Models for Code Translation
von: Macedo, Marcos, et al.
Veröffentlicht: (2024)
von: Macedo, Marcos, et al.
Veröffentlicht: (2024)
exLong: Generating Exceptional Behavior Tests with Large Language Models
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
CrossPL: Evaluating Large Language Models on Cross Programming Language Code Generation
von: Xiong, Zhanhang, et al.
Veröffentlicht: (2025)
von: Xiong, Zhanhang, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
A System for Automated Unit Test Generation Using Large Language Models and Assessment of Generated Test Suites
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
Navigating the Labyrinth: Path-Sensitive Unit Test Generation with Large Language Models
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
Metamorphic Testing of Large Language Models for Natural Language Processing
von: Cho, Steven, et al.
Veröffentlicht: (2025)
von: Cho, Steven, et al.
Veröffentlicht: (2025)
Legal Compliance Evaluation of Smart Contracts Generated By Large Language Models
von: Wijayakoon, Chanuka, et al.
Veröffentlicht: (2025)
von: Wijayakoon, Chanuka, et al.
Veröffentlicht: (2025)
A Comprehensive Framework for Evaluating API-oriented Code Generation in Large Language Models
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for Functional and Maintainable Code in Industrial Settings: A Case Study at ASML
von: Mundhra, Yash, et al.
Veröffentlicht: (2025)
von: Mundhra, Yash, et al.
Veröffentlicht: (2025)
RobuNFR: Evaluating the Robustness of Large Language Models on Non-Functional Requirements Aware Code Generation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Impact of Code Context and Prompting Strategies on Automated Unit Test Generation with Modern General-Purpose Large Language Models
von: Walczak, Jakub, et al.
Veröffentlicht: (2025)
von: Walczak, Jakub, et al.
Veröffentlicht: (2025)
Large Language Models for Software Testing: A Research Roadmap
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
COMPASS: A Multi-Dimensional Benchmark for Evaluating Code Generation in Large Language Models
von: Meaden, James, et al.
Veröffentlicht: (2025)
von: Meaden, James, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index
von: Christakis, Nicholas, et al.
Veröffentlicht: (2025)
von: Christakis, Nicholas, et al.
Veröffentlicht: (2025)
Domain Adaptation for Code Model-based Unit Test Case Generation
von: Shin, Jiho, et al.
Veröffentlicht: (2023)
von: Shin, Jiho, et al.
Veröffentlicht: (2023)
Evaluating Large Language Models for Code Review
von: Cihan, Umut, et al.
Veröffentlicht: (2025)
von: Cihan, Umut, et al.
Veröffentlicht: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
von: Baqar, Mohammad, et al.
Veröffentlicht: (2024)
von: Baqar, Mohammad, et al.
Veröffentlicht: (2024)
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
von: He, Mengliang, et al.
Veröffentlicht: (2025)
von: He, Mengliang, et al.
Veröffentlicht: (2025)
The Performance of the LSTM-based Code Generated by Large Language Models (LLMs) in Forecasting Time Series Data
von: Gopali, Saroj, et al.
Veröffentlicht: (2024)
von: Gopali, Saroj, et al.
Veröffentlicht: (2024)
PromptPex: Automatic Test Generation for Language Model Prompts
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
GenAI-powered Multi-Agent Paradigm for Smart Urban Mobility: Opportunities and Challenges for Integrating Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) with Intelligent Transportation Systems
von: Xu, Haowen, et al.
Veröffentlicht: (2024)
von: Xu, Haowen, et al.
Veröffentlicht: (2024)
Beyond Autoregression: An Empirical Study of Diffusion Large Language Models for Code Generation
von: Li, Chengze, et al.
Veröffentlicht: (2025)
von: Li, Chengze, et al.
Veröffentlicht: (2025)
CodeGolf Bench: A Multi-Language Benchmark for Evaluating Concise Code Generation Capabilities of Large Language Models
von: Padwal, Vedant
Veröffentlicht: (2026)
von: Padwal, Vedant
Veröffentlicht: (2026)
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
von: Taherkhani, Hamed, et al.
Veröffentlicht: (2024)
von: Taherkhani, Hamed, et al.
Veröffentlicht: (2024)
Testing and Evaluation of Large Language Models: Correctness, Non-Toxicity, and Fairness
von: Wang, Wenxuan
Veröffentlicht: (2024)
von: Wang, Wenxuan
Veröffentlicht: (2024)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025)
von: Fu, Jia, et al.
Veröffentlicht: (2025)
LangBiTe: A Platform for Testing Bias in Large Language Models
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
Turbulence: Systematically and Automatically Testing Instruction-Tuned Large Language Models for Code
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
Automated Test-Case Generation for REST APIs Using Model Inference Search Heuristic
von: Cao, Clinton, et al.
Veröffentlicht: (2024)
von: Cao, Clinton, et al.
Veröffentlicht: (2024)
Worst-Case Symbolic Constraints Analysis and Generalisation with Large Language Models
von: Koh, Daniel, et al.
Veröffentlicht: (2025)
von: Koh, Daniel, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Detecting Architectural Decision Violations
von: Su, Ruoyu, et al.
Veröffentlicht: (2026)
von: Su, Ruoyu, et al.
Veröffentlicht: (2026)
CAKE: Cloud Architecture Knowledge Evaluation of Large Language Models
von: Adam, Tim Lukas, et al.
Veröffentlicht: (2026)
von: Adam, Tim Lukas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CoCoEvo: Co-Evolution of Programs and Test Cases to Enhance Code Generation
von: Li, Kefan, et al.
Veröffentlicht: (2025) -
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025) -
Automated User Story Generation with Test Case Specification Using Large Language Model
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024) -
Evaluating Large Language Models for the Generation of Unit Tests with Equivalence Partitions and Boundary Values
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025) -
Task Abstention for Large Language Models in Code Generation
von: Zhou, Yanke, et al.
Veröffentlicht: (2026)