MultiFileTest: A Multi-File-Level LLM Unit Test Generation Benchmark and Impact of Error Fixing Mechanisms
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yibo, Xia, Congying, Zhao, Wenting, Du, Jiangshu, Miao, Chunyu, Deng, Zhongfen, Yu, Philip S., Xing, Chen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Synthesizing File-Level Data for Unit Test Generation with Chain-of-Thoughts via Self-Debugging
by: Hua, Ziyue, et al.
Published: (2026)
by: Hua, Ziyue, et al.
Published: (2026)
Decomposing God Header File via Multi-View Graph Clustering
by: Wang, Yue, et al.
Published: (2024)
by: Wang, Yue, et al.
Published: (2024)
Automated Unit Test Refactoring
by: Gao, Yi, et al.
Published: (2024)
by: Gao, Yi, et al.
Published: (2024)
Test smells in LLM-Generated Unit Tests
by: Ouédraogo, Wendkûuni C., et al.
Published: (2024)
by: Ouédraogo, Wendkûuni C., et al.
Published: (2024)
Encrypted Container File: Design and Implementation of a Hybrid-Encrypted Multi-Recipient File Structure
by: Bauer, Tobias J., et al.
Published: (2024)
by: Bauer, Tobias J., et al.
Published: (2024)
BLAgent: Agentic RAG for File-Level Bug Localization
by: Mamun, Md Afif Al, et al.
Published: (2026)
by: Mamun, Md Afif Al, et al.
Published: (2026)
Testing Multi-Subroutine Quantum Programs: From Unit Testing to Integration Testing
by: Long, Peixun, et al.
Published: (2023)
by: Long, Peixun, et al.
Published: (2023)
SparseCoder: Identifier-Aware Sparse Transformer for File-Level Code Summarization
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
Sharpen the Spec, Cut the Code: A Case for Generative File System with SYSSPEC
by: Liu, Qingyuan, et al.
Published: (2025)
by: Liu, Qingyuan, et al.
Published: (2025)
Unit Test Update through LLM-Driven Context Collection and Error-Type-Aware Refinement
by: Zhang, Yuanhe, et al.
Published: (2025)
by: Zhang, Yuanhe, et al.
Published: (2025)
TestGenEval: A Real World Unit Test Generation and Test Completion Benchmark
by: Jain, Kush, et al.
Published: (2024)
by: Jain, Kush, et al.
Published: (2024)
YATE: The Role of Test Repair in LLM-Based Unit Test Generation
by: Konstantinou, Michael, et al.
Published: (2025)
by: Konstantinou, Michael, et al.
Published: (2025)
Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study
by: Chand, Sivajeet, et al.
Published: (2026)
by: Chand, Sivajeet, et al.
Published: (2026)
ASTER: Natural and Multi-language Unit Test Generation with LLMs
by: Pan, Rangeet, et al.
Published: (2024)
by: Pan, Rangeet, et al.
Published: (2024)
Attention Mechanism and Heuristic Approach: Context-Aware File Ranking Using Multi-Head Self-Attention
by: Sharma, Pradeep Kumar, et al.
Published: (2026)
by: Sharma, Pradeep Kumar, et al.
Published: (2026)
Multi-Level Testing of Conversational AI Systems
by: Masserini, Elena
Published: (2026)
by: Masserini, Elena
Published: (2026)
Test vs Mutant: Adversarial LLM Agents for Robust Unit Test Generation
by: Chang, Pengyu, et al.
Published: (2026)
by: Chang, Pengyu, et al.
Published: (2026)
A FAIR File Format for Mathematical Software
by: Della Vecchia, Antony, et al.
Published: (2023)
by: Della Vecchia, Antony, et al.
Published: (2023)
Instruction Adherence in Coding Agent Configuration Files: A Factorial Study of Four File-Structure Variables
by: McMillan, Damon
Published: (2026)
by: McMillan, Damon
Published: (2026)
Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?
by: Gloaguen, Thibaud, et al.
Published: (2026)
by: Gloaguen, Thibaud, et al.
Published: (2026)
Benchmarking LLMs for Unit Test Generation from Real-World Functions
by: Huang, Dong, et al.
Published: (2025)
by: Huang, Dong, et al.
Published: (2025)
Does My README File Need To Be Updated? Exploring LLM-Based README Maintenance
by: Gao, Haoyu, et al.
Published: (2026)
by: Gao, Haoyu, et al.
Published: (2026)
Less is More: On the Importance of Data Quality for Unit Test Generation
by: Zhang, Junwei, et al.
Published: (2025)
by: Zhang, Junwei, et al.
Published: (2025)
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
by: Yang, Ruozhao, et al.
Published: (2025)
by: Yang, Ruozhao, et al.
Published: (2025)
The Divine Software Engineering Comedy -- Inferno: The Okinawa Files
by: Lanza, Michele
Published: (2025)
by: Lanza, Michele
Published: (2025)
From Illusion to Insight: Change-Aware File-Level Software Defect Prediction Using Agentic AI
by: Hesamolhokama, Mohsen, et al.
Published: (2025)
by: Hesamolhokama, Mohsen, et al.
Published: (2025)
UTFix: Change Aware Unit Test Repairing using LLM
by: Rahman, Shanto, et al.
Published: (2025)
by: Rahman, Shanto, et al.
Published: (2025)
LLM-based Unit Test Generation via Property Retrieval
by: Zhang, Zhe, et al.
Published: (2024)
by: Zhang, Zhe, et al.
Published: (2024)
Mellum: Production-Grade in-IDE Contextual Code Completion with Multi-File Project Understanding
by: Pavlichenko, Nikita, et al.
Published: (2025)
by: Pavlichenko, Nikita, et al.
Published: (2025)
Fixing Function-Level Code Generation Errors for Foundation Large Language Models
by: Wen, Hao, et al.
Published: (2024)
by: Wen, Hao, et al.
Published: (2024)
Reflective Unit Test Generation for Precise Type Error Detection with Large Language Models
by: Yang, Chen, et al.
Published: (2025)
by: Yang, Chen, et al.
Published: (2025)
Fine-grained Testing for Autonomous Driving Software: a Study on Autoware with LLM-driven Unit Testing
by: Wang, Wenhan, et al.
Published: (2025)
by: Wang, Wenhan, et al.
Published: (2025)
LogiAgent: Automated Logical Testing for REST Systems with LLM-Based Multi-Agents
by: Zhang, Ke, et al.
Published: (2025)
by: Zhang, Ke, et al.
Published: (2025)
Everything is Context: Agentic File System Abstraction for Context Engineering
by: Xu, Xiwei, et al.
Published: (2025)
by: Xu, Xiwei, et al.
Published: (2025)
Agent READMEs: An Empirical Study of Context Files for Agentic Coding
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
An LLM-based Readability Measurement for Unit Tests' Context-aware Inputs
by: Zhou, Zhichao, et al.
Published: (2024)
by: Zhou, Zhichao, et al.
Published: (2024)
TestART: Improving LLM-based Unit Testing via Co-evolution of Automated Generation and Repair Iteration
by: Gu, Siqi, et al.
Published: (2024)
by: Gu, Siqi, et al.
Published: (2024)
Test Wars: A Comparative Study of SBST, Symbolic Execution, and LLM-Based Approaches to Unit Test Generation
by: Abdullin, Azat, et al.
Published: (2025)
by: Abdullin, Azat, et al.
Published: (2025)
LLM-based Content Classification Approach for GitHub Repositories by the README Files
by: Mehmood, Malik Uzair, et al.
Published: (2025)
by: Mehmood, Malik Uzair, et al.
Published: (2025)
AutoMT: A Multi-Agent LLM Framework for Automated Metamorphic Testing of Autonomous Driving Systems
by: Liang, Linfeng, et al.
Published: (2025)
by: Liang, Linfeng, et al.
Published: (2025)
Similar Items
-
Synthesizing File-Level Data for Unit Test Generation with Chain-of-Thoughts via Self-Debugging
by: Hua, Ziyue, et al.
Published: (2026) -
Decomposing God Header File via Multi-View Graph Clustering
by: Wang, Yue, et al.
Published: (2024) -
Automated Unit Test Refactoring
by: Gao, Yi, et al.
Published: (2024) -
Test smells in LLM-Generated Unit Tests
by: Ouédraogo, Wendkûuni C., et al.
Published: (2024) -
Encrypted Container File: Design and Implementation of a Hybrid-Encrypted Multi-Recipient File Structure
by: Bauer, Tobias J., et al.
Published: (2024)