Hallucination to Consensus: Multi-Agent LLMs for End-to-End JUnit Test Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Qinghua, Wang, Guancheng, Briand, Lionel, Liu, Kui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mutation-Guided Unit Test Generation with a Large Language Model
by: Wang, Guancheng, et al.
Published: (2025)
by: Wang, Guancheng, et al.
Published: (2025)
Call-Chain-Aware LLM-Based Test Generation for Java Projects
by: Wang, Guancheng, et al.
Published: (2026)
by: Wang, Guancheng, et al.
Published: (2026)
Uncertainty-Guided Label Rebalancing for CPS Safety Monitoring
by: Ayotunde, John, et al.
Published: (2026)
by: Ayotunde, John, et al.
Published: (2026)
Characterizing the Failure Modes of LLMs in Resolving Real-World GitHub Issues
by: Jiang, Yanjie, et al.
Published: (2026)
by: Jiang, Yanjie, et al.
Published: (2026)
LLM meets ML: Data-efficient Anomaly Detection on Unstable Logs
by: Hadadi, Fatemeh, et al.
Published: (2024)
by: Hadadi, Fatemeh, et al.
Published: (2024)
Feature-Driven End-To-End Test Generation
by: Alian, Parsa, et al.
Published: (2024)
by: Alian, Parsa, et al.
Published: (2024)
CUBETESTERAI: Automated JUnit Test Generation using the LLaMA Model
by: Gorla, Daniele, et al.
Published: (2025)
by: Gorla, Daniele, et al.
Published: (2025)
Using Large Language Models to Generate JUnit Tests: An Empirical Study
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
Fuzzing-based Mutation Testing of C/C++ Software in Cyber-Physical Systems
by: Lee, Jaekwon, et al.
Published: (2025)
by: Lee, Jaekwon, et al.
Published: (2025)
Drivora: A Unified and Extensible Infrastructure for Search-based Autonomous Driving Testing
by: Cheng, Mingfei, et al.
Published: (2026)
by: Cheng, Mingfei, et al.
Published: (2026)
Testing Updated Apps by Adapting Learned Models
by: Ngo, Chanh-Duc, et al.
Published: (2023)
by: Ngo, Chanh-Duc, et al.
Published: (2023)
End-to-End Automated Logging via Multi-Agent Framework
by: Zhong, Renyi, et al.
Published: (2025)
by: Zhong, Renyi, et al.
Published: (2025)
WEFix: Intelligent Automatic Generation of Explicit Waits for Efficient Web End-to-End Flaky Tests
by: Liu, Xinyue, et al.
Published: (2024)
by: Liu, Xinyue, et al.
Published: (2024)
Search-based DNN Testing and Retraining with GAN-enhanced Simulations
by: Attaoui, Mohammed Oualid, et al.
Published: (2024)
by: Attaoui, Mohammed Oualid, et al.
Published: (2024)
Multi-Agent End-to-End Vulnerability Management for Mitigating Recurring Vulnerabilities
by: Zheng, Zelong, et al.
Published: (2026)
by: Zheng, Zelong, et al.
Published: (2026)
MOTIF: A tool for Mutation Testing with Fuzzing
by: Lee, Jaekwon, et al.
Published: (2024)
by: Lee, Jaekwon, et al.
Published: (2024)
LTM: Scalable and Black-box Similarity-based Test Suite Minimization based on Language Models
by: Pan, Rongqi, et al.
Published: (2023)
by: Pan, Rongqi, et al.
Published: (2023)
Learning-Guided Fuzzing for Testing Stateful SDN Controllers
by: Ollando, Raphaël, et al.
Published: (2024)
by: Ollando, Raphaël, et al.
Published: (2024)
TEASMA: A Practical Methodology for Test Adequacy Assessment of Deep Neural Networks
by: Abbasishahkoo, Amin, et al.
Published: (2023)
by: Abbasishahkoo, Amin, et al.
Published: (2023)
Supporting System Testing with a Multi-Agent LLM-based Framework for Knowledge Graph Extraction: A Case Study with Ethernet Switch Systems
by: Pan, Rongqi, et al.
Published: (2026)
by: Pan, Rongqi, et al.
Published: (2026)
Automated Test Case Repair Using Language Models
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2024)
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2024)
Requirements Coverage-Guided Minimization for Natural Language Test Cases
by: Pan, Rongqi, et al.
Published: (2025)
by: Pan, Rongqi, et al.
Published: (2025)
Scalable and Accurate Test Case Prioritization in Continuous Integration Contexts
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2021)
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2021)
VISCA: Inferring Component Abstractions for Automated End-to-End Testing
by: Alian, Parsa, et al.
Published: (2025)
by: Alian, Parsa, et al.
Published: (2025)
Constrained Co-evolutionary Metamorphic Differential Testing for Autonomous Systems with an Interpretability Approach
by: Yousefizadeh, Hossein, et al.
Published: (2025)
by: Yousefizadeh, Hossein, et al.
Published: (2025)
RepoGenesis: Benchmarking End-to-End Microservice Generation from Readme to Repository
by: Peng, Zhiyuan, et al.
Published: (2026)
by: Peng, Zhiyuan, et al.
Published: (2026)
DeepGD: A Multi-Objective Black-Box Test Selection Approach for Deep Neural Networks
by: Aghababaeyan, Zohreh, et al.
Published: (2023)
by: Aghababaeyan, Zohreh, et al.
Published: (2023)
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair
by: Fatima, Sakina, et al.
Published: (2023)
by: Fatima, Sakina, et al.
Published: (2023)
GenIA-E2ETest: A Generative AI-Based Approach for End-to-End Test Automation
by: Júnior, Elvis, et al.
Published: (2025)
by: Júnior, Elvis, et al.
Published: (2025)
Stress Testing Control Loops in Cyber-Physical Systems
by: Mandrioli, Claudio, et al.
Published: (2023)
by: Mandrioli, Claudio, et al.
Published: (2023)
Benchmarking and Studying the LLM-based Agent System in End-to-End Software Development
by: Zeng, Zhengran, et al.
Published: (2025)
by: Zeng, Zhengran, et al.
Published: (2025)
MetaSel: A Test Selection Approach for Fine-tuned DNN Models
by: Abbasishahkoo, Amin, et al.
Published: (2025)
by: Abbasishahkoo, Amin, et al.
Published: (2025)
TVR: Automotive System Requirement Traceability Validation and Recovery Through Retrieval-Augmented Generation
by: Niu, Feifei, et al.
Published: (2025)
by: Niu, Feifei, et al.
Published: (2025)
Exploring the Effectiveness of LLMs in Automated Logging Generation: An Empirical Study
by: Li, Yichen, et al.
Published: (2023)
by: Li, Yichen, et al.
Published: (2023)
DiffGAN: A Test Generation Approach for Differential Testing of Deep Neural Networks for Image Analysis
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
by: Aghababaeyan, Zohreh, et al.
Published: (2024)
Similar Pattern Annotation via Retrieval Knowledge for LLM-Based Test Code Fault Localization
by: Gharachorlu, Golnaz, et al.
Published: (2026)
by: Gharachorlu, Golnaz, et al.
Published: (2026)
Using Cooperative Co-evolutionary Search to Generate Metamorphic Test Cases for Autonomous Driving Systems
by: Yousefizadeh, Hossein, et al.
Published: (2024)
by: Yousefizadeh, Hossein, et al.
Published: (2024)
EvoDev: An Iterative Feature-Driven Framework for End-to-End Software Development with LLM-based Agents
by: Liu, Junwei, et al.
Published: (2025)
by: Liu, Junwei, et al.
Published: (2025)
Testing CPS with Design Assumptions-Based Metamorphic Relations and Genetic Programming
by: Mandrioli, Claudio, et al.
Published: (2024)
by: Mandrioli, Claudio, et al.
Published: (2024)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026)
by: Yu, Boxi, et al.
Published: (2026)
Similar Items
-
Mutation-Guided Unit Test Generation with a Large Language Model
by: Wang, Guancheng, et al.
Published: (2025) -
Call-Chain-Aware LLM-Based Test Generation for Java Projects
by: Wang, Guancheng, et al.
Published: (2026) -
Uncertainty-Guided Label Rebalancing for CPS Safety Monitoring
by: Ayotunde, John, et al.
Published: (2026) -
Characterizing the Failure Modes of LLMs in Resolving Real-World GitHub Issues
by: Jiang, Yanjie, et al.
Published: (2026) -
LLM meets ML: Data-efficient Anomaly Detection on Unstable Logs
by: Hadadi, Fatemeh, et al.
Published: (2024)