Using Large Language Models to Generate JUnit Tests: An Empirical Study
Fuente:
arXiv
Saved in:
| Main Authors: | Siddiq, Mohammed Latif, Santos, Joanna C. S., Tanvir, Ridwanul Hasan, Ulfat, Noshin, Rifat, Fahmid Al, Lopes, Vinicius Carvalho |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Security in the Age of AI Teammates: An Empirical Study of Agentic Pull Requests on GitHub
by: Siddiq, Mohammed Latif, et al.
Published: (2026)
by: Siddiq, Mohammed Latif, et al.
Published: (2026)
Large Language Models for Software Engineering: A Reproducibility Crisis
by: Siddiq, Mohammed Latif, et al.
Published: (2025)
by: Siddiq, Mohammed Latif, et al.
Published: (2025)
An Empirical Study on Remote Code Execution in Machine Learning Model Hosting Ecosystems
by: Siddiq, Mohammed Latif, et al.
Published: (2026)
by: Siddiq, Mohammed Latif, et al.
Published: (2026)
FRANC: A Lightweight Framework for High-Quality Code Generation
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks
by: Siddiq, Mohammed Latif, et al.
Published: (2024)
by: Siddiq, Mohammed Latif, et al.
Published: (2024)
SALLM: Security Assessment of Generated Code
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
Measuring LLM Trust Allocation Across Conflicting Software Artifacts
by: Ulfat, Noshin, et al.
Published: (2026)
by: Ulfat, Noshin, et al.
Published: (2026)
Assessing the Software Security Comprehension of Large Language Models
by: Siddiq, Mohammed Latif, et al.
Published: (2025)
by: Siddiq, Mohammed Latif, et al.
Published: (2025)
Hallucination to Consensus: Multi-Agent LLMs for End-to-End JUnit Test Generation
by: Xu, Qinghua, et al.
Published: (2025)
by: Xu, Qinghua, et al.
Published: (2025)
CUBETESTERAI: Automated JUnit Test Generation using the LLaMA Model
by: Gorla, Daniele, et al.
Published: (2025)
by: Gorla, Daniele, et al.
Published: (2025)
Large Language Models for Automated Web-Form-Test Generation: An Empirical Study
by: Li, Tao, et al.
Published: (2024)
by: Li, Tao, et al.
Published: (2024)
Crash Report Enhancement with Large Language Models: An Empirical Study
by: Fahim, S M Farah Al, et al.
Published: (2025)
by: Fahim, S M Farah Al, et al.
Published: (2025)
A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing
by: Shang, Ye, et al.
Published: (2024)
by: Shang, Ye, et al.
Published: (2024)
An Empirical Study of Safetensors' Usage Trends and Developers' Perceptions
by: Casey, Beatrice, et al.
Published: (2025)
by: Casey, Beatrice, et al.
Published: (2025)
An Empirical Study of Python Library Migration Using Large Language Models
by: Islam, Md Mohayeminul, et al.
Published: (2025)
by: Islam, Md Mohayeminul, et al.
Published: (2025)
Testing the Untestable? An Empirical Study on the Testing Process of LLM-Powered Software Systems
by: Magalhaes, Cleyton, et al.
Published: (2025)
by: Magalhaes, Cleyton, et al.
Published: (2025)
An Empirical Study of Sustainability in Prompt-driven Test Script Generation Using Small Language Models
by: Kumari, Pragati, et al.
Published: (2026)
by: Kumari, Pragati, et al.
Published: (2026)
An Empirical Study of Testing Practices in Open Source AI Agent Frameworks and Agentic Applications
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
An Empirical Evaluation of Pre-trained Large Language Models for Repairing Declarative Formal Specifications
by: Alhanahnah, Mohannad, et al.
Published: (2024)
by: Alhanahnah, Mohannad, et al.
Published: (2024)
Software Testing with Large Language Models: An Interview Study with Practitioners
by: Santana, Maria Deolinda, et al.
Published: (2025)
by: Santana, Maria Deolinda, et al.
Published: (2025)
Large Language Models for Fault Localization: An Empirical Study
by: Xiao, YingJian, et al.
Published: (2025)
by: Xiao, YingJian, et al.
Published: (2025)
Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks
by: Hasan, Md Mahade, et al.
Published: (2025)
by: Hasan, Md Mahade, et al.
Published: (2025)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
by: Berndt, Alexander, et al.
Published: (2026)
by: Berndt, Alexander, et al.
Published: (2026)
An Empirical Study on the Code Refactoring Capability of Large Language Models
by: Cordeiro, Jonathan, et al.
Published: (2024)
by: Cordeiro, Jonathan, et al.
Published: (2024)
Secret Breach Detection in Source Code with Large Language Models
by: Rahman, Md Nafiu, et al.
Published: (2025)
by: Rahman, Md Nafiu, et al.
Published: (2025)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
Guidelines for Empirical Studies in Software Engineering involving Large Language Models
by: Baltes, Sebastian, et al.
Published: (2025)
by: Baltes, Sebastian, et al.
Published: (2025)
Automatic High-Level Test Case Generation using Large Language Models
by: Hasan, Navid Bin, et al.
Published: (2025)
by: Hasan, Navid Bin, et al.
Published: (2025)
Automated Commit Message Generation with Large Language Models: An Empirical Study and Beyond
by: Xue, Pengyu, et al.
Published: (2024)
by: Xue, Pengyu, et al.
Published: (2024)
Large Language Models for Mobile GUI Text Input Generation: An Empirical Study
by: Cui, Chenhui, et al.
Published: (2024)
by: Cui, Chenhui, et al.
Published: (2024)
Testing with AI Agents: An Empirical Study of Test Generation Frequency, Quality, and Coverage
by: Yoshimoto, Suzuka, et al.
Published: (2026)
by: Yoshimoto, Suzuka, et al.
Published: (2026)
What Makes Software Bugs Escape Testing? Evidence from a Large-Scale Empirical Study
by: Cotroneo, Domenico, et al.
Published: (2026)
by: Cotroneo, Domenico, et al.
Published: (2026)
LLPut: Investigating Large Language Models for Bug Report-Based Input Generation
by: Hasan, Alif Al, et al.
Published: (2025)
by: Hasan, Alif Al, et al.
Published: (2025)
Applications and Implications of Large Language Models in Qualitative Analysis: A New Frontier for Empirical Software Engineering
by: Leça, Matheus de Morais, et al.
Published: (2024)
by: Leça, Matheus de Morais, et al.
Published: (2024)
An Empirical Study on the Amount of Changes Required for Merge Request Acceptance
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
Understanding API Usage and Testing: An Empirical Study of C Libraries
by: Zaki, Ahmed, et al.
Published: (2025)
by: Zaki, Ahmed, et al.
Published: (2025)
Understanding Bug-Reproducing Tests: A First Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
Deciphering Refactoring Branch Dynamics in Modern Code Review: An Empirical Study on Qt
by: AlOmar, Eman Abdullah
Published: (2024)
by: AlOmar, Eman Abdullah
Published: (2024)
An Empirical Study on the Impact of Code Duplication-aware Refactoring Practices on Quality Metrics
by: AlOmar, Eman Abdullah
Published: (2025)
by: AlOmar, Eman Abdullah
Published: (2025)
Similar Items
-
Security in the Age of AI Teammates: An Empirical Study of Agentic Pull Requests on GitHub
by: Siddiq, Mohammed Latif, et al.
Published: (2026) -
Large Language Models for Software Engineering: A Reproducibility Crisis
by: Siddiq, Mohammed Latif, et al.
Published: (2025) -
An Empirical Study on Remote Code Execution in Machine Learning Model Hosting Ecosystems
by: Siddiq, Mohammed Latif, et al.
Published: (2026) -
FRANC: A Lightweight Framework for High-Quality Code Generation
by: Siddiq, Mohammed Latif, et al.
Published: (2023) -
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks
by: Siddiq, Mohammed Latif, et al.
Published: (2024)