LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Kaibo, Chen, Zhenpeng, Liu, Yiyang, Zhang, Jie M., Harman, Mark, Han, Yudong, Ma, Yun, Dong, Yihong, Li, Ge, Huang, Gang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A First Look at Bugs in LLM Inference Engines
by: Liu, Mugeng, et al.
Published: (2025)
by: Liu, Mugeng, et al.
Published: (2025)
Fairness Testing: A Comprehensive Survey and Analysis of Trends
by: Chen, Zhenpeng, et al.
Published: (2022)
by: Chen, Zhenpeng, et al.
Published: (2022)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025)
by: Hanna, Carol, et al.
Published: (2025)
HITS: High-coverage LLM-based Unit Test Generation via Method Slicing
by: Wang, Zejun, et al.
Published: (2024)
by: Wang, Zejun, et al.
Published: (2024)
YATE: The Role of Test Repair in LLM-Based Unit Test Generation
by: Konstantinou, Michael, et al.
Published: (2025)
by: Konstantinou, Michael, et al.
Published: (2025)
GUIPilot: A Consistency-based Mobile GUI Testing Approach for Detecting Application-specific Bugs
by: Liu, Ruofan, et al.
Published: (2025)
by: Liu, Ruofan, et al.
Published: (2025)
EET: Experience-Driven Early Termination for Cost-Efficient Software Engineering Agents
by: Guo, Yaoqi, et al.
Published: (2026)
by: Guo, Yaoqi, et al.
Published: (2026)
Python Symbolic Execution with LLM-powered Code Generation
by: Wang, Wenhan, et al.
Published: (2024)
by: Wang, Wenhan, et al.
Published: (2024)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
by: Li, Dawei, et al.
Published: (2026)
by: Li, Dawei, et al.
Published: (2026)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
Measuring the Influence of Incorrect Code on Test Generation
by: Huang, Dong, et al.
Published: (2024)
by: Huang, Dong, et al.
Published: (2024)
Fairness Improvement with Multiple Protected Attributes: How Far Are We?
by: Chen, Zhenpeng, et al.
Published: (2023)
by: Chen, Zhenpeng, et al.
Published: (2023)
Can Large Language Models Solve Path Constraints in Symbolic Execution?
by: Wang, Wenhan, et al.
Published: (2025)
by: Wang, Wenhan, et al.
Published: (2025)
Harden and Catch for Just-in-Time Assured LLM-Based Software Testing: Open Research Challenges
by: Harman, Mark, et al.
Published: (2025)
by: Harman, Mark, et al.
Published: (2025)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
LLM-Powered Silent Bug Fuzzing in Deep Learning Libraries via Versatile and Controlled Bug Transfer
by: Zhang, Kunpeng, et al.
Published: (2026)
by: Zhang, Kunpeng, et al.
Published: (2026)
ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
by: Jiang, Xue, et al.
Published: (2024)
by: Jiang, Xue, et al.
Published: (2024)
LLM-Based Misconfiguration Detection for AWS Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2024)
by: Wen, Jinfeng, et al.
Published: (2024)
Testing Refactoring Engine via Historical Bug Report driven LLM
by: Wang, Haibo, et al.
Published: (2025)
by: Wang, Haibo, et al.
Published: (2025)
Personality-Guided Code Generation Using Large Language Models
by: Guo, Yaoqi, et al.
Published: (2024)
by: Guo, Yaoqi, et al.
Published: (2024)
Duplicate Bug Report Detection: How Far Are We?
by: Zhang, Ting, et al.
Published: (2022)
by: Zhang, Ting, et al.
Published: (2022)
Generalizing Test Cases for Comprehensive Test Scenario Coverage
by: Qi, Binhang, et al.
Published: (2026)
by: Qi, Binhang, et al.
Published: (2026)
Enriching Automatic Test Case Generation by Extracting Relevant Test Inputs from Bug Reports
by: Ouédraogo, Wendkûuni C., et al.
Published: (2023)
by: Ouédraogo, Wendkûuni C., et al.
Published: (2023)
Does Programming Language Matter? An Empirical Study of Fuzzing Bug Detection
by: Shirai, Tatsuya, et al.
Published: (2026)
by: Shirai, Tatsuya, et al.
Published: (2026)
AssertFlip: Reproducing Bugs via Inversion of LLM-Generated Passing Tests
by: Khatib, Lara, et al.
Published: (2025)
by: Khatib, Lara, et al.
Published: (2025)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
TestExplora: Benchmarking LLMs for Proactive Bug Discovery via Repository-Level Test Generation
by: Liu, Steven, et al.
Published: (2026)
by: Liu, Steven, et al.
Published: (2026)
Benchmarking LLMs for Unit Test Generation from Real-World Functions
by: Huang, Dong, et al.
Published: (2025)
by: Huang, Dong, et al.
Published: (2025)
An Exploratory Study on Just-in-Time Multi-Programming-Language Bug Prediction
by: Li, Zengyang, et al.
Published: (2024)
by: Li, Zengyang, et al.
Published: (2024)
An Empirical Study of Bugs in Modern LLM Agent Frameworks
by: Zhu, Xinxue, et al.
Published: (2026)
by: Zhu, Xinxue, et al.
Published: (2026)
Feature Slice Matching for Precise Bug Detection
by: Ma, Ke, et al.
Published: (2025)
by: Ma, Ke, et al.
Published: (2025)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
by: Jiang, Weipeng, et al.
Published: (2025)
by: Jiang, Weipeng, et al.
Published: (2025)
Bug-locating Method based on Statistical Testing for Quantum Programs
by: Sato, Naoto, et al.
Published: (2024)
by: Sato, Naoto, et al.
Published: (2024)
Dynamic Cogeneration of Bug Reproduction Test in Agentic Program Repair
by: Cheng, Runxiang, et al.
Published: (2026)
by: Cheng, Runxiang, et al.
Published: (2026)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
by: Xue, Pengyu, et al.
Published: (2024)
by: Xue, Pengyu, et al.
Published: (2024)
iCoRe: An Iterative Correlation-Aware Retriever for Bug Reproduction Test Generation
by: Wang, Junyi, et al.
Published: (2026)
by: Wang, Junyi, et al.
Published: (2026)
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
by: Ge, Yifei, et al.
Published: (2026)
by: Ge, Yifei, et al.
Published: (2026)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
by: Ouyang, Shuyin, et al.
Published: (2023)
by: Ouyang, Shuyin, et al.
Published: (2023)
PanicFI: An Infrastructure for Fixing Panic Bugs in Real-World Rust Programs
by: Ni, Yunbo, et al.
Published: (2024)
by: Ni, Yunbo, et al.
Published: (2024)
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
Similar Items
-
A First Look at Bugs in LLM Inference Engines
by: Liu, Mugeng, et al.
Published: (2025) -
Fairness Testing: A Comprehensive Survey and Analysis of Trends
by: Chen, Zhenpeng, et al.
Published: (2022) -
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025) -
HITS: High-coverage LLM-based Unit Test Generation via Method Slicing
by: Wang, Zejun, et al.
Published: (2024) -
YATE: The Role of Test Repair in LLM-Based Unit Test Generation
by: Konstantinou, Michael, et al.
Published: (2025)