LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Kaibo, Chen, Zhenpeng, Liu, Yiyang, Zhang, Jie M., Harman, Mark, Han, Yudong, Ma, Yun, Dong, Yihong, Li, Ge, Huang, Gang |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A First Look at Bugs in LLM Inference Engines
par: Liu, Mugeng, et autres
Publié: (2025)
par: Liu, Mugeng, et autres
Publié: (2025)
Fairness Testing: A Comprehensive Survey and Analysis of Trends
par: Chen, Zhenpeng, et autres
Publié: (2022)
par: Chen, Zhenpeng, et autres
Publié: (2022)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
par: Hanna, Carol, et autres
Publié: (2025)
par: Hanna, Carol, et autres
Publié: (2025)
HITS: High-coverage LLM-based Unit Test Generation via Method Slicing
par: Wang, Zejun, et autres
Publié: (2024)
par: Wang, Zejun, et autres
Publié: (2024)
YATE: The Role of Test Repair in LLM-Based Unit Test Generation
par: Konstantinou, Michael, et autres
Publié: (2025)
par: Konstantinou, Michael, et autres
Publié: (2025)
GUIPilot: A Consistency-based Mobile GUI Testing Approach for Detecting Application-specific Bugs
par: Liu, Ruofan, et autres
Publié: (2025)
par: Liu, Ruofan, et autres
Publié: (2025)
EET: Experience-Driven Early Termination for Cost-Efficient Software Engineering Agents
par: Guo, Yaoqi, et autres
Publié: (2026)
par: Guo, Yaoqi, et autres
Publié: (2026)
Python Symbolic Execution with LLM-powered Code Generation
par: Wang, Wenhan, et autres
Publié: (2024)
par: Wang, Wenhan, et autres
Publié: (2024)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
par: Li, Dawei, et autres
Publié: (2026)
par: Li, Dawei, et autres
Publié: (2026)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
par: Rathnasuriya, Ravishka, et autres
Publié: (2026)
par: Rathnasuriya, Ravishka, et autres
Publié: (2026)
Measuring the Influence of Incorrect Code on Test Generation
par: Huang, Dong, et autres
Publié: (2024)
par: Huang, Dong, et autres
Publié: (2024)
Fairness Improvement with Multiple Protected Attributes: How Far Are We?
par: Chen, Zhenpeng, et autres
Publié: (2023)
par: Chen, Zhenpeng, et autres
Publié: (2023)
Can Large Language Models Solve Path Constraints in Symbolic Execution?
par: Wang, Wenhan, et autres
Publié: (2025)
par: Wang, Wenhan, et autres
Publié: (2025)
Harden and Catch for Just-in-Time Assured LLM-Based Software Testing: Open Research Challenges
par: Harman, Mark, et autres
Publié: (2025)
par: Harman, Mark, et autres
Publié: (2025)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
par: Widyasari, Ratnadira, et autres
Publié: (2024)
par: Widyasari, Ratnadira, et autres
Publié: (2024)
LLM-Powered Silent Bug Fuzzing in Deep Learning Libraries via Versatile and Controlled Bug Transfer
par: Zhang, Kunpeng, et autres
Publié: (2026)
par: Zhang, Kunpeng, et autres
Publié: (2026)
ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
par: Jiang, Xue, et autres
Publié: (2024)
par: Jiang, Xue, et autres
Publié: (2024)
LLM-Based Misconfiguration Detection for AWS Serverless Computing
par: Wen, Jinfeng, et autres
Publié: (2024)
par: Wen, Jinfeng, et autres
Publié: (2024)
Testing Refactoring Engine via Historical Bug Report driven LLM
par: Wang, Haibo, et autres
Publié: (2025)
par: Wang, Haibo, et autres
Publié: (2025)
Personality-Guided Code Generation Using Large Language Models
par: Guo, Yaoqi, et autres
Publié: (2024)
par: Guo, Yaoqi, et autres
Publié: (2024)
Duplicate Bug Report Detection: How Far Are We?
par: Zhang, Ting, et autres
Publié: (2022)
par: Zhang, Ting, et autres
Publié: (2022)
Generalizing Test Cases for Comprehensive Test Scenario Coverage
par: Qi, Binhang, et autres
Publié: (2026)
par: Qi, Binhang, et autres
Publié: (2026)
Enriching Automatic Test Case Generation by Extracting Relevant Test Inputs from Bug Reports
par: Ouédraogo, Wendkûuni C., et autres
Publié: (2023)
par: Ouédraogo, Wendkûuni C., et autres
Publié: (2023)
Does Programming Language Matter? An Empirical Study of Fuzzing Bug Detection
par: Shirai, Tatsuya, et autres
Publié: (2026)
par: Shirai, Tatsuya, et autres
Publié: (2026)
AssertFlip: Reproducing Bugs via Inversion of LLM-Generated Passing Tests
par: Khatib, Lara, et autres
Publié: (2025)
par: Khatib, Lara, et autres
Publié: (2025)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
par: Zheng, Mingwei, et autres
Publié: (2025)
par: Zheng, Mingwei, et autres
Publié: (2025)
TestExplora: Benchmarking LLMs for Proactive Bug Discovery via Repository-Level Test Generation
par: Liu, Steven, et autres
Publié: (2026)
par: Liu, Steven, et autres
Publié: (2026)
Benchmarking LLMs for Unit Test Generation from Real-World Functions
par: Huang, Dong, et autres
Publié: (2025)
par: Huang, Dong, et autres
Publié: (2025)
An Exploratory Study on Just-in-Time Multi-Programming-Language Bug Prediction
par: Li, Zengyang, et autres
Publié: (2024)
par: Li, Zengyang, et autres
Publié: (2024)
An Empirical Study of Bugs in Modern LLM Agent Frameworks
par: Zhu, Xinxue, et autres
Publié: (2026)
par: Zhu, Xinxue, et autres
Publié: (2026)
Feature Slice Matching for Precise Bug Detection
par: Ma, Ke, et autres
Publié: (2025)
par: Ma, Ke, et autres
Publié: (2025)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
par: Jiang, Weipeng, et autres
Publié: (2025)
par: Jiang, Weipeng, et autres
Publié: (2025)
Bug-locating Method based on Statistical Testing for Quantum Programs
par: Sato, Naoto, et autres
Publié: (2024)
par: Sato, Naoto, et autres
Publié: (2024)
Dynamic Cogeneration of Bug Reproduction Test in Agentic Program Repair
par: Cheng, Runxiang, et autres
Publié: (2026)
par: Cheng, Runxiang, et autres
Publié: (2026)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
par: Xue, Pengyu, et autres
Publié: (2024)
par: Xue, Pengyu, et autres
Publié: (2024)
iCoRe: An Iterative Correlation-Aware Retriever for Bug Reproduction Test Generation
par: Wang, Junyi, et autres
Publié: (2026)
par: Wang, Junyi, et autres
Publié: (2026)
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
par: Ge, Yifei, et autres
Publié: (2026)
par: Ge, Yifei, et autres
Publié: (2026)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
par: Ouyang, Shuyin, et autres
Publié: (2023)
par: Ouyang, Shuyin, et autres
Publié: (2023)
PanicFI: An Infrastructure for Fixing Panic Bugs in Real-World Rust Programs
par: Ni, Yunbo, et autres
Publié: (2024)
par: Ni, Yunbo, et autres
Publié: (2024)
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
par: Li, Xinyue, et autres
Publié: (2026)
par: Li, Xinyue, et autres
Publié: (2026)
Documents similaires
-
A First Look at Bugs in LLM Inference Engines
par: Liu, Mugeng, et autres
Publié: (2025) -
Fairness Testing: A Comprehensive Survey and Analysis of Trends
par: Chen, Zhenpeng, et autres
Publié: (2022) -
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
par: Hanna, Carol, et autres
Publié: (2025) -
HITS: High-coverage LLM-based Unit Test Generation via Method Slicing
par: Wang, Zejun, et autres
Publié: (2024) -
YATE: The Role of Test Repair in LLM-Based Unit Test Generation
par: Konstantinou, Michael, et autres
Publié: (2025)