Test Case Generation from Bug Reports via Large Language Models: A Cognitive Layered Evaluation Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qureshi, Irtaza Sajid, Ming, Zhen, Jiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enriching Automatic Test Case Generation by Extracting Relevant Test Inputs from Bug Reports
von: Ouédraogo, Wendkûuni C., et al.
Veröffentlicht: (2023)
von: Ouédraogo, Wendkûuni C., et al.
Veröffentlicht: (2023)
LLPut: Investigating Large Language Models for Bug Report-Based Input Generation
von: Hasan, Alif Al, et al.
Veröffentlicht: (2025)
von: Hasan, Alif Al, et al.
Veröffentlicht: (2025)
Isolating Compiler Bugs by Generating Effective Witness Programs with Large Language Models
von: Tu, Haoxin, et al.
Veröffentlicht: (2023)
von: Tu, Haoxin, et al.
Veröffentlicht: (2023)
TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
von: Zhang, Quanjun, et al.
Veröffentlicht: (2024)
von: Zhang, Quanjun, et al.
Veröffentlicht: (2024)
Bridging the Interpretation Gap in Accessibility Testing: Empathetic and Legal-Aware Bug Report Generation via Large Language Models
von: Koyama, Ryoya, et al.
Veröffentlicht: (2026)
von: Koyama, Ryoya, et al.
Veröffentlicht: (2026)
Improving Compiler Bug Isolation by Leveraging Large Language Models
von: Qi, Yixian, et al.
Veröffentlicht: (2025)
von: Qi, Yixian, et al.
Veröffentlicht: (2025)
Testing Refactoring Engine via Historical Bug Report driven LLM
von: Wang, Haibo, et al.
Veröffentlicht: (2025)
von: Wang, Haibo, et al.
Veröffentlicht: (2025)
Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
English Please: Evaluating Machine Translation with Large Language Models for Multilingual Bug Reports
von: Patil, Avinash, et al.
Veröffentlicht: (2025)
von: Patil, Avinash, et al.
Veröffentlicht: (2025)
On the Evaluation of Large Language Models in Unit Test Generation
von: Yang, Lin, et al.
Veröffentlicht: (2024)
von: Yang, Lin, et al.
Veröffentlicht: (2024)
A Deep Dive into Large Language Models for Automated Bug Localization and Repair
von: Hossain, Soneya Binta, et al.
Veröffentlicht: (2024)
von: Hossain, Soneya Binta, et al.
Veröffentlicht: (2024)
TESTEVAL: Benchmarking Large Language Models for Test Case Generation
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
Large Language Models as Test Case Generators: Performance Evaluation and Enhancement
von: Li, Kefan, et al.
Veröffentlicht: (2024)
von: Li, Kefan, et al.
Veröffentlicht: (2024)
Bridging Bug Localization and Issue Fixing: A Hierarchical Localization Framework Leveraging Large Language Models
von: Chang, Jianming, et al.
Veröffentlicht: (2025)
von: Chang, Jianming, et al.
Veröffentlicht: (2025)
Comparative Evaluation of Large Language Models for Test-Skeleton Generation
von: Boorlagadda, Subhang, et al.
Veröffentlicht: (2025)
von: Boorlagadda, Subhang, et al.
Veröffentlicht: (2025)
A Tool for Test Case Scenarios Generation Using Large Language Models
von: Sami, Abdul Malik, et al.
Veröffentlicht: (2024)
von: Sami, Abdul Malik, et al.
Veröffentlicht: (2024)
LLMs as Evaluators: A Novel Approach to Evaluate Bug Report Summarization
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
Acceptance Test Generation with Large Language Models: An Industrial Case Study
von: Ferreira, Margarida, et al.
Veröffentlicht: (2025)
von: Ferreira, Margarida, et al.
Veröffentlicht: (2025)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
von: Li, Dawei, et al.
Veröffentlicht: (2026)
von: Li, Dawei, et al.
Veröffentlicht: (2026)
Bug Whispering: Towards Audio Bug Reporting
von: Masserini, Elena, et al.
Veröffentlicht: (2025)
von: Masserini, Elena, et al.
Veröffentlicht: (2025)
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
von: Gheyi, Rohit, et al.
Veröffentlicht: (2025)
von: Gheyi, Rohit, et al.
Veröffentlicht: (2025)
AssertFlip: Reproducing Bugs via Inversion of LLM-Generated Passing Tests
von: Khatib, Lara, et al.
Veröffentlicht: (2025)
von: Khatib, Lara, et al.
Veröffentlicht: (2025)
Uncovering Business Logic Bugs via Semantics-Driven Unit Test Generation
von: Yang, Chen, et al.
Veröffentlicht: (2026)
von: Yang, Chen, et al.
Veröffentlicht: (2026)
HAFix: History-Augmented Large Language Models for Bug Fixing
von: Shi, Yu, et al.
Veröffentlicht: (2025)
von: Shi, Yu, et al.
Veröffentlicht: (2025)
Bugs in Large Language Models Generated Code: An Empirical Study
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
GitBugs: Bug Reports for Duplicate Detection, Retrieval Augmented Generation, Triage, and More
von: Patil, Avinash, et al.
Veröffentlicht: (2025)
von: Patil, Avinash, et al.
Veröffentlicht: (2025)
Automatic High-Level Test Case Generation using Large Language Models
von: Hasan, Navid Bin, et al.
Veröffentlicht: (2025)
von: Hasan, Navid Bin, et al.
Veröffentlicht: (2025)
Automated Control Logic Test Case Generation using Large Language Models
von: Koziolek, Heiko, et al.
Veröffentlicht: (2024)
von: Koziolek, Heiko, et al.
Veröffentlicht: (2024)
LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
von: Liu, Kaibo, et al.
Veröffentlicht: (2024)
von: Liu, Kaibo, et al.
Veröffentlicht: (2024)
Analyzing the Instability of Large Language Models in Automated Bug Injection and Correction
von: Er, Mehmet Bilal, et al.
Veröffentlicht: (2025)
von: Er, Mehmet Bilal, et al.
Veröffentlicht: (2025)
Testing Framework Migration with Large Language Models
von: Alves, Altino, et al.
Veröffentlicht: (2026)
von: Alves, Altino, et al.
Veröffentlicht: (2026)
SysPro: Reproducing System-level Concurrency Bugs from Bug Reports
von: Zaman, Tarannum Shaila, et al.
Veröffentlicht: (2026)
von: Zaman, Tarannum Shaila, et al.
Veröffentlicht: (2026)
Identifying Concurrency Bug Reports via Linguistic Patterns
von: Shao, Shuai, et al.
Veröffentlicht: (2026)
von: Shao, Shuai, et al.
Veröffentlicht: (2026)
Toward Understanding Deep Learning Framework Bugs
von: Chen, Junjie, et al.
Veröffentlicht: (2022)
von: Chen, Junjie, et al.
Veröffentlicht: (2022)
Automated Duplicate Bug Report Detection in Large Open Bug Repositories
von: Laney, Clare E., et al.
Veröffentlicht: (2025)
von: Laney, Clare E., et al.
Veröffentlicht: (2025)
BugMentor: Generating Answers to Follow-up Questions from Software Bug Reports using Structured Information Retrieval and Neural Text Generation
von: Mukherjee, Usmi, et al.
Veröffentlicht: (2023)
von: Mukherjee, Usmi, et al.
Veröffentlicht: (2023)
Evaluating Large Language Models in Detecting Test Smells
von: Lucas, Keila, et al.
Veröffentlicht: (2024)
von: Lucas, Keila, et al.
Veröffentlicht: (2024)
Automated Generation of High-Quality Bug Reports for Android Applications
von: Saha, Antu, et al.
Veröffentlicht: (2026)
von: Saha, Antu, et al.
Veröffentlicht: (2026)
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
ViBR: Automated Bug Replay from Video-based Reports using Vision-Language Models
von: Feng, Sidong, et al.
Veröffentlicht: (2026)
von: Feng, Sidong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enriching Automatic Test Case Generation by Extracting Relevant Test Inputs from Bug Reports
von: Ouédraogo, Wendkûuni C., et al.
Veröffentlicht: (2023) -
LLPut: Investigating Large Language Models for Bug Report-Based Input Generation
von: Hasan, Alif Al, et al.
Veröffentlicht: (2025) -
Isolating Compiler Bugs by Generating Effective Witness Programs with Large Language Models
von: Tu, Haoxin, et al.
Veröffentlicht: (2023) -
TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
von: Zhang, Quanjun, et al.
Veröffentlicht: (2024) -
Bridging the Interpretation Gap in Accessibility Testing: Empathetic and Legal-Aware Bug Report Generation via Large Language Models
von: Koyama, Ryoya, et al.
Veröffentlicht: (2026)