Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
Fuente:
arXiv
Guardado en:
| Autores principales: | Nirujan, Hinduja, Patil, Shreyas, Ayoub, Abdallah, Latif, Ahmad Abdel, Ginde, Gouri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
por: Acharya, Jagrit, et al.
Publicado: (2025)
por: Acharya, Jagrit, et al.
Publicado: (2025)
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
por: Acharya, Jagrit, et al.
Publicado: (2025)
por: Acharya, Jagrit, et al.
Publicado: (2025)
What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook
por: Huo, Junyu, et al.
Publicado: (2026)
por: Huo, Junyu, et al.
Publicado: (2026)
Beyond Keywords: A Context-based Hybrid Approach to Mining Ethical Concern-related App Reviews
por: Sorathiya, Aakash, et al.
Publicado: (2024)
por: Sorathiya, Aakash, et al.
Publicado: (2024)
Towards Energy-aware Requirements Dependency Classification: Knowledge-Graph vs. Vector-Retrieval Augmented Inference with SLMs
por: Patil, Shreyas, et al.
Publicado: (2026)
por: Patil, Shreyas, et al.
Publicado: (2026)
"So what if I used GenAI?" -- Implications of Using Cloud-based GenAI in Software Engineering Research
por: Ginde, Gouri
Publicado: (2024)
por: Ginde, Gouri
Publicado: (2024)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
por: Meng, Xiangxin, et al.
Publicado: (2024)
por: Meng, Xiangxin, et al.
Publicado: (2024)
Exploring Ethical Concerns of Mobile Applications from App Reviews: A Literature Survey
por: Sorathiya, Aakash, et al.
Publicado: (2026)
por: Sorathiya, Aakash, et al.
Publicado: (2026)
Towards Extracting Ethical Concerns-related Software Requirements from App Reviews
por: Sorathiya, Aakash, et al.
Publicado: (2024)
por: Sorathiya, Aakash, et al.
Publicado: (2024)
CMER: A Context-Aware Approach for Mining Ethical Concern-related App Reviews
por: Sorathiya, Aakash, et al.
Publicado: (2025)
por: Sorathiya, Aakash, et al.
Publicado: (2025)
SAGE: A Context-Aware Approach for Mining Privacy Requirements Relevant Reviews from Mental Health Apps
por: Sorathiya, Aakash, et al.
Publicado: (2025)
por: Sorathiya, Aakash, et al.
Publicado: (2025)
Towards Extracting Software Requirements from App Reviews using Seq2seq Framework
por: Sorathiya, Aakash, et al.
Publicado: (2025)
por: Sorathiya, Aakash, et al.
Publicado: (2025)
Automated Duplicate Bug Report Detection in Large Open Bug Repositories
por: Laney, Clare E., et al.
Publicado: (2025)
por: Laney, Clare E., et al.
Publicado: (2025)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
por: Khati, Dipin, et al.
Publicado: (2026)
por: Khati, Dipin, et al.
Publicado: (2026)
Bugs in Large Language Models Generated Code: An Empirical Study
por: Tambon, Florian, et al.
Publicado: (2024)
por: Tambon, Florian, et al.
Publicado: (2024)
Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry
por: Du, Xueying, et al.
Publicado: (2026)
por: Du, Xueying, et al.
Publicado: (2026)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
por: Zheng, Mingwei, et al.
Publicado: (2025)
por: Zheng, Mingwei, et al.
Publicado: (2025)
Bug Analysis Towards Bug Resolution Time Prediction
por: Ozkan, Hasan Yagiz, et al.
Publicado: (2024)
por: Ozkan, Hasan Yagiz, et al.
Publicado: (2024)
Ethical software requirements from user reviews: A systematic literature review
por: Sorathiya, Aakash, et al.
Publicado: (2024)
por: Sorathiya, Aakash, et al.
Publicado: (2024)
ETF: An Entity Tracing Framework for Hallucination Detection in Code Summaries
por: Maharaj, Kishan, et al.
Publicado: (2024)
por: Maharaj, Kishan, et al.
Publicado: (2024)
Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents
por: Liu, Xiang, et al.
Publicado: (2026)
por: Liu, Xiang, et al.
Publicado: (2026)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
por: Vulićević, Jelena Ilić
Publicado: (2026)
por: Vulićević, Jelena Ilić
Publicado: (2026)
ImproBR: Bug Report Improver Using LLMs
por: Akyol, Emre Furkan, et al.
Publicado: (2026)
por: Akyol, Emre Furkan, et al.
Publicado: (2026)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
por: Liu, Fang, et al.
Publicado: (2024)
por: Liu, Fang, et al.
Publicado: (2024)
Hallucination in LLM-Based Code Generation: An Automotive Case Study
por: Pavel, Marc, et al.
Publicado: (2025)
por: Pavel, Marc, et al.
Publicado: (2025)
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns
por: Samsonau, Sergey V.
Publicado: (2026)
por: Samsonau, Sergey V.
Publicado: (2026)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
por: Yu, Jiongchi, et al.
Publicado: (2025)
por: Yu, Jiongchi, et al.
Publicado: (2025)
Evaluating Generative AI for CS1 Code Grading: Direct vs Reverse Methods
por: Memon, Ahmad, et al.
Publicado: (2025)
por: Memon, Ahmad, et al.
Publicado: (2025)
Automated Bug Report Prioritization in Large Open-Source Projects
por: Pierson, Riley, et al.
Publicado: (2025)
por: Pierson, Riley, et al.
Publicado: (2025)
Eliminating Hallucination-Induced Errors in LLM Code Generation with Functional Clustering
por: Ravuri, Chaitanya, et al.
Publicado: (2025)
por: Ravuri, Chaitanya, et al.
Publicado: (2025)
A Survey of Bugs in AI-Generated Code
por: Gao, Ruofan, et al.
Publicado: (2025)
por: Gao, Ruofan, et al.
Publicado: (2025)
GitBugs: Bug Reports for Duplicate Detection, Retrieval Augmented Generation, Triage, and More
por: Patil, Avinash, et al.
Publicado: (2025)
por: Patil, Avinash, et al.
Publicado: (2025)
Towards Specification-Driven LLM-Based Generation of Embedded Automotive Software
por: Patil, Minal Suresh, et al.
Publicado: (2024)
por: Patil, Minal Suresh, et al.
Publicado: (2024)
HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions
por: Wang, Jing, et al.
Publicado: (2025)
por: Wang, Jing, et al.
Publicado: (2025)
BugSpotter: Automated Generation of Code Debugging Exercises
por: Pădurean, Victor-Alexandru, et al.
Publicado: (2024)
por: Pădurean, Victor-Alexandru, et al.
Publicado: (2024)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
por: Jiang, Weipeng, et al.
Publicado: (2025)
por: Jiang, Weipeng, et al.
Publicado: (2025)
One Bug, Hundreds Behind: LLMs for Large-Scale Bug Discovery
por: Wu, Qiushi, et al.
Publicado: (2025)
por: Wu, Qiushi, et al.
Publicado: (2025)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
por: Vitale, Antonio, et al.
Publicado: (2026)
por: Vitale, Antonio, et al.
Publicado: (2026)
Fine-Tuning Code Language Models to Detect Cross-Language Bugs
por: Li, Zengyang, et al.
Publicado: (2025)
por: Li, Zengyang, et al.
Publicado: (2025)
Past, Present, and Future of Bug Tracking in the Generative AI Era
por: Torun, Utku Boran, et al.
Publicado: (2025)
por: Torun, Utku Boran, et al.
Publicado: (2025)
Ejemplares similares
-
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
por: Acharya, Jagrit, et al.
Publicado: (2025) -
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
por: Acharya, Jagrit, et al.
Publicado: (2025) -
What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook
por: Huo, Junyu, et al.
Publicado: (2026) -
Beyond Keywords: A Context-based Hybrid Approach to Mining Ethical Concern-related App Reviews
por: Sorathiya, Aakash, et al.
Publicado: (2024) -
Towards Energy-aware Requirements Dependency Classification: Knowledge-Graph vs. Vector-Retrieval Augmented Inference with SLMs
por: Patil, Shreyas, et al.
Publicado: (2026)