LLMs as Evaluators: A Novel Approach to Evaluate Bug Report Summarization
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Abhishek, Haiduc, Sonia, Das, Partha Pratim, Chakrabarti, Partha Pratim |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Design-Aware Context in Large Language Models for Code Comment Generation
by: Mitra, Aritra, et al.
Published: (2025)
by: Mitra, Aritra, et al.
Published: (2025)
RLocator: Reinforcement Learning for Bug Localization
by: Chakraborty, Partha, et al.
Published: (2023)
by: Chakraborty, Partha, et al.
Published: (2023)
Amalgamation of Physics-Informed Neural Network and LBM for the Prediction of Unsteady Fluid Flows in Fractal-Rough Microchannels
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
Lattice-Boltzmann-Driven Physics-Informed Neural Networks for Droplet Wettability on Rough Surfaces
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
Extending deep learning U-Net architecture for predicting unsteady fluid flows in textured microchannels
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
$μ$-FlowNet: A Deep Learning Approach for Mapping Flow Fields in Irregular Microchannels Using an Attention-based U-Net Encoder-Decoder Architecture
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)
BLAZE: Cross-Language and Cross-Project Bug Localization via Dynamic Chunking and Hard Example Learning
by: Chakraborty, Partha, et al.
Published: (2024)
by: Chakraborty, Partha, et al.
Published: (2024)
Hierarchical Knowledge Injection for Improving LLM-based Program Repair
by: Ehsani, Ramtin, et al.
Published: (2025)
by: Ehsani, Ramtin, et al.
Published: (2025)
Investigating Developers' Preferences for Learning and Issue Resolution Resources in the ChatGPT Era
by: Tayeb, Ahmad, et al.
Published: (2024)
by: Tayeb, Ahmad, et al.
Published: (2024)
Towards A Sustainable Future for Peer Review in Software Engineering
by: Parra, Esteban, et al.
Published: (2026)
by: Parra, Esteban, et al.
Published: (2026)
An Empirical Study on the Capability of LLMs in Decomposing Bug Reports
by: Chen, Zhiyuan, et al.
Published: (2025)
by: Chen, Zhiyuan, et al.
Published: (2025)
What Characteristics Make ChatGPT Effective for Software Issue Resolution? An Empirical Study of Task, Project, and Conversational Signals in GitHub Issues
by: Ehsani, Ramtin, et al.
Published: (2025)
by: Ehsani, Ramtin, et al.
Published: (2025)
Perspective Chapter: MOOCs in India: Evolution, Innovation, Impact, and Roadmap
by: Das, Partha Pratim
Published: (2025)
by: Das, Partha Pratim
Published: (2025)
A Formal Verification Approach to Safeguard Controller Variables from Single Event Upset
by: Ganesha, et al.
Published: (2025)
by: Ganesha, et al.
Published: (2025)
Progressive Code Integration for Abstractive Bug Report Summarization
by: Karim, Shaira Sadia, et al.
Published: (2025)
by: Karim, Shaira Sadia, et al.
Published: (2025)
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
by: Hu, Haichuan, et al.
Published: (2024)
by: Hu, Haichuan, et al.
Published: (2024)
Bug Whispering: Towards Audio Bug Reporting
by: Masserini, Elena, et al.
Published: (2025)
by: Masserini, Elena, et al.
Published: (2025)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
by: Guo, Liwei, et al.
Published: (2025)
by: Guo, Liwei, et al.
Published: (2025)
Deep Learning-based Code Reviews: A Paradigm Shift or a Double-Edged Sword?
by: Tufano, Rosalia, et al.
Published: (2024)
by: Tufano, Rosalia, et al.
Published: (2024)
Aligning Programming Language and Natural Language: Exploring Design Choices in Multi-Modal Transformer-Based Embedding for Bug Localization
by: Chakraborty, Partha, et al.
Published: (2024)
by: Chakraborty, Partha, et al.
Published: (2024)
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
by: Gheyi, Rohit, et al.
Published: (2025)
by: Gheyi, Rohit, et al.
Published: (2025)
ImproBR: Bug Report Improver Using LLMs
by: Akyol, Emre Furkan, et al.
Published: (2026)
by: Akyol, Emre Furkan, et al.
Published: (2026)
Test Case Generation from Bug Reports via Large Language Models: A Cognitive Layered Evaluation Framework
by: Qureshi, Irtaza Sajid, et al.
Published: (2025)
by: Qureshi, Irtaza Sajid, et al.
Published: (2025)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
by: Acharya, Jagrit, et al.
Published: (2025)
by: Acharya, Jagrit, et al.
Published: (2025)
Automated Generation of High-Quality Bug Reports for Android Applications
by: Saha, Antu, et al.
Published: (2026)
by: Saha, Antu, et al.
Published: (2026)
SysPro: Reproducing System-level Concurrency Bugs from Bug Reports
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
GitBugs: Bug Reports for Duplicate Detection, Retrieval Augmented Generation, Triage, and More
by: Patil, Avinash, et al.
Published: (2025)
by: Patil, Avinash, et al.
Published: (2025)
Coding in a Bubble? Evaluating LLMs in Resolving Context Adaptation Bugs During Code Adaptation
by: Zhang, Tanghaoran, et al.
Published: (2026)
by: Zhang, Tanghaoran, et al.
Published: (2026)
English Please: Evaluating Machine Translation with Large Language Models for Multilingual Bug Reports
by: Patil, Avinash, et al.
Published: (2025)
by: Patil, Avinash, et al.
Published: (2025)
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
by: Acharya, Jagrit, et al.
Published: (2025)
by: Acharya, Jagrit, et al.
Published: (2025)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
by: Vulićević, Jelena Ilić
Published: (2026)
by: Vulićević, Jelena Ilić
Published: (2026)
Human Side of Smart Contract Fuzzing: An Empirical Study
by: Qiao, Guanming, et al.
Published: (2025)
by: Qiao, Guanming, et al.
Published: (2025)
Can Large Language Models Serve as Evaluators for Code Summarization?
by: Wu, Yang, et al.
Published: (2024)
by: Wu, Yang, et al.
Published: (2024)
Model See, Model Do? Exposure-Aware Evaluation of Bug-vs-Fix Preference in Code LLMs
by: Al-Kaswan, Ali, et al.
Published: (2026)
by: Al-Kaswan, Ali, et al.
Published: (2026)
Can LLMs Find Bugs in Code? An Evaluation from Beginner Errors to Security Vulnerabilities in Python and C++
by: Mhatre, Akshay, et al.
Published: (2025)
by: Mhatre, Akshay, et al.
Published: (2025)
AI-Powered Commit Explorer (APCE)
by: Grees, Yousab, et al.
Published: (2025)
by: Grees, Yousab, et al.
Published: (2025)
One Bug, Hundreds Behind: LLMs for Large-Scale Bug Discovery
by: Wu, Qiushi, et al.
Published: (2025)
by: Wu, Qiushi, et al.
Published: (2025)
ImageR: Enhancing Bug Report Clarity by Screenshots
by: Tan, Xuchen, et al.
Published: (2025)
by: Tan, Xuchen, et al.
Published: (2025)
Identifying Concurrency Bug Reports via Linguistic Patterns
by: Shao, Shuai, et al.
Published: (2026)
by: Shao, Shuai, et al.
Published: (2026)
Duplicate Bug Report Detection: How Far Are We?
by: Zhang, Ting, et al.
Published: (2022)
by: Zhang, Ting, et al.
Published: (2022)
Similar Items
-
Leveraging Design-Aware Context in Large Language Models for Code Comment Generation
by: Mitra, Aritra, et al.
Published: (2025) -
RLocator: Reinforcement Learning for Bug Localization
by: Chakraborty, Partha, et al.
Published: (2023) -
Amalgamation of Physics-Informed Neural Network and LBM for the Prediction of Unsteady Fluid Flows in Fractal-Rough Microchannels
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026) -
Lattice-Boltzmann-Driven Physics-Informed Neural Networks for Droplet Wettability on Rough Surfaces
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026) -
Extending deep learning U-Net architecture for predicting unsteady fluid flows in textured microchannels
by: Meshram, Ganesh Sahadeo, et al.
Published: (2026)