BloomAPR: A Bloom's Taxonomy-based Framework for Assessing the Capabilities of LLM-Powered APR Solutions
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Yinghang, Shin, Jiho, Da Silva, Leuson, Ming, Zhen, Jiang, Wang, Song, Khomh, Foutse, Tan, Shin Hwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Taxonomy of Inefficiencies in LLM-Generated Python Code
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
by: Da Silva, Leuson, et al.
Published: (2024)
by: Da Silva, Leuson, et al.
Published: (2024)
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
by: P, Akilesh, et al.
Published: (2026)
by: P, Akilesh, et al.
Published: (2026)
Performance Smells in ML and Non-ML Python Projects: A Comparative Study
by: Belias, François, et al.
Published: (2025)
by: Belias, François, et al.
Published: (2025)
Exploring Security Practices in Infrastructure as Code: An Empirical Study
by: Verdet, Alexandre, et al.
Published: (2023)
by: Verdet, Alexandre, et al.
Published: (2023)
An Empirical Study of Policy-as-Code Adoption in Open-Source Software Projects
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
ReCatcher: Towards LLMs Regression Testing for Code Generation
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
Continuously Learning Bug Locations
by: Mindom, Paulina Stevia Nouwou, et al.
Published: (2024)
by: Mindom, Paulina Stevia Nouwou, et al.
Published: (2024)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Logging Requirement for Continuous Auditing of Responsible Machine Learning-based Applications
by: Foalem, Patrick Loic, et al.
Published: (2025)
by: Foalem, Patrick Loic, et al.
Published: (2025)
Structured Safety Auditing for Balancing Code Correctness and Content Safety in LLM-Generated Code
by: Tan, Honghao, et al.
Published: (2026)
by: Tan, Honghao, et al.
Published: (2026)
Real Faults in Model Context Protocol (MCP) Software: a Comprehensive Taxonomy
by: Taraghi, Mina, et al.
Published: (2026)
by: Taraghi, Mina, et al.
Published: (2026)
Historian: Reducing Manual Validation in APR Benchmarking via Evidence-Based Assessment
by: Moslemi, Sahand, et al.
Published: (2026)
by: Moslemi, Sahand, et al.
Published: (2026)
Beyond Localization: Recoverable Headroom and Residual Frontier in Repository-Level RAG-APR
by: Zhao, Pengtao, et al.
Published: (2026)
by: Zhao, Pengtao, et al.
Published: (2026)
Structural Anchors and Reasoning Fragility:Understanding CoT Robustness in LLM4Code
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Testing Refactoring Engine via Historical Bug Report driven LLM
by: Wang, Haibo, et al.
Published: (2025)
by: Wang, Haibo, et al.
Published: (2025)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
by: Oueslati, Khouloud, et al.
Published: (2025)
by: Oueslati, Khouloud, et al.
Published: (2025)
T5APR: Empowering Automated Program Repair across Languages through Checkpoint Ensemble
by: Gharibi, Reza, et al.
Published: (2023)
by: Gharibi, Reza, et al.
Published: (2023)
Dissecting Bug Triggers and Failure Modes in Modern Agentic Frameworks: An Empirical Study
by: Zhang, Xiaowen, et al.
Published: (2026)
by: Zhang, Xiaowen, et al.
Published: (2026)
Impact of LLM-based Review Comment Generation in Practice: A Mixed Open-/Closed-source User Study
by: Olewicki, Doriane, et al.
Published: (2024)
by: Olewicki, Doriane, et al.
Published: (2024)
LLM-Guided Issue Generation from Uncovered Code Segments
by: Pressato, Diany, et al.
Published: (2026)
by: Pressato, Diany, et al.
Published: (2026)
Characterizing Faults in Agentic AI: A Taxonomy of Types, Symptoms, and Root Causes
by: Shah, Mehil B, et al.
Published: (2026)
by: Shah, Mehil B, et al.
Published: (2026)
Refactoring with LLMs: Bridging Human Expertise and Machine Understanding
by: Piao, Yonnel Chen Kuang, et al.
Published: (2025)
by: Piao, Yonnel Chen Kuang, et al.
Published: (2025)
Investigating Code Reuse in Software Redesign: A Case Study
by: Zhang, Xiaowen, et al.
Published: (2026)
by: Zhang, Xiaowen, et al.
Published: (2026)
Guiding ChatGPT to Fix Web UI Tests via Explanation-Consistency Checking
by: Xu, Zhuolin, et al.
Published: (2023)
by: Xu, Zhuolin, et al.
Published: (2023)
Assessing Evaluation Metrics for Neural Test Oracle Generation
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
Aligning the Objective of LLM-based Program Repair
by: Xu, Junjielong, et al.
Published: (2024)
by: Xu, Junjielong, et al.
Published: (2024)
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
by: Li, Yuanhao, et al.
Published: (2026)
by: Li, Yuanhao, et al.
Published: (2026)
Ethics Testing: Proactive Identification of Generative AI System Harms
by: Tan, Shin Hwei, et al.
Published: (2026)
by: Tan, Shin Hwei, et al.
Published: (2026)
Protecting Privacy in Software Logs: What Should Be Anonymized?
by: Aghili, Roozbeh, et al.
Published: (2024)
by: Aghili, Roozbeh, et al.
Published: (2024)
COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation
by: Dau, Anh T. V., et al.
Published: (2026)
by: Dau, Anh T. V., et al.
Published: (2026)
SDLog: A Deep Learning Framework for Detecting Sensitive Information in Software Logs
by: Aghili, Roozbeh, et al.
Published: (2025)
by: Aghili, Roozbeh, et al.
Published: (2025)
Machine Learning Robustness: A Primer
by: Braiek, Houssem Ben, et al.
Published: (2024)
by: Braiek, Houssem Ben, et al.
Published: (2024)
Understanding and Detecting Annotation-Induced Faults of Static Analyzers
by: Zhang, Huaien, et al.
Published: (2024)
by: Zhang, Huaien, et al.
Published: (2024)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
by: Wu, Xingfang, et al.
Published: (2023)
by: Wu, Xingfang, et al.
Published: (2023)
GIST: Generated Inputs Sets Transferability in Deep Learning
by: Tambon, Florian, et al.
Published: (2023)
by: Tambon, Florian, et al.
Published: (2023)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024)
by: Abukhalaf, Seif, et al.
Published: (2024)
Similar Items
-
A Taxonomy of Inefficiencies in LLM-Generated Python Code
by: Abbassi, Altaf Allah, et al.
Published: (2025) -
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
by: Da Silva, Leuson, et al.
Published: (2024) -
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
by: P, Akilesh, et al.
Published: (2026) -
Performance Smells in ML and Non-ML Python Projects: A Comparative Study
by: Belias, François, et al.
Published: (2025) -
Exploring Security Practices in Infrastructure as Code: An Empirical Study
by: Verdet, Alexandre, et al.
Published: (2023)