HEJ-Robust: A Robustness Benchmark for LLM-Based Automated Program Repair
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rabbi, Fazle, Yang, Jinqiu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Multi-Language Perspective on the Robustness of LLM Code Generation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
Bias Unveiled: Investigating Social Bias in LLM-Generated Code
von: Ling, Lin, et al.
Veröffentlicht: (2024)
von: Ling, Lin, et al.
Veröffentlicht: (2024)
Social Bias in LLM-Generated Code: Benchmark and Mitigation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
Specification-Driven Code Translation Powered by Large Language Models: How Far Are We?
von: Saha, Soumit Kanti, et al.
Veröffentlicht: (2024)
von: Saha, Soumit Kanti, et al.
Veröffentlicht: (2024)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
von: Xue, Pengyu, et al.
Veröffentlicht: (2024)
von: Xue, Pengyu, et al.
Veröffentlicht: (2024)
An Exploratory Study on Fine-Tuning Large Language Models for Secure Code Generation
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
A Task-Level Evaluation of AI Agents in Open-Source Projects
von: Rahman, Shojibur, et al.
Veröffentlicht: (2026)
von: Rahman, Shojibur, et al.
Veröffentlicht: (2026)
BabelCoder: Agentic Code Translation with Specification Alignment
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
The Quiet Contributions: Insights into AI-Generated Silent Pull Requests
von: Hasan, S M Mahedy, et al.
Veröffentlicht: (2026)
von: Hasan, S M Mahedy, et al.
Veröffentlicht: (2026)
SWE-Refactor: A Repository-Level Benchmark for Real-World LLM-Based Code Refactoring
von: Xu, Yisen, et al.
Veröffentlicht: (2026)
von: Xu, Yisen, et al.
Veröffentlicht: (2026)
DebugRepair: Enhancing LLM-Based Automated Program Repair via Self-Directed Debugging
von: Wu, Linhao, et al.
Veröffentlicht: (2026)
von: Wu, Linhao, et al.
Veröffentlicht: (2026)
Chasing the Clock: How Fast Are Vulnerabilities Fixed in the Maven Ecosystem?
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2025)
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2025)
Understanding Software Vulnerabilities in the Maven Ecosystem: Patterns, Timelines, and Risks
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2025)
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2025)
Insights into Dependency Maintenance Trends in the Maven Ecosystem
von: Chowdhury, Barisha, et al.
Veröffentlicht: (2025)
von: Chowdhury, Barisha, et al.
Veröffentlicht: (2025)
Faster Releases, Fewer Risks: A Study on Maven Artifact Vulnerabilities and Lifecycle Management
von: Shafin, Md Shafiullah, et al.
Veröffentlicht: (2025)
von: Shafin, Md Shafiullah, et al.
Veröffentlicht: (2025)
On the Robustness Evaluation of 3D Obstacle Detection Against Specifications in Autonomous Driving
von: Pham, Tri Minh Triet, et al.
Veröffentlicht: (2024)
von: Pham, Tri Minh Triet, et al.
Veröffentlicht: (2024)
A Survey of LLM-based Automated Program Repair: Taxonomies, Design Paradigms, and Applications
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
ThinkRepair: Self-Directed Automated Program Repair
von: Yin, Xin, et al.
Veröffentlicht: (2024)
von: Yin, Xin, et al.
Veröffentlicht: (2024)
What's in a Benchmark? The Case of SWE-Bench in Automated Program Repair
von: Martinez, Matias, et al.
Veröffentlicht: (2026)
von: Martinez, Matias, et al.
Veröffentlicht: (2026)
Insights into Security-Related AI-Generated Pull Requests
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2026)
von: Rabbi, Md Fazle, et al.
Veröffentlicht: (2026)
The Fact Selection Problem in LLM-Based Program Repair
von: Parasaram, Nikhil, et al.
Veröffentlicht: (2024)
von: Parasaram, Nikhil, et al.
Veröffentlicht: (2024)
Semantic Evolution over Populations for LLM-Guided Automated Program Repair
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
The Impact of Program Reduction on Automated Program Repair
von: Vidziunas, Linas, et al.
Veröffentlicht: (2024)
von: Vidziunas, Linas, et al.
Veröffentlicht: (2024)
ASAP-Repair: API-Specific Automated Program Repair Based on API Usage Graphs
von: Nielebock, Sebastian, et al.
Veröffentlicht: (2024)
von: Nielebock, Sebastian, et al.
Veröffentlicht: (2024)
RePurr: Automated Repair of Block-Based Learners' Programs
von: Schweikl, Sebastian, et al.
Veröffentlicht: (2025)
von: Schweikl, Sebastian, et al.
Veröffentlicht: (2025)
Specification Vibing for Automated Program Repair
von: Zhu, Taohong, et al.
Veröffentlicht: (2026)
von: Zhu, Taohong, et al.
Veröffentlicht: (2026)
Energy Consumption of Automated Program Repair
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
MANTRA: Enhancing Automated Method-Level Refactoring with Contextual RAG and Multi-Agent LLM Collaboration
von: Xu, Yisen, et al.
Veröffentlicht: (2025)
von: Xu, Yisen, et al.
Veröffentlicht: (2025)
BUGSPHP: A dataset for Automated Program Repair in PHP
von: Pramod, K. D., et al.
Veröffentlicht: (2024)
von: Pramod, K. D., et al.
Veröffentlicht: (2024)
ContrastRepair: Enhancing Conversation-Based Automated Program Repair via Contrastive Test Case Pairs
von: Kong, Jiaolong, et al.
Veröffentlicht: (2024)
von: Kong, Jiaolong, et al.
Veröffentlicht: (2024)
A Systematic Literature Review on Large Language Models for Automated Program Repair
von: Zhang, Quanjun, et al.
Veröffentlicht: (2024)
von: Zhang, Quanjun, et al.
Veröffentlicht: (2024)
Automated Repair of Ambiguous Problem Descriptions for LLM-Based Code Generation
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
Input Reduction Enhanced LLM-based Program Repair
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
von: Bouzenia, Islem, et al.
Veröffentlicht: (2024)
von: Bouzenia, Islem, et al.
Veröffentlicht: (2024)
Enhancing Automated Program Repair with Solution Design
von: Zhao, Jiuang, et al.
Veröffentlicht: (2024)
von: Zhao, Jiuang, et al.
Veröffentlicht: (2024)
TSAPR: A Tree Search Framework For Automated Program Repair
von: Hu, Haichuan, et al.
Veröffentlicht: (2025)
von: Hu, Haichuan, et al.
Veröffentlicht: (2025)
PathFix: Automated Program Repair with Expected Path
von: He, Xu, et al.
Veröffentlicht: (2025)
von: He, Xu, et al.
Veröffentlicht: (2025)
On The Effectiveness of Dynamic Reduction Techniques in Automated Program Repair
von: Al-Bataineh, Omar I.
Veröffentlicht: (2024)
von: Al-Bataineh, Omar I.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Multi-Language Perspective on the Robustness of LLM Code Generation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025) -
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026) -
Bias Unveiled: Investigating Social Bias in LLM-Generated Code
von: Ling, Lin, et al.
Veröffentlicht: (2024) -
Social Bias in LLM-Generated Code: Benchmark and Mitigation
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026) -
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
von: Li, Junjie, et al.
Veröffentlicht: (2025)