Rethinking Kernel Program Repair: Benchmarking and Enhancing LLMs with RGym
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shehada, Kareem, Wu, Yifan, Feng, Wyatt D., Iyer, Adithya, Kumfert, Gryphon, Ding, Yangruibo, Qian, Zhiyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Patch-to-PoC: A Systematic Study of Agentic LLM Systems for Linux Kernel N-Day Reproduction
von: Pu, Juefei, et al.
Veröffentlicht: (2026)
von: Pu, Juefei, et al.
Veröffentlicht: (2026)
The Hitchhiker's Guide to Program Analysis, Part II: Deep Thoughts by LLMs
von: Li, Haonan, et al.
Veröffentlicht: (2025)
von: Li, Haonan, et al.
Veröffentlicht: (2025)
Logging Like Humans for LLMs: Rethinking Logging via Execution and Runtime Feedback
von: Wang, Xin, et al.
Veröffentlicht: (2026)
von: Wang, Xin, et al.
Veröffentlicht: (2026)
An Investigation of Patch Porting Practices of the Linux Kernel Ecosystem
von: Li, Xingyu, et al.
Veröffentlicht: (2024)
von: Li, Xingyu, et al.
Veröffentlicht: (2024)
DebugRepair: Enhancing LLM-Based Automated Program Repair via Self-Directed Debugging
von: Wu, Linhao, et al.
Veröffentlicht: (2026)
von: Wu, Linhao, et al.
Veröffentlicht: (2026)
CYCLE: Learning to Self-Refine the Code Generation
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
TritonForge: Profiling-Guided Framework for Automated Triton Kernel Optimization
von: Li, Haonan, et al.
Veröffentlicht: (2025)
von: Li, Haonan, et al.
Veröffentlicht: (2025)
Knowledge-Enhanced Program Repair for Data Science Code
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2025)
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2025)
Input Reduction Enhanced LLM-based Program Repair
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
von: Yang, Boyang, et al.
Veröffentlicht: (2025)
What's in a Benchmark? The Case of SWE-Bench in Automated Program Repair
von: Martinez, Matias, et al.
Veröffentlicht: (2026)
von: Martinez, Matias, et al.
Veröffentlicht: (2026)
Specification-Guided Repair of Arithmetic Errors in Dafny Programs using LLMs
von: Wu, Valentina, et al.
Veröffentlicht: (2025)
von: Wu, Valentina, et al.
Veröffentlicht: (2025)
Rethinking the Evaluation of Microservice RCA with a Fault Propagation-Aware Benchmark
von: Fang, Aoyang, et al.
Veröffentlicht: (2025)
von: Fang, Aoyang, et al.
Veröffentlicht: (2025)
Evaluating the Generalizability of LLMs in Automated Program Repair
von: Li, Fengjie, et al.
Veröffentlicht: (2025)
von: Li, Fengjie, et al.
Veröffentlicht: (2025)
MergeRepair: An Exploratory Study on Merging Task-Specific Adapters in Code LLMs for Automated Program Repair
von: Dehghan, Meghdad, et al.
Veröffentlicht: (2024)
von: Dehghan, Meghdad, et al.
Veröffentlicht: (2024)
HEJ-Robust: A Robustness Benchmark for LLM-Based Automated Program Repair
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2026)
RESTORE: Retrospective Fault Localization Enhancing Automated Program Repair
von: Xu, Tongtong, et al.
Veröffentlicht: (2019)
von: Xu, Tongtong, et al.
Veröffentlicht: (2019)
Adapting Knowledge Prompt Tuning for Enhanced Automated Program Repair
von: Cai, Xuemeng, et al.
Veröffentlicht: (2025)
von: Cai, Xuemeng, et al.
Veröffentlicht: (2025)
Enhancing Program Repair with Specification Guidance and Intermediate Behavioral Signals
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
ContrastRepair: Enhancing Conversation-Based Automated Program Repair via Contrastive Test Case Pairs
von: Kong, Jiaolong, et al.
Veröffentlicht: (2024)
von: Kong, Jiaolong, et al.
Veröffentlicht: (2024)
ThinkRepair: Self-Directed Automated Program Repair
von: Yin, Xin, et al.
Veröffentlicht: (2024)
von: Yin, Xin, et al.
Veröffentlicht: (2024)
CigaR: Cost-efficient Program Repair with LLMs
von: Hidvégi, Dávid, et al.
Veröffentlicht: (2024)
von: Hidvégi, Dávid, et al.
Veröffentlicht: (2024)
Towards Practical and Useful Automated Program Repair for Debugging
von: Xin, Qi, et al.
Veröffentlicht: (2024)
von: Xin, Qi, et al.
Veröffentlicht: (2024)
The Impact of Program Reduction on Automated Program Repair
von: Vidziunas, Linas, et al.
Veröffentlicht: (2024)
von: Vidziunas, Linas, et al.
Veröffentlicht: (2024)
Beyond Crash-to-Patch: Patch Evolution for Linux Kernel Repair
von: Bai, Luyao, et al.
Veröffentlicht: (2026)
von: Bai, Luyao, et al.
Veröffentlicht: (2026)
Invariant-based Program Repair
von: Al-Bataineh, Omar I.
Veröffentlicht: (2023)
von: Al-Bataineh, Omar I.
Veröffentlicht: (2023)
Execution-free Program Repair
von: Huang, Li, et al.
Veröffentlicht: (2024)
von: Huang, Li, et al.
Veröffentlicht: (2024)
EXPEREPAIR: Dual-Memory Enhanced LLM-based Repository-Level Program Repair
von: Mu, Fangwen, et al.
Veröffentlicht: (2025)
von: Mu, Fangwen, et al.
Veröffentlicht: (2025)
PALM: Synergizing Program Analysis and LLMs to Enhance Rust Unit Test Coverage
von: Chu, Bei, et al.
Veröffentlicht: (2025)
von: Chu, Bei, et al.
Veröffentlicht: (2025)
SWE-Bench+: Enhanced Coding Benchmark for LLMs
von: Aleithan, Reem, et al.
Veröffentlicht: (2024)
von: Aleithan, Reem, et al.
Veröffentlicht: (2024)
Automated Code Editing with Search-Generate-Modify
von: Liu, Changshu, et al.
Veröffentlicht: (2023)
von: Liu, Changshu, et al.
Veröffentlicht: (2023)
Enhancing Automated Program Repair with Solution Design
von: Zhao, Jiuang, et al.
Veröffentlicht: (2024)
von: Zhao, Jiuang, et al.
Veröffentlicht: (2024)
Benchmark Dataset Generation and Evaluation for Excel Formula Repair with LLMs
von: Singha, Ananya, et al.
Veröffentlicht: (2025)
von: Singha, Ananya, et al.
Veröffentlicht: (2025)
From Benchmark Data To Applicable Program Repair: An Experience Report
von: Chandramohan, Mahinthan, et al.
Veröffentlicht: (2025)
von: Chandramohan, Mahinthan, et al.
Veröffentlicht: (2025)
Boosting Open-Source LLMs for Program Repair via Reasoning Transfer and LLM-Guided Reinforcement Learning
von: Tang, Xunzhu, et al.
Veröffentlicht: (2025)
von: Tang, Xunzhu, et al.
Veröffentlicht: (2025)
BUGSPHP: A dataset for Automated Program Repair in PHP
von: Pramod, K. D., et al.
Veröffentlicht: (2024)
von: Pramod, K. D., et al.
Veröffentlicht: (2024)
Is Measurement Enough? Rethinking Output Validation in Quantum Program Testing
von: Ye, Jiaming, et al.
Veröffentlicht: (2025)
von: Ye, Jiaming, et al.
Veröffentlicht: (2025)
Rethinking the Capability of Fine-Tuned Language Models for Automated Vulnerability Repair
von: Han, Woorim, et al.
Veröffentlicht: (2025)
von: Han, Woorim, et al.
Veröffentlicht: (2025)
Specification Vibing for Automated Program Repair
von: Zhu, Taohong, et al.
Veröffentlicht: (2026)
von: Zhu, Taohong, et al.
Veröffentlicht: (2026)
Energy Consumption of Automated Program Repair
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs
von: Haque, Mirazul, et al.
Veröffentlicht: (2025)
von: Haque, Mirazul, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Patch-to-PoC: A Systematic Study of Agentic LLM Systems for Linux Kernel N-Day Reproduction
von: Pu, Juefei, et al.
Veröffentlicht: (2026) -
The Hitchhiker's Guide to Program Analysis, Part II: Deep Thoughts by LLMs
von: Li, Haonan, et al.
Veröffentlicht: (2025) -
Logging Like Humans for LLMs: Rethinking Logging via Execution and Runtime Feedback
von: Wang, Xin, et al.
Veröffentlicht: (2026) -
An Investigation of Patch Porting Practices of the Linux Kernel Ecosystem
von: Li, Xingyu, et al.
Veröffentlicht: (2024) -
DebugRepair: Enhancing LLM-Based Automated Program Repair via Self-Directed Debugging
von: Wu, Linhao, et al.
Veröffentlicht: (2026)