Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Garg, Spandan, Huang, Yufan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BugSpotter: Automated Generation of Code Debugging Exercises
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2024)
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2024)
RAPGen: An Approach for Fixing Code Inefficiencies in Zero-Shot
von: Garg, Spandan, et al.
Veröffentlicht: (2023)
von: Garg, Spandan, et al.
Veröffentlicht: (2023)
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
von: Adnan, Muntasir, et al.
Veröffentlicht: (2025)
von: Adnan, Muntasir, et al.
Veröffentlicht: (2025)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
Saving SWE-Bench: A Benchmark Mutation Approach for Realistic Agent Evaluation
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
AgentStepper: Interactive Debugging of Software Development Agents
von: Hutter, Robert, et al.
Veröffentlicht: (2026)
von: Hutter, Robert, et al.
Veröffentlicht: (2026)
MarsCode Agent: AI-native Automated Bug Fixing
von: Liu, Yizhou, et al.
Veröffentlicht: (2024)
von: Liu, Yizhou, et al.
Veröffentlicht: (2024)
Towards Adaptive Software Agents for Debugging
von: Majdoub, Yacine, et al.
Veröffentlicht: (2025)
von: Majdoub, Yacine, et al.
Veröffentlicht: (2025)
MLDebugging: Towards Benchmarking Code Debugging Across Multi-Library Scenarios
von: Huang, Jinyang, et al.
Veröffentlicht: (2025)
von: Huang, Jinyang, et al.
Veröffentlicht: (2025)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
von: Meng, Xiangxin, et al.
Veröffentlicht: (2024)
von: Meng, Xiangxin, et al.
Veröffentlicht: (2024)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
von: Chen, Xiancai, et al.
Veröffentlicht: (2025)
von: Chen, Xiancai, et al.
Veröffentlicht: (2025)
Large Language Model Guided Self-Debugging Code Generation
von: Adnan, Muntasir, et al.
Veröffentlicht: (2025)
von: Adnan, Muntasir, et al.
Veröffentlicht: (2025)
Can Agents Fix Agent Issues?
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
von: Yang, Weiqing, et al.
Veröffentlicht: (2024)
von: Yang, Weiqing, et al.
Veröffentlicht: (2024)
DebugBench: Evaluating Debugging Capability of Large Language Models
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models
von: Liu, Jingjing, et al.
Veröffentlicht: (2025)
von: Liu, Jingjing, et al.
Veröffentlicht: (2025)
SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents
von: Mündler, Niels, et al.
Veröffentlicht: (2024)
von: Mündler, Niels, et al.
Veröffentlicht: (2024)
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents
von: Zhao, Chenyu, et al.
Veröffentlicht: (2026)
von: Zhao, Chenyu, et al.
Veröffentlicht: (2026)
DREAM: Debugging and Repairing AutoML Pipelines
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2023)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
von: Vitale, Antonio, et al.
Veröffentlicht: (2026)
von: Vitale, Antonio, et al.
Veröffentlicht: (2026)
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
von: Ma, Ming, et al.
Veröffentlicht: (2025)
von: Ma, Ming, et al.
Veröffentlicht: (2025)
DeepFix: Debugging and Fixing Machine Learning Workflow using Agentic AI
von: Seydou, Fadel Mamar, et al.
Veröffentlicht: (2026)
von: Seydou, Fadel Mamar, et al.
Veröffentlicht: (2026)
TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated Code
von: Huang, Jiangping, et al.
Veröffentlicht: (2026)
von: Huang, Jiangping, et al.
Veröffentlicht: (2026)
Enhancing Debugging Skills with AI-Powered Assistance: A Real-Time Tool for Debugging Support
von: Artser, Elizaveta, et al.
Veröffentlicht: (2026)
von: Artser, Elizaveta, et al.
Veröffentlicht: (2026)
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction
von: Wang, Ning, et al.
Veröffentlicht: (2025)
von: Wang, Ning, et al.
Veröffentlicht: (2025)
Post-hoc LLM-Supported Debugging of Distributed Processes
von: Schiese, Dennis, et al.
Veröffentlicht: (2025)
von: Schiese, Dennis, et al.
Veröffentlicht: (2025)
Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging
von: Liu, Zhilin, et al.
Veröffentlicht: (2026)
von: Liu, Zhilin, et al.
Veröffentlicht: (2026)
Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2
von: Qiu, Yu-Ning, et al.
Veröffentlicht: (2026)
von: Qiu, Yu-Ning, et al.
Veröffentlicht: (2026)
CODESIM: Multi-Agent Code Generation and Problem Solving through Simulation-Driven Planning and Debugging
von: Islam, Md. Ashraful, et al.
Veröffentlicht: (2025)
von: Islam, Md. Ashraful, et al.
Veröffentlicht: (2025)
Model See, Model Do? Exposure-Aware Evaluation of Bug-vs-Fix Preference in Code LLMs
von: Al-Kaswan, Ali, et al.
Veröffentlicht: (2026)
von: Al-Kaswan, Ali, et al.
Veröffentlicht: (2026)
LeDex: Training LLMs to Better Self-Debug and Explain Code
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports
von: Gon, Mahmut Furkan, et al.
Veröffentlicht: (2026)
von: Gon, Mahmut Furkan, et al.
Veröffentlicht: (2026)
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
von: Nandal, Deeksha, et al.
Veröffentlicht: (2026)
von: Nandal, Deeksha, et al.
Veröffentlicht: (2026)
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
von: Pham, Minh V. T., et al.
Veröffentlicht: (2025)
von: Pham, Minh V. T., et al.
Veröffentlicht: (2025)
AtPatch: Debugging Transformers via Hot-Fixing Over-Attention
von: Weng, Shihao, et al.
Veröffentlicht: (2026)
von: Weng, Shihao, et al.
Veröffentlicht: (2026)
Automated Multi-Source Debugging and Natural Language Error Explanation for Dashboard Applications
von: Tata, Devendra, et al.
Veröffentlicht: (2026)
von: Tata, Devendra, et al.
Veröffentlicht: (2026)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
Exploring Interaction Patterns for Debugging: Enhancing Conversational Capabilities of AI-assistants
von: Chopra, Bhavya, et al.
Veröffentlicht: (2024)
von: Chopra, Bhavya, et al.
Veröffentlicht: (2024)
debug-gym: A Text-Based Environment for Interactive Debugging
von: Yuan, Xingdi, et al.
Veröffentlicht: (2025)
von: Yuan, Xingdi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BugSpotter: Automated Generation of Code Debugging Exercises
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2024) -
RAPGen: An Approach for Fixing Code Inefficiencies in Zero-Shot
von: Garg, Spandan, et al.
Veröffentlicht: (2023) -
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
von: Adnan, Muntasir, et al.
Veröffentlicht: (2025) -
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026) -
PerfBench: Can Agents Resolve Real-World Performance Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2025)