Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiu, Yu-Ning, Zou, Lin-Feng, Wang, Jiong-Da, Yuan, Xue-Rong, Dai, Wang-Zhou |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction
di: Wang, Ning, et al.
Pubblicazione: (2025)
di: Wang, Ning, et al.
Pubblicazione: (2025)
DebugTA: An LLM-Based Agent for Simplifying Debugging and Teaching in Programming Education
di: Fu, Lingyue, et al.
Pubblicazione: (2025)
di: Fu, Lingyue, et al.
Pubblicazione: (2025)
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
di: Ma, Ming, et al.
Pubblicazione: (2025)
di: Ma, Ming, et al.
Pubblicazione: (2025)
Post-hoc LLM-Supported Debugging of Distributed Processes
di: Schiese, Dennis, et al.
Pubblicazione: (2025)
di: Schiese, Dennis, et al.
Pubblicazione: (2025)
RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models
di: Liu, Jingjing, et al.
Pubblicazione: (2025)
di: Liu, Jingjing, et al.
Pubblicazione: (2025)
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
di: Adnan, Muntasir, et al.
Pubblicazione: (2025)
di: Adnan, Muntasir, et al.
Pubblicazione: (2025)
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
di: Nandal, Deeksha, et al.
Pubblicazione: (2026)
di: Nandal, Deeksha, et al.
Pubblicazione: (2026)
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
di: Yang, Weiqing, et al.
Pubblicazione: (2024)
di: Yang, Weiqing, et al.
Pubblicazione: (2024)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
di: Bai, Yunsheng, et al.
Pubblicazione: (2025)
di: Bai, Yunsheng, et al.
Pubblicazione: (2025)
MLDebugging: Towards Benchmarking Code Debugging Across Multi-Library Scenarios
di: Huang, Jinyang, et al.
Pubblicazione: (2025)
di: Huang, Jinyang, et al.
Pubblicazione: (2025)
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
di: Garg, Spandan, et al.
Pubblicazione: (2026)
di: Garg, Spandan, et al.
Pubblicazione: (2026)
DebugBench: Evaluating Debugging Capability of Large Language Models
di: Tian, Runchu, et al.
Pubblicazione: (2024)
di: Tian, Runchu, et al.
Pubblicazione: (2024)
DebugRepair: Enhancing LLM-Based Automated Program Repair via Self-Directed Debugging
di: Wu, Linhao, et al.
Pubblicazione: (2026)
di: Wu, Linhao, et al.
Pubblicazione: (2026)
Towards Adaptive Software Agents for Debugging
di: Majdoub, Yacine, et al.
Pubblicazione: (2025)
di: Majdoub, Yacine, et al.
Pubblicazione: (2025)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
di: Chen, Xiancai, et al.
Pubblicazione: (2025)
di: Chen, Xiancai, et al.
Pubblicazione: (2025)
DREAM: Debugging and Repairing AutoML Pipelines
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2023)
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2023)
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents
di: Zhao, Chenyu, et al.
Pubblicazione: (2026)
di: Zhao, Chenyu, et al.
Pubblicazione: (2026)
Enhancing Debugging Skills with AI-Powered Assistance: A Real-Time Tool for Debugging Support
di: Artser, Elizaveta, et al.
Pubblicazione: (2026)
di: Artser, Elizaveta, et al.
Pubblicazione: (2026)
LLM as an Execution Estimator: Recovering Missing Dependency for Practical Time-travelling Debugging
di: Pei, Yunrui, et al.
Pubblicazione: (2025)
di: Pei, Yunrui, et al.
Pubblicazione: (2025)
BugSpotter: Automated Generation of Code Debugging Exercises
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
AgentStepper: Interactive Debugging of Software Development Agents
di: Hutter, Robert, et al.
Pubblicazione: (2026)
di: Hutter, Robert, et al.
Pubblicazione: (2026)
Large Language Model Guided Self-Debugging Code Generation
di: Adnan, Muntasir, et al.
Pubblicazione: (2025)
di: Adnan, Muntasir, et al.
Pubblicazione: (2025)
LeDex: Training LLMs to Better Self-Debug and Explain Code
di: Jiang, Nan, et al.
Pubblicazione: (2024)
di: Jiang, Nan, et al.
Pubblicazione: (2024)
DebugHarness: Emulating Human Dynamic Debugging for Autonomous Program Repair
di: Sun, Maolin, et al.
Pubblicazione: (2026)
di: Sun, Maolin, et al.
Pubblicazione: (2026)
Options, Not Clicks: Lattice Refinement for Consent-Driven MCP Authorization
di: Li, Ying, et al.
Pubblicazione: (2026)
di: Li, Ying, et al.
Pubblicazione: (2026)
Accelerating Delta Debugging through Probabilistic Monotonicity Assessment
di: Tao, Yonggang, et al.
Pubblicazione: (2025)
di: Tao, Yonggang, et al.
Pubblicazione: (2025)
Automated Multi-Source Debugging and Natural Language Error Explanation for Dashboard Applications
di: Tata, Devendra, et al.
Pubblicazione: (2026)
di: Tata, Devendra, et al.
Pubblicazione: (2026)
Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution
di: Yu, Kai, et al.
Pubblicazione: (2026)
di: Yu, Kai, et al.
Pubblicazione: (2026)
Guided Debugging of Auto-Translated Code Using Differential Testing
di: Wu, Shengnan, et al.
Pubblicazione: (2025)
di: Wu, Shengnan, et al.
Pubblicazione: (2025)
ScriptSmith: A Unified LLM Framework for Enhancing IT Operations via Automated Bash Script Generation, Assessment, and Refinement
di: Chatterjee, Oishik, et al.
Pubblicazione: (2024)
di: Chatterjee, Oishik, et al.
Pubblicazione: (2024)
PropertyGPT: LLM-driven Formal Verification of Smart Contracts through Retrieval-Augmented Property Generation
di: Liu, Ye, et al.
Pubblicazione: (2024)
di: Liu, Ye, et al.
Pubblicazione: (2024)
TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated Code
di: Huang, Jiangping, et al.
Pubblicazione: (2026)
di: Huang, Jiangping, et al.
Pubblicazione: (2026)
InfCode: Adversarial Iterative Refinement of Tests and Patches for Reliable Software Issue Resolution
di: Li, KeFan, et al.
Pubblicazione: (2025)
di: Li, KeFan, et al.
Pubblicazione: (2025)
Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
Effective LLM Code Refinement via Property-Oriented and Structurally Minimal Feedback
di: He, Lehan, et al.
Pubblicazione: (2025)
di: He, Lehan, et al.
Pubblicazione: (2025)
Blueprint First, Model Second: A Framework for Deterministic LLM Workflow
di: Qiu, Libin, et al.
Pubblicazione: (2025)
di: Qiu, Libin, et al.
Pubblicazione: (2025)
Vul-RAG: Enhancing LLM-based Vulnerability Detection via Knowledge-level RAG
di: Du, Xueying, et al.
Pubblicazione: (2024)
di: Du, Xueying, et al.
Pubblicazione: (2024)
Uncertainty Quantification for LLM-based Code Generation
di: Xu, Senrong, et al.
Pubblicazione: (2026)
di: Xu, Senrong, et al.
Pubblicazione: (2026)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
di: Bouchoucha, Rached, et al.
Pubblicazione: (2024)
di: Bouchoucha, Rached, et al.
Pubblicazione: (2024)
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
di: Tao, Wei, et al.
Pubblicazione: (2024)
di: Tao, Wei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction
di: Wang, Ning, et al.
Pubblicazione: (2025) -
DebugTA: An LLM-Based Agent for Simplifying Debugging and Teaching in Programming Education
di: Fu, Lingyue, et al.
Pubblicazione: (2025) -
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
di: Ma, Ming, et al.
Pubblicazione: (2025) -
Post-hoc LLM-Supported Debugging of Distributed Processes
di: Schiese, Dennis, et al.
Pubblicazione: (2025) -
RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models
di: Liu, Jingjing, et al.
Pubblicazione: (2025)