TGPR: Tree-Guided Policy Refinement for Robust Self-Debugging of LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Ozerova, Daria, Trofimova, Ekaterina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CodeRefine: A Pipeline for Enhancing LLM-Generated Code Implementations of Research Papers
by: Trofimova, Ekaterina, et al.
Published: (2024)
by: Trofimova, Ekaterina, et al.
Published: (2024)
What Makes Cryptic Crosswords Challenging for LLMs?
by: Sadallah, Abdelrahman, et al.
Published: (2024)
by: Sadallah, Abdelrahman, et al.
Published: (2024)
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
by: Adnan, Muntasir, et al.
Published: (2025)
by: Adnan, Muntasir, et al.
Published: (2025)
Large Language Model Guided Self-Debugging Code Generation
by: Adnan, Muntasir, et al.
Published: (2025)
by: Adnan, Muntasir, et al.
Published: (2025)
Are LLMs Good Cryptic Crossword Solvers?
by: Sadallah, Abdelrahman, et al.
Published: (2024)
by: Sadallah, Abdelrahman, et al.
Published: (2024)
Self-Abstraction from Grounded Experience for Plan-Guided Policy Refinement
by: Hayashi, Hiroaki, et al.
Published: (2025)
by: Hayashi, Hiroaki, et al.
Published: (2025)
Scientific Reasoning: Assessment of Multimodal Generative LLMs
by: Dreyer, Florian, et al.
Published: (2025)
by: Dreyer, Florian, et al.
Published: (2025)
LeDex: Training LLMs to Better Self-Debug and Explain Code
by: Jiang, Nan, et al.
Published: (2024)
by: Jiang, Nan, et al.
Published: (2024)
Linguacodus: A Synergistic Framework for Transformative Code Generation in Machine Learning Pipelines
by: Trofimova, Ekaterina, et al.
Published: (2024)
by: Trofimova, Ekaterina, et al.
Published: (2024)
Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2
by: Qiu, Yu-Ning, et al.
Published: (2026)
by: Qiu, Yu-Ning, et al.
Published: (2026)
Enhancing Debugging Skills with AI-Powered Assistance: A Real-Time Tool for Debugging Support
by: Artser, Elizaveta, et al.
Published: (2026)
by: Artser, Elizaveta, et al.
Published: (2026)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
by: Chen, Xiancai, et al.
Published: (2025)
by: Chen, Xiancai, et al.
Published: (2025)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
by: Wang, Qibin, et al.
Published: (2025)
by: Wang, Qibin, et al.
Published: (2025)
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction
by: Wang, Ning, et al.
Published: (2025)
by: Wang, Ning, et al.
Published: (2025)
Beyond Text-to-SQL: Can LLMs Really Debug Enterprise ETL SQL?
by: Ye, Jing, et al.
Published: (2026)
by: Ye, Jing, et al.
Published: (2026)
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
by: Jin, Chen, et al.
Published: (2026)
by: Jin, Chen, et al.
Published: (2026)
Evolutionary Guided Decoding: Iterative Value Refinement for LLMs
by: Liu, Zhenhua, et al.
Published: (2025)
by: Liu, Zhenhua, et al.
Published: (2025)
Bridging Interpretability and Robustness Using LIME-Guided Model Refinement
by: Nayyem, Navid, et al.
Published: (2024)
by: Nayyem, Navid, et al.
Published: (2024)
Divide-Verify-Refine: Can LLMs Self-Align with Complex Instructions?
by: Zhang, Xianren, et al.
Published: (2024)
by: Zhang, Xianren, et al.
Published: (2024)
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
by: Ayalew, Tewodros, et al.
Published: (2024)
by: Ayalew, Tewodros, et al.
Published: (2024)
DebugBench: Evaluating Debugging Capability of Large Language Models
by: Tian, Runchu, et al.
Published: (2024)
by: Tian, Runchu, et al.
Published: (2024)
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
by: Yang, Weiqing, et al.
Published: (2024)
by: Yang, Weiqing, et al.
Published: (2024)
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
by: Garg, Spandan, et al.
Published: (2026)
by: Garg, Spandan, et al.
Published: (2026)
MAHL: Multi-Agent LLM-Guided Hierarchical Chiplet Design with Adaptive Debugging
by: Tang, Jinwei, et al.
Published: (2025)
by: Tang, Jinwei, et al.
Published: (2025)
RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models
by: Liu, Jingjing, et al.
Published: (2025)
by: Liu, Jingjing, et al.
Published: (2025)
One-Way Policy Optimization for Self-Evolving LLMs
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
RefineRL: Advancing Competitive Programming with Self-Refinement Reinforcement Learning
by: Fu, Shaopeng, et al.
Published: (2026)
by: Fu, Shaopeng, et al.
Published: (2026)
A Systematic Approach for Large Language Models Debugging
by: Shbita, Basel, et al.
Published: (2026)
by: Shbita, Basel, et al.
Published: (2026)
Generative Large Language Models (gLLMs) in Content Analysis: A Practical Guide for Communication Research
by: Kravets-Meinke, Daria, et al.
Published: (2025)
by: Kravets-Meinke, Daria, et al.
Published: (2025)
Enhancing the Medical Context-Awareness Ability of LLMs via Multifaceted Self-Refinement Learning
by: Zhou, Yuxuan, et al.
Published: (2025)
by: Zhou, Yuxuan, et al.
Published: (2025)
Disentangle-then-Refine: LLM-Guided Decoupling and Structure-Aware Refinement for Graph Contrastive Learning
by: Li, Zhaoxing, et al.
Published: (2026)
by: Li, Zhaoxing, et al.
Published: (2026)
Towards Adaptive Software Agents for Debugging
by: Majdoub, Yacine, et al.
Published: (2025)
by: Majdoub, Yacine, et al.
Published: (2025)
Are LLMs Better GNN Helpers? Rethinking Robust Graph Learning under Deficiencies with Iterative Refinement
by: Wang, Zhaoyan, et al.
Published: (2025)
by: Wang, Zhaoyan, et al.
Published: (2025)
World Models for Policy Refinement in StarCraft II
by: Zhang, Yixin, et al.
Published: (2026)
by: Zhang, Yixin, et al.
Published: (2026)
Policy Abstraction and Nash Refinement in Tree-Exploiting PSRO
by: Konicki, Christine, et al.
Published: (2025)
by: Konicki, Christine, et al.
Published: (2025)
LLMs for High-Frequency Decision-Making: Normalized Action Reward-Guided Consistency Policy Optimization
by: Zhao, Yang, et al.
Published: (2026)
by: Zhao, Yang, et al.
Published: (2026)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
by: Bai, Yunsheng, et al.
Published: (2025)
by: Bai, Yunsheng, et al.
Published: (2025)
Refining Positive and Toxic Samples for Dual Safety Self-Alignment of LLMs with Minimal Human Interventions
by: Xu, Jingxin, et al.
Published: (2025)
by: Xu, Jingxin, et al.
Published: (2025)
Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness
by: Wan, Hanwen, et al.
Published: (2025)
by: Wan, Hanwen, et al.
Published: (2025)
Can LLMs Assist Expert Elicitation for Probabilistic Causal Modeling?
by: Shaposhnyk, Olha, et al.
Published: (2025)
by: Shaposhnyk, Olha, et al.
Published: (2025)
Similar Items
-
CodeRefine: A Pipeline for Enhancing LLM-Generated Code Implementations of Research Papers
by: Trofimova, Ekaterina, et al.
Published: (2024) -
What Makes Cryptic Crosswords Challenging for LLMs?
by: Sadallah, Abdelrahman, et al.
Published: (2024) -
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
by: Adnan, Muntasir, et al.
Published: (2025) -
Large Language Model Guided Self-Debugging Code Generation
by: Adnan, Muntasir, et al.
Published: (2025) -
Are LLMs Good Cryptic Crossword Solvers?
by: Sadallah, Abdelrahman, et al.
Published: (2024)