Repairing Tool Calls Using Post-tool Execution Reflection and RAG
Fuente:
arXiv
Saved in:
| Main Authors: | Tsay, Jason, Wright, Zidane, Fang, Gaodan, Kate, Kiran, Jha, Saurabh, Rizk, Yara |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Agents go Astray: Course-Correcting SWE Agents with PRMs
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
Live API-Bench: 2500+ Live APIs for Testing Multi-Step Tool Calling
by: Elder, Benjamin, et al.
Published: (2025)
by: Elder, Benjamin, et al.
Published: (2025)
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025)
by: Jain, Kush, et al.
Published: (2025)
Automating Android Build Repair: Bridging the Reasoning-Execution Gap in LLM Agents with Domain-Specific Tools
by: Son, Ha Min, et al.
Published: (2025)
by: Son, Ha Min, et al.
Published: (2025)
Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments
by: Crouse, Maxwell, et al.
Published: (2026)
by: Crouse, Maxwell, et al.
Published: (2026)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025)
by: Stennett, Tyler, et al.
Published: (2025)
OASBuilder: Generating OpenAPI Specifications from Online API Documentation with Large Language Models
by: Lazar, Koren, et al.
Published: (2025)
by: Lazar, Koren, et al.
Published: (2025)
An Executable Benchmarking Suite for Tool-Using Agents
by: Zhong, Zhiqing, et al.
Published: (2026)
by: Zhong, Zhiqing, et al.
Published: (2026)
ASA: Training-Free Representation Engineering for Tool-Calling Agents
by: Wang, Youjin, et al.
Published: (2026)
by: Wang, Youjin, et al.
Published: (2026)
Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action
by: Pujar, Saurabh, et al.
Published: (2025)
by: Pujar, Saurabh, et al.
Published: (2025)
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment
by: Yu, Shasha, et al.
Published: (2026)
by: Yu, Shasha, et al.
Published: (2026)
DynaFix: Iterative Automated Program Repair Driven by Execution-Level Dynamic Information
by: Huang, Zhili, et al.
Published: (2025)
by: Huang, Zhili, et al.
Published: (2025)
JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents
by: Ghoshal, Sandip, et al.
Published: (2026)
by: Ghoshal, Sandip, et al.
Published: (2026)
Verification-Guided Context Optimization for Tool Calling via Hierarchical LLMs-as-Editors
by: Li, Henger, et al.
Published: (2025)
by: Li, Henger, et al.
Published: (2025)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
by: Bhat, Vishvesh, et al.
Published: (2025)
by: Bhat, Vishvesh, et al.
Published: (2025)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
by: Gao, Xingjie, et al.
Published: (2026)
by: Gao, Xingjie, et al.
Published: (2026)
CyberRAG: An Agentic RAG cyber attack classification and reporting tool
by: Blefari, Francesco, et al.
Published: (2025)
by: Blefari, Francesco, et al.
Published: (2025)
Project Prometheus: Bridging the Intent Gap in Agentic Program Repair via Reverse-Engineered Executable Specifications
by: Wang, Yongchao, et al.
Published: (2026)
by: Wang, Yongchao, et al.
Published: (2026)
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
by: Li, Yuanhao, et al.
Published: (2026)
by: Li, Yuanhao, et al.
Published: (2026)
DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells
by: Cherief, Houcine Abdelkader, et al.
Published: (2026)
by: Cherief, Houcine Abdelkader, et al.
Published: (2026)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
by: Gan, Tiantian, et al.
Published: (2025)
by: Gan, Tiantian, et al.
Published: (2025)
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study
by: Ceka, Ira, et al.
Published: (2025)
by: Ceka, Ira, et al.
Published: (2025)
Meta-RAG on Large Codebases Using Code Summarization
by: Tawosi, Vali, et al.
Published: (2025)
by: Tawosi, Vali, et al.
Published: (2025)
Peer-aided Repairer: Empowering Large Language Models to Repair Advanced Student Assignments
by: Zhao, Qianhui, et al.
Published: (2024)
by: Zhao, Qianhui, et al.
Published: (2024)
An Empirically-grounded tool for Automatic Prompt Linting and Repair: A Case Study on Bias, Vulnerability, and Optimization in Developer Prompts
by: Rzig, Dhia Elhaq, et al.
Published: (2025)
by: Rzig, Dhia Elhaq, et al.
Published: (2025)
Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair
by: de-Fitero-Dominguez, David, et al.
Published: (2025)
by: de-Fitero-Dominguez, David, et al.
Published: (2025)
How Good Are LLMs at Processing Tool Outputs?
by: Kate, Kiran, et al.
Published: (2025)
by: Kate, Kiran, et al.
Published: (2025)
Enhancing Automated Program Repair with Solution Design
by: Zhao, Jiuang, et al.
Published: (2024)
by: Zhao, Jiuang, et al.
Published: (2024)
Repair-R1: Better Test Before Repair
by: Hu, Haichuan, et al.
Published: (2025)
by: Hu, Haichuan, et al.
Published: (2025)
Out of style: Misadventures with LLMs and code style transfer
by: Munson, Karl, et al.
Published: (2024)
by: Munson, Karl, et al.
Published: (2024)
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
by: Bouzenia, Islem, et al.
Published: (2024)
by: Bouzenia, Islem, et al.
Published: (2024)
Vul-RAG: Enhancing LLM-based Vulnerability Detection via Knowledge-level RAG
by: Du, Xueying, et al.
Published: (2024)
by: Du, Xueying, et al.
Published: (2024)
RAG Does Not Work for Enterprises
by: Bruckhaus, Tilmann
Published: (2024)
by: Bruckhaus, Tilmann
Published: (2024)
RepoRepair: Leveraging Code Documentation for Repository-Level Automated Program Repair
by: Pan, Zhongqiang, et al.
Published: (2026)
by: Pan, Zhongqiang, et al.
Published: (2026)
Towards Requirements Engineering for RAG Systems
by: Sporsem, Tor, et al.
Published: (2025)
by: Sporsem, Tor, et al.
Published: (2025)
EigenData: A Self-Evolving Multi-Agent Platform for Function-Calling Data Synthesis, Auditing, and Repair
by: Chen, Jiaao, et al.
Published: (2026)
by: Chen, Jiaao, et al.
Published: (2026)
RA-Gen: A Controllable Code Generation Framework Using ReAct for Multi-Agent Task Execution
by: Liu, Aofan, et al.
Published: (2025)
by: Liu, Aofan, et al.
Published: (2025)
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization
by: Orvalho, Pedro, et al.
Published: (2024)
by: Orvalho, Pedro, et al.
Published: (2024)
Agent Lifecycle Toolkit (ALTK): Reusable Middleware Components for Robust AI Agents
by: Wright, Zidane, et al.
Published: (2026)
by: Wright, Zidane, et al.
Published: (2026)
MemRepair: Hierarchical Memory for Agentic Repository-Level Vulnerability Repair
by: Liu, Simiao, et al.
Published: (2026)
by: Liu, Simiao, et al.
Published: (2026)
Similar Items
-
When Agents go Astray: Course-Correcting SWE Agents with PRMs
by: Gandhi, Shubham, et al.
Published: (2025) -
Live API-Bench: 2500+ Live APIs for Testing Multi-Step Tool Calling
by: Elder, Benjamin, et al.
Published: (2025) -
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025) -
Automating Android Build Repair: Bridging the Reasoning-Execution Gap in LLM Agents with Domain-Specific Tools
by: Son, Ha Min, et al.
Published: (2025) -
Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments
by: Crouse, Maxwell, et al.
Published: (2026)