Investigating Tool-Memory Conflicts in Tool-Augmented LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Jiali, Pan, Rui, Amiri, Hadi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tool Unlearning for Tool-Augmented LLMs
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
ParaTool: Shifting Tool Representations from Context to Parameters
von: Yu, Zekai, et al.
Veröffentlicht: (2026)
von: Yu, Zekai, et al.
Veröffentlicht: (2026)
Solver-Aided Verification of Policy Compliance in Tool-Augmented LLM Agents
von: Winston, Cailin, et al.
Veröffentlicht: (2026)
von: Winston, Cailin, et al.
Veröffentlicht: (2026)
ToolFuzz -- Automated Agent Tool Testing
von: Milev, Ivan, et al.
Veröffentlicht: (2025)
von: Milev, Ivan, et al.
Veröffentlicht: (2025)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
Verification-Guided Context Optimization for Tool Calling via Hierarchical LLMs-as-Editors
von: Li, Henger, et al.
Veröffentlicht: (2025)
von: Li, Henger, et al.
Veröffentlicht: (2025)
CodeWatcher: IDE Telemetry Data Extraction Tool for Understanding Coding Interactions with LLMs
von: Basha, Manaal, et al.
Veröffentlicht: (2025)
von: Basha, Manaal, et al.
Veröffentlicht: (2025)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
von: Gan, Tiantian, et al.
Veröffentlicht: (2025)
von: Gan, Tiantian, et al.
Veröffentlicht: (2025)
ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
von: Li, Dawei, et al.
Veröffentlicht: (2026)
von: Li, Dawei, et al.
Veröffentlicht: (2026)
EvolveTool-Bench: Evaluating the Quality of LLM-Generated Tool Libraries as Software Artifacts
von: Kaliyev, Alibek T., et al.
Veröffentlicht: (2026)
von: Kaliyev, Alibek T., et al.
Veröffentlicht: (2026)
ToolMisuseBench: An Offline Deterministic Benchmark for Tool Misuse and Recovery in Agentic Systems
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
von: Stennett, Tyler, et al.
Veröffentlicht: (2025)
von: Stennett, Tyler, et al.
Veröffentlicht: (2025)
CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision
von: Lu, Yifei, et al.
Veröffentlicht: (2025)
von: Lu, Yifei, et al.
Veröffentlicht: (2025)
ASA: Training-Free Representation Engineering for Tool-Calling Agents
von: Wang, Youjin, et al.
Veröffentlicht: (2026)
von: Wang, Youjin, et al.
Veröffentlicht: (2026)
The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
Trajectory Supervision for Continual Tool-Use Learning in LLMs
von: Reddy, Vishnu Vardhan, et al.
Veröffentlicht: (2026)
von: Reddy, Vishnu Vardhan, et al.
Veröffentlicht: (2026)
Semantic Tool Discovery for Large Language Models: A Vector-Based Approach to MCP Tool Selection
von: Mudunuri, Sarat, et al.
Veröffentlicht: (2026)
von: Mudunuri, Sarat, et al.
Veröffentlicht: (2026)
Breaking the Illusion of Identity in LLM Tooling
von: Miller, Marek
Veröffentlicht: (2026)
von: Miller, Marek
Veröffentlicht: (2026)
Schema First Tool APIs for LLM Agents: A Controlled Study of Tool Misuse, Recovery, and Budgeted Performance
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
von: Sigdel, Akshey, et al.
Veröffentlicht: (2026)
User Centric Evaluation of Code Generation Tools
von: Miah, Tanha, et al.
Veröffentlicht: (2024)
von: Miah, Tanha, et al.
Veröffentlicht: (2024)
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
von: Bartlett, Antony, et al.
Veröffentlicht: (2025)
von: Bartlett, Antony, et al.
Veröffentlicht: (2025)
Do Generative AI Tools Ensure Green Code? An Investigative Study
von: Sikand, Samarth, et al.
Veröffentlicht: (2025)
von: Sikand, Samarth, et al.
Veröffentlicht: (2025)
LLMs Integration in Software Engineering Team Projects: Roles, Impact, and a Pedagogical Design Space for AI Tools in Computing Education
von: Kharrufa, Ahmed, et al.
Veröffentlicht: (2024)
von: Kharrufa, Ahmed, et al.
Veröffentlicht: (2024)
VQA support to Arabic Language Learning Educational Tool
von: Delassi, Khaled Bachir, et al.
Veröffentlicht: (2025)
von: Delassi, Khaled Bachir, et al.
Veröffentlicht: (2025)
Tool-integrated Reinforcement Learning for Repo Deep Search
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells
von: Cherief, Houcine Abdelkader, et al.
Veröffentlicht: (2026)
von: Cherief, Houcine Abdelkader, et al.
Veröffentlicht: (2026)
ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs
von: Ding, Peng, et al.
Veröffentlicht: (2025)
von: Ding, Peng, et al.
Veröffentlicht: (2025)
Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents
von: Kovács, Ádám
Veröffentlicht: (2026)
von: Kovács, Ádám
Veröffentlicht: (2026)
Applying an Agentic Coding Tool for Improving Published Algorithm Implementations
von: Suwannik, Worasait
Veröffentlicht: (2026)
von: Suwannik, Worasait
Veröffentlicht: (2026)
ReXCL: A Tool for Requirement Document Extraction and Classification
von: Bhattacharya, Paheli, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Paheli, et al.
Veröffentlicht: (2025)
AISysRev -- LLM-based Tool for Title-abstract Screening
von: Huotala, Aleksi, et al.
Veröffentlicht: (2025)
von: Huotala, Aleksi, et al.
Veröffentlicht: (2025)
MCP-Zero: Active Tool Discovery for Autonomous LLM Agents
von: Fei, Xiang, et al.
Veröffentlicht: (2025)
von: Fei, Xiang, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Test Generation: How Far Are We?
von: Shin, Jiho, et al.
Veröffentlicht: (2024)
von: Shin, Jiho, et al.
Veröffentlicht: (2024)
Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
SynthTools: A Framework for Scaling Synthetic Tools for Agent Development
von: Castellani, Tommaso, et al.
Veröffentlicht: (2025)
von: Castellani, Tommaso, et al.
Veröffentlicht: (2025)
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
von: Chen, Aili, et al.
Veröffentlicht: (2026)
von: Chen, Aili, et al.
Veröffentlicht: (2026)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
von: Agarwal, Prerna, et al.
Veröffentlicht: (2025)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
von: Tsay, Jason, et al.
Veröffentlicht: (2025)
von: Tsay, Jason, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tool Unlearning for Tool-Augmented LLMs
von: Cheng, Jiali, et al.
Veröffentlicht: (2025) -
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
von: Kokane, Shirley, et al.
Veröffentlicht: (2024) -
ParaTool: Shifting Tool Representations from Context to Parameters
von: Yu, Zekai, et al.
Veröffentlicht: (2026) -
Solver-Aided Verification of Policy Compliance in Tool-Augmented LLM Agents
von: Winston, Cailin, et al.
Veröffentlicht: (2026) -
ToolFuzz -- Automated Agent Tool Testing
von: Milev, Ivan, et al.
Veröffentlicht: (2025)