Butterfly Effects in Toolchains: A Comprehensive Analysis of Failed Parameter Filling in LLM Tool-Agent Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Qian, Huang, Yuekai, Jiang, Ziyou, Chang, Zhiyuan, Zheng, Yujia, Li, Tianhao, Li, Mingyang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VulRTex: A Reasoning-Guided Approach to Identify Vulnerabilities from Rich-Text Issue Report
by: Jiang, Ziyou, et al.
Published: (2025)
by: Jiang, Ziyou, et al.
Published: (2025)
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
by: Islam, Niful, et al.
Published: (2026)
by: Islam, Niful, et al.
Published: (2026)
Integrating Static Code Analysis Toolchains
by: Kern, Matthias, et al.
Published: (2024)
by: Kern, Matthias, et al.
Published: (2024)
Why Does My Transaction Fail? A First Look at Failed Transactions on the Solana Blockchain
by: Zheng, Xiaoye, et al.
Published: (2025)
by: Zheng, Xiaoye, et al.
Published: (2025)
The Evolution of Tool Use in LLM Agents: From Single-Tool Call to Multi-Tool Orchestration
by: Xu, Haoyuan, et al.
Published: (2026)
by: Xu, Haoyuan, et al.
Published: (2026)
Emerging from Ground: Addressing Intent Deviation in Tool-Using Agents via Deriving Real Calls into Virtual Trajectories
by: Xiong, Qian, et al.
Published: (2026)
by: Xiong, Qian, et al.
Published: (2026)
From Logic to Toolchains: An Empirical Study of Bugs in the TypeScript Ecosystem
by: Tang, TianYi, et al.
Published: (2026)
by: Tang, TianYi, et al.
Published: (2026)
Empirical Investigation of Quantum Computing Toolchains and Algorithms : Mining Stack Overflow Repository
by: Sabzevari, Maryam Tavassoli, et al.
Published: (2026)
by: Sabzevari, Maryam Tavassoli, et al.
Published: (2026)
Toolchain for Faster Iterations in Quantum Software Development
by: Kinanen, Otso, et al.
Published: (2025)
by: Kinanen, Otso, et al.
Published: (2025)
M, Toolchain and Language for Reusable Model Compilation
by: Trinh, Hiep Hong, et al.
Published: (2025)
by: Trinh, Hiep Hong, et al.
Published: (2025)
CompileAgent: Automated Real-World Repo-Level Compilation with Tool-Integrated LLM-based Agent System
by: Hu, Li, et al.
Published: (2025)
by: Hu, Li, et al.
Published: (2025)
WebSuite: Systematically Evaluating Why Web Agents Fail
by: Li, Eric, et al.
Published: (2024)
by: Li, Eric, et al.
Published: (2024)
Towards Verifiably Safe Tool Use for LLM Agents
by: Doshi, Aarya, et al.
Published: (2026)
by: Doshi, Aarya, et al.
Published: (2026)
RepoTransAgent: Multi-Agent LLM Framework for Repository-Aware Code Translation
by: Guan, Ziqi, et al.
Published: (2025)
by: Guan, Ziqi, et al.
Published: (2025)
MCP-Zero: Active Tool Discovery for Autonomous LLM Agents
by: Fei, Xiang, et al.
Published: (2025)
by: Fei, Xiang, et al.
Published: (2025)
Prompt-with-Me: in-IDE Structured Prompt Management for LLM-Driven Software Engineering
by: Li, Ziyou, et al.
Published: (2025)
by: Li, Ziyou, et al.
Published: (2025)
LLM-based Vulnerability Detection at Project Scale: An Empirical Study
by: Li, Fengjie, et al.
Published: (2026)
by: Li, Fengjie, et al.
Published: (2026)
OptiLoop: Coordination-in-the-Loop Verification and Repair for LLM-Generated Optimization Agents
by: Xu, Yujia, et al.
Published: (2026)
by: Xu, Yujia, et al.
Published: (2026)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
Liberal Entity Matching as a Compound AI Toolchain
by: Fu, Silvery D., et al.
Published: (2024)
by: Fu, Silvery D., et al.
Published: (2024)
A Comprehensive Evaluation of Parameter-Efficient Fine-Tuning on Code Smell Detection
by: Zhang, Beiqi, et al.
Published: (2024)
by: Zhang, Beiqi, et al.
Published: (2024)
Microservices and Real-Time Processing in Retail IT: A Review of Open-Source Toolchains and Deployment Strategies
by: Vashisht, Aaditaa, et al.
Published: (2025)
by: Vashisht, Aaditaa, et al.
Published: (2025)
Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive Filtering
by: Xiong, Yunpeng, et al.
Published: (2026)
by: Xiong, Yunpeng, et al.
Published: (2026)
Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks
by: Lu, Ruofan, et al.
Published: (2025)
by: Lu, Ruofan, et al.
Published: (2025)
SkillCraft: Can LLM Agents Learn to Use Tools Skillfully?
by: Chen, Shiqi, et al.
Published: (2026)
by: Chen, Shiqi, et al.
Published: (2026)
A Comprehensive Study on Static Application Security Testing (SAST) Tools for Android
by: Zhu, Jingyun, et al.
Published: (2024)
by: Zhu, Jingyun, et al.
Published: (2024)
ChainFuzzer: Greybox Fuzzing for Workflow-Level Multi-Tool Vulnerabilities in LLM Agents
by: Wu, Jiangrong, et al.
Published: (2026)
by: Wu, Jiangrong, et al.
Published: (2026)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
by: Zhang, Kechi, et al.
Published: (2024)
by: Zhang, Kechi, et al.
Published: (2024)
DADL: A Declarative Description Language for Enterprise Tool Libraries in LLM Agent Systems
by: Dunkel, Axel
Published: (2026)
by: Dunkel, Axel
Published: (2026)
A High-level Synthesis Toolchain for the Julia Language
by: Short, Benedict, et al.
Published: (2025)
by: Short, Benedict, et al.
Published: (2025)
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
by: Li, Yuanyang, et al.
Published: (2026)
by: Li, Yuanyang, et al.
Published: (2026)
SoK: Comprehensive Analysis of Rug Pull Causes, Datasets, and Detection Tools in DeFi
by: Sun, Dianxiang, et al.
Published: (2024)
by: Sun, Dianxiang, et al.
Published: (2024)
HarnessAgent: Scaling Automatic Fuzzing Harness Construction with Tool-Augmented LLM Pipelines
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
SLICEMATE: Accurate and Scalable Static Program Slicing via LLM-Powered Agents
by: Chang, Jianming, et al.
Published: (2025)
by: Chang, Jianming, et al.
Published: (2025)
Knowledge-Guided Multi-Agent Framework for Application-Level Software Code Generation
by: Xiong, Qian, et al.
Published: (2025)
by: Xiong, Qian, et al.
Published: (2025)
ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering
by: Liu, Marianne Menglin, et al.
Published: (2025)
by: Liu, Marianne Menglin, et al.
Published: (2025)
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
by: Ehsani, Ramtin, et al.
Published: (2026)
by: Ehsani, Ramtin, et al.
Published: (2026)
Evaluating LLM Agents on Automated Software Analysis Tasks
by: Bouzenia, Islem, et al.
Published: (2026)
by: Bouzenia, Islem, et al.
Published: (2026)
AgentRaft: Automated Detection of Data Over-Exposure in LLM Agents
by: Lin, Yixi, et al.
Published: (2026)
by: Lin, Yixi, et al.
Published: (2026)
Bridging Safety and Security in Complex Systems: A Model-Based Approach with SAFT-GT Toolchain
by: Pekaric, Irdin, et al.
Published: (2026)
by: Pekaric, Irdin, et al.
Published: (2026)
Similar Items
-
VulRTex: A Reasoning-Guided Approach to Identify Vulnerabilities from Rich-Text Issue Report
by: Jiang, Ziyou, et al.
Published: (2025) -
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
by: Islam, Niful, et al.
Published: (2026) -
Integrating Static Code Analysis Toolchains
by: Kern, Matthias, et al.
Published: (2024) -
Why Does My Transaction Fail? A First Look at Failed Transactions on the Solana Blockchain
by: Zheng, Xiaoye, et al.
Published: (2025) -
The Evolution of Tool Use in LLM Agents: From Single-Tool Call to Multi-Tool Orchestration
by: Xu, Haoyuan, et al.
Published: (2026)