An Empirical Study of Bugs in Modern LLM Agent Frameworks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Xinxue, Wu, Jiacong, Zhang, Xiaoyu, Li, Tianlin, Mu, Yanzhou, Zhai, Juan, Shen, Chao, Fang, Chunrong, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding LLM-Centric Challenges for Deep Learning Frameworks: An Empirical Analysis
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
Deep Learning Framework Testing via Heuristic Guidance Based on Multiple Model Measurements
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
CITADEL: Context Similarity Based Deep Learning Framework Bug Finding
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
Improving Deep Learning Framework Testing with Model-Level Metamorphic Testing
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
DevMuT: Testing Deep Learning Framework via Developer Expertise-Based Mutation
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
Deep Learning Framework Testing via Model Mutation: How Far Are We?
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
von: Meng, Xiangxin, et al.
Veröffentlicht: (2024)
von: Meng, Xiangxin, et al.
Veröffentlicht: (2024)
Dissecting Bug Triggers and Failure Modes in Modern Agentic Frameworks: An Empirical Study
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
Scalpel: Automotive Deep Learning Framework Testing via Assembling Model Components
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
SelfHeal: Empirical Fix Pattern Analysis and Bug Repair in LLM Agents
von: Islam, Niful, et al.
Veröffentlicht: (2026)
von: Islam, Niful, et al.
Veröffentlicht: (2026)
GPU Temperature Simulation-Based Testing for In-Vehicle Deep Learning Frameworks
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
von: Zou, Yinglong, et al.
Veröffentlicht: (2025)
Mutation-Based Deep Learning Framework Testing Method in JavaScript Environment
von: Zou, Yinglong, et al.
Veröffentlicht: (2024)
von: Zou, Yinglong, et al.
Veröffentlicht: (2024)
Rethinking Technology Stack Selection with AI Coding Proficiency
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
An Empirical Study of Refactoring Engine Bugs
von: Wang, Haibo, et al.
Veröffentlicht: (2024)
von: Wang, Haibo, et al.
Veröffentlicht: (2024)
An Empirical Study on Noisy Label Learning for Program Understanding
von: Wang, Wenhan, et al.
Veröffentlicht: (2023)
von: Wang, Wenhan, et al.
Veröffentlicht: (2023)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
von: Guo, Liwei, et al.
Veröffentlicht: (2025)
von: Guo, Liwei, et al.
Veröffentlicht: (2025)
Bug Priority Change: An Empirical Study on Apache Projects
von: Li, Zengyang, et al.
Veröffentlicht: (2024)
von: Li, Zengyang, et al.
Veröffentlicht: (2024)
Empirical Research on Utilizing LLM-based Agents for Automated Bug Fixing via LangGraph
von: Wang, Jialin, et al.
Veröffentlicht: (2025)
von: Wang, Jialin, et al.
Veröffentlicht: (2025)
Generate Realistic Test Scenes for V2X Communication Systems
von: Guo, An, et al.
Veröffentlicht: (2025)
von: Guo, An, et al.
Veröffentlicht: (2025)
Exploring the Power of Diffusion Large Language Models for Software Engineering: An Empirical Investigation
von: Zhang, Jingyao, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyao, et al.
Veröffentlicht: (2025)
Characterizing Bugs in Login Processes of Android Applications: An Empirical Study
von: Zhou, Zixu, et al.
Veröffentlicht: (2025)
von: Zhou, Zixu, et al.
Veröffentlicht: (2025)
SGAgent: Suggestion-Guided LLM-Based Multi-Agent Framework for Repository-Level Software Repair
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
An Empirical Study on the Capability of LLMs in Decomposing Bug Reports
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2025)
Exploring the Jupyter Ecosystem: An Empirical Study of Bugs and Vulnerabilities
von: Jiang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Wenyuan, et al.
Veröffentlicht: (2025)
An Empirical Study on the Classification of Bug Reports with Machine Learning
von: Andrade, Renato, et al.
Veröffentlicht: (2025)
von: Andrade, Renato, et al.
Veröffentlicht: (2025)
An Empirical Study of Interaction Bugs in ROS-based Software
von: Chen, Zhixiang, et al.
Veröffentlicht: (2025)
von: Chen, Zhixiang, et al.
Veröffentlicht: (2025)
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
von: Islam, Niful, et al.
Veröffentlicht: (2026)
von: Islam, Niful, et al.
Veröffentlicht: (2026)
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
An Empirical Study on the Characteristics of Database Access Bugs in Java Applications
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing
von: Shang, Ye, et al.
Veröffentlicht: (2024)
von: Shang, Ye, et al.
Veröffentlicht: (2024)
Understanding Bug-Reproducing Tests: A First Empirical Study
von: Hora, Andre, et al.
Veröffentlicht: (2026)
von: Hora, Andre, et al.
Veröffentlicht: (2026)
An Empirical Study on Leveraging Images in Automated Bug Report Reproduction
von: Wang, Dingbang, et al.
Veröffentlicht: (2025)
von: Wang, Dingbang, et al.
Veröffentlicht: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
A Comprehensive Study of Bugs in Modern Distributed Deep Learning Systems
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2025)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2025)
An Empirical Evaluation of Modern MLOps Frameworks
von: Marcos-Mercadé, Jon, et al.
Veröffentlicht: (2026)
von: Marcos-Mercadé, Jon, et al.
Veröffentlicht: (2026)
Train in Vain: Functionality-Preserving Poisoning to Prevent Unauthorized Use of Code Datasets
von: Xiao, Yuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yuan, et al.
Veröffentlicht: (2026)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
From Logic to Toolchains: An Empirical Study of Bugs in the TypeScript Ecosystem
von: Tang, TianYi, et al.
Veröffentlicht: (2026)
von: Tang, TianYi, et al.
Veröffentlicht: (2026)
Does Programming Language Matter? An Empirical Study of Fuzzing Bug Detection
von: Shirai, Tatsuya, et al.
Veröffentlicht: (2026)
von: Shirai, Tatsuya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Understanding LLM-Centric Challenges for Deep Learning Frameworks: An Empirical Analysis
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025) -
Deep Learning Framework Testing via Heuristic Guidance Based on Multiple Model Measurements
von: Zou, Yinglong, et al.
Veröffentlicht: (2025) -
CITADEL: Context Similarity Based Deep Learning Framework Bug Finding
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024) -
Improving Deep Learning Framework Testing with Model-Level Metamorphic Testing
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025) -
DevMuT: Testing Deep Learning Framework via Developer Expertise-Based Mutation
von: Mu, Yanzhou, et al.
Veröffentlicht: (2025)