A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Daniel, Upadhyay, Krishna, Chhetri, Vinaik, Siddique, A. B., Farooq, Umar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
by: Chhetri, Vinaik, et al.
Published: (2025)
by: Chhetri, Vinaik, et al.
Published: (2025)
Analyzing the Evolution and Maintenance of Quantum Software Repositories
by: Upadhyay, Krishna, et al.
Published: (2025)
by: Upadhyay, Krishna, et al.
Published: (2025)
Understanding Robustness of Model Editing in Code LLMs
by: Chhetri, Vinaik, et al.
Published: (2025)
by: Chhetri, Vinaik, et al.
Published: (2025)
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
by: Fakorede, Moshood A., et al.
Published: (2026)
by: Fakorede, Moshood A., et al.
Published: (2026)
Looking into Black Box Code Language Models
by: Haider, Muhammad Umair, et al.
Published: (2024)
by: Haider, Muhammad Umair, et al.
Published: (2024)
Assessing and Enhancing Quantum Readiness in Mobile Apps
by: Strauss, Joseph, et al.
Published: (2025)
by: Strauss, Joseph, et al.
Published: (2025)
Understanding Bugs in Quantum Simulators: An Empirical Study
by: Upadhyay, Krishna, et al.
Published: (2026)
by: Upadhyay, Krishna, et al.
Published: (2026)
An Empirical Study of Agent Developer Practices in AI Agent Frameworks
by: Wang, Yanlin, et al.
Published: (2025)
by: Wang, Yanlin, et al.
Published: (2025)
Skilled AI Agents for Embedded and IoT Systems Development
by: Li, Yiming, et al.
Published: (2026)
by: Li, Yiming, et al.
Published: (2026)
WhatsCode: Large-Scale GenAI Deployment for Developer Efficiency at WhatsApp
by: Mao, Ke, et al.
Published: (2025)
by: Mao, Ke, et al.
Published: (2025)
Facilitating Trustworthy Human-Agent Collaboration in LLM-based Multi-Agent System oriented Software Engineering
by: Ronanki, Krishna
Published: (2025)
by: Ronanki, Krishna
Published: (2025)
AgentMesh: A Cooperative Multi-Agent Generative AI Framework for Software Development Automation
by: Khanzadeh, Sourena
Published: (2025)
by: Khanzadeh, Sourena
Published: (2025)
Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development
by: Stalnaker, Trevor, et al.
Published: (2024)
by: Stalnaker, Trevor, et al.
Published: (2024)
EmbedAgent: Benchmarking Large Language Models in Embedded System Development
by: Xu, Ruiyang, et al.
Published: (2025)
by: Xu, Ruiyang, et al.
Published: (2025)
Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects
by: Kashif, Syed Mohammad, et al.
Published: (2026)
by: Kashif, Syed Mohammad, et al.
Published: (2026)
FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning
by: Ding, Haoran, et al.
Published: (2026)
by: Ding, Haoran, et al.
Published: (2026)
A Case Study on AI Engineering Practices: Developing an Autonomous Stock Trading System
by: Grote, Marcel, et al.
Published: (2023)
by: Grote, Marcel, et al.
Published: (2023)
Agentsway -- Software Development Methodology for AI Agents-based Teams
by: Bandara, Eranga, et al.
Published: (2025)
by: Bandara, Eranga, et al.
Published: (2025)
Can Agents Fix Agent Issues?
by: Rahardja, Alfin Wijaya, et al.
Published: (2025)
by: Rahardja, Alfin Wijaya, et al.
Published: (2025)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
by: Lu, Pengrui, et al.
Published: (2026)
by: Lu, Pengrui, et al.
Published: (2026)
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
by: Wang, Kaixin, et al.
Published: (2025)
by: Wang, Kaixin, et al.
Published: (2025)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Code Researcher: Deep Research Agent for Large Systems Code and Commit History
by: Singh, Ramneet, et al.
Published: (2025)
by: Singh, Ramneet, et al.
Published: (2025)
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
by: Zhu, Yuecai, et al.
Published: (2026)
by: Zhu, Yuecai, et al.
Published: (2026)
On the Adoption of AI Coding Agents in Open-source Android and iOS Development
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
RE-centric Recommendations for the Development of Trustworthy(er) Autonomous Systems
by: Ronanki, Krishna, et al.
Published: (2023)
by: Ronanki, Krishna, et al.
Published: (2023)
GenAI-powered Multi-Agent Paradigm for Smart Urban Mobility: Opportunities and Challenges for Integrating Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) with Intelligent Transportation Systems
by: Xu, Haowen, et al.
Published: (2024)
by: Xu, Haowen, et al.
Published: (2024)
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
by: Li, Yuanyang, et al.
Published: (2026)
by: Li, Yuanyang, et al.
Published: (2026)
Understanding and Bridging the Planner-Coder Gap: A Systematic Study on the Robustness of Multi-Agent Systems for Code Generation
by: Lyu, Zongyi, et al.
Published: (2025)
by: Lyu, Zongyi, et al.
Published: (2025)
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents
by: Shen, Haiyang, et al.
Published: (2024)
by: Shen, Haiyang, et al.
Published: (2024)
The AI Agent Index
by: Casper, Stephen, et al.
Published: (2025)
by: Casper, Stephen, et al.
Published: (2025)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
by: Trae Research Team, et al.
Published: (2025)
by: Trae Research Team, et al.
Published: (2025)
Improving Performance of Commercially Available AI Products in a Multi-Agent Configuration
by: Hymel, Cory, et al.
Published: (2024)
by: Hymel, Cory, et al.
Published: (2024)
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
by: Tao, Wei, et al.
Published: (2024)
by: Tao, Wei, et al.
Published: (2024)
Scaling Coding Agents via Atomic Skills
by: Ma, Yingwei, et al.
Published: (2026)
by: Ma, Yingwei, et al.
Published: (2026)
The AI-Native Large-Scale Agile Software Development Manifesto
by: Britto, Ricardo, et al.
Published: (2026)
by: Britto, Ricardo, et al.
Published: (2026)
On Using Agent-based Modeling and Simulation for Studying Blockchain Systems
by: Gürcan, Önder
Published: (2024)
by: Gürcan, Önder
Published: (2024)
AgentGuard: Runtime Verification of AI Agents
by: Koohestani, Roham
Published: (2025)
by: Koohestani, Roham
Published: (2025)
Efficient Failure Management for Multi-Agent Systems with Reasoning Trace Representation
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
AgentStepper: Interactive Debugging of Software Development Agents
by: Hutter, Robert, et al.
Published: (2026)
by: Hutter, Robert, et al.
Published: (2026)
Similar Items
-
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
by: Chhetri, Vinaik, et al.
Published: (2025) -
Analyzing the Evolution and Maintenance of Quantum Software Repositories
by: Upadhyay, Krishna, et al.
Published: (2025) -
Understanding Robustness of Model Editing in Code LLMs
by: Chhetri, Vinaik, et al.
Published: (2025) -
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
by: Fakorede, Moshood A., et al.
Published: (2026) -
Looking into Black Box Code Language Models
by: Haider, Muhammad Umair, et al.
Published: (2024)