How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval
Fuente:
arXiv
Saved in:
| Main Author: | Ashrafi, Nazmus |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Collaborative LLM Agents for C4 Software Architecture Design Automation
by: Szczepanik, Kamil, et al.
Published: (2025)
by: Szczepanik, Kamil, et al.
Published: (2025)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
by: Wang, Yuchen, et al.
Published: (2026)
by: Wang, Yuchen, et al.
Published: (2026)
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
by: Szczepanik, Kamil, et al.
Published: (2025)
by: Szczepanik, Kamil, et al.
Published: (2025)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
by: Kohl, Jens, et al.
Published: (2024)
by: Kohl, Jens, et al.
Published: (2024)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
by: Trooskens, Geert, et al.
Published: (2026)
by: Trooskens, Geert, et al.
Published: (2026)
Contrastive Learning-Enhanced Large Language Models for Monolith-to-Microservice Decomposition
by: Sellami, Khaled, et al.
Published: (2025)
by: Sellami, Khaled, et al.
Published: (2025)
Can AI Assist in Olympiad Coding
by: Ren, Samuel
Published: (2025)
by: Ren, Samuel
Published: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
by: Cohen, Philip R., et al.
Published: (2023)
by: Cohen, Philip R., et al.
Published: (2023)
Who Writes the Docs in SE 3.0? Agent vs. Human Documentation Pull Requests
by: Yamasaki, Kazuma, et al.
Published: (2026)
by: Yamasaki, Kazuma, et al.
Published: (2026)
Multi-Agent LLM Orchestration Achieves Deterministic, High-Quality Decision Support for Incident Response
by: Drammeh, Philip
Published: (2025)
by: Drammeh, Philip
Published: (2025)
Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures
by: Rombaut, Benjamin
Published: (2026)
by: Rombaut, Benjamin
Published: (2026)
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents
by: Mazaheri, Parsa, et al.
Published: (2026)
by: Mazaheri, Parsa, et al.
Published: (2026)
Beyond Prompt Engineering: Neuro-Symbolic-Causal Architecture for Robust Multi-Objective AI Agents
by: Akarlar, Gokturk Aytug
Published: (2025)
by: Akarlar, Gokturk Aytug
Published: (2025)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
by: Ray, Aninda
Published: (2026)
by: Ray, Aninda
Published: (2026)
AutoBench: Automating LLM Evaluation through Reciprocal Peer Assessment
by: Loi, Dario, et al.
Published: (2025)
by: Loi, Dario, et al.
Published: (2025)
AgentSpawn: Adaptive Multi-Agent Collaboration Through Dynamic Spawning for Long-Horizon Code Generation
by: Costa, Igor
Published: (2026)
by: Costa, Igor
Published: (2026)
Privacy Preserving Multi Agent Path Finding
by: Lehman, Rotem Lev, et al.
Published: (2026)
by: Lehman, Rotem Lev, et al.
Published: (2026)
A Tale of Two Systems: Characterizing Architectural Complexity on Machine Learning-Enabled Systems
by: Ferreira, Renato Cordeiro
Published: (2025)
by: Ferreira, Renato Cordeiro
Published: (2025)
BACE: LLM-based Code Generation through Bayesian Anchored Co-Evolution of Code and Test Populations
by: Silva, Kaushitha, et al.
Published: (2026)
by: Silva, Kaushitha, et al.
Published: (2026)
A Metrics-Oriented Architectural Model to Characterize Complexity on Machine Learning-Enabled Systems
by: Ferreira, Renato Cordeiro
Published: (2025)
by: Ferreira, Renato Cordeiro
Published: (2025)
SPIRA: Building an Intelligent System for Respiratory Insufficiency Detection
by: Ferreira, Renato Cordeiro, et al.
Published: (2025)
by: Ferreira, Renato Cordeiro, et al.
Published: (2025)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
by: Alshaikh, Moaath, et al.
Published: (2026)
by: Alshaikh, Moaath, et al.
Published: (2026)
Applying Cognitive Design Patterns to General LLM Agents
by: Wray, Robert E., et al.
Published: (2025)
by: Wray, Robert E., et al.
Published: (2025)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
by: Sidik, Bronislav, et al.
Published: (2026)
by: Sidik, Bronislav, et al.
Published: (2026)
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
by: Wang, Zixu, et al.
Published: (2026)
by: Wang, Zixu, et al.
Published: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
by: Kohl, Jens, et al.
Published: (2026)
by: Kohl, Jens, et al.
Published: (2026)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
by: Wu, Yalun, et al.
Published: (2026)
by: Wu, Yalun, et al.
Published: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
by: Annapureddy, Sasank
Published: (2026)
by: Annapureddy, Sasank
Published: (2026)
SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair
by: Dinu, Ion George, et al.
Published: (2026)
by: Dinu, Ion George, et al.
Published: (2026)
Making a Pipeline Production-Ready: Challenges and Lessons Learned in the Healthcare Domain
by: Lawand, Daniel Angelo Esteves, et al.
Published: (2025)
by: Lawand, Daniel Angelo Esteves, et al.
Published: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
by: Palacios, Diego Cabezas
Published: (2026)
by: Palacios, Diego Cabezas
Published: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
by: Costa, Rimom
Published: (2025)
by: Costa, Rimom
Published: (2025)
Multi-Agent Design Assistant for the Simulation of Inertial Fusion Energy
by: Shachar, Meir H., et al.
Published: (2025)
by: Shachar, Meir H., et al.
Published: (2025)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
by: Bradbury, Jeremy S., et al.
Published: (2024)
by: Bradbury, Jeremy S., et al.
Published: (2024)
Procedural Game Level Design with Deep Reinforcement Learning
by: Özkan, Miraç Buğra
Published: (2025)
by: Özkan, Miraç Buğra
Published: (2025)
MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization
by: Tanjim, Md Mehrab, et al.
Published: (2026)
by: Tanjim, Md Mehrab, et al.
Published: (2026)
Deployment-Time Reliability of Learned Robot Policies
by: Agia, Christopher
Published: (2026)
by: Agia, Christopher
Published: (2026)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
by: Ueda, Keisuke, et al.
Published: (2025)
by: Ueda, Keisuke, et al.
Published: (2025)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
by: Chang, Jiale, et al.
Published: (2026)
by: Chang, Jiale, et al.
Published: (2026)
A Language for Describing Agentic LLM Contexts
by: Pelc, Noga Peleg, et al.
Published: (2026)
by: Pelc, Noga Peleg, et al.
Published: (2026)
Similar Items
-
Collaborative LLM Agents for C4 Software Architecture Design Automation
by: Szczepanik, Kamil, et al.
Published: (2025) -
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
by: Wang, Yuchen, et al.
Published: (2026) -
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
by: Szczepanik, Kamil, et al.
Published: (2025) -
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
by: Kohl, Jens, et al.
Published: (2024) -
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
by: Trooskens, Geert, et al.
Published: (2026)