Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuzhe, Liu, Feiran, Shan, Yi, Huang, Xinyi, Yang, Xin, Zhu, Yueqi, Cheng, Xuxin, Liu, Cao, Zeng, Ke, Zhang, Terry Jingchen, Jiang, Wenyuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpecBench: Evaluating Specification-Level Reasoning for Software Engineering LLM Agents
by: Hamblin, Grant, et al.
Published: (2026)
by: Hamblin, Grant, et al.
Published: (2026)
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
by: Zhong, Shanshan, et al.
Published: (2026)
by: Zhong, Shanshan, et al.
Published: (2026)
CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs
by: Zou, Chelsea, et al.
Published: (2026)
by: Zou, Chelsea, et al.
Published: (2026)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026)
by: Yang, Shan, et al.
Published: (2026)
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models
by: Agashe, Saaket, et al.
Published: (2023)
by: Agashe, Saaket, et al.
Published: (2023)
IndoorR2X: Indoor Robot-to-Everything Coordination with LLM-Driven Planning
by: Yang, Fan, et al.
Published: (2026)
by: Yang, Fan, et al.
Published: (2026)
MLC-Agent: Cognitive Model based on Memory-Learning Collaboration in LLM Empowered Agent Simulation Environment
by: Zhang, Ming, et al.
Published: (2025)
by: Zhang, Ming, et al.
Published: (2025)
TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination
by: Xie, Yi, et al.
Published: (2026)
by: Xie, Yi, et al.
Published: (2026)
OpenCity: A Scalable Platform to Simulate Urban Activities with Massive LLM Agents
by: Yan, Yuwei, et al.
Published: (2024)
by: Yan, Yuwei, et al.
Published: (2024)
MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
by: Jiang, Yixing, et al.
Published: (2025)
by: Jiang, Yixing, et al.
Published: (2025)
AgentNet: Decentralized Evolutionary Coordination for LLM-based Multi-Agent Systems
by: Yang, Yingxuan, et al.
Published: (2025)
by: Yang, Yingxuan, et al.
Published: (2025)
Strategic Communication and Language Bias in Multi-Agent LLM Coordination
by: Buscemi, Alessio, et al.
Published: (2025)
by: Buscemi, Alessio, et al.
Published: (2025)
StackPilot: Autonomous Function Agents for Scalable and Environment-Free Code Execution
by: Zhao, Xinkui, et al.
Published: (2025)
by: Zhao, Xinkui, et al.
Published: (2025)
Distributed online constrained convex optimization with event-triggered communication
by: Zhang, Kunpeng, et al.
Published: (2023)
by: Zhang, Kunpeng, et al.
Published: (2023)
EconAI: Dynamic Persona Evolution and Memory-Aware Agents in Evolving Economic Environments
by: Liu, Annie, et al.
Published: (2026)
by: Liu, Annie, et al.
Published: (2026)
Distributed Task Allocation for Multi-Agent Systems: A Submodular Optimization Approach
by: Liu, Jing, et al.
Published: (2024)
by: Liu, Jing, et al.
Published: (2024)
DAO-Agent: Zero Knowledge-Verified Incentives for Decentralized Multi-Agent Coordination
by: Xia, Yihan, et al.
Published: (2025)
by: Xia, Yihan, et al.
Published: (2025)
Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants
by: Nathani, Deepak, et al.
Published: (2026)
by: Nathani, Deepak, et al.
Published: (2026)
HYGMA: Hypergraph Coordination Networks with Dynamic Grouping for Multi-Agent Reinforcement Learning
by: Liu, Chiqiang, et al.
Published: (2025)
by: Liu, Chiqiang, et al.
Published: (2025)
ClimateAgents: A Multi-Agent Research Assistant for Social-Climate Dynamics Analysis
by: Shan, Shan
Published: (2026)
by: Shan, Shan
Published: (2026)
WMAS: A Multi-Agent System Towards Intelligent and Customized Wireless Networks
by: Peng, Jingchen, et al.
Published: (2025)
by: Peng, Jingchen, et al.
Published: (2025)
Emergent Coordinated Behaviors in Networked LLM Agents: Modeling the Strategic Dynamics of Information Operations
by: Orlando, Gian Marco, et al.
Published: (2025)
by: Orlando, Gian Marco, et al.
Published: (2025)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
by: Zhu, Kunlun, et al.
Published: (2025)
by: Zhu, Kunlun, et al.
Published: (2025)
Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents
by: Zhang, Yuxin, et al.
Published: (2026)
by: Zhang, Yuxin, et al.
Published: (2026)
PRO-SPECT: Probabilistically Safe Scalable Planning for Energy-Aware Coordinated UAV-UGV Teams in Stochastic Environments
by: Fowler, Roger, et al.
Published: (2026)
by: Fowler, Roger, et al.
Published: (2026)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
by: Cobben, Pepijn, et al.
Published: (2026)
by: Cobben, Pepijn, et al.
Published: (2026)
GameChat: Multi-LLM Dialogue for Safe, Agile, and Socially Optimal Multi-Agent Navigation in Constrained Environments
by: Mahadevan, Vagul, et al.
Published: (2025)
by: Mahadevan, Vagul, et al.
Published: (2025)
LLM-ABM for Transportation: Assessing the Potential of LLM Agents in System Analysis
by: Liu, Tianming, et al.
Published: (2025)
by: Liu, Tianming, et al.
Published: (2025)
Evaluating Collective Behaviour of Hundreds of LLM Agents
by: Willis, Richard, et al.
Published: (2026)
by: Willis, Richard, et al.
Published: (2026)
Making Teams and Influencing Agents: Efficiently Coordinating Decision Trees for Interpretable Multi-Agent Reinforcement Learning
by: Chen, Rex, et al.
Published: (2025)
by: Chen, Rex, et al.
Published: (2025)
Aegis: Taxonomy and Optimizations for Overcoming Agent-Environment Failures in LLM Agents
by: Song, Kevin, et al.
Published: (2025)
by: Song, Kevin, et al.
Published: (2025)
Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game
by: Jian, HuaDong, et al.
Published: (2026)
by: Jian, HuaDong, et al.
Published: (2026)
Symphony-Coord: Adaptive Routing for Multi-Agent LLM Systems
by: Guan, Zhaoyang, et al.
Published: (2026)
by: Guan, Zhaoyang, et al.
Published: (2026)
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
by: Gho, Brendan, et al.
Published: (2025)
by: Gho, Brendan, et al.
Published: (2025)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
by: Schipper, Olivier, et al.
Published: (2025)
by: Schipper, Olivier, et al.
Published: (2025)
Beyond Browsing: API-Based Web Agents
by: Song, Yueqi, et al.
Published: (2024)
by: Song, Yueqi, et al.
Published: (2024)
HiveMind: Contribution-Guided Online Prompt Optimization of LLM Multi-Agent Systems
by: Xia, Yihan, et al.
Published: (2025)
by: Xia, Yihan, et al.
Published: (2025)
First Field-Trial Demonstration of L4 Autonomous Optical Network for Distributed AI Training Communication: An LLM-Powered Multi-AI-Agent Solution
by: Zhang, Yihao, et al.
Published: (2025)
by: Zhang, Yihao, et al.
Published: (2025)
Hierarchical Decentralized Multi-Agent Coordination with Privacy-Preserving Knowledge Sharing: Extending AgentNet for Scalable Autonomous Systems
by: Nalagatla, Goutham
Published: (2025)
by: Nalagatla, Goutham
Published: (2025)
Evaluating Multi-Agent LLM Architectures for Rare Disease Diagnosis
by: Almasoud, Ahmed
Published: (2026)
by: Almasoud, Ahmed
Published: (2026)
Similar Items
-
SpecBench: Evaluating Specification-Level Reasoning for Software Engineering LLM Agents
by: Hamblin, Grant, et al.
Published: (2026) -
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
by: Zhong, Shanshan, et al.
Published: (2026) -
CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs
by: Zou, Chelsea, et al.
Published: (2026) -
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026) -
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models
by: Agashe, Saaket, et al.
Published: (2023)