MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhu, Kunlun, Du, Hongyi, Hong, Zhaochen, Yang, Xiaocheng, Guo, Shuyi, Wang, Zhe, Wang, Zhenhailong, Qian, Cheng, Tang, Xiangru, Ji, Heng, You, Jiaxuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning
por: White, Isadora, et al.
Publicado: (2025)
por: White, Isadora, et al.
Publicado: (2025)
Cross-Task Experiential Learning on LLM-based Multi-Agent Collaboration
por: Li, Yilong, et al.
Publicado: (2025)
por: Li, Yilong, et al.
Publicado: (2025)
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling
por: Zhou, Jialong, et al.
Publicado: (2025)
por: Zhou, Jialong, et al.
Publicado: (2025)
AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research
por: Xia, Tong, et al.
Publicado: (2025)
por: Xia, Tong, et al.
Publicado: (2025)
When AI Agents Collude Online: Financial Fraud Risks by Collaborative LLM Agents on Social Platforms
por: Ren, Qibing, et al.
Publicado: (2025)
por: Ren, Qibing, et al.
Publicado: (2025)
Competition and Cooperation of LLM Agents in Games
por: Yao, Jiayi, et al.
Publicado: (2026)
por: Yao, Jiayi, et al.
Publicado: (2026)
Transforming Competition into Collaboration: The Revolutionary Role of Multi-Agent Systems and Language Models in Modern Organizations
por: Cruz, Carlos Jose Xavier
Publicado: (2024)
por: Cruz, Carlos Jose Xavier
Publicado: (2024)
CACA Agent: Capability Collaboration based AI Agent
por: Xu, Peng, et al.
Publicado: (2024)
por: Xu, Peng, et al.
Publicado: (2024)
Is Your LLM-as-a-Recommender Agent Trustable? LLMs' Recommendation is Easily Hacked by Biases (Preferences)
por: Tang, Zichen, et al.
Publicado: (2026)
por: Tang, Zichen, et al.
Publicado: (2026)
Scaling Large Language Model-based Multi-Agent Collaboration
por: Qian, Chen, et al.
Publicado: (2024)
por: Qian, Chen, et al.
Publicado: (2024)
Communication to Completion: Modeling Collaborative Workflows with Intelligent Multi-Agent Communication
por: Lu, Yiming, et al.
Publicado: (2025)
por: Lu, Yiming, et al.
Publicado: (2025)
Dive into the Agent Matrix: A Realistic Evaluation of Self-Replication Risk in LLM Agents
por: Zhang, Boxuan, et al.
Publicado: (2025)
por: Zhang, Boxuan, et al.
Publicado: (2025)
DDO: Dual-Decision Optimization for LLM-Based Medical Consultation via Multi-Agent Collaboration
por: Jia, Zhihao, et al.
Publicado: (2025)
por: Jia, Zhihao, et al.
Publicado: (2025)
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models
por: Agashe, Saaket, et al.
Publicado: (2023)
por: Agashe, Saaket, et al.
Publicado: (2023)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
por: Liu, Zijun, et al.
Publicado: (2023)
por: Liu, Zijun, et al.
Publicado: (2023)
How to Steer Your Multi-Agent System: Human-LLM Collaborative Planning
por: He, Zeyu, et al.
Publicado: (2026)
por: He, Zeyu, et al.
Publicado: (2026)
Preventing Rogue Agents Improves Multi-Agent Collaboration
por: Barbi, Ohav, et al.
Publicado: (2025)
por: Barbi, Ohav, et al.
Publicado: (2025)
Autonomous Agents for Collaborative Task under Information Asymmetry
por: Liu, Wei, et al.
Publicado: (2024)
por: Liu, Wei, et al.
Publicado: (2024)
A Survey on Trustworthy LLM Agents: Threats and Countermeasures
por: Yu, Miao, et al.
Publicado: (2025)
por: Yu, Miao, et al.
Publicado: (2025)
Inherent and emergent liability issues in LLM-based agentic systems: a principal-agent perspective
por: Gabison, Garry A., et al.
Publicado: (2025)
por: Gabison, Garry A., et al.
Publicado: (2025)
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems
por: Zhang, Shaokun, et al.
Publicado: (2025)
por: Zhang, Shaokun, et al.
Publicado: (2025)
MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems
por: Ye, Rui, et al.
Publicado: (2025)
por: Ye, Rui, et al.
Publicado: (2025)
AstroVLM: Expert Multi-agent Collaborative Reasoning for Astronomical Imaging Quality Diagnosis
por: Han, Yaohui, et al.
Publicado: (2026)
por: Han, Yaohui, et al.
Publicado: (2026)
Multi-Agent Collaboration via Cross-Team Orchestration
por: Du, Zhuoyun, et al.
Publicado: (2024)
por: Du, Zhuoyun, et al.
Publicado: (2024)
The Subtle Art of Defection: Understanding Uncooperative Behaviors in LLM based Multi-Agent Systems
por: Kulshreshtha, Devang, et al.
Publicado: (2025)
por: Kulshreshtha, Devang, et al.
Publicado: (2025)
The Consensus Trap: Rescuing Multi-Agent LLMs from Adversarial Majorities via Token-Level Collaboration
por: Liu, Jiayuan, et al.
Publicado: (2026)
por: Liu, Jiayuan, et al.
Publicado: (2026)
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
por: Gho, Brendan, et al.
Publicado: (2025)
por: Gho, Brendan, et al.
Publicado: (2025)
A Simulation-Based Method for Testing Collaborative Learning Scaffolds Using LLM-Based Multi-Agent Systems
por: Wua, Han, et al.
Publicado: (2026)
por: Wua, Han, et al.
Publicado: (2026)
Multi-Stakeholder Alignment in LLM-Powered Collaborative AI Systems: A Multi-Agent Framework for Intelligent Tutoring
por: Uchoa, Alexandre P, et al.
Publicado: (2025)
por: Uchoa, Alexandre P, et al.
Publicado: (2025)
Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration
por: Zhang, Yang, et al.
Publicado: (2024)
por: Zhang, Yang, et al.
Publicado: (2024)
GTAlign: Game-Theoretic Alignment of LLM Assistants for Social Welfare
por: Zhu, Siqi, et al.
Publicado: (2025)
por: Zhu, Siqi, et al.
Publicado: (2025)
Multi-Agent Collaboration via Evolving Orchestration
por: Dang, Yufan, et al.
Publicado: (2025)
por: Dang, Yufan, et al.
Publicado: (2025)
What Makes Good Collaborative Views? Contrastive Mutual Information Maximization for Multi-Agent Perception
por: Su, Wanfang, et al.
Publicado: (2024)
por: Su, Wanfang, et al.
Publicado: (2024)
MultiAgent Collaboration Attack: Investigating Adversarial Attacks in Large Language Model Collaborations via Debate
por: Amayuelas, Alfonso, et al.
Publicado: (2024)
por: Amayuelas, Alfonso, et al.
Publicado: (2024)
APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents
por: Zheng, Quan, et al.
Publicado: (2026)
por: Zheng, Quan, et al.
Publicado: (2026)
A Text-to-Game Engine for UGC-Based Role-Playing Games
por: Zhang, Lei, et al.
Publicado: (2024)
por: Zhang, Lei, et al.
Publicado: (2024)
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents
por: Zhao, Qinlin, et al.
Publicado: (2023)
por: Zhao, Qinlin, et al.
Publicado: (2023)
LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis
por: Xu, Shihao, et al.
Publicado: (2026)
por: Xu, Shihao, et al.
Publicado: (2026)
SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System
por: Pan, Shuai, et al.
Publicado: (2026)
por: Pan, Shuai, et al.
Publicado: (2026)
HARBOR: Exploring Persona Dynamics in Multi-Agent Competition
por: Jiang, Kenan, et al.
Publicado: (2025)
por: Jiang, Kenan, et al.
Publicado: (2025)
Ejemplares similares
-
Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning
por: White, Isadora, et al.
Publicado: (2025) -
Cross-Task Experiential Learning on LLM-based Multi-Agent Collaboration
por: Li, Yilong, et al.
Publicado: (2025) -
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling
por: Zhou, Jialong, et al.
Publicado: (2025) -
AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research
por: Xia, Tong, et al.
Publicado: (2025) -
When AI Agents Collude Online: Financial Fraud Risks by Collaborative LLM Agents on Social Platforms
por: Ren, Qibing, et al.
Publicado: (2025)