ZSC-Eval: An Evaluation Toolkit and Benchmark for Multi-agent Zero-shot Coordination
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xihuai, Zhang, Shao, Zhang, Wenhao, Dong, Wentao, Chen, Jingxiao, Wen, Ying, Zhang, Weinan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mutual Theory of Mind in Human-AI Collaboration: An Empirical Study with LLM-driven AI Agents in a Real-time Shared Workspace Task
by: Zhang, Shao, et al.
Published: (2024)
by: Zhang, Shao, et al.
Published: (2024)
Leveraging Dual Process Theory in Language Agent Framework for Real-time Simultaneous Human-AI Collaboration
by: Zhang, Shao, et al.
Published: (2025)
by: Zhang, Shao, et al.
Published: (2025)
FlowEval: Reference-based Evaluation of Generated User Interfaces
by: Wu, Jason, et al.
Published: (2026)
by: Wu, Jason, et al.
Published: (2026)
AIPOM: Agent-aware Interactive Planning for Multi-Agent Systems
by: Kim, Hannah, et al.
Published: (2025)
by: Kim, Hannah, et al.
Published: (2025)
Too Many Specialists: Emergent Inefficiencies and Bottlenecks for Multi-agent Ad-hoc Collaboration
by: Panny, Benjamin, et al.
Published: (2026)
by: Panny, Benjamin, et al.
Published: (2026)
On the Utility of External Agent Intention Predictor for Human-AI Coordination
by: Wang, Chenxu, et al.
Published: (2024)
by: Wang, Chenxu, et al.
Published: (2024)
Model-based Multi-agent Reinforcement Learning: Recent Progress and Prospects
by: Wang, Xihuai, et al.
Published: (2022)
by: Wang, Xihuai, et al.
Published: (2022)
A Simulation-Based Method for Testing Collaborative Learning Scaffolds Using LLM-Based Multi-Agent Systems
by: Wua, Han, et al.
Published: (2026)
by: Wua, Han, et al.
Published: (2026)
How to Steer Your Multi-Agent System: Human-LLM Collaborative Planning
by: He, Zeyu, et al.
Published: (2026)
by: He, Zeyu, et al.
Published: (2026)
Auto-Slides: An Interactive Multi-Agent System for Creating and Customizing Research Presentations
by: Yang, Yuheng, et al.
Published: (2025)
by: Yang, Yuheng, et al.
Published: (2025)
NarrativeLoom: Enhancing Creative Storytelling through Multi-Persona Collaborative Improvisation
by: Ma, Yuxi, et al.
Published: (2026)
by: Ma, Yuxi, et al.
Published: (2026)
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026)
by: Zhang, Shuyu, et al.
Published: (2026)
MAxPrototyper: A Multi-Agent Generation System for Interactive User Interface Prototyping
by: Yuan, Mingyue, et al.
Published: (2024)
by: Yuan, Mingyue, et al.
Published: (2024)
Modeling multi-agent motion dynamics in immersive rooms
by: Huang, Jerry M., et al.
Published: (2025)
by: Huang, Jerry M., et al.
Published: (2025)
Multi-Stakeholder Alignment in LLM-Powered Collaborative AI Systems: A Multi-Agent Framework for Intelligent Tutoring
by: Uchoa, Alexandre P, et al.
Published: (2025)
by: Uchoa, Alexandre P, et al.
Published: (2025)
Do LLMs Need to See Everything? A Benchmark and Study of Failures in LLM-driven Smartphone Automation using Screentext vs. Screenshots
by: Zhang, Shiquan, et al.
Published: (2026)
by: Zhang, Shiquan, et al.
Published: (2026)
Context-Mediated Domain Adaptation in Multi-Agent Sensemaking Systems
by: Wolter, Anton, et al.
Published: (2026)
by: Wolter, Anton, et al.
Published: (2026)
CartoAgent: a multimodal large language model-powered multi-agent cartographic framework for map style transfer and evaluation
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
FOCAL: Filtered On-device Continuous Activity Logging for Efficient Personal Desktop Summarization
by: Yin, Haoran, et al.
Published: (2026)
by: Yin, Haoran, et al.
Published: (2026)
Inject, Fork, Compare: Defining an Interaction Vocabulary for Multi-Agent Simulation Platforms
by: Lee, HwiJoon, et al.
Published: (2025)
by: Lee, HwiJoon, et al.
Published: (2025)
Proteus: Shapeshifting Desktop Visualizations for Mobile via Multi-level Intelligent Adaptation
by: Liu, Can, et al.
Published: (2026)
by: Liu, Can, et al.
Published: (2026)
Decoupled Intelligence: A Multi-Agent LLM Framework for Controllable Traffic Scenario Generation in SUMO
by: Li, Shuyang, et al.
Published: (2026)
by: Li, Shuyang, et al.
Published: (2026)
FACET: Teacher-Centred LLM-Based Multi-Agent Systems-Towards Personalized Educational Worksheets
by: Gonnermann-Müller, Jana, et al.
Published: (2025)
by: Gonnermann-Müller, Jana, et al.
Published: (2025)
Ad-Hoc Human-AI Coordination Challenge
by: Dizdarević, Tin, et al.
Published: (2025)
by: Dizdarević, Tin, et al.
Published: (2025)
Enabling Multi-Robot Collaboration from Single-Human Guidance
by: Ji, Zhengran, et al.
Published: (2024)
by: Ji, Zhengran, et al.
Published: (2024)
DialogGuard: Multi-Agent Psychosocial Safety Evaluation of Sensitive LLM Responses
by: Luo, Han, et al.
Published: (2025)
by: Luo, Han, et al.
Published: (2025)
TransLaw: A Large-Scale Dataset and Multi-Agent Benchmark Simulating Professional Translation of Hong Kong Case Law
by: Xuan, Xi, et al.
Published: (2025)
by: Xuan, Xi, et al.
Published: (2025)
CE-MRS: Contrastive Explanations for Multi-Robot Systems
by: Schneider, Ethan, et al.
Published: (2024)
by: Schneider, Ethan, et al.
Published: (2024)
Narrative Memory in Machines: Multi-Agent Arc Extraction in Serialized TV
by: Balestri, Roberto, et al.
Published: (2025)
by: Balestri, Roberto, et al.
Published: (2025)
SoDA: An Efficient Interaction Paradigm for the Agentic Web
by: Cui, Zicai, et al.
Published: (2025)
by: Cui, Zicai, et al.
Published: (2025)
Intermittent Rendezvous Plans with Mixed Integer Linear Program for Large-Scale Multi-Robot Exploration
by: da Silva, Alysson Ribeiro, et al.
Published: (2025)
by: da Silva, Alysson Ribeiro, et al.
Published: (2025)
GPT versus Humans: Uncovering Ethical Concerns in Conversational Generative AI-empowered Multi-Robot Systems
by: Rousi, Rebekah, et al.
Published: (2024)
by: Rousi, Rebekah, et al.
Published: (2024)
Conversational Self-Play for Discovering and Understanding Psychotherapy Approaches
by: Kampman, Onno P, et al.
Published: (2025)
by: Kampman, Onno P, et al.
Published: (2025)
Data assimilation approach for addressing imperfections in people flow measurement techniques using particle filter
by: Murata, Ryo, et al.
Published: (2024)
by: Murata, Ryo, et al.
Published: (2024)
EmBARDiment: an Embodied AI Agent for Productivity in XR
by: Bovo, Riccardo, et al.
Published: (2024)
by: Bovo, Riccardo, et al.
Published: (2024)
CandorMD: An AI-Assisted Audio Simulation and Feedback System for Training Clinicians for Medical Error Disclosure
by: Lin, Inna Wanyin, et al.
Published: (2026)
by: Lin, Inna Wanyin, et al.
Published: (2026)
Cloud and IoT based Smart Agent-driven Simulation of Human Gait for Detecting Muscles Disorder
by: Saadati, Sina, et al.
Published: (2024)
by: Saadati, Sina, et al.
Published: (2024)
LLM-Powered Virtual Patient Agents for Interactive Clinical Skills Training with Automated Feedback
by: Voigt, Henrik, et al.
Published: (2025)
by: Voigt, Henrik, et al.
Published: (2025)
A2H: Agent-to-Human Protocol for AI Agent
by: Liang, Zhiyuan, et al.
Published: (2025)
by: Liang, Zhiyuan, et al.
Published: (2025)
CoCre-Sam (Kokkuri-san): Modeling Ouija Board as Collective Langevin Dynamics Sampling from Fused Language Models
by: Taniguchi, Tadahiro, et al.
Published: (2025)
by: Taniguchi, Tadahiro, et al.
Published: (2025)
Similar Items
-
Mutual Theory of Mind in Human-AI Collaboration: An Empirical Study with LLM-driven AI Agents in a Real-time Shared Workspace Task
by: Zhang, Shao, et al.
Published: (2024) -
Leveraging Dual Process Theory in Language Agent Framework for Real-time Simultaneous Human-AI Collaboration
by: Zhang, Shao, et al.
Published: (2025) -
FlowEval: Reference-based Evaluation of Generated User Interfaces
by: Wu, Jason, et al.
Published: (2026) -
AIPOM: Agent-aware Interactive Planning for Multi-Agent Systems
by: Kim, Hannah, et al.
Published: (2025) -
Too Many Specialists: Emergent Inefficiencies and Bottlenecks for Multi-agent Ad-hoc Collaboration
by: Panny, Benjamin, et al.
Published: (2026)