Can Lessons From Human Teams Be Applied to Multi-Agent Systems? The Role of Structure, Diversity, and Interaction Dynamics
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Muralidharan, Rasika, Kwak, Haewoon, An, Jisun |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
par: Kachwala, Zoher, et autres
Publié: (2026)
par: Kachwala, Zoher, et autres
Publié: (2026)
XChoice: Explainable Evaluation of AI-Human Alignment in LLM-based Constrained Choice Decision Making
par: Qi, Weihong, et autres
Publié: (2026)
par: Qi, Weihong, et autres
Publié: (2026)
Vulnerability of LLMs' Stated Beliefs? LLMs Belief Resistance Check Through Strategic Persuasive Conversation Interventions
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
par: Huang, Fan, et autres
Publié: (2024)
par: Huang, Fan, et autres
Publié: (2024)
Understanding Moral Reasoning Trajectories in Large Language Models: Toward Probing-Based Explainability
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
ChatGPT Rates Natural Language Explanation Quality Like Humans: But on Which Scales?
par: Huang, Fan, et autres
Publié: (2024)
par: Huang, Fan, et autres
Publié: (2024)
Can we trust the evaluation on ChatGPT?
par: Aiyappa, Rachith, et autres
Publié: (2023)
par: Aiyappa, Rachith, et autres
Publié: (2023)
Benchmarking zero-shot stance detection with FlanT5-XXL: Insights from training data, prompting, and decoding strategies into its near-SoTA performance
par: Aiyappa, Rachith, et autres
Publié: (2024)
par: Aiyappa, Rachith, et autres
Publié: (2024)
A Cross-Cultural Comparison of LLM-based Public Opinion Simulation: Evaluating Chinese and U.S. Models on Diverse Societies
par: Qi, Weihong, et autres
Publié: (2025)
par: Qi, Weihong, et autres
Publié: (2025)
Rematch: Robust and Efficient Matching of Local Knowledge Graphs to Improve Structural and Semantic Similarity
par: Kachwala, Zoher, et autres
Publié: (2024)
par: Kachwala, Zoher, et autres
Publié: (2024)
CogBias: Measuring and Mitigating Cognitive Bias in Large Language Models
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs
par: Sakai, Shintaro, et autres
Publié: (2025)
par: Sakai, Shintaro, et autres
Publié: (2025)
LLMs Can Infer Political Alignment from Online Conversations
par: Lee, Byunghwee, et autres
Publié: (2026)
par: Lee, Byunghwee, et autres
Publié: (2026)
What Helps Language Models Predict Human Beliefs: Demographics or Prior Stances?
par: Malone, Joseph, et autres
Publié: (2025)
par: Malone, Joseph, et autres
Publié: (2025)
Dynamic Role Assignment for Multi-Agent Debate
par: Zhang, Miao, et autres
Publié: (2026)
par: Zhang, Miao, et autres
Publié: (2026)
SIRAJ: Diverse and Efficient Red-Teaming for LLM Agents via Distilled Structured Reasoning
par: Zhou, Kaiwen, et autres
Publié: (2025)
par: Zhou, Kaiwen, et autres
Publié: (2025)
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
par: Song, Maojia, et autres
Publié: (2025)
par: Song, Maojia, et autres
Publié: (2025)
TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks
par: Wang, Xiangyu, et autres
Publié: (2026)
par: Wang, Xiangyu, et autres
Publié: (2026)
TopoDIM: One-shot Topology Generation of Diverse Interaction Modes for Multi-Agent Systems
par: Sun, Rui, et autres
Publié: (2026)
par: Sun, Rui, et autres
Publié: (2026)
A semantic embedding space based on large language models for modelling human beliefs
par: Lee, Byunghwee, et autres
Publié: (2024)
par: Lee, Byunghwee, et autres
Publié: (2024)
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
par: Xu, Zhenhua, et autres
Publié: (2026)
par: Xu, Zhenhua, et autres
Publié: (2026)
Lessons from the Field: An Adaptable Lifecycle Approach to Applied Dialogue Summarization
par: Chawla, Kushal, et autres
Publié: (2026)
par: Chawla, Kushal, et autres
Publié: (2026)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
par: Dhamaskar, Mohammed Amaan, et autres
Publié: (2025)
par: Dhamaskar, Mohammed Amaan, et autres
Publié: (2025)
RareAgents: Autonomous Multi-disciplinary Team for Rare Disease Diagnosis and Treatment
par: Chen, Xuanzhong, et autres
Publié: (2024)
par: Chen, Xuanzhong, et autres
Publié: (2024)
OpenDeception: Learning Deception and Trust in Human-AI Interaction via Multi-Agent Simulation
par: Wu, Yichen, et autres
Publié: (2025)
par: Wu, Yichen, et autres
Publié: (2025)
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team
par: Hosain, Md Tanzib, et autres
Publié: (2025)
par: Hosain, Md Tanzib, et autres
Publié: (2025)
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
par: Wang, Qiyao, et autres
Publié: (2026)
par: Wang, Qiyao, et autres
Publié: (2026)
CoMAS: Co-Evolving Multi-Agent Systems via Interaction Rewards
par: Xue, Xiangyuan, et autres
Publié: (2025)
par: Xue, Xiangyuan, et autres
Publié: (2025)
Can an Individual Manipulate the Collective Decisions of Multi-Agents?
par: Liu, Fengyuan, et autres
Publié: (2025)
par: Liu, Fengyuan, et autres
Publié: (2025)
Beyond Static Testbeds: An Interaction-Centric Agent Simulation Platform for Dynamic Recommender Systems
par: Jin, Song, et autres
Publié: (2025)
par: Jin, Song, et autres
Publié: (2025)
Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation
par: Chen, Nuo, et autres
Publié: (2026)
par: Chen, Nuo, et autres
Publié: (2026)
AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation
par: Wang, Ming, et autres
Publié: (2025)
par: Wang, Ming, et autres
Publié: (2025)
WebLists: Extracting Structured Information From Complex Interactive Websites Using Executable LLM Agents
par: Bohra, Arth, et autres
Publié: (2025)
par: Bohra, Arth, et autres
Publié: (2025)
TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
par: Long, Qian, et autres
Publié: (2024)
par: Long, Qian, et autres
Publié: (2024)
SMART-Editor: A Multi-Agent Framework for Human-Like Design Editing with Structural Integrity
par: Mondal, Ishani, et autres
Publié: (2025)
par: Mondal, Ishani, et autres
Publié: (2025)
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
par: Mushtaq, Abdullah, et autres
Publié: (2025)
par: Mushtaq, Abdullah, et autres
Publié: (2025)
RCR-Router: Efficient Role-Aware Context Routing for Multi-Agent LLM Systems with Structured Memory
par: Liu, Jun, et autres
Publié: (2025)
par: Liu, Jun, et autres
Publié: (2025)
CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming
par: Chiu, Yu Ying, et autres
Publié: (2024)
par: Chiu, Yu Ying, et autres
Publié: (2024)
BenchAgents: Multi-Agent Systems for Structured Benchmark Creation
par: Butt, Natasha, et autres
Publié: (2024)
par: Butt, Natasha, et autres
Publié: (2024)
CRMArena-Pro: Holistic Assessment of LLM Agents Across Diverse Business Scenarios and Interactions
par: Huang, Kung-Hsiang, et autres
Publié: (2025)
par: Huang, Kung-Hsiang, et autres
Publié: (2025)
Documents similaires
-
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
par: Kachwala, Zoher, et autres
Publié: (2026) -
XChoice: Explainable Evaluation of AI-Human Alignment in LLM-based Constrained Choice Decision Making
par: Qi, Weihong, et autres
Publié: (2026) -
Vulnerability of LLMs' Stated Beliefs? LLMs Belief Resistance Check Through Strategic Persuasive Conversation Interventions
par: Huang, Fan, et autres
Publié: (2026) -
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
par: Huang, Fan, et autres
Publié: (2024) -
Understanding Moral Reasoning Trajectories in Large Language Models: Toward Probing-Based Explainability
par: Huang, Fan, et autres
Publié: (2026)