Toward Evaluation Frameworks for Multi-Agent Scientific AI Systems
Fuente:
arXiv
Saved in:
| Main Author: | Abram, Marcin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024)
by: Jiang, Yuan-Hao, et al.
Published: (2024)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
Towards Computational Social Dynamics of Semi-Autonomous AI Agents
by: Lidarity, S. O., et al.
Published: (2026)
by: Lidarity, S. O., et al.
Published: (2026)
Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits
by: Ye, Jinyi, et al.
Published: (2026)
by: Ye, Jinyi, et al.
Published: (2026)
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
Soft-Label Governance for Distributional Safety in Multi-Agent Systems
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
Scaling Behavior of Single LLM-Driven Multi-Agent Systems
by: Li, Jialing, et al.
Published: (2026)
by: Li, Jialing, et al.
Published: (2026)
Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics
by: Haklidir, Mehmet
Published: (2026)
by: Haklidir, Mehmet
Published: (2026)
Mapping Human Anti-collusion Mechanisms to Multi-agent AI Systems
by: Idowu, Jamiu, et al.
Published: (2026)
by: Idowu, Jamiu, et al.
Published: (2026)
Preserving Cultural Identity with Context-Aware Translation Through Multi-Agent AI Systems
by: Anik, Mahfuz Ahmed, et al.
Published: (2025)
by: Anik, Mahfuz Ahmed, et al.
Published: (2025)
Urban-MAS: Human-Centered Urban Prediction with LLM-Based Multi-Agent System
by: Lou, Shangyu
Published: (2025)
by: Lou, Shangyu
Published: (2025)
AI Agents as Policymakers in Simulated Epidemics
by: Aoki, Goshi, et al.
Published: (2026)
by: Aoki, Goshi, et al.
Published: (2026)
CogniPair: From LLM Chatbots to Conscious AI Agents -- GNWT-Based Multi-Agent Digital Twins for Social Pairing -- Dating & Hiring Applications
by: Ye, Wanghao, et al.
Published: (2025)
by: Ye, Wanghao, et al.
Published: (2025)
AgentTutor: Empowering Personalized Learning with Multi-Turn Interactive Teaching in Intelligent Education Systems
by: Liu, Yuxin, et al.
Published: (2025)
by: Liu, Yuxin, et al.
Published: (2025)
Bit-politeia: An AI Agent Community in Blockchain
by: Yang, Xing
Published: (2026)
by: Yang, Xing
Published: (2026)
CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs
by: Nasarian, Elham, et al.
Published: (2026)
by: Nasarian, Elham, et al.
Published: (2026)
Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline
by: Dietrich, Juergen
Published: (2026)
by: Dietrich, Juergen
Published: (2026)
Invisible Orchestrators Suppress Protective Behavior and Dissociate Power-Holders: Safety Risks in Multi-Agent LLM Systems
by: Fukui, Hiroki
Published: (2026)
by: Fukui, Hiroki
Published: (2026)
Let's Get You Hired: A Job Seeker's Perspective on Multi-Agent Recruitment Systems for Explaining Hiring Decisions
by: Bhattacharya, Aditya, et al.
Published: (2025)
by: Bhattacharya, Aditya, et al.
Published: (2025)
Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure
by: Butt, Talal Ashraf, et al.
Published: (2026)
by: Butt, Talal Ashraf, et al.
Published: (2026)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
by: Zhu, Kunlun, et al.
Published: (2025)
by: Zhu, Kunlun, et al.
Published: (2025)
GGBond: Growing Graph-Based AI-Agent Society for Socially-Aware Recommender Simulation
by: Zhong, Hailin, et al.
Published: (2025)
by: Zhong, Hailin, et al.
Published: (2025)
Towards Simulating Social Influence Dynamics with LLM-based Multi-agents
by: Lin, Hsien-Tsung, et al.
Published: (2025)
by: Lin, Hsien-Tsung, et al.
Published: (2025)
From Narrative to Action: A Hierarchical LLM-Agent Framework for Human Mobility Generation
by: Li, Qiumeng, et al.
Published: (2025)
by: Li, Qiumeng, et al.
Published: (2025)
CardAIc-Agents: A Multimodal Framework with Hierarchical Adaptation for Cardiac Care Support
by: Zhang, Yuting, et al.
Published: (2025)
by: Zhang, Yuting, et al.
Published: (2025)
Socio-technical aspects of Agentic AI
by: Donta, Praveen Kumar, et al.
Published: (2025)
by: Donta, Praveen Kumar, et al.
Published: (2025)
TACLA: An LLM-Based Multi-Agent Tool for Transactional Analysis Training in Education
by: Zamojska, Monika, et al.
Published: (2025)
by: Zamojska, Monika, et al.
Published: (2025)
SentinelAI: A Multi-Agent Framework for Structuring and Linking NG9-1-1 Emergency Incident Data
by: Ho, Kliment, et al.
Published: (2026)
by: Ho, Kliment, et al.
Published: (2026)
Advancing Responsible Innovation in Agentic AI: A study of Ethical Frameworks for Household Automation
by: Chandra, Joydeep, et al.
Published: (2025)
by: Chandra, Joydeep, et al.
Published: (2025)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
by: Hu, Yueqing, et al.
Published: (2026)
by: Hu, Yueqing, et al.
Published: (2026)
Human-AI Collaboration for Estimating Scientific Replicability
by: Chakravorti, Tatiana, et al.
Published: (2026)
by: Chakravorti, Tatiana, et al.
Published: (2026)
The Hidden Strength of Disagreement: Unraveling the Consensus-Diversity Tradeoff in Adaptive Multi-Agent Systems
by: Wu, Zengqing, et al.
Published: (2025)
by: Wu, Zengqing, et al.
Published: (2025)
Multi-Agent Quantum Reinforcement Learning using Evolutionary Optimization
by: Kölle, Michael, et al.
Published: (2023)
by: Kölle, Michael, et al.
Published: (2023)
Will Agents Replace Us? Perceptions of Autonomous Multi-Agent AI
by: Balic, Nikola
Published: (2025)
by: Balic, Nikola
Published: (2025)
Transforming Competition into Collaboration: The Revolutionary Role of Multi-Agent Systems and Language Models in Modern Organizations
by: Cruz, Carlos Jose Xavier
Published: (2024)
by: Cruz, Carlos Jose Xavier
Published: (2024)
H2-MARL: Multi-Agent Reinforcement Learning for Pareto Optimality in Hospital Capacity Strain and Human Mobility during Epidemic
by: Luo, Xueting, et al.
Published: (2025)
by: Luo, Xueting, et al.
Published: (2025)
AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power
by: Ruan, Anbang, et al.
Published: (2026)
by: Ruan, Anbang, et al.
Published: (2026)
Group size effects and collective misalignment in LLM multi-agent systems
by: Flint, Ariel, et al.
Published: (2025)
by: Flint, Ariel, et al.
Published: (2025)
Emergent social conventions and collective bias in LLM populations
by: Ashery, Ariel Flint, et al.
Published: (2024)
by: Ashery, Ariel Flint, et al.
Published: (2024)
FlockVote: LLM-Empowered Agent-Based Modeling for Simulating U.S. Presidential Elections
by: Zhou, Lingfeng, et al.
Published: (2025)
by: Zhou, Lingfeng, et al.
Published: (2025)
Similar Items
-
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024) -
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
by: Ranjan, Rajesh, et al.
Published: (2025) -
Towards Computational Social Dynamics of Semi-Autonomous AI Agents
by: Lidarity, S. O., et al.
Published: (2026) -
Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits
by: Ye, Jinyi, et al.
Published: (2026) -
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems
by: Ranjan, Rajesh, et al.
Published: (2025)