Invisible Orchestrators Suppress Protective Behavior and Dissociate Power-Holders: Safety Risks in Multi-Agent LLM Systems
Fuente:
arXiv
Saved in:
| Main Author: | Fukui, Hiroki |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Behavior of Single LLM-Driven Multi-Agent Systems
by: Li, Jialing, et al.
Published: (2026)
by: Li, Jialing, et al.
Published: (2026)
Soft-Label Governance for Distributional Safety in Multi-Agent Systems
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms
by: Hu, Botao Amber, et al.
Published: (2026)
by: Hu, Botao Amber, et al.
Published: (2026)
Urban-MAS: Human-Centered Urban Prediction with LLM-Based Multi-Agent System
by: Lou, Shangyu
Published: (2025)
by: Lou, Shangyu
Published: (2025)
A Safety-Aware Role-Orchestrated Multi-Agent LLM Framework for Behavioral Health Communication Simulation
by: Cho, Ha Na
Published: (2026)
by: Cho, Ha Na
Published: (2026)
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024)
by: Jiang, Yuan-Hao, et al.
Published: (2024)
TACLA: An LLM-Based Multi-Agent Tool for Transactional Analysis Training in Education
by: Zamojska, Monika, et al.
Published: (2025)
by: Zamojska, Monika, et al.
Published: (2025)
CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs
by: Nasarian, Elham, et al.
Published: (2026)
by: Nasarian, Elham, et al.
Published: (2026)
AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power
by: Ruan, Anbang, et al.
Published: (2026)
by: Ruan, Anbang, et al.
Published: (2026)
CogniPair: From LLM Chatbots to Conscious AI Agents -- GNWT-Based Multi-Agent Digital Twins for Social Pairing -- Dating & Hiring Applications
by: Ye, Wanghao, et al.
Published: (2025)
by: Ye, Wanghao, et al.
Published: (2025)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
Evaluating Online Moderation Via LLM-Powered Counterfactual Simulations
by: Fidone, Giacomo, et al.
Published: (2025)
by: Fidone, Giacomo, et al.
Published: (2025)
Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline
by: Dietrich, Juergen
Published: (2026)
by: Dietrich, Juergen
Published: (2026)
AgentTutor: Empowering Personalized Learning with Multi-Turn Interactive Teaching in Intelligent Education Systems
by: Liu, Yuxin, et al.
Published: (2025)
by: Liu, Yuxin, et al.
Published: (2025)
When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation
by: Andric, Sandro
Published: (2026)
by: Andric, Sandro
Published: (2026)
Toward Personalizing Quantum Computing Education: An Evolutionary LLM-Powered Approach
by: Elhaimeur, Iizalaarab, et al.
Published: (2025)
by: Elhaimeur, Iizalaarab, et al.
Published: (2025)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
by: Zhu, Kunlun, et al.
Published: (2025)
by: Zhu, Kunlun, et al.
Published: (2025)
Emergent Social Dynamics of LLM Agents in the El Farol Bar Problem
by: Takata, Ryosuke, et al.
Published: (2025)
by: Takata, Ryosuke, et al.
Published: (2025)
Let's Get You Hired: A Job Seeker's Perspective on Multi-Agent Recruitment Systems for Explaining Hiring Decisions
by: Bhattacharya, Aditya, et al.
Published: (2025)
by: Bhattacharya, Aditya, et al.
Published: (2025)
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
by: Dietrich, Juergen
Published: (2026)
by: Dietrich, Juergen
Published: (2026)
Why Agents Compromise Safety Under Pressure
by: Jiang, Hengle, et al.
Published: (2026)
by: Jiang, Hengle, et al.
Published: (2026)
From Narrative to Action: A Hierarchical LLM-Agent Framework for Human Mobility Generation
by: Li, Qiumeng, et al.
Published: (2025)
by: Li, Qiumeng, et al.
Published: (2025)
Towards Simulating Social Influence Dynamics with LLM-based Multi-agents
by: Lin, Hsien-Tsung, et al.
Published: (2025)
by: Lin, Hsien-Tsung, et al.
Published: (2025)
MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems
by: Chen, Kai, et al.
Published: (2025)
by: Chen, Kai, et al.
Published: (2025)
Multimodal Safety Evaluation in Generative Agent Social Simulations
by: Vera, Alhim, et al.
Published: (2025)
by: Vera, Alhim, et al.
Published: (2025)
Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics
by: Haklidir, Mehmet
Published: (2026)
by: Haklidir, Mehmet
Published: (2026)
HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems
by: Hou, Zhipeng, et al.
Published: (2025)
by: Hou, Zhipeng, et al.
Published: (2025)
I Want to Break Free! Persuasion and Anti-Social Behavior of LLMs in Multi-Agent Settings with Social Hierarchy
by: Campedelli, Gian Maria, et al.
Published: (2024)
by: Campedelli, Gian Maria, et al.
Published: (2024)
Learning Latency-Aware Orchestration for Parallel Multi-Agent Systems
by: Shi, Xi, et al.
Published: (2026)
by: Shi, Xi, et al.
Published: (2026)
MAEBE: Multi-Agent Emergent Behavior Framework
by: Erisken, Sinem, et al.
Published: (2025)
by: Erisken, Sinem, et al.
Published: (2025)
The Hidden Strength of Disagreement: Unraveling the Consensus-Diversity Tradeoff in Adaptive Multi-Agent Systems
by: Wu, Zengqing, et al.
Published: (2025)
by: Wu, Zengqing, et al.
Published: (2025)
Preserving Cultural Identity with Context-Aware Translation Through Multi-Agent AI Systems
by: Anik, Mahfuz Ahmed, et al.
Published: (2025)
by: Anik, Mahfuz Ahmed, et al.
Published: (2025)
Mapping Human Anti-collusion Mechanisms to Multi-agent AI Systems
by: Idowu, Jamiu, et al.
Published: (2026)
by: Idowu, Jamiu, et al.
Published: (2026)
Transforming Competition into Collaboration: The Revolutionary Role of Multi-Agent Systems and Language Models in Modern Organizations
by: Cruz, Carlos Jose Xavier
Published: (2024)
by: Cruz, Carlos Jose Xavier
Published: (2024)
Toward Evaluation Frameworks for Multi-Agent Scientific AI Systems
by: Abram, Marcin
Published: (2026)
by: Abram, Marcin
Published: (2026)
Embodied LLM Agents Learn to Cooperate in Organized Teams
by: Guo, Xudong, et al.
Published: (2024)
by: Guo, Xudong, et al.
Published: (2024)
H2-MARL: Multi-Agent Reinforcement Learning for Pareto Optimality in Hospital Capacity Strain and Human Mobility during Epidemic
by: Luo, Xueting, et al.
Published: (2025)
by: Luo, Xueting, et al.
Published: (2025)
Law in Silico: Simulating Legal Society with LLM-Based Agents
by: Wang, Yiding, et al.
Published: (2025)
by: Wang, Yiding, et al.
Published: (2025)
Dynamics of Moral Behavior in Heterogeneous Populations of Learning Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Gradientsys: A Multi-Agent LLM Scheduler with ReAct Orchestration
by: Song, Xinyuan, et al.
Published: (2025)
by: Song, Xinyuan, et al.
Published: (2025)
Similar Items
-
Scaling Behavior of Single LLM-Driven Multi-Agent Systems
by: Li, Jialing, et al.
Published: (2026) -
Soft-Label Governance for Distributional Safety in Multi-Agent Systems
by: Aiersilan, Aizierjiang, et al.
Published: (2026) -
Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms
by: Hu, Botao Amber, et al.
Published: (2026) -
Urban-MAS: Human-Centered Urban Prediction with LLM-Based Multi-Agent System
by: Lou, Shangyu
Published: (2025) -
A Safety-Aware Role-Orchestrated Multi-Agent LLM Framework for Behavioral Health Communication Simulation
by: Cho, Ha Na
Published: (2026)