GAMBIT: A Three-Mode Benchmark for Adversarial Robustness in Multi-Agent LLM Collectives
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mercier, Alexandre Le, Develder, Chris, Demeester, Thomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hidden State Poisoning Attacks against Mamba-based Language Models
von: Mercier, Alexandre Le, et al.
Veröffentlicht: (2026)
von: Mercier, Alexandre Le, et al.
Veröffentlicht: (2026)
An Adversary-Resistant Multi-Agent LLM System via Credibility Scoring
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2025)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2025)
Unleashing Diverse Thinking Modes in LLMs through Multi-Agent Collaboration
von: He, Zhixuan, et al.
Veröffentlicht: (2025)
von: He, Zhixuan, et al.
Veröffentlicht: (2025)
MAGIC: A Co-Evolving Attacker-Defender Adversarial Game for Robust LLM Safety
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2026)
Symphony: A Decentralized Multi-Agent Framework for Scalable Collective Intelligence
von: Wang, Ji, et al.
Veröffentlicht: (2025)
von: Wang, Ji, et al.
Veröffentlicht: (2025)
Context, Reasoning, and Hierarchy: A Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP
von: Bogdanov, Igor, et al.
Veröffentlicht: (2026)
von: Bogdanov, Igor, et al.
Veröffentlicht: (2026)
MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks
von: Zhu, Yinghao, et al.
Veröffentlicht: (2025)
von: Zhu, Yinghao, et al.
Veröffentlicht: (2025)
AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems
von: Wang, Zhexuan, et al.
Veröffentlicht: (2026)
von: Wang, Zhexuan, et al.
Veröffentlicht: (2026)
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
von: Khan, Rana Muhammad Shahroz, et al.
Veröffentlicht: (2025)
von: Khan, Rana Muhammad Shahroz, et al.
Veröffentlicht: (2025)
LLM Agents Making Agent Tools
von: Wölflein, Georg, et al.
Veröffentlicht: (2025)
von: Wölflein, Georg, et al.
Veröffentlicht: (2025)
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
von: Mushtaq, Abdullah, et al.
Veröffentlicht: (2025)
von: Mushtaq, Abdullah, et al.
Veröffentlicht: (2025)
PLANET: A Collection of Benchmarks for Evaluating LLMs' Planning Capabilities
von: Li, Haoming, et al.
Veröffentlicht: (2025)
von: Li, Haoming, et al.
Veröffentlicht: (2025)
Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control
von: Rezazadeh, Alireza, et al.
Veröffentlicht: (2025)
von: Rezazadeh, Alireza, et al.
Veröffentlicht: (2025)
RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation
von: Shen, Chengzhi, et al.
Veröffentlicht: (2026)
von: Shen, Chengzhi, et al.
Veröffentlicht: (2026)
Opponent Shaping in LLM Agents
von: Segura, Marta Emili Garcia, et al.
Veröffentlicht: (2025)
von: Segura, Marta Emili Garcia, et al.
Veröffentlicht: (2025)
TrustAgent: Towards Safe and Trustworthy LLM-based Agents
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
SPIO: Ensemble and Selective Strategies via LLM-Based Multi-Agent Planning in Automated Data Science
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
von: Nourzad, Narjes, et al.
Veröffentlicht: (2025)
von: Nourzad, Narjes, et al.
Veröffentlicht: (2025)
Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration
von: Zhang, Yang, et al.
Veröffentlicht: (2024)
von: Zhang, Yang, et al.
Veröffentlicht: (2024)
Dive into the Agent Matrix: A Realistic Evaluation of Self-Replication Risk in LLM Agents
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
Multi-Agent Design: Optimizing Agents with Better Prompts and Topologies
von: Zhou, Han, et al.
Veröffentlicht: (2025)
von: Zhou, Han, et al.
Veröffentlicht: (2025)
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions
von: Sun, Chuanneng, et al.
Veröffentlicht: (2024)
von: Sun, Chuanneng, et al.
Veröffentlicht: (2024)
MAC: Multi-Agent Constitution Learning
von: Thareja, Rushil, et al.
Veröffentlicht: (2026)
von: Thareja, Rushil, et al.
Veröffentlicht: (2026)
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
von: Yu, Zhuoyun, et al.
Veröffentlicht: (2026)
von: Yu, Zhuoyun, et al.
Veröffentlicht: (2026)
DroidSpeak: KV Cache Sharing for Cross-LLM Communication and Multi-LLM Serving
von: Liu, Yuhan, et al.
Veröffentlicht: (2024)
von: Liu, Yuhan, et al.
Veröffentlicht: (2024)
Verification-Aware Planning for Multi-Agent Systems
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
MARCO: Multi-Agent Real-time Chat Orchestration
von: Shrimal, Anubhav, et al.
Veröffentlicht: (2024)
von: Shrimal, Anubhav, et al.
Veröffentlicht: (2024)
Towards Reliable ML Feature Engineering via Planning in Constrained-Topology of LLM Agents
von: Thakur, Himanshu, et al.
Veröffentlicht: (2026)
von: Thakur, Himanshu, et al.
Veröffentlicht: (2026)
Multi-Agent Constraint Factorization Reveals Latent Invariant Solution Structure
von: Scofield, Christopher
Veröffentlicht: (2026)
von: Scofield, Christopher
Veröffentlicht: (2026)
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
von: Sarkar, Bidipta, et al.
Veröffentlicht: (2025)
von: Sarkar, Bidipta, et al.
Veröffentlicht: (2025)
MinionsLLM: a Task-adaptive Framework For The Training and Control of Multi-Agent Systems Through Natural Language
von: Rincon, Andres Garcia, et al.
Veröffentlicht: (2025)
von: Rincon, Andres Garcia, et al.
Veröffentlicht: (2025)
ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
von: Wan, Ziyu, et al.
Veröffentlicht: (2025)
von: Wan, Ziyu, et al.
Veröffentlicht: (2025)
MLZero: A Multi-Agent System for End-to-end Machine Learning Automation
von: Fang, Haoyang, et al.
Veröffentlicht: (2025)
von: Fang, Haoyang, et al.
Veröffentlicht: (2025)
LENS: Learning Ensemble Confidence from Neural States for Multi-LLM Answer Integration
von: Guo, Jizhou
Veröffentlicht: (2025)
von: Guo, Jizhou
Veröffentlicht: (2025)
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings
von: Taylor, Russell, et al.
Veröffentlicht: (2025)
von: Taylor, Russell, et al.
Veröffentlicht: (2025)
Harnessing Multi-Agent LLMs for Complex Engineering Problem-Solving: A Framework for Senior Design Projects
von: Mushtaq, Abdullah, et al.
Veröffentlicht: (2025)
von: Mushtaq, Abdullah, et al.
Veröffentlicht: (2025)
LLM-Based Multi-Agent Blackboard System for Information Discovery in Data Science
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning
von: Lee, Sunwoo, et al.
Veröffentlicht: (2026)
von: Lee, Sunwoo, et al.
Veröffentlicht: (2026)
Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hidden State Poisoning Attacks against Mamba-based Language Models
von: Mercier, Alexandre Le, et al.
Veröffentlicht: (2026) -
An Adversary-Resistant Multi-Agent LLM System via Credibility Scoring
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2025) -
Unleashing Diverse Thinking Modes in LLMs through Multi-Agent Collaboration
von: He, Zhixuan, et al.
Veröffentlicht: (2025) -
MAGIC: A Co-Evolving Attacker-Defender Adversarial Game for Robust LLM Safety
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2026) -
Symphony: A Decentralized Multi-Agent Framework for Scalable Collective Intelligence
von: Wang, Ji, et al.
Veröffentlicht: (2025)