BlindGuard: Safeguarding LLM-based Multi-Agent Systems under Unknown Attacks
Fuente:
arXiv
Saved in:
| Main Authors: | Miao, Rui, Liu, Yixin, Wang, Yili, Shen, Xu, Tan, Yue, Dai, Yiwei, Pan, Shirui, Wang, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding the Information Propagation Effects of Communication Topologies in LLM-based Multi-Agent Systems
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
Raising the Bar in Graph OOD Generalization: Invariant Learning Beyond Explicit Environment Modeling
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
by: Pan, Junjun, et al.
Published: (2025)
by: Pan, Junjun, et al.
Published: (2025)
Unifying Unsupervised Graph-Level Anomaly Detection and Out-of-Distribution Detection: A Benchmark
by: Wang, Yili, et al.
Published: (2024)
by: Wang, Yili, et al.
Published: (2024)
TrinityGuard: A Unified Framework for Safeguarding Multi-Agent Systems
by: Wang, Kai, et al.
Published: (2026)
by: Wang, Kai, et al.
Published: (2026)
GuardReasoner: Towards Reasoning-based LLM Safeguards
by: Liu, Yue, et al.
Published: (2025)
by: Liu, Yue, et al.
Published: (2025)
GoAgent: Group-of-Agents Communication Topology Generation for LLM-based Multi-Agent Systems
by: Chen, Hongjiang, et al.
Published: (2026)
by: Chen, Hongjiang, et al.
Published: (2026)
Self-Guard: Empower the LLM to Safeguard Itself
by: Wang, Zezhong, et al.
Published: (2023)
by: Wang, Zezhong, et al.
Published: (2023)
Optimizing OOD Detection in Molecular Graphs: A Novel Approach with Diffusion Models
by: Shen, Xu, et al.
Published: (2024)
by: Shen, Xu, et al.
Published: (2024)
GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
by: Xiang, Zhen, et al.
Published: (2024)
by: Xiang, Zhen, et al.
Published: (2024)
Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems
by: Pai, Aaditya
Published: (2026)
by: Pai, Aaditya
Published: (2026)
Dual Mamba for Node-Specific Representation Learning: Tackling Over-Smoothing with Selective State Space Modeling
by: He, Xin, et al.
Published: (2025)
by: He, Xin, et al.
Published: (2025)
CoopGuard: Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Round Attacks
by: Li, Siyuan, et al.
Published: (2026)
by: Li, Siyuan, et al.
Published: (2026)
INFA-Guard: Mitigating Malicious Propagation via Infection-Aware Safeguarding in LLM-Based Multi-Agent Systems
by: Zhou, Yijin, et al.
Published: (2026)
by: Zhou, Yijin, et al.
Published: (2026)
DynHD: Hallucination Detection for Diffusion Large Language Models via Denoising Dynamics Deviation Learning
by: Qian, Yanyu, et al.
Published: (2026)
by: Qian, Yanyu, et al.
Published: (2026)
CIA: Inferring the Communication Topology from LLM-based Multi-Agent Systems
by: Wu, Yongxuan, et al.
Published: (2026)
by: Wu, Yongxuan, et al.
Published: (2026)
G-Safeguard: A Topology-Guided Security Lens and Treatment on LLM-based Multi-agent Systems
by: Wang, Shilong, et al.
Published: (2025)
by: Wang, Shilong, et al.
Published: (2025)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
by: Lin, Lixing, et al.
Published: (2026)
by: Lin, Lixing, et al.
Published: (2026)
GLiGuard: Schema-Conditioned Classification for LLM Safeguard
by: Zaratiana, Urchade, et al.
Published: (2026)
by: Zaratiana, Urchade, et al.
Published: (2026)
HyperD: Hybrid Periodicity Decoupling Framework for Traffic Forecasting
by: Shao, Minlan, et al.
Published: (2025)
by: Shao, Minlan, et al.
Published: (2025)
RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting
by: Jiang, Yilei, et al.
Published: (2024)
by: Jiang, Yilei, et al.
Published: (2024)
LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models
by: Yu, Miao, et al.
Published: (2024)
by: Yu, Miao, et al.
Published: (2024)
Balancing User Preferences by Social Networks: A Condition-Guided Social Recommendation Model for Mitigating Popularity Bias
by: He, Xin, et al.
Published: (2024)
by: He, Xin, et al.
Published: (2024)
GroupGuard: A Framework for Modeling and Defending Collusive Attacks in Multi-Agent Systems
by: Tao, Yiling, et al.
Published: (2026)
by: Tao, Yiling, et al.
Published: (2026)
OFA-MAS: One-for-All Multi-Agent System Topology Design based on Mixture-of-Experts Graph Generative Models
by: Li, Shiyuan, et al.
Published: (2026)
by: Li, Shiyuan, et al.
Published: (2026)
GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning
by: Liu, Yue, et al.
Published: (2025)
by: Liu, Yue, et al.
Published: (2025)
Red-Teaming LLM Multi-Agent Systems via Communication Attacks
by: He, Pengfei, et al.
Published: (2025)
by: He, Pengfei, et al.
Published: (2025)
CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Graph-Augmented Large Language Model Agents: Current Progress and Future Prospects
by: Liu, Yixin, et al.
Published: (2025)
by: Liu, Yixin, et al.
Published: (2025)
Prompt-Unknown Promotion Attacks against LLM-based Sequential Recommender Systems
by: Zhao, Yuchuan, et al.
Published: (2026)
by: Zhao, Yuchuan, et al.
Published: (2026)
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling
by: Zhou, Jialong, et al.
Published: (2025)
by: Zhou, Jialong, et al.
Published: (2025)
AgentGuard: An Attribute-Based Access Control Framework for Tool-Use LLM-Based Agent
by: Luo, Jiaqi, et al.
Published: (2026)
by: Luo, Jiaqi, et al.
Published: (2026)
Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space
by: He, Xin, et al.
Published: (2025)
by: He, Xin, et al.
Published: (2025)
AgentSafe: Safeguarding Large Language Model-based Multi-agent Systems via Hierarchical Data Management
by: Mao, Junyuan, et al.
Published: (2025)
by: Mao, Junyuan, et al.
Published: (2025)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
by: Wu, Yixin, et al.
Published: (2025)
by: Wu, Yixin, et al.
Published: (2025)
DrugPilot: LLM-based Parameterized Reasoning Agent for Drug Discovery
by: Li, Kun, et al.
Published: (2025)
by: Li, Kun, et al.
Published: (2025)
PandaGuard: Systematic Evaluation of LLM Safety against Jailbreaking Attacks
by: Shen, Guobin, et al.
Published: (2025)
by: Shen, Guobin, et al.
Published: (2025)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
by: Zhou, Xueyang, et al.
Published: (2025)
by: Zhou, Xueyang, et al.
Published: (2025)
NExT-Guard: Training-Free Streaming Safeguard without Token-Level Labels
by: Fang, Junfeng, et al.
Published: (2026)
by: Fang, Junfeng, et al.
Published: (2026)
PropGuard: Safeguarding LLM-MAS via Propagation-Aware Exploration and Remediation
by: Yan, Bingyu, et al.
Published: (2026)
by: Yan, Bingyu, et al.
Published: (2026)
Similar Items
-
Understanding the Information Propagation Effects of Communication Topologies in LLM-based Multi-Agent Systems
by: Shen, Xu, et al.
Published: (2025) -
Raising the Bar in Graph OOD Generalization: Invariant Learning Beyond Explicit Environment Modeling
by: Shen, Xu, et al.
Published: (2025) -
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
by: Pan, Junjun, et al.
Published: (2025) -
Unifying Unsupervised Graph-Level Anomaly Detection and Out-of-Distribution Detection: A Benchmark
by: Wang, Yili, et al.
Published: (2024) -
TrinityGuard: A Unified Framework for Safeguarding Multi-Agent Systems
by: Wang, Kai, et al.
Published: (2026)