SentinelNet: Safeguarding Multi-Agent Collaboration Through Credit-Based Dynamic Threat Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Yang, Pan, Xudong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2024)
by: Song, Chengyu, et al.
Published: (2024)
StruPhantom: Evolutionary Injection Attacks on Black-Box Tabular Agents Powered by Large Language Models
by: Feng, Yang, et al.
Published: (2025)
by: Feng, Yang, et al.
Published: (2025)
CyberSentinel: An Emergent Threat Detection System for AI Security
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
RPM-Net Reciprocal Point MLP Network for Unknown Network Security Threat Detection
by: Zhang, Jiachen, et al.
Published: (2026)
by: Zhang, Jiachen, et al.
Published: (2026)
Invisible Threats from Model Context Protocol: Generating Stealthy Injection Payload via Tree-based Adaptive Search
by: Shen, Yulin, et al.
Published: (2026)
by: Shen, Yulin, et al.
Published: (2026)
MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory
by: Wang, Yuhui, et al.
Published: (2026)
by: Wang, Yuhui, et al.
Published: (2026)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
by: Pan, Junjun, et al.
Published: (2025)
by: Pan, Junjun, et al.
Published: (2025)
Safeguarding AI Agents: Developing and Analyzing Safety Architectures
by: Domkundwar, Ishaan, et al.
Published: (2024)
by: Domkundwar, Ishaan, et al.
Published: (2024)
CoopGuard: Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Round Attacks
by: Li, Siyuan, et al.
Published: (2026)
by: Li, Siyuan, et al.
Published: (2026)
Integrating Multi-Agent Simulation, Behavioral Forensics, and Trust-Aware Machine Learning for Adaptive Insider Threat Detection
by: Kausar, Firdous, et al.
Published: (2026)
by: Kausar, Firdous, et al.
Published: (2026)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
by: Lin, Zheng, et al.
Published: (2026)
by: Lin, Zheng, et al.
Published: (2026)
AgenticCyber: A GenAI-Powered Multi-Agent System for Multimodal Threat Detection and Adaptive Response in Cybersecurity
by: Roy, Shovan
Published: (2025)
by: Roy, Shovan
Published: (2025)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
by: Wang, Xilong, et al.
Published: (2026)
by: Wang, Xilong, et al.
Published: (2026)
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
D-CIPHER: Dynamic Collaborative Intelligent Multi-Agent System with Planner and Heterogeneous Executors for Offensive Security
by: Udeshi, Meet, et al.
Published: (2025)
by: Udeshi, Meet, et al.
Published: (2025)
CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly
by: Fan, Yihe, et al.
Published: (2026)
by: Fan, Yihe, et al.
Published: (2026)
WebTrap Park: An Automated Platform for Systematic Security Evaluation of Web Agents
by: Wu, Xinyi, et al.
Published: (2026)
by: Wu, Xinyi, et al.
Published: (2026)
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
by: Patil, KrishnaSaiReddy
Published: (2026)
by: Patil, KrishnaSaiReddy
Published: (2026)
Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents
by: Yang, Wenkai, et al.
Published: (2024)
by: Yang, Wenkai, et al.
Published: (2024)
Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions
by: Hu, Qinnan, et al.
Published: (2025)
by: Hu, Qinnan, et al.
Published: (2025)
Backdoor Sentinel: Detecting and Detoxifying Backdoors in Diffusion Models via Temporal Noise Consistency
by: Wang, Bingzheng, et al.
Published: (2026)
by: Wang, Bingzheng, et al.
Published: (2026)
ATAG: AI-Agent Application Threat Assessment with Attack Graphs
by: Gandhi, Parth Atulbhai, et al.
Published: (2025)
by: Gandhi, Parth Atulbhai, et al.
Published: (2025)
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
by: Beurer-Kellner, Luca, et al.
Published: (2026)
by: Beurer-Kellner, Luca, et al.
Published: (2026)
AI Agents Under Threat: A Survey of Key Security Challenges and Future Pathways
by: Deng, Zehang, et al.
Published: (2024)
by: Deng, Zehang, et al.
Published: (2024)
Graph Neural Network Based Adaptive Threat Detection for Cloud Identity and Access Management Logs
by: Madireddy, Venkata Tanuja
Published: (2025)
by: Madireddy, Venkata Tanuja
Published: (2025)
AgentSys: Secure and Dynamic LLM Agents Through Explicit Hierarchical Memory Management
by: Wen, Ruoyao, et al.
Published: (2026)
by: Wen, Ruoyao, et al.
Published: (2026)
Argus: A Multi-Agent Sensitive Information Leakage Detection Framework Based on Hierarchical Reference Relationships
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
Foe for Fraud: Transferable Adversarial Attacks in Credit Card Fraud Detection
by: Fok, Jan Lum, et al.
Published: (2025)
by: Fok, Jan Lum, et al.
Published: (2025)
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
by: Deng, Xinhao, et al.
Published: (2026)
by: Deng, Xinhao, et al.
Published: (2026)
MCP Guardian: A Security-First Layer for Safeguarding MCP-Based AI System
by: Kumar, Sonu, et al.
Published: (2025)
by: Kumar, Sonu, et al.
Published: (2025)
A Formal Security Framework for MCP-Based AI Agents: Threat Taxonomy, Verification Models, and Defense Mechanisms
by: Acharya, Nirajan, et al.
Published: (2026)
by: Acharya, Nirajan, et al.
Published: (2026)
An AI Agent Execution Environment to Safeguard User Data
by: Stanley, Robert, et al.
Published: (2026)
by: Stanley, Robert, et al.
Published: (2026)
On Evaluating the Durability of Safeguards for Open-Weight LLMs
by: Qi, Xiangyu, et al.
Published: (2024)
by: Qi, Xiangyu, et al.
Published: (2024)
Safeguarding Large Language Models: A Survey
by: Dong, Yi, et al.
Published: (2024)
by: Dong, Yi, et al.
Published: (2024)
PromptKeeper: Safeguarding System Prompts for LLMs
by: Jiang, Zhifeng, et al.
Published: (2024)
by: Jiang, Zhifeng, et al.
Published: (2024)
Do Not Merge My Model! Safeguarding Open-Source LLMs Against Unauthorized Model Merging
by: Li, Qinfeng, et al.
Published: (2025)
by: Li, Qinfeng, et al.
Published: (2025)
Insider Threat Detection Using GCN and Bi-LSTM with Explicit and Implicit Graph Representations
by: Yumlembam, Rahul, et al.
Published: (2025)
by: Yumlembam, Rahul, et al.
Published: (2025)
Mobile GUI Agents under Real-world Threats: Are We There Yet?
by: Liu, Guohong, et al.
Published: (2025)
by: Liu, Guohong, et al.
Published: (2025)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
by: Li, Zhiyuan, et al.
Published: (2026)
by: Li, Zhiyuan, et al.
Published: (2026)
WIPI: A New Web Threat for LLM-Driven Web Agents
by: Wu, Fangzhou, et al.
Published: (2024)
by: Wu, Fangzhou, et al.
Published: (2024)
Similar Items
-
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2024) -
StruPhantom: Evolutionary Injection Attacks on Black-Box Tabular Agents Powered by Large Language Models
by: Feng, Yang, et al.
Published: (2025) -
CyberSentinel: An Emergent Threat Detection System for AI Security
by: Tallam, Krti
Published: (2025) -
RPM-Net Reciprocal Point MLP Network for Unknown Network Security Threat Detection
by: Zhang, Jiachen, et al.
Published: (2026) -
Invisible Threats from Model Context Protocol: Generating Stealthy Injection Payload via Tree-based Adaptive Search
by: Shen, Yulin, et al.
Published: (2026)