Towards Unifying Quantitative Security Benchmarking for Multi Agent Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Sharma, Gauri, Kulkarni, Vidhi, King, Miles, Huang, Ken |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
di: Narajala, Vineeth Sai, et al.
Pubblicazione: (2025)
di: Narajala, Vineeth Sai, et al.
Pubblicazione: (2025)
Building A Secure Agentic AI Application Leveraging A2A Protocol
di: Habler, Idan, et al.
Pubblicazione: (2025)
di: Habler, Idan, et al.
Pubblicazione: (2025)
Toward a Unified Security Framework for AI Agents: Trust, Risk, and Liability
di: Mo, Jiayun, et al.
Pubblicazione: (2025)
di: Mo, Jiayun, et al.
Pubblicazione: (2025)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
di: de Witt, Christian Schroeder, et al.
Pubblicazione: (2025)
di: de Witt, Christian Schroeder, et al.
Pubblicazione: (2025)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
SkillTester: Benchmarking Utility and Security of Agent Skills
di: Wang, Leye, et al.
Pubblicazione: (2026)
di: Wang, Leye, et al.
Pubblicazione: (2026)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
di: Li, Zhiyuan, et al.
Pubblicazione: (2026)
di: Li, Zhiyuan, et al.
Pubblicazione: (2026)
Agent Security is a Systems Problem
di: Christodorescu, Mihai, et al.
Pubblicazione: (2026)
di: Christodorescu, Mihai, et al.
Pubblicazione: (2026)
Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark
di: Shao, Minghao, et al.
Pubblicazione: (2025)
di: Shao, Minghao, et al.
Pubblicazione: (2025)
The Authorization-Execution Gap Is a Major Safety and Security Problem in Open-World Agents
di: Wu, Baoyuan, et al.
Pubblicazione: (2026)
di: Wu, Baoyuan, et al.
Pubblicazione: (2026)
Security Concerns for Large Language Models: A Survey
di: Li, Miles Q., et al.
Pubblicazione: (2025)
di: Li, Miles Q., et al.
Pubblicazione: (2025)
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
di: Evtimov, Ivan, et al.
Pubblicazione: (2025)
di: Evtimov, Ivan, et al.
Pubblicazione: (2025)
Measuring Security Without Fooling Ourselves: Why Benchmarking Agents Is Hard
di: Abdelnabi, Sahar, et al.
Pubblicazione: (2026)
di: Abdelnabi, Sahar, et al.
Pubblicazione: (2026)
Agent Audit: A Security Analysis System for LLM Agent Applications
di: Zhang, Haiyue, et al.
Pubblicazione: (2026)
di: Zhang, Haiyue, et al.
Pubblicazione: (2026)
Latent Adversarial Detection: Adaptive Probing of LLM Activations for Multi-Turn Attack Detection
di: Kulkarni, Prashant
Pubblicazione: (2026)
di: Kulkarni, Prashant
Pubblicazione: (2026)
D-CIPHER: Dynamic Collaborative Intelligent Multi-Agent System with Planner and Heterogeneous Executors for Offensive Security
di: Udeshi, Meet, et al.
Pubblicazione: (2025)
di: Udeshi, Meet, et al.
Pubblicazione: (2025)
Towards Secure Retrieval-Augmented Generation: A Comprehensive Review of Threats, Defenses and Benchmarks
di: Mu, Yanming, et al.
Pubblicazione: (2026)
di: Mu, Yanming, et al.
Pubblicazione: (2026)
Enhancing TinyML Security: Study of Adversarial Attack Transferability
di: Shah, Parin, et al.
Pubblicazione: (2024)
di: Shah, Parin, et al.
Pubblicazione: (2024)
Agent Operating Systems (AOS): Integrating Agentic Control Planes into, and Beyond, Traditional Operating Systems
di: Sharma, Ankur, et al.
Pubblicazione: (2026)
di: Sharma, Ankur, et al.
Pubblicazione: (2026)
SecRepoBench: Benchmarking Code Agents for Secure Code Completion in Real-World Repositories
di: Shen, Chihao, et al.
Pubblicazione: (2025)
di: Shen, Chihao, et al.
Pubblicazione: (2025)
MCP Security Bench (MSB): Benchmarking Attacks Against Model Context Protocol in LLM Agents
di: Zhang, Dongsen, et al.
Pubblicazione: (2025)
di: Zhang, Dongsen, et al.
Pubblicazione: (2025)
RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments
di: Fu, Yuchuan, et al.
Pubblicazione: (2025)
di: Fu, Yuchuan, et al.
Pubblicazione: (2025)
Security of AI Agents
di: He, Yifeng, et al.
Pubblicazione: (2024)
di: He, Yifeng, et al.
Pubblicazione: (2024)
Provably Secure Agent Guardrail
di: Wu, Benlong, et al.
Pubblicazione: (2026)
di: Wu, Benlong, et al.
Pubblicazione: (2026)
Joint Optimization of Prompt Security and System Performance in Edge-Cloud LLM Systems
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
di: Yuan, He Yang, et al.
Pubblicazione: (2026)
di: Yuan, He Yang, et al.
Pubblicazione: (2026)
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
di: Patil, KrishnaSaiReddy
Pubblicazione: (2026)
di: Patil, KrishnaSaiReddy
Pubblicazione: (2026)
Security of Internet of Agents: Attacks and Countermeasures
di: Wang, Yuntao, et al.
Pubblicazione: (2025)
di: Wang, Yuntao, et al.
Pubblicazione: (2025)
From Thinker to Society: Security in Hierarchical Autonomy Evolution of AI Agents
di: Zhang, Xiaolei, et al.
Pubblicazione: (2026)
di: Zhang, Xiaolei, et al.
Pubblicazione: (2026)
Towards Understanding and Enhancing Security of Proof-of-Training for DNN Model Ownership Verification
di: Chang, Yijia, et al.
Pubblicazione: (2024)
di: Chang, Yijia, et al.
Pubblicazione: (2024)
The PBSAI Governance Ecosystem: A Multi-Agent AI Reference Architecture for Securing Enterprise AI Estates
di: Willis, John M.
Pubblicazione: (2026)
di: Willis, John M.
Pubblicazione: (2026)
Agent-Fence: Mapping Security Vulnerabilities Across Deep Research Agents
di: Puppala, Sai, et al.
Pubblicazione: (2026)
di: Puppala, Sai, et al.
Pubblicazione: (2026)
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
di: Zhang, Yixiang, et al.
Pubblicazione: (2026)
di: Zhang, Yixiang, et al.
Pubblicazione: (2026)
LLM Agents Should Employ Security Principles
di: Zhang, Kaiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Kaiyuan, et al.
Pubblicazione: (2025)
Securing AI Agents with Information-Flow Control
di: Costa, Manuel, et al.
Pubblicazione: (2025)
di: Costa, Manuel, et al.
Pubblicazione: (2025)
Progent: Securing AI Agents with Privilege Control
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
Mind the Web: The Security of Web Use Agents
di: Shapira, Avishag, et al.
Pubblicazione: (2025)
di: Shapira, Avishag, et al.
Pubblicazione: (2025)
A Framework for Formalizing LLM Agent Security
di: Siu, Vincent, et al.
Pubblicazione: (2026)
di: Siu, Vincent, et al.
Pubblicazione: (2026)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
di: Xiang, Chong, et al.
Pubblicazione: (2026)
di: Xiang, Chong, et al.
Pubblicazione: (2026)
The System Prompt Is the Attack Surface: How LLM Agent Configuration Shapes Security and Creates Exploitable Vulnerabilities
di: Litvak, Ron
Pubblicazione: (2026)
di: Litvak, Ron
Pubblicazione: (2026)
Documenti analoghi
-
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
di: Narajala, Vineeth Sai, et al.
Pubblicazione: (2025) -
Building A Secure Agentic AI Application Leveraging A2A Protocol
di: Habler, Idan, et al.
Pubblicazione: (2025) -
Toward a Unified Security Framework for AI Agents: Trust, Risk, and Liability
di: Mo, Jiayun, et al.
Pubblicazione: (2025) -
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
di: de Witt, Christian Schroeder, et al.
Pubblicazione: (2025) -
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)