A Comparative Evaluation of AI Agent Security Guardrails
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Qi, Li, Jiu, Wei, Pingtao, Xu, Jianjun, Wei, Xueyi, Shi, Jiwei, Zhang, Xuan, Yang, Yanhui, Hui, Xiaodong, Xu, Peng, Zhou, Lingquan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Provably Secure Agent Guardrail
by: Wu, Benlong, et al.
Published: (2026)
by: Wu, Benlong, et al.
Published: (2026)
Beyond Input Guardrails: Reconstructing Cross-Agent Semantic Flows for Execution-Aware Attack Detection
by: Wei, Yangyang, et al.
Published: (2026)
by: Wei, Yangyang, et al.
Published: (2026)
DeepKnown-Guard: A Proprietary Model-Based Safety Response Framework for AI Agents
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
Enhancing Guardrails for Safe and Secure Healthcare AI
by: Gangavarapu, Ananya
Published: (2024)
by: Gangavarapu, Ananya
Published: (2024)
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
by: Zhang, Yixiang, et al.
Published: (2026)
by: Zhang, Yixiang, et al.
Published: (2026)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
by: Li, Xueyi, et al.
Published: (2026)
by: Li, Xueyi, et al.
Published: (2026)
Re-Evaluating EVMBench: Are AI Agents Ready for Smart Contract Security?
by: Peng, Chaoyuan, et al.
Published: (2026)
by: Peng, Chaoyuan, et al.
Published: (2026)
The Trust Paradox in LLM-Based Multi-Agent Systems: When Collaboration Becomes a Security Vulnerability
by: Xu, Zijie, et al.
Published: (2025)
by: Xu, Zijie, et al.
Published: (2025)
Aegis: Towards Governance, Integrity, and Security of AI Voice Agents
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
PoseGuard: Pose-Guided Generation with Safety Guardrails
by: Wang, Kongxin, et al.
Published: (2025)
by: Wang, Kongxin, et al.
Published: (2025)
Toward Web 4.0: Bidirectional Trust between AI Agents and Blockchain
by: Xia, Yunfeng, et al.
Published: (2026)
by: Xia, Yunfeng, et al.
Published: (2026)
Security Analysis of Agentic AI Communication Protocols: A Comparative Evaluation
by: Louck, Yedidel, et al.
Published: (2025)
by: Louck, Yedidel, et al.
Published: (2025)
Investigating red packet fraud in Android applications: Insights from user reviews
by: Cheng, Yu, et al.
Published: (2025)
by: Cheng, Yu, et al.
Published: (2025)
From Thinker to Society: Security in Hierarchical Autonomy Evolution of AI Agents
by: Zhang, Xiaolei, et al.
Published: (2026)
by: Zhang, Xiaolei, et al.
Published: (2026)
CIBER: A Comprehensive Benchmark for Security Evaluation of Code Interpreter Agents
by: Ba, Lei, et al.
Published: (2026)
by: Ba, Lei, et al.
Published: (2026)
A Quantitative Method for Evaluating Security Boundaries in Quantum Key Distribution Combined with Block Ciphers
by: Chen, Xiaoming, et al.
Published: (2025)
by: Chen, Xiaoming, et al.
Published: (2025)
RulePilot: An LLM-Powered Agent for Security Rule Generation
by: Wang, Hongtai, et al.
Published: (2025)
by: Wang, Hongtai, et al.
Published: (2025)
Poly-Guard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
by: Kang, Mintong, et al.
Published: (2025)
by: Kang, Mintong, et al.
Published: (2025)
SecureNT: Smart Topology Obfuscation for Privacy-Aware Network Monitoring
by: Du, Chengze, et al.
Published: (2024)
by: Du, Chengze, et al.
Published: (2024)
Progent: Securing AI Agents with Privilege Control
by: Shi, Tianneng, et al.
Published: (2025)
by: Shi, Tianneng, et al.
Published: (2025)
Proof-of-Guardrail in AI Agents and What (Not) to Trust from It
by: Jin, Xisen, et al.
Published: (2026)
by: Jin, Xisen, et al.
Published: (2026)
OpenGuardrails: A Configurable, Unified, and Scalable Guardrails Platform for Large Language Models
by: Wang, Thomas, et al.
Published: (2025)
by: Wang, Thomas, et al.
Published: (2025)
DeFeed: Secure Decentralized Cross-Contract Data Feed in Web 3.0 for Connected Autonomous Vehicles
by: Sun, Xingchen, et al.
Published: (2025)
by: Sun, Xingchen, et al.
Published: (2025)
The Trojan Knowledge: Bypassing Commercial LLM Guardrails via Harmless Prompt Weaving and Adaptive Tree Search
by: Wei, Rongzhe, et al.
Published: (2025)
by: Wei, Rongzhe, et al.
Published: (2025)
SoK: Evaluating Jailbreak Guardrails for Large Language Models
by: Wang, Xunguang, et al.
Published: (2025)
by: Wang, Xunguang, et al.
Published: (2025)
The Agent Economy: A Blockchain-Based Foundation for Autonomous AI Agents
by: Xu, Minghui
Published: (2026)
by: Xu, Minghui
Published: (2026)
EVMbench: Evaluating AI Agents on Smart Contract Security
by: Wang, Justin, et al.
Published: (2026)
by: Wang, Justin, et al.
Published: (2026)
SecureT2I: No More Unauthorized Manipulation on AI Generated Images from Prompts
by: Wu, Xiaodong, et al.
Published: (2025)
by: Wu, Xiaodong, et al.
Published: (2025)
From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
Kangaroo: A Private and Amortized Inference Framework over WAN for Large-Scale Decision Tree Evaluation
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
A Large Scale Study of AI-based Binary Function Similarity Detection Techniques for Security Researchers and Practitioners
by: Shi, Jingyi, et al.
Published: (2025)
by: Shi, Jingyi, et al.
Published: (2025)
QDBFT: A Dynamic Consensus Algorithm for Quantum-Secured Blockchain
by: Xu, Fei, et al.
Published: (2026)
by: Xu, Fei, et al.
Published: (2026)
Robust-Wide: Robust Watermarking against Instruction-driven Image Editing
by: Hu, Runyi, et al.
Published: (2024)
by: Hu, Runyi, et al.
Published: (2024)
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration
by: Karthikeyan, Harish, et al.
Published: (2025)
by: Karthikeyan, Harish, et al.
Published: (2025)
GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio
by: Zhu, Zhenhao, et al.
Published: (2026)
by: Zhu, Zhenhao, et al.
Published: (2026)
SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents
by: Cheng, Hao, et al.
Published: (2026)
by: Cheng, Hao, et al.
Published: (2026)
Emerging Cyber Attack Risks of Medical AI Agents
by: Qiu, Jianing, et al.
Published: (2025)
by: Qiu, Jianing, et al.
Published: (2025)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
by: Wang, Peiran, et al.
Published: (2026)
by: Wang, Peiran, et al.
Published: (2026)
Similar Items
-
Provably Secure Agent Guardrail
by: Wu, Benlong, et al.
Published: (2026) -
Beyond Input Guardrails: Reconstructing Cross-Agent Semantic Flows for Execution-Aware Attack Detection
by: Wei, Yangyang, et al.
Published: (2026) -
DeepKnown-Guard: A Proprietary Model-Based Safety Response Framework for AI Agents
by: Li, Qi, et al.
Published: (2025) -
Enhancing Guardrails for Safe and Secure Healthcare AI
by: Gangavarapu, Ananya
Published: (2024) -
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
by: Zhang, Yixiang, et al.
Published: (2026)