F2A: An Innovative Approach for Prompt Injection by Utilizing Feign Security Detection Agents
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Ren, Yupeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Securing AI Agents Against Prompt Injection Attacks
von: Ramakrishnan, Badrinath, et al.
Veröffentlicht: (2025)
von: Ramakrishnan, Badrinath, et al.
Veröffentlicht: (2025)
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025)
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025)
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
von: Xiang, Chong, et al.
Veröffentlicht: (2026)
von: Xiang, Chong, et al.
Veröffentlicht: (2026)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
von: Maloyan, Narek, et al.
Veröffentlicht: (2026)
von: Maloyan, Narek, et al.
Veröffentlicht: (2026)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
Defending against Indirect Prompt Injection by Instruction Detection
von: Wen, Tongyu, et al.
Veröffentlicht: (2025)
von: Wen, Tongyu, et al.
Veröffentlicht: (2025)
The Cognitive Firewall:Securing Browser Based AI Agents Against Indirect Prompt Injection Via Hybrid Edge Cloud Defense
von: Lan, Qianlong, et al.
Veröffentlicht: (2026)
von: Lan, Qianlong, et al.
Veröffentlicht: (2026)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
Signed-Prompt: A New Approach to Prevent Prompt Injection Attacks Against LLM-Integrated Applications
von: Suo, Xuchen
Veröffentlicht: (2024)
von: Suo, Xuchen
Veröffentlicht: (2024)
VPI-Bench: Visual Prompt Injection Attacks for Computer-Use Agents
von: Cao, Tri, et al.
Veröffentlicht: (2025)
von: Cao, Tri, et al.
Veröffentlicht: (2025)
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
von: Zhong, Peter Yong, et al.
Veröffentlicht: (2025)
von: Zhong, Peter Yong, et al.
Veröffentlicht: (2025)
QueryIPI: Query-agnostic Indirect Prompt Injection on Coding Agents
von: Xie, Yuchong, et al.
Veröffentlicht: (2025)
von: Xie, Yuchong, et al.
Veröffentlicht: (2025)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
von: Cao, Tri, et al.
Veröffentlicht: (2026)
von: Cao, Tri, et al.
Veröffentlicht: (2026)
Bypassing Prompt Injection Detectors through Evasive Injections
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
AgentTypo: Adaptive Typographic Prompt Injection Attacks against Black-box Multimodal Agents
von: Li, Yanjie, et al.
Veröffentlicht: (2025)
von: Li, Yanjie, et al.
Veröffentlicht: (2025)
PromptLocate: Localizing Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection
von: Zhao, Lei, et al.
Veröffentlicht: (2026)
von: Zhao, Lei, et al.
Veröffentlicht: (2026)
ICON: Indirect Prompt Injection Defense for Agents based on Inference-Time Correction
von: Wang, Che, et al.
Veröffentlicht: (2026)
von: Wang, Che, et al.
Veröffentlicht: (2026)
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
Defeating Prompt Injections by Design
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
AgentVigil: Generic Black-Box Red-teaming for Indirect Prompt Injection against LLM Agents
von: Wang, Zhun, et al.
Veröffentlicht: (2025)
von: Wang, Zhun, et al.
Veröffentlicht: (2025)
Prompt Injection 2.0: Hybrid AI Threats
von: McHugh, Jeremy, et al.
Veröffentlicht: (2025)
von: McHugh, Jeremy, et al.
Veröffentlicht: (2025)
Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection
von: Debi, Tanusree, et al.
Veröffentlicht: (2026)
von: Debi, Tanusree, et al.
Veröffentlicht: (2026)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
PromptArmor: Simple yet Effective Prompt Injection Defenses
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
CASCADE: A Cascaded Hybrid Defense Architecture for Prompt Injection Detection in MCP-Based Systems
von: Turgut, İpek Abasıkeleş, et al.
Veröffentlicht: (2026)
von: Turgut, İpek Abasıkeleş, et al.
Veröffentlicht: (2026)
Trust No AI: Prompt Injection Along The CIA Security Triad
von: Rehberger, Johann
Veröffentlicht: (2024)
von: Rehberger, Johann
Veröffentlicht: (2024)
Manipulating LLM Web Agents with Indirect Prompt Injection Attack via HTML Accessibility Tree
von: Johnson, Sam, et al.
Veröffentlicht: (2025)
von: Johnson, Sam, et al.
Veröffentlicht: (2025)
MUZZLE: Adaptive Agentic Red-Teaming of Web Agents Against Indirect Prompt Injection Attacks
von: Syros, Georgios, et al.
Veröffentlicht: (2026)
von: Syros, Georgios, et al.
Veröffentlicht: (2026)
Silent Egress: When Implicit Prompt Injection Makes LLM Agents Leak Without a Trace
von: Lan, Qianlong, et al.
Veröffentlicht: (2026)
von: Lan, Qianlong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Securing AI Agents Against Prompt Injection Attacks
von: Ramakrishnan, Badrinath, et al.
Veröffentlicht: (2025) -
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025) -
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
von: Du, Mengyao, et al.
Veröffentlicht: (2026) -
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
von: Zhao, Wei, et al.
Veröffentlicht: (2026) -
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
von: Xiang, Chong, et al.
Veröffentlicht: (2026)