PROMPTFUZZ: Harnessing Fuzzing Techniques for Robust Testing of Prompt Injection in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Jiahao, Shao, Yangguang, Miao, Hanwen, Shi, Junzheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
In-Browser LLM-Guided Fuzzing for Real-Time Prompt Injection Testing in Agentic AI Browsers
di: Cohen, Avihay
Pubblicazione: (2025)
di: Cohen, Avihay
Pubblicazione: (2025)
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
di: Cai, Jiacheng, et al.
Pubblicazione: (2024)
di: Cai, Jiacheng, et al.
Pubblicazione: (2024)
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
di: Guo, Xuyang, et al.
Pubblicazione: (2025)
di: Guo, Xuyang, et al.
Pubblicazione: (2025)
Assessing Prompt Injection Risks in 200+ Custom GPTs
di: Yu, Jiahao, et al.
Pubblicazione: (2023)
di: Yu, Jiahao, et al.
Pubblicazione: (2023)
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs
di: Gong, Xueluan, et al.
Pubblicazione: (2024)
di: Gong, Xueluan, et al.
Pubblicazione: (2024)
PromptLocate: Localizing Prompt Injection Attacks
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
Analysis of LLMs Against Prompt Injection and Jailbreak Attacks
di: Jaiswal, Piyush, et al.
Pubblicazione: (2026)
di: Jaiswal, Piyush, et al.
Pubblicazione: (2026)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
di: Yeo, Andrew, et al.
Pubblicazione: (2025)
di: Yeo, Andrew, et al.
Pubblicazione: (2025)
Defeating Prompt Injections by Design
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2025)
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2025)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
di: Cao, Tri, et al.
Pubblicazione: (2026)
di: Cao, Tri, et al.
Pubblicazione: (2026)
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
di: Du, Mengyao, et al.
Pubblicazione: (2026)
di: Du, Mengyao, et al.
Pubblicazione: (2026)
A Critical Evaluation of Defenses against Prompt Injection Attacks
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
Attention Tracker: Detecting Prompt Injection Attacks in LLMs
di: Hung, Kuo-Han, et al.
Pubblicazione: (2024)
di: Hung, Kuo-Han, et al.
Pubblicazione: (2024)
PromptArmor: Simple yet Effective Prompt Injection Defenses
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
Bypassing Prompt Injection Detectors through Evasive Injections
di: Rahman, Md Jahedur, et al.
Pubblicazione: (2026)
di: Rahman, Md Jahedur, et al.
Pubblicazione: (2026)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
Decoding Latent Attack Surfaces in LLMs: Prompt Injection via HTML in Web Summarization
di: Verma, Ishaan, et al.
Pubblicazione: (2025)
di: Verma, Ishaan, et al.
Pubblicazione: (2025)
AdapTools: Adaptive Tool-based Indirect Prompt Injection Attacks on Agentic LLMs
di: Wang, Che, et al.
Pubblicazione: (2026)
di: Wang, Che, et al.
Pubblicazione: (2026)
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
di: Shi, Jiawen, et al.
Pubblicazione: (2024)
di: Shi, Jiawen, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
Locus: Agentic Predicate Synthesis for Directed Fuzzing
di: Zhu, Jie, et al.
Pubblicazione: (2025)
di: Zhu, Jie, et al.
Pubblicazione: (2025)
Overcoming the Retrieval Barrier: Indirect Prompt Injection in the Wild for LLM Systems
di: Chang, Hongyan, et al.
Pubblicazione: (2026)
di: Chang, Hongyan, et al.
Pubblicazione: (2026)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
di: Wang, Zhilong, et al.
Pubblicazione: (2025)
di: Wang, Zhilong, et al.
Pubblicazione: (2025)
Evaluation of Prompt Injection Defenses in Large Language Models
di: Deep, Priyal, et al.
Pubblicazione: (2026)
di: Deep, Priyal, et al.
Pubblicazione: (2026)
Prompt Injection 2.0: Hybrid AI Threats
di: McHugh, Jeremy, et al.
Pubblicazione: (2025)
di: McHugh, Jeremy, et al.
Pubblicazione: (2025)
Securing AI Agents Against Prompt Injection Attacks
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
Defending against Indirect Prompt Injection by Instruction Detection
di: Wen, Tongyu, et al.
Pubblicazione: (2025)
di: Wen, Tongyu, et al.
Pubblicazione: (2025)
LLMpatronous: Harnessing the Power of LLMs For Vulnerability Detection
di: Yarra, Rajesh
Pubblicazione: (2025)
di: Yarra, Rajesh
Pubblicazione: (2025)
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
di: Bhatt, Manish, et al.
Pubblicazione: (2026)
di: Bhatt, Manish, et al.
Pubblicazione: (2026)
CourtGuard: A Local, Multiagent Prompt Injection Classifier
di: Wu, Isaac, et al.
Pubblicazione: (2025)
di: Wu, Isaac, et al.
Pubblicazione: (2025)
Token-Efficient Prompt Injection Attack: Provoking Cessation in LLM Reasoning via Adaptive Token Compression
di: Cui, Yu, et al.
Pubblicazione: (2025)
di: Cui, Yu, et al.
Pubblicazione: (2025)
Soft Begging: Modular and Efficient Shielding of LLMs against Prompt Injection and Jailbreaking based on Prompt Tuning
di: Ostermann, Simon, et al.
Pubblicazione: (2024)
di: Ostermann, Simon, et al.
Pubblicazione: (2024)
Know Thy Enemy: Securing LLMs Against Prompt Injection via Diverse Data Synthesis and Instruction-Level Chain-of-Thought Learning
di: Chang, Zhiyuan, et al.
Pubblicazione: (2026)
di: Chang, Zhiyuan, et al.
Pubblicazione: (2026)
AgentVigil: Generic Black-Box Red-teaming for Indirect Prompt Injection against LLM Agents
di: Wang, Zhun, et al.
Pubblicazione: (2025)
di: Wang, Zhun, et al.
Pubblicazione: (2025)
PromptKeeper: Safeguarding System Prompts for LLMs
di: Jiang, Zhifeng, et al.
Pubblicazione: (2024)
di: Jiang, Zhifeng, et al.
Pubblicazione: (2024)
Signed-Prompt: A New Approach to Prevent Prompt Injection Attacks Against LLM-Integrated Applications
di: Suo, Xuchen
Pubblicazione: (2024)
di: Suo, Xuchen
Pubblicazione: (2024)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
di: Liu, Yupei, et al.
Pubblicazione: (2025)
di: Liu, Yupei, et al.
Pubblicazione: (2025)
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
di: Liu, Yupei, et al.
Pubblicazione: (2025)
di: Liu, Yupei, et al.
Pubblicazione: (2025)
VPI-Bench: Visual Prompt Injection Attacks for Computer-Use Agents
di: Cao, Tri, et al.
Pubblicazione: (2025)
di: Cao, Tri, et al.
Pubblicazione: (2025)
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
di: Zhong, Peter Yong, et al.
Pubblicazione: (2025)
di: Zhong, Peter Yong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
In-Browser LLM-Guided Fuzzing for Real-Time Prompt Injection Testing in Agentic AI Browsers
di: Cohen, Avihay
Pubblicazione: (2025) -
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
di: Cai, Jiacheng, et al.
Pubblicazione: (2024) -
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
di: Guo, Xuyang, et al.
Pubblicazione: (2025) -
Assessing Prompt Injection Risks in 200+ Custom GPTs
di: Yu, Jiahao, et al.
Pubblicazione: (2023) -
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs
di: Gong, Xueluan, et al.
Pubblicazione: (2024)