Knowledge Return Oriented Prompting (KROP)
Fuente:
arXiv
Saved in:
| Main Authors: | Martin, Jason, Yeung, Kenneth |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TokenBreak: Bypassing Text Classification Models Through Token Manipulation
by: Schulz, Kasimir, et al.
Published: (2025)
by: Schulz, Kasimir, et al.
Published: (2025)
zkLLM: Zero Knowledge Proofs for Large Language Models
by: Sun, Haochen, et al.
Published: (2024)
by: Sun, Haochen, et al.
Published: (2024)
Neural Exec: Learning (and Learning from) Execution Triggers for Prompt Injection Attacks
by: Pasquini, Dario, et al.
Published: (2024)
by: Pasquini, Dario, et al.
Published: (2024)
Bypassing Prompt Guards in Production with Controlled-Release Prompting
by: Fairoze, Jaiden, et al.
Published: (2025)
by: Fairoze, Jaiden, et al.
Published: (2025)
Leveraging Soft Prompts for Privacy Attacks in Federated Prompt Tuning
by: Nguyen, Quan Minh, et al.
Published: (2026)
by: Nguyen, Quan Minh, et al.
Published: (2026)
DPack: Efficiency-Oriented Privacy Budget Scheduling
by: Tholoniat, Pierre, et al.
Published: (2022)
by: Tholoniat, Pierre, et al.
Published: (2022)
Are You Using Reliable Graph Prompts? Trojan Prompt Attacks on Graph Neural Networks
by: Lin, Minhua, et al.
Published: (2024)
by: Lin, Minhua, et al.
Published: (2024)
Adversarial Prompt Evaluation: Systematic Benchmarking of Guardrails Against Prompt Input Attacks on LLMs
by: Zizzo, Giulio, et al.
Published: (2025)
by: Zizzo, Giulio, et al.
Published: (2025)
Prompt Obfuscation for Large Language Models
by: Pape, David, et al.
Published: (2024)
by: Pape, David, et al.
Published: (2024)
JANUS: A Difference-Oriented Analyzer For Financial Centralization Risks in Smart Contracts
by: Wang, Wansen, et al.
Published: (2024)
by: Wang, Wansen, et al.
Published: (2024)
Truth in Text: A Meta-Analysis of ML-Based Cyber Information Influence Detection Approaches
by: Pittman, Jason M.
Published: (2025)
by: Pittman, Jason M.
Published: (2025)
Stealix: Model Stealing via Prompt Evolution
by: Zhuang, Zhixiong, et al.
Published: (2025)
by: Zhuang, Zhixiong, et al.
Published: (2025)
HoGS: Homophily-Oriented Graph Synthesis for Local Differentially Private GNN Training
by: Xu, Wen, et al.
Published: (2026)
by: Xu, Wen, et al.
Published: (2026)
An Early Categorization of Prompt Injection Attacks on Large Language Models
by: Rossi, Sippo, et al.
Published: (2024)
by: Rossi, Sippo, et al.
Published: (2024)
Defending Jailbreak Prompts via In-Context Adversarial Game
by: Zhou, Yujun, et al.
Published: (2024)
by: Zhou, Yujun, et al.
Published: (2024)
Preventing Prompt Injection with Type-Directed Privilege Separation
by: Jacob, Dennis, et al.
Published: (2025)
by: Jacob, Dennis, et al.
Published: (2025)
Pr$εε$mpt: Sanitizing Sensitive Prompts for LLMs
by: Chowdhury, Amrita Roy, et al.
Published: (2025)
by: Chowdhury, Amrita Roy, et al.
Published: (2025)
SecAlign: Defending Against Prompt Injection with Preference Optimization
by: Chen, Sizhe, et al.
Published: (2024)
by: Chen, Sizhe, et al.
Published: (2024)
Krait: A Backdoor Attack Against Graph Prompt Tuning
by: Song, Ying, et al.
Published: (2024)
by: Song, Ying, et al.
Published: (2024)
Cape: Context-Aware Prompt Perturbation Mechanism with Differential Privacy
by: Wu, Haoqi, et al.
Published: (2025)
by: Wu, Haoqi, et al.
Published: (2025)
Death by a Thousand Prompts: Open Model Vulnerability Analysis
by: Chang, Amy, et al.
Published: (2025)
by: Chang, Amy, et al.
Published: (2025)
Design Patterns for Securing LLM Agents against Prompt Injections
by: Beurer-Kellner, Luca, et al.
Published: (2025)
by: Beurer-Kellner, Luca, et al.
Published: (2025)
CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs
by: Fahey, Ryan
Published: (2026)
by: Fahey, Ryan
Published: (2026)
Lessons from Defending Gemini Against Indirect Prompt Injections
by: Shi, Chongyang, et al.
Published: (2025)
by: Shi, Chongyang, et al.
Published: (2025)
Prompt Stealing Attacks Against Text-to-Image Generation Models
by: Shen, Xinyue, et al.
Published: (2023)
by: Shen, Xinyue, et al.
Published: (2023)
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
by: Xiang, Zhen, et al.
Published: (2024)
by: Xiang, Zhen, et al.
Published: (2024)
LLM Security and Safety: Insights from Homotopy-Inspired Prompt Obfuscation
by: Lazo, Luis, et al.
Published: (2026)
by: Lazo, Luis, et al.
Published: (2026)
Efficient and Adaptable Detection of Malicious LLM Prompts via Bootstrap Aggregation
by: Hassan, Shayan Ali, et al.
Published: (2026)
by: Hassan, Shayan Ali, et al.
Published: (2026)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
by: Yin, Chenlong, et al.
Published: (2026)
by: Yin, Chenlong, et al.
Published: (2026)
Mitigating Indirect Prompt Injection via Instruction-Following Intent Analysis
by: Kang, Mintong, et al.
Published: (2025)
by: Kang, Mintong, et al.
Published: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
by: Zou, Wei, et al.
Published: (2025)
by: Zou, Wei, et al.
Published: (2025)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
by: Hossain, S M Asif, et al.
Published: (2025)
by: Hossain, S M Asif, et al.
Published: (2025)
Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs
by: Lu, Yiyang, et al.
Published: (2026)
by: Lu, Yiyang, et al.
Published: (2026)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
by: Clop, Cody, et al.
Published: (2024)
by: Clop, Cody, et al.
Published: (2024)
AutoAdv: Automated Adversarial Prompting for Multi-Turn Jailbreaking of Large Language Models
by: Reddy, Aashray, et al.
Published: (2025)
by: Reddy, Aashray, et al.
Published: (2025)
Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills
by: Hsu, Chia-Yi, et al.
Published: (2026)
by: Hsu, Chia-Yi, et al.
Published: (2026)
"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
by: Shen, Xinyue, et al.
Published: (2023)
by: Shen, Xinyue, et al.
Published: (2023)
Persona-Conditioned Adversarial Prompting: Multi-Identity Red-Teaming for Adversarial Discovery and Mitigation
by: Morasso, Cristian, et al.
Published: (2026)
by: Morasso, Cristian, et al.
Published: (2026)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
by: Zhan, Qiusi, et al.
Published: (2025)
by: Zhan, Qiusi, et al.
Published: (2025)
Similar Items
-
TokenBreak: Bypassing Text Classification Models Through Token Manipulation
by: Schulz, Kasimir, et al.
Published: (2025) -
zkLLM: Zero Knowledge Proofs for Large Language Models
by: Sun, Haochen, et al.
Published: (2024) -
Neural Exec: Learning (and Learning from) Execution Triggers for Prompt Injection Attacks
by: Pasquini, Dario, et al.
Published: (2024) -
Bypassing Prompt Guards in Production with Controlled-Release Prompting
by: Fairoze, Jaiden, et al.
Published: (2025) -
Leveraging Soft Prompts for Privacy Attacks in Federated Prompt Tuning
by: Nguyen, Quan Minh, et al.
Published: (2026)