Palisade -- Prompt Injection Detection Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Kokkula, Sahasra, R, Somanathan, R, Nandavardhan, Aashishkumar, Divya, G |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
by: Gosmar, Diego, et al.
Published: (2025)
by: Gosmar, Diego, et al.
Published: (2025)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
by: Rahman, Md Abdur, et al.
Published: (2024)
by: Rahman, Md Abdur, et al.
Published: (2024)
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
by: Ji, Yi, et al.
Published: (2025)
by: Ji, Yi, et al.
Published: (2025)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
by: Liu, Yinuo, et al.
Published: (2025)
by: Liu, Yinuo, et al.
Published: (2025)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
by: Wang, Xilong, et al.
Published: (2026)
by: Wang, Xilong, et al.
Published: (2026)
Prompt Injection as Role Confusion
by: Ye, Charles, et al.
Published: (2026)
by: Ye, Charles, et al.
Published: (2026)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
by: Kholkar, Gauri, et al.
Published: (2025)
by: Kholkar, Gauri, et al.
Published: (2025)
Multilingual Hidden Prompt Injection Attacks on LLM-Based Academic Reviewing
by: Theocharopoulos, Panagiotis, et al.
Published: (2025)
by: Theocharopoulos, Panagiotis, et al.
Published: (2025)
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
by: Yi, Jingwei, et al.
Published: (2023)
by: Yi, Jingwei, et al.
Published: (2023)
Comparative Analysis of Machine Learning Approaches for Bone Age Assessment: A Comprehensive Study on Three Distinct Models
by: R., Nandavardhan, et al.
Published: (2024)
by: R., Nandavardhan, et al.
Published: (2024)
Concept-Based Interpretability for Toxicity Detection
by: Garg, Samarth, et al.
Published: (2025)
by: Garg, Samarth, et al.
Published: (2025)
A General Knowledge Injection Framework for ICD Coding
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Jatmo: Prompt Injection Defense by Task-Specific Finetuning
by: Piet, Julien, et al.
Published: (2023)
by: Piet, Julien, et al.
Published: (2023)
ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Detecting Errors through Ensembling Prompts (DEEP): An End-to-End LLM Framework for Detecting Factual Errors
by: Chandler, Alex, et al.
Published: (2024)
by: Chandler, Alex, et al.
Published: (2024)
Semantics as a Shield: Label Disguise Defense (LDD) against Prompt Injection in LLM Sentiment Classification
by: Li, Yanxi, et al.
Published: (2025)
by: Li, Yanxi, et al.
Published: (2025)
Soft Begging: Modular and Efficient Shielding of LLMs against Prompt Injection and Jailbreaking based on Prompt Tuning
by: Ostermann, Simon, et al.
Published: (2024)
by: Ostermann, Simon, et al.
Published: (2024)
UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models
by: Lin, Huawei, et al.
Published: (2025)
by: Lin, Huawei, et al.
Published: (2025)
Goal-guided Generative Prompt Injection Attack on Large Language Models
by: Zhang, Chong, et al.
Published: (2024)
by: Zhang, Chong, et al.
Published: (2024)
Future-Proofing IoT: Unleashing the Power of AWS Greengrass in Propelling Smart Devices to New Heights
by: Kokkula, Sahasra, et al.
Published: (2024)
by: Kokkula, Sahasra, et al.
Published: (2024)
PIArena: A Platform for Prompt Injection Evaluation
by: Geng, Runpeng, et al.
Published: (2026)
by: Geng, Runpeng, et al.
Published: (2026)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
by: Liu, Yupei, et al.
Published: (2023)
by: Liu, Yupei, et al.
Published: (2023)
Federated Learning Under Temporal Drift -- Mitigating Catastrophic Forgetting via Experience Replay
by: Kokkula, Sahasra, et al.
Published: (2026)
by: Kokkula, Sahasra, et al.
Published: (2026)
InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Prompt Recursive Search: A Living Framework with Adaptive Growth in LLM Auto-Prompting
by: Zhao, Xiangyu, et al.
Published: (2024)
by: Zhao, Xiangyu, et al.
Published: (2024)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation
by: Sharma, Divyam, et al.
Published: (2024)
by: Sharma, Divyam, et al.
Published: (2024)
The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure
by: Cook, David A.
Published: (2026)
by: Cook, David A.
Published: (2026)
Prompt Injection attack against LLM-integrated Applications
by: Liu, Yi, et al.
Published: (2023)
by: Liu, Yi, et al.
Published: (2023)
Context-Aware Pragmatic Metacognitive Prompting for Sarcasm Detection
by: Iskandardinata, Michael, et al.
Published: (2025)
by: Iskandardinata, Michael, et al.
Published: (2025)
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
by: Duc, Do Minh, et al.
Published: (2025)
by: Duc, Do Minh, et al.
Published: (2025)
SA-MDKIF: A Scalable and Adaptable Medical Domain Knowledge Injection Framework for Large Language Models
by: Xu, Tianhan, et al.
Published: (2024)
by: Xu, Tianhan, et al.
Published: (2024)
WebInject: Prompt Injection Attack to Web Agents
by: Wang, Xilong, et al.
Published: (2025)
by: Wang, Xilong, et al.
Published: (2025)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
by: Murthy, Rithesh, et al.
Published: (2025)
by: Murthy, Rithesh, et al.
Published: (2025)
Systematically Analyzing Prompt Injection Vulnerabilities in Diverse LLM Architectures
by: Benjamin, Victoria, et al.
Published: (2024)
by: Benjamin, Victoria, et al.
Published: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
by: Shao, Zedian, et al.
Published: (2024)
by: Shao, Zedian, et al.
Published: (2024)
PromptWizard: Task-Aware Prompt Optimization Framework
by: Agarwal, Eshaan, et al.
Published: (2024)
by: Agarwal, Eshaan, et al.
Published: (2024)
Memorization and Knowledge Injection in Gated LLMs
by: Pan, Xu, et al.
Published: (2025)
by: Pan, Xu, et al.
Published: (2025)
MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content
by: Guo, Ruoqi, et al.
Published: (2026)
by: Guo, Ruoqi, et al.
Published: (2026)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
by: Sahoo, Devanshu, et al.
Published: (2025)
by: Sahoo, Devanshu, et al.
Published: (2025)
Similar Items
-
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
by: Gosmar, Diego, et al.
Published: (2025) -
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
by: Rahman, Md Abdur, et al.
Published: (2024) -
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
by: Ji, Yi, et al.
Published: (2025) -
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
by: Liu, Yinuo, et al.
Published: (2025) -
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
by: Wang, Xilong, et al.
Published: (2026)