Palisade -- Prompt Injection Detection Framework
Fuente:
arXiv
Guardado en:
| Autores principales: | Kokkula, Sahasra, R, Somanathan, R, Nandavardhan, Aashishkumar, Divya, G |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
por: Gosmar, Diego, et al.
Publicado: (2025)
por: Gosmar, Diego, et al.
Publicado: (2025)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
por: Rahman, Md Abdur, et al.
Publicado: (2024)
por: Rahman, Md Abdur, et al.
Publicado: (2024)
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
por: Ji, Yi, et al.
Publicado: (2025)
por: Ji, Yi, et al.
Publicado: (2025)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
por: Liu, Yinuo, et al.
Publicado: (2025)
por: Liu, Yinuo, et al.
Publicado: (2025)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
por: Wang, Xilong, et al.
Publicado: (2026)
por: Wang, Xilong, et al.
Publicado: (2026)
Prompt Injection as Role Confusion
por: Ye, Charles, et al.
Publicado: (2026)
por: Ye, Charles, et al.
Publicado: (2026)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
por: Kholkar, Gauri, et al.
Publicado: (2025)
por: Kholkar, Gauri, et al.
Publicado: (2025)
Multilingual Hidden Prompt Injection Attacks on LLM-Based Academic Reviewing
por: Theocharopoulos, Panagiotis, et al.
Publicado: (2025)
por: Theocharopoulos, Panagiotis, et al.
Publicado: (2025)
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
por: Yi, Jingwei, et al.
Publicado: (2023)
por: Yi, Jingwei, et al.
Publicado: (2023)
Comparative Analysis of Machine Learning Approaches for Bone Age Assessment: A Comprehensive Study on Three Distinct Models
por: R., Nandavardhan, et al.
Publicado: (2024)
por: R., Nandavardhan, et al.
Publicado: (2024)
Concept-Based Interpretability for Toxicity Detection
por: Garg, Samarth, et al.
Publicado: (2025)
por: Garg, Samarth, et al.
Publicado: (2025)
A General Knowledge Injection Framework for ICD Coding
por: Zhang, Xu, et al.
Publicado: (2025)
por: Zhang, Xu, et al.
Publicado: (2025)
Jatmo: Prompt Injection Defense by Task-Specific Finetuning
por: Piet, Julien, et al.
Publicado: (2023)
por: Piet, Julien, et al.
Publicado: (2023)
ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
por: Li, Hao, et al.
Publicado: (2026)
por: Li, Hao, et al.
Publicado: (2026)
Detecting Errors through Ensembling Prompts (DEEP): An End-to-End LLM Framework for Detecting Factual Errors
por: Chandler, Alex, et al.
Publicado: (2024)
por: Chandler, Alex, et al.
Publicado: (2024)
Semantics as a Shield: Label Disguise Defense (LDD) against Prompt Injection in LLM Sentiment Classification
por: Li, Yanxi, et al.
Publicado: (2025)
por: Li, Yanxi, et al.
Publicado: (2025)
Soft Begging: Modular and Efficient Shielding of LLMs against Prompt Injection and Jailbreaking based on Prompt Tuning
por: Ostermann, Simon, et al.
Publicado: (2024)
por: Ostermann, Simon, et al.
Publicado: (2024)
UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models
por: Lin, Huawei, et al.
Publicado: (2025)
por: Lin, Huawei, et al.
Publicado: (2025)
Goal-guided Generative Prompt Injection Attack on Large Language Models
por: Zhang, Chong, et al.
Publicado: (2024)
por: Zhang, Chong, et al.
Publicado: (2024)
Future-Proofing IoT: Unleashing the Power of AWS Greengrass in Propelling Smart Devices to New Heights
por: Kokkula, Sahasra, et al.
Publicado: (2024)
por: Kokkula, Sahasra, et al.
Publicado: (2024)
PIArena: A Platform for Prompt Injection Evaluation
por: Geng, Runpeng, et al.
Publicado: (2026)
por: Geng, Runpeng, et al.
Publicado: (2026)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
por: Liu, Yupei, et al.
Publicado: (2023)
por: Liu, Yupei, et al.
Publicado: (2023)
Federated Learning Under Temporal Drift -- Mitigating Catastrophic Forgetting via Experience Replay
por: Kokkula, Sahasra, et al.
Publicado: (2026)
por: Kokkula, Sahasra, et al.
Publicado: (2026)
InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
por: Li, Hao, et al.
Publicado: (2024)
por: Li, Hao, et al.
Publicado: (2024)
Prompt Recursive Search: A Living Framework with Adaptive Growth in LLM Auto-Prompting
por: Zhao, Xiangyu, et al.
Publicado: (2024)
por: Zhao, Xiangyu, et al.
Publicado: (2024)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
por: Geng, Runpeng, et al.
Publicado: (2025)
por: Geng, Runpeng, et al.
Publicado: (2025)
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation
por: Sharma, Divyam, et al.
Publicado: (2024)
por: Sharma, Divyam, et al.
Publicado: (2024)
The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure
por: Cook, David A.
Publicado: (2026)
por: Cook, David A.
Publicado: (2026)
Prompt Injection attack against LLM-integrated Applications
por: Liu, Yi, et al.
Publicado: (2023)
por: Liu, Yi, et al.
Publicado: (2023)
Context-Aware Pragmatic Metacognitive Prompting for Sarcasm Detection
por: Iskandardinata, Michael, et al.
Publicado: (2025)
por: Iskandardinata, Michael, et al.
Publicado: (2025)
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
por: Duc, Do Minh, et al.
Publicado: (2025)
por: Duc, Do Minh, et al.
Publicado: (2025)
SA-MDKIF: A Scalable and Adaptable Medical Domain Knowledge Injection Framework for Large Language Models
por: Xu, Tianhan, et al.
Publicado: (2024)
por: Xu, Tianhan, et al.
Publicado: (2024)
WebInject: Prompt Injection Attack to Web Agents
por: Wang, Xilong, et al.
Publicado: (2025)
por: Wang, Xilong, et al.
Publicado: (2025)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
por: Murthy, Rithesh, et al.
Publicado: (2025)
por: Murthy, Rithesh, et al.
Publicado: (2025)
Systematically Analyzing Prompt Injection Vulnerabilities in Diverse LLM Architectures
por: Benjamin, Victoria, et al.
Publicado: (2024)
por: Benjamin, Victoria, et al.
Publicado: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
por: Shao, Zedian, et al.
Publicado: (2024)
por: Shao, Zedian, et al.
Publicado: (2024)
PromptWizard: Task-Aware Prompt Optimization Framework
por: Agarwal, Eshaan, et al.
Publicado: (2024)
por: Agarwal, Eshaan, et al.
Publicado: (2024)
Memorization and Knowledge Injection in Gated LLMs
por: Pan, Xu, et al.
Publicado: (2025)
por: Pan, Xu, et al.
Publicado: (2025)
MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content
por: Guo, Ruoqi, et al.
Publicado: (2026)
por: Guo, Ruoqi, et al.
Publicado: (2026)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
por: Sahoo, Devanshu, et al.
Publicado: (2025)
por: Sahoo, Devanshu, et al.
Publicado: (2025)
Ejemplares similares
-
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
por: Gosmar, Diego, et al.
Publicado: (2025) -
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
por: Rahman, Md Abdur, et al.
Publicado: (2024) -
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
por: Ji, Yi, et al.
Publicado: (2025) -
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
por: Liu, Yinuo, et al.
Publicado: (2025) -
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
por: Wang, Xilong, et al.
Publicado: (2026)