Salvato in:
| Autori principali: | Kokkula, Sahasra, R, Somanathan, R, Nandavardhan, Aashishkumar, Divya, G |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2410.21146 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
di: Gosmar, Diego, et al.
Pubblicazione: (2025)
di: Gosmar, Diego, et al.
Pubblicazione: (2025)
Comparative Analysis of Machine Learning Approaches for Bone Age Assessment: A Comprehensive Study on Three Distinct Models
di: R., Nandavardhan, et al.
Pubblicazione: (2024)
di: R., Nandavardhan, et al.
Pubblicazione: (2024)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
di: Ji, Yi, et al.
Pubblicazione: (2025)
di: Ji, Yi, et al.
Pubblicazione: (2025)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
di: Liu, Yinuo, et al.
Pubblicazione: (2025)
di: Liu, Yinuo, et al.
Pubblicazione: (2025)
Prompt Injection as Role Confusion
di: Ye, Charles, et al.
Pubblicazione: (2026)
di: Ye, Charles, et al.
Pubblicazione: (2026)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
di: Wang, Xilong, et al.
Pubblicazione: (2026)
di: Wang, Xilong, et al.
Pubblicazione: (2026)
Future-Proofing IoT: Unleashing the Power of AWS Greengrass in Propelling Smart Devices to New Heights
di: Kokkula, Sahasra, et al.
Pubblicazione: (2024)
di: Kokkula, Sahasra, et al.
Pubblicazione: (2024)
Federated Learning Under Temporal Drift -- Mitigating Catastrophic Forgetting via Experience Replay
di: Kokkula, Sahasra, et al.
Pubblicazione: (2026)
di: Kokkula, Sahasra, et al.
Pubblicazione: (2026)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
di: Kholkar, Gauri, et al.
Pubblicazione: (2025)
di: Kholkar, Gauri, et al.
Pubblicazione: (2025)
Multilingual Hidden Prompt Injection Attacks on LLM-Based Academic Reviewing
di: Theocharopoulos, Panagiotis, et al.
Pubblicazione: (2025)
di: Theocharopoulos, Panagiotis, et al.
Pubblicazione: (2025)
Concept-Based Interpretability for Toxicity Detection
di: Garg, Samarth, et al.
Pubblicazione: (2025)
di: Garg, Samarth, et al.
Pubblicazione: (2025)
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
di: Yi, Jingwei, et al.
Pubblicazione: (2023)
di: Yi, Jingwei, et al.
Pubblicazione: (2023)
ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
di: Li, Hao, et al.
Pubblicazione: (2026)
di: Li, Hao, et al.
Pubblicazione: (2026)
Jatmo: Prompt Injection Defense by Task-Specific Finetuning
di: Piet, Julien, et al.
Pubblicazione: (2023)
di: Piet, Julien, et al.
Pubblicazione: (2023)
A General Knowledge Injection Framework for ICD Coding
di: Zhang, Xu, et al.
Pubblicazione: (2025)
di: Zhang, Xu, et al.
Pubblicazione: (2025)
Detecting Errors through Ensembling Prompts (DEEP): An End-to-End LLM Framework for Detecting Factual Errors
di: Chandler, Alex, et al.
Pubblicazione: (2024)
di: Chandler, Alex, et al.
Pubblicazione: (2024)
Soft Begging: Modular and Efficient Shielding of LLMs against Prompt Injection and Jailbreaking based on Prompt Tuning
di: Ostermann, Simon, et al.
Pubblicazione: (2024)
di: Ostermann, Simon, et al.
Pubblicazione: (2024)
Semantics as a Shield: Label Disguise Defense (LDD) against Prompt Injection in LLM Sentiment Classification
di: Li, Yanxi, et al.
Pubblicazione: (2025)
di: Li, Yanxi, et al.
Pubblicazione: (2025)
PIArena: A Platform for Prompt Injection Evaluation
di: Geng, Runpeng, et al.
Pubblicazione: (2026)
di: Geng, Runpeng, et al.
Pubblicazione: (2026)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
di: Liu, Yupei, et al.
Pubblicazione: (2023)
di: Liu, Yupei, et al.
Pubblicazione: (2023)
UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models
di: Lin, Huawei, et al.
Pubblicazione: (2025)
di: Lin, Huawei, et al.
Pubblicazione: (2025)
Goal-guided Generative Prompt Injection Attack on Large Language Models
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
di: Geng, Runpeng, et al.
Pubblicazione: (2025)
di: Geng, Runpeng, et al.
Pubblicazione: (2025)
InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Prompt Injection attack against LLM-integrated Applications
di: Liu, Yi, et al.
Pubblicazione: (2023)
di: Liu, Yi, et al.
Pubblicazione: (2023)
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation
di: Sharma, Divyam, et al.
Pubblicazione: (2024)
di: Sharma, Divyam, et al.
Pubblicazione: (2024)
WebInject: Prompt Injection Attack to Web Agents
di: Wang, Xilong, et al.
Pubblicazione: (2025)
di: Wang, Xilong, et al.
Pubblicazione: (2025)
Prompt Recursive Search: A Living Framework with Adaptive Growth in LLM Auto-Prompting
di: Zhao, Xiangyu, et al.
Pubblicazione: (2024)
di: Zhao, Xiangyu, et al.
Pubblicazione: (2024)
Systematically Analyzing Prompt Injection Vulnerabilities in Diverse LLM Architectures
di: Benjamin, Victoria, et al.
Pubblicazione: (2024)
di: Benjamin, Victoria, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure
di: Cook, David A.
Pubblicazione: (2026)
di: Cook, David A.
Pubblicazione: (2026)
PromptWizard: Task-Aware Prompt Optimization Framework
di: Agarwal, Eshaan, et al.
Pubblicazione: (2024)
di: Agarwal, Eshaan, et al.
Pubblicazione: (2024)
Context-Aware Pragmatic Metacognitive Prompting for Sarcasm Detection
di: Iskandardinata, Michael, et al.
Pubblicazione: (2025)
di: Iskandardinata, Michael, et al.
Pubblicazione: (2025)
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
di: Duc, Do Minh, et al.
Pubblicazione: (2025)
di: Duc, Do Minh, et al.
Pubblicazione: (2025)
SA-MDKIF: A Scalable and Adaptable Medical Domain Knowledge Injection Framework for Large Language Models
di: Xu, Tianhan, et al.
Pubblicazione: (2024)
di: Xu, Tianhan, et al.
Pubblicazione: (2024)
MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content
di: Guo, Ruoqi, et al.
Pubblicazione: (2026)
di: Guo, Ruoqi, et al.
Pubblicazione: (2026)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
DAdEE: Unsupervised Domain Adaptation in Early Exit PLMs
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
di: Murthy, Rithesh, et al.
Pubblicazione: (2025)
di: Murthy, Rithesh, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Prompt Injection Detection and Mitigation via AI Multi-Agent NLP Frameworks
di: Gosmar, Diego, et al.
Pubblicazione: (2025) -
Comparative Analysis of Machine Learning Approaches for Bone Age Assessment: A Comprehensive Study on Three Distinct Models
di: R., Nandavardhan, et al.
Pubblicazione: (2024) -
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024) -
Detection Method for Prompt Injection by Integrating Pre-trained Model and Heuristic Feature Engineering
di: Ji, Yi, et al.
Pubblicazione: (2025) -
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
di: Liu, Yinuo, et al.
Pubblicazione: (2025)