Salvato in:
| Autori principali: | Pandey, Rohan, Bhujang, Archit |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2605.24421 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
di: Hossain, S M Asif, et al.
Pubblicazione: (2025)
di: Hossain, S M Asif, et al.
Pubblicazione: (2025)
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
di: Paracha, Anum, et al.
Pubblicazione: (2025)
di: Paracha, Anum, et al.
Pubblicazione: (2025)
Secure Retrieval-Augmented Generation against Poisoning Attacks
di: Cheng, Zirui, et al.
Pubblicazione: (2025)
di: Cheng, Zirui, et al.
Pubblicazione: (2025)
Defending Against Indirect Prompt Injection Attacks With Spotlighting
di: Hines, Keegan, et al.
Pubblicazione: (2024)
di: Hines, Keegan, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
Design Patterns for Securing LLM Agents against Prompt Injections
di: Beurer-Kellner, Luca, et al.
Pubblicazione: (2025)
di: Beurer-Kellner, Luca, et al.
Pubblicazione: (2025)
Adversarial Prompt Evaluation: Systematic Benchmarking of Guardrails Against Prompt Input Attacks on LLMs
di: Zizzo, Giulio, et al.
Pubblicazione: (2025)
di: Zizzo, Giulio, et al.
Pubblicazione: (2025)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
di: Liu, Yinuo, et al.
Pubblicazione: (2025)
di: Liu, Yinuo, et al.
Pubblicazione: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
di: Zou, Wei, et al.
Pubblicazione: (2025)
di: Zou, Wei, et al.
Pubblicazione: (2025)
Comments on "Privacy-Enhanced Federated Learning Against Poisoning Adversaries"
di: Schneider, Thomas, et al.
Pubblicazione: (2024)
di: Schneider, Thomas, et al.
Pubblicazione: (2024)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
di: Clop, Cody, et al.
Pubblicazione: (2024)
di: Clop, Cody, et al.
Pubblicazione: (2024)
The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
di: Nasr, Milad, et al.
Pubblicazione: (2025)
di: Nasr, Milad, et al.
Pubblicazione: (2025)
GenTel-Safe: A Unified Benchmark and Shielding Framework for Defending Against Prompt Injection Attacks
di: Li, Rongchang, et al.
Pubblicazione: (2024)
di: Li, Rongchang, et al.
Pubblicazione: (2024)
A Systematic Review of Poisoning Attacks Against Large Language Models
di: Fendley, Neil, et al.
Pubblicazione: (2025)
di: Fendley, Neil, et al.
Pubblicazione: (2025)
Securing Large Language Models (LLMs) from Prompt Injection Attacks
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
LogJack: Indirect Prompt Injection Through Cloud Logs Against LLM Debugging Agents
di: Shah, Harsh
Pubblicazione: (2026)
di: Shah, Harsh
Pubblicazione: (2026)
Exploring Secure Machine Learning Through Payload Injection and FGSM Attacks on ResNet-50
di: Yadav, Umesh, et al.
Pubblicazione: (2025)
di: Yadav, Umesh, et al.
Pubblicazione: (2025)
PoisonedParrot: Subtle Data Poisoning Attacks to Elicit Copyright-Infringing Content from Large Language Models
di: Panaitescu-Liess, Michael-Andrei, et al.
Pubblicazione: (2025)
di: Panaitescu-Liess, Michael-Andrei, et al.
Pubblicazione: (2025)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
di: Rathbun, Ethan, et al.
Pubblicazione: (2024)
di: Rathbun, Ethan, et al.
Pubblicazione: (2024)
Provable Robustness of (Graph) Neural Networks Against Data Poisoning and Backdoor Attacks
di: Gosch, Lukas, et al.
Pubblicazione: (2024)
di: Gosch, Lukas, et al.
Pubblicazione: (2024)
Defending Against Sophisticated Poisoning Attacks with RL-based Aggregation in Federated Learning
di: Wang, Yujing, et al.
Pubblicazione: (2024)
di: Wang, Yujing, et al.
Pubblicazione: (2024)
Online Poisoning Attack Against Reinforcement Learning under Black-box Environments
di: Li, Jianhui, et al.
Pubblicazione: (2024)
di: Li, Jianhui, et al.
Pubblicazione: (2024)
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2024)
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2024)
SecAlign: Defending Against Prompt Injection with Preference Optimization
di: Chen, Sizhe, et al.
Pubblicazione: (2024)
di: Chen, Sizhe, et al.
Pubblicazione: (2024)
Lessons from Defending Gemini Against Indirect Prompt Injections
di: Shi, Chongyang, et al.
Pubblicazione: (2025)
di: Shi, Chongyang, et al.
Pubblicazione: (2025)
LeakSealer: A Semisupervised Defense for LLMs Against Prompt Injection and Leakage Attacks
di: Panebianco, Francesco, et al.
Pubblicazione: (2025)
di: Panebianco, Francesco, et al.
Pubblicazione: (2025)
Traceback of Poisoning Attacks to Retrieval-Augmented Generation
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
Robustness Against Adversarial Attacks via Learning Confined Adversarial Polytopes
di: Hamidi, Shayan Mohajer, et al.
Pubblicazione: (2024)
di: Hamidi, Shayan Mohajer, et al.
Pubblicazione: (2024)
Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks
di: Halloran, John T., et al.
Pubblicazione: (2026)
di: Halloran, John T., et al.
Pubblicazione: (2026)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
di: Xu, Yinglun, et al.
Pubblicazione: (2024)
di: Xu, Yinglun, et al.
Pubblicazione: (2024)
Practical Poisoning Attacks against Retrieval-Augmented Generation
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
di: Zhang, Baolei, et al.
Pubblicazione: (2025)
Transferable Availability Poisoning Attacks
di: Liu, Yiyong, et al.
Pubblicazione: (2023)
di: Liu, Yiyong, et al.
Pubblicazione: (2023)
Secure Aggregation is Not Private Against Membership Inference Attacks
di: Ngo, Khac-Hoang, et al.
Pubblicazione: (2024)
di: Ngo, Khac-Hoang, et al.
Pubblicazione: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
di: Zou, Wei, et al.
Pubblicazione: (2024)
di: Zou, Wei, et al.
Pubblicazione: (2024)
How to Defend Against Large-scale Model Poisoning Attacks in Federated Learning: A Vertical Solution
di: Wang, Jinbo, et al.
Pubblicazione: (2024)
di: Wang, Jinbo, et al.
Pubblicazione: (2024)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
di: Xu, Yuancheng, et al.
Pubblicazione: (2024)
di: Xu, Yuancheng, et al.
Pubblicazione: (2024)
Poison Attacks and Adversarial Prompts Against an Informed University Virtual Assistant
di: Fernandez, Ivan A., et al.
Pubblicazione: (2024)
di: Fernandez, Ivan A., et al.
Pubblicazione: (2024)
Krait: A Backdoor Attack Against Graph Prompt Tuning
di: Song, Ying, et al.
Pubblicazione: (2024)
di: Song, Ying, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
di: Zhan, Qiusi, et al.
Pubblicazione: (2025) -
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
di: Hossain, S M Asif, et al.
Pubblicazione: (2025) -
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
di: Paracha, Anum, et al.
Pubblicazione: (2025) -
Secure Retrieval-Augmented Generation against Poisoning Attacks
di: Cheng, Zirui, et al.
Pubblicazione: (2025) -
Defending Against Indirect Prompt Injection Attacks With Spotlighting
di: Hines, Keegan, et al.
Pubblicazione: (2024)