Diffusion-Guided Adversarial Perturbation Injection for Generalizable Defense Against Facial Manipulations
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yue, Xue, Linying, Lin, Kaiqing, Quan, Hanyu, Lin, Dongdong, Tian, Hui, Wang, Hongxia, Wang, Bin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Imperceptible Adversarial Defense: A Gradient-Driven Shield against Facial Manipulations
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
Protecting Your Voice: Temporal-aware Robust Watermarking
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
di: Gong, Yuyang, et al.
Pubblicazione: (2026)
di: Gong, Yuyang, et al.
Pubblicazione: (2026)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
di: Cao, Tri, et al.
Pubblicazione: (2026)
di: Cao, Tri, et al.
Pubblicazione: (2026)
Defending Against Prompt Injection With a Few DefensiveTokens
di: Chen, Sizhe, et al.
Pubblicazione: (2025)
di: Chen, Sizhe, et al.
Pubblicazione: (2025)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
di: Wang, Yihan, et al.
Pubblicazione: (2025)
di: Wang, Yihan, et al.
Pubblicazione: (2025)
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
di: Chen, Yulin, et al.
Pubblicazione: (2024)
di: Chen, Yulin, et al.
Pubblicazione: (2024)
TriniMark: A Robust Generative Speech Watermarking Method for Trinity-Level Traceability
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
Behavior-Aware and Generalizable Defense Against Black-Box Adversarial Attacks for ML-Based IDS
di: Ennaji, Sabrine, et al.
Pubblicazione: (2025)
di: Ennaji, Sabrine, et al.
Pubblicazione: (2025)
From Allies to Adversaries: Manipulating LLM Tool-Calling through Adversarial Injection
di: Wang, Haowei, et al.
Pubblicazione: (2024)
di: Wang, Haowei, et al.
Pubblicazione: (2024)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
di: Zhu, Kaijie, et al.
Pubblicazione: (2025)
di: Zhu, Kaijie, et al.
Pubblicazione: (2025)
Prediction Inconsistency Helps Achieve Generalizable Detection of Adversarial Examples
di: Han, Sicong, et al.
Pubblicazione: (2025)
di: Han, Sicong, et al.
Pubblicazione: (2025)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
di: Du, Mengyao, et al.
Pubblicazione: (2026)
di: Du, Mengyao, et al.
Pubblicazione: (2026)
LESSON: Multi-Label Adversarial False Data Injection Attack for Deep Learning Locational Detection
di: Tian, Jiwei, et al.
Pubblicazione: (2024)
di: Tian, Jiwei, et al.
Pubblicazione: (2024)
Struggle with Adversarial Defense? Try Diffusion
di: Li, Yujie, et al.
Pubblicazione: (2024)
di: Li, Yujie, et al.
Pubblicazione: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
di: Brown, Hannah, et al.
Pubblicazione: (2024)
di: Brown, Hannah, et al.
Pubblicazione: (2024)
FedBAP: Backdoor Defense via Benign Adversarial Perturbation in Federated Learning
di: Yan, Xinhai, et al.
Pubblicazione: (2025)
di: Yan, Xinhai, et al.
Pubblicazione: (2025)
Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense
di: Hao, Shuyang, et al.
Pubblicazione: (2025)
di: Hao, Shuyang, et al.
Pubblicazione: (2025)
SLIP: Soft Label Mechanism and Key-Extraction-Guided CoT-based Defense Against Instruction Backdoor in APIs
di: Wu, Zhengxian, et al.
Pubblicazione: (2025)
di: Wu, Zhengxian, et al.
Pubblicazione: (2025)
DUAP: Dual-task Universal Adversarial Perturbations Against Voice Control Systems
di: Sun, Suyang, et al.
Pubblicazione: (2026)
di: Sun, Suyang, et al.
Pubblicazione: (2026)
Covert Communication for Untrusted UAV-Assisted Wireless Systems
di: Gao, Chan, et al.
Pubblicazione: (2024)
di: Gao, Chan, et al.
Pubblicazione: (2024)
Dynamic Dual-level Defense Routing for Continual Adversarial Training
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
Defending Against Prompt Injection with DataFilter
di: Wang, Yizhu, et al.
Pubblicazione: (2025)
di: Wang, Yizhu, et al.
Pubblicazione: (2025)
Prompt Control-Flow Integrity: A Priority-Aware Runtime Defense Against Prompt Injection in LLM Systems
di: Alam, Md Takrim Ul, et al.
Pubblicazione: (2026)
di: Alam, Md Takrim Ul, et al.
Pubblicazione: (2026)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
di: Hossain, S M Asif, et al.
Pubblicazione: (2025)
di: Hossain, S M Asif, et al.
Pubblicazione: (2025)
PP3D: An In-Browser Vision-Based Defense Against Web Behavior Manipulation Attacks
di: King, Spencer, et al.
Pubblicazione: (2025)
di: King, Spencer, et al.
Pubblicazione: (2025)
A Case Study on the Use of Representativeness Bias as a Defense Against Adversarial Cyber Threats
di: Hitaj, Briland, et al.
Pubblicazione: (2025)
di: Hitaj, Briland, et al.
Pubblicazione: (2025)
GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis
di: Liu, Weizhi, et al.
Pubblicazione: (2024)
di: Liu, Weizhi, et al.
Pubblicazione: (2024)
A No-Defense Defense Against Gradient-Based Adversarial Attacks on ML-NIDS: Is Less More?
di: elShehaby, Mohamed, et al.
Pubblicazione: (2026)
di: elShehaby, Mohamed, et al.
Pubblicazione: (2026)
SUAD: Solid-Channel Ultrasound Injection Attack and Defense to Voice Assistants
di: Liu, Chao, et al.
Pubblicazione: (2025)
di: Liu, Chao, et al.
Pubblicazione: (2025)
DYNAMITE: Dynamic Defense Selection for Enhancing Machine Learning-based Intrusion Detection Against Adversarial Attacks
di: Chen, Jing, et al.
Pubblicazione: (2025)
di: Chen, Jing, et al.
Pubblicazione: (2025)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
di: Cheng, Darren, et al.
Pubblicazione: (2026)
di: Cheng, Darren, et al.
Pubblicazione: (2026)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
CycleGANWM: A CycleGAN watermarking method for ownership verification
di: Lin, Dongdong, et al.
Pubblicazione: (2022)
di: Lin, Dongdong, et al.
Pubblicazione: (2022)
Dynamic Probabilistic Noise Injection for Membership Inference Defense
di: Forough, Javad, et al.
Pubblicazione: (2025)
di: Forough, Javad, et al.
Pubblicazione: (2025)
IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
di: An, Hengyu, et al.
Pubblicazione: (2025)
di: An, Hengyu, et al.
Pubblicazione: (2025)
Minimal Cascade Gradient Smoothing for Fast Transferable Preemptive Adversarial Defense
di: Wang, Hanrui, et al.
Pubblicazione: (2024)
di: Wang, Hanrui, et al.
Pubblicazione: (2024)
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs
di: Sekar, Anirudh, et al.
Pubblicazione: (2026)
di: Sekar, Anirudh, et al.
Pubblicazione: (2026)
Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
di: Pasquini, Dario, et al.
Pubblicazione: (2024)
di: Pasquini, Dario, et al.
Pubblicazione: (2024)
I Don't Know You, But I Can Catch You: Real-Time Defense against Diverse Adversarial Patches for Object Detectors
di: Lin, Zijin, et al.
Pubblicazione: (2024)
di: Lin, Zijin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Imperceptible Adversarial Defense: A Gradient-Driven Shield against Facial Manipulations
di: Li, Yue, et al.
Pubblicazione: (2025) -
Protecting Your Voice: Temporal-aware Robust Watermarking
di: Li, Yue, et al.
Pubblicazione: (2025) -
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
di: Gong, Yuyang, et al.
Pubblicazione: (2026) -
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
di: Cao, Tri, et al.
Pubblicazione: (2026) -
Defending Against Prompt Injection With a Few DefensiveTokens
di: Chen, Sizhe, et al.
Pubblicazione: (2025)