Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
Fuente:
arXiv
Saved in:
| Main Authors: | Jha, Abha, Aravindan, Ashwath Vaithinathan, Salaway, Matthew, Bhide, Atharva Sandeep, Yaldiz, Duygu Nur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
Do VLMs Have Bad Eyes? Diagnosing Compositional Failures via Mechanistic Interpretability
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025)
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024)
by: Zhang, Yimeng, et al.
Published: (2024)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
SATBA: An Invisible Backdoor Attack Based On Spatial Attention
by: Zhou, Huasong, et al.
Published: (2023)
by: Zhou, Huasong, et al.
Published: (2023)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling
by: Li, Zida, et al.
Published: (2026)
by: Li, Zida, et al.
Published: (2026)
Unveiling and Mitigating Backdoor Vulnerabilities based on Unlearning Weight Changes and Backdoor Activeness
by: Lin, Weilin, et al.
Published: (2024)
by: Lin, Weilin, et al.
Published: (2024)
Nearest is Not Dearest: Towards Practical Defense against Quantization-conditioned Backdoor Attacks
by: Li, Boheng, et al.
Published: (2024)
by: Li, Boheng, et al.
Published: (2024)
Test-Time Attention Purification for Backdoored Large Vision Language Models
by: Zhang, Zhifang, et al.
Published: (2026)
by: Zhang, Zhifang, et al.
Published: (2026)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
by: Chen, Yukun, et al.
Published: (2025)
by: Chen, Yukun, et al.
Published: (2025)
Efficient Backdoor Defense in Multimodal Contrastive Learning: A Token-Level Unlearning Method for Mitigating Threats
by: Liu, Kuanrong, et al.
Published: (2024)
by: Liu, Kuanrong, et al.
Published: (2024)
Gungnir: Exploiting Stylistic Features in Images for Backdoor Attacks on Diffusion Models
by: Zhang, Lei, et al.
Published: (2025)
by: Zhang, Lei, et al.
Published: (2025)
REDEditing: Relationship-Driven Precise Backdoor Poisoning on Text-to-Image Diffusion Models
by: Guo, Chongye, et al.
Published: (2025)
by: Guo, Chongye, et al.
Published: (2025)
Invisible Backdoor Attacks on Diffusion Models
by: Li, Sen, et al.
Published: (2024)
by: Li, Sen, et al.
Published: (2024)
Struggle with Adversarial Defense? Try Diffusion
by: Li, Yujie, et al.
Published: (2024)
by: Li, Yujie, et al.
Published: (2024)
Model X-ray:Detecting Backdoored Models via Decision Boundary
by: Su, Yanghao, et al.
Published: (2024)
by: Su, Yanghao, et al.
Published: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
by: Xu, Yukai, et al.
Published: (2024)
by: Xu, Yukai, et al.
Published: (2024)
BadBlocks: Low-Cost and Stealthy Backdoor Attacks Tailored for Text-to-Image Diffusion Models
by: Wu, Jia, et al.
Published: (2025)
by: Wu, Jia, et al.
Published: (2025)
VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models
by: Chou, Sheng-Yen, et al.
Published: (2023)
by: Chou, Sheng-Yen, et al.
Published: (2023)
Revocable Backdoor for Deep Model Trading
by: Xu, Yiran, et al.
Published: (2024)
by: Xu, Yiran, et al.
Published: (2024)
Removing the Trigger, Not the Backdoor: Alternative Triggers and Latent Backdoors
by: Abad, Gorka, et al.
Published: (2026)
by: Abad, Gorka, et al.
Published: (2026)
Training-Free Color-Aware Adversarial Diffusion Sanitization for Diffusion Stegomalware Defense at Security Gateways
by: Frants, Vladimir, et al.
Published: (2025)
by: Frants, Vladimir, et al.
Published: (2025)
DisDet: Exploring Detectability of Backdoor Attack on Diffusion Models
by: Sui, Yang, et al.
Published: (2024)
by: Sui, Yang, et al.
Published: (2024)
Revealing Vulnerabilities of Neural Networks in Parameter Learning and Defense Against Explanation-Aware Backdoors
by: Kadir, Md Abdul, et al.
Published: (2024)
by: Kadir, Md Abdul, et al.
Published: (2024)
Backdoor Attack Against Vision Transformers via Attention Gradient-Based Image Erosion
by: Guo, Ji, et al.
Published: (2024)
by: Guo, Ji, et al.
Published: (2024)
Backdoor Attacks against No-Reference Image Quality Assessment Models via a Scalable Trigger
by: Yu, Yi, et al.
Published: (2024)
by: Yu, Yi, et al.
Published: (2024)
Unstable Unlearning: The Hidden Risk of Concept Resurgence in Diffusion Models
by: Suriyakumar, Vinith M., et al.
Published: (2024)
by: Suriyakumar, Vinith M., et al.
Published: (2024)
Backdoor Mitigation in Object Detection via Adversarial Fine-Tuning
by: Dunnett, Kealan, et al.
Published: (2026)
by: Dunnett, Kealan, et al.
Published: (2026)
UNIT: Backdoor Mitigation via Automated Neural Distribution Tightening
by: Cheng, Siyuan, et al.
Published: (2024)
by: Cheng, Siyuan, et al.
Published: (2024)
UnlearnShield: Shielding Forgotten Privacy against Unlearning Inversion
by: Xue, Lulu, et al.
Published: (2026)
by: Xue, Lulu, et al.
Published: (2026)
HoneypotNet: Backdoor Attacks Against Model Extraction
by: Wang, Yixu, et al.
Published: (2025)
by: Wang, Yixu, et al.
Published: (2025)
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
by: Gao, Hongcheng, et al.
Published: (2024)
by: Gao, Hongcheng, et al.
Published: (2024)
FedDefender: Backdoor Attack Defense in Federated Learning
by: Gill, Waris, et al.
Published: (2023)
by: Gill, Waris, et al.
Published: (2023)
Let the Noise Speak: Harnessing Noise for a Unified Defense Against Adversarial and Backdoor Attacks
by: Shahriar, Md Hasan, et al.
Published: (2024)
by: Shahriar, Md Hasan, et al.
Published: (2024)
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation
by: Zhong, Zhiyuan, et al.
Published: (2025)
by: Zhong, Zhiyuan, et al.
Published: (2025)
Mitigating Backdoor Attacks using Activation-Guided Model Editing
by: Hsieh, Felix, et al.
Published: (2024)
by: Hsieh, Felix, et al.
Published: (2024)
Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models
by: Zhang, Zongmin, et al.
Published: (2025)
by: Zhang, Zongmin, et al.
Published: (2025)
InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning
by: Sun, Mengyuan, et al.
Published: (2025)
by: Sun, Mengyuan, et al.
Published: (2025)
Clean-image Backdoor Attacks
by: Rong, Dazhong, et al.
Published: (2024)
by: Rong, Dazhong, et al.
Published: (2024)
Similar Items
-
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025) -
Do VLMs Have Bad Eyes? Diagnosing Compositional Failures via Mechanistic Interpretability
by: Aravindan, Ashwath Vaithinathan, et al.
Published: (2025) -
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024) -
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024) -
SATBA: An Invisible Backdoor Attack Based On Spatial Attention
by: Zhou, Huasong, et al.
Published: (2023)