Rethinking Robust Adversarial Concept Erasure in Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Qinghong, Tian, Yu, Yang, Heming, Chen, Xiang, Zhang, Xianlin, Li, Xueming, Zhan, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024)
by: Zhang, Yimeng, et al.
Published: (2024)
Neighbor-Aware Localized Concept Erasure in Text-to-Image Diffusion Models
by: Shi, Zhuan, et al.
Published: (2026)
by: Shi, Zhuan, et al.
Published: (2026)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
by: Xu, Naen, et al.
Published: (2025)
by: Xu, Naen, et al.
Published: (2025)
Guarding the Gate: ConceptGuard Battles Concept-Level Backdoors in Concept Bottleneck Models
by: Lai, Songning, et al.
Published: (2024)
by: Lai, Songning, et al.
Published: (2024)
CGCE: Classifier-Guided Concept Erasure in Generative Models
by: Nguyen, Viet, et al.
Published: (2025)
by: Nguyen, Viet, et al.
Published: (2025)
Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration
by: Li, Jun, et al.
Published: (2026)
by: Li, Jun, et al.
Published: (2026)
Gaussian Shading++: Rethinking the Realistic Deployment Challenge of Performance-Lossless Image Watermark for Diffusion Models
by: Yang, Zijin, et al.
Published: (2025)
by: Yang, Zijin, et al.
Published: (2025)
Six-CD: Benchmarking Concept Removals for Benign Text-to-image Diffusion Models
by: Ren, Jie, et al.
Published: (2024)
by: Ren, Jie, et al.
Published: (2024)
CAT: Concept-level backdoor ATtacks for Concept Bottleneck Models
by: Lai, Songning, et al.
Published: (2024)
by: Lai, Songning, et al.
Published: (2024)
SAP-DIFF: Semantic Adversarial Patch Generation for Black-Box Face Recognition Models via Diffusion Models
by: Wang, Mingsi, et al.
Published: (2025)
by: Wang, Mingsi, et al.
Published: (2025)
ROBIN: Robust and Invisible Watermarks for Diffusion Models with Adversarial Optimization
by: Huang, Huayang, et al.
Published: (2024)
by: Huang, Huayang, et al.
Published: (2024)
Espresso: Robust Concept Filtering in Text-to-Image Models
by: Das, Anudeep, et al.
Published: (2024)
by: Das, Anudeep, et al.
Published: (2024)
Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
PatchCURE: Improving Certifiable Robustness, Model Utility, and Computation Efficiency of Adversarial Patch Defenses
by: Xiang, Chong, et al.
Published: (2023)
by: Xiang, Chong, et al.
Published: (2023)
TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models
by: Zhang, Chaoshuo, et al.
Published: (2026)
by: Zhang, Chaoshuo, et al.
Published: (2026)
Adversarial Examples are Misaligned in Diffusion Model Manifolds
by: Lorenz, Peter, et al.
Published: (2024)
by: Lorenz, Peter, et al.
Published: (2024)
Improving Adversarial Robustness via Feature Pattern Consistency Constraint
by: Hu, Jiacong, et al.
Published: (2024)
by: Hu, Jiacong, et al.
Published: (2024)
Controllable Adversarial Makeup for Privacy via Text-Guided Diffusion
by: Kwon, Youngjin, et al.
Published: (2025)
by: Kwon, Youngjin, et al.
Published: (2025)
What Concepts Lie Within? Detecting and Suppressing Risky Content in Diffusion Transformers
by: Zhang, Chenyu
Published: (2026)
by: Zhang, Chenyu
Published: (2026)
Who Can See Through You? Adversarial Shielding Against VLM-Based Attribute Inference Attacks
by: Fan, Yucheng, et al.
Published: (2025)
by: Fan, Yucheng, et al.
Published: (2025)
Struggle with Adversarial Defense? Try Diffusion
by: Li, Yujie, et al.
Published: (2024)
by: Li, Yujie, et al.
Published: (2024)
Boosting Adversarial Transferability with Spatial Adversarial Alignment
by: Chen, Zhaoyu, et al.
Published: (2025)
by: Chen, Zhaoyu, et al.
Published: (2025)
Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models
by: Yang, Zijin, et al.
Published: (2024)
by: Yang, Zijin, et al.
Published: (2024)
VLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained Models
by: Yin, Ziyi, et al.
Published: (2023)
by: Yin, Ziyi, et al.
Published: (2023)
Rethinking and Red-Teaming Protective Perturbation in Personalized Diffusion Models
by: Liu, Yixin, et al.
Published: (2024)
by: Liu, Yixin, et al.
Published: (2024)
RoMA: Robust Malware Attribution via Byte-level Adversarial Training with Global Perturbations and Adversarial Consistency Regularization
by: Sun, Yuxia, et al.
Published: (2025)
by: Sun, Yuxia, et al.
Published: (2025)
DiffProtect: Generate Adversarial Examples with Diffusion Models for Facial Privacy Protection
by: Liu, Jiang, et al.
Published: (2023)
by: Liu, Jiang, et al.
Published: (2023)
Is RobustBench/AutoAttack a suitable Benchmark for Adversarial Robustness?
by: Lorenz, Peter, et al.
Published: (2021)
by: Lorenz, Peter, et al.
Published: (2021)
On the Robustness of Kolmogorov-Arnold Networks: An Adversarial Perspective
by: Alter, Tal, et al.
Published: (2024)
by: Alter, Tal, et al.
Published: (2024)
Distilling Adversarial Robustness Using Heterogeneous Teachers
by: Deng, Jieren, et al.
Published: (2024)
by: Deng, Jieren, et al.
Published: (2024)
Stealthy Multi-Task Adversarial Attacks
by: Guo, Jiacheng, et al.
Published: (2024)
by: Guo, Jiacheng, et al.
Published: (2024)
BSPA: Exploring Black-box Stealthy Prompt Attacks against Image Generators
by: Tian, Yu, et al.
Published: (2024)
by: Tian, Yu, et al.
Published: (2024)
LightPure: Realtime Adversarial Image Purification for Mobile Devices Using Diffusion Models
by: Khalili, Hossein, et al.
Published: (2024)
by: Khalili, Hossein, et al.
Published: (2024)
CAPAA: Classifier-Agnostic Projector-Based Adversarial Attack
by: Li, Zhan, et al.
Published: (2025)
by: Li, Zhan, et al.
Published: (2025)
Intriguing Properties of Diffusion Models: An Empirical Study of the Natural Attack Capability in Text-to-Image Generative Models
by: Sato, Takami, et al.
Published: (2023)
by: Sato, Takami, et al.
Published: (2023)
Robust Watermarks Leak: Channel-Aware Feature Extraction Enables Adversarial Watermark Manipulation
by: Ba, Zhongjie, et al.
Published: (2025)
by: Ba, Zhongjie, et al.
Published: (2025)
Exploring Adversarial Attacks against Latent Diffusion Model from the Perspective of Adversarial Transferability
by: Chen, Junxi, et al.
Published: (2024)
by: Chen, Junxi, et al.
Published: (2024)
Natias: Neuron Attribution based Transferable Image Adversarial Steganography
by: Fan, Zexin, et al.
Published: (2024)
by: Fan, Zexin, et al.
Published: (2024)
GaussMarker: Robust Dual-Domain Watermark for Diffusion Models
by: Li, Kecen, et al.
Published: (2025)
by: Li, Kecen, et al.
Published: (2025)
Deciphering the Definition of Adversarial Robustness for post-hoc OOD Detectors
by: Lorenz, Peter, et al.
Published: (2024)
by: Lorenz, Peter, et al.
Published: (2024)
Similar Items
-
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024) -
Neighbor-Aware Localized Concept Erasure in Text-to-Image Diffusion Models
by: Shi, Zhuan, et al.
Published: (2026) -
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
by: Xu, Naen, et al.
Published: (2025) -
Guarding the Gate: ConceptGuard Battles Concept-Level Backdoors in Concept Bottleneck Models
by: Lai, Songning, et al.
Published: (2024) -
CGCE: Classifier-Guided Concept Erasure in Generative Models
by: Nguyen, Viet, et al.
Published: (2025)