Adversarial Attacks and Defenses on Text-to-Image Diffusion Models: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Chenyu, Hu, Mingwang, Li, Wenhui, Wang, Lanjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
Metaphor-based Jailbreak Attacks on Text-to-Image Models
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025)
Beyond Vulnerabilities: A Survey of Adversarial Attacks as Both Threats and Defenses in Computer Vision Systems
von: Guo, Zhongliang, et al.
Veröffentlicht: (2025)
von: Guo, Zhongliang, et al.
Veröffentlicht: (2025)
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
von: Wei, Xingxing, et al.
Veröffentlicht: (2024)
von: Wei, Xingxing, et al.
Veröffentlicht: (2024)
DIFFender: Diffusion-Based Adversarial Defense against Patch Attacks
von: Kang, Caixin, et al.
Veröffentlicht: (2023)
von: Kang, Caixin, et al.
Veröffentlicht: (2023)
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
von: Li, Xiao, et al.
Veröffentlicht: (2026)
von: Li, Xiao, et al.
Veröffentlicht: (2026)
Towards Dataset Copyright Evasion Attack against Personalized Text-to-Image Diffusion Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2025)
von: Gao, Kuofeng, et al.
Veröffentlicht: (2025)
Silent Branding Attack: Trigger-free Data Poisoning Attack on Text-to-Image Diffusion Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
Autoencoder-based Denoising Defense against Adversarial Attacks on Object Detection
von: Song, Min Geun, et al.
Veröffentlicht: (2025)
von: Song, Min Geun, et al.
Veröffentlicht: (2025)
CtrlAttack: A Unified Attack on World-Model Control in Diffusion Models
von: Xu, Shuhan, et al.
Veröffentlicht: (2026)
von: Xu, Shuhan, et al.
Veröffentlicht: (2026)
PLA: Prompt Learning Attack against Text-to-Image Generative Models
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
von: Lyu, Xinqi, et al.
Veröffentlicht: (2025)
PromptLA: Towards Integrity Verification of Black-box Text-to-Image Diffusion Models
von: Zhang, Zhuomeng, et al.
Veröffentlicht: (2024)
von: Zhang, Zhuomeng, et al.
Veröffentlicht: (2024)
DiffZOO: A Purely Query-Based Black-Box Attack for Red-teaming Text-to-Image Generative Model via Zeroth Order Optimization
von: Dang, Pucheng, et al.
Veröffentlicht: (2024)
von: Dang, Pucheng, et al.
Veröffentlicht: (2024)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
SteerDiff: Steering towards Safe Text-to-Image Diffusion Models
von: Zhang, Hongxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Hongxiang, et al.
Veröffentlicht: (2024)
Superpixel Attack: Enhancing Black-box Adversarial Attack with Image-driven Division Areas
von: Oe, Issa, et al.
Veröffentlicht: (2025)
von: Oe, Issa, et al.
Veröffentlicht: (2025)
CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models
von: Liu, Renyang, et al.
Veröffentlicht: (2026)
von: Liu, Renyang, et al.
Veröffentlicht: (2026)
Undermining Image and Text Classification Algorithms Using Adversarial Attacks
von: Lunga, Langalibalele, et al.
Veröffentlicht: (2024)
von: Lunga, Langalibalele, et al.
Veröffentlicht: (2024)
A Survey on Physical Adversarial Attacks against Face Recognition Systems
von: Wang, Mingsi, et al.
Veröffentlicht: (2024)
von: Wang, Mingsi, et al.
Veröffentlicht: (2024)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
von: Jha, Abha, et al.
Veröffentlicht: (2025)
von: Jha, Abha, et al.
Veröffentlicht: (2025)
DREAM: Scalable Red Teaming for Text-to-Image Generative Systems via Distribution Modeling
von: Li, Boheng, et al.
Veröffentlicht: (2025)
von: Li, Boheng, et al.
Veröffentlicht: (2025)
PuriDefense: Randomized Local Implicit Adversarial Purification for Defending Black-box Query-based Attacks
von: Guo, Ping, et al.
Veröffentlicht: (2024)
von: Guo, Ping, et al.
Veröffentlicht: (2024)
Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
ROBIN: Robust and Invisible Watermarks for Diffusion Models with Adversarial Optimization
von: Huang, Huayang, et al.
Veröffentlicht: (2024)
von: Huang, Huayang, et al.
Veröffentlicht: (2024)
Backdoor Attacks against Image-to-Image Networks
von: Jiang, Wenbo, et al.
Veröffentlicht: (2024)
von: Jiang, Wenbo, et al.
Veröffentlicht: (2024)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
von: Xu, Naen, et al.
Veröffentlicht: (2025)
von: Xu, Naen, et al.
Veröffentlicht: (2025)
Sparse Autoencoder as a Zero-Shot Classifier for Concept Erasing in Text-to-Image Diffusion Models
von: Tian, Zhihua, et al.
Veröffentlicht: (2025)
von: Tian, Zhihua, et al.
Veröffentlicht: (2025)
On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
DeMark: A Query-Free Black-Box Attack on Deepfake Watermarking Defenses
von: Song, Wei, et al.
Veröffentlicht: (2026)
von: Song, Wei, et al.
Veröffentlicht: (2026)
Watertox: The Art of Simplicity in Universal Attacks A Cross-Model Framework for Robust Adversarial Generation
von: Gao, Zhenghao, et al.
Veröffentlicht: (2024)
von: Gao, Zhenghao, et al.
Veröffentlicht: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models
von: Li, Xiang, et al.
Veröffentlicht: (2023)
von: Li, Xiang, et al.
Veröffentlicht: (2023)
Edge-Only Universal Adversarial Attacks in Distributed Learning
von: Rossolini, Giulio, et al.
Veröffentlicht: (2024)
von: Rossolini, Giulio, et al.
Veröffentlicht: (2024)
SKeDA: A Generative Watermarking Framework for Text-to-video Diffusion Models
von: Yang, Yang, et al.
Veröffentlicht: (2026)
von: Yang, Yang, et al.
Veröffentlicht: (2026)
Black-Box Forgery Attacks on Semantic Watermarks for Diffusion Models
von: Müller, Andreas, et al.
Veröffentlicht: (2024)
von: Müller, Andreas, et al.
Veröffentlicht: (2024)
Security Risk of Misalignment between Text and Image in Multi-modal Model
von: Wang, Xiaosen, et al.
Veröffentlicht: (2025)
von: Wang, Xiaosen, et al.
Veröffentlicht: (2025)
AR-GAN: Generative Adversarial Network-Based Defense Method Against Adversarial Attacks on the Traffic Sign Classification System of Autonomous Vehicles
von: Salek, M Sabbir, et al.
Veröffentlicht: (2023)
von: Salek, M Sabbir, et al.
Veröffentlicht: (2023)
Diffusion Soup: Model Merging for Text-to-Image Diffusion Models
von: Biggs, Benjamin, et al.
Veröffentlicht: (2024)
von: Biggs, Benjamin, et al.
Veröffentlicht: (2024)
Struggle with Adversarial Defense? Try Diffusion
von: Li, Yujie, et al.
Veröffentlicht: (2024)
von: Li, Yujie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025) -
Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025) -
Metaphor-based Jailbreak Attacks on Text-to-Image Models
von: Zhang, Chenyu, et al.
Veröffentlicht: (2025) -
Beyond Vulnerabilities: A Survey of Adversarial Attacks as Both Threats and Defenses in Computer Vision Systems
von: Guo, Zhongliang, et al.
Veröffentlicht: (2025) -
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
von: Wei, Xingxing, et al.
Veröffentlicht: (2024)