Responsible Diffusion: A Comprehensive Survey on Safety, Ethics, and Trust in Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Kang, Yuan, Xin, Huo, Fushuo, Ma, Chuan, Yuan, Long, Li, Songze, Ding, Ming, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
by: Ye, Mang, et al.
Published: (2025)
by: Ye, Mang, et al.
Published: (2025)
Beauty and the Beast: Imperceptible Perturbations Against Diffusion-Based Face Swapping via Directional Attribute Editing
by: Huang, Yilong, et al.
Published: (2026)
by: Huang, Yilong, et al.
Published: (2026)
Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks
by: Ying, Zonghao, et al.
Published: (2024)
by: Ying, Zonghao, et al.
Published: (2024)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
by: Ma, Xingjun, et al.
Published: (2025)
by: Ma, Xingjun, et al.
Published: (2025)
Where the Devil Hides: Deepfake Detectors Can No Longer Be Trusted
by: Yuan, Shuaiwei, et al.
Published: (2025)
by: Yuan, Shuaiwei, et al.
Published: (2025)
Is Diffusion Model Safe? Severe Data Leakage via Gradient-Guided Diffusion Model
by: Meng, Jiayang, et al.
Published: (2024)
by: Meng, Jiayang, et al.
Published: (2024)
CtrlAttack: A Unified Attack on World-Model Control in Diffusion Models
by: Xu, Shuhan, et al.
Published: (2026)
by: Xu, Shuhan, et al.
Published: (2026)
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024)
by: Zhang, Yimeng, et al.
Published: (2024)
HMARK: Radioactive Multi-Bit Semantic-Latent Watermarking for Diffusion Models
by: Li, Kexin, et al.
Published: (2025)
by: Li, Kexin, et al.
Published: (2025)
Struggle with Adversarial Defense? Try Diffusion
by: Li, Yujie, et al.
Published: (2024)
by: Li, Yujie, et al.
Published: (2024)
MMA-Diffusion: MultiModal Attack on Diffusion Models
by: Yang, Yijun, et al.
Published: (2023)
by: Yang, Yijun, et al.
Published: (2023)
Cert-LAS: Toward Certified Model Ownership Verification for Text-to-Image Diffusion Models via Layer-Adaptive Smoothing
by: Qi, Leyi, et al.
Published: (2026)
by: Qi, Leyi, et al.
Published: (2026)
TimeGuard: Channel-wise Pool Training for Backdoor Defense in Time Series Forecasting
by: Nguyen, Quang Duc, et al.
Published: (2026)
by: Nguyen, Quang Duc, et al.
Published: (2026)
An Inversion-based Measure of Memorization for Diffusion Models
by: Ma, Zhe, et al.
Published: (2024)
by: Ma, Zhe, et al.
Published: (2024)
LoyalDiffusion: A Diffusion Model Guarding Against Data Replication
by: Li, Chenghao, et al.
Published: (2024)
by: Li, Chenghao, et al.
Published: (2024)
What Lurks Within? Concept Auditing for Shared Diffusion Models at Scale
by: Yuan, Xiaoyong, et al.
Published: (2025)
by: Yuan, Xiaoyong, et al.
Published: (2025)
Guidance Watermarking for Diffusion Models
by: Gesny, Enoal, et al.
Published: (2025)
by: Gesny, Enoal, et al.
Published: (2025)
Training-Free Color-Aware Adversarial Diffusion Sanitization for Diffusion Stegomalware Defense at Security Gateways
by: Frants, Vladimir, et al.
Published: (2025)
by: Frants, Vladimir, et al.
Published: (2025)
Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
Secure and Robust Watermarking for AI-generated Images: A Comprehensive Survey
by: Cao, Jie, et al.
Published: (2025)
by: Cao, Jie, et al.
Published: (2025)
Gaussian Shading++: Rethinking the Realistic Deployment Challenge of Performance-Lossless Image Watermark for Diffusion Models
by: Yang, Zijin, et al.
Published: (2025)
by: Yang, Zijin, et al.
Published: (2025)
Adversarial Attacks and Defenses on Text-to-Image Diffusion Models: A Survey
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Time Is All It Takes: Spike-Retiming Attacks on Event-Driven Spiking Neural Networks
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
Invisible Backdoor Attacks on Diffusion Models
by: Li, Sen, et al.
Published: (2024)
by: Li, Sen, et al.
Published: (2024)
Adversarial Examples are Misaligned in Diffusion Model Manifolds
by: Lorenz, Peter, et al.
Published: (2024)
by: Lorenz, Peter, et al.
Published: (2024)
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
by: Ying, Zonghao, et al.
Published: (2024)
by: Ying, Zonghao, et al.
Published: (2024)
SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery Attacks
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
DisDet: Exploring Detectability of Backdoor Attack on Diffusion Models
by: Sui, Yang, et al.
Published: (2024)
by: Sui, Yang, et al.
Published: (2024)
Rethinking Robust Adversarial Concept Erasure in Diffusion Models
by: Yin, Qinghong, et al.
Published: (2025)
by: Yin, Qinghong, et al.
Published: (2025)
Video Signature: Implicit Watermarking for Video Diffusion Models
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
by: Li, Xiao, et al.
Published: (2026)
by: Li, Xiao, et al.
Published: (2026)
DP-RDM: Adapting Diffusion Models to Private Domains Without Fine-Tuning
by: Lebensold, Jonathan, et al.
Published: (2024)
by: Lebensold, Jonathan, et al.
Published: (2024)
Controllable Adversarial Makeup for Privacy via Text-Guided Diffusion
by: Kwon, Youngjin, et al.
Published: (2025)
by: Kwon, Youngjin, et al.
Published: (2025)
Universal Adversarial Purification with DDIM Metric Loss for Stable Diffusion
by: Zheng, Li, et al.
Published: (2026)
by: Zheng, Li, et al.
Published: (2026)
GaussMarker: Robust Dual-Domain Watermark for Diffusion Models
by: Li, Kecen, et al.
Published: (2025)
by: Li, Kecen, et al.
Published: (2025)
Embedding Watermarks in Diffusion Process for Model Intellectual Property Protection
by: Yang, Jijia, et al.
Published: (2024)
by: Yang, Jijia, et al.
Published: (2024)
Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
CLIP-Flow: A Universal Discriminator for AI-Generated Images Inspired by Anomaly Detection
by: Yuan, Zhipeng, et al.
Published: (2025)
by: Yuan, Zhipeng, et al.
Published: (2025)
Gaussian Shannon: High-Precision Diffusion Model Watermarking Based on Communication
by: Zhang, Yi, et al.
Published: (2026)
by: Zhang, Yi, et al.
Published: (2026)
Similar Items
-
Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models
by: Chen, Hao, et al.
Published: (2025) -
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
by: Ye, Mang, et al.
Published: (2025) -
Beauty and the Beast: Imperceptible Perturbations Against Diffusion-Based Face Swapping via Directional Attribute Editing
by: Huang, Yilong, et al.
Published: (2026) -
Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks
by: Ying, Zonghao, et al.
Published: (2024) -
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
by: Ma, Xingjun, et al.
Published: (2025)