Membership Inference on Text-to-Image Diffusion Models via Conditional Likelihood Discrepancy

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhai, Shengfang, Chen, Huanran, Dong, Yinpeng, Li, Jiajun, Shen, Qingni, Gao, Yansong, Su, Hang, Liu, Yang
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912089130401792
author Zhai, Shengfang
Chen, Huanran
Dong, Yinpeng
Li, Jiajun
Shen, Qingni
Gao, Yansong
Su, Hang
Liu, Yang
author_facet Zhai, Shengfang
Chen, Huanran
Dong, Yinpeng
Li, Jiajun
Shen, Qingni
Gao, Yansong
Su, Hang
Liu, Yang
contents Text-to-image diffusion models have achieved tremendous success in the field of controllable image generation, while also coming along with issues of privacy leakage and data copyrights. Membership inference arises in these contexts as a potential auditing method for detecting unauthorized data usage. While some efforts have been made on diffusion models, they are not applicable to text-to-image diffusion models due to the high computation overhead and enhanced generalization capabilities. In this paper, we first identify a conditional overfitting phenomenon in text-to-image diffusion models, indicating that these models tend to overfit the conditional distribution of images given the corresponding text rather than the marginal distribution of images only. Based on this observation, we derive an analytical indicator, namely Conditional Likelihood Discrepancy (CLiD), to perform membership inference, which reduces the stochasticity in estimating memorization of individual samples. Experimental results demonstrate that our method significantly outperforms previous methods across various data distributions and dataset scales. Additionally, our method shows superior resistance to overfitting mitigation strategies, such as early stopping and data augmentation.
format Preprint
id arxiv_https___arxiv_org_abs_2405_14800
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Membership Inference on Text-to-Image Diffusion Models via Conditional Likelihood Discrepancy
Zhai, Shengfang
Chen, Huanran
Dong, Yinpeng
Li, Jiajun
Shen, Qingni
Gao, Yansong
Su, Hang
Liu, Yang
Cryptography and Security
Computer Vision and Pattern Recognition
Text-to-image diffusion models have achieved tremendous success in the field of controllable image generation, while also coming along with issues of privacy leakage and data copyrights. Membership inference arises in these contexts as a potential auditing method for detecting unauthorized data usage. While some efforts have been made on diffusion models, they are not applicable to text-to-image diffusion models due to the high computation overhead and enhanced generalization capabilities. In this paper, we first identify a conditional overfitting phenomenon in text-to-image diffusion models, indicating that these models tend to overfit the conditional distribution of images given the corresponding text rather than the marginal distribution of images only. Based on this observation, we derive an analytical indicator, namely Conditional Likelihood Discrepancy (CLiD), to perform membership inference, which reduces the stochasticity in estimating memorization of individual samples. Experimental results demonstrate that our method significantly outperforms previous methods across various data distributions and dataset scales. Additionally, our method shows superior resistance to overfitting mitigation strategies, such as early stopping and data augmentation.
title Membership Inference on Text-to-Image Diffusion Models via Conditional Likelihood Discrepancy
topic Cryptography and Security
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2405.14800