A Quality-Centric Framework for Generic Deepfake Detection

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Song, Wentang, Yan, Zhiyuan, Lin, Yuzhen, Yao, Taiping, Chen, Changsheng, Chen, Shen, Zhao, Yandan, Ding, Shouhong, Li, Bin
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909613722435584
author Song, Wentang
Yan, Zhiyuan
Lin, Yuzhen
Yao, Taiping
Chen, Changsheng
Chen, Shen
Zhao, Yandan
Ding, Shouhong
Li, Bin
author_facet Song, Wentang
Yan, Zhiyuan
Lin, Yuzhen
Yao, Taiping
Chen, Changsheng
Chen, Shen
Zhao, Yandan
Ding, Shouhong
Li, Bin
contents Detecting AI-generated images, particularly deepfakes, has become increasingly crucial, with the primary challenge being the generalization to previously unseen manipulation methods. This paper tackles this issue by leveraging the forgery quality of training data to improve the generalization performance of existing deepfake detectors. Generally, the forgery quality of different deepfakes varies: some have easily recognizable forgery clues, while others are highly realistic. Existing works often train detectors on a mix of deepfakes with varying forgery qualities, potentially leading detectors to short-cut the easy-to-spot artifacts from low-quality forgery samples, thereby hurting generalization performance. To tackle this issue, we propose a novel quality-centric framework for generic deepfake detection, which is composed of a Quality Evaluator, a low-quality data enhancement module, and a learning pacing strategy that explicitly incorporates forgery quality into the training process. Our framework is inspired by curriculum learning, which is designed to gradually enable the detector to learn more challenging deepfake samples, starting with easier samples and progressing to more realistic ones. We employ both static and dynamic assessments to assess the forgery quality, combining their scores to produce a final rating for each training sample. The rating score guides the selection of deepfake samples for training, with higher-rated samples having a higher probability of being chosen. Furthermore, we propose a novel frequency data augmentation method specifically designed for low-quality forgery samples, which helps to reduce obvious forgery traces and improve their overall realism. Extensive experiments demonstrate that our proposed framework can be applied plug-and-play to existing detection models and significantly enhance their generalization performance in detection.
format Preprint
id arxiv_https___arxiv_org_abs_2411_05335
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle A Quality-Centric Framework for Generic Deepfake Detection
Song, Wentang
Yan, Zhiyuan
Lin, Yuzhen
Yao, Taiping
Chen, Changsheng
Chen, Shen
Zhao, Yandan
Ding, Shouhong
Li, Bin
Computer Vision and Pattern Recognition
Cryptography and Security
Machine Learning
Detecting AI-generated images, particularly deepfakes, has become increasingly crucial, with the primary challenge being the generalization to previously unseen manipulation methods. This paper tackles this issue by leveraging the forgery quality of training data to improve the generalization performance of existing deepfake detectors. Generally, the forgery quality of different deepfakes varies: some have easily recognizable forgery clues, while others are highly realistic. Existing works often train detectors on a mix of deepfakes with varying forgery qualities, potentially leading detectors to short-cut the easy-to-spot artifacts from low-quality forgery samples, thereby hurting generalization performance. To tackle this issue, we propose a novel quality-centric framework for generic deepfake detection, which is composed of a Quality Evaluator, a low-quality data enhancement module, and a learning pacing strategy that explicitly incorporates forgery quality into the training process. Our framework is inspired by curriculum learning, which is designed to gradually enable the detector to learn more challenging deepfake samples, starting with easier samples and progressing to more realistic ones. We employ both static and dynamic assessments to assess the forgery quality, combining their scores to produce a final rating for each training sample. The rating score guides the selection of deepfake samples for training, with higher-rated samples having a higher probability of being chosen. Furthermore, we propose a novel frequency data augmentation method specifically designed for low-quality forgery samples, which helps to reduce obvious forgery traces and improve their overall realism. Extensive experiments demonstrate that our proposed framework can be applied plug-and-play to existing detection models and significantly enhance their generalization performance in detection.
title A Quality-Centric Framework for Generic Deepfake Detection
topic Computer Vision and Pattern Recognition
Cryptography and Security
Machine Learning
url https://arxiv.org/abs/2411.05335