Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhang, Haifeng, He, Qinghui, Bi, Xiuli, Liu, Bo, Pun, Chi-Man, Xiao, Bin
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908962137309184
author Zhang, Haifeng
He, Qinghui
Bi, Xiuli
Liu, Bo
Pun, Chi-Man
Xiao, Bin
author_facet Zhang, Haifeng
He, Qinghui
Bi, Xiuli
Liu, Bo
Pun, Chi-Man
Xiao, Bin
contents In recent years, the rapid development of generative artificial intelligence technology has significantly lowered the barrier to creating high-quality fake images, posing a serious challenge to information authenticity and credibility. Existing generated image detection methods typically enhance generalization through model architecture or network design. However, their generalization performance remains susceptible to data bias, as the training data may drive models to fit specific generative patterns and content rather than the common features shared by images from different generative models (asymmetric bias learning). To address this issue, we propose a Multi-dimensional Adversarial Feature Learning (MAFL) framework. The framework adopts a pretrained multimodal image encoder as the feature extraction backbone, constructs a real-fake feature learning network, and designs an adversarial bias-learning branch equipped with a multi-dimensional adversarial loss, forming an adversarial training mechanism between authenticity-discriminative feature learning and bias feature learning. By suppressing generation-pattern and content biases, MAFL guides the model to focus on the generative features shared across different generative models, thereby effectively capturing the fundamental differences between real and generated images, enhancing cross-model generalization, and substantially reducing the reliance on large-scale training data. Through extensive experimental validation, our method outperforms existing state-of-the-art approaches by 10.89% in accuracy and 8.57% in Average Precision (AP). Notably, even when trained with only 320 images, it can still achieve over 80% detection accuracy on public datasets.
format Preprint
id arxiv_https___arxiv_org_abs_2604_12353
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection
Zhang, Haifeng
He, Qinghui
Bi, Xiuli
Liu, Bo
Pun, Chi-Man
Xiao, Bin
Computer Vision and Pattern Recognition
In recent years, the rapid development of generative artificial intelligence technology has significantly lowered the barrier to creating high-quality fake images, posing a serious challenge to information authenticity and credibility. Existing generated image detection methods typically enhance generalization through model architecture or network design. However, their generalization performance remains susceptible to data bias, as the training data may drive models to fit specific generative patterns and content rather than the common features shared by images from different generative models (asymmetric bias learning). To address this issue, we propose a Multi-dimensional Adversarial Feature Learning (MAFL) framework. The framework adopts a pretrained multimodal image encoder as the feature extraction backbone, constructs a real-fake feature learning network, and designs an adversarial bias-learning branch equipped with a multi-dimensional adversarial loss, forming an adversarial training mechanism between authenticity-discriminative feature learning and bias feature learning. By suppressing generation-pattern and content biases, MAFL guides the model to focus on the generative features shared across different generative models, thereby effectively capturing the fundamental differences between real and generated images, enhancing cross-model generalization, and substantially reducing the reliance on large-scale training data. Through extensive experimental validation, our method outperforms existing state-of-the-art approaches by 10.89% in accuracy and 8.57% in Average Precision (AP). Notably, even when trained with only 320 images, it can still achieve over 80% detection accuracy on public datasets.
title Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2604.12353