Can Large Vision-Language Models Detect Images Copyright Infringement from GenAI?

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Xu, Qipan, Wang, Zhenting, He, Xiaoxiao, Han, Ligong, Tang, Ruixiang
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866913703439368192
author Xu, Qipan
Wang, Zhenting
He, Xiaoxiao
Han, Ligong
Tang, Ruixiang
author_facet Xu, Qipan
Wang, Zhenting
He, Xiaoxiao
Han, Ligong
Tang, Ruixiang
contents Generative AI models, renowned for their ability to synthesize high-quality content, have sparked growing concerns over the improper generation of copyright-protected material. While recent studies have proposed various approaches to address copyright issues, the capability of large vision-language models (LVLMs) to detect copyright infringements remains largely unexplored. In this work, we focus on evaluating the copyright detection abilities of state-of-the-art LVLMs using a various set of image samples. Recognizing the absence of a comprehensive dataset that includes both IP-infringement samples and ambiguous non-infringement negative samples, we construct a benchmark dataset comprising positive samples that violate the copyright protection of well-known IP figures, as well as negative samples that resemble these figures but do not raise copyright concerns. This dataset is created using advanced prompt engineering techniques. We then evaluate leading LVLMs using our benchmark dataset. Our experimental results reveal that LVLMs are prone to overfitting, leading to the misclassification of some negative samples as IP-infringement cases. In the final section, we analyze these failure cases and propose potential solutions to mitigate the overfitting problem.
format Preprint
id arxiv_https___arxiv_org_abs_2502_16618
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Can Large Vision-Language Models Detect Images Copyright Infringement from GenAI?
Xu, Qipan
Wang, Zhenting
He, Xiaoxiao
Han, Ligong
Tang, Ruixiang
Computer Vision and Pattern Recognition
Artificial Intelligence
Computation and Language
Generative AI models, renowned for their ability to synthesize high-quality content, have sparked growing concerns over the improper generation of copyright-protected material. While recent studies have proposed various approaches to address copyright issues, the capability of large vision-language models (LVLMs) to detect copyright infringements remains largely unexplored. In this work, we focus on evaluating the copyright detection abilities of state-of-the-art LVLMs using a various set of image samples. Recognizing the absence of a comprehensive dataset that includes both IP-infringement samples and ambiguous non-infringement negative samples, we construct a benchmark dataset comprising positive samples that violate the copyright protection of well-known IP figures, as well as negative samples that resemble these figures but do not raise copyright concerns. This dataset is created using advanced prompt engineering techniques. We then evaluate leading LVLMs using our benchmark dataset. Our experimental results reveal that LVLMs are prone to overfitting, leading to the misclassification of some negative samples as IP-infringement cases. In the final section, we analyze these failure cases and propose potential solutions to mitigate the overfitting problem.
title Can Large Vision-Language Models Detect Images Copyright Infringement from GenAI?
topic Computer Vision and Pattern Recognition
Artificial Intelligence
Computation and Language
url https://arxiv.org/abs/2502.16618