XAI-Based Detection of Adversarial Attacks on Deepfake Detectors

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Pinhasov, Ben, Lapid, Raz, Ohayon, Rony, Sipper, Moshe, Aperstein, Yehudit
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910568543158272
author Pinhasov, Ben
Lapid, Raz
Ohayon, Rony
Sipper, Moshe
Aperstein, Yehudit
author_facet Pinhasov, Ben
Lapid, Raz
Ohayon, Rony
Sipper, Moshe
Aperstein, Yehudit
contents We introduce a novel methodology for identifying adversarial attacks on deepfake detectors using eXplainable Artificial Intelligence (XAI). In an era characterized by digital advancement, deepfakes have emerged as a potent tool, creating a demand for efficient detection systems. However, these systems are frequently targeted by adversarial attacks that inhibit their performance. We address this gap, developing a defensible deepfake detector by leveraging the power of XAI. The proposed methodology uses XAI to generate interpretability maps for a given method, providing explicit visualizations of decision-making factors within the AI models. We subsequently employ a pretrained feature extractor that processes both the input image and its corresponding XAI image. The feature embeddings extracted from this process are then used for training a simple yet effective classifier. Our approach contributes not only to the detection of deepfakes but also enhances the understanding of possible adversarial attacks, pinpointing potential vulnerabilities. Furthermore, this approach does not change the performance of the deepfake detector. The paper demonstrates promising results suggesting a potential pathway for future deepfake detection mechanisms. We believe this study will serve as a valuable contribution to the community, sparking much-needed discourse on safeguarding deepfake detectors.
format Preprint
id arxiv_https___arxiv_org_abs_2403_02955
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle XAI-Based Detection of Adversarial Attacks on Deepfake Detectors
Pinhasov, Ben
Lapid, Raz
Ohayon, Rony
Sipper, Moshe
Aperstein, Yehudit
Cryptography and Security
Computer Vision and Pattern Recognition
We introduce a novel methodology for identifying adversarial attacks on deepfake detectors using eXplainable Artificial Intelligence (XAI). In an era characterized by digital advancement, deepfakes have emerged as a potent tool, creating a demand for efficient detection systems. However, these systems are frequently targeted by adversarial attacks that inhibit their performance. We address this gap, developing a defensible deepfake detector by leveraging the power of XAI. The proposed methodology uses XAI to generate interpretability maps for a given method, providing explicit visualizations of decision-making factors within the AI models. We subsequently employ a pretrained feature extractor that processes both the input image and its corresponding XAI image. The feature embeddings extracted from this process are then used for training a simple yet effective classifier. Our approach contributes not only to the detection of deepfakes but also enhances the understanding of possible adversarial attacks, pinpointing potential vulnerabilities. Furthermore, this approach does not change the performance of the deepfake detector. The paper demonstrates promising results suggesting a potential pathway for future deepfake detection mechanisms. We believe this study will serve as a valuable contribution to the community, sparking much-needed discourse on safeguarding deepfake detectors.
title XAI-Based Detection of Adversarial Attacks on Deepfake Detectors
topic Cryptography and Security
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2403.02955