XAI-Based Detection of Adversarial Attacks on Deepfake Detectors
Fuente:
arXiv
Saved in:
| Main Authors: | Pinhasov, Ben, Lapid, Raz, Ohayon, Rony, Sipper, Moshe, Aperstein, Yehudit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Robustness of Kolmogorov-Arnold Networks: An Adversarial Perspective
by: Alter, Tal, et al.
Published: (2024)
by: Alter, Tal, et al.
Published: (2024)
Patch of Invisibility: Naturalistic Physical Black-Box Adversarial Attacks on Object Detectors
by: Lapid, Raz, et al.
Published: (2023)
by: Lapid, Raz, et al.
Published: (2023)
Fortify the Guardian, Not the Treasure: Resilient Adversarial Detectors
by: Lapid, Raz, et al.
Published: (2024)
by: Lapid, Raz, et al.
Published: (2024)
Pulling Back the Curtain: Unsupervised Adversarial Detection via Contrastive Auxiliary Networks
by: Mizrahi, Eylon, et al.
Published: (2025)
by: Mizrahi, Eylon, et al.
Published: (2025)
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
by: Ziv, Roee, et al.
Published: (2025)
by: Ziv, Roee, et al.
Published: (2025)
Open Sesame! Universal Black Box Jailbreaking of Large Language Models
by: Lapid, Raz, et al.
Published: (2023)
by: Lapid, Raz, et al.
Published: (2023)
AttackBench: Evaluating Gradient-based Attacks for Adversarial Examples
by: Cinà, Antonio Emanuele, et al.
Published: (2024)
by: Cinà, Antonio Emanuele, et al.
Published: (2024)
Improving the Perturbation-Based Explanation of Deepfake Detectors Through the Use of Adversarially-Generated Samples
by: Tsigos, Konstantinos, et al.
Published: (2025)
by: Tsigos, Konstantinos, et al.
Published: (2025)
How Generalizable are Deepfake Image Detectors? An Empirical Study
by: Li, Boquan, et al.
Published: (2023)
by: Li, Boquan, et al.
Published: (2023)
MVPatch: More Vivid Patch for Adversarial Camouflaged Attacks on Object Detectors in the Physical World
by: Zhou, Zheng, et al.
Published: (2023)
by: Zhou, Zheng, et al.
Published: (2023)
Where the Devil Hides: Deepfake Detectors Can No Longer Be Trusted
by: Yuan, Shuaiwei, et al.
Published: (2025)
by: Yuan, Shuaiwei, et al.
Published: (2025)
FRIDAY: Mitigating Unintentional Facial Identity in Deepfake Detectors Guided by Facial Recognizers
by: Kim, Younhun, et al.
Published: (2024)
by: Kim, Younhun, et al.
Published: (2024)
Vulnerabilities in AI-generated Image Detection: The Challenge of Adversarial Attacks
by: Diao, Yunfeng, et al.
Published: (2024)
by: Diao, Yunfeng, et al.
Published: (2024)
AdvART: Adversarial Art for Camouflaged Object Detection Attacks
by: Guesmi, Amira, et al.
Published: (2023)
by: Guesmi, Amira, et al.
Published: (2023)
CAPAA: Classifier-Agnostic Projector-Based Adversarial Attack
by: Li, Zhan, et al.
Published: (2025)
by: Li, Zhan, et al.
Published: (2025)
Improving Deepfake Detection with Reinforcement Learning-Based Adaptive Data Augmentation
by: Zhou, Yuxuan, et al.
Published: (2025)
by: Zhou, Yuxuan, et al.
Published: (2025)
Enhancing Remote Adversarial Patch Attacks on Face Detectors with Tiling and Scaling
by: Okano, Masora, et al.
Published: (2024)
by: Okano, Masora, et al.
Published: (2024)
Deciphering the Definition of Adversarial Robustness for post-hoc OOD Detectors
by: Lorenz, Peter, et al.
Published: (2024)
by: Lorenz, Peter, et al.
Published: (2024)
Detecting Adversarial Attacks in Semantic Segmentation via Uncertainty Estimation: A Deep Analysis
by: Maag, Kira, et al.
Published: (2024)
by: Maag, Kira, et al.
Published: (2024)
Signal Adversarial Examples Generation for Signal Detection Network via White-Box Attack
by: Li, Dongyang, et al.
Published: (2024)
by: Li, Dongyang, et al.
Published: (2024)
AVA: Inconspicuous Attribute Variation-based Adversarial Attack bypassing DeepFake Detection
by: Meng, Xiangtao, et al.
Published: (2023)
by: Meng, Xiangtao, et al.
Published: (2023)
Stealthy Multi-Task Adversarial Attacks
by: Guo, Jiacheng, et al.
Published: (2024)
by: Guo, Jiacheng, et al.
Published: (2024)
DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection
by: Zhao, Kangran, et al.
Published: (2025)
by: Zhao, Kangran, et al.
Published: (2025)
Attacking Transformers with Feature Diversity Adversarial Perturbation
by: Gao, Chenxing, et al.
Published: (2024)
by: Gao, Chenxing, et al.
Published: (2024)
Snowball Adversarial Attack on Traffic Sign Classification
by: Etim, Anthony, et al.
Published: (2025)
by: Etim, Anthony, et al.
Published: (2025)
Enhancing Adversarial Attacks: The Similar Target Method
by: Zhang, Shuo, et al.
Published: (2023)
by: Zhang, Shuo, et al.
Published: (2023)
2D-Malafide: Adversarial Attacks Against Face Deepfake Detection Systems
by: Galdi, Chiara, et al.
Published: (2024)
by: Galdi, Chiara, et al.
Published: (2024)
Bounding-box Watermarking: Defense against Model Extraction Attacks on Object Detectors
by: Koda, Satoru, et al.
Published: (2024)
by: Koda, Satoru, et al.
Published: (2024)
Query-Efficient Video Adversarial Attack with Stylized Logo
by: Tang, Duoxun, et al.
Published: (2024)
by: Tang, Duoxun, et al.
Published: (2024)
Fall Leaf Adversarial Attack on Traffic Sign Classification
by: Etim, Anthony, et al.
Published: (2024)
by: Etim, Anthony, et al.
Published: (2024)
Boosting the Transferability of Adversarial Attacks with Global Momentum Initialization
by: Wang, Jiafeng, et al.
Published: (2022)
by: Wang, Jiafeng, et al.
Published: (2022)
Delocate: Detection and Localization for Deepfake Videos with Randomly-Located Tampered Traces
by: Hu, Juan, et al.
Published: (2024)
by: Hu, Juan, et al.
Published: (2024)
Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
Reduced Spatial Dependency for More General Video-level Deepfake Detection
by: Chu, Beilin, et al.
Published: (2025)
by: Chu, Beilin, et al.
Published: (2025)
Benchmarking Adversarial Patch Selection and Location
by: Kimhi, Shai, et al.
Published: (2025)
by: Kimhi, Shai, et al.
Published: (2025)
SoK: Systematization and Benchmarking of Deepfake Detectors in a Unified Framework
by: Le, Binh M., et al.
Published: (2024)
by: Le, Binh M., et al.
Published: (2024)
Transferable Dual-Domain Feature Importance Attack against AI-Generated Image Detector
by: Zhu, Weiheng, et al.
Published: (2025)
by: Zhu, Weiheng, et al.
Published: (2025)
DeiTFake: Deepfake Detection Model using DeiT Multi-Stage Training
by: Kumar, Saksham, et al.
Published: (2025)
by: Kumar, Saksham, et al.
Published: (2025)
Revisiting Adversarial Patch Defenses on Object Detectors: Unified Evaluation, Large-Scale Dataset, and New Insights
by: Zheng, Junhao, et al.
Published: (2025)
by: Zheng, Junhao, et al.
Published: (2025)
Anomaly Unveiled: Securing Image Classification against Adversarial Patch Attacks
by: Chattopadhyay, Nandish, et al.
Published: (2024)
by: Chattopadhyay, Nandish, et al.
Published: (2024)
Similar Items
-
On the Robustness of Kolmogorov-Arnold Networks: An Adversarial Perspective
by: Alter, Tal, et al.
Published: (2024) -
Patch of Invisibility: Naturalistic Physical Black-Box Adversarial Attacks on Object Detectors
by: Lapid, Raz, et al.
Published: (2023) -
Fortify the Guardian, Not the Treasure: Resilient Adversarial Detectors
by: Lapid, Raz, et al.
Published: (2024) -
Pulling Back the Curtain: Unsupervised Adversarial Detection via Contrastive Auxiliary Networks
by: Mizrahi, Eylon, et al.
Published: (2025) -
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
by: Ziv, Roee, et al.
Published: (2025)