DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Kangran, Chen, Yupeng, Zhang, Xiaoyu, Chen, Yize, Guan, Weinan, Chen, Baicheng, Sun, Chengzhe, Datta, Soumyya Kanti, Liu, Qingshan, Lyu, Siwei, Wu, Baoyuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PIA: Deepfake Detection Using Phoneme-Temporal and Identity-Dynamic Analysis
by: Datta, Soumyya Kanti, et al.
Published: (2025)
by: Datta, Soumyya Kanti, et al.
Published: (2025)
Cyber Vaccine for Deepfake Immunity
by: Chang, Ching-Chun, et al.
Published: (2023)
by: Chang, Ching-Chun, et al.
Published: (2023)
Exposing Lip-syncing Deepfakes from Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2024)
by: Datta, Soumyya Kanti, et al.
Published: (2024)
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2025)
by: Datta, Soumyya Kanti, et al.
Published: (2025)
SafeEar: Content Privacy-Preserving Audio Deepfake Detection
by: Li, Xinfeng, et al.
Published: (2024)
by: Li, Xinfeng, et al.
Published: (2024)
X2-DFD: A framework for eXplainable and eXtendable Deepfake Detection
by: Chen, Yize, et al.
Published: (2024)
by: Chen, Yize, et al.
Published: (2024)
A Survey on Speech Deepfake Detection
by: Li, Menglu, et al.
Published: (2024)
by: Li, Menglu, et al.
Published: (2024)
Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
Hindi audio-video-Deepfake (HAV-DF): A Hindi language-based Audio-video Deepfake Dataset
by: Kaur, Sukhandeep, et al.
Published: (2024)
by: Kaur, Sukhandeep, et al.
Published: (2024)
Every Breath You Don't Take: Deepfake Speech Detection Using Breath
by: Layton, Seth, et al.
Published: (2024)
by: Layton, Seth, et al.
Published: (2024)
Can ChatGPT Detect DeepFakes? A Study of Using Multimodal Large Language Models for Media Forensics
by: Jia, Shan, et al.
Published: (2024)
by: Jia, Shan, et al.
Published: (2024)
VocalCrypt: Novel Active Defense Against Deepfake Voice Based on Masking Effect
by: Fei, Qingyuan, et al.
Published: (2025)
by: Fei, Qingyuan, et al.
Published: (2025)
A Unified Framework for Modality-Agnostic Deepfakes Detection
by: Yu, Cai, et al.
Published: (2023)
by: Yu, Cai, et al.
Published: (2023)
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies
by: Astrid, Marcella, et al.
Published: (2025)
by: Astrid, Marcella, et al.
Published: (2025)
Unveiling Covert Toxicity in Multimodal Data via Toxicity Association Graphs: A Graph-Based Metric and Interpretable Detection Framework
by: Wu, Guanzong, et al.
Published: (2026)
by: Wu, Guanzong, et al.
Published: (2026)
ARGUS: Defending Against Multimodal Indirect Prompt Injection via Steering Instruction-Following Behavior
by: Lu, Weikai, et al.
Published: (2025)
by: Lu, Weikai, et al.
Published: (2025)
Multimodal Unlearnable Examples: Protecting Data against Multimodal Contrastive Learning
by: Liu, Xinwei, et al.
Published: (2024)
by: Liu, Xinwei, et al.
Published: (2024)
PRISM-XR: Empowering Privacy-Aware XR Collaboration with Multimodal Large Language Models
by: Chen, Jiangong, et al.
Published: (2026)
by: Chen, Jiangong, et al.
Published: (2026)
In Anticipation of Perfect Deepfake: Identity-anchored Artifact-agnostic Detection under Rebalanced Deepfake Detection Protocol
by: Wang, Wei-Han, et al.
Published: (2024)
by: Wang, Wei-Han, et al.
Published: (2024)
Seeing, Hearing, and Knowing Together: Multimodal Strategies in Deepfake Videos Detection
by: Chen, Chen, et al.
Published: (2026)
by: Chen, Chen, et al.
Published: (2026)
Digital Fingerprinting on Multimedia: A Survey
by: Chen, Wendi, et al.
Published: (2024)
by: Chen, Wendi, et al.
Published: (2024)
Social Media Authentication and Combating Deepfakes using Semi-fragile Invisible Image Watermarking
by: Nadimpalli, Aakash Varma, et al.
Published: (2024)
by: Nadimpalli, Aakash Varma, et al.
Published: (2024)
Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models
by: Wen, Jinming, et al.
Published: (2025)
by: Wen, Jinming, et al.
Published: (2025)
SEA: Low-Resource Safety Alignment for Multimodal Large Language Models via Synthetic Embeddings
by: Lu, Weikai, et al.
Published: (2025)
by: Lu, Weikai, et al.
Published: (2025)
RoboKA: KAN Informed Multimodal Learning for RoboCall Surveillance System
by: Choudhury, Nitin, et al.
Published: (2026)
by: Choudhury, Nitin, et al.
Published: (2026)
Security Analysis of Thumbnail-Preserving Image Encryption and a New Framework
by: Xie, Dong, et al.
Published: (2025)
by: Xie, Dong, et al.
Published: (2025)
Intelligent Carrier Allocation: A Cross-Modal Reasoning Framework for Adaptive Multimodal Steganography
by: Das, Abhirup, et al.
Published: (2025)
by: Das, Abhirup, et al.
Published: (2025)
DeepFake-O-Meter v2.0: An Open Platform for DeepFake Detection
by: Ju, Yan, et al.
Published: (2024)
by: Ju, Yan, et al.
Published: (2024)
Beyond Text: Multimodal Jailbreaking of Vision-Language and Audio Models through Perceptually Simple Transformations
by: Kumar, Divyanshu, et al.
Published: (2025)
by: Kumar, Divyanshu, et al.
Published: (2025)
Divide and Conquer: Multimodal Video Deepfake Detection via Cross-Modal Fusion and Localization
by: Li, Qingcao, et al.
Published: (2026)
by: Li, Qingcao, et al.
Published: (2026)
Deepfake Detection: A Comprehensive Survey from the Reliability Perspective
by: Wang, Tianyi, et al.
Published: (2022)
by: Wang, Tianyi, et al.
Published: (2022)
MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio
by: Li, Qingcao, et al.
Published: (2026)
by: Li, Qingcao, et al.
Published: (2026)
Persistence of Backdoor-based Watermarks for Neural Networks: A Comprehensive Evaluation
by: Ngo, Anh Tu, et al.
Published: (2025)
by: Ngo, Anh Tu, et al.
Published: (2025)
Enkidu: Universal Frequential Perturbation for Real-Time Audio Privacy Protection against Voice Deepfakes
by: Feng, Zhou, et al.
Published: (2025)
by: Feng, Zhou, et al.
Published: (2025)
Multimodal Reasoning with LLM for Encrypted Traffic Interpretation: A Benchmark
by: Zhang, Longgang, et al.
Published: (2026)
by: Zhang, Longgang, et al.
Published: (2026)
BlackboxBench: A Comprehensive Benchmark of Black-box Adversarial Attacks
by: Zheng, Meixi, et al.
Published: (2023)
by: Zheng, Meixi, et al.
Published: (2023)
Optimization-Free Universal Watermark Forgery with Regenerative Diffusion Models
by: Zhu, Chaoyi, et al.
Published: (2025)
by: Zhu, Chaoyi, et al.
Published: (2025)
Compressed Deepfake Video Detection Based on 3D Spatiotemporal Trajectories
by: Chen, Zongmei, et al.
Published: (2024)
by: Chen, Zongmei, et al.
Published: (2024)
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction
by: Fu, Jiyuan, et al.
Published: (2024)
by: Fu, Jiyuan, et al.
Published: (2024)
Towards Source Attribution of Singing Voice Deepfake with Multimodal Foundation Models
by: Phukan, Orchid Chetia, et al.
Published: (2025)
by: Phukan, Orchid Chetia, et al.
Published: (2025)
Similar Items
-
PIA: Deepfake Detection Using Phoneme-Temporal and Identity-Dynamic Analysis
by: Datta, Soumyya Kanti, et al.
Published: (2025) -
Cyber Vaccine for Deepfake Immunity
by: Chang, Ching-Chun, et al.
Published: (2023) -
Exposing Lip-syncing Deepfakes from Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2024) -
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2025) -
SafeEar: Content Privacy-Preserving Audio Deepfake Detection
by: Li, Xinfeng, et al.
Published: (2024)