Evolving from Single-modal to Multi-modal Facial Deepfake Detection: Progress and Challenges
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Ping, Tao, Qiqi, Zhou, Joey Tianyi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-modal Deepfake Detection and Localization with FPN-Transformer
di: Zheng, Chende, et al.
Pubblicazione: (2025)
di: Zheng, Chende, et al.
Pubblicazione: (2025)
PUDD: Towards Robust Multi-modal Prototype-based Deepfake Detection
di: Pellcier, Alvaro Lopez, et al.
Pubblicazione: (2024)
di: Pellcier, Alvaro Lopez, et al.
Pubblicazione: (2024)
Learning Real Facial Concepts for Independent Deepfake Detection
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
Progressive Multi-modal Conditional Prompt Tuning
di: Qiu, Xiaoyu, et al.
Pubblicazione: (2024)
di: Qiu, Xiaoyu, et al.
Pubblicazione: (2024)
Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey
di: Nguyen-Le, Hong-Hanh, et al.
Pubblicazione: (2024)
di: Nguyen-Le, Hong-Hanh, et al.
Pubblicazione: (2024)
EvolveReason: Self-Evolving Reasoning Paradigm for Explainable Deepfake Facial Image Identification
di: Zhou, Binjia, et al.
Pubblicazione: (2026)
di: Zhou, Binjia, et al.
Pubblicazione: (2026)
MDMP: Multi-modal Diffusion for supervised Motion Predictions with uncertainty
di: Bringer, Leo, et al.
Pubblicazione: (2024)
di: Bringer, Leo, et al.
Pubblicazione: (2024)
Hierarchical Deep Fusion Framework for Multi-dimensional Facial Forgery Detection -- The 2024 Global Deepfake Image Detection Challenge
di: Wang, Kohou, et al.
Pubblicazione: (2025)
di: Wang, Kohou, et al.
Pubblicazione: (2025)
FakeOut: Leveraging Out-of-domain Self-supervision for Multi-modal Video Deepfake Detection
di: Knafo, Gil, et al.
Pubblicazione: (2022)
di: Knafo, Gil, et al.
Pubblicazione: (2022)
Cross-modal Active Complementary Learning with Self-refining Correspondence
di: Qin, Yang, et al.
Pubblicazione: (2023)
di: Qin, Yang, et al.
Pubblicazione: (2023)
MMHead: Towards Fine-grained Multi-modal 3D Facial Animation
di: Wu, Sijing, et al.
Pubblicazione: (2024)
di: Wu, Sijing, et al.
Pubblicazione: (2024)
Investigating the Viability of Employing Multi-modal Large Language Models in the Context of Audio Deepfake Detection
di: Chuchra, Akanksha, et al.
Pubblicazione: (2026)
di: Chuchra, Akanksha, et al.
Pubblicazione: (2026)
Multi-modal Transfer Learning for Dynamic Facial Emotion Recognition in the Wild
di: Engel, Ezra, et al.
Pubblicazione: (2025)
di: Engel, Ezra, et al.
Pubblicazione: (2025)
Cross-modal Proxy Evolving for OOD Detection with Vision-Language Models
di: Tang, Hao, et al.
Pubblicazione: (2026)
di: Tang, Hao, et al.
Pubblicazione: (2026)
Bridging the Gap between Multi-focus and Multi-modal: A Focused Integration Framework for Multi-modal Image Fusion
di: Li, Xilai, et al.
Pubblicazione: (2023)
di: Li, Xilai, et al.
Pubblicazione: (2023)
Facial Attractiveness Prediction in Live Streaming: A New Benchmark and Multi-modal Method
di: Li, Hui, et al.
Pubblicazione: (2025)
di: Li, Hui, et al.
Pubblicazione: (2025)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
di: Zhang, Yaning, et al.
Pubblicazione: (2024)
di: Zhang, Yaning, et al.
Pubblicazione: (2024)
TokenSwap: Backdoor Attack on the Compositional Understanding of Large Vision-Language Models
di: Zhang, Zhifang, et al.
Pubblicazione: (2025)
di: Zhang, Zhifang, et al.
Pubblicazione: (2025)
Enhancing Incomplete Multi-modal Brain Tumor Segmentation with Intra-modal Asymmetry and Inter-modal Dependency
di: Liu, Weide, et al.
Pubblicazione: (2024)
di: Liu, Weide, et al.
Pubblicazione: (2024)
Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
di: Wen, Zichen, et al.
Pubblicazione: (2025)
di: Wen, Zichen, et al.
Pubblicazione: (2025)
RGB-T Object Detection via Group Shuffled Multi-receptive Attention and Multi-modal Supervision
di: Wang, Jinzhong, et al.
Pubblicazione: (2024)
di: Wang, Jinzhong, et al.
Pubblicazione: (2024)
Beyond Deepfake vs Real: Facial Deepfake Detection in the Open-Set Paradigm
di: Bahavan, Nadarasar, et al.
Pubblicazione: (2025)
di: Bahavan, Nadarasar, et al.
Pubblicazione: (2025)
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis
di: Wang, Gui, et al.
Pubblicazione: (2026)
di: Wang, Gui, et al.
Pubblicazione: (2026)
Detecting Deepfake Talking Heads from Facial Biometric Anomalies
di: Norman, Justin D., et al.
Pubblicazione: (2025)
di: Norman, Justin D., et al.
Pubblicazione: (2025)
Large Multi-modal Models Can Interpret Features in Large Multi-modal Models
di: Zhang, Kaichen, et al.
Pubblicazione: (2024)
di: Zhang, Kaichen, et al.
Pubblicazione: (2024)
Multi-modal Attribute Prompting for Vision-Language Models
di: Liu, Xin, et al.
Pubblicazione: (2024)
di: Liu, Xin, et al.
Pubblicazione: (2024)
Personalizing Causal Audio-Driven Facial Motion via Dynamic Multi-modal Retrieval
di: Chu, Xuangeng, et al.
Pubblicazione: (2026)
di: Chu, Xuangeng, et al.
Pubblicazione: (2026)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
di: He, Wen-Jue, et al.
Pubblicazione: (2025)
di: He, Wen-Jue, et al.
Pubblicazione: (2025)
FauForensics: Boosting Audio-Visual Deepfake Detection with Facial Action Units
di: Wang, Jian, et al.
Pubblicazione: (2025)
di: Wang, Jian, et al.
Pubblicazione: (2025)
SimDistill: Simulated Multi-modal Distillation for BEV 3D Object Detection
di: Zhao, Haimei, et al.
Pubblicazione: (2023)
di: Zhao, Haimei, et al.
Pubblicazione: (2023)
DDL: A Large-Scale Datasets for Deepfake Detection and Localization in Diversified Real-World Scenarios
di: Miao, Changtao, et al.
Pubblicazione: (2025)
di: Miao, Changtao, et al.
Pubblicazione: (2025)
Modelship Attribution: Tracing Multi-Stage Manipulations Across Generative Models
di: Tan, Zhiya, et al.
Pubblicazione: (2025)
di: Tan, Zhiya, et al.
Pubblicazione: (2025)
mmWalk: Towards Multi-modal Multi-view Walking Assistance
di: Ying, Kedi, et al.
Pubblicazione: (2025)
di: Ying, Kedi, et al.
Pubblicazione: (2025)
Awesome Multi-modal Object Tracking
di: Zhang, Chunhui, et al.
Pubblicazione: (2024)
di: Zhang, Chunhui, et al.
Pubblicazione: (2024)
Multi-modality Affinity Inference for Weakly Supervised 3D Semantic Segmentation
di: Li, Xiawei, et al.
Pubblicazione: (2023)
di: Li, Xiawei, et al.
Pubblicazione: (2023)
Knowledge-Guided Prompt Learning for Deepfake Facial Image Detection
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
di: Li, Kunchang, et al.
Pubblicazione: (2023)
di: Li, Kunchang, et al.
Pubblicazione: (2023)
A Spatial-Frequency Aware Multi-Scale Fusion Network for Real-Time Deepfake Detection
di: Lv, Libo, et al.
Pubblicazione: (2025)
di: Lv, Libo, et al.
Pubblicazione: (2025)
Human Knowledge Integrated Multi-modal Learning for Single Source Domain Generalization
di: Banerjee, Ayan, et al.
Pubblicazione: (2026)
di: Banerjee, Ayan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Multi-modal Deepfake Detection and Localization with FPN-Transformer
di: Zheng, Chende, et al.
Pubblicazione: (2025) -
PUDD: Towards Robust Multi-modal Prototype-based Deepfake Detection
di: Pellcier, Alvaro Lopez, et al.
Pubblicazione: (2024) -
Learning Real Facial Concepts for Independent Deepfake Detection
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025) -
Progressive Multi-modal Conditional Prompt Tuning
di: Qiu, Xiaoyu, et al.
Pubblicazione: (2024) -
Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey
di: Nguyen-Le, Hong-Hanh, et al.
Pubblicazione: (2024)