Unsupervised Multimodal Deepfake Detection Using Intra- and Cross-Modal Inconsistencies
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Mulin, Khayatkhoei, Mahyar, Mathai, Joe, AbdAlmageed, Wael |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ManiFPT: Defining and Analyzing Fingerprints of Generative Models
by: Song, Hae Jin, et al.
Published: (2024)
by: Song, Hae Jin, et al.
Published: (2024)
An Investigation on The Position Encoding in Vision-Based Dynamics Prediction
by: Zhu, Jiageng, et al.
Published: (2024)
by: Zhu, Jiageng, et al.
Published: (2024)
Look, Learn and Leverage (L$^3$): Mitigating Visual-Domain Shift and Discovering Intrinsic Relations via Symbolic Alignment
by: Xie, Hanchen, et al.
Published: (2024)
by: Xie, Hanchen, et al.
Published: (2024)
A Neuro-Symbolic Framework Combining Inductive and Deductive Reasoning for Autonomous Driving Planning
by: Wei, Hongyan, et al.
Published: (2026)
by: Wei, Hongyan, et al.
Published: (2026)
TRIGS: Trojan Identification from Gradient-based Signatures
by: Hussein, Mohamed E., et al.
Published: (2023)
by: Hussein, Mohamed E., et al.
Published: (2023)
AS2 -- Attention-Based Soft Answer Sets: An End-to-End Differentiable Neuro-Soft-Symbolic Reasoning Architecture
by: AbdAlmageed, Wael
Published: (2026)
by: AbdAlmageed, Wael
Published: (2026)
Towards Perceiving Small Visual Details in Zero-shot Visual Question Answering with Multimodal LLMs
by: Zhang, Jiarui, et al.
Published: (2023)
by: Zhang, Jiarui, et al.
Published: (2023)
MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs
by: Zhang, Jiarui, et al.
Published: (2025)
by: Zhang, Jiarui, et al.
Published: (2025)
Inconsistency-aware Multimodal Schrödinger Bridge for Deepfake Localization
by: Xiong, Jiayu, et al.
Published: (2026)
by: Xiong, Jiayu, et al.
Published: (2026)
Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints
by: Ahmad, Wasim, et al.
Published: (2026)
by: Ahmad, Wasim, et al.
Published: (2026)
Causal Representation Learning on High-Dimensional Data: Benchmarks, Reproducibility, and Evaluation Metrics
by: Sadeghi, Alireza, et al.
Published: (2026)
by: Sadeghi, Alireza, et al.
Published: (2026)
Towards Generalizable Deepfake Detection with Spatial-Frequency Collaborative Learning and Hierarchical Cross-Modal Fusion
by: Qiao, Mengyu, et al.
Published: (2025)
by: Qiao, Mengyu, et al.
Published: (2025)
Exploring Perceptual Limitation of Multimodal Large Language Models
by: Zhang, Jiarui, et al.
Published: (2024)
by: Zhang, Jiarui, et al.
Published: (2024)
Beyond Flicker: Detecting Kinematic Inconsistencies for Generalizable Deepfake Video Detection
by: Cobo, Alejandro, et al.
Published: (2025)
by: Cobo, Alejandro, et al.
Published: (2025)
CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation
by: Du, Yuxuan, et al.
Published: (2025)
by: Du, Yuxuan, et al.
Published: (2025)
Explicit Correlation Learning for Generalizable Cross-Modal Deepfake Detection
by: Yu, Cai, et al.
Published: (2024)
by: Yu, Cai, et al.
Published: (2024)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
by: Wei, Fangda, et al.
Published: (2026)
by: Wei, Fangda, et al.
Published: (2026)
Learning Spatiotemporal Inconsistency via Thumbnail Layout for Face Deepfake Detection
by: Xu, Yuting, et al.
Published: (2024)
by: Xu, Yuting, et al.
Published: (2024)
A Critical Review of Predominant Bias in Neural Networks
by: Li, Jiazhi, et al.
Published: (2025)
by: Li, Jiazhi, et al.
Published: (2025)
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2025)
by: Datta, Soumyya Kanti, et al.
Published: (2025)
Exposing Lip-syncing Deepfakes from Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2024)
by: Datta, Soumyya Kanti, et al.
Published: (2024)
IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
UMCL: Unimodal-generated Multimodal Contrastive Learning for Cross-compression-rate Deepfake Detection
by: Lai, Ching-Yi, et al.
Published: (2025)
by: Lai, Ching-Yi, et al.
Published: (2025)
Detecting Audio-Visual Deepfakes with Fine-Grained Inconsistencies
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Compositional Understanding
by: Zhang, Le, et al.
Published: (2023)
by: Zhang, Le, et al.
Published: (2023)
Cross-Domain Object Detection Using Unsupervised Image Translation
by: Arruda, Vinicius F., et al.
Published: (2026)
by: Arruda, Vinicius F., et al.
Published: (2026)
Pyramidal Adaptive Cross-Gating for Multimodal Detection
by: Gu, Zidong, et al.
Published: (2025)
by: Gu, Zidong, et al.
Published: (2025)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Incomplete Multimodal Industrial Anomaly Detection via Cross-Modal Distillation
by: Sui, Wenbo, et al.
Published: (2024)
by: Sui, Wenbo, et al.
Published: (2024)
Bridging Cross-task Protocol Inconsistency for Distillation in Dense Object Detection
by: Yang, Longrong, et al.
Published: (2023)
by: Yang, Longrong, et al.
Published: (2023)
CrossDF: Improving Cross-Domain Deepfake Detection with Deep Information Decomposition
by: Yang, Shanmin, et al.
Published: (2023)
by: Yang, Shanmin, et al.
Published: (2023)
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
by: Liu, Yuxin, et al.
Published: (2026)
by: Liu, Yuxin, et al.
Published: (2026)
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies
by: Astrid, Marcella, et al.
Published: (2025)
by: Astrid, Marcella, et al.
Published: (2025)
WWW: Where, Which and Whatever Enhancing Interpretability in Multimodal Deepfake Detection
by: Jung, Juho, et al.
Published: (2024)
by: Jung, Juho, et al.
Published: (2024)
Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection
by: Li, Tianxiao, et al.
Published: (2026)
by: Li, Tianxiao, et al.
Published: (2026)
Next-Frame Feature Prediction for Multimodal Deepfake Detection and Temporal Localization
by: Anshul, Ashutosh, et al.
Published: (2025)
by: Anshul, Ashutosh, et al.
Published: (2025)
Fact or Fake? Assessing the Role of Deepfake Detectors in Multimodal Misinformation Detection
by: Sagar, A S M Sharifuzzaman, et al.
Published: (2026)
by: Sagar, A S M Sharifuzzaman, et al.
Published: (2026)
DFBench: Benchmarking Deepfake Image Detection Capability of Large Multimodal Models
by: Wang, Jiarui, et al.
Published: (2025)
by: Wang, Jiarui, et al.
Published: (2025)
Cross-Modal Prototype Allocation: Unsupervised Slide Representation Learning via Patch-Text Contrast in Computational Pathology
by: Chen, Yuxuan, et al.
Published: (2025)
by: Chen, Yuxuan, et al.
Published: (2025)
Reevaluating the Intra-Modal Misalignment Hypothesis in CLIP
by: Herzog, Jonas, et al.
Published: (2026)
by: Herzog, Jonas, et al.
Published: (2026)
Similar Items
-
ManiFPT: Defining and Analyzing Fingerprints of Generative Models
by: Song, Hae Jin, et al.
Published: (2024) -
An Investigation on The Position Encoding in Vision-Based Dynamics Prediction
by: Zhu, Jiageng, et al.
Published: (2024) -
Look, Learn and Leverage (L$^3$): Mitigating Visual-Domain Shift and Discovering Intrinsic Relations via Symbolic Alignment
by: Xie, Hanchen, et al.
Published: (2024) -
A Neuro-Symbolic Framework Combining Inductive and Deductive Reasoning for Autonomous Driving Planning
by: Wei, Hongyan, et al.
Published: (2026) -
TRIGS: Trojan Identification from Gradient-based Signatures
by: Hussein, Mohamed E., et al.
Published: (2023)