MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Wenbo, Lu, Wei, Luo, Xiangyang, Zhou, Jiantao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Weakly Supervised Multimodal Temporal Forgery Localization via Multitask Learning
von: Xu, Wenbo, et al.
Veröffentlicht: (2025)
von: Xu, Wenbo, et al.
Veröffentlicht: (2025)
GLCF: A Global-Local Multimodal Coherence Analysis Framework for Talking Face Generation Detection
von: Chen, Xiaocan, et al.
Veröffentlicht: (2024)
von: Chen, Xiaocan, et al.
Veröffentlicht: (2024)
RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News Detection
von: Yu, Xinquan, et al.
Veröffentlicht: (2024)
von: Yu, Xinquan, et al.
Veröffentlicht: (2024)
A Multimodal Deviation Perceiving Framework for Weakly-Supervised Temporal Forgery Localization
von: Xu, Wenbo, et al.
Veröffentlicht: (2025)
von: Xu, Wenbo, et al.
Veröffentlicht: (2025)
Unlocking the Capabilities of Large Vision-Language Models for Generalizable and Explainable Deepfake Detection
von: Yu, Peipeng, et al.
Veröffentlicht: (2025)
von: Yu, Peipeng, et al.
Veröffentlicht: (2025)
Towards Open-world Generalized Deepfake Detection: General Feature Extraction via Unsupervised Domain Adaptation
von: Guo, Midou, et al.
Veröffentlicht: (2025)
von: Guo, Midou, et al.
Veröffentlicht: (2025)
SLIP: Structural-aware Language-Image Pretraining for Vision-Language Alignment
von: Lu, Wenbo
Veröffentlicht: (2025)
von: Lu, Wenbo
Veröffentlicht: (2025)
EDVD-LLaMA: Explainable Deepfake Video Detection via Multimodal Large Language Model Reasoning
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Weakly-Supervised Image Forgery Localization via Vision-Language Collaborative Reasoning Framework
von: Sheng, Ziqi, et al.
Veröffentlicht: (2025)
von: Sheng, Ziqi, et al.
Veröffentlicht: (2025)
CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation
von: Du, Yuxuan, et al.
Veröffentlicht: (2025)
von: Du, Yuxuan, et al.
Veröffentlicht: (2025)
SUMI-IFL: An Information-Theoretic Framework for Image Forgery Localization with Sufficiency and Minimality Constraints
von: Sheng, Ziqi, et al.
Veröffentlicht: (2024)
von: Sheng, Ziqi, et al.
Veröffentlicht: (2024)
Generalizable Synthetic Image Detection via Language-guided Contrastive Learning
von: Wu, Haiwei, et al.
Veröffentlicht: (2023)
von: Wu, Haiwei, et al.
Veröffentlicht: (2023)
Vision-Language Feature Alignment for Road Anomaly Segmentation
von: He, Zhuolin, et al.
Veröffentlicht: (2026)
von: He, Zhuolin, et al.
Veröffentlicht: (2026)
Fine-grained Multiple Supervisory Network for Multi-modal Manipulation Detecting and Grounding
von: Yu, Xinquan, et al.
Veröffentlicht: (2025)
von: Yu, Xinquan, et al.
Veröffentlicht: (2025)
CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection
von: Khan, Sohail Ahmed, et al.
Veröffentlicht: (2024)
von: Khan, Sohail Ahmed, et al.
Veröffentlicht: (2024)
AIGI-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models
von: Zhou, Ziyin, et al.
Veröffentlicht: (2025)
von: Zhou, Ziyin, et al.
Veröffentlicht: (2025)
Unleashing Vision-Language Semantics for Deepfake Video Detection
von: Zhu, Jiawen, et al.
Veröffentlicht: (2026)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2026)
Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints
von: Ahmad, Wasim, et al.
Veröffentlicht: (2026)
von: Ahmad, Wasim, et al.
Veröffentlicht: (2026)
From Prediction to Explanation: Multimodal, Explainable, and Interactive Deepfake Detection Framework for Non-Expert Users
von: Tariq, Shahroz, et al.
Veröffentlicht: (2025)
von: Tariq, Shahroz, et al.
Veröffentlicht: (2025)
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
CIEC: Coupling Implicit and Explicit Cues for Multimodal Weakly Supervised Manipulation Localization
von: Yu, Xinquan, et al.
Veröffentlicht: (2026)
von: Yu, Xinquan, et al.
Veröffentlicht: (2026)
Large Vision-Language Model Alignment and Misalignment: A Survey Through the Lens of Explainability
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
Training-Free Multimodal Deepfake Detection via Graph Reasoning
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
Semantic Alignment for Multimodal Large Language Models
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
VQAThinker: Exploring Generalizable and Explainable Video Quality Assessment via Reinforcement Learning
von: Cao, Linhan, et al.
Veröffentlicht: (2025)
von: Cao, Linhan, et al.
Veröffentlicht: (2025)
AuthGuard: Generalizable Deepfake Detection via Language Guidance
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
A Timely Survey on Vision Transformer for Deepfake Detection
von: Wang, Zhikan, et al.
Veröffentlicht: (2024)
von: Wang, Zhikan, et al.
Veröffentlicht: (2024)
Explainable Deepfake Detection with RL Enhanced Self-Blended Images
von: Jiang, Ning, et al.
Veröffentlicht: (2026)
von: Jiang, Ning, et al.
Veröffentlicht: (2026)
BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
Accurate and Scalable Multimodal Pathology Retrieval via Attentive Vision-Language Alignment
von: Wang, Hongyi, et al.
Veröffentlicht: (2025)
von: Wang, Hongyi, et al.
Veröffentlicht: (2025)
PaAgent: Portrait-Aware Image Restoration Agent via Subjective-Objective Reinforcement Learning
von: Wang, Yijian, et al.
Veröffentlicht: (2026)
von: Wang, Yijian, et al.
Veröffentlicht: (2026)
EvolveReason: Self-Evolving Reasoning Paradigm for Explainable Deepfake Facial Image Identification
von: Zhou, Binjia, et al.
Veröffentlicht: (2026)
von: Zhou, Binjia, et al.
Veröffentlicht: (2026)
Capture Artifacts via Progressive Disentangling and Purifying Blended Identities for Deepfake Detection
von: Zhou, Weijie, et al.
Veröffentlicht: (2024)
von: Zhou, Weijie, et al.
Veröffentlicht: (2024)
Towards Generalizable Deepfake Detection via Real Distribution Bias Correction
von: Liu, Ming-Hui, et al.
Veröffentlicht: (2026)
von: Liu, Ming-Hui, et al.
Veröffentlicht: (2026)
AlignMMBench: Evaluating Chinese Multimodal Alignment in Large Vision-Language Models
von: Wu, Yuhang, et al.
Veröffentlicht: (2024)
von: Wu, Yuhang, et al.
Veröffentlicht: (2024)
Modest-Align: Data-Efficient Alignment for Vision-Language Models
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
DFBench: Benchmarking Deepfake Image Detection Capability of Large Multimodal Models
von: Wang, Jiarui, et al.
Veröffentlicht: (2025)
von: Wang, Jiarui, et al.
Veröffentlicht: (2025)
Suppressing Gradient Conflict for Generalizable Deepfake Detection
von: Liu, Ming-Hui, et al.
Veröffentlicht: (2025)
von: Liu, Ming-Hui, et al.
Veröffentlicht: (2025)
Lossless Copyright Protection via Intrinsic Model Fingerprinting
von: Chen, Lingxiao, et al.
Veröffentlicht: (2026)
von: Chen, Lingxiao, et al.
Veröffentlicht: (2026)
Safeguarding Facial Identity against Diffusion-based Face Swapping via Cascading Pathway Disruption
von: Wang, Liqin, et al.
Veröffentlicht: (2026)
von: Wang, Liqin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Weakly Supervised Multimodal Temporal Forgery Localization via Multitask Learning
von: Xu, Wenbo, et al.
Veröffentlicht: (2025) -
GLCF: A Global-Local Multimodal Coherence Analysis Framework for Talking Face Generation Detection
von: Chen, Xiaocan, et al.
Veröffentlicht: (2024) -
RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News Detection
von: Yu, Xinquan, et al.
Veröffentlicht: (2024) -
A Multimodal Deviation Perceiving Framework for Weakly-Supervised Temporal Forgery Localization
von: Xu, Wenbo, et al.
Veröffentlicht: (2025) -
Unlocking the Capabilities of Large Vision-Language Models for Generalizable and Explainable Deepfake Detection
von: Yu, Peipeng, et al.
Veröffentlicht: (2025)