The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ai, Wei, Tan, Yilong, Shou, Yuntao, Meng, Tao, Chen, Haowen, He, Zhixiong, Li, Keqin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CILF-CIAE: CLIP-driven Image-Language Fusion for Correcting Inverse Age Estimation
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
LRA-GNN: Latent Relation-Aware Graph Neural Network with Initial and Dynamic Residual for Facial Age Estimation
von: Zhang, Yiping, et al.
Veröffentlicht: (2025)
von: Zhang, Yiping, et al.
Veröffentlicht: (2025)
GroupFace: Imbalanced Age Estimation Based on Multi-hop Attention Graph Convolutional Network and Group-aware Margin Optimization
von: Zhang, Yiping, et al.
Veröffentlicht: (2024)
von: Zhang, Yiping, et al.
Veröffentlicht: (2024)
Multimodal Large Language Models Meet Multimodal Emotion Recognition and Reasoning: A Survey
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
A Multi-view Mask Contrastive Learning Graph Convolutional Neural Network for Age Estimation
von: Zhang, Yiping, et al.
Veröffentlicht: (2024)
von: Zhang, Yiping, et al.
Veröffentlicht: (2024)
A Flow Model with Low-Rank Transformers for Incomplete Multimodal Survival Analysis
von: Yin, Yi, et al.
Veröffentlicht: (2025)
von: Yin, Yi, et al.
Veröffentlicht: (2025)
Graph Information Bottleneck for Remote Sensing Segmentation
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
von: Wang, Yunsheng, et al.
Veröffentlicht: (2026)
von: Wang, Yunsheng, et al.
Veröffentlicht: (2026)
Dynamic Fusion-Aware Graph Convolutional Neural Network for Multimodal Emotion Recognition in Conversations
von: Meng, Tao, et al.
Veröffentlicht: (2026)
von: Meng, Tao, et al.
Veröffentlicht: (2026)
FBLNet: FeedBack Loop Network for Driver Attention Prediction
von: Chen, Yilong, et al.
Veröffentlicht: (2022)
von: Chen, Yilong, et al.
Veröffentlicht: (2022)
On-Road Object Importance Estimation: A New Dataset and A Model with Multi-Fold Top-Down Guidance
von: Nan, Zhixiong, et al.
Veröffentlicht: (2024)
von: Nan, Zhixiong, et al.
Veröffentlicht: (2024)
Hallucination of Multimodal Large Language Models: A Survey
von: Bai, Zechen, et al.
Veröffentlicht: (2024)
von: Bai, Zechen, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on Multi-modal Conversational Emotion Recognition with Deep Learning
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
ThinkFake: Reasoning in Multimodal Large Language Models for AI-Generated Image Detection
von: Huang, Tai-Ming, et al.
Veröffentlicht: (2025)
von: Huang, Tai-Ming, et al.
Veröffentlicht: (2025)
Dynamic Graph Neural ODE Network for Multi-modal Emotion Recognition in Conversation
von: Shou, Yuntao, et al.
Veröffentlicht: (2024)
von: Shou, Yuntao, et al.
Veröffentlicht: (2024)
Adversarial Representation with Intra-Modal and Inter-Modal Graph Contrastive Learning for Multimodal Emotion Recognition
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
VisionNVS: Self-Supervised Inpainting for Novel View Synthesis under the Virtual-Shift Paradigm
von: Lu, Hongbo, et al.
Veröffentlicht: (2026)
von: Lu, Hongbo, et al.
Veröffentlicht: (2026)
COUNTS: Benchmarking Object Detectors and Multimodal Large Language Models under Distribution Shifts
von: Li, Jiansheng, et al.
Veröffentlicht: (2025)
von: Li, Jiansheng, et al.
Veröffentlicht: (2025)
DER-GCN: Dialogue and Event Relation-Aware Graph Convolutional Neural Network for Multimodal Dialogue Emotion Recognition
von: Ai, Wei, et al.
Veröffentlicht: (2023)
von: Ai, Wei, et al.
Veröffentlicht: (2023)
FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models
von: Li, Yixuan, et al.
Veröffentlicht: (2024)
von: Li, Yixuan, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts
von: Liang, Jian, et al.
Veröffentlicht: (2023)
von: Liang, Jian, et al.
Veröffentlicht: (2023)
VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism
von: Zhang, Congzhi, et al.
Veröffentlicht: (2025)
von: Zhang, Congzhi, et al.
Veröffentlicht: (2025)
DIVER: Dynamic Iterative Visual Evidence Reasoning for Multimodal Fake News Detection
von: Zhou, Weilin, et al.
Veröffentlicht: (2026)
von: Zhou, Weilin, et al.
Veröffentlicht: (2026)
Spot the Fake: Large Multimodal Model-Based Synthetic Image Detection with Artifact Explanation
von: Wen, Siwei, et al.
Veröffentlicht: (2025)
von: Wen, Siwei, et al.
Veröffentlicht: (2025)
Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms
von: Cao, Zhixiang, et al.
Veröffentlicht: (2026)
von: Cao, Zhixiang, et al.
Veröffentlicht: (2026)
Layout-Guided Controllable Pathology Image Generation with In-Context Diffusion Transformers
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
Multimodal Fake News Detection: MFND Dataset and Shallow-Deep Multitask Learning
von: Zhu, Ye, et al.
Veröffentlicht: (2025)
von: Zhu, Ye, et al.
Veröffentlicht: (2025)
Triple Path Enhanced Neural Architecture Search for Multimodal Fake News Detection
von: Xu, Bo, et al.
Veröffentlicht: (2025)
von: Xu, Bo, et al.
Veröffentlicht: (2025)
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
A Survey on Video Temporal Grounding with Multimodal Large Language Model
von: Wu, Jianlong, et al.
Veröffentlicht: (2025)
von: Wu, Jianlong, et al.
Veröffentlicht: (2025)
Masked Graph Learning with Recurrent Alignment for Multimodal Emotion Recognition in Conversation
von: Meng, Tao, et al.
Veröffentlicht: (2024)
von: Meng, Tao, et al.
Veröffentlicht: (2024)
Revisiting Multimodal Emotion Recognition in Conversation from the Perspective of Graph Spectrum
von: Meng, Tao, et al.
Veröffentlicht: (2024)
von: Meng, Tao, et al.
Veröffentlicht: (2024)
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
von: Han, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Han, Xiaofeng, et al.
Veröffentlicht: (2025)
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation
von: Yu, Hong-Tao, et al.
Veröffentlicht: (2025)
von: Yu, Hong-Tao, et al.
Veröffentlicht: (2025)
Evaluating Attribute Comprehension in Large Vision-Language Models
von: Zhang, Haiwen, et al.
Veröffentlicht: (2024)
von: Zhang, Haiwen, et al.
Veröffentlicht: (2024)
CoVLM: Leveraging Consensus from Vision-Language Models for Semi-supervised Multi-modal Fake News Detection
von: Devank, et al.
Veröffentlicht: (2024)
von: Devank, et al.
Veröffentlicht: (2024)
GSDNet: Revisiting Incomplete Multimodal-Diffusion from Graph Spectrum Perspective for Conversation Emotion Recognition
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
TruthLens:A Training-Free Paradigm for DeepFake Detection
von: Chakraborty, Ritabrata, et al.
Veröffentlicht: (2025)
von: Chakraborty, Ritabrata, et al.
Veröffentlicht: (2025)
Efficient Multimodal Large Language Models: A Survey
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection
von: Jiang, Changjiang, et al.
Veröffentlicht: (2026)
von: Jiang, Changjiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CILF-CIAE: CLIP-driven Image-Language Fusion for Correcting Inverse Age Estimation
von: Shou, Yuntao, et al.
Veröffentlicht: (2023) -
LRA-GNN: Latent Relation-Aware Graph Neural Network with Initial and Dynamic Residual for Facial Age Estimation
von: Zhang, Yiping, et al.
Veröffentlicht: (2025) -
GroupFace: Imbalanced Age Estimation Based on Multi-hop Attention Graph Convolutional Network and Group-aware Margin Optimization
von: Zhang, Yiping, et al.
Veröffentlicht: (2024) -
Multimodal Large Language Models Meet Multimodal Emotion Recognition and Reasoning: A Survey
von: Shou, Yuntao, et al.
Veröffentlicht: (2025) -
A Multi-view Mask Contrastive Learning Graph Convolutional Neural Network for Age Estimation
von: Zhang, Yiping, et al.
Veröffentlicht: (2024)