SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Qi, Peng, Yan, Zehong, Hsu, Wynne, Lee, Mong Li |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
por: Yan, Zehong, et al.
Publicado: (2025)
por: Yan, Zehong, et al.
Publicado: (2025)
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection
por: Yan, Zehong, et al.
Publicado: (2025)
por: Yan, Zehong, et al.
Publicado: (2025)
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
por: Wang, Bing, et al.
Publicado: (2024)
por: Wang, Bing, et al.
Publicado: (2024)
Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality Perspective
por: Wang, Bing, et al.
Publicado: (2025)
por: Wang, Bing, et al.
Publicado: (2025)
LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection
por: Wu, Lanhu, et al.
Publicado: (2025)
por: Wu, Lanhu, et al.
Publicado: (2025)
CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection
por: Li, Fanxiao, et al.
Publicado: (2025)
por: Li, Fanxiao, et al.
Publicado: (2025)
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
por: Cheng, Zebang, et al.
Publicado: (2024)
por: Cheng, Zebang, et al.
Publicado: (2024)
MaVEn: An Effective Multi-granularity Hybrid Visual Encoding Framework for Multimodal Large Language Model
por: Jiang, Chaoya, et al.
Publicado: (2024)
por: Jiang, Chaoya, et al.
Publicado: (2024)
FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models
por: Li, Yixuan, et al.
Publicado: (2024)
por: Li, Yixuan, et al.
Publicado: (2024)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
por: Bin, Yi, et al.
Publicado: (2024)
por: Bin, Yi, et al.
Publicado: (2024)
FakingRecipe: Detecting Fake News on Short Video Platforms from the Perspective of Creative Process
por: Bu, Yuyan, et al.
Publicado: (2024)
por: Bu, Yuyan, et al.
Publicado: (2024)
WordArt Designer API: User-Driven Artistic Typography Synthesis with Large Language Models on ModelScope
por: He, Jun-Yan, et al.
Publicado: (2024)
por: He, Jun-Yan, et al.
Publicado: (2024)
Latent Reconstruction from Generated Data for Multimodal Misinformation Detection
por: Papadopoulos, Stefanos-Iordanis, et al.
Publicado: (2025)
por: Papadopoulos, Stefanos-Iordanis, et al.
Publicado: (2025)
What Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models
por: Baraldi, Lorenzo, et al.
Publicado: (2025)
por: Baraldi, Lorenzo, et al.
Publicado: (2025)
MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
por: Lin, Xiao, et al.
Publicado: (2025)
por: Lin, Xiao, et al.
Publicado: (2025)
Knowledge Acquisition Disentanglement for Knowledge-based Visual Question Answering with Large Language Models
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs
por: Chen, Qian, et al.
Publicado: (2026)
por: Chen, Qian, et al.
Publicado: (2026)
The Revolution of Multimodal Large Language Models: A Survey
por: Caffagni, Davide, et al.
Publicado: (2024)
por: Caffagni, Davide, et al.
Publicado: (2024)
Anim-Director: A Large Multimodal Model Powered Agent for Controllable Animation Video Generation
por: Li, Yunxin, et al.
Publicado: (2024)
por: Li, Yunxin, et al.
Publicado: (2024)
Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models
por: Wu, Jiaying, et al.
Publicado: (2025)
por: Wu, Jiaying, et al.
Publicado: (2025)
ChronusOmni: Improving Time Awareness of Omni Large Language Models
por: Chen, Yijing, et al.
Publicado: (2025)
por: Chen, Yijing, et al.
Publicado: (2025)
Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis
por: Bucciarelli, Davide, et al.
Publicado: (2024)
por: Bucciarelli, Davide, et al.
Publicado: (2024)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
por: Fazli, Mehrdad, et al.
Publicado: (2025)
por: Fazli, Mehrdad, et al.
Publicado: (2025)
Dynamic Self-adaptive Multiscale Distillation from Pre-trained Multimodal Large Model for Efficient Cross-modal Representation Learning
por: Liang, Zhengyang, et al.
Publicado: (2024)
por: Liang, Zhengyang, et al.
Publicado: (2024)
"Humor, Art, or Misinformation?": A Multimodal Dataset for Intent-Aware Synthetic Image Detection
por: Skoularikis, Anastasios, et al.
Publicado: (2025)
por: Skoularikis, Anastasios, et al.
Publicado: (2025)
How to Bridge the Gap between Modalities: Survey on Multimodal Large Language Model
por: Song, Shezheng, et al.
Publicado: (2023)
por: Song, Shezheng, et al.
Publicado: (2023)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
por: Wang, Xiao, et al.
Publicado: (2025)
por: Wang, Xiao, et al.
Publicado: (2025)
Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models
por: Ye, Weihao, et al.
Publicado: (2024)
por: Ye, Weihao, et al.
Publicado: (2024)
Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding
por: Luo, Chuwei, et al.
Publicado: (2022)
por: Luo, Chuwei, et al.
Publicado: (2022)
RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models
por: Mattioli, Gabriele, et al.
Publicado: (2026)
por: Mattioli, Gabriele, et al.
Publicado: (2026)
Seeing Beyond Words: Self-Supervised Visual Learning for Multimodal Large Language Models
por: Caffagni, Davide, et al.
Publicado: (2025)
por: Caffagni, Davide, et al.
Publicado: (2025)
A Survey of Multimodal Large Language Model from A Data-centric Perspective
por: Bai, Tianyi, et al.
Publicado: (2024)
por: Bai, Tianyi, et al.
Publicado: (2024)
Visual Authority and the Rhetoric of Health Misinformation: A Multimodal Analysis of Social Media Videos
por: Zarei, Mohammad Reza, et al.
Publicado: (2025)
por: Zarei, Mohammad Reza, et al.
Publicado: (2025)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
por: Fu, Yuhan, et al.
Publicado: (2024)
por: Fu, Yuhan, et al.
Publicado: (2024)
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
por: Wu, Qiong, et al.
Publicado: (2024)
por: Wu, Qiong, et al.
Publicado: (2024)
Joint Modeling of Big Five and HEXACO for Multimodal Apparent Personality-trait Recognition
por: Masumura, Ryo, et al.
Publicado: (2025)
por: Masumura, Ryo, et al.
Publicado: (2025)
Towards Robust and Realible Multimodal Misinformation Recognition with Incomplete Modality
por: Zhou, Hengyang, et al.
Publicado: (2025)
por: Zhou, Hengyang, et al.
Publicado: (2025)
Watch Out Your Album! On the Inadvertent Privacy Memorization in Multi-Modal Large Language Models
por: Ju, Tianjie, et al.
Publicado: (2025)
por: Ju, Tianjie, et al.
Publicado: (2025)
Incorporating Visual Experts to Resolve the Information Loss in Multimodal Large Language Models
por: He, Xin, et al.
Publicado: (2024)
por: He, Xin, et al.
Publicado: (2024)
Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring
por: Zhang, Dongxu, et al.
Publicado: (2026)
por: Zhang, Dongxu, et al.
Publicado: (2026)
Ejemplares similares
-
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
por: Yan, Zehong, et al.
Publicado: (2025) -
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection
por: Yan, Zehong, et al.
Publicado: (2025) -
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
por: Wang, Bing, et al.
Publicado: (2024) -
Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality Perspective
por: Wang, Bing, et al.
Publicado: (2025) -
LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection
por: Wu, Lanhu, et al.
Publicado: (2025)