Multi-modal Misinformation Detection: Approaches, Challenges and Opportunities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdali, Sara, shaham, Sina, Krishnamachari, Bhaskar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
von: Qi, Peng, et al.
Veröffentlicht: (2024)
von: Qi, Peng, et al.
Veröffentlicht: (2024)
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
von: Wu, Bo, et al.
Veröffentlicht: (2024)
von: Wu, Bo, et al.
Veröffentlicht: (2024)
Detecting Misinformation in Multimedia Content through Cross-Modal Entity Consistency: A Dual Learning Approach
von: Fu, Zhe, et al.
Veröffentlicht: (2024)
von: Fu, Zhe, et al.
Veröffentlicht: (2024)
Personality Analysis from Online Short Video Platforms with Multi-domain Adaptation
von: An, Sixu, et al.
Veröffentlicht: (2024)
von: An, Sixu, et al.
Veröffentlicht: (2024)
NativE: Multi-modal Knowledge Graph Completion in the Wild
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
CrisisViT: A Robust Vision Transformer for Crisis Image Classification
von: Long, Zijun, et al.
Veröffentlicht: (2024)
von: Long, Zijun, et al.
Veröffentlicht: (2024)
Embedding an Ethical Mind: Aligning Text-to-Image Synthesis via Lightweight Value Optimization
von: Wang, Xingqi, et al.
Veröffentlicht: (2024)
von: Wang, Xingqi, et al.
Veröffentlicht: (2024)
Unmasking Illusions: Understanding Human Perception of Audiovisual Deepfakes
von: Hashmi, Ammarah, et al.
Veröffentlicht: (2024)
von: Hashmi, Ammarah, et al.
Veröffentlicht: (2024)
Counteracting temporal attacks in Video Copy Detection
von: Fojcik, Katarzyna, et al.
Veröffentlicht: (2025)
von: Fojcik, Katarzyna, et al.
Veröffentlicht: (2025)
MultiWay-Adapater: Adapting large-scale multi-modal models for scalable image-text retrieval
von: Long, Zijun, et al.
Veröffentlicht: (2023)
von: Long, Zijun, et al.
Veröffentlicht: (2023)
Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Explore the Limits of Omni-modal Pretraining at Scale
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2024)
How Do Images Align and Complement LiDAR? Towards a Harmonized Multi-modal 3D Panoptic Segmentation
von: Pan, Yining, et al.
Veröffentlicht: (2025)
von: Pan, Yining, et al.
Veröffentlicht: (2025)
LLM-based Fusion of Multi-modal Features for Commercial Memorability Prediction
von: Pramov, Aleksandar
Veröffentlicht: (2025)
von: Pramov, Aleksandar
Veröffentlicht: (2025)
Visual Authority and the Rhetoric of Health Misinformation: A Multimodal Analysis of Social Media Videos
von: Zarei, Mohammad Reza, et al.
Veröffentlicht: (2025)
von: Zarei, Mohammad Reza, et al.
Veröffentlicht: (2025)
Can LLMs Create Legally Relevant Summaries and Analyses of Videos?
von: Hoeben-Kuil, Lyra, et al.
Veröffentlicht: (2025)
von: Hoeben-Kuil, Lyra, et al.
Veröffentlicht: (2025)
KI-Bilder und die Widerständigkeit der Medienkonvergenz: Von primärer zu sekundärer Intermedialität?
von: Wilde, Lukas R. A.
Veröffentlicht: (2024)
von: Wilde, Lukas R. A.
Veröffentlicht: (2024)
Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study
von: Bačić, Boris, et al.
Veröffentlicht: (2024)
von: Bačić, Boris, et al.
Veröffentlicht: (2024)
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
AI-based System for Transforming text and sound to Educational Videos
von: ElAlami, M. E., et al.
Veröffentlicht: (2026)
von: ElAlami, M. E., et al.
Veröffentlicht: (2026)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2023)
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2023)
Enhancing multimodal cooperation via sample-level modality valuation
von: Wei, Yake, et al.
Veröffentlicht: (2023)
von: Wei, Yake, et al.
Veröffentlicht: (2023)
Multi-modal Segment Assemblage Network for Ad Video Editing with Importance-Coherence Reward
von: Tang, Yolo Yunlong, et al.
Veröffentlicht: (2022)
von: Tang, Yolo Yunlong, et al.
Veröffentlicht: (2022)
PointCoT: A Multi-modal Benchmark for Explicit 3D Geometric Reasoning
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
A Rate-Distortion-Classification Approach for Lossy Image Compression
von: Zhang, Yuefeng
Veröffentlicht: (2024)
von: Zhang, Yuefeng
Veröffentlicht: (2024)
Art2Music: Generating Music for Art Images with Multi-modal Feeling Alignment
von: Hong, Jiaying, et al.
Veröffentlicht: (2025)
von: Hong, Jiaying, et al.
Veröffentlicht: (2025)
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
von: Bhaskar, Paramananda, et al.
Veröffentlicht: (2026)
von: Bhaskar, Paramananda, et al.
Veröffentlicht: (2026)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
Distilling Neuro-Symbolic Programs into 3D Multi-modal LLMs
von: Mo, Wentao, et al.
Veröffentlicht: (2026)
von: Mo, Wentao, et al.
Veröffentlicht: (2026)
TelcoAI: Advancing 3GPP Technical Specification Search through Agentic Multi-Modal Retrieval-Augmented Generation
von: Ghosh, Rahul, et al.
Veröffentlicht: (2025)
von: Ghosh, Rahul, et al.
Veröffentlicht: (2025)
Cross-modal Causal Intervention for Alzheimer's Disease Prediction
von: Jin, Yutao, et al.
Veröffentlicht: (2025)
von: Jin, Yutao, et al.
Veröffentlicht: (2025)
CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection
von: Li, Fanxiao, et al.
Veröffentlicht: (2025)
von: Li, Fanxiao, et al.
Veröffentlicht: (2025)
Image Complexity-Aware Adaptive Retrieval for Efficient Vision-Language Models
von: Williams-Lekuona, Mikel, et al.
Veröffentlicht: (2025)
von: Williams-Lekuona, Mikel, et al.
Veröffentlicht: (2025)
Infinite Video Understanding
von: Zhang, Dell, et al.
Veröffentlicht: (2025)
von: Zhang, Dell, et al.
Veröffentlicht: (2025)
PC$^2$: Pseudo-Classification Based Pseudo-Captioning for Noisy Correspondence Learning in Cross-Modal Retrieval
von: Duan, Yue, et al.
Veröffentlicht: (2024)
von: Duan, Yue, et al.
Veröffentlicht: (2024)
CalliffusionV2: Personalized Natural Calligraphy Generation with Flexible Multi-modal Control
von: Liao, Qisheng, et al.
Veröffentlicht: (2024)
von: Liao, Qisheng, et al.
Veröffentlicht: (2024)
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
von: Tian, Zeyue, et al.
Veröffentlicht: (2026)
von: Tian, Zeyue, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
von: Qi, Peng, et al.
Veröffentlicht: (2024) -
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
von: Wu, Bo, et al.
Veröffentlicht: (2024) -
Detecting Misinformation in Multimedia Content through Cross-Modal Entity Consistency: A Dual Learning Approach
von: Fu, Zhe, et al.
Veröffentlicht: (2024) -
Personality Analysis from Online Short Video Platforms with Multi-domain Adaptation
von: An, Sixu, et al.
Veröffentlicht: (2024) -
NativE: Multi-modal Knowledge Graph Completion in the Wild
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)