Image-Text Out-Of-Context Detection Using Synthetic Multimodal Misinformation
Fuente:
arXiv
Salvato in:
| Autori principali: | Shalabi, Fatma, Nguyen, Huy H., Felouat, Hichem, Chang, Ching-Chun, Echizen, Isao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Leveraging Chat-Based Large Vision Language Models for Multimodal Out-Of-Context Detection
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
GFT-GCN: Privacy-Preserving 3D Face Mesh Recognition with Spectral Diffusion
di: Felouat, Hichem, et al.
Pubblicazione: (2025)
di: Felouat, Hichem, et al.
Pubblicazione: (2025)
Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
Enhancing Robustness of LLM-Synthetic Text Detectors for Academic Writing: A Comprehensive Analysis
di: Dou, Zhicheng, et al.
Pubblicazione: (2024)
di: Dou, Zhicheng, et al.
Pubblicazione: (2024)
Exploring Self-Supervised Vision Transformers for Deepfake Detection: A Comparative Analysis
di: Nguyen, Huy H., et al.
Pubblicazione: (2024)
di: Nguyen, Huy H., et al.
Pubblicazione: (2024)
Zero-Shot Warning Generation for Misinformative Multimodal Content
di: Delvecchio, Giovanni Pio, et al.
Pubblicazione: (2025)
di: Delvecchio, Giovanni Pio, et al.
Pubblicazione: (2025)
Fine-Tuning Text-To-Image Diffusion Models for Class-Wise Spurious Feature Generation
di: MaungMaung, AprilPyone, et al.
Pubblicazione: (2024)
di: MaungMaung, AprilPyone, et al.
Pubblicazione: (2024)
EvoGuard: An Extensible Agentic RL-based Framework for Practical and Evolving AI-Generated Image Detection
di: Zhu, Chenyang, et al.
Pubblicazione: (2026)
di: Zhu, Chenyang, et al.
Pubblicazione: (2026)
AnimeDL-2M: Million-Scale AI-Generated Anime Image Detection and Localization in Diffusion Era
di: Zhu, Chenyang, et al.
Pubblicazione: (2025)
di: Zhu, Chenyang, et al.
Pubblicazione: (2025)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
di: Yan, Zehong, et al.
Pubblicazione: (2025)
di: Yan, Zehong, et al.
Pubblicazione: (2025)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
di: Qi, Peng, et al.
Pubblicazione: (2024)
di: Qi, Peng, et al.
Pubblicazione: (2024)
Defending Against Physical Adversarial Patch Attacks on Infrared Human Detection
di: Strack, Lukas, et al.
Pubblicazione: (2023)
di: Strack, Lukas, et al.
Pubblicazione: (2023)
Exploring Active Data Selection Strategies for Continuous Training in Deepfake Detection
di: Furuhashi, Yoshihiko, et al.
Pubblicazione: (2025)
di: Furuhashi, Yoshihiko, et al.
Pubblicazione: (2025)
LookupForensics: A Large-Scale Multi-Task Dataset for Multi-Phase Image-Based Fact Verification
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
Measuring Human Involvement in AI-Generated Text: A Case Study on Academic Writing
di: Guo, Yuchen, et al.
Pubblicazione: (2025)
di: Guo, Yuchen, et al.
Pubblicazione: (2025)
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
di: Wang, Bing, et al.
Pubblicazione: (2024)
di: Wang, Bing, et al.
Pubblicazione: (2024)
Uncolorable Examples: Preventing Unauthorized AI Colorization via Perception-Aware Chroma-Restrictive Perturbation
di: Nii, Yuki, et al.
Pubblicazione: (2025)
di: Nii, Yuki, et al.
Pubblicazione: (2025)
Agentic Copyright Watermarking against Adversarial Evidence Forgery with Purification-Agnostic Curriculum Proxy Learning
di: Bao, Erjin, et al.
Pubblicazione: (2024)
di: Bao, Erjin, et al.
Pubblicazione: (2024)
A Controllable 3D Deepfake Generation Framework with Gaussian Splatting
di: Liu, Wending, et al.
Pubblicazione: (2025)
di: Liu, Wending, et al.
Pubblicazione: (2025)
Probabilistic Concept Graph Reasoning for Multimodal Misinformation Detection
di: Yang, Ruichao, et al.
Pubblicazione: (2026)
di: Yang, Ruichao, et al.
Pubblicazione: (2026)
Insight-A: Attribution-aware for Multimodal Misinformation Detection
di: Wu, Junjie, et al.
Pubblicazione: (2025)
di: Wu, Junjie, et al.
Pubblicazione: (2025)
Cross-Attention Watermarking of Large Language Models
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality Perspective
di: Wang, Bing, et al.
Pubblicazione: (2025)
di: Wang, Bing, et al.
Pubblicazione: (2025)
Quality Text, Robust Vision: The Role of Language in Enhancing Visual Robustness of Vision-Language Models
di: Waseda, Futa, et al.
Pubblicazione: (2025)
di: Waseda, Futa, et al.
Pubblicazione: (2025)
TI-JEPA: An Innovative Energy-based Joint Embedding Strategy for Text-Image Multimodal Systems
di: Vo, Khang H. N., et al.
Pubblicazione: (2025)
di: Vo, Khang H. N., et al.
Pubblicazione: (2025)
MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs
di: Liu, Xuannan, et al.
Pubblicazione: (2024)
di: Liu, Xuannan, et al.
Pubblicazione: (2024)
Physics-Based Adversarial Attack on Near-Infrared Human Detector for Nighttime Surveillance Camera Systems
di: Niu, Muyao, et al.
Pubblicazione: (2024)
di: Niu, Muyao, et al.
Pubblicazione: (2024)
Beyond Standard Benchmarks: A Systematic Audit of Vision-Language Model's Robustness to Natural Semantic Variation Across Diverse Tasks
di: Chengyu, Jia, et al.
Pubblicazione: (2026)
di: Chengyu, Jia, et al.
Pubblicazione: (2026)
Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation
di: Yang, Yue, et al.
Pubblicazione: (2025)
di: Yang, Yue, et al.
Pubblicazione: (2025)
Linguistically Informed Multimodal Fusion for Vietnamese Scene-Text Image Captioning: Dataset, Graph Framework, and Phonological Attention
di: Nguyen, Nhi Ngoc-Yen, et al.
Pubblicazione: (2026)
di: Nguyen, Nhi Ngoc-Yen, et al.
Pubblicazione: (2026)
GreedyPixel: Fine-Grained Black-Box Adversarial Attack Via Greedy Algorithm
di: Wang, Hanrui, et al.
Pubblicazione: (2025)
di: Wang, Hanrui, et al.
Pubblicazione: (2025)
Surface Normal Estimation with Transformers
di: Hu, Barry Shichen, et al.
Pubblicazione: (2024)
di: Hu, Barry Shichen, et al.
Pubblicazione: (2024)
Common Objects Out of Context (COOCo): Investigating Multimodal Context and Semantic Scene Violations in Referential Communication
di: Merlo, Filippo, et al.
Pubblicazione: (2025)
di: Merlo, Filippo, et al.
Pubblicazione: (2025)
Online Misinformation Detection in Live Streaming Videos
di: Cao, Rui
Pubblicazione: (2025)
di: Cao, Rui
Pubblicazione: (2025)
Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2023)
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2023)
Mitigating Backdoor Attacks using Activation-Guided Model Editing
di: Hsieh, Felix, et al.
Pubblicazione: (2024)
di: Hsieh, Felix, et al.
Pubblicazione: (2024)
Beyond Image-Text Matching: Verb Understanding in Multimodal Transformers Using Guided Masking
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
Support or Refute: Analyzing the Stance of Evidence to Detect Out-of-Context Mis- and Disinformation
di: Yuan, Xin, et al.
Pubblicazione: (2023)
di: Yuan, Xin, et al.
Pubblicazione: (2023)
SK-VQA: Synthetic Knowledge Generation at Scale for Training Context-Augmented Multimodal LLMs
di: Su, Xin, et al.
Pubblicazione: (2024)
di: Su, Xin, et al.
Pubblicazione: (2024)
Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think
di: Chen, Liang, et al.
Pubblicazione: (2025)
di: Chen, Liang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Leveraging Chat-Based Large Vision Language Models for Multimodal Out-Of-Context Detection
di: Shalabi, Fatma, et al.
Pubblicazione: (2024) -
GFT-GCN: Privacy-Preserving 3D Face Mesh Recognition with Spectral Diffusion
di: Felouat, Hichem, et al.
Pubblicazione: (2025) -
Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025) -
Enhancing Robustness of LLM-Synthetic Text Detectors for Academic Writing: A Comprehensive Analysis
di: Dou, Zhicheng, et al.
Pubblicazione: (2024) -
Exploring Self-Supervised Vision Transformers for Deepfake Detection: A Comparative Analysis
di: Nguyen, Huy H., et al.
Pubblicazione: (2024)