Common Sense Reasoning for Deepfake Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yue, Colman, Ben, Guo, Xiao, Shahriyari, Ali, Bharaj, Gaurav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Common-Sense Bias Modeling for Classification Tasks
von: Zhang, Miao, et al.
Veröffentlicht: (2024)
von: Zhang, Miao, et al.
Veröffentlicht: (2024)
AVFF: Audio-Visual Feature Fusion for Video Deepfake Detection
von: Oorloff, Trevine, et al.
Veröffentlicht: (2024)
von: Oorloff, Trevine, et al.
Veröffentlicht: (2024)
X-Edit: Detecting and Localizing Edits in Images Altered by Text-Guided Diffusion Models
von: Bazyleva, Valentina, et al.
Veröffentlicht: (2025)
von: Bazyleva, Valentina, et al.
Veröffentlicht: (2025)
FaceLift: Semi-supervised 3D Facial Landmark Localization
von: Ferman, David, et al.
Veröffentlicht: (2024)
von: Ferman, David, et al.
Veröffentlicht: (2024)
Through the Looking Glass: Common Sense Consistency Evaluation of Weird Images
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
Probabilistic Concept Graph Reasoning for Multimodal Misinformation Detection
von: Yang, Ruichao, et al.
Veröffentlicht: (2026)
von: Yang, Ruichao, et al.
Veröffentlicht: (2026)
Structured and Abstractive Reasoning on Multi-modal Relational Knowledge Images
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
Customizing Visual-Language Foundation Models for Multi-modal Anomaly Detection and Reasoning
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
AIP: Subverting Retrieval-Augmented Generation via Adversarial Instructional Prompt
von: Chaturvedi, Saket S., et al.
Veröffentlicht: (2025)
von: Chaturvedi, Saket S., et al.
Veröffentlicht: (2025)
3ViewSense: Spatial and Mental Perspective Reasoning from Orthographic Views in Vision-Language Models
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2026)
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2026)
MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale
von: Guo, Jarvis, et al.
Veröffentlicht: (2024)
von: Guo, Jarvis, et al.
Veröffentlicht: (2024)
ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction
von: Guo, Zichun, et al.
Veröffentlicht: (2026)
von: Guo, Zichun, et al.
Veröffentlicht: (2026)
PhonemeFake: Redefining Deepfake Realism with Language-Driven Segmental Manipulation and Adaptive Bilevel Detection
von: Baser, Oguzhan, et al.
Veröffentlicht: (2025)
von: Baser, Oguzhan, et al.
Veröffentlicht: (2025)
Multimodal Event Detection: Current Approaches and Defining the New Playground through LLMs and VLMs
von: Dey, Abhishek, et al.
Veröffentlicht: (2025)
von: Dey, Abhishek, et al.
Veröffentlicht: (2025)
Scaling Agentic Reinforcement Learning for Tool-Integrated Reasoning in VLMs
von: Lu, Meng, et al.
Veröffentlicht: (2025)
von: Lu, Meng, et al.
Veröffentlicht: (2025)
GRAM: Global Reasoning for Multi-Page VQA
von: Blau, Tsachi, et al.
Veröffentlicht: (2024)
von: Blau, Tsachi, et al.
Veröffentlicht: (2024)
MMGR: Multi-Modal Generative Reasoning
von: Cai, Zefan, et al.
Veröffentlicht: (2025)
von: Cai, Zefan, et al.
Veröffentlicht: (2025)
A Hitchhikers Guide to Fine-Grained Face Forgery Detection Using Common Sense Reasoning
von: Foteinopoulou, Niki Maria, et al.
Veröffentlicht: (2024)
von: Foteinopoulou, Niki Maria, et al.
Veröffentlicht: (2024)
Scaling Laws for Deepfake Detection
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
Play to Generalize: Learning to Reason Through Game Play
von: Xie, Yunfei, et al.
Veröffentlicht: (2025)
von: Xie, Yunfei, et al.
Veröffentlicht: (2025)
IV-Bench: A Benchmark for Image-Grounded Video Perception and Reasoning in Multimodal LLMs
von: Ma, David, et al.
Veröffentlicht: (2025)
von: Ma, David, et al.
Veröffentlicht: (2025)
LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning
von: Dai, Yifan, et al.
Veröffentlicht: (2026)
von: Dai, Yifan, et al.
Veröffentlicht: (2026)
VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection
von: Li, Xinghan, et al.
Veröffentlicht: (2026)
von: Li, Xinghan, et al.
Veröffentlicht: (2026)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
von: Chang, Yue, et al.
Veröffentlicht: (2024)
von: Chang, Yue, et al.
Veröffentlicht: (2024)
MSR-Align: Policy-Grounded Multimodal Alignment for Safety-Aware Reasoning in Vision-Language Models
von: Xia, Yinan, et al.
Veröffentlicht: (2025)
von: Xia, Yinan, et al.
Veröffentlicht: (2025)
Embodied-Reasoner: Synergizing Visual Search, Reasoning, and Action for Embodied Interactive Tasks
von: Zhang, Wenqi, et al.
Veröffentlicht: (2025)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2025)
Visual Reasoning at Urban Intersections: FineTuning GPT-4o for Traffic Conflict Detection
von: Masri, Sari, et al.
Veröffentlicht: (2025)
von: Masri, Sari, et al.
Veröffentlicht: (2025)
Training-Free Multimodal Deepfake Detection via Graph Reasoning
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
ClimateViz: A Benchmark for Statistical Reasoning and Fact Verification on Scientific Charts
von: Su, Ruiran, et al.
Veröffentlicht: (2025)
von: Su, Ruiran, et al.
Veröffentlicht: (2025)
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models
von: Li, Yunxin, et al.
Veröffentlicht: (2025)
von: Li, Yunxin, et al.
Veröffentlicht: (2025)
Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning
von: Wei, Yana, et al.
Veröffentlicht: (2025)
von: Wei, Yana, et al.
Veröffentlicht: (2025)
RSCC: A Large-Scale Remote Sensing Change Caption Dataset for Disaster Events
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2025)
Deepfake Detection via Knowledge Injection
von: Li, Tonghui, et al.
Veröffentlicht: (2025)
von: Li, Tonghui, et al.
Veröffentlicht: (2025)
MiRAGeNews: Multimodal Realistic AI-Generated News Detection
von: Huang, Runsheng, et al.
Veröffentlicht: (2024)
von: Huang, Runsheng, et al.
Veröffentlicht: (2024)
Head-wise Modality Specialization within MLLMs for Robust Fake News Detection under Missing Modality
von: Qian, Kai, et al.
Veröffentlicht: (2026)
von: Qian, Kai, et al.
Veröffentlicht: (2026)
Cerberus: Real-Time Video Anomaly Detection via Cascaded Vision-Language Models
von: Zheng, Yue, et al.
Veröffentlicht: (2025)
von: Zheng, Yue, et al.
Veröffentlicht: (2025)
Actial: Activate Spatial Reasoning Ability of Multimodal Large Language Models
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Common-Sense Bias Modeling for Classification Tasks
von: Zhang, Miao, et al.
Veröffentlicht: (2024) -
AVFF: Audio-Visual Feature Fusion for Video Deepfake Detection
von: Oorloff, Trevine, et al.
Veröffentlicht: (2024) -
X-Edit: Detecting and Localizing Edits in Images Altered by Text-Guided Diffusion Models
von: Bazyleva, Valentina, et al.
Veröffentlicht: (2025) -
FaceLift: Semi-supervised 3D Facial Landmark Localization
von: Ferman, David, et al.
Veröffentlicht: (2024) -
Through the Looking Glass: Common Sense Consistency Evaluation of Weird Images
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)