Tell me Habibi, is it Real or Fake?
Fuente:
arXiv
Saved in:
| Main Authors: | Kuckreja, Kartik, Gupta, Parul, Hamed, Injy, Solorio, Thamar, Khan, Muhammad Haris, Dhall, Abhinav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026)
by: Kuckreja, Kartik, et al.
Published: (2026)
AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations
by: Cai, Zhixi, et al.
Published: (2025)
by: Cai, Zhixi, et al.
Published: (2025)
Real, fake and synthetic faces -- does the coin have three sides?
by: Naeem, Shahzeb, et al.
Published: (2024)
by: Naeem, Shahzeb, et al.
Published: (2024)
Improving Personalized Image Generation through Social Context Feedback
by: Gupta, Parul, et al.
Published: (2025)
by: Gupta, Parul, et al.
Published: (2025)
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
by: Gupta, Parul, et al.
Published: (2025)
by: Gupta, Parul, et al.
Published: (2025)
Generation and Detection of Sign Language Deepfakes - A Linguistic and Visual Analysis
by: Naeem, Shahzeb, et al.
Published: (2024)
by: Naeem, Shahzeb, et al.
Published: (2024)
LayLens: Improving Deepfake Understanding through Simplified Explanations
by: Narang, Abhijeet, et al.
Published: (2025)
by: Narang, Abhijeet, et al.
Published: (2025)
Conditional Distribution Modelling for Few-Shot Image Synthesis with Diffusion Models
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
by: Jahangard, Simindokht, et al.
Published: (2025)
by: Jahangard, Simindokht, et al.
Published: (2025)
Question-Instructed Visual Descriptions for Zero-Shot Video Question Answering
by: Romero, David, et al.
Published: (2024)
by: Romero, David, et al.
Published: (2024)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
by: Silva, Sathira, et al.
Published: (2025)
by: Silva, Sathira, et al.
Published: (2025)
Unsupervised Deep Graph Matching Based on Cycle Consistency
by: Tourani, Siddharth, et al.
Published: (2023)
by: Tourani, Siddharth, et al.
Published: (2023)
CLIP Architecture for Abdominal CT Image-Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scaling
by: Shivika, et al.
Published: (2026)
by: Shivika, et al.
Published: (2026)
A Survey of Deep Learning for Group-level Emotion Recognition
by: Huang, Xiaohua, et al.
Published: (2024)
by: Huang, Xiaohua, et al.
Published: (2024)
DiffAugment: Diffusion based Long-Tailed Visual Relationship Recognition
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
SFANet: Spatial-Frequency Attention Network for Deepfake Detection
by: Ahire, Vrushank, et al.
Published: (2025)
by: Ahire, Vrushank, et al.
Published: (2025)
CapsFake: A Multimodal Capsule Network for Detecting Instruction-Guided Deepfakes
by: Nguyen, Tuan, et al.
Published: (2025)
by: Nguyen, Tuan, et al.
Published: (2025)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
by: Wei, Yibing, et al.
Published: (2024)
by: Wei, Yibing, et al.
Published: (2024)
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
by: Tavella, Federico, et al.
Published: (2025)
by: Tavella, Federico, et al.
Published: (2025)
CLOFAI: A Dataset of Real And Fake Image Classification Tasks for Continual Learning
by: Doherty, William, et al.
Published: (2025)
by: Doherty, William, et al.
Published: (2025)
Real-Time Object Detection in Occluded Environment with Background Cluttering Effects Using Deep Learning
by: Aamir, Syed Muhammad, et al.
Published: (2024)
by: Aamir, Syed Muhammad, et al.
Published: (2024)
TrueFake: A Real World Case Dataset of Last Generation Fake Images also Shared on Social Networks
by: Dell'Anna, Stefano, et al.
Published: (2025)
by: Dell'Anna, Stefano, et al.
Published: (2025)
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network
by: Ahire, Vrushank, et al.
Published: (2025)
by: Ahire, Vrushank, et al.
Published: (2025)
Fake-in-Facext: Towards Fine-Grained Explainable DeepFake Analysis
by: Qin, Lixiong, et al.
Published: (2025)
by: Qin, Lixiong, et al.
Published: (2025)
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
by: Khan, Ufaq, et al.
Published: (2026)
by: Khan, Ufaq, et al.
Published: (2026)
MedROV: Towards Real-Time Open-Vocabulary Detection Across Diverse Medical Imaging Modalities
by: Sheikh, Tooba Tehreem, et al.
Published: (2025)
by: Sheikh, Tooba Tehreem, et al.
Published: (2025)
Real Risks of Fake Data: Synthetic Data, Diversity-Washing and Consent Circumvention
by: Whitney, Cedric Deslandes, et al.
Published: (2024)
by: Whitney, Cedric Deslandes, et al.
Published: (2024)
Multi-modal Medical Image Fusion For Non-Small Cell Lung Cancer Classification
by: Hassan, Salma, et al.
Published: (2024)
by: Hassan, Salma, et al.
Published: (2024)
Zero-Shot and Supervised Bird Image Segmentation Using Foundation Models: A Dual-Pipeline Approach with Grounding DINO~1.5, YOLOv11, and SAM~2.1
by: Munagala, Abhinav
Published: (2026)
by: Munagala, Abhinav
Published: (2026)
Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments
by: Ali, Muhammad, et al.
Published: (2025)
by: Ali, Muhammad, et al.
Published: (2025)
Dynamic Weight Adjustment for Knowledge Distillation: Leveraging Vision Transformer for High-Accuracy Lung Cancer Detection and Real-Time Deployment
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
FakeParts: a New Family of AI-Generated DeepFakes
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
Adaptive Image Restoration for Video Surveillance: A Real-Time Approach
by: Amin, Muhammad Awais, et al.
Published: (2025)
by: Amin, Muhammad Awais, et al.
Published: (2025)
Do DeepFake Attribution Models Generalize?
by: Baxavanakis, Spiros, et al.
Published: (2025)
by: Baxavanakis, Spiros, et al.
Published: (2025)
See then Tell: Enhancing Key Information Extraction with Vision Grounding
by: Liu, Shuhang, et al.
Published: (2024)
by: Liu, Shuhang, et al.
Published: (2024)
GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing
by: Hasan, Maram, et al.
Published: (2026)
by: Hasan, Maram, et al.
Published: (2026)
Weakly Supervised Teacher-Student Framework with Progressive Pseudo-mask Refinement for Gland Segmentation
by: Khan, Hikmat, et al.
Published: (2026)
by: Khan, Hikmat, et al.
Published: (2026)
VelocityNet: Real-Time Crowd Anomaly Detection via Person-Specific Velocity Analysis
by: AlGhamdi, Fatima, et al.
Published: (2025)
by: AlGhamdi, Fatima, et al.
Published: (2025)
Similar Items
-
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026) -
AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations
by: Cai, Zhixi, et al.
Published: (2025) -
Real, fake and synthetic faces -- does the coin have three sides?
by: Naeem, Shahzeb, et al.
Published: (2024) -
Improving Personalized Image Generation through Social Context Feedback
by: Gupta, Parul, et al.
Published: (2025) -
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
by: Gupta, Parul, et al.
Published: (2025)