Exploring Self-Supervised Vision Transformers for Deepfake Detection: A Comparative Analysis
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Huy H., Yamagishi, Junichi, Echizen, Isao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring Active Data Selection Strategies for Continuous Training in Deepfake Detection
di: Furuhashi, Yoshihiko, et al.
Pubblicazione: (2025)
di: Furuhashi, Yoshihiko, et al.
Pubblicazione: (2025)
A Controllable 3D Deepfake Generation Framework with Gaussian Splatting
di: Liu, Wending, et al.
Pubblicazione: (2025)
di: Liu, Wending, et al.
Pubblicazione: (2025)
Leveraging Chat-Based Large Vision Language Models for Multimodal Out-Of-Context Detection
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
Defending Against Physical Adversarial Patch Attacks on Infrared Human Detection
di: Strack, Lukas, et al.
Pubblicazione: (2023)
di: Strack, Lukas, et al.
Pubblicazione: (2023)
Beyond Standard Benchmarks: A Systematic Audit of Vision-Language Model's Robustness to Natural Semantic Variation Across Diverse Tasks
di: Chengyu, Jia, et al.
Pubblicazione: (2026)
di: Chengyu, Jia, et al.
Pubblicazione: (2026)
Surface Normal Estimation with Transformers
di: Hu, Barry Shichen, et al.
Pubblicazione: (2024)
di: Hu, Barry Shichen, et al.
Pubblicazione: (2024)
Fine-Tuning Text-To-Image Diffusion Models for Class-Wise Spurious Feature Generation
di: MaungMaung, AprilPyone, et al.
Pubblicazione: (2024)
di: MaungMaung, AprilPyone, et al.
Pubblicazione: (2024)
Quality Text, Robust Vision: The Role of Language in Enhancing Visual Robustness of Vision-Language Models
di: Waseda, Futa, et al.
Pubblicazione: (2025)
di: Waseda, Futa, et al.
Pubblicazione: (2025)
Image-Text Out-Of-Context Detection Using Synthetic Multimodal Misinformation
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
di: Shalabi, Fatma, et al.
Pubblicazione: (2024)
Physics-Based Adversarial Attack on Near-Infrared Human Detector for Nighttime Surveillance Camera Systems
di: Niu, Muyao, et al.
Pubblicazione: (2024)
di: Niu, Muyao, et al.
Pubblicazione: (2024)
LookupForensics: A Large-Scale Multi-Task Dataset for Multi-Phase Image-Based Fact Verification
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
Mitigating Backdoor Attacks using Activation-Guided Model Editing
di: Hsieh, Felix, et al.
Pubblicazione: (2024)
di: Hsieh, Felix, et al.
Pubblicazione: (2024)
GFT-GCN: Privacy-Preserving 3D Face Mesh Recognition with Spectral Diffusion
di: Felouat, Hichem, et al.
Pubblicazione: (2025)
di: Felouat, Hichem, et al.
Pubblicazione: (2025)
SAVe: Self-Supervised Audio-visual Deepfake Detection Exploiting Visual Artifacts and Audio-visual Misalignment
di: Shahzad, Sahibzada Adil, et al.
Pubblicazione: (2026)
di: Shahzad, Sahibzada Adil, et al.
Pubblicazione: (2026)
A Timely Survey on Vision Transformer for Deepfake Detection
di: Wang, Zhikan, et al.
Pubblicazione: (2024)
di: Wang, Zhikan, et al.
Pubblicazione: (2024)
Cyber Vaccine for Deepfake Immunity
di: Chang, Ching-Chun, et al.
Pubblicazione: (2023)
di: Chang, Ching-Chun, et al.
Pubblicazione: (2023)
EditSleuth: A Dataset of Grounded Reasoning Chains for Image-Edit Forensics
di: Nguyen, Van-Loc, et al.
Pubblicazione: (2026)
di: Nguyen, Van-Loc, et al.
Pubblicazione: (2026)
EvoGuard: An Extensible Agentic RL-based Framework for Practical and Evolving AI-Generated Image Detection
di: Zhu, Chenyang, et al.
Pubblicazione: (2026)
di: Zhu, Chenyang, et al.
Pubblicazione: (2026)
Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
AnimeDL-2M: Million-Scale AI-Generated Anime Image Detection and Localization in Diffusion Era
di: Zhu, Chenyang, et al.
Pubblicazione: (2025)
di: Zhu, Chenyang, et al.
Pubblicazione: (2025)
Multimodal Adversarial Defense for Vision-Language Models by Leveraging One-To-Many Relationships
di: Waseda, Futa, et al.
Pubblicazione: (2024)
di: Waseda, Futa, et al.
Pubblicazione: (2024)
CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection
di: Khan, Sohail Ahmed, et al.
Pubblicazione: (2024)
di: Khan, Sohail Ahmed, et al.
Pubblicazione: (2024)
VisionGuard: Synergistic Framework for Helmet Violation Detection
di: Nguyen, Lam-Huy, et al.
Pubblicazione: (2025)
di: Nguyen, Lam-Huy, et al.
Pubblicazione: (2025)
FakeFormer: Efficient Vulnerability-Driven Transformers for Generalisable Deepfake Detection
di: Nguyen, Dat, et al.
Pubblicazione: (2024)
di: Nguyen, Dat, et al.
Pubblicazione: (2024)
DifAttack++: Query-Efficient Black-Box Adversarial Attack via Hierarchical Disentangled Feature Space in Cross-Domain
di: Liu, Jun, et al.
Pubblicazione: (2024)
di: Liu, Jun, et al.
Pubblicazione: (2024)
SDPA++: A General Framework for Self-Supervised Denoising with Patch Aggregation
di: Nguyen, Huy Minh Nhat, et al.
Pubblicazione: (2025)
di: Nguyen, Huy Minh Nhat, et al.
Pubblicazione: (2025)
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
di: Datta, Soumyya Kanti, et al.
Pubblicazione: (2025)
di: Datta, Soumyya Kanti, et al.
Pubblicazione: (2025)
GUNNEL: Guided Mixup Augmentation and Multi-Model Fusion for Aquatic Animal Segmentation
di: Le, Minh-Quan, et al.
Pubblicazione: (2021)
di: Le, Minh-Quan, et al.
Pubblicazione: (2021)
Exploring the Effect of Dataset Diversity in Self-Supervised Learning for Surgical Computer Vision
di: Jaspers, Tim J. M., et al.
Pubblicazione: (2024)
di: Jaspers, Tim J. M., et al.
Pubblicazione: (2024)
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer
di: Deressa, Deressa Wodajo, et al.
Pubblicazione: (2023)
di: Deressa, Deressa Wodajo, et al.
Pubblicazione: (2023)
Comparative Analysis of Deep Convolutional Neural Networks for Detecting Medical Image Deepfakes
di: Alsabbagh, Abdel Rahman, et al.
Pubblicazione: (2024)
di: Alsabbagh, Abdel Rahman, et al.
Pubblicazione: (2024)
Liveness Detection in Computer Vision: Transformer-based Self-Supervised Learning for Face Anti-Spoofing
di: Keresh, Arman, et al.
Pubblicazione: (2024)
di: Keresh, Arman, et al.
Pubblicazione: (2024)
TAB: Transformer Attention Bottlenecks enable User Intervention and Debugging in Vision-Language Models
di: Rahmanzadehgervi, Pooyan, et al.
Pubblicazione: (2024)
di: Rahmanzadehgervi, Pooyan, et al.
Pubblicazione: (2024)
Self-Supervised Vision Transformer for Enhanced Virtual Clothes Try-On
di: Lu, Lingxiao, et al.
Pubblicazione: (2024)
di: Lu, Lingxiao, et al.
Pubblicazione: (2024)
Agentic Copyright Watermarking against Adversarial Evidence Forgery with Purification-Agnostic Curriculum Proxy Learning
di: Bao, Erjin, et al.
Pubblicazione: (2024)
di: Bao, Erjin, et al.
Pubblicazione: (2024)
Uncolorable Examples: Preventing Unauthorized AI Colorization via Perception-Aware Chroma-Restrictive Perturbation
di: Nii, Yuki, et al.
Pubblicazione: (2025)
di: Nii, Yuki, et al.
Pubblicazione: (2025)
DRAGON: A Large-Scale Dataset of Realistic Images Generated by Diffusion Models
di: Bertazzini, Giulia, et al.
Pubblicazione: (2025)
di: Bertazzini, Giulia, et al.
Pubblicazione: (2025)
Comparative Analysis of Deepfake Detection Models: New Approaches and Perspectives
di: Batista, Matheus Martins
Pubblicazione: (2025)
di: Batista, Matheus Martins
Pubblicazione: (2025)
Unleashing Vision-Language Semantics for Deepfake Video Detection
di: Zhu, Jiawen, et al.
Pubblicazione: (2026)
di: Zhu, Jiawen, et al.
Pubblicazione: (2026)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
di: Lee, Seungho, et al.
Pubblicazione: (2024)
di: Lee, Seungho, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Exploring Active Data Selection Strategies for Continuous Training in Deepfake Detection
di: Furuhashi, Yoshihiko, et al.
Pubblicazione: (2025) -
A Controllable 3D Deepfake Generation Framework with Gaussian Splatting
di: Liu, Wending, et al.
Pubblicazione: (2025) -
Leveraging Chat-Based Large Vision Language Models for Multimodal Out-Of-Context Detection
di: Shalabi, Fatma, et al.
Pubblicazione: (2024) -
Defending Against Physical Adversarial Patch Attacks on Infrared Human Detection
di: Strack, Lukas, et al.
Pubblicazione: (2023) -
Beyond Standard Benchmarks: A Systematic Audit of Vision-Language Model's Robustness to Natural Semantic Variation Across Diverse Tasks
di: Chengyu, Jia, et al.
Pubblicazione: (2026)