AI-Powered Deepfake Detection Using CNN and Vision Transformer Architectures
Fuente:
arXiv
Guardado en:
| Autores principales: | Urmi, Sifatullah Sheikh, Arthi, Kirtonia Nuzath Tabassum, Al-Imran, Md |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Generalizable Deepfake Image Detection with Vision Transformers
por: Srinanda, Kaliki V, et al.
Publicado: (2026)
por: Srinanda, Kaliki V, et al.
Publicado: (2026)
Placenta Accreta Spectrum Detection Using an MRI-based Hybrid CNN-Transformer Model
por: Ali, Sumaiya, et al.
Publicado: (2025)
por: Ali, Sumaiya, et al.
Publicado: (2025)
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
por: Mia, Shakil, et al.
Publicado: (2026)
por: Mia, Shakil, et al.
Publicado: (2026)
Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
por: Das, Dabbrata, et al.
Publicado: (2025)
por: Das, Dabbrata, et al.
Publicado: (2025)
Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection
por: Delibasoglu, Ibrahim
Publicado: (2026)
por: Delibasoglu, Ibrahim
Publicado: (2026)
Leaf Angle Estimation using Mask R-CNN and LETR Vision Transformer
por: Margapuri, Venkat, et al.
Publicado: (2024)
por: Margapuri, Venkat, et al.
Publicado: (2024)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
por: ALBarqawi, Ahmad, et al.
Publicado: (2025)
por: ALBarqawi, Ahmad, et al.
Publicado: (2025)
Optimizing CNN Architectures for Advanced Thoracic Disease Classification
por: Mirthipati, Tejas
Publicado: (2025)
por: Mirthipati, Tejas
Publicado: (2025)
SynthEnsemble: A Fusion of CNN, Vision Transformer, and Hybrid Models for Multi-Label Chest X-Ray Classification
por: Ashraf, S. M. Nabil, et al.
Publicado: (2023)
por: Ashraf, S. M. Nabil, et al.
Publicado: (2023)
WiTUnet: A U-Shaped Architecture Integrating CNN and Transformer for Improved Feature Alignment and Local Information Fusion
por: Wang, Bin, et al.
Publicado: (2024)
por: Wang, Bin, et al.
Publicado: (2024)
Uncovering Critical Features for Deepfake Detection through the Lottery Ticket Hypothesis
por: Amin, Lisan Al, et al.
Publicado: (2025)
por: Amin, Lisan Al, et al.
Publicado: (2025)
Flying By ML -- CNN Inversion of Affine Transforms
por: Van Warren, L.
Publicado: (2023)
por: Van Warren, L.
Publicado: (2023)
Intriguing Equivalence Structures of the Embedding Space of Vision Transformers
por: Salman, Shaeke, et al.
Publicado: (2024)
por: Salman, Shaeke, et al.
Publicado: (2024)
RS-CA-HSICT: A Residual and Spatial Channel Augmented CNN Transformer Framework for Monkeypox Detection
por: Iqbal, Rashid, et al.
Publicado: (2025)
por: Iqbal, Rashid, et al.
Publicado: (2025)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
por: Li, Zhengang, et al.
Publicado: (2024)
por: Li, Zhengang, et al.
Publicado: (2024)
Multi-modal Deepfake Detection and Localization with FPN-Transformer
por: Zheng, Chende, et al.
Publicado: (2025)
por: Zheng, Chende, et al.
Publicado: (2025)
Real Time American Sign Language Detection Using Yolo-v9
por: Imran, Amna, et al.
Publicado: (2024)
por: Imran, Amna, et al.
Publicado: (2024)
Is It Certainly a Deepfake? Reliability Analysis in Detection & Generation Ecosystem
por: Kose, Neslihan, et al.
Publicado: (2025)
por: Kose, Neslihan, et al.
Publicado: (2025)
Wavelet-Driven Generalizable Framework for Deepfake Face Forgery Detection
por: Baru, Lalith Bharadwaj, et al.
Publicado: (2024)
por: Baru, Lalith Bharadwaj, et al.
Publicado: (2024)
Enhancing Vision Transformer Explainability Using Artificial Astrocytes
por: Echevarrieta-Catalan, Nicolas, et al.
Publicado: (2025)
por: Echevarrieta-Catalan, Nicolas, et al.
Publicado: (2025)
Feature Fusion for Improved Classification: Combining Dempster-Shafer Theory and Multiple CNN Architectures
por: Alzahem, Ayyub, et al.
Publicado: (2024)
por: Alzahem, Ayyub, et al.
Publicado: (2024)
Balancing Accuracy and Efficiency: CNN Fusion Models for Diabetic Retinopathy Screening
por: Islam, Md Rafid, et al.
Publicado: (2025)
por: Islam, Md Rafid, et al.
Publicado: (2025)
X-AVDT: Audio-Visual Cross-Attention for Robust Deepfake Detection
por: Kim, Youngseo, et al.
Publicado: (2026)
por: Kim, Youngseo, et al.
Publicado: (2026)
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition
por: Wang, Xiao, et al.
Publicado: (2023)
por: Wang, Xiao, et al.
Publicado: (2023)
Pre-Trained CNN Architecture for Transformer-Based Image Caption Generation Model
por: Dufera, Amanuel Tafese
Publicado: (2025)
por: Dufera, Amanuel Tafese
Publicado: (2025)
When Does Supervised Training Pay Off? The Hidden Economics of Object Detection in the Era of Vision-Language Models
por: Al-Hamadani, Samer
Publicado: (2025)
por: Al-Hamadani, Samer
Publicado: (2025)
AI-Powered Intracranial Hemorrhage Detection: A Co-Scale Convolutional Attention Model with Uncertainty-Based Fuzzy Integral Operator and Feature Screening
por: Chagahi, Mehdi Hosseini, et al.
Publicado: (2024)
por: Chagahi, Mehdi Hosseini, et al.
Publicado: (2024)
UPDP: A Unified Progressive Depth Pruner for CNN and Vision Transformer
por: Liu, Ji, et al.
Publicado: (2024)
por: Liu, Ji, et al.
Publicado: (2024)
Intriguing Differences Between Zero-Shot and Systematic Evaluations of Vision-Language Transformer Models
por: Salman, Shaeke, et al.
Publicado: (2024)
por: Salman, Shaeke, et al.
Publicado: (2024)
Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection
por: Tsigos, Konstantinos, et al.
Publicado: (2024)
por: Tsigos, Konstantinos, et al.
Publicado: (2024)
Advancing AI-Powered Medical Image Synthesis: Insights from MedVQA-GI Challenge Using CLIP, Fine-Tuned Stable Diffusion, and Dream-Booth + LoRA
por: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Publicado: (2025)
por: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Publicado: (2025)
Wildfire Detection Using Vision Transformer with the Wildfire Dataset
por: Vuppari, Gowtham Raj, et al.
Publicado: (2025)
por: Vuppari, Gowtham Raj, et al.
Publicado: (2025)
OmniPatch: A Universal Adversarial Patch for ViT-CNN Cross-Architecture Transfer in Semantic Segmentation
por: Aggarwal, Aarush, et al.
Publicado: (2026)
por: Aggarwal, Aarush, et al.
Publicado: (2026)
A Comparative Study of Adversarial Robustness in CNN and CNN-ANFIS Architectures
por: Shankar, Kaaustaaub, et al.
Publicado: (2026)
por: Shankar, Kaaustaaub, et al.
Publicado: (2026)
ZAYAN: Disentangled Contrastive Transformer for Tabular Remote Sensing Data
por: Habib, Al Zadid Sultan Bin, et al.
Publicado: (2026)
por: Habib, Al Zadid Sultan Bin, et al.
Publicado: (2026)
Revisiting Deepfake Detection: Chronological Continual Learning and the Limits of Generalization
por: Fontana, Federico, et al.
Publicado: (2025)
por: Fontana, Federico, et al.
Publicado: (2025)
Evaluating Deepfake Detectors in the Wild
por: Pirogov, Viacheslav, et al.
Publicado: (2025)
por: Pirogov, Viacheslav, et al.
Publicado: (2025)
ViTs are Everywhere: A Comprehensive Study Showcasing Vision Transformers in Different Domain
por: Mia, Md Sohag, et al.
Publicado: (2023)
por: Mia, Md Sohag, et al.
Publicado: (2023)
A Clinically Interpretable Deep CNN Framework for Early Chronic Kidney Disease Prediction Using Grad-CAM-Based Explainable AI
por: Ayub, Anas Bin, et al.
Publicado: (2025)
por: Ayub, Anas Bin, et al.
Publicado: (2025)
Surveillance Video-Based Traffic Accident Detection Using Transformer Architecture
por: Singh, Tanu, et al.
Publicado: (2025)
por: Singh, Tanu, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Generalizable Deepfake Image Detection with Vision Transformers
por: Srinanda, Kaliki V, et al.
Publicado: (2026) -
Placenta Accreta Spectrum Detection Using an MRI-based Hybrid CNN-Transformer Model
por: Ali, Sumaiya, et al.
Publicado: (2025) -
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
por: Mia, Shakil, et al.
Publicado: (2026) -
Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
por: Das, Dabbrata, et al.
Publicado: (2025) -
Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection
por: Delibasoglu, Ibrahim
Publicado: (2026)