Towards Generalizable Deepfake Image Detection with Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Srinanda, Kaliki V, Prabhu, M Manvith, Mogilipalem, Hemanth K, Abhinai, Jayavarapu S, Santhosh, Vaibhav, Herur, Aryan, Vijayasenan, Deepu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CASCRNet: An Atrous Spatial Pyramid Pooling and Shared Channel Residual based Network for Capsule Endoscopy
von: Srinanda, K V, et al.
Veröffentlicht: (2024)
von: Srinanda, K V, et al.
Veröffentlicht: (2024)
CoVid-19 Detection leveraging Vision Transformers and Explainable AI
von: Kumar, Pangoth Santhosh, et al.
Veröffentlicht: (2023)
von: Kumar, Pangoth Santhosh, et al.
Veröffentlicht: (2023)
Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026)
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026)
Image Provenance Analysis via Graph Encoding with Vision Transformer
von: Zhang, Keyang, et al.
Veröffentlicht: (2024)
von: Zhang, Keyang, et al.
Veröffentlicht: (2024)
GIGA: Generalizable Sparse Image-driven Gaussian Humans
von: Zubekhin, Anton, et al.
Veröffentlicht: (2025)
von: Zubekhin, Anton, et al.
Veröffentlicht: (2025)
Global and Local Attention-Based Transformer for Hyperspectral Image Change Detection
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
GANESH: Generalizable NeRF for Lensless Imaging
von: Madavan, Rakesh Raj, et al.
Veröffentlicht: (2024)
von: Madavan, Rakesh Raj, et al.
Veröffentlicht: (2024)
Towards Robust and Generalizable Lensless Imaging with Modular Learned Reconstruction
von: Bezzam, Eric, et al.
Veröffentlicht: (2025)
von: Bezzam, Eric, et al.
Veröffentlicht: (2025)
Learning Scan-Adaptive MRI Undersampling Patterns with Pre-Optimized Mask Supervision
von: Dhar, Aryan, et al.
Veröffentlicht: (2025)
von: Dhar, Aryan, et al.
Veröffentlicht: (2025)
Improved Generalizability of CNN Based Lane Detection in Challenging Weather Using Adaptive Preprocessing Parameter Tuning
von: Sang, I-Chen, et al.
Veröffentlicht: (2024)
von: Sang, I-Chen, et al.
Veröffentlicht: (2024)
FEFormer: Frequency-enhanced Vision Transformer for Generic Knowledge Extraction and Adaptive Feature Fusion in Volumetric Medical Image Segmentation
von: Yang, Jin, et al.
Veröffentlicht: (2026)
von: Yang, Jin, et al.
Veröffentlicht: (2026)
ELFATT: Efficient Linear Fast Attention for Vision Transformers
von: Wu, Chong, et al.
Veröffentlicht: (2025)
von: Wu, Chong, et al.
Veröffentlicht: (2025)
Towards Zero-Shot Task-Generalizable Learning on fMRI
von: Wang, Jiyao, et al.
Veröffentlicht: (2025)
von: Wang, Jiyao, et al.
Veröffentlicht: (2025)
Towards Generalizable Tumor Synthesis
von: Chen, Qi, et al.
Veröffentlicht: (2024)
von: Chen, Qi, et al.
Veröffentlicht: (2024)
Convolutional Transformer-Based Image Compression
von: Arezki, Bouzid, et al.
Veröffentlicht: (2024)
von: Arezki, Bouzid, et al.
Veröffentlicht: (2024)
Exploring Autoregressive Vision Foundation Models for Image Compression
von: Phung, Huu-Tai, et al.
Veröffentlicht: (2025)
von: Phung, Huu-Tai, et al.
Veröffentlicht: (2025)
Transforming Single Photon Camera Images to Color High Dynamic Range Images
von: Sharma, Sumit, et al.
Veröffentlicht: (2024)
von: Sharma, Sumit, et al.
Veröffentlicht: (2024)
Versatile Volumetric Medical Image Coding for Human-Machine Vision
von: Chen, Jietao, et al.
Veröffentlicht: (2024)
von: Chen, Jietao, et al.
Veröffentlicht: (2024)
Memory Efficient Corner Detection for Event-driven Dynamic Vision Sensors
von: Sun, Pao-Sheng Vincent, et al.
Veröffentlicht: (2024)
von: Sun, Pao-Sheng Vincent, et al.
Veröffentlicht: (2024)
A Method for Target Detection Based on Mmw Radar and Vision Fusion
von: Zong, Ming, et al.
Veröffentlicht: (2024)
von: Zong, Ming, et al.
Veröffentlicht: (2024)
Towards Backward-Compatible Continual Learning of Image Compression
von: Duan, Zhihao, et al.
Veröffentlicht: (2024)
von: Duan, Zhihao, et al.
Veröffentlicht: (2024)
Transfer CLIP for Generalizable Image Denoising
von: Cheng, Jun, et al.
Veröffentlicht: (2024)
von: Cheng, Jun, et al.
Veröffentlicht: (2024)
Fine-tuned Transformer Models for Breast Cancer Detection and Classification
von: Osman, Showkat, et al.
Veröffentlicht: (2025)
von: Osman, Showkat, et al.
Veröffentlicht: (2025)
Transformer-Based Local Feature Matching for Multimodal Image Registration
von: Delaunay, Remi, et al.
Veröffentlicht: (2024)
von: Delaunay, Remi, et al.
Veröffentlicht: (2024)
Transformer-based Learned Image Compression for Joint Decoding and Denoising
von: Chen, Yi-Hsin, et al.
Veröffentlicht: (2024)
von: Chen, Yi-Hsin, et al.
Veröffentlicht: (2024)
CompSRT: Quantization and Pruning for Image Super Resolution Transformers
von: Zeinali, Dorsa, et al.
Veröffentlicht: (2026)
von: Zeinali, Dorsa, et al.
Veröffentlicht: (2026)
Efficient and Accurate Tuberculosis Diagnosis: Attention Residual U-Net and Vision Transformer Based Detection Framework
von: K, Greeshma, et al.
Veröffentlicht: (2025)
von: K, Greeshma, et al.
Veröffentlicht: (2025)
Low Latency and Generalizable Dynamic MRI via L+S Alternating GD and Minimization
von: Babu, Silpa, et al.
Veröffentlicht: (2025)
von: Babu, Silpa, et al.
Veröffentlicht: (2025)
Explainable Detection of AI-Generated Images with Artifact Localization Using Faster-Than-Lies and Vision-Language Models for Edge Devices
von: Mathur, Aryan, et al.
Veröffentlicht: (2025)
von: Mathur, Aryan, et al.
Veröffentlicht: (2025)
Ship Detection in SAR Images with Human-in-the-Loop
von: Jia, Hecheng, et al.
Veröffentlicht: (2024)
von: Jia, Hecheng, et al.
Veröffentlicht: (2024)
Vision-Inspired Image Quality Assessment for Radar-Based Human Activity Representations
von: Trinh, Huy, et al.
Veröffentlicht: (2026)
von: Trinh, Huy, et al.
Veröffentlicht: (2026)
Invited Paper: BitMedViT: Ternary-Quantized Vision Transformer for Medical AI Assistants on the Edge
von: Walczak, Mikolaj, et al.
Veröffentlicht: (2025)
von: Walczak, Mikolaj, et al.
Veröffentlicht: (2025)
Exploring Strengths and Weaknesses of Super-Resolution Attack in Deepfake Detection
von: Coccomini, Davide Alessandro, et al.
Veröffentlicht: (2024)
von: Coccomini, Davide Alessandro, et al.
Veröffentlicht: (2024)
Improving Food Image Recognition with Noisy Vision Transformer
von: Ghosh, Tonmoy, et al.
Veröffentlicht: (2025)
von: Ghosh, Tonmoy, et al.
Veröffentlicht: (2025)
Toward Generalizable Multiple Sclerosis Lesion Segmentation Models
von: Badea, Liviu, et al.
Veröffentlicht: (2024)
von: Badea, Liviu, et al.
Veröffentlicht: (2024)
Generalized User-Oriented Image Semantic Coding Empowered by Large Vision-Language Model
von: Huang, Sin-Yu, et al.
Veröffentlicht: (2025)
von: Huang, Sin-Yu, et al.
Veröffentlicht: (2025)
Context-Aware Vision Language Foundation Models for Ocular Disease Screening in Retinal Images
von: Berger, Lucie, et al.
Veröffentlicht: (2025)
von: Berger, Lucie, et al.
Veröffentlicht: (2025)
ATFusion: An Alternate Cross-Attention Transformer Network for Infrared and Visible Image Fusion
von: Yan, Han, et al.
Veröffentlicht: (2024)
von: Yan, Han, et al.
Veröffentlicht: (2024)
RoTIR: Rotation-Equivariant Network and Transformers for Fish Scale Image Registration
von: Wang, Ruixiong, et al.
Veröffentlicht: (2024)
von: Wang, Ruixiong, et al.
Veröffentlicht: (2024)
HDRSplat: Gaussian Splatting for High Dynamic Range 3D Scene Reconstruction from Raw Images
von: Singh, Shreyas, et al.
Veröffentlicht: (2024)
von: Singh, Shreyas, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CASCRNet: An Atrous Spatial Pyramid Pooling and Shared Channel Residual based Network for Capsule Endoscopy
von: Srinanda, K V, et al.
Veröffentlicht: (2024) -
CoVid-19 Detection leveraging Vision Transformers and Explainable AI
von: Kumar, Pangoth Santhosh, et al.
Veröffentlicht: (2023) -
Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026) -
Image Provenance Analysis via Graph Encoding with Vision Transformer
von: Zhang, Keyang, et al.
Veröffentlicht: (2024) -
GIGA: Generalizable Sparse Image-driven Gaussian Humans
von: Zubekhin, Anton, et al.
Veröffentlicht: (2025)