Cross multiscale vision transformer for deep fake detection
Fuente:
arXiv
Saved in:
| Main Authors: | P, Akhshan, Sanjay, Taneti, S, Chandrakala |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fact-checking based fake news detection: a review
by: Yang, Yuzhou, et al.
Published: (2024)
by: Yang, Yuzhou, et al.
Published: (2024)
Deepfake detection in videos with multiple faces using geometric-fakeness features
by: Vyshegorodtsev, Kirill, et al.
Published: (2024)
by: Vyshegorodtsev, Kirill, et al.
Published: (2024)
Continuous fake media detection: adapting deepfake detectors to new generative techniques
by: Tassone, Francesco, et al.
Published: (2024)
by: Tassone, Francesco, et al.
Published: (2024)
METER: a mobile vision transformer architecture for monocular depth estimation
by: Papa, L., et al.
Published: (2024)
by: Papa, L., et al.
Published: (2024)
BFA-YOLO: A balanced multiscale object detection network for building façade attachments detection
by: Chen, Yangguang, et al.
Published: (2024)
by: Chen, Yangguang, et al.
Published: (2024)
Understanding vision transformer robustness through the lens of out-of-distribution detection
by: Kuang, Joey, et al.
Published: (2026)
by: Kuang, Joey, et al.
Published: (2026)
Generalized Face Liveness Detection via De-fake Face Generator
by: Long, Xingming, et al.
Published: (2024)
by: Long, Xingming, et al.
Published: (2024)
an interpretable vision transformer framework for automated brain tumor classification
by: Mbonu, Chinedu Emmanuel, et al.
Published: (2026)
by: Mbonu, Chinedu Emmanuel, et al.
Published: (2026)
Beyond the final layer: Attentive multilayer fusion for vision transformers
by: Ciernik, Laure, et al.
Published: (2026)
by: Ciernik, Laure, et al.
Published: (2026)
Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG
by: Madavan, Rakesh Raj, et al.
Published: (2025)
by: Madavan, Rakesh Raj, et al.
Published: (2025)
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking
by: Papa, Lorenzo, et al.
Published: (2023)
by: Papa, Lorenzo, et al.
Published: (2023)
Self-supervised pretraining for an iterative image size agnostic vision transformer
by: Prisadnikov, Nedyalko, et al.
Published: (2026)
by: Prisadnikov, Nedyalko, et al.
Published: (2026)
SAFER: Sharpness Aware layer-selective Finetuning for Enhanced Robustness in vision transformers
by: Gopal, Bhavna, et al.
Published: (2025)
by: Gopal, Bhavna, et al.
Published: (2025)
Low-latency vision transformers via large-scale multi-head attention
by: Gross, Ronit D., et al.
Published: (2025)
by: Gross, Ronit D., et al.
Published: (2025)
Small data deep learning methodology for in-field disease detection
by: Herrera-Poyato, David, et al.
Published: (2024)
by: Herrera-Poyato, David, et al.
Published: (2024)
Change detection needs change information: improving deep 3D point cloud change detection
by: de Gélis, Iris, et al.
Published: (2023)
by: de Gélis, Iris, et al.
Published: (2023)
Enhancing cross-domain detection: adaptive class-aware contrastive transformer
by: Zeng, Ziru, et al.
Published: (2024)
by: Zeng, Ziru, et al.
Published: (2024)
PTQ4ViT: Post-training quantization for vision transformers with twin uniform quantization
by: Yuan, Zhihang, et al.
Published: (2021)
by: Yuan, Zhihang, et al.
Published: (2021)
EdgeFormer: local patch-based edge detection transformer on point clouds
by: Xie, Yifei, et al.
Published: (2026)
by: Xie, Yifei, et al.
Published: (2026)
SpecDETR: A transformer-based hyperspectral point object detection network
by: Li, Zhaoxu, et al.
Published: (2024)
by: Li, Zhaoxu, et al.
Published: (2024)
Shallow- and Deep-fake Image Manipulation Localization Using Vision Mamba and Guided Graph Neural Network
by: Zhang, Junbin, et al.
Published: (2026)
by: Zhang, Junbin, et al.
Published: (2026)
Anatomy of a failure: When, how, and why deep vision fails in scientific domains
by: Oh, Ji-Hun, et al.
Published: (2026)
by: Oh, Ji-Hun, et al.
Published: (2026)
A comprehensive review of datasets and deep learning techniques for vision in Unmanned Surface Vehicles
by: Trinh, Linh, et al.
Published: (2024)
by: Trinh, Linh, et al.
Published: (2024)
A novel method to enhance pneumonia detection via a model-level ensembling of CNN and vision transformer
by: Angara, Sandeep, et al.
Published: (2024)
by: Angara, Sandeep, et al.
Published: (2024)
Beyond flattening: a geometrically principled positional encoding for vision transformers with Weierstrass elliptic functions
by: Xin, Zhihang, et al.
Published: (2025)
by: Xin, Zhihang, et al.
Published: (2025)
Contrastive learning-based video quality assessment-jointed video vision transformer for video recognition
by: Sun, Jian, et al.
Published: (2026)
by: Sun, Jian, et al.
Published: (2026)
A self-supervised text-vision framework for automated brain abnormality detection
by: Wood, David A., et al.
Published: (2024)
by: Wood, David A., et al.
Published: (2024)
Modeling Cross-vision Synergy for Unified Large Vision Model
by: Wu, Shengqiong, et al.
Published: (2026)
by: Wu, Shengqiong, et al.
Published: (2026)
Interpreting vision transformers via residual replacement model
by: Kim, Jinyeong, et al.
Published: (2025)
by: Kim, Jinyeong, et al.
Published: (2025)
First qualitative observations on deep learning vision model YOLO and DETR for automated driving in Austria
by: Schoder, Stefan
Published: (2023)
by: Schoder, Stefan
Published: (2023)
A hierarchical semantic segmentation framework for computer vision-based bridge damage detection
by: Liu, Jingxiao, et al.
Published: (2022)
by: Liu, Jingxiao, et al.
Published: (2022)
A multi-center analysis of deep learning methods for video polyp detection and segmentation
by: Ghatwary, Noha, et al.
Published: (2026)
by: Ghatwary, Noha, et al.
Published: (2026)
X-ray illicit object detection using hybrid CNN-transformer neural network architectures
by: Cani, Jorgen, et al.
Published: (2025)
by: Cani, Jorgen, et al.
Published: (2025)
Real, fake and synthetic faces -- does the coin have three sides?
by: Naeem, Shahzeb, et al.
Published: (2024)
by: Naeem, Shahzeb, et al.
Published: (2024)
Automatic location detection based on deep learning
by: Karangiya, Anjali, et al.
Published: (2024)
by: Karangiya, Anjali, et al.
Published: (2024)
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
Octave-YOLO: Cross frequency detection network with octave convolution
by: Shin, Sangjune, et al.
Published: (2024)
by: Shin, Sangjune, et al.
Published: (2024)
Oriented object detection in optical remote sensing images using deep learning: a survey
by: Wang, Kun, et al.
Published: (2023)
by: Wang, Kun, et al.
Published: (2023)
CapST: Leveraging Capsule Networks and Temporal Attention for Accurate Model Attribution in Deep-fake Videos
by: Ahmad, Wasim, et al.
Published: (2023)
by: Ahmad, Wasim, et al.
Published: (2023)
Illicit object detection in X-ray imaging using deep learning techniques: A comparative evaluation
by: Cani, Jorgen, et al.
Published: (2025)
by: Cani, Jorgen, et al.
Published: (2025)
Similar Items
-
Fact-checking based fake news detection: a review
by: Yang, Yuzhou, et al.
Published: (2024) -
Deepfake detection in videos with multiple faces using geometric-fakeness features
by: Vyshegorodtsev, Kirill, et al.
Published: (2024) -
Continuous fake media detection: adapting deepfake detectors to new generative techniques
by: Tassone, Francesco, et al.
Published: (2024) -
METER: a mobile vision transformer architecture for monocular depth estimation
by: Papa, L., et al.
Published: (2024) -
BFA-YOLO: A balanced multiscale object detection network for building façade attachments detection
by: Chen, Yangguang, et al.
Published: (2024)