Tex-ViT: A Generalizable, Robust, Texture-based dual-branch cross-attention deepfake detector
Fuente:
arXiv
Guardado en:
| Autores principales: | Dagar, Deepak, Vishwakarma, Dinesh Kumar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Noise and Edge extraction-based dual-branch method for Shallowfake and Deepfake Localization
por: Dagar, Deepak, et al.
Publicado: (2024)
por: Dagar, Deepak, et al.
Publicado: (2024)
ViTCAE: ViT-based Class-conditioned Autoencoder
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
por: Chen, Lu, et al.
Publicado: (2025)
por: Chen, Lu, et al.
Publicado: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
How to train your ViT for OOD Detection
por: Mueller, Maximilian, et al.
Publicado: (2024)
por: Mueller, Maximilian, et al.
Publicado: (2024)
InsTex: Indoor Scenes Stylized Texture Synthesis
por: Zhang, Yunfan, et al.
Publicado: (2025)
por: Zhang, Yunfan, et al.
Publicado: (2025)
Triamese-ViT: A 3D-Aware Method for Robust Brain Age Estimation from MRIs
por: Zhang, Zhaonian, et al.
Publicado: (2024)
por: Zhang, Zhaonian, et al.
Publicado: (2024)
HydraViT: Stacking Heads for a Scalable ViT
por: Haberer, Janek, et al.
Publicado: (2024)
por: Haberer, Janek, et al.
Publicado: (2024)
FlashTex: Fast Relightable Mesh Texturing with LightControlNet
por: Deng, Kangle, et al.
Publicado: (2024)
por: Deng, Kangle, et al.
Publicado: (2024)
ZACH-ViT: A Zero-Token Vision Transformer with ShuffleStrides Data Augmentation for Robust Lung Ultrasound Classification
por: Angelakis, Athanasios, et al.
Publicado: (2025)
por: Angelakis, Athanasios, et al.
Publicado: (2025)
ProtoS-ViT: Visual foundation models for sparse self-explainable classifications
por: Turbé, Hugues, et al.
Publicado: (2024)
por: Turbé, Hugues, et al.
Publicado: (2024)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
por: Kang, Ben, et al.
Publicado: (2025)
por: Kang, Ben, et al.
Publicado: (2025)
Token Cropr: Faster ViTs for Quite a Few Tasks
por: Bergner, Benjamin, et al.
Publicado: (2024)
por: Bergner, Benjamin, et al.
Publicado: (2024)
CubistMerge: Spatial-Preserving Token Merging For Diverse ViT Backbones
por: Gong, Wenyi, et al.
Publicado: (2025)
por: Gong, Wenyi, et al.
Publicado: (2025)
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
por: Sinha, Sanchit, et al.
Publicado: (2025)
por: Sinha, Sanchit, et al.
Publicado: (2025)
Exploring the Synergies of Hybrid CNNs and ViTs Architectures for Computer Vision: A survey
por: Yunusa, Haruna, et al.
Publicado: (2024)
por: Yunusa, Haruna, et al.
Publicado: (2024)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
por: Nag, Shashank, et al.
Publicado: (2025)
por: Nag, Shashank, et al.
Publicado: (2025)
Layer by layer, module by module: Choose both for optimal OOD probing of ViT
por: Odonnat, Ambroise, et al.
Publicado: (2026)
por: Odonnat, Ambroise, et al.
Publicado: (2026)
S-E Pipeline: A Vision Transformer (ViT) based Resilient Classification Pipeline for Medical Imaging Against Adversarial Attacks
por: S, Neha A, et al.
Publicado: (2024)
por: S, Neha A, et al.
Publicado: (2024)
ViT-MUL: A Baseline Study on Recent Machine Unlearning Methods Applied to Vision Transformers
por: Cho, Ikhyun, et al.
Publicado: (2024)
por: Cho, Ikhyun, et al.
Publicado: (2024)
Communication Efficient Split Learning of ViTs with Attention-based Double Compression
por: Alvetreti, Federico, et al.
Publicado: (2025)
por: Alvetreti, Federico, et al.
Publicado: (2025)
Causality $\neq$ Decodability, and Vice Versa: Lessons from Interpreting Counting ViTs
por: Huang, Lianghuan, et al.
Publicado: (2025)
por: Huang, Lianghuan, et al.
Publicado: (2025)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
por: Hwang, Dongyoon, et al.
Publicado: (2024)
por: Hwang, Dongyoon, et al.
Publicado: (2024)
MobilePlantViT: A Mobile-friendly Hybrid ViT for Generalized Plant Disease Image Classification
por: Tonmoy, Moshiur Rahman, et al.
Publicado: (2025)
por: Tonmoy, Moshiur Rahman, et al.
Publicado: (2025)
Register and [CLS] tokens yield a decoupling of local and global features in large ViTs
por: Lappe, Alexander, et al.
Publicado: (2025)
por: Lappe, Alexander, et al.
Publicado: (2025)
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
por: Bourriez, Nicolas, et al.
Publicado: (2023)
por: Bourriez, Nicolas, et al.
Publicado: (2023)
GrowTAS: Progressive Expansion from Small to Large Subnets for Efficient ViT Architecture Search
por: Lee, Hyunju, et al.
Publicado: (2025)
por: Lee, Hyunju, et al.
Publicado: (2025)
Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness
por: Elisha, Yehonatan, et al.
Publicado: (2026)
por: Elisha, Yehonatan, et al.
Publicado: (2026)
Training-Free Acceleration of ViTs with Delayed Spatial Merging
por: Heo, Jung Hwan, et al.
Publicado: (2023)
por: Heo, Jung Hwan, et al.
Publicado: (2023)
Octic Vision Transformers: Quicker ViTs Through Equivariance
por: Nordström, David, et al.
Publicado: (2025)
por: Nordström, David, et al.
Publicado: (2025)
DeCLIP: Decoding CLIP representations for deepfake localization
por: Smeu, Stefan, et al.
Publicado: (2024)
por: Smeu, Stefan, et al.
Publicado: (2024)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
por: Han, Donghoon, et al.
Publicado: (2023)
por: Han, Donghoon, et al.
Publicado: (2023)
ViT Enhanced Privacy-Preserving Secure Medical Data Sharing and Classification
por: Amin, Al, et al.
Publicado: (2024)
por: Amin, Al, et al.
Publicado: (2024)
Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring
por: Huang, Youhan, et al.
Publicado: (2026)
por: Huang, Youhan, et al.
Publicado: (2026)
ViT-ProtoNet for Few-Shot Image Classification: A Multi-Benchmark Evaluation
por: Mutlu, Abdulvahap, et al.
Publicado: (2025)
por: Mutlu, Abdulvahap, et al.
Publicado: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
por: Salzmann, Tim, et al.
Publicado: (2024)
por: Salzmann, Tim, et al.
Publicado: (2024)
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs
por: Kim, Donghyun, et al.
Publicado: (2024)
por: Kim, Donghyun, et al.
Publicado: (2024)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
por: Li, Zhengang, et al.
Publicado: (2024)
por: Li, Zhengang, et al.
Publicado: (2024)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
por: Wang, Zhibo, et al.
Publicado: (2026)
por: Wang, Zhibo, et al.
Publicado: (2026)
MC-ViViT: Multi-branch Classifier-ViViT to detect Mild Cognitive Impairment in older adults using facial videos
por: Sun, Jian, et al.
Publicado: (2023)
por: Sun, Jian, et al.
Publicado: (2023)
Ejemplares similares
-
A Noise and Edge extraction-based dual-branch method for Shallowfake and Deepfake Localization
por: Dagar, Deepak, et al.
Publicado: (2024) -
ViTCAE: ViT-based Class-conditioned Autoencoder
por: Jebraeeli, Vahid, et al.
Publicado: (2025) -
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
por: Chen, Lu, et al.
Publicado: (2025) -
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
por: Chattopadhyay, Nandish, et al.
Publicado: (2026) -
How to train your ViT for OOD Detection
por: Mueller, Maximilian, et al.
Publicado: (2024)