Combined CNN and ViT features off-the-shelf: Another astounding baseline for recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alonso-Fernandez, Fernando, Hernandez-Diaz, Kevin, Tiwari, Prayag, Bigun, Josef |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Network Pruning: A Comparative Study on CNNs in Face Recognition
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2024)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2024)
Understanding and Improving CNNs with Complex Structure Tensor: A Biometrics Study
von: Hernandez-Diaz, Kevin, et al.
Veröffentlicht: (2024)
von: Hernandez-Diaz, Kevin, et al.
Veröffentlicht: (2024)
SqueezeFacePoseNet: Lightweight Face Verification Across Different Poses for Mobile Platforms
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2020)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2020)
Total Least Square Optimal Analytic Signal by Structure Tensor for N-D images
von: Bigun, Josef, et al.
Veröffentlicht: (2020)
von: Bigun, Josef, et al.
Veröffentlicht: (2020)
Leveraging Large-Scale Face Datasets for Deep Periocular Recognition via Ocular Cropping
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
Note on the Construction of Structure Tensor
von: Bigun, Josef, et al.
Veröffentlicht: (2025)
von: Bigun, Josef, et al.
Veröffentlicht: (2025)
Exploring Complementarity and Explainability in CNNs for Periocular Verification Across Acquisition Distances
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Overtake Detection in Trucks Using CAN Bus Signals: A Comparative Study of Machine Learning Methods
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)
Exploring the correlation between the type of music and the emotions evoked: A study using subjective questionnaires and EEG
von: Jankowska, Jelizaveta, et al.
Veröffentlicht: (2025)
von: Jankowska, Jelizaveta, et al.
Veröffentlicht: (2025)
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets
von: Amangeldi, Aidar, et al.
Veröffentlicht: (2025)
von: Amangeldi, Aidar, et al.
Veröffentlicht: (2025)
OmniPatch: A Universal Adversarial Patch for ViT-CNN Cross-Architecture Transfer in Semantic Segmentation
von: Aggarwal, Aarush, et al.
Veröffentlicht: (2026)
von: Aggarwal, Aarush, et al.
Veröffentlicht: (2026)
Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
Hybrid CNN-ViT Framework for Motion-Blurred Scene Text Restoration
von: Rashid, Umar, et al.
Veröffentlicht: (2025)
von: Rashid, Umar, et al.
Veröffentlicht: (2025)
Learning CNN on ViT: A Hybrid Model to Explicitly Class-specific Boundaries for Domain Adaptation
von: Ngo, Ba Hung, et al.
Veröffentlicht: (2024)
von: Ngo, Ba Hung, et al.
Veröffentlicht: (2024)
A Hybrid Framework Bridging CNN and ViT based on Theory of Evidence for Diabetic Retinopathy Grading
von: Qiu, Junlai, et al.
Veröffentlicht: (2025)
von: Qiu, Junlai, et al.
Veröffentlicht: (2025)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
ViT-Explainer: An Interactive Walkthrough of the Vision Transformer Pipeline
von: Hernandez, Juan Manuel, et al.
Veröffentlicht: (2026)
von: Hernandez, Juan Manuel, et al.
Veröffentlicht: (2026)
Pretrained ViTs Yield Versatile Representations For Medical Images
von: Matsoukas, Christos, et al.
Veröffentlicht: (2023)
von: Matsoukas, Christos, et al.
Veröffentlicht: (2023)
A Hybrid CNN-ViT-GNN Framework with GAN-Based Augmentation for Intelligent Weed Detection in Precision Agriculture
von: V, Pandiyaraju, et al.
Veröffentlicht: (2025)
von: V, Pandiyaraju, et al.
Veröffentlicht: (2025)
Mobile U-ViT: Revisiting large kernel and U-shaped ViT for efficient medical image segmentation
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
Rethinking Random Masking in Self-Distillation on ViT
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026)
von: Wang, Feng, et al.
Veröffentlicht: (2026)
CNN-ViT Hybrid for Pneumonia Detection: Theory and Empiric on Limited Data without Pretraining
von: Basnet, Prashant Singh, et al.
Veröffentlicht: (2025)
von: Basnet, Prashant Singh, et al.
Veröffentlicht: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
PaW-ViT: A Patch-based Warping Vision Transformer for Robust Ear Verification
von: Arun, Deeksha, et al.
Veröffentlicht: (2026)
von: Arun, Deeksha, et al.
Veröffentlicht: (2026)
ViTCAE: ViT-based Class-conditioned Autoencoder
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
Applying ViT in Generalized Few-shot Semantic Segmentation
von: Geng, Liyuan, et al.
Veröffentlicht: (2024)
von: Geng, Liyuan, et al.
Veröffentlicht: (2024)
One-Shot Multilingual Font Generation Via ViT
von: Wang, Zhiheng, et al.
Veröffentlicht: (2024)
von: Wang, Zhiheng, et al.
Veröffentlicht: (2024)
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
Register and [CLS] tokens yield a decoupling of local and global features in large ViTs
von: Lappe, Alexander, et al.
Veröffentlicht: (2025)
von: Lappe, Alexander, et al.
Veröffentlicht: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
H-CNN-ViT: A Hierarchical Gated Attention Multi-Branch Model for Bladder Cancer Recurrence Prediction
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
U-REPA: Aligning Diffusion U-Nets to ViTs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Deep Network Pruning: A Comparative Study on CNNs in Face Recognition
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2024) -
Understanding and Improving CNNs with Complex Structure Tensor: A Biometrics Study
von: Hernandez-Diaz, Kevin, et al.
Veröffentlicht: (2024) -
SqueezeFacePoseNet: Lightweight Face Verification Across Different Poses for Mobile Platforms
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2020) -
Total Least Square Optimal Analytic Signal by Structure Tensor for N-D images
von: Bigun, Josef, et al.
Veröffentlicht: (2020) -
Leveraging Large-Scale Face Datasets for Deep Periocular Recognition via Ocular Cropping
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2025)