Do Transformers Understand Ancient Roman Coin Motifs Better than CNNs?
Fuente:
arXiv
Saved in:
| Main Authors: | Reid, David, Arandjelovic, Ognjen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A White-Box False Positive Adversarial Attack Method on Contrastive Loss Based Offline Handwritten Signature Verification Models
by: Guo, Zhongliang, et al.
Published: (2023)
by: Guo, Zhongliang, et al.
Published: (2023)
This Looks Better than That: Better Interpretable Models with ProtoPNeXt
by: Willard, Frank, et al.
Published: (2024)
by: Willard, Frank, et al.
Published: (2024)
TraNCE: Transformative Non-linear Concept Explainer for CNNs
by: Akpudo, Ugochukwu Ejike, et al.
Published: (2025)
by: Akpudo, Ugochukwu Ejike, et al.
Published: (2025)
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2024)
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2024)
A Gray-box Attack against Latent Diffusion Model-based Image Editing by Posterior Collapse
by: Guo, Zhongliang, et al.
Published: (2024)
by: Guo, Zhongliang, et al.
Published: (2024)
RNNs, CNNs and Transformers in Human Action Recognition: A Survey and a Hybrid Model
by: Alomar, Khaled, et al.
Published: (2024)
by: Alomar, Khaled, et al.
Published: (2024)
VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
by: Scabini, Leonardo, et al.
Published: (2025)
by: Scabini, Leonardo, et al.
Published: (2025)
Efficient Hyperparameter Importance Assessment for CNNs
by: Wang, Ruinan, et al.
Published: (2024)
by: Wang, Ruinan, et al.
Published: (2024)
Steering the LoCoMotif: Using Domain Knowledge in Time Series Motif Discovery
by: Yurtman, Aras, et al.
Published: (2025)
by: Yurtman, Aras, et al.
Published: (2025)
Detecção da Psoríase Utilizando Visão Computacional: Uma Abordagem Comparativa Entre CNNs e Vision Transformers
by: Lucena, Natanael, et al.
Published: (2025)
by: Lucena, Natanael, et al.
Published: (2025)
CNNs Avoid Curse of Dimensionality by Learning on Patches
by: Madala, Vamshi C., et al.
Published: (2022)
by: Madala, Vamshi C., et al.
Published: (2022)
Explaning with trees: interpreting CNNs using hierarchies
by: Rodrigues, Caroline Mazini, et al.
Published: (2024)
by: Rodrigues, Caroline Mazini, et al.
Published: (2024)
From Ground to Air: Noise Robustness in Vision Transformers and CNNs for Event-Based Vehicle Classification with Potential UAV Applications
by: Almesafri, Nouf, et al.
Published: (2025)
by: Almesafri, Nouf, et al.
Published: (2025)
Better Understanding Differences in Attribution Methods via Systematic Evaluations
by: Rao, Sukrut, et al.
Published: (2023)
by: Rao, Sukrut, et al.
Published: (2023)
Explaining Model Overfitting in CNNs via GMM Clustering
by: Dou, Hui, et al.
Published: (2024)
by: Dou, Hui, et al.
Published: (2024)
Evaluating the Stability of Semantic Concept Representations in CNNs for Robust Explainability
by: Mikriukov, Georgii, et al.
Published: (2023)
by: Mikriukov, Georgii, et al.
Published: (2023)
Do Language Models Understand Time?
by: Ding, Xi, et al.
Published: (2024)
by: Ding, Xi, et al.
Published: (2024)
Reliable Evaluation of Attribution Maps in CNNs: A Perturbation-Based Approach
by: Nieradzik, Lars, et al.
Published: (2024)
by: Nieradzik, Lars, et al.
Published: (2024)
Stacked Ensemble of Fine-Tuned CNNs for Knee Osteoarthritis Severity Grading
by: Gupta, Adarsh, et al.
Published: (2025)
by: Gupta, Adarsh, et al.
Published: (2025)
Understanding Multi-View Transformers
by: Stary, Michal, et al.
Published: (2025)
by: Stary, Michal, et al.
Published: (2025)
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
by: Mehri, Faridoun, et al.
Published: (2024)
by: Mehri, Faridoun, et al.
Published: (2024)
Data or Language Supervision: What Makes CLIP Better than DINO?
by: Liu, Yiming, et al.
Published: (2025)
by: Liu, Yiming, et al.
Published: (2025)
ScribFormer: Transformer Makes CNN Work Better for Scribble-based Medical Image Segmentation
by: Li, Zihan, et al.
Published: (2024)
by: Li, Zihan, et al.
Published: (2024)
ImageNet-trained CNNs are not biased towards texture: Revisiting feature reliance through controlled suppression
by: Burgert, Tom, et al.
Published: (2025)
by: Burgert, Tom, et al.
Published: (2025)
Frugal Federated Learning for Violence Detection: A Comparison of LoRA-Tuned VLMs and Personalized CNNs
by: Thuau, Sébastien, et al.
Published: (2025)
by: Thuau, Sébastien, et al.
Published: (2025)
A Comparative Study of Custom CNNs, Pre-trained Models, and Transfer Learning Across Multiple Visual Datasets
by: Akhand, Annoor Sharara
Published: (2026)
by: Akhand, Annoor Sharara
Published: (2026)
ImageDDI: Image-enhanced Molecular Motif Sequence Representation for Drug-Drug Interaction Prediction
by: He, Yuqin, et al.
Published: (2025)
by: He, Yuqin, et al.
Published: (2025)
Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models
by: Rao, Abinav, et al.
Published: (2026)
by: Rao, Abinav, et al.
Published: (2026)
Drawing the Line: Deep Segmentation for Extracting Art from Ancient Etruscan Mirrors
by: Sterzinger, Rafael, et al.
Published: (2024)
by: Sterzinger, Rafael, et al.
Published: (2024)
Federated Learning for Video Violence Detection: Complementary Roles of Lightweight CNNs and Vision-Language Models for Energy-Efficient Use
by: Thuau, Sébastien, et al.
Published: (2025)
by: Thuau, Sébastien, et al.
Published: (2025)
Role of Locality and Weight Sharing in Image-Based Tasks: A Sample Complexity Separation between CNNs, LCNs, and FCNs
by: Lahoti, Aakash, et al.
Published: (2024)
by: Lahoti, Aakash, et al.
Published: (2024)
An accuracy-aware extension to LRP-based pruning for CNNs to prevent cascading accuracy degradation in data-scarce transfer learning
by: Yasui, Daisuke, et al.
Published: (2025)
by: Yasui, Daisuke, et al.
Published: (2025)
Efficient CNNs via Passive Filter Pruning
by: Singh, Arshdeep, et al.
Published: (2023)
by: Singh, Arshdeep, et al.
Published: (2023)
Generalization of CNNs on Relational Reasoning with Bar Charts
by: Cui, Zhenxing, et al.
Published: (2025)
by: Cui, Zhenxing, et al.
Published: (2025)
ULTra: Unveiling Latent Token Interpretability in Transformer-Based Understanding and Segmentation
by: Hosseini, Hesam, et al.
Published: (2024)
by: Hosseini, Hesam, et al.
Published: (2024)
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
by: Yang, Wenhan, et al.
Published: (2023)
by: Yang, Wenhan, et al.
Published: (2023)
How Do I Do That? Synthesizing 3D Hand Motion and Contacts for Everyday Interactions
by: Prakash, Aditya, et al.
Published: (2025)
by: Prakash, Aditya, et al.
Published: (2025)
Learning the RoPEs: Better 2D and 3D Position Encodings with STRING
by: Schenck, Connor, et al.
Published: (2025)
by: Schenck, Connor, et al.
Published: (2025)
PlantDiseaseNet-RT50: A Fine-tuned ResNet50 Architecture for High-Accuracy Plant Disease Detection Beyond Standard CNNs
by: Sagnika, Santwana, et al.
Published: (2025)
by: Sagnika, Santwana, et al.
Published: (2025)
Provably Better Explanations with Optimized Aggregation of Feature Attributions
by: Decker, Thomas, et al.
Published: (2024)
by: Decker, Thomas, et al.
Published: (2024)
Similar Items
-
A White-Box False Positive Adversarial Attack Method on Contrastive Loss Based Offline Handwritten Signature Verification Models
by: Guo, Zhongliang, et al.
Published: (2023) -
This Looks Better than That: Better Interpretable Models with ProtoPNeXt
by: Willard, Frank, et al.
Published: (2024) -
TraNCE: Transformative Non-linear Concept Explainer for CNNs
by: Akpudo, Ugochukwu Ejike, et al.
Published: (2025) -
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2024) -
A Gray-box Attack against Latent Diffusion Model-based Image Editing by Posterior Collapse
by: Guo, Zhongliang, et al.
Published: (2024)