A Saccade-inspired Approach to Image Classification using Vision Transformer Attention Maps
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dallain, Matthis, Rodriguez, Laurent, Perrinet, Laurent Udo, Miramond, Benoît |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Foveated Retinotopy Improves Classification and Localization in Convolutional Neural Networks
von: Jérémie, Jean-Nicolas, et al.
Veröffentlicht: (2024)
von: Jérémie, Jean-Nicolas, et al.
Veröffentlicht: (2024)
Saccadic Vision for Fine-Grained Visual Classification
von: Schmidt, Johann, et al.
Veröffentlicht: (2025)
von: Schmidt, Johann, et al.
Veröffentlicht: (2025)
Spiking monocular event based 6D pose estimation for space application
von: Courtois, Jonathan, et al.
Veröffentlicht: (2025)
von: Courtois, Jonathan, et al.
Veröffentlicht: (2025)
Emergence of Fixational and Saccadic Movements in a Multi-Level Recurrent Attention Model for Vision
von: Pan, Pengcheng, et al.
Veröffentlicht: (2025)
von: Pan, Pengcheng, et al.
Veröffentlicht: (2025)
Lightweight Vision Transformer with Window and Spatial Attention for Food Image Classification
von: Gao, Xinle, et al.
Veröffentlicht: (2025)
von: Gao, Xinle, et al.
Veröffentlicht: (2025)
On Reducing Activity with Distillation and Regularization for Energy Efficient Spiking Neural Networks
von: Louis, Thomas, et al.
Veröffentlicht: (2024)
von: Louis, Thomas, et al.
Veröffentlicht: (2024)
Enhanced Neuromorphic Semantic Segmentation Latency through Stream Event
von: Hareb, D., et al.
Veröffentlicht: (2025)
von: Hareb, D., et al.
Veröffentlicht: (2025)
Reading Is Believing: Revisiting Language Bottleneck Models for Image Classification
von: Udo, Honori, et al.
Veröffentlicht: (2024)
von: Udo, Honori, et al.
Veröffentlicht: (2024)
SaccadeDet: A Novel Dual-Stage Architecture for Rapid and Accurate Detection in Gigapixel Images
von: Li, Wenxi, et al.
Veröffentlicht: (2024)
von: Li, Wenxi, et al.
Veröffentlicht: (2024)
Vision Transformer for Classification of Breast Ultrasound Images
von: Gheflati, Behnaz, et al.
Veröffentlicht: (2021)
von: Gheflati, Behnaz, et al.
Veröffentlicht: (2021)
IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models
von: Madinei, Parsa, et al.
Veröffentlicht: (2026)
von: Madinei, Parsa, et al.
Veröffentlicht: (2026)
Triplet-Watershed for Hyperspectral Image Classification
von: Challa, Aditya, et al.
Veröffentlicht: (2021)
von: Challa, Aditya, et al.
Veröffentlicht: (2021)
Improved EATFormer: A Vision Transformer for Medical Image Classification
von: Shisu, Yulong, et al.
Veröffentlicht: (2024)
von: Shisu, Yulong, et al.
Veröffentlicht: (2024)
Interpretable Vision Transformers in Image Classification via SVDA
von: Arampatzakis, Vasileios, et al.
Veröffentlicht: (2026)
von: Arampatzakis, Vasileios, et al.
Veröffentlicht: (2026)
Graph Attention Transformer Network for Multi-Label Image Classification
von: Yuan, Jin, et al.
Veröffentlicht: (2022)
von: Yuan, Jin, et al.
Veröffentlicht: (2022)
SaFiRe: Saccade-Fixation Reiteration with Mamba for Referring Image Segmentation
von: Mao, Zhenjie, et al.
Veröffentlicht: (2025)
von: Mao, Zhenjie, et al.
Veröffentlicht: (2025)
Embedded event based object detection with spiking neural network
von: Courtois, Jonathan, et al.
Veröffentlicht: (2024)
von: Courtois, Jonathan, et al.
Veröffentlicht: (2024)
Medical Image Classification with KAN-Integrated Transformers and Dilated Neighborhood Attention
von: Manzari, Omid Nejati, et al.
Veröffentlicht: (2025)
von: Manzari, Omid Nejati, et al.
Veröffentlicht: (2025)
Interpretable Image Classification with Adaptive Prototype-based Vision Transformers
von: Ma, Chiyu, et al.
Veröffentlicht: (2024)
von: Ma, Chiyu, et al.
Veröffentlicht: (2024)
Hierarchical Vision Transformer with Prototypes for Interpretable Medical Image Classification
von: Gallée, Luisa, et al.
Veröffentlicht: (2025)
von: Gallée, Luisa, et al.
Veröffentlicht: (2025)
Sensitive Image Classification by Vision Transformers
von: He, Hanxian, et al.
Veröffentlicht: (2024)
von: He, Hanxian, et al.
Veröffentlicht: (2024)
Representative Attention For Vision Transformers
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
Vision Transformers with Hierarchical Attention
von: Liu, Yun, et al.
Veröffentlicht: (2021)
von: Liu, Yun, et al.
Veröffentlicht: (2021)
A Synergistic CNN-Transformer Network with Pooling Attention Fusion for Hyperspectral Image Classification
von: Chen, Peng, et al.
Veröffentlicht: (2026)
von: Chen, Peng, et al.
Veröffentlicht: (2026)
A Novel Vision Transformer with Residual in Self-attention for Biomedical Image Classification
von: Sharma, Arun K., et al.
Veröffentlicht: (2023)
von: Sharma, Arun K., et al.
Veröffentlicht: (2023)
Federated Vision Transformer with Adaptive Focal Loss for Medical Image Classification
von: Zhao, Xinyuan, et al.
Veröffentlicht: (2026)
von: Zhao, Xinyuan, et al.
Veröffentlicht: (2026)
Fusion of Foundation and Vision Transformer Model Features for Dermatoscopic Image Classification
von: Mahbod, Amirreza, et al.
Veröffentlicht: (2025)
von: Mahbod, Amirreza, et al.
Veröffentlicht: (2025)
Class-Discriminative Attention Maps for Vision Transformers
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification
von: Kim, Ga-Eun, et al.
Veröffentlicht: (2023)
von: Kim, Ga-Eun, et al.
Veröffentlicht: (2023)
EVA: Bridging Performance and Human Alignment in Hard-Attention Vision Models for Image Classification
von: Pan, Pengcheng, et al.
Veröffentlicht: (2026)
von: Pan, Pengcheng, et al.
Veröffentlicht: (2026)
HAViT: Historical Attention Vision Transformer
von: Banik, Swarnendu, et al.
Veröffentlicht: (2026)
von: Banik, Swarnendu, et al.
Veröffentlicht: (2026)
Structured Initialization for Attention in Vision Transformers
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
Vision Transformers are Circulant Attention Learners
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
Multi-manifold Attention for Vision Transformers
von: Konstantinidis, Dimitrios, et al.
Veröffentlicht: (2022)
von: Konstantinidis, Dimitrios, et al.
Veröffentlicht: (2022)
Skin Cancer Detection utilizing Deep Learning: Classification of Skin Lesion Images using a Vision Transformer
von: Flosdorf, Carolin, et al.
Veröffentlicht: (2024)
von: Flosdorf, Carolin, et al.
Veröffentlicht: (2024)
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
von: Tosato, Lucrezia, et al.
Veröffentlicht: (2024)
TCSAFormer: Efficient Vision Transformer with Token Compression and Sparse Attention for Medical Image Segmentation
von: Xia, Zunhui, et al.
Veröffentlicht: (2025)
von: Xia, Zunhui, et al.
Veröffentlicht: (2025)
Multi-Modal Vision Transformers for Crop Mapping from Satellite Image Time Series
von: Follath, Theresa, et al.
Veröffentlicht: (2024)
von: Follath, Theresa, et al.
Veröffentlicht: (2024)
Polyline Path Masked Attention for Vision Transformer
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025) -
Foveated Retinotopy Improves Classification and Localization in Convolutional Neural Networks
von: Jérémie, Jean-Nicolas, et al.
Veröffentlicht: (2024) -
Saccadic Vision for Fine-Grained Visual Classification
von: Schmidt, Johann, et al.
Veröffentlicht: (2025) -
Spiking monocular event based 6D pose estimation for space application
von: Courtois, Jonathan, et al.
Veröffentlicht: (2025) -
Emergence of Fixational and Saccadic Movements in a Multi-Level Recurrent Attention Model for Vision
von: Pan, Pengcheng, et al.
Veröffentlicht: (2025)