The Loupe: A Plug-and-Play Attention Module for Amplifying Discriminative Features in Vision Transformers
Fuente:
arXiv
Salvato in:
| Autore principale: | Sengodan, Naren |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Breast Cancer Histopathology Classification using CBAM-EfficientNetV2 with Transfer Learning
di: Sengodan, Naren
Pubblicazione: (2024)
di: Sengodan, Naren
Pubblicazione: (2024)
Class-Discriminative Attention Maps for Vision Transformers
di: Brocki, Lennart, et al.
Pubblicazione: (2023)
di: Brocki, Lennart, et al.
Pubblicazione: (2023)
Taming Score-Based Denoisers in ADMM: A Convergent Plug-and-Play Framework
di: Shrestha, Rajesh, et al.
Pubblicazione: (2026)
di: Shrestha, Rajesh, et al.
Pubblicazione: (2026)
HiGS: History-Guided Sampling for Plug-and-Play Enhancement of Diffusion Models
di: Sadat, Seyedmorteza, et al.
Pubblicazione: (2025)
di: Sadat, Seyedmorteza, et al.
Pubblicazione: (2025)
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
di: Ye, Xin, et al.
Pubblicazione: (2025)
di: Ye, Xin, et al.
Pubblicazione: (2025)
PromptLoop: Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model Alignment
di: Lee, Suhyeon, et al.
Pubblicazione: (2025)
di: Lee, Suhyeon, et al.
Pubblicazione: (2025)
Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs
di: Wang, Hao, et al.
Pubblicazione: (2026)
di: Wang, Hao, et al.
Pubblicazione: (2026)
FasterViT: Fast Vision Transformers with Hierarchical Attention
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
Enhancing Online Continual Learning with Plug-and-Play State Space Model and Class-Conditional Mixture of Discretization
di: Liu, Sihao, et al.
Pubblicazione: (2024)
di: Liu, Sihao, et al.
Pubblicazione: (2024)
Fairness-aware Vision Transformer via Debiased Self-Attention
di: Qiang, Yao, et al.
Pubblicazione: (2023)
di: Qiang, Yao, et al.
Pubblicazione: (2023)
Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models
di: Ryan, Yuriel, et al.
Pubblicazione: (2026)
di: Ryan, Yuriel, et al.
Pubblicazione: (2026)
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
di: Schaumlöffel, Timothy, et al.
Pubblicazione: (2023)
di: Schaumlöffel, Timothy, et al.
Pubblicazione: (2023)
Regularizing Neural Network Training via Identity-wise Discriminative Feature Suppression
di: Chapman, Avraham, et al.
Pubblicazione: (2022)
di: Chapman, Avraham, et al.
Pubblicazione: (2022)
GLoG-CSUnet: Enhancing Vision Transformers with Adaptable Radiomic Features for Medical Image Segmentation
di: Eghbali, Niloufar, et al.
Pubblicazione: (2025)
di: Eghbali, Niloufar, et al.
Pubblicazione: (2025)
LaVIDE: A Language-Vision Discriminator for Detecting Changes in Satellite Image with Map References
di: Jiang, Shuguo, et al.
Pubblicazione: (2024)
di: Jiang, Shuguo, et al.
Pubblicazione: (2024)
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
di: Colagrande, Alex, et al.
Pubblicazione: (2025)
di: Colagrande, Alex, et al.
Pubblicazione: (2025)
PCaM: A Progressive Focus Attention-Based Information Fusion Method for Improving Vision Transformer Domain Adaptation
di: Zang, Zelin, et al.
Pubblicazione: (2025)
di: Zang, Zelin, et al.
Pubblicazione: (2025)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
di: Ntrougkas, Mariano V., et al.
Pubblicazione: (2024)
di: Ntrougkas, Mariano V., et al.
Pubblicazione: (2024)
On the Surprising Effectiveness of Attention Transfer for Vision Transformers
di: Li, Alexander C., et al.
Pubblicazione: (2024)
di: Li, Alexander C., et al.
Pubblicazione: (2024)
Transforming Game Play: A Comparative Study of DCQN and DTQN Architectures in Reinforcement Learning
di: Stigall, William A.
Pubblicazione: (2024)
di: Stigall, William A.
Pubblicazione: (2024)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
di: Choi, Kanghyun, et al.
Pubblicazione: (2024)
di: Choi, Kanghyun, et al.
Pubblicazione: (2024)
AttentionDrop: A Novel Regularization Method for Transformer Models
di: Baig, Mirza Samad Ahmed, et al.
Pubblicazione: (2025)
di: Baig, Mirza Samad Ahmed, et al.
Pubblicazione: (2025)
Scratching Visual Transformer's Back with Uniform Attention
di: Hyeon-Woo, Nam, et al.
Pubblicazione: (2022)
di: Hyeon-Woo, Nam, et al.
Pubblicazione: (2022)
Evaluating the Posterior Sampling Ability of Plug&Play Diffusion Methods in Sparse-View CT
di: Moroy, Liam, et al.
Pubblicazione: (2024)
di: Moroy, Liam, et al.
Pubblicazione: (2024)
From $\mathcal{O}(n^{2})$ to $\mathcal{O}(n)$ Parameters: Quantum Self-Attention in Vision Transformers for Biomedical Image Classification
di: Boucher, Thomas, et al.
Pubblicazione: (2025)
di: Boucher, Thomas, et al.
Pubblicazione: (2025)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
di: Chiu, Pin-Yen, et al.
Pubblicazione: (2025)
di: Chiu, Pin-Yen, et al.
Pubblicazione: (2025)
Block-Recurrent Dynamics in Vision Transformers
di: Jacobs, Mozes, et al.
Pubblicazione: (2025)
di: Jacobs, Mozes, et al.
Pubblicazione: (2025)
Improving Interpretation Faithfulness for Vision Transformers
di: Hu, Lijie, et al.
Pubblicazione: (2023)
di: Hu, Lijie, et al.
Pubblicazione: (2023)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
di: Miyato, Takeru, et al.
Pubblicazione: (2023)
di: Miyato, Takeru, et al.
Pubblicazione: (2023)
Precipitation Nowcasting Using Diffusion Transformer with Causal Attention
di: Li, ChaoRong, et al.
Pubblicazione: (2024)
di: Li, ChaoRong, et al.
Pubblicazione: (2024)
On Diversity in Discriminative Neural Networks
di: Oubaha, Brahim, et al.
Pubblicazione: (2024)
di: Oubaha, Brahim, et al.
Pubblicazione: (2024)
Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching
di: Zou, Chang, et al.
Pubblicazione: (2024)
di: Zou, Chang, et al.
Pubblicazione: (2024)
VariViT: A Vision Transformer for Variable Image Sizes
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
A Survey of the Self Supervised Learning Mechanisms for Vision Transformers
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
VisPlay: Self-Evolving Vision-Language Models from Images
di: He, Yicheng, et al.
Pubblicazione: (2025)
di: He, Yicheng, et al.
Pubblicazione: (2025)
Mechanisms of Non-Monotonic Scaling in Vision Transformers
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Discovering Influential Neuron Path in Vision Transformers
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
Accelerating Vision Transformers with Adaptive Patch Sizes
di: Choudhury, Rohan, et al.
Pubblicazione: (2025)
di: Choudhury, Rohan, et al.
Pubblicazione: (2025)
Continual Adaptation of Vision Transformers for Federated Learning
di: Halbe, Shaunak, et al.
Pubblicazione: (2023)
di: Halbe, Shaunak, et al.
Pubblicazione: (2023)
DiffiT: Diffusion Vision Transformers for Image Generation
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Breast Cancer Histopathology Classification using CBAM-EfficientNetV2 with Transfer Learning
di: Sengodan, Naren
Pubblicazione: (2024) -
Class-Discriminative Attention Maps for Vision Transformers
di: Brocki, Lennart, et al.
Pubblicazione: (2023) -
Taming Score-Based Denoisers in ADMM: A Convergent Plug-and-Play Framework
di: Shrestha, Rajesh, et al.
Pubblicazione: (2026) -
HiGS: History-Guided Sampling for Plug-and-Play Enhancement of Diffusion Models
di: Sadat, Seyedmorteza, et al.
Pubblicazione: (2025) -
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
di: Ye, Xin, et al.
Pubblicazione: (2025)