STEAM: Squeeze and Transform Enhanced Attention Module
Fuente:
arXiv
Guardado en:
| Autores principales: | Sabharwal, Rishabh, B, Ram Samarth B, Rathore, Parikshit Singh, Rathore, Punit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DeepVAT: A Self-Supervised Technique for Cluster Assessment in Image Datasets
por: Mazumder, Alokendu, et al.
Publicado: (2023)
por: Mazumder, Alokendu, et al.
Publicado: (2023)
Generating Part-Based Global Explanations Via Correspondence
por: Rathore, Kunal, et al.
Publicado: (2025)
por: Rathore, Kunal, et al.
Publicado: (2025)
Echo-DND: A dual noise diffusion model for robust and precise left ventricle segmentation in echocardiography
por: Rahman, Abdur, et al.
Publicado: (2025)
por: Rahman, Abdur, et al.
Publicado: (2025)
HIDISC: A Hyperbolic Framework for Domain Generalization with Generalized Category Discovery
por: Rathore, Vaibhav, et al.
Publicado: (2025)
por: Rathore, Vaibhav, et al.
Publicado: (2025)
Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition
por: Chittersu, Raghu Vamsi, et al.
Publicado: (2025)
por: Chittersu, Raghu Vamsi, et al.
Publicado: (2025)
HipyrNet: Hypernet-Guided Feature Pyramid network for mixed-exposure correction
por: Rathore, Shaurya Singh, et al.
Publicado: (2025)
por: Rathore, Shaurya Singh, et al.
Publicado: (2025)
GatedLexiconNet: A Comprehensive End-to-End Handwritten Paragraph Text Recognition System
por: Kumari, Lalita, et al.
Publicado: (2024)
por: Kumari, Lalita, et al.
Publicado: (2024)
HAViT: Historical Attention Vision Transformer
por: Banik, Swarnendu, et al.
Publicado: (2026)
por: Banik, Swarnendu, et al.
Publicado: (2026)
When Domain Generalization meets Generalized Category Discovery: An Adaptive Task-Arithmetic Driven Approach
por: Rathore, Vaibhav, et al.
Publicado: (2025)
por: Rathore, Vaibhav, et al.
Publicado: (2025)
FOCUS: Bridging Fine-Grained Recognition and Open-World Discovery across Domains
por: Rathore, Vaibhav, et al.
Publicado: (2026)
por: Rathore, Vaibhav, et al.
Publicado: (2026)
DEAL-YOLO: Drone-based Efficient Animal Localization using YOLO
por: Naidu, Aditya Prashant, et al.
Publicado: (2025)
por: Naidu, Aditya Prashant, et al.
Publicado: (2025)
BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities
por: Sharma, Akash, et al.
Publicado: (2026)
por: Sharma, Akash, et al.
Publicado: (2026)
MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware
por: Palanisamy, Senthil, et al.
Publicado: (2026)
por: Palanisamy, Senthil, et al.
Publicado: (2026)
Interpretable Plant Leaf Disease Detection Using Attention-Enhanced CNN
por: Singh, Balram, et al.
Publicado: (2025)
por: Singh, Balram, et al.
Publicado: (2025)
The Urban Vision Hackathon Dataset and Models: Towards Image Annotations and Accurate Vision Models for Indian Traffic
por: Sharma, Akash, et al.
Publicado: (2025)
por: Sharma, Akash, et al.
Publicado: (2025)
EfficientSign: An Attention-Enhanced Lightweight Architecture for Indian Sign Language Recognition
por: Gupta, Rishabh, et al.
Publicado: (2026)
por: Gupta, Rishabh, et al.
Publicado: (2026)
Squeezed Diffusion Models
por: Singh, Jyotirmai, et al.
Publicado: (2025)
por: Singh, Jyotirmai, et al.
Publicado: (2025)
BuckTales : A multi-UAV dataset for multi-object tracking and re-identification of wild antelopes
por: Naik, Hemal, et al.
Publicado: (2024)
por: Naik, Hemal, et al.
Publicado: (2024)
Deep Clustering with Associative Memories
por: Saha, Bishwajit, et al.
Publicado: (2026)
por: Saha, Bishwajit, et al.
Publicado: (2026)
Enhancing Vehicle Make and Model Recognition with 3D Attention Modules
por: Semiromizadeh, Narges, et al.
Publicado: (2025)
por: Semiromizadeh, Narges, et al.
Publicado: (2025)
Efficient Epistemic Uncertainty Estimation in Cerebrovascular Segmentation
por: Rathore, Omini, et al.
Publicado: (2025)
por: Rathore, Omini, et al.
Publicado: (2025)
SeaFormer++: Squeeze-enhanced Axial Transformer for Mobile Visual Recognition
por: Wan, Qiang, et al.
Publicado: (2023)
por: Wan, Qiang, et al.
Publicado: (2023)
DSXFormer: Dual-Pooling Spectral Squeeze-Expansion and Dynamic Context Attention Transformer for Hyperspectral Image Classification
por: Ullah, Farhan, et al.
Publicado: (2026)
por: Ullah, Farhan, et al.
Publicado: (2026)
A Visual-Analytical Approach for Automatic Detection of Cyclonic Events in Satellite Observations
por: Agrawal, Akash, et al.
Publicado: (2024)
por: Agrawal, Akash, et al.
Publicado: (2024)
AbracADDbra: Touch-Guided Object Addition by Decoupling Placement and Editing Subtasks
por: Swami, Kunal, et al.
Publicado: (2026)
por: Swami, Kunal, et al.
Publicado: (2026)
Review and Evaluation of Point-Cloud based Leaf Surface Reconstruction Methods for Agricultural Applications
por: Ahmed, Arif, et al.
Publicado: (2026)
por: Ahmed, Arif, et al.
Publicado: (2026)
Attention Based Encoder Decoder Model for Video Captioning in Nepali (2023)
por: Parajuli, Kabita, et al.
Publicado: (2023)
por: Parajuli, Kabita, et al.
Publicado: (2023)
Towards Blind and Low-Vision Accessibility of Lightweight VLMs and Custom LLM-Evals
por: Baghel, Shruti Singh, et al.
Publicado: (2025)
por: Baghel, Shruti Singh, et al.
Publicado: (2025)
Transformer-based Clipped Contrastive Quantization Learning for Unsupervised Image Retrieval
por: Dubey, Ayush, et al.
Publicado: (2024)
por: Dubey, Ayush, et al.
Publicado: (2024)
There is More to Attention: Statistical Filtering Enhances Explanations in Vision Transformers
por: Ayyar, Meghna P, et al.
Publicado: (2025)
por: Ayyar, Meghna P, et al.
Publicado: (2025)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
por: Mishra, Samarth, et al.
Publicado: (2025)
por: Mishra, Samarth, et al.
Publicado: (2025)
ToSA: Token Selective Attention for Efficient Vision Transformers
por: Singh, Manish Kumar, et al.
Publicado: (2024)
por: Singh, Manish Kumar, et al.
Publicado: (2024)
Attention (as Discrete-Time Markov) Chains
por: Erel, Yotam, et al.
Publicado: (2025)
por: Erel, Yotam, et al.
Publicado: (2025)
SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers
por: Rajabi, Javad, et al.
Publicado: (2026)
por: Rajabi, Javad, et al.
Publicado: (2026)
HIRE: Lightweight High-Resolution Image Feature Enrichment for Multimodal LLMs
por: SR, Nikitha, et al.
Publicado: (2025)
por: SR, Nikitha, et al.
Publicado: (2025)
ReViT: Enhancing Vision Transformers Feature Diversity with Attention Residual Connections
por: Diko, Anxhelo, et al.
Publicado: (2024)
por: Diko, Anxhelo, et al.
Publicado: (2024)
HINT: High-quality INPainting Transformer with Mask-Aware Encoding and Enhanced Attention
por: Chen, Shuang, et al.
Publicado: (2024)
por: Chen, Shuang, et al.
Publicado: (2024)
Align-cDAE: Alzheimer's Disease Progression Modeling with Attention-Aligned Conditional Diffusion Auto-Encoder
por: Das, Ayantika, et al.
Publicado: (2026)
por: Das, Ayantika, et al.
Publicado: (2026)
Influence of Geometry, Class Imbalance and Alignment on Reconstruction Accuracy -- A Micro-CT Phantom-Based Evaluation
por: M, Avinash Kumar K, et al.
Publicado: (2026)
por: M, Avinash Kumar K, et al.
Publicado: (2026)
Scalable Methods for Brick Kiln Detection and Compliance Monitoring from Satellite Imagery: A Deployment Case Study in India
por: Mondal, Rishabh, et al.
Publicado: (2024)
por: Mondal, Rishabh, et al.
Publicado: (2024)
Ejemplares similares
-
DeepVAT: A Self-Supervised Technique for Cluster Assessment in Image Datasets
por: Mazumder, Alokendu, et al.
Publicado: (2023) -
Generating Part-Based Global Explanations Via Correspondence
por: Rathore, Kunal, et al.
Publicado: (2025) -
Echo-DND: A dual noise diffusion model for robust and precise left ventricle segmentation in echocardiography
por: Rahman, Abdur, et al.
Publicado: (2025) -
HIDISC: A Hyperbolic Framework for Domain Generalization with Generalized Category Discovery
por: Rathore, Vaibhav, et al.
Publicado: (2025) -
Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition
por: Chittersu, Raghu Vamsi, et al.
Publicado: (2025)