SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph Attention for Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Venkatraman, Shravan, Walia, Jaskaran Singh, R, Joe Dhanith P |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
by: Khan, Md Ashik, et al.
Published: (2025)
by: Khan, Md Ashik, et al.
Published: (2025)
Attend, Distill, Detect: Attention-aware Entropy Distillation for Anomaly Detection
by: Jena, Sushovan, et al.
Published: (2024)
by: Jena, Sushovan, et al.
Published: (2024)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
by: Kambhatla, Akhila, et al.
Published: (2025)
by: Kambhatla, Akhila, et al.
Published: (2025)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
by: Alanazi, Ahmed, et al.
Published: (2025)
by: Alanazi, Ahmed, et al.
Published: (2025)
GIIM: Graph-based Learning of Inter- and Intra-view Dependencies for Multi-view Medical Image Diagnosis
by: Sam, Tran Bao, et al.
Published: (2026)
by: Sam, Tran Bao, et al.
Published: (2026)
MIPHEI-ViT: Multiplex Immunofluorescence Prediction from H&E Images using ViT Foundation Models
by: Balezo, Guillaume, et al.
Published: (2025)
by: Balezo, Guillaume, et al.
Published: (2025)
MB-DSMIL-CL-PL: Scalable Weakly Supervised Ovarian Cancer Subtype Classification and Localisation Using Contrastive and Prototype Learning with Frozen Patch Features
by: Jenkins, Marcus, et al.
Published: (2026)
by: Jenkins, Marcus, et al.
Published: (2026)
Self-Attention And Beyond the Infinite: Towards Linear Transformers with Infinite Self-Attention
by: Roffo, Giorgio, et al.
Published: (2026)
by: Roffo, Giorgio, et al.
Published: (2026)
Banana Ripeness Level Classification using a Simple CNN Model Trained with Real and Synthetic Datasets
by: Chuquimarca, Luis, et al.
Published: (2025)
by: Chuquimarca, Luis, et al.
Published: (2025)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
by: Viveiros, André G., et al.
Published: (2025)
by: Viveiros, André G., et al.
Published: (2025)
Deep Learning From Routine Histology Improves Risk Stratification for Biochemical Recurrence in Prostate Cancer
by: Grisi, Clément, et al.
Published: (2026)
by: Grisi, Clément, et al.
Published: (2026)
Detection of Intracranial Hemorrhage for Trauma Patients
by: Sanner, Antoine P., et al.
Published: (2024)
by: Sanner, Antoine P., et al.
Published: (2024)
Conterfactual Generative Zero-Shot Semantic Segmentation
by: Shen, Feihong, et al.
Published: (2021)
by: Shen, Feihong, et al.
Published: (2021)
Voxel Scene Graph for Intracranial Hemorrhage
by: Sanner, Antoine P., et al.
Published: (2024)
by: Sanner, Antoine P., et al.
Published: (2024)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
by: Jin, Hang, et al.
Published: (2025)
by: Jin, Hang, et al.
Published: (2025)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
by: Rashid, Muhammad, et al.
Published: (2026)
by: Rashid, Muhammad, et al.
Published: (2026)
SAFformer:Improving Spiking Transformer via Active Predictive Filtering
by: Xie, Zequan, et al.
Published: (2026)
by: Xie, Zequan, et al.
Published: (2026)
JVLGS: Joint Vision-Language Gas Leak Segmentation
by: Zhao, Xinlong, et al.
Published: (2025)
by: Zhao, Xinlong, et al.
Published: (2025)
A Novel Adaptive Hybrid Focal-Entropy Loss for Enhancing Diabetic Retinopathy Detection Using Convolutional Neural Networks
by: Malarvannan, Santhosh, et al.
Published: (2024)
by: Malarvannan, Santhosh, et al.
Published: (2024)
Learning Sign Language Representation using CNN LSTM, 3DCNN, CNN RNN LSTM and CCN TD
by: Louison, Nikita, et al.
Published: (2024)
by: Louison, Nikita, et al.
Published: (2024)
A Lightweight Multi-Module Fusion Approach for Korean Character Recognition
by: Park, Inho Jake, et al.
Published: (2025)
by: Park, Inho Jake, et al.
Published: (2025)
Attention Gathers, MLPs Compose: A Causal Analysis of an Action-Outcome Circuit in VideoViT
by: Chereddy, Sai V R
Published: (2026)
by: Chereddy, Sai V R
Published: (2026)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
Advancing Brain Tumor Segmentation via Attention-based 3D U-Net Architecture and Digital Image Processing
by: Gad, Eyad, et al.
Published: (2025)
by: Gad, Eyad, et al.
Published: (2025)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
by: Nauen, Tobias Christian, et al.
Published: (2024)
by: Nauen, Tobias Christian, et al.
Published: (2024)
GLL: A Differentiable Graph Learning Layer for Neural Networks
by: Brown, Jason, et al.
Published: (2024)
by: Brown, Jason, et al.
Published: (2024)
Parameter-efficient fine-tuning (PEFT) of Vision Foundation Models for Atypical Mitotic Figure Classification
by: Ramchandani, Lavish, et al.
Published: (2025)
by: Ramchandani, Lavish, et al.
Published: (2025)
Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
by: Kang, Xueyang, et al.
Published: (2026)
by: Kang, Xueyang, et al.
Published: (2026)
Exploiting Precision Mapping and Component-Specific Feature Enhancement for Breast Cancer Segmentation and Identification
by: V, Pandiyaraju, et al.
Published: (2024)
by: V, Pandiyaraju, et al.
Published: (2024)
HuMoCon: Concept Discovery for Human Motion Understanding
by: Fang, Qihang, et al.
Published: (2025)
by: Fang, Qihang, et al.
Published: (2025)
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
by: Chowdhury, Arindam, et al.
Published: (2025)
by: Chowdhury, Arindam, et al.
Published: (2025)
NFIG: Multi-Scale Autoregressive Image Generation via Frequency Ordering
by: Huang, Zhihao, et al.
Published: (2025)
by: Huang, Zhihao, et al.
Published: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
by: Rychkovskiy, Denis
Published: (2025)
by: Rychkovskiy, Denis
Published: (2025)
Which Transformer to Favor: A Comparative Analysis of Efficiency in Vision Transformers
by: Nauen, Tobias Christian, et al.
Published: (2023)
by: Nauen, Tobias Christian, et al.
Published: (2023)
Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible
by: Zhao, Lepeng, et al.
Published: (2026)
by: Zhao, Lepeng, et al.
Published: (2026)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
by: Qesaraku, Bjorna, et al.
Published: (2025)
by: Qesaraku, Bjorna, et al.
Published: (2025)
MRI Brain Tumor Detection with Computer Vision
by: Krolik, Jack, et al.
Published: (2025)
by: Krolik, Jack, et al.
Published: (2025)
On Memory: A comparison of memory mechanisms in world models
by: Laird, Eli J., et al.
Published: (2025)
by: Laird, Eli J., et al.
Published: (2025)
LRVS-Fashion: Extending Visual Search with Referring Instructions
by: Lepage, Simon, et al.
Published: (2023)
by: Lepage, Simon, et al.
Published: (2023)
E Pluribus Unum Interpretable Convolutional Neural Networks
by: Dimas, George, et al.
Published: (2022)
by: Dimas, George, et al.
Published: (2022)
Similar Items
-
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
by: Khan, Md Ashik, et al.
Published: (2025) -
Attend, Distill, Detect: Attention-aware Entropy Distillation for Anomaly Detection
by: Jena, Sushovan, et al.
Published: (2024) -
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
by: Kambhatla, Akhila, et al.
Published: (2025) -
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
by: Alanazi, Ahmed, et al.
Published: (2025) -
GIIM: Graph-based Learning of Inter- and Intra-view Dependencies for Multi-view Medical Image Diagnosis
by: Sam, Tran Bao, et al.
Published: (2026)