ELSA: Exact Linear-Scan Attention for Fast and Memory-Light Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hsu, Chih-Chung, Ma, Xin-Di, Liao, Wo-Ting, Lee, Chia-Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MISS: Memory-efficient Instance Segmentation Framework By Visual Inductive Priors Flow Propagation
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
Augment Before Copy-Paste: Data and Memory Efficiency-Oriented Instance Segmentation Framework for Sport-scenes
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
CSAKD: Knowledge Distillation with Cross Self-Attention for Hyperspectral and Multispectral Image Fusion
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
DenseSR: Image Shadow Removal as Dense Prediction
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2025)
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2025)
PromptHSI: Universal Hyperspectral Image Restoration with Vision-Language Modulated Frequency Adaptation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2024)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2024)
UMCL: Unimodal-generated Multimodal Contrastive Learning for Cross-compression-rate Deepfake Detection
von: Lai, Ching-Yi, et al.
Veröffentlicht: (2025)
von: Lai, Ching-Yi, et al.
Veröffentlicht: (2025)
DRCT: Saving Image Super-resolution away from Information Bottleneck
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
WWE-UIE: A Wavelet & White Balance Efficient Network for Underwater Image Enhancement
von: Cheng, Ching-Heng, et al.
Veröffentlicht: (2025)
von: Cheng, Ching-Heng, et al.
Veröffentlicht: (2025)
Divide and Conquer: Grounding a Bleeding Areas in Gastrointestinal Image with Two-Stage Model
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2024)
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2024)
CANDLE: Illumination-Invariant Semantic Priors for Color Ambient Lighting Normalization
von: Jian, Rong-Lin, et al.
Veröffentlicht: (2026)
von: Jian, Rong-Lin, et al.
Veröffentlicht: (2026)
Progressive Alignment with VLM-LLM Feature to Augment Defect Classification for the ASE Dataset
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
ELSA: Exploiting Layer-wise N:M Sparsity for Vision Transformer Acceleration
von: Huang, Ning-Chi, et al.
Veröffentlicht: (2024)
von: Huang, Ning-Chi, et al.
Veröffentlicht: (2024)
HyFusion: Enhanced Reception Field Transformer for Hyperspectral Image Fusion
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
Fast-COS: A Fast One-Stage Object Detector Based on Reparameterized Attention Vision Transformer for Autonomous Driving
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
OCR is All you need: Importing Multi-Modality into Image-based Defect Detection System
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
Enhancing Learnable Descriptive Convolutional Vision Transformer for Face Anti-Spoofing
von: Huanga, Pei-Kai, et al.
Veröffentlicht: (2025)
von: Huanga, Pei-Kai, et al.
Veröffentlicht: (2025)
MicroViT: A Vision Transformer with Low Complexity Self Attention for Edge Device
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
Towards Robust DeepFake Detection under Unstable Face Sequences: Adaptive Sparse Graph Embedding with Order-Free Representation and Explicit Laplacian Spectral Prior
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2025)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2025)
Robust Hyperspectral Image Panshapring via Sparse Spatial-Spectral Representation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
FaceLiVT: Face Recognition using Linear Vision Transformer with Structural Reparameterization For Mobile Device
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection -- after competition results
von: Qiu, Bo-Cheng, et al.
Veröffentlicht: (2026)
von: Qiu, Bo-Cheng, et al.
Veröffentlicht: (2026)
PolaFormer: Polarity-aware Linear Attention for Vision Transformers
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
HSSDCT: Factorized Spatial-Spectral Correlation for Hyperspectral Image Fusion
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
The Linear Attention Resurrection in Vision Transformer
von: Zheng, Chuanyang
Veröffentlicht: (2025)
von: Zheng, Chuanyang
Veröffentlicht: (2025)
A Hierarchical Slice Attention Network for Appendicitis Classification in 3D CT Scans
von: Huang, Chia-Wen, et al.
Veröffentlicht: (2025)
von: Huang, Chia-Wen, et al.
Veröffentlicht: (2025)
Partial Ring Scan: Revisiting Scan Order in Vision State Space Models
von: Hsieh, Yi-Kuan, et al.
Veröffentlicht: (2026)
von: Hsieh, Yi-Kuan, et al.
Veröffentlicht: (2026)
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
GTATrack: Winner Solution to SoccerTrack 2025 with Deep-EIoU and Global Tracklet Association
von: Jian, Rong-Lin, et al.
Veröffentlicht: (2026)
von: Jian, Rong-Lin, et al.
Veröffentlicht: (2026)
Vision Transformer with Sparse Scan Prior
von: Zhang, Yuguang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuguang, et al.
Veröffentlicht: (2024)
GRACE: Graph-Regularized Attentive Convolutional Entanglement with Laplacian Smoothing for Robust DeepFake Video Detection
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
PhaSR: Generalized Image Shadow Removal with Physically Aligned Priors
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
Real-Time Compressed Sensing for Joint Hyperspectral Image Transmission and Restoration for CubeSat
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
ChangeDINO: DINOv3-Driven Building Change Detection in Optical Remote Sensing Imagery
von: Cheng, Ching-Heng, et al.
Veröffentlicht: (2025)
von: Cheng, Ching-Heng, et al.
Veröffentlicht: (2025)
Artifacts and Attention Sinks: Structured Approximations for Efficient Vision Transformers
von: Lu, Andrew, et al.
Veröffentlicht: (2025)
von: Lu, Andrew, et al.
Veröffentlicht: (2025)
Taming Domain Shift in Multi-source CT-Scan Classification via Input-Space Standardization
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
Castling-ViT: Compressing Self-Attention via Switching Towards Linear-Angular Attention at Vision Transformer Inference
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
Reasoning-VLA: A Fast and General Vision-Language-Action Reasoning Model for Autonomous Driving
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
Self-supervised Fusarium Head Blight Detection with Hyperspectral Image and Feature Mining
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2024)
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2024)
ReflexSplit: Single Image Reflection Separation via Layer Fusion-Separation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2026)
Representative Attention For Vision Transformers
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MISS: Memory-efficient Instance Segmentation Framework By Visual Inductive Priors Flow Propagation
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024) -
Augment Before Copy-Paste: Data and Memory Efficiency-Oriented Instance Segmentation Framework for Sport-scenes
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024) -
CSAKD: Knowledge Distillation with Cross Self-Attention for Hyperspectral and Multispectral Image Fusion
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024) -
DenseSR: Image Shadow Removal as Dense Prediction
von: Lin, Yu-Fan, et al.
Veröffentlicht: (2025) -
PromptHSI: Universal Hyperspectral Image Restoration with Vision-Language Modulated Frequency Adaptation
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2024)