S2AFormer: Strip Self-Attention for Efficient Vision Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Guoan, Huang, Wenfeng, Jia, Wenjing, Li, Jiamao, Gao, Guangwei, Qi, Guo-Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
ReviveDiff: A Universal Diffusion Model for Restoring Images in Adverse Weather Conditions
von: Huang, Wenfeng, et al.
Veröffentlicht: (2024)
von: Huang, Wenfeng, et al.
Veröffentlicht: (2024)
RSGMamba: Reliability-Aware Self-Gated State Space Model for Multimodal Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2026)
von: Xu, Guoan, et al.
Veröffentlicht: (2026)
WaveSeg: Enhancing Segmentation Precision via High-Frequency Prior and Mamba-Driven Spectrum Decomposition
von: Xu, Guoan, et al.
Veröffentlicht: (2025)
von: Xu, Guoan, et al.
Veröffentlicht: (2025)
MacFormer: Semantic Segmentation with Fine Object Boundaries
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
HAFormer: Unleashing the Power of Hierarchy-Aware Features for Lightweight Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
Efficient Semantic Segmentation via Lightweight Multiple-Information Interaction Network
von: Qiu, Yangyang, et al.
Veröffentlicht: (2024)
von: Qiu, Yangyang, et al.
Veröffentlicht: (2024)
Versatile and Efficient Medical Image Super-Resolution Via Frequency-Gated Mamba
von: Huang, Wenfeng, et al.
Veröffentlicht: (2025)
von: Huang, Wenfeng, et al.
Veröffentlicht: (2025)
Vision Transformers are Circulant Attention Learners
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
JointSplat: Probabilistic Joint Flow-Depth Optimization for Sparse-View Gaussian Splatting
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Representative Attention For Vision Transformers
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
von: Li, Yuntong, et al.
Veröffentlicht: (2026)
Cross Paradigm Representation and Alignment Transformer for Image Deraining
von: Zou, Shun, et al.
Veröffentlicht: (2025)
von: Zou, Shun, et al.
Veröffentlicht: (2025)
Efficient Image Super-Resolution with Feature Interaction Weighted Hybrid Network
von: Li, Wenjie, et al.
Veröffentlicht: (2022)
von: Li, Wenjie, et al.
Veröffentlicht: (2022)
SAGS: Self-Adaptive Alias-Free Gaussian Splatting for Dynamic Surgical Endoscopic Reconstruction
von: Huang, Wenfeng, et al.
Veröffentlicht: (2025)
von: Huang, Wenfeng, et al.
Veröffentlicht: (2025)
FADPNet: Frequency-Aware Dual-Path Network for Face Super-Resolution
von: Xu, Siyu, et al.
Veröffentlicht: (2025)
von: Xu, Siyu, et al.
Veröffentlicht: (2025)
Self-Supervised Selective-Guided Diffusion Model for Old-Photo Face Restoration
von: Li, Wenjie, et al.
Veröffentlicht: (2025)
von: Li, Wenjie, et al.
Veröffentlicht: (2025)
Efficient Diffusion Transformer with Step-wise Dynamic Attention Mediators
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
MicroViT: A Vision Transformer with Low Complexity Self Attention for Edge Device
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
von: Setyawan, Novendra, et al.
Veröffentlicht: (2025)
Multiple-Exit Tuning: Towards Inference-Efficient Adaptation for Vision Transformer
von: Liu, Zheng, et al.
Veröffentlicht: (2024)
von: Liu, Zheng, et al.
Veröffentlicht: (2024)
EViT: An Eagle Vision Transformer with Bi-Fovea Self-Attention
von: Shi, Yulong, et al.
Veröffentlicht: (2023)
von: Shi, Yulong, et al.
Veröffentlicht: (2023)
Polyline Path Masked Attention for Vision Transformer
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
von: Zhao, Zhongchen, et al.
Veröffentlicht: (2025)
GeminiFusion: Efficient Pixel-wise Multimodal Fusion for Vision Transformer
von: Jia, Ding, et al.
Veröffentlicht: (2024)
von: Jia, Ding, et al.
Veröffentlicht: (2024)
Transformer-Progressive Mamba Network for Lightweight Image Super-Resolution
von: Guo, Sichen, et al.
Veröffentlicht: (2025)
von: Guo, Sichen, et al.
Veröffentlicht: (2025)
ToSA: Token Selective Attention for Efficient Vision Transformers
von: Singh, Manish Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Manish Kumar, et al.
Veröffentlicht: (2024)
Artifacts and Attention Sinks: Structured Approximations for Efficient Vision Transformers
von: Lu, Andrew, et al.
Veröffentlicht: (2025)
von: Lu, Andrew, et al.
Veröffentlicht: (2025)
Calibration Attention: Learning Reliability-Aware Representations for Vision Transformers
von: Liang, Wenhao, et al.
Veröffentlicht: (2025)
von: Liang, Wenhao, et al.
Veröffentlicht: (2025)
Structured Initialization for Attention in Vision Transformers
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
Improving Vision Transformers by Overlapping Heads in Multi-Head Self-Attention
von: Zhang, Tianxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Tianxiao, et al.
Veröffentlicht: (2024)
Parallel Cross Strip Attention Network for Single Image Dehazing
von: Tong, Lihan, et al.
Veröffentlicht: (2024)
von: Tong, Lihan, et al.
Veröffentlicht: (2024)
TCSAFormer: Efficient Vision Transformer with Token Compression and Sparse Attention for Medical Image Segmentation
von: Xia, Zunhui, et al.
Veröffentlicht: (2025)
von: Xia, Zunhui, et al.
Veröffentlicht: (2025)
StripDet: Strip Attention-Based Lightweight 3D Object Detection from Point Cloud
von: Wang, Weichao, et al.
Veröffentlicht: (2025)
von: Wang, Weichao, et al.
Veröffentlicht: (2025)
Vision Transformers with Hierarchical Attention
von: Liu, Yun, et al.
Veröffentlicht: (2021)
von: Liu, Yun, et al.
Veröffentlicht: (2021)
Lightweight Vision Transformer with Window and Spatial Attention for Food Image Classification
von: Gao, Xinle, et al.
Veröffentlicht: (2025)
von: Gao, Xinle, et al.
Veröffentlicht: (2025)
Attention-Guided Multi-scale Interaction Network for Face Super-Resolution
von: Wan, Xujie, et al.
Veröffentlicht: (2024)
von: Wan, Xujie, et al.
Veröffentlicht: (2024)
MambaMIC: An Efficient Baseline for Microscopic Image Classification with State Space Models
von: Zou, Shun, et al.
Veröffentlicht: (2024)
von: Zou, Shun, et al.
Veröffentlicht: (2024)
Efficient Self-supervised Vision Pretraining with Local Masked Reconstruction
von: Chen, Jun, et al.
Veröffentlicht: (2022)
von: Chen, Jun, et al.
Veröffentlicht: (2022)
Efficient Generation of Targeted and Transferable Adversarial Examples for Vision-Language Models Via Diffusion Models
von: Guo, Qi, et al.
Veröffentlicht: (2024)
von: Guo, Qi, et al.
Veröffentlicht: (2024)
EDIT: Enhancing Vision Transformers by Mitigating Attention Sink through an Encoder-Decoder Architecture
von: Feng, Wenfeng, et al.
Veröffentlicht: (2025)
von: Feng, Wenfeng, et al.
Veröffentlicht: (2025)
PADRe: A Unifying Polynomial Attention Drop-in Replacement for Efficient Vision Transformer
von: Letourneau, Pierre-David, et al.
Veröffentlicht: (2024)
von: Letourneau, Pierre-David, et al.
Veröffentlicht: (2024)
Castling-ViT: Compressing Self-Attention via Switching Towards Linear-Angular Attention at Vision Transformer Inference
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2024) -
ReviveDiff: A Universal Diffusion Model for Restoring Images in Adverse Weather Conditions
von: Huang, Wenfeng, et al.
Veröffentlicht: (2024) -
RSGMamba: Reliability-Aware Self-Gated State Space Model for Multimodal Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2026) -
WaveSeg: Enhancing Segmentation Precision via High-Frequency Prior and Mamba-Driven Spectrum Decomposition
von: Xu, Guoan, et al.
Veröffentlicht: (2025) -
MacFormer: Semantic Segmentation with Fine Object Boundaries
von: Xu, Guoan, et al.
Veröffentlicht: (2024)