vGamba: Attentive State Space Bottleneck for efficient Long-range Dependencies in Visual Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Haruna, Yunusa, Lawan, Adamu, Muhammad, Shamsuddeen Hassan, Zhang, Jiaquan, Zhang, Chaoning |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias Redistribution in Visual Machine Unlearning: Does Forgetting One Group Harm Another?
by: Haruna, Yunusa, et al.
Published: (2026)
by: Haruna, Yunusa, et al.
Published: (2026)
iiANET: Inception Inspired Attention Hybrid Network for efficient Long-Range Dependency
by: Yunusa, Haruna, et al.
Published: (2024)
by: Yunusa, Haruna, et al.
Published: (2024)
Enhancing Long-Range Dependency with State Space Model and Kolmogorov-Arnold Networks for Aspect-Based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2024)
by: Lawan, Adamu, et al.
Published: (2024)
KonvLiNA: Integrating Kolmogorov-Arnold Network with Linear Nyström Attention for feature fusion in Crop Field Detection
by: Yunusa, Haruna, et al.
Published: (2024)
by: Yunusa, Haruna, et al.
Published: (2024)
GateMABSA: Aspect-Image Gated Fusion for Multimodal Aspect-based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2025)
by: Lawan, Adamu, et al.
Published: (2025)
SaRPFF: A Self-Attention with Register-based Pyramid Feature Fusion module for enhanced RLD detection
by: Haruna, Yunusa, et al.
Published: (2024)
by: Haruna, Yunusa, et al.
Published: (2024)
Exploring the Synergies of Hybrid CNNs and ViTs Architectures for Computer Vision: A survey
by: Yunusa, Haruna, et al.
Published: (2024)
by: Yunusa, Haruna, et al.
Published: (2024)
AF-MAT: Aspect-aware Flip-and-Fuse xLSTM for Aspect-based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2025)
by: Lawan, Adamu, et al.
Published: (2025)
Amplifying Aspect-Sentence Awareness: A Novel Approach for Aspect-Based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2024)
by: Lawan, Adamu, et al.
Published: (2024)
DualKanbaFormer: An Efficient Selective Sparse Framework for Multimodal Aspect-based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2024)
by: Lawan, Adamu, et al.
Published: (2024)
M2ANET: Mobile Malaria Attention Network for efficient classification of plasmodium parasites in blood cells
by: Ali, Salam Ahmed, et al.
Published: (2024)
by: Ali, Salam Ahmed, et al.
Published: (2024)
Mixture of Balanced Information Bottlenecks for Long-Tailed Visual Recognition
by: Lan, Yifan, et al.
Published: (2025)
by: Lan, Yifan, et al.
Published: (2025)
Spatial Information Bottleneck for Interpretable Visual Recognition
by: Shu, Kaixiang, et al.
Published: (2025)
by: Shu, Kaixiang, et al.
Published: (2025)
Mamba-Adaptor: State Space Model Adaptor for Visual Recognition
by: Xie, Fei, et al.
Published: (2025)
by: Xie, Fei, et al.
Published: (2025)
MambaIRv2: Attentive State Space Restoration
by: Guo, Hang, et al.
Published: (2024)
by: Guo, Hang, et al.
Published: (2024)
Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction
by: Shen, Qiuhong, et al.
Published: (2024)
by: Shen, Qiuhong, et al.
Published: (2024)
Fast SAM2 with Text-Driven Token Pruning
by: Mandal, Avilasha, et al.
Published: (2025)
by: Mandal, Avilasha, et al.
Published: (2025)
Toward Next-generation Medical Vision Backbones: Modeling Finer-grained Long-range Visual Dependency
by: Meng, Mingyuan
Published: (2025)
by: Meng, Mingyuan
Published: (2025)
Black-box Targeted Adversarial Attack on Segment Anything (SAM)
by: Zheng, Sheng, et al.
Published: (2023)
by: Zheng, Sheng, et al.
Published: (2023)
LLaVA-FA: Learning Fourier Approximation for Compressing Large Multimodal Models
by: Zheng, Pengcheng, et al.
Published: (2026)
by: Zheng, Pengcheng, et al.
Published: (2026)
Long-Tailed Continual Learning For Visual Food Recognition
by: He, Jiangpeng, et al.
Published: (2023)
by: He, Jiangpeng, et al.
Published: (2023)
A2Mamba: Attention-augmented State Space Models for Visual Recognition
by: Lou, Meng, et al.
Published: (2025)
by: Lou, Meng, et al.
Published: (2025)
Holistic Surgical Phase Recognition with Hierarchical Input Dependent State Space Models
by: Wu, Haoyang, et al.
Published: (2025)
by: Wu, Haoyang, et al.
Published: (2025)
TRACE: Thermal Recognition Attentive-Framework for CO2 Emissions from Livestock
by: Islam, Taminul, et al.
Published: (2026)
by: Islam, Taminul, et al.
Published: (2026)
Improving Visual Prompt Tuning by Gaussian Neighborhood Minimization for Long-Tailed Visual Recognition
by: Li, Mengke, et al.
Published: (2024)
by: Li, Mengke, et al.
Published: (2024)
Semantic-guided Fine-tuning of Foundation Model for Long-tailed Visual Recognition
by: Peng, Yufei, et al.
Published: (2025)
by: Peng, Yufei, et al.
Published: (2025)
MV-GMN: State Space Model for Multi-View Action Recognition
by: Lin, Yuhui, et al.
Published: (2025)
by: Lin, Yuhui, et al.
Published: (2025)
PASS: Path-selective State Space Model for Event-based Recognition
by: Zhou, Jiazhou, et al.
Published: (2024)
by: Zhou, Jiazhou, et al.
Published: (2024)
Representation Learning for Compressed Video Action Recognition via Attentive Cross-modal Interaction with Motion Enhancement
by: Li, Bing, et al.
Published: (2022)
by: Li, Bing, et al.
Published: (2022)
D$^{2}$-VPR: A Parameter-efficient Visual-foundation-model-based Visual Place Recognition Method via Knowledge Distillation and Deformable Aggregation
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
Spatial-Mamba: Effective Visual State Space Models via Structure-aware State Fusion
by: Xiao, Chaodong, et al.
Published: (2024)
by: Xiao, Chaodong, et al.
Published: (2024)
Deformable Attentive Visual Enhancement for Referring Segmentation Using Vision-Language Model
by: Dalaq, Alaa, et al.
Published: (2025)
by: Dalaq, Alaa, et al.
Published: (2025)
PCBEAR: Pose Concept Bottleneck for Explainable Action Recognition
by: Lee, Jongseo, et al.
Published: (2025)
by: Lee, Jongseo, et al.
Published: (2025)
Long-Context State-Space Video World Models
by: Po, Ryan, et al.
Published: (2025)
by: Po, Ryan, et al.
Published: (2025)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024)
by: Zhang, Zaiwei, et al.
Published: (2024)
Scalable Visual State Space Model with Fractal Scanning
by: Tang, Lv, et al.
Published: (2024)
by: Tang, Lv, et al.
Published: (2024)
DefMamba: Deformable Visual State Space Model
by: Liu, Leiye, et al.
Published: (2025)
by: Liu, Leiye, et al.
Published: (2025)
Learning Pyramid-structured Long-range Dependencies for 3D Human Pose Estimation
by: Wei, Mingjie, et al.
Published: (2025)
by: Wei, Mingjie, et al.
Published: (2025)
VMambaCC: A Visual State Space Model for Crowd Counting
by: Ma, Hao-Yuan, et al.
Published: (2024)
by: Ma, Hao-Yuan, et al.
Published: (2024)
Regressing Transformers for Data-efficient Visual Place Recognition
by: Leyva-Vallina, María, et al.
Published: (2024)
by: Leyva-Vallina, María, et al.
Published: (2024)
Similar Items
-
Bias Redistribution in Visual Machine Unlearning: Does Forgetting One Group Harm Another?
by: Haruna, Yunusa, et al.
Published: (2026) -
iiANET: Inception Inspired Attention Hybrid Network for efficient Long-Range Dependency
by: Yunusa, Haruna, et al.
Published: (2024) -
Enhancing Long-Range Dependency with State Space Model and Kolmogorov-Arnold Networks for Aspect-Based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2024) -
KonvLiNA: Integrating Kolmogorov-Arnold Network with Linear Nyström Attention for feature fusion in Crop Field Detection
by: Yunusa, Haruna, et al.
Published: (2024) -
GateMABSA: Aspect-Image Gated Fusion for Multimodal Aspect-based Sentiment Analysis
by: Lawan, Adamu, et al.
Published: (2025)