EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Sanghyeok, Choi, Joonmyung, Kim, Hyunwoo J. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
by: Lee, Sanghyeok, et al.
Published: (2024)
by: Lee, Sanghyeok, et al.
Published: (2024)
vid-TLDR: Training Free Token merging for Light-weight Video Transformer
by: Choi, Joonmyung, et al.
Published: (2024)
by: Choi, Joonmyung, et al.
Published: (2024)
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
by: Choi, Joonmyung, et al.
Published: (2026)
by: Choi, Joonmyung, et al.
Published: (2026)
Representation Shift: Unifying Token Compression with FlashAttention
by: Choi, Joonmyung, et al.
Published: (2025)
by: Choi, Joonmyung, et al.
Published: (2025)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
by: Park, Jihwan, et al.
Published: (2025)
by: Park, Jihwan, et al.
Published: (2025)
VSSD: Vision Mamba with Non-Causal State Space Duality
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
TabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token Focusing
by: Kim, Jongha, et al.
Published: (2025)
by: Kim, Jongha, et al.
Published: (2025)
MoE-GRPO: Optimizing Mixture-of-Experts via Reinforcement Learning in Vision-Language Models
by: Ko, Dohwan, et al.
Published: (2026)
by: Ko, Dohwan, et al.
Published: (2026)
MambaMixer: Efficient Selective State Space Models with Dual Token and Channel Selection
by: Behrouz, Ali, et al.
Published: (2024)
by: Behrouz, Ali, et al.
Published: (2024)
PhysMamba: State Space Duality Model for Remote Physiological Measurement
by: Yan, Zhixin, et al.
Published: (2024)
by: Yan, Zhixin, et al.
Published: (2024)
VideoMamba: State Space Model for Efficient Video Understanding
by: Li, Kunchang, et al.
Published: (2024)
by: Li, Kunchang, et al.
Published: (2024)
Mamba-ST: State Space Model for Efficient Style Transfer
by: Botti, Filippo, et al.
Published: (2024)
by: Botti, Filippo, et al.
Published: (2024)
MambaVF: State Space Model for Efficient Video Fusion
by: Zhao, Zixiang, et al.
Published: (2026)
by: Zhao, Zixiang, et al.
Published: (2026)
Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models
by: Bai, Dewei, et al.
Published: (2026)
by: Bai, Dewei, et al.
Published: (2026)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
by: Zhu, Lianghui, et al.
Published: (2024)
by: Zhu, Lianghui, et al.
Published: (2024)
Efficient multi-view training for 3D Gaussian Splatting
by: Choi, Minhyuk, et al.
Published: (2025)
by: Choi, Minhyuk, et al.
Published: (2025)
MixerCSeg: An Efficient Mixer Architecture for Crack Segmentation via Decoupled Mamba Attention
by: Zhao, Zilong, et al.
Published: (2026)
by: Zhao, Zilong, et al.
Published: (2026)
ViM-UNet: Vision Mamba for Biomedical Segmentation
by: Archit, Anwai, et al.
Published: (2024)
by: Archit, Anwai, et al.
Published: (2024)
GroupMamba: Efficient Group-Based Visual State Space Model
by: Shaker, Abdelrahman, et al.
Published: (2024)
by: Shaker, Abdelrahman, et al.
Published: (2024)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
by: Cha, Juhan, et al.
Published: (2024)
by: Cha, Juhan, et al.
Published: (2024)
Towards Efficient Vision State Space Models via Token Merging
by: Park, Jinyoung, et al.
Published: (2025)
by: Park, Jinyoung, et al.
Published: (2025)
Gather-Scatter Mamba: Accelerating Propagation with Efficient State Space Model
by: Ko, Hyun-kyu, et al.
Published: (2025)
by: Ko, Hyun-kyu, et al.
Published: (2025)
ViM-VQ: Efficient Post-Training Vector Quantization for Visual Mamba
by: Deng, Juncan, et al.
Published: (2025)
by: Deng, Juncan, et al.
Published: (2025)
DensePercept-NCSSD: Vision Mamba towards Real-time Dense Visual Perception with Non-Causal State Space Duality
by: Anand, Tushar, et al.
Published: (2025)
by: Anand, Tushar, et al.
Published: (2025)
MambaMIC: An Efficient Baseline for Microscopic Image Classification with State Space Models
by: Zou, Shun, et al.
Published: (2024)
by: Zou, Shun, et al.
Published: (2024)
HTD-Mamba: Efficient Hyperspectral Target Detection with Pyramid State Space Model
by: Shen, Dunbin, et al.
Published: (2024)
by: Shen, Dunbin, et al.
Published: (2024)
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
by: Yun, Seokju, et al.
Published: (2024)
by: Yun, Seokju, et al.
Published: (2024)
Surface Vision Mamba: Leveraging Bidirectional State Space Model for Efficient Spherical Manifold Representation
by: He, Rongzhao, et al.
Published: (2025)
by: He, Rongzhao, et al.
Published: (2025)
OmniMamba: Efficient and Unified Multimodal Understanding and Generation via State Space Models
by: Zou, Jialv, et al.
Published: (2025)
by: Zou, Jialv, et al.
Published: (2025)
VideoMamba: Spatio-Temporal Selective State Space Model
by: Park, Jinyoung, et al.
Published: (2024)
by: Park, Jinyoung, et al.
Published: (2024)
TinyViM: Frequency Decoupling for Tiny Hybrid Vision Mamba
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
ChebMixer: Efficient Graph Representation Learning with MLP Mixer
by: Kui, Xiaoyan, et al.
Published: (2024)
by: Kui, Xiaoyan, et al.
Published: (2024)
MambaLoc: Efficient Camera Localisation via State Space Model
by: Wang, Jialu, et al.
Published: (2024)
by: Wang, Jialu, et al.
Published: (2024)
EAMamba: Efficient All-Around Vision State Space Model for Image Restoration
by: Lin, Yu-Cheng, et al.
Published: (2025)
by: Lin, Yu-Cheng, et al.
Published: (2025)
Directing Mamba to Complex Textures: An Efficient Texture-Aware State Space Model for Image Restoration
by: Peng, Long, et al.
Published: (2025)
by: Peng, Long, et al.
Published: (2025)
SF-Mamba: Rethinking State Space Model for Vision
by: Yoshimura, Masakazu, et al.
Published: (2026)
by: Yoshimura, Masakazu, et al.
Published: (2026)
MambaTalk: Efficient Holistic Gesture Synthesis with Selective State Space Models
by: Xu, Zunnan, et al.
Published: (2024)
by: Xu, Zunnan, et al.
Published: (2024)
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
Mamba2MIL: State Space Duality Based Multiple Instance Learning for Computational Pathology
by: Zhang, Yuqi, et al.
Published: (2024)
by: Zhang, Yuqi, et al.
Published: (2024)
Efficient Unstructured Pruning of Mamba State-Space Models for Resource-Constrained Environments
by: Shihab, Ibne Farabi, et al.
Published: (2025)
by: Shihab, Ibne Farabi, et al.
Published: (2025)
Similar Items
-
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
by: Lee, Sanghyeok, et al.
Published: (2024) -
vid-TLDR: Training Free Token merging for Light-weight Video Transformer
by: Choi, Joonmyung, et al.
Published: (2024) -
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
by: Choi, Joonmyung, et al.
Published: (2026) -
Representation Shift: Unifying Token Compression with FlashAttention
by: Choi, Joonmyung, et al.
Published: (2025) -
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
by: Park, Jihwan, et al.
Published: (2025)