Memba: Membrane-driven Parameter-Efficient Fine-Tuning for Mamba

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lee, Donghyun, Li, Yuhang, Yin, Ruokai, Xiao, Shiting, Panda, Priyadarshini
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912935749615616
author Lee, Donghyun
Li, Yuhang
Yin, Ruokai
Xiao, Shiting
Panda, Priyadarshini
author_facet Lee, Donghyun
Li, Yuhang
Yin, Ruokai
Xiao, Shiting
Panda, Priyadarshini
contents State Space Models (SSMs) have emerged as powerful alternatives to attention-based Transformers, with Mamba demonstrating impressive efficiency and scalability. As these models grow increasingly larger, the need for Parameter-Efficient Fine-Tuning (PEFT) methods becomes critical to adapt pre-trained Mamba to downstream tasks without prohibitive computational costs. However, previous approaches simply apply traditional Transformer-tailored PEFT methods without addressing the unique temporal processing dynamics of SSMs. To address this limitation, we propose Memba, a membrane-driven PEFT approach specifically designed for Mamba. Memba introduces Leaky Integrate Membrane (LIM) neurons as bio-inspired gating mechanisms that naturally accumulate membrane potentials over time, enhancing selective information retention. By strategically combining LIM neurons with Low-Rank Adaptations (LoRA) and cross-layer membrane transfer, our approach significantly improves Mamba's temporal modeling capabilities. Extensive experiments across language and vision tasks demonstrate that Memba achieves substantial improvements over existing PEFT methods. The code is available at https://github.com/Intelligent-Computing-Lab-Yale/Memba.
format Preprint
id arxiv_https___arxiv_org_abs_2506_18184
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Memba: Membrane-driven Parameter-Efficient Fine-Tuning for Mamba
Lee, Donghyun
Li, Yuhang
Yin, Ruokai
Xiao, Shiting
Panda, Priyadarshini
Machine Learning
State Space Models (SSMs) have emerged as powerful alternatives to attention-based Transformers, with Mamba demonstrating impressive efficiency and scalability. As these models grow increasingly larger, the need for Parameter-Efficient Fine-Tuning (PEFT) methods becomes critical to adapt pre-trained Mamba to downstream tasks without prohibitive computational costs. However, previous approaches simply apply traditional Transformer-tailored PEFT methods without addressing the unique temporal processing dynamics of SSMs. To address this limitation, we propose Memba, a membrane-driven PEFT approach specifically designed for Mamba. Memba introduces Leaky Integrate Membrane (LIM) neurons as bio-inspired gating mechanisms that naturally accumulate membrane potentials over time, enhancing selective information retention. By strategically combining LIM neurons with Low-Rank Adaptations (LoRA) and cross-layer membrane transfer, our approach significantly improves Mamba's temporal modeling capabilities. Extensive experiments across language and vision tasks demonstrate that Memba achieves substantial improvements over existing PEFT methods. The code is available at https://github.com/Intelligent-Computing-Lab-Yale/Memba.
title Memba: Membrane-driven Parameter-Efficient Fine-Tuning for Mamba
topic Machine Learning
url https://arxiv.org/abs/2506.18184