TransMamba: Fast Universal Architecture Adaption from Transformers to Mamba
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xiuwei, Hu, Wentao, Dong, Xiao, Lin, Sihao, Chen, Zisheng, Cao, Meng, Zhuang, Yina, Han, Jianhua, Xu, Hang, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TransMamba: A Sequence-Level Hybrid Transformer-Mamba Language Model
by: Li, Yixing, et al.
Published: (2025)
by: Li, Yixing, et al.
Published: (2025)
C2-Evo: Co-Evolving Multimodal Data and Model for Self-Improving Reasoning
by: Chen, Xiuwei, et al.
Published: (2025)
by: Chen, Xiuwei, et al.
Published: (2025)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
by: Chen, Zisheng, et al.
Published: (2025)
by: Chen, Zisheng, et al.
Published: (2025)
GeoMamba: A Geometry-driven MambaVision Framework and Dataset for Fine-grained Optical-SAR Object Retrieval
by: Fang, Tiantong, et al.
Published: (2026)
by: Fang, Tiantong, et al.
Published: (2026)
AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
by: Zhang, Likui, et al.
Published: (2026)
by: Zhang, Likui, et al.
Published: (2026)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
by: Chen, Wenchao, et al.
Published: (2024)
by: Chen, Wenchao, et al.
Published: (2024)
InfoMamba: An Attention-Free Hybrid Mamba-Transformer Model
by: Wang, Youjin, et al.
Published: (2026)
by: Wang, Youjin, et al.
Published: (2026)
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
by: Tang, Haoran, et al.
Published: (2024)
by: Tang, Haoran, et al.
Published: (2024)
VAMamba: An Efficient Visual Adaptive Mamba for Image Restoration
by: Hu, Han, et al.
Published: (2025)
by: Hu, Han, et al.
Published: (2025)
OccluDentNet : Self‐Supervised 3D Tooth Segmentation via MambaTrans Architecture
by: Hong‐an Li, et al.
Published: (2026)
by: Hong‐an Li, et al.
Published: (2026)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
by: Ren, Weiming, et al.
Published: (2025)
by: Ren, Weiming, et al.
Published: (2025)
MedSegMamba: 3D CNN-Mamba Hybrid Architecture for Brain Segmentation
by: Cao, Aaron, et al.
Published: (2024)
by: Cao, Aaron, et al.
Published: (2024)
Decision Mamba Architectures
by: Correia, André, et al.
Published: (2024)
by: Correia, André, et al.
Published: (2024)
MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation
by: Du, Juntian, et al.
Published: (2025)
by: Du, Juntian, et al.
Published: (2025)
MambaVesselNet++: A Hybrid CNN-Mamba Architecture for Medical Image Segmentation
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
MambaGlue: Fast and Robust Local Feature Matching With Mamba
by: Ryoo, Kihwan, et al.
Published: (2025)
by: Ryoo, Kihwan, et al.
Published: (2025)
Laplace-Mamba: Laplace Frequency Prior-Guided Mamba-CNN Fusion Network for Image Dehazing
by: Wang, Yongzhen, et al.
Published: (2025)
by: Wang, Yongzhen, et al.
Published: (2025)
InceptionMamba: An Efficient Hybrid Network with Large Band Convolution and Bottleneck Mamba
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
InfiniMotion: Mamba Boosts Memory in Transformer for Arbitrary Long Motion Generation
by: Zhang, Zeyu, et al.
Published: (2024)
by: Zhang, Zeyu, et al.
Published: (2024)
Mamba Modulation: On the Length Generalization of Mamba
by: Lu, Peng, et al.
Published: (2025)
by: Lu, Peng, et al.
Published: (2025)
MedMamba: Recasting Mamba for Medical Time Series Classification
by: He, ZhengXiao, et al.
Published: (2026)
by: He, ZhengXiao, et al.
Published: (2026)
TextMamba: Scene Text Detector with Mamba
by: Zhao, Qiyan, et al.
Published: (2025)
by: Zhao, Qiyan, et al.
Published: (2025)
MambaVision: A Hybrid Mamba-Transformer Vision Backbone
by: Hatamizadeh, Ali, et al.
Published: (2024)
by: Hatamizadeh, Ali, et al.
Published: (2024)
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods
by: Xu, Zukang, et al.
Published: (2025)
by: Xu, Zukang, et al.
Published: (2025)
Hazy Pedestrian Trajectory Prediction via Physical Priors and Graph-Mamba
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
SMILES-Mamba: Chemical Mamba Foundation Models for Drug ADMET Prediction
by: Xu, Bohao, et al.
Published: (2024)
by: Xu, Bohao, et al.
Published: (2024)
MambaNet: Mamba-assisted Channel Estimation Neural Network With Attention Mechanism
by: Luan, Dianxin, et al.
Published: (2026)
by: Luan, Dianxin, et al.
Published: (2026)
MobileMamba: Lightweight Multi-Receptive Visual Mamba Network
by: He, Haoyang, et al.
Published: (2024)
by: He, Haoyang, et al.
Published: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
by: Wang, Junyu, et al.
Published: (2024)
by: Wang, Junyu, et al.
Published: (2024)
MxT: Mamba x Transformer for Image Inpainting
by: Chen, Shuang, et al.
Published: (2024)
by: Chen, Shuang, et al.
Published: (2024)
RankMamba: Benchmarking Mamba's Document Ranking Performance in the Era of Transformers
by: Xu, Zhichao
Published: (2024)
by: Xu, Zhichao
Published: (2024)
GTR-Mamba: Geometry-to-Tangent Routing Mamba for Hyperbolic POI Recommendation
by: Li, Zhuoxuan, et al.
Published: (2025)
by: Li, Zhuoxuan, et al.
Published: (2025)
Bi-Mamba+: Bidirectional Mamba for Time Series Forecasting
by: Liang, Aobo, et al.
Published: (2024)
by: Liang, Aobo, et al.
Published: (2024)
Hybrid Transformer-Mamba Architecture for Weakly Supervised Volumetric Medical Segmentation
by: Lyu, Yiheng, et al.
Published: (2025)
by: Lyu, Yiheng, et al.
Published: (2025)
MambaSCI: Efficient Mamba-UNet for Quad-Bayer Patterned Video Snapshot Compressive Imaging
by: Pan, Zhenghao, et al.
Published: (2024)
by: Pan, Zhenghao, et al.
Published: (2024)
TC-BiMamba: Trans-Chunk bidirectionally within BiMamba for unified streaming and non-streaming ASR
by: She, Qingshun, et al.
Published: (2026)
by: She, Qingshun, et al.
Published: (2026)
MM-DETR: An Efficient Multimodal Detection Transformer with Mamba-Driven Dual-Granularity Fusion and Frequency-Aware Modality Adapters
by: Han, Jianhong, et al.
Published: (2025)
by: Han, Jianhong, et al.
Published: (2025)
Mamba4KT:An Efficient and Effective Mamba-based Knowledge Tracing Model
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
FastMamba: A High-Speed and Efficient Mamba Accelerator on FPGA with Accurate Quantization
by: Wang, Aotao, et al.
Published: (2025)
by: Wang, Aotao, et al.
Published: (2025)
PhysMamba: Efficient Remote Physiological Measurement with SlowFast Temporal Difference Mamba
by: Luo, Chaoqi, et al.
Published: (2024)
by: Luo, Chaoqi, et al.
Published: (2024)
Similar Items
-
TransMamba: A Sequence-Level Hybrid Transformer-Mamba Language Model
by: Li, Yixing, et al.
Published: (2025) -
C2-Evo: Co-Evolving Multimodal Data and Model for Self-Improving Reasoning
by: Chen, Xiuwei, et al.
Published: (2025) -
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
by: Chen, Zisheng, et al.
Published: (2025) -
GeoMamba: A Geometry-driven MambaVision Framework and Dataset for Fine-grained Optical-SAR Object Retrieval
by: Fang, Tiantong, et al.
Published: (2026) -
AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
by: Zhang, Likui, et al.
Published: (2026)