TransMamba: Fast Universal Architecture Adaption from Transformers to Mamba
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Xiuwei, Hu, Wentao, Dong, Xiao, Lin, Sihao, Chen, Zisheng, Cao, Meng, Zhuang, Yina, Han, Jianhua, Xu, Hang, Liang, Xiaodan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TransMamba: A Sequence-Level Hybrid Transformer-Mamba Language Model
di: Li, Yixing, et al.
Pubblicazione: (2025)
di: Li, Yixing, et al.
Pubblicazione: (2025)
C2-Evo: Co-Evolving Multimodal Data and Model for Self-Improving Reasoning
di: Chen, Xiuwei, et al.
Pubblicazione: (2025)
di: Chen, Xiuwei, et al.
Pubblicazione: (2025)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
di: Chen, Zisheng, et al.
Pubblicazione: (2025)
di: Chen, Zisheng, et al.
Pubblicazione: (2025)
GeoMamba: A Geometry-driven MambaVision Framework and Dataset for Fine-grained Optical-SAR Object Retrieval
di: Fang, Tiantong, et al.
Pubblicazione: (2026)
di: Fang, Tiantong, et al.
Pubblicazione: (2026)
AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
di: Zhang, Likui, et al.
Pubblicazione: (2026)
di: Zhang, Likui, et al.
Pubblicazione: (2026)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
di: Chen, Wenchao, et al.
Pubblicazione: (2024)
di: Chen, Wenchao, et al.
Pubblicazione: (2024)
InfoMamba: An Attention-Free Hybrid Mamba-Transformer Model
di: Wang, Youjin, et al.
Pubblicazione: (2026)
di: Wang, Youjin, et al.
Pubblicazione: (2026)
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
di: Tang, Haoran, et al.
Pubblicazione: (2024)
di: Tang, Haoran, et al.
Pubblicazione: (2024)
VAMamba: An Efficient Visual Adaptive Mamba for Image Restoration
di: Hu, Han, et al.
Pubblicazione: (2025)
di: Hu, Han, et al.
Pubblicazione: (2025)
OccluDentNet : Self‐Supervised 3D Tooth Segmentation via MambaTrans Architecture
di: Hong‐an Li, et al.
Pubblicazione: (2026)
di: Hong‐an Li, et al.
Pubblicazione: (2026)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
di: Ren, Weiming, et al.
Pubblicazione: (2025)
di: Ren, Weiming, et al.
Pubblicazione: (2025)
MedSegMamba: 3D CNN-Mamba Hybrid Architecture for Brain Segmentation
di: Cao, Aaron, et al.
Pubblicazione: (2024)
di: Cao, Aaron, et al.
Pubblicazione: (2024)
Decision Mamba Architectures
di: Correia, André, et al.
Pubblicazione: (2024)
di: Correia, André, et al.
Pubblicazione: (2024)
MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation
di: Du, Juntian, et al.
Pubblicazione: (2025)
di: Du, Juntian, et al.
Pubblicazione: (2025)
MambaVesselNet++: A Hybrid CNN-Mamba Architecture for Medical Image Segmentation
di: Xu, Qing, et al.
Pubblicazione: (2025)
di: Xu, Qing, et al.
Pubblicazione: (2025)
MambaGlue: Fast and Robust Local Feature Matching With Mamba
di: Ryoo, Kihwan, et al.
Pubblicazione: (2025)
di: Ryoo, Kihwan, et al.
Pubblicazione: (2025)
Laplace-Mamba: Laplace Frequency Prior-Guided Mamba-CNN Fusion Network for Image Dehazing
di: Wang, Yongzhen, et al.
Pubblicazione: (2025)
di: Wang, Yongzhen, et al.
Pubblicazione: (2025)
InceptionMamba: An Efficient Hybrid Network with Large Band Convolution and Bottleneck Mamba
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
InfiniMotion: Mamba Boosts Memory in Transformer for Arbitrary Long Motion Generation
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
Mamba Modulation: On the Length Generalization of Mamba
di: Lu, Peng, et al.
Pubblicazione: (2025)
di: Lu, Peng, et al.
Pubblicazione: (2025)
MedMamba: Recasting Mamba for Medical Time Series Classification
di: He, ZhengXiao, et al.
Pubblicazione: (2026)
di: He, ZhengXiao, et al.
Pubblicazione: (2026)
TextMamba: Scene Text Detector with Mamba
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
MambaVision: A Hybrid Mamba-Transformer Vision Backbone
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2024)
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2024)
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods
di: Xu, Zukang, et al.
Pubblicazione: (2025)
di: Xu, Zukang, et al.
Pubblicazione: (2025)
Hazy Pedestrian Trajectory Prediction via Physical Priors and Graph-Mamba
di: Chen, Jian, et al.
Pubblicazione: (2025)
di: Chen, Jian, et al.
Pubblicazione: (2025)
SMILES-Mamba: Chemical Mamba Foundation Models for Drug ADMET Prediction
di: Xu, Bohao, et al.
Pubblicazione: (2024)
di: Xu, Bohao, et al.
Pubblicazione: (2024)
MambaNet: Mamba-assisted Channel Estimation Neural Network With Attention Mechanism
di: Luan, Dianxin, et al.
Pubblicazione: (2026)
di: Luan, Dianxin, et al.
Pubblicazione: (2026)
MobileMamba: Lightweight Multi-Receptive Visual Mamba Network
di: He, Haoyang, et al.
Pubblicazione: (2024)
di: He, Haoyang, et al.
Pubblicazione: (2024)
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement
di: Wang, Junyu, et al.
Pubblicazione: (2024)
di: Wang, Junyu, et al.
Pubblicazione: (2024)
MxT: Mamba x Transformer for Image Inpainting
di: Chen, Shuang, et al.
Pubblicazione: (2024)
di: Chen, Shuang, et al.
Pubblicazione: (2024)
RankMamba: Benchmarking Mamba's Document Ranking Performance in the Era of Transformers
di: Xu, Zhichao
Pubblicazione: (2024)
di: Xu, Zhichao
Pubblicazione: (2024)
GTR-Mamba: Geometry-to-Tangent Routing Mamba for Hyperbolic POI Recommendation
di: Li, Zhuoxuan, et al.
Pubblicazione: (2025)
di: Li, Zhuoxuan, et al.
Pubblicazione: (2025)
Bi-Mamba+: Bidirectional Mamba for Time Series Forecasting
di: Liang, Aobo, et al.
Pubblicazione: (2024)
di: Liang, Aobo, et al.
Pubblicazione: (2024)
Hybrid Transformer-Mamba Architecture for Weakly Supervised Volumetric Medical Segmentation
di: Lyu, Yiheng, et al.
Pubblicazione: (2025)
di: Lyu, Yiheng, et al.
Pubblicazione: (2025)
MambaSCI: Efficient Mamba-UNet for Quad-Bayer Patterned Video Snapshot Compressive Imaging
di: Pan, Zhenghao, et al.
Pubblicazione: (2024)
di: Pan, Zhenghao, et al.
Pubblicazione: (2024)
TC-BiMamba: Trans-Chunk bidirectionally within BiMamba for unified streaming and non-streaming ASR
di: She, Qingshun, et al.
Pubblicazione: (2026)
di: She, Qingshun, et al.
Pubblicazione: (2026)
MM-DETR: An Efficient Multimodal Detection Transformer with Mamba-Driven Dual-Granularity Fusion and Frequency-Aware Modality Adapters
di: Han, Jianhong, et al.
Pubblicazione: (2025)
di: Han, Jianhong, et al.
Pubblicazione: (2025)
Mamba4KT:An Efficient and Effective Mamba-based Knowledge Tracing Model
di: Cao, Yang, et al.
Pubblicazione: (2024)
di: Cao, Yang, et al.
Pubblicazione: (2024)
FastMamba: A High-Speed and Efficient Mamba Accelerator on FPGA with Accurate Quantization
di: Wang, Aotao, et al.
Pubblicazione: (2025)
di: Wang, Aotao, et al.
Pubblicazione: (2025)
PhysMamba: Efficient Remote Physiological Measurement with SlowFast Temporal Difference Mamba
di: Luo, Chaoqi, et al.
Pubblicazione: (2024)
di: Luo, Chaoqi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
TransMamba: A Sequence-Level Hybrid Transformer-Mamba Language Model
di: Li, Yixing, et al.
Pubblicazione: (2025) -
C2-Evo: Co-Evolving Multimodal Data and Model for Self-Improving Reasoning
di: Chen, Xiuwei, et al.
Pubblicazione: (2025) -
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
di: Chen, Zisheng, et al.
Pubblicazione: (2025) -
GeoMamba: A Geometry-driven MambaVision Framework and Dataset for Fine-grained Optical-SAR Object Retrieval
di: Fang, Tiantong, et al.
Pubblicazione: (2026) -
AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
di: Zhang, Likui, et al.
Pubblicazione: (2026)