M1: Towards Scalable Test-Time Compute with Mamba Reasoning Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Junxiong, Li, Wen-Ding, Paliotta, Daniele, Ritter, Daniel, Rush, Alexander M., Dao, Tri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Mamba in the Llama: Distilling and Accelerating Hybrid Models
by: Wang, Junxiong, et al.
Published: (2024)
by: Wang, Junxiong, et al.
Published: (2024)
Mamba: Linear-Time Sequence Modeling with Selective State Spaces
by: Gu, Albert, et al.
Published: (2023)
by: Gu, Albert, et al.
Published: (2023)
MambaByte: Token-free Selective State Space Model
by: Wang, Junxiong, et al.
Published: (2024)
by: Wang, Junxiong, et al.
Published: (2024)
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners
by: Paliotta, Daniele, et al.
Published: (2025)
by: Paliotta, Daniele, et al.
Published: (2025)
M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling
by: Mishra, Mayank, et al.
Published: (2026)
by: Mishra, Mayank, et al.
Published: (2026)
Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality
by: Dao, Tri, et al.
Published: (2024)
by: Dao, Tri, et al.
Published: (2024)
Mamba-3: Improved Sequence Modeling using State Space Principles
by: Lahoti, Aakash, et al.
Published: (2026)
by: Lahoti, Aakash, et al.
Published: (2026)
Understanding and Minimising Outlier Features in Neural Network Training
by: He, Bobby, et al.
Published: (2024)
by: He, Bobby, et al.
Published: (2024)
Stabilizing Recurrent Dynamics for Test-Time Scalable Latent Reasoning in Looped Language Models
by: Yang, Xiao-Wen, et al.
Published: (2026)
by: Yang, Xiao-Wen, et al.
Published: (2026)
Leveraging the true depth of LLMs
by: González, Ramón Calvo, et al.
Published: (2025)
by: González, Ramón Calvo, et al.
Published: (2025)
Opportunistic Expert Activation: Batch-Aware Expert Routing for Faster Decode Without Retraining
by: Oncescu, Costin-Andrei, et al.
Published: (2025)
by: Oncescu, Costin-Andrei, et al.
Published: (2025)
Reward Model Generalization for Compute-Aware Test-Time Reasoning
by: Song, Zeen, et al.
Published: (2025)
by: Song, Zeen, et al.
Published: (2025)
Speculative Speculative Decoding
by: Kumar, Tanishq, et al.
Published: (2026)
by: Kumar, Tanishq, et al.
Published: (2026)
Adaptive Test-Time Compute Allocation for Reasoning LLMs via Constrained Policy Optimization
by: Zhai, Zhiyuan, et al.
Published: (2026)
by: Zhai, Zhiyuan, et al.
Published: (2026)
Hardware-Efficient Attention for Fast Decoding
by: Zadouri, Ted, et al.
Published: (2025)
by: Zadouri, Ted, et al.
Published: (2025)
Compute-Constrained Data Selection
by: Yin, Junjie Oscar, et al.
Published: (2024)
by: Yin, Junjie Oscar, et al.
Published: (2024)
Hydra: Bidirectional State Space Models Through Generalized Matrix Mixers
by: Hwang, Sukjun, et al.
Published: (2024)
by: Hwang, Sukjun, et al.
Published: (2024)
An Empirical Study of Mamba-based Language Models
by: Waleffe, Roger, et al.
Published: (2024)
by: Waleffe, Roger, et al.
Published: (2024)
Toward Scalable and Valid Conditional Independence Testing with Spectral Representations
by: Frohlich, Alek, et al.
Published: (2025)
by: Frohlich, Alek, et al.
Published: (2025)
FR-Mamba: Time-Series Physical Field Reconstruction Based on State Space Model
by: Long, Jiahuan, et al.
Published: (2025)
by: Long, Jiahuan, et al.
Published: (2025)
Bi-Mamba+: Bidirectional Mamba for Time Series Forecasting
by: Liang, Aobo, et al.
Published: (2024)
by: Liang, Aobo, et al.
Published: (2024)
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
by: MiniMax, et al.
Published: (2025)
by: MiniMax, et al.
Published: (2025)
FEMBA: Efficient and Scalable EEG Analysis with a Bidirectional Mamba Foundation Model
by: Tegon, Anna, et al.
Published: (2025)
by: Tegon, Anna, et al.
Published: (2025)
LLMs Can Learn to Reason Via Off-Policy RL
by: Ritter, Daniel, et al.
Published: (2026)
by: Ritter, Daniel, et al.
Published: (2026)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
by: Chen, Yanxi, et al.
Published: (2024)
by: Chen, Yanxi, et al.
Published: (2024)
MedMamba: Recasting Mamba for Medical Time Series Classification
by: He, ZhengXiao, et al.
Published: (2026)
by: He, ZhengXiao, et al.
Published: (2026)
$\nabla$-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space
by: Wang, Peihao, et al.
Published: (2026)
by: Wang, Peihao, et al.
Published: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
by: You, Runyang, et al.
Published: (2025)
by: You, Runyang, et al.
Published: (2025)
Towards Scalable and Robust Model Versioning
by: Ding, Wenxin, et al.
Published: (2024)
by: Ding, Wenxin, et al.
Published: (2024)
eMamba: Efficient Acceleration Framework for Mamba Models in Edge Computing
by: Kim, Jiyong, et al.
Published: (2025)
by: Kim, Jiyong, et al.
Published: (2025)
VFScale: Intrinsic Reasoning through Verifier-Free Test-time Scalable Diffusion Model
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
Is Mamba Effective for Time Series Forecasting?
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs
by: Alomrani, Mohammad Ali, et al.
Published: (2025)
by: Alomrani, Mohammad Ali, et al.
Published: (2025)
Exploration-Driven Optimization for Test-Time Large Language Model Reasoning
by: Li, Changhao, et al.
Published: (2026)
by: Li, Changhao, et al.
Published: (2026)
PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning
by: Hu, Jingcheng, et al.
Published: (2026)
by: Hu, Jingcheng, et al.
Published: (2026)
HybriDNA: A Hybrid Transformer-Mamba2 Long-Range DNA Language Model
by: Ma, Mingqian, et al.
Published: (2025)
by: Ma, Mingqian, et al.
Published: (2025)
Towards Scalable Backpropagation-Free Gradient Estimation
by: Wang, Daniel, et al.
Published: (2025)
by: Wang, Daniel, et al.
Published: (2025)
Learning to Reason from Feedback at Test-Time
by: Li, Yanyang, et al.
Published: (2025)
by: Li, Yanyang, et al.
Published: (2025)
Caduceus: Bi-Directional Equivariant Long-Range DNA Sequence Modeling
by: Schiff, Yair, et al.
Published: (2024)
by: Schiff, Yair, et al.
Published: (2024)
Graph Mamba: Towards Learning on Graphs with State Space Models
by: Behrouz, Ali, et al.
Published: (2024)
by: Behrouz, Ali, et al.
Published: (2024)
Similar Items
-
The Mamba in the Llama: Distilling and Accelerating Hybrid Models
by: Wang, Junxiong, et al.
Published: (2024) -
Mamba: Linear-Time Sequence Modeling with Selective State Spaces
by: Gu, Albert, et al.
Published: (2023) -
MambaByte: Token-free Selective State Space Model
by: Wang, Junxiong, et al.
Published: (2024) -
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners
by: Paliotta, Daniele, et al.
Published: (2025) -
M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling
by: Mishra, Mayank, et al.
Published: (2026)