Mamba or Transformer for Time Series Forecasting? Mixture of Universals (MoU) Is All You Need
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Sijia, Xiong, Yun, Zhu, Yangyong, Shen, Zhiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attention Smoothing Is All You Need For Unlearning
by: Zade, Saleh Zare, et al.
Published: (2026)
by: Zade, Saleh Zare, et al.
Published: (2026)
Unified Training of Universal Time Series Forecasting Transformers
by: Woo, Gerald, et al.
Published: (2024)
by: Woo, Gerald, et al.
Published: (2024)
Attention Is All You Need for KV Cache in Diffusion LLMs
by: Nguyen-Tri, Quan, et al.
Published: (2025)
by: Nguyen-Tri, Quan, et al.
Published: (2025)
Wavelet Mixture of Experts for Time Series Forecasting
by: Zhou, Zheng, et al.
Published: (2025)
by: Zhou, Zheng, et al.
Published: (2025)
Seg-MoE: Multi-Resolution Segment-wise Mixture-of-Experts for Time Series Forecasting Transformers
by: Ortigossa, Evandro S., et al.
Published: (2026)
by: Ortigossa, Evandro S., et al.
Published: (2026)
ShapeCond: Fast Shapelet-Guided Dataset Condensation for Time Series Classification
by: Peng, Sijia, et al.
Published: (2026)
by: Peng, Sijia, et al.
Published: (2026)
MoFE-Time: Mixture of Frequency Domain Experts for Time-Series Forecasting Models
by: Liu, Yiwen, et al.
Published: (2025)
by: Liu, Yiwen, et al.
Published: (2025)
ms-Mamba: Multi-scale Mamba for Time-Series Forecasting
by: Karadag, Yusuf Meric, et al.
Published: (2025)
by: Karadag, Yusuf Meric, et al.
Published: (2025)
MoHETS: Long-term Time Series Forecasting with Mixture-of-Heterogeneous-Experts
by: Ortigossa, Evandro S., et al.
Published: (2026)
by: Ortigossa, Evandro S., et al.
Published: (2026)
MoGU: Mixture-of-Gaussians with Uncertainty-based Gating for Time Series Forecasting
by: Aviv, Gilad, et al.
Published: (2025)
by: Aviv, Gilad, et al.
Published: (2025)
DB2-TransF: All You Need Is Learnable Daubechies Wavelets for Time Series Forecasting
by: Gupta, Moulik, et al.
Published: (2025)
by: Gupta, Moulik, et al.
Published: (2025)
SST: Multi-Scale Hybrid Mamba-Transformer Experts for Time Series Forecasting
by: Xu, Xiongxiao, et al.
Published: (2024)
by: Xu, Xiongxiao, et al.
Published: (2024)
LeMoLE: LLM-Enhanced Mixture of Linear Experts for Time Series Forecasting
by: Zhang, Lingzheng, et al.
Published: (2024)
by: Zhang, Lingzheng, et al.
Published: (2024)
HTMformer: Hybrid Time and Multivariate Transformer for Time Series Forecasting
by: Wang, Tan, et al.
Published: (2025)
by: Wang, Tan, et al.
Published: (2025)
DMamba: Decomposition-enhanced Mamba for Time Series Forecasting
by: Chen, Ruxuan, et al.
Published: (2026)
by: Chen, Ruxuan, et al.
Published: (2026)
Sequential Order-Robust Mamba for Time Series Forecasting
by: Lee, Seunghan, et al.
Published: (2024)
by: Lee, Seunghan, et al.
Published: (2024)
A Mamba Foundation Model for Time Series Forecasting
by: Ma, Haoyu, et al.
Published: (2024)
by: Ma, Haoyu, et al.
Published: (2024)
MoDEx: Mixture of Depth-specific Experts for Multivariate Long-term Time Series Forecasting
by: Yoon, Hyekyung, et al.
Published: (2026)
by: Yoon, Hyekyung, et al.
Published: (2026)
STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction
by: Chen, Haolong, et al.
Published: (2025)
by: Chen, Haolong, et al.
Published: (2025)
Attention is All You Need Until You Need Retention
by: Yaslioglu, M. Murat
Published: (2025)
by: Yaslioglu, M. Murat
Published: (2025)
WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting
by: Wu, Shunyu, et al.
Published: (2026)
by: Wu, Shunyu, et al.
Published: (2026)
Context is All You Need
by: Delanois, Jean Erik, et al.
Published: (2026)
by: Delanois, Jean Erik, et al.
Published: (2026)
Integration of Mamba and Transformer -- MAT for Long-Short Range Time Series Forecasting with Application to Weather Dynamics
by: Zhang, Wenqing, et al.
Published: (2024)
by: Zhang, Wenqing, et al.
Published: (2024)
MambaTS: Improved Selective State Space Models for Long-term Time Series Forecasting
by: Cai, Xiuding, et al.
Published: (2024)
by: Cai, Xiuding, et al.
Published: (2024)
Time-o1: Time-Series Forecasting Needs Transformed Label Alignment
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
ASGMamba: Adaptive Spectral Gating Mamba for Multivariate Time Series Forecasting
by: Li, Qianyang, et al.
Published: (2026)
by: Li, Qianyang, et al.
Published: (2026)
Small Graph Is All You Need: DeepStateGNN for Scalable Traffic Forecasting
by: Wölker, Yannick, et al.
Published: (2025)
by: Wölker, Yannick, et al.
Published: (2025)
Rethinking Time Encoding via Learnable Transformation Functions
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
HDT: Hierarchical Discrete Transformer for Multivariate Time Series Forecasting
by: Feng, Shibo, et al.
Published: (2025)
by: Feng, Shibo, et al.
Published: (2025)
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025)
by: Rentschler, Micah, et al.
Published: (2025)
AME-TS: Anchored Mixture-of-Experts for Time Series Forecasting
by: Wang, Rui, et al.
Published: (2026)
by: Wang, Rui, et al.
Published: (2026)
Mixture-of-Linear-Experts for Long-term Time Series Forecasting
by: Ni, Ronghao, et al.
Published: (2023)
by: Ni, Ronghao, et al.
Published: (2023)
The Residual Stream Is All You Need: On the Redundancy of the KV Cache in Transformer Inference
by: Qasim, Kaleem Ullah, et al.
Published: (2026)
by: Qasim, Kaleem Ullah, et al.
Published: (2026)
Hierarchical Information-Guided Spatio-Temporal Mamba for Stock Time Series Forecasting
by: Yan, Wenbo, et al.
Published: (2025)
by: Yan, Wenbo, et al.
Published: (2025)
Efficient Deep Learning Board: Training Feedback Is Not All You Need
by: Gong, Lina, et al.
Published: (2024)
by: Gong, Lina, et al.
Published: (2024)
Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
by: Shi, Xiaoming, et al.
Published: (2024)
by: Shi, Xiaoming, et al.
Published: (2024)
All You Need Is Synthetic Task Augmentation
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
Element-wise Attention Is All You Need
by: Feng, Guoxin
Published: (2025)
by: Feng, Guoxin
Published: (2025)
VARMA-Enhanced Transformer for Time Series Forecasting
by: Song, Jiajun, et al.
Published: (2025)
by: Song, Jiajun, et al.
Published: (2025)
ISMRNN: An Implicitly Segmented RNN Method with Mamba for Long-Term Time Series Forecasting
by: Zhao, GaoXiang, et al.
Published: (2024)
by: Zhao, GaoXiang, et al.
Published: (2024)
Similar Items
-
Attention Smoothing Is All You Need For Unlearning
by: Zade, Saleh Zare, et al.
Published: (2026) -
Unified Training of Universal Time Series Forecasting Transformers
by: Woo, Gerald, et al.
Published: (2024) -
Attention Is All You Need for KV Cache in Diffusion LLMs
by: Nguyen-Tri, Quan, et al.
Published: (2025) -
Wavelet Mixture of Experts for Time Series Forecasting
by: Zhou, Zheng, et al.
Published: (2025) -
Seg-MoE: Multi-Resolution Segment-wise Mixture-of-Experts for Time Series Forecasting Transformers
by: Ortigossa, Evandro S., et al.
Published: (2026)