xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
Fuente:
arXiv
Saved in:
| Main Authors: | Beck, Maximilian, Schweighofer, Kajetan, Böck, Sebastian, Lehner, Sebastian, Hochreiter, Sepp |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
Vision-LSTM: xLSTM as Generic Vision Backbone
by: Alkin, Benedikt, et al.
Published: (2024)
by: Alkin, Benedikt, et al.
Published: (2024)
Effective Distillation to Hybrid xLSTM Architectures
by: Hauzenberger, Lukas, et al.
Published: (2026)
by: Hauzenberger, Lukas, et al.
Published: (2026)
xLSTM 7B: A Recurrent LLM for Fast and Efficient Inference
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
by: Aichberger, Lukas, et al.
Published: (2024)
by: Aichberger, Lukas, et al.
Published: (2024)
xLSTM: Extended Long Short-Term Memory
by: Beck, Maximilian, et al.
Published: (2024)
by: Beck, Maximilian, et al.
Published: (2024)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
by: Ielanskyi, Mykyta, et al.
Published: (2025)
by: Ielanskyi, Mykyta, et al.
Published: (2025)
On Information-Theoretic Measures of Predictive Uncertainty
by: Schweighofer, Kajetan, et al.
Published: (2024)
by: Schweighofer, Kajetan, et al.
Published: (2024)
Improving Uncertainty Estimation through Semantically Diverse Language Generation
by: Aichberger, Lukas, et al.
Published: (2024)
by: Aichberger, Lukas, et al.
Published: (2024)
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
The Disparate Benefits of Deep Ensembles
by: Schweighofer, Kajetan, et al.
Published: (2024)
by: Schweighofer, Kajetan, et al.
Published: (2024)
A Diffusion Model Framework for Unsupervised Neural Combinatorial Optimization
by: Sanokowski, Sebastian, et al.
Published: (2024)
by: Sanokowski, Sebastian, et al.
Published: (2024)
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences
by: Schmidinger, Niklas, et al.
Published: (2024)
by: Schmidinger, Niklas, et al.
Published: (2024)
Rethinking Losses for Diffusion Bridge Samplers
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
by: Pöppel, Korbinian, et al.
Published: (2024)
by: Pöppel, Korbinian, et al.
Published: (2024)
xLSTM-ECG: Multi-label ECG Classification via Feature Fusion with xLSTM
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
xLSTMTime : Long-term Time Series Forecasting With xLSTM
by: Alharthi, Musleh, et al.
Published: (2024)
by: Alharthi, Musleh, et al.
Published: (2024)
pLSTM: parallelizable Linear Source Transition Mark networks
by: Pöppel, Korbinian, et al.
Published: (2025)
by: Pöppel, Korbinian, et al.
Published: (2025)
Benchmarking Transformer and xLSTM for Time-Series Forecasting of Heat Consumption
by: Wahl, Marja, et al.
Published: (2026)
by: Wahl, Marja, et al.
Published: (2026)
Pre-trained Forecasting Models: Strong Zero-Shot Feature Extractors for Time Series Classification
by: Auer, Andreas, et al.
Published: (2025)
by: Auer, Andreas, et al.
Published: (2025)
Seg-LSTM: Performance of xLSTM for Semantic Segmentation of Remotely Sensed Images
by: Zhu, Qinfeng, et al.
Published: (2024)
by: Zhu, Qinfeng, et al.
Published: (2024)
MolGraph-xLSTM: A graph-based dual-level xLSTM framework with multi-head mixture-of-experts for enhanced molecular representation and interpretability
by: Sun, Yan, et al.
Published: (2025)
by: Sun, Yan, et al.
Published: (2025)
TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
by: Auer, Andreas, et al.
Published: (2025)
by: Auer, Andreas, et al.
Published: (2025)
Scalable Discrete Diffusion Samplers: Combinatorial Optimization and Statistical Physics
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
xLSTMAD: A Powerful xLSTM-based Method for Anomaly Detection
by: Faber, Kamil, et al.
Published: (2025)
by: Faber, Kamil, et al.
Published: (2025)
xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar Memories
by: Kraus, Maurice, et al.
Published: (2024)
by: Kraus, Maurice, et al.
Published: (2024)
xLSTM-PINN: Memory-Gated Spectral Remodeling for Physics-Informed Learning
by: Tao, Ze, et al.
Published: (2025)
by: Tao, Ze, et al.
Published: (2025)
Overcoming Saturation in Density Ratio Estimation by Iterated Regularization
by: Gruber, Lukas, et al.
Published: (2024)
by: Gruber, Lukas, et al.
Published: (2024)
Enhancing Spatiotemporal Networks with xLSTM: A Scalar LSTM Approach for Cellular Traffic Forecasting
by: Ali, Khalid, et al.
Published: (2025)
by: Ali, Khalid, et al.
Published: (2025)
Distil-xLSTM: Learning Attention Mechanisms through Recurrent Structures
by: Thiombiano, Abdoul Majid O., et al.
Published: (2025)
by: Thiombiano, Abdoul Majid O., et al.
Published: (2025)
Energy-based Hopfield Boosting for Out-of-Distribution Detection
by: Hofmann, Claus, et al.
Published: (2024)
by: Hofmann, Claus, et al.
Published: (2024)
AP-OOD: Attention Pooling for Out-of-Distribution Detection
by: Hofmann, Claus, et al.
Published: (2026)
by: Hofmann, Claus, et al.
Published: (2026)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
by: Thiombiano, Abdoul Majid O., et al.
Published: (2025)
by: Thiombiano, Abdoul Majid O., et al.
Published: (2025)
Safe and Certifiable AI Systems: Concepts, Challenges, and Lessons Learned
by: Schweighofer, Kajetan, et al.
Published: (2025)
by: Schweighofer, Kajetan, et al.
Published: (2025)
Uncertainty Quantification for Regression using Proper Scoring Rules
by: Fishkov, Alexander, et al.
Published: (2025)
by: Fishkov, Alexander, et al.
Published: (2025)
Linear Alignment of Vision-language Models for Image Captioning
by: Paischer, Fabian, et al.
Published: (2023)
by: Paischer, Fabian, et al.
Published: (2023)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
by: Schweighofer, Kajetan, et al.
Published: (2026)
by: Schweighofer, Kajetan, et al.
Published: (2026)
A Deep Reinforcement Learning Approach to Automated Stock Trading, using xLSTM Networks
by: Sarlakifar, Faezeh, et al.
Published: (2025)
by: Sarlakifar, Faezeh, et al.
Published: (2025)
xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
by: Kühne, Nikolai Lund, et al.
Published: (2025)
by: Kühne, Nikolai Lund, et al.
Published: (2025)
Efficient Pre-Training of LLMs through Truncated SVD Layers
by: Kamali, Kaivan, et al.
Published: (2026)
by: Kamali, Kaivan, et al.
Published: (2026)
Similar Items
-
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
by: Beck, Maximilian, et al.
Published: (2025) -
Vision-LSTM: xLSTM as Generic Vision Backbone
by: Alkin, Benedikt, et al.
Published: (2024) -
Effective Distillation to Hybrid xLSTM Architectures
by: Hauzenberger, Lukas, et al.
Published: (2026) -
xLSTM 7B: A Recurrent LLM for Fast and Efficient Inference
by: Beck, Maximilian, et al.
Published: (2025) -
Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
by: Aichberger, Lukas, et al.
Published: (2024)