Effective Distillation to Hybrid xLSTM Architectures
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hauzenberger, Lukas, Schmidinger, Niklas, Schmied, Thomas, Hartl, Anamaria-Roberta, Stap, David, Hoedt, Pieter-Jan, Beck, Maximilian, Böck, Sebastian, Klambauer, Günter, Hochreiter, Sepp |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences
von: Schmidinger, Niklas, et al.
Veröffentlicht: (2024)
von: Schmidinger, Niklas, et al.
Veröffentlicht: (2024)
xLSTM 7B: A Recurrent LLM for Fast and Efficient Inference
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
xLSTM: Extended Long Short-Term Memory
von: Beck, Maximilian, et al.
Veröffentlicht: (2024)
von: Beck, Maximilian, et al.
Veröffentlicht: (2024)
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
Vision-LSTM: xLSTM as Generic Vision Backbone
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
Adaptive Retrieval helps Reasoning in LLMs -- but mostly if it's not used
von: Shakya, Srijan, et al.
Veröffentlicht: (2026)
von: Shakya, Srijan, et al.
Veröffentlicht: (2026)
Parameter Efficient Fine-tuning via Explained Variance Adaptation
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
pLSTM: parallelizable Linear Source Transition Mark networks
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2025)
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2025)
xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
xLSTM-ECG: Multi-label ECG Classification via Feature Fusion with xLSTM
von: Kang, Lei, et al.
Veröffentlicht: (2025)
von: Kang, Lei, et al.
Veröffentlicht: (2025)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2024)
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2024)
VN-EGNN: E(3)-Equivariant Graph Neural Networks with Virtual Nodes Enhance Protein Binding Site Identification
von: Sestak, Florian, et al.
Veröffentlicht: (2024)
von: Sestak, Florian, et al.
Veröffentlicht: (2024)
Distil-xLSTM: Learning Attention Mechanisms through Recurrent Structures
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
When Mamba Meets xLSTM: An Efficient and Precise Method with the xLSTM-VMUNet Model for Skin lesion Segmentation
von: Fang, Zhuoyi, et al.
Veröffentlicht: (2024)
von: Fang, Zhuoyi, et al.
Veröffentlicht: (2024)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
xLSTMTime : Long-term Time Series Forecasting With xLSTM
von: Alharthi, Musleh, et al.
Veröffentlicht: (2024)
von: Alharthi, Musleh, et al.
Veröffentlicht: (2024)
Seg-LSTM: Performance of xLSTM for Semantic Segmentation of Remotely Sensed Images
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2024)
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2024)
MolGraph-xLSTM: A graph-based dual-level xLSTM framework with multi-head mixture-of-experts for enhanced molecular representation and interpretability
von: Sun, Yan, et al.
Veröffentlicht: (2025)
von: Sun, Yan, et al.
Veröffentlicht: (2025)
xLSTMAD: A Powerful xLSTM-based Method for Anomaly Detection
von: Faber, Kamil, et al.
Veröffentlicht: (2025)
von: Faber, Kamil, et al.
Veröffentlicht: (2025)
Benchmarking Transformer and xLSTM for Time-Series Forecasting of Heat Consumption
von: Wahl, Marja, et al.
Veröffentlicht: (2026)
von: Wahl, Marja, et al.
Veröffentlicht: (2026)
Pre-trained Forecasting Models: Strong Zero-Shot Feature Extractors for Time Series Classification
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
Enhancing Spatiotemporal Networks with xLSTM: A Scalar LSTM Approach for Cellular Traffic Forecasting
von: Ali, Khalid, et al.
Veröffentlicht: (2025)
von: Ali, Khalid, et al.
Veröffentlicht: (2025)
Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
xLSTM-PINN: Memory-Gated Spectral Remodeling for Physics-Informed Learning
von: Tao, Ze, et al.
Veröffentlicht: (2025)
von: Tao, Ze, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar Memories
von: Kraus, Maurice, et al.
Veröffentlicht: (2024)
von: Kraus, Maurice, et al.
Veröffentlicht: (2024)
MAL: Cluster-Masked and Multi-Task Pretraining for Enhanced xLSTM Vision Performance
von: Huang, Wenjun, et al.
Veröffentlicht: (2024)
von: Huang, Wenjun, et al.
Veröffentlicht: (2024)
AF-MAT: Aspect-aware Flip-and-Fuse xLSTM for Aspect-based Sentiment Analysis
von: Lawan, Adamu, et al.
Veröffentlicht: (2025)
von: Lawan, Adamu, et al.
Veröffentlicht: (2025)
Are Vision xLSTM Embedded UNet More Reliable in Medical 3D Image Segmentation?
von: Dutta, Pallabi, et al.
Veröffentlicht: (2024)
von: Dutta, Pallabi, et al.
Veröffentlicht: (2024)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
A Deep Reinforcement Learning Approach to Automated Stock Trading, using xLSTM Networks
von: Sarlakifar, Faezeh, et al.
Veröffentlicht: (2025)
von: Sarlakifar, Faezeh, et al.
Veröffentlicht: (2025)
On Information-Theoretic Measures of Predictive Uncertainty
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
MIM-Refiner: A Contrastive Learning Boost from Intermediate Pre-Trained Representations
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
von: Ielanskyi, Mykyta, et al.
Veröffentlicht: (2025)
von: Ielanskyi, Mykyta, et al.
Veröffentlicht: (2025)
Improving Uncertainty Estimation through Semantically Diverse Language Generation
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
xLSTM-UNet can be an Effective 2D & 3D Medical Image Segmentation Backbone with Vision-LSTM (ViL) better than its Mamba Counterpart
von: Chen, Tianrun, et al.
Veröffentlicht: (2024)
von: Chen, Tianrun, et al.
Veröffentlicht: (2024)
xLSTM-FER: Enhancing Student Expression Recognition with Extended Vision Long Short-Term Memory Network
von: Huang, Qionghao, et al.
Veröffentlicht: (2024)
von: Huang, Qionghao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences
von: Schmidinger, Niklas, et al.
Veröffentlicht: (2024) -
xLSTM 7B: A Recurrent LLM for Fast and Efficient Inference
von: Beck, Maximilian, et al.
Veröffentlicht: (2025) -
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
von: Schmied, Thomas, et al.
Veröffentlicht: (2024) -
xLSTM: Extended Long Short-Term Memory
von: Beck, Maximilian, et al.
Veröffentlicht: (2024) -
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)