pLSTM: parallelizable Linear Source Transition Mark networks
Fuente:
arXiv
Saved in:
| Main Authors: | Pöppel, Korbinian, Freinschlag, Richard, Schmied, Thomas, Lin, Wei, Hochreiter, Sepp |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
Vision-LSTM: xLSTM as Generic Vision Backbone
by: Alkin, Benedikt, et al.
Published: (2024)
by: Alkin, Benedikt, et al.
Published: (2024)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
by: Pöppel, Korbinian, et al.
Published: (2024)
by: Pöppel, Korbinian, et al.
Published: (2024)
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Adaptive Retrieval helps Reasoning in LLMs -- but mostly if it's not used
by: Shakya, Srijan, et al.
Published: (2026)
by: Shakya, Srijan, et al.
Published: (2026)
xLSTM 7B: A Recurrent LLM for Fast and Efficient Inference
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
xLSTM: Extended Long Short-Term Memory
by: Beck, Maximilian, et al.
Published: (2024)
by: Beck, Maximilian, et al.
Published: (2024)
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
Effective Distillation to Hybrid xLSTM Architectures
by: Hauzenberger, Lukas, et al.
Published: (2026)
by: Hauzenberger, Lukas, et al.
Published: (2026)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Parameter Efficient Fine-tuning via Explained Variance Adaptation
by: Paischer, Fabian, et al.
Published: (2024)
by: Paischer, Fabian, et al.
Published: (2024)
Linear Alignment of Vision-language Models for Image Captioning
by: Paischer, Fabian, et al.
Published: (2023)
by: Paischer, Fabian, et al.
Published: (2023)
Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
by: Aichberger, Lukas, et al.
Published: (2024)
by: Aichberger, Lukas, et al.
Published: (2024)
A Diffusion Model Framework for Unsupervised Neural Combinatorial Optimization
by: Sanokowski, Sebastian, et al.
Published: (2024)
by: Sanokowski, Sebastian, et al.
Published: (2024)
Pre-trained Forecasting Models: Strong Zero-Shot Feature Extractors for Time Series Classification
by: Auer, Andreas, et al.
Published: (2025)
by: Auer, Andreas, et al.
Published: (2025)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
by: Ielanskyi, Mykyta, et al.
Published: (2025)
by: Ielanskyi, Mykyta, et al.
Published: (2025)
On Information-Theoretic Measures of Predictive Uncertainty
by: Schweighofer, Kajetan, et al.
Published: (2024)
by: Schweighofer, Kajetan, et al.
Published: (2024)
Improving Uncertainty Estimation through Semantically Diverse Language Generation
by: Aichberger, Lukas, et al.
Published: (2024)
by: Aichberger, Lukas, et al.
Published: (2024)
Contrastive Abstraction for Reinforcement Learning
by: Patil, Vihang, et al.
Published: (2024)
by: Patil, Vihang, et al.
Published: (2024)
Overcoming Saturation in Density Ratio Estimation by Iterated Regularization
by: Gruber, Lukas, et al.
Published: (2024)
by: Gruber, Lukas, et al.
Published: (2024)
Simplified priors for Object-Centric Learning
by: Patil, Vihang, et al.
Published: (2024)
by: Patil, Vihang, et al.
Published: (2024)
The Disparate Benefits of Deep Ensembles
by: Schweighofer, Kajetan, et al.
Published: (2024)
by: Schweighofer, Kajetan, et al.
Published: (2024)
Symbol-Equivariant Recurrent Reasoning Models
by: Freinschlag, Richard, et al.
Published: (2026)
by: Freinschlag, Richard, et al.
Published: (2026)
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences
by: Schmidinger, Niklas, et al.
Published: (2024)
by: Schmidinger, Niklas, et al.
Published: (2024)
The Offline-Frontier Shift: Diagnosing Distributional Limits in Generative Multi-Objective Optimization
by: Holly, Stephanie, et al.
Published: (2026)
by: Holly, Stephanie, et al.
Published: (2026)
MIM-Refiner: A Contrastive Learning Boost from Intermediate Pre-Trained Representations
by: Alkin, Benedikt, et al.
Published: (2024)
by: Alkin, Benedikt, et al.
Published: (2024)
Rethinking Losses for Diffusion Bridge Samplers
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
by: Auer, Andreas, et al.
Published: (2025)
by: Auer, Andreas, et al.
Published: (2025)
AP-OOD: Attention Pooling for Out-of-Distribution Detection
by: Hofmann, Claus, et al.
Published: (2026)
by: Hofmann, Claus, et al.
Published: (2026)
Energy-based Hopfield Boosting for Out-of-Distribution Detection
by: Hofmann, Claus, et al.
Published: (2024)
by: Hofmann, Claus, et al.
Published: (2024)
GNN-VPA: A Variance-Preserving Aggregation Strategy for Graph Neural Networks
by: Schneckenreiter, Lisa, et al.
Published: (2024)
by: Schneckenreiter, Lisa, et al.
Published: (2024)
VN-EGNN: E(3)-Equivariant Graph Neural Networks with Virtual Nodes Enhance Protein Binding Site Identification
by: Sestak, Florian, et al.
Published: (2024)
by: Sestak, Florian, et al.
Published: (2024)
Geometry-Informed Neural Networks
by: Berzins, Arturs, et al.
Published: (2024)
by: Berzins, Arturs, et al.
Published: (2024)
SymbolicAI: A framework for logic-based approaches combining generative models and solvers
by: Dinu, Marius-Constantin, et al.
Published: (2024)
by: Dinu, Marius-Constantin, et al.
Published: (2024)
Scalable Discrete Diffusion Samplers: Combinatorial Optimization and Statistical Physics
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
Large Language Models Can Self-Improve At Web Agent Tasks
by: Patel, Ajay, et al.
Published: (2024)
by: Patel, Ajay, et al.
Published: (2024)
An Open-Source and Reproducible Implementation of LSTM and GRU Networks for Time Series Forecasting
by: Velarde, Gissel, et al.
Published: (2025)
by: Velarde, Gissel, et al.
Published: (2025)
Towards Improved Research Methodologies for Industrial AI: A case study of false call reduction
by: Pfab, Korbinian, et al.
Published: (2025)
by: Pfab, Korbinian, et al.
Published: (2025)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
by: Schmied, Thomas, et al.
Published: (2025)
by: Schmied, Thomas, et al.
Published: (2025)
StrADiff: A Structured Source-Wise Adaptive Diffusion Framework for Linear and Nonlinear Blind Source Separation
by: Wei, Yuan-Hao
Published: (2026)
by: Wei, Yuan-Hao
Published: (2026)
Similar Items
-
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
by: Beck, Maximilian, et al.
Published: (2025) -
Vision-LSTM: xLSTM as Generic Vision Backbone
by: Alkin, Benedikt, et al.
Published: (2024) -
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
by: Pöppel, Korbinian, et al.
Published: (2024) -
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
by: Schmied, Thomas, et al.
Published: (2024) -
Adaptive Retrieval helps Reasoning in LLMs -- but mostly if it's not used
by: Shakya, Srijan, et al.
Published: (2026)