Gespeichert in:
| Hauptverfasser: | Ebrahimi, M. Reza, Memisevic, Roland |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2505.21749 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the "Induction Bias" in Sequence Models
von: Ebrahimi, M. Reza, et al.
Veröffentlicht: (2026)
von: Ebrahimi, M. Reza, et al.
Veröffentlicht: (2026)
Multi-Draft Speculative Sampling: Canonical Decomposition and Theoretical Limits
von: Khisti, Ashish, et al.
Veröffentlicht: (2024)
von: Khisti, Ashish, et al.
Veröffentlicht: (2024)
Your Context Is Not an Array: Unveiling Random Access Limitations in Transformers
von: Ebrahimi, MohammadReza, et al.
Veröffentlicht: (2024)
von: Ebrahimi, MohammadReza, et al.
Veröffentlicht: (2024)
Delayed Attention Training Improves Length Generalization in Transformer--RNN Hybrids
von: Phan, Buu, et al.
Veröffentlicht: (2025)
von: Phan, Buu, et al.
Veröffentlicht: (2025)
Advancing Regular Language Reasoning in Linear Recurrent Neural Networks
von: Fan, Ting-Han, et al.
Veröffentlicht: (2023)
von: Fan, Ting-Han, et al.
Veröffentlicht: (2023)
Optimal Decay Spectra for Linear Recurrences
von: Cao, Yang
Veröffentlicht: (2026)
von: Cao, Yang
Veröffentlicht: (2026)
On the Representational Capacity of Recurrent Neural Language Models
von: Nowak, Franz, et al.
Veröffentlicht: (2023)
von: Nowak, Franz, et al.
Veröffentlicht: (2023)
Hybrid Quantum-Classical Recurrent Neural Networks
von: Xu, Wenduan
Veröffentlicht: (2025)
von: Xu, Wenduan
Veröffentlicht: (2025)
Automated SNOMED CT Concept Annotation in Clinical Text Using Bi-GRU Neural Networks
von: Noori, Ali, et al.
Veröffentlicht: (2025)
von: Noori, Ali, et al.
Veröffentlicht: (2025)
Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models
von: De, Soham, et al.
Veröffentlicht: (2024)
von: De, Soham, et al.
Veröffentlicht: (2024)
VisualRWKV: Exploring Recurrent Neural Networks for Visual Language Models
von: Hou, Haowen, et al.
Veröffentlicht: (2024)
von: Hou, Haowen, et al.
Veröffentlicht: (2024)
Depression detection from Social Media Bangla Text Using Recurrent Neural Networks
von: Ahmed, Sultan, et al.
Veröffentlicht: (2024)
von: Ahmed, Sultan, et al.
Veröffentlicht: (2024)
Dissecting Linear Recurrent Models: How Different Gating Strategies Drive Selectivity and Generalization
von: Bouhadjar, Younes, et al.
Veröffentlicht: (2026)
von: Bouhadjar, Younes, et al.
Veröffentlicht: (2026)
Look, Remember and Reason: Grounded reasoning in videos with language models
von: Bhattacharyya, Apratim, et al.
Veröffentlicht: (2023)
von: Bhattacharyya, Apratim, et al.
Veröffentlicht: (2023)
A Novel Recurrent Neural Network Framework for Prediction and Treatment of Oncogenic Mutation Progression
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
Liger: Linearizing Large Language Models to Gated Recurrent Structures
von: Lan, Disen, et al.
Veröffentlicht: (2025)
von: Lan, Disen, et al.
Veröffentlicht: (2025)
Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence
von: Xiao, Liu
Veröffentlicht: (2026)
von: Xiao, Liu
Veröffentlicht: (2026)
Rethinking State Tracking in Recurrent Models Through Error Control Dynamics
von: Chung, Jiwan, et al.
Veröffentlicht: (2026)
von: Chung, Jiwan, et al.
Veröffentlicht: (2026)
Scaling Linear Attention with Sparse State Expansion
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
Exploring Major Transitions in the Evolution of Biological Cognition With Artificial Neural Networks
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2025)
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2025)
State Stream Transformer (SST) V2: Parallel Training of Nonlinear Recurrence for Latent Space Reasoning
von: Aviss, Thea
Veröffentlicht: (2026)
von: Aviss, Thea
Veröffentlicht: (2026)
On The Expressivity of Recurrent Neural Cascades
von: Knorozova, Nadezda Alexandrovna, et al.
Veröffentlicht: (2023)
von: Knorozova, Nadezda Alexandrovna, et al.
Veröffentlicht: (2023)
From Out-of-Distribution Detection to Hallucination Detection: A Geometric View
von: Liu, Litian, et al.
Veröffentlicht: (2026)
von: Liu, Litian, et al.
Veröffentlicht: (2026)
Detection of Opioid Users from Reddit Posts via an Attention-based Bidirectional Recurrent Neural Network
von: Wang, Yuchen, et al.
Veröffentlicht: (2024)
von: Wang, Yuchen, et al.
Veröffentlicht: (2024)
Learning State-Tracking from Code Using Linear RNNs
von: Siems, Julien, et al.
Veröffentlicht: (2026)
von: Siems, Julien, et al.
Veröffentlicht: (2026)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
von: Yu, Xuemin, et al.
Veröffentlicht: (2026)
von: Yu, Xuemin, et al.
Veröffentlicht: (2026)
BiHRNN -- Bi-Directional Hierarchical Recurrent Neural Network for Inflation Forecasting
von: Vilenko, Maya
Veröffentlicht: (2025)
von: Vilenko, Maya
Veröffentlicht: (2025)
Neural Isomorphic Fields: A Transformer-based Algebraic Numerical Embedding
von: Sadeghi, Hamidreza, et al.
Veröffentlicht: (2026)
von: Sadeghi, Hamidreza, et al.
Veröffentlicht: (2026)
Attention-Based Recurrent Neural Network For Automatic Behavior Laying Hen Recognition
von: Laleye, Fréjus A. A., et al.
Veröffentlicht: (2024)
von: Laleye, Fréjus A. A., et al.
Veröffentlicht: (2024)
Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models
von: Deng, Difan, et al.
Veröffentlicht: (2026)
von: Deng, Difan, et al.
Veröffentlicht: (2026)
Max-pooling Network Revisited: Analyzing the Role of Semantic Probability in Multiple Instance Learning for Hallucination Detection
von: Fujikawa, Shota, et al.
Veröffentlicht: (2026)
von: Fujikawa, Shota, et al.
Veröffentlicht: (2026)
BiDoRA: Bi-level Optimization-Based Weight-Decomposed Low-Rank Adaptation
von: Qin, Peijia, et al.
Veröffentlicht: (2024)
von: Qin, Peijia, et al.
Veröffentlicht: (2024)
Unlocking State-Tracking in Linear RNNs Through Negative Eigenvalues
von: Grazzi, Riccardo, et al.
Veröffentlicht: (2024)
von: Grazzi, Riccardo, et al.
Veröffentlicht: (2024)
REQUAL-LM: Reliability and Equity through Aggregation in Large Language Models
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling
von: Katsch, Tobias
Veröffentlicht: (2023)
von: Katsch, Tobias
Veröffentlicht: (2023)
A New Method for Cross-Lingual-based Semantic Role Labeling
von: Ebrahimi, Mohammad, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Mohammad, et al.
Veröffentlicht: (2024)
Replacing thinking with tool usage enables reasoning in small language models
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
In-context Learning and Gradient Descent Revisited
von: Deutch, Gilad, et al.
Veröffentlicht: (2023)
von: Deutch, Gilad, et al.
Veröffentlicht: (2023)
Structured Recurrent Mixers for Massively Parallelized Sequence Generation
von: Badger, Benjamin L.
Veröffentlicht: (2026)
von: Badger, Benjamin L.
Veröffentlicht: (2026)
Convolutional Neural Networks for Toxic Comment Classification
von: Georgakopoulos, Spiros V., et al.
Veröffentlicht: (2018)
von: Georgakopoulos, Spiros V., et al.
Veröffentlicht: (2018)
Ähnliche Einträge
-
On the "Induction Bias" in Sequence Models
von: Ebrahimi, M. Reza, et al.
Veröffentlicht: (2026) -
Multi-Draft Speculative Sampling: Canonical Decomposition and Theoretical Limits
von: Khisti, Ashish, et al.
Veröffentlicht: (2024) -
Your Context Is Not an Array: Unveiling Random Access Limitations in Transformers
von: Ebrahimi, MohammadReza, et al.
Veröffentlicht: (2024) -
Delayed Attention Training Improves Length Generalization in Transformer--RNN Hybrids
von: Phan, Buu, et al.
Veröffentlicht: (2025) -
Advancing Regular Language Reasoning in Linear Recurrent Neural Networks
von: Fan, Ting-Han, et al.
Veröffentlicht: (2023)