Balancing Sparse RNNs with Hyperparameterization Benefiting Meta-Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Hershey, Quincy, Paffenroth, Randy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking the Relationship between Recurrent and Non-Recurrent Neural Networks: A Study in Sparsity
by: Hershey, Quincy, et al.
Published: (2024)
by: Hershey, Quincy, et al.
Published: (2024)
Principled Curriculum Learning using Parameter Continuation Methods
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
Solo Connection: A Parameter Efficient Fine-Tuning Technique for Transformers
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
Implicit Language Models are RNNs: Balancing Parallelization and Expressivity
by: Schöne, Mark, et al.
Published: (2025)
by: Schöne, Mark, et al.
Published: (2025)
HadamRNN: Binary and Sparse Ternary Orthogonal RNNs
by: Foucault, Armand, et al.
Published: (2025)
by: Foucault, Armand, et al.
Published: (2025)
Memory Caching: RNNs with Growing Memory
by: Behrouz, Ali, et al.
Published: (2026)
by: Behrouz, Ali, et al.
Published: (2026)
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
Learning Dynamics of RNNs in Closed-Loop Environments
by: Ger, Yoav, et al.
Published: (2025)
by: Ger, Yoav, et al.
Published: (2025)
$\texttt{lrnnx}$: A library for Linear RNNs
by: Bania, Karan, et al.
Published: (2026)
by: Bania, Karan, et al.
Published: (2026)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2024)
by: Galesloot, Maris F. L., et al.
Published: (2024)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
by: Xu, Hongtao, et al.
Published: (2026)
by: Xu, Hongtao, et al.
Published: (2026)
Learning to (Learn at Test Time): RNNs with Expressive Hidden States
by: Sun, Yu, et al.
Published: (2024)
by: Sun, Yu, et al.
Published: (2024)
Learning reveals invisible structure in low-rank RNNs
by: Ger, Yoav, et al.
Published: (2026)
by: Ger, Yoav, et al.
Published: (2026)
Does Transformer Interpretability Transfer to RNNs?
by: Paulo, Gonçalo, et al.
Published: (2024)
by: Paulo, Gonçalo, et al.
Published: (2024)
Paradoxical noise preference in RNNs
by: Eckstein, Noah, et al.
Published: (2026)
by: Eckstein, Noah, et al.
Published: (2026)
Meta Additive Model: Interpretable Sparse Learning With Auto Weighting
by: Zhang, Xuelin, et al.
Published: (2026)
by: Zhang, Xuelin, et al.
Published: (2026)
Compositional Reasoning with Transformers, RNNs, and Chain of Thought
by: Yehudai, Gilad, et al.
Published: (2025)
by: Yehudai, Gilad, et al.
Published: (2025)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
by: Pöppel, Korbinian, et al.
Published: (2024)
by: Pöppel, Korbinian, et al.
Published: (2024)
Balanced Direction from Multifarious Choices: Arithmetic Meta-Learning for Domain Generalization
by: Wang, Xiran, et al.
Published: (2025)
by: Wang, Xiran, et al.
Published: (2025)
Renaissance of RNNs in Streaming Clinical Time Series: Compact Recurrence Remains Competitive with Transformers
by: Tong, Ran, et al.
Published: (2025)
by: Tong, Ran, et al.
Published: (2025)
Detecting Invariant Manifolds in ReLU-Based RNNs
by: Eisenmann, Lukas, et al.
Published: (2025)
by: Eisenmann, Lukas, et al.
Published: (2025)
Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications
by: Brandoit, Julien, et al.
Published: (2026)
by: Brandoit, Julien, et al.
Published: (2026)
TempoPFN: Synthetic Pre-training of Linear RNNs for Zero-shot Time Series Forecasting
by: Moroshan, Vladyslav, et al.
Published: (2025)
by: Moroshan, Vladyslav, et al.
Published: (2025)
M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling
by: Mishra, Mayank, et al.
Published: (2026)
by: Mishra, Mayank, et al.
Published: (2026)
A Framework for Predicting the Impact of Game Balance Changes through Meta Discovery
by: Saravanan, Akash, et al.
Published: (2024)
by: Saravanan, Akash, et al.
Published: (2024)
Linear RNNs for autoregressive generation of long music samples
by: Szewczyk, Konrad, et al.
Published: (2025)
by: Szewczyk, Konrad, et al.
Published: (2025)
Provable Benefits of In-Tool Learning for Large Language Models
by: Houliston, Sam, et al.
Published: (2025)
by: Houliston, Sam, et al.
Published: (2025)
Provable Benefit of Cutout and CutMix for Feature Learning
by: Oh, Junsoo, et al.
Published: (2024)
by: Oh, Junsoo, et al.
Published: (2024)
Graph Coordinates and Conventional Neural Networks -- An Alternative for Graph Neural Networks
by: Qin, Zheyi, et al.
Published: (2023)
by: Qin, Zheyi, et al.
Published: (2023)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
by: Hartman, Max, et al.
Published: (2025)
by: Hartman, Max, et al.
Published: (2025)
Dual-Balancing for Multi-Task Learning
by: Lin, Baijiong, et al.
Published: (2023)
by: Lin, Baijiong, et al.
Published: (2023)
Integrating Meta-Features with Knowledge Graph Embeddings for Meta-Learning
by: Klironomos, Antonis, et al.
Published: (2026)
by: Klironomos, Antonis, et al.
Published: (2026)
Set-based Meta-Interpolation for Few-Task Meta-Learning
by: Lee, Seanie, et al.
Published: (2022)
by: Lee, Seanie, et al.
Published: (2022)
Preference-based Reinforcement Learning beyond Pairwise Comparisons: Benefits of Multiple Options
by: Lee, Joongkyu, et al.
Published: (2025)
by: Lee, Joongkyu, et al.
Published: (2025)
A Framework for Quantifying How Pre-Training and Context Benefit In-Context Learning
by: Song, Bingqing, et al.
Published: (2025)
by: Song, Bingqing, et al.
Published: (2025)
The Disparate Benefits of Deep Ensembles
by: Schweighofer, Kajetan, et al.
Published: (2024)
by: Schweighofer, Kajetan, et al.
Published: (2024)
A Metric for the Balance of Information in Graph Learning
by: Davies, Alex O., et al.
Published: (2025)
by: Davies, Alex O., et al.
Published: (2025)
Any-Way Meta Learning
by: Lee, Junhoo, et al.
Published: (2024)
by: Lee, Junhoo, et al.
Published: (2024)
Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent
by: Hoppmann, Björn, et al.
Published: (2026)
by: Hoppmann, Björn, et al.
Published: (2026)
Data Whitening Improves Sparse Autoencoder Learning
by: Saraswatula, Ashwin, et al.
Published: (2025)
by: Saraswatula, Ashwin, et al.
Published: (2025)
Similar Items
-
Rethinking the Relationship between Recurrent and Non-Recurrent Neural Networks: A Study in Sparsity
by: Hershey, Quincy, et al.
Published: (2024) -
Principled Curriculum Learning using Parameter Continuation Methods
by: Pathak, Harsh Nilesh, et al.
Published: (2025) -
Solo Connection: A Parameter Efficient Fine-Tuning Technique for Transformers
by: Pathak, Harsh Nilesh, et al.
Published: (2025) -
Implicit Language Models are RNNs: Balancing Parallelization and Expressivity
by: Schöne, Mark, et al.
Published: (2025) -
HadamRNN: Binary and Sparse Ternary Orthogonal RNNs
by: Foucault, Armand, et al.
Published: (2025)