Fast Training of Recurrent Neural Networks with Stationary State Feedbacks
Fuente:
arXiv
Saved in:
| Main Authors: | Caillon, Paul, Fagnou, Erwan, Allauzen, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Forward Only Learning for Orthogonal Neural Networks of any Depth
by: Caillon, Paul, et al.
Published: (2025)
by: Caillon, Paul, et al.
Published: (2025)
Accelerated Training through Iterative Gradient Propagation Along the Residual Path
by: Fagnou, Erwan, et al.
Published: (2025)
by: Fagnou, Erwan, et al.
Published: (2025)
Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexity
by: Zhao, Hangyue, et al.
Published: (2026)
by: Zhao, Hangyue, et al.
Published: (2026)
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
by: Fagnou, Erwan, et al.
Published: (2026)
by: Fagnou, Erwan, et al.
Published: (2026)
Chain and Causal Attention for Efficient Entity Tracking
by: Fagnou, Erwan, et al.
Published: (2024)
by: Fagnou, Erwan, et al.
Published: (2024)
Bridging the Theoretical Gap in Randomized Smoothing
by: Delattre, Blaise, et al.
Published: (2025)
by: Delattre, Blaise, et al.
Published: (2025)
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
by: Colagrande, Alex, et al.
Published: (2025)
by: Colagrande, Alex, et al.
Published: (2025)
Limits of Resolution Equivariance in Fourier Neural Operators
by: Colagrande, Alex, et al.
Published: (2026)
by: Colagrande, Alex, et al.
Published: (2026)
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
by: Galashov, Alexandre, et al.
Published: (2024)
by: Galashov, Alexandre, et al.
Published: (2024)
DASH: Warm-Starting Neural Network Training in Stationary Settings without Loss of Plasticity
by: Shin, Baekrok, et al.
Published: (2024)
by: Shin, Baekrok, et al.
Published: (2024)
Feed-Forward Optimization With Delayed Feedback for Neural Network Training
by: Flügel, Katharina, et al.
Published: (2023)
by: Flügel, Katharina, et al.
Published: (2023)
Understanding the differences in Foundation Models: Attention, State Space Models, and Recurrent Neural Networks
by: Sieber, Jerome, et al.
Published: (2024)
by: Sieber, Jerome, et al.
Published: (2024)
Interactive Training: Feedback-Driven Neural Network Optimization
by: Zhang, Wentao, et al.
Published: (2025)
by: Zhang, Wentao, et al.
Published: (2025)
On the MIA Vulnerability Gap Between Private GANs and Diffusion Models
by: Sebag, Ilana, et al.
Published: (2025)
by: Sebag, Ilana, et al.
Published: (2025)
Generative System Dynamics in Recurrent Neural Networks
by: Casoni, Michele, et al.
Published: (2025)
by: Casoni, Michele, et al.
Published: (2025)
BARNN: A Bayesian Autoregressive and Recurrent Neural Network
by: Coscia, Dario, et al.
Published: (2025)
by: Coscia, Dario, et al.
Published: (2025)
A Scalable Hybrid Training Approach for Recurrent Spiking Neural Networks
by: Baronig, Maximilian, et al.
Published: (2025)
by: Baronig, Maximilian, et al.
Published: (2025)
Modular Boundaries in Recurrent Neural Networks
by: Tanner, Jacob, et al.
Published: (2023)
by: Tanner, Jacob, et al.
Published: (2023)
Fast-DataShapley: Neural Modeling for Training Data Valuation
by: Sun, Haifeng, et al.
Published: (2025)
by: Sun, Haifeng, et al.
Published: (2025)
Fast Training of Sinusoidal Neural Fields via Scaling Initialization
by: Yeom, Taesun, et al.
Published: (2024)
by: Yeom, Taesun, et al.
Published: (2024)
Training Spiking Neural Networks via Augmented Direct Feedback Alignment
by: Zhang, Yongbo, et al.
Published: (2024)
by: Zhang, Yongbo, et al.
Published: (2024)
Spectral Theory for Edge Pruning in Asynchronous Recurrent Graph Neural Networks
by: Bessone, Nicolas
Published: (2025)
by: Bessone, Nicolas
Published: (2025)
ENOT: Expectile Regularization for Fast and Accurate Training of Neural Optimal Transport
by: Buzun, Nazar, et al.
Published: (2024)
by: Buzun, Nazar, et al.
Published: (2024)
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization
by: Chaubard, Francois, et al.
Published: (2025)
by: Chaubard, Francois, et al.
Published: (2025)
Kalman Filter for Online Classification of Non-Stationary Data
by: Titsias, Michalis K., et al.
Published: (2023)
by: Titsias, Michalis K., et al.
Published: (2023)
GARNN: An Interpretable Graph Attentive Recurrent Neural Network for Predicting Blood Glucose Levels via Multivariate Time Series
by: Piao, Chengzhe, et al.
Published: (2024)
by: Piao, Chengzhe, et al.
Published: (2024)
Recurrent Aggregators in Neural Algorithmic Reasoning
by: Xu, Kaijia, et al.
Published: (2024)
by: Xu, Kaijia, et al.
Published: (2024)
Scale-Consistent State-Space Dynamics via Fractal of Stationary Transformations
by: Yu, Geunhyeok, et al.
Published: (2026)
by: Yu, Geunhyeok, et al.
Published: (2026)
Identifying Information-Transfer Nodes in a Recurrent Neural Network Reveals Dynamic Representations
by: Hintze, Arend, et al.
Published: (2025)
by: Hintze, Arend, et al.
Published: (2025)
IRNN: Innovation-driven Recurrent Neural Network for Time-Series Data Modeling and Prediction
by: Zhou, Yifan, et al.
Published: (2025)
by: Zhou, Yifan, et al.
Published: (2025)
Uncertainty-Aware Deep Attention Recurrent Neural Network for Heterogeneous Time Series Imputation
by: Qian, Linglong, et al.
Published: (2024)
by: Qian, Linglong, et al.
Published: (2024)
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
by: Yin, Bojian, et al.
Published: (2024)
by: Yin, Bojian, et al.
Published: (2024)
Probabilistic Multi-Regional Solar Power Forecasting with Any-Quantile Recurrent Neural Networks
by: Smyl, Slawek, et al.
Published: (2026)
by: Smyl, Slawek, et al.
Published: (2026)
Signal-Adaptive Trust Regions for Gradient-Free Optimization of Recurrent Spiking Neural Networks
by: Li, Jinhao, et al.
Published: (2026)
by: Li, Jinhao, et al.
Published: (2026)
Deep Neural Networks for Predicting Recurrence and Survival in Patients with Esophageal Cancer After Surgery
by: Zheng, Yuhan, et al.
Published: (2024)
by: Zheng, Yuhan, et al.
Published: (2024)
Z-Error Loss for Training Neural Networks
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
Energy Consumption in Parallel Neural Network Training
by: Huber, Philipp, et al.
Published: (2025)
by: Huber, Philipp, et al.
Published: (2025)
Deep Residual Echo State Networks: exploring residual orthogonal connections in untrained Recurrent Neural Networks
by: Pinna, Matteo, et al.
Published: (2025)
by: Pinna, Matteo, et al.
Published: (2025)
Training Neural Networks for Modularity aids Interpretability
by: Golechha, Satvik, et al.
Published: (2024)
by: Golechha, Satvik, et al.
Published: (2024)
Automatic Stability and Recovery for Neural Network Training
by: Or, Barak
Published: (2026)
by: Or, Barak
Published: (2026)
Similar Items
-
Forward Only Learning for Orthogonal Neural Networks of any Depth
by: Caillon, Paul, et al.
Published: (2025) -
Accelerated Training through Iterative Gradient Propagation Along the Residual Path
by: Fagnou, Erwan, et al.
Published: (2025) -
Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexity
by: Zhao, Hangyue, et al.
Published: (2026) -
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
by: Fagnou, Erwan, et al.
Published: (2026) -
Chain and Causal Attention for Efficient Entity Tracking
by: Fagnou, Erwan, et al.
Published: (2024)