Towards Scalable and Stable Parallelization of Nonlinear RNNs
Fuente:
arXiv
Saved in:
| Main Authors: | Gonzalez, Xavier, Warrington, Andrew, Smith, Jimmy T. H., Linderman, Scott W. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
Learning Stochastic Nonlinear Dynamics with Embedded Latent Transfer Operators
by: Ke, Naichang, et al.
Published: (2025)
by: Ke, Naichang, et al.
Published: (2025)
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
by: Saxena, Udit
Published: (2025)
by: Saxena, Udit
Published: (2025)
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
by: Moghadasi, Mahdi Naser, et al.
Published: (2026)
by: Moghadasi, Mahdi Naser, et al.
Published: (2026)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
BOND: License to Train with Black-Box Functions
by: Clark, Andrew, et al.
Published: (2025)
by: Clark, Andrew, et al.
Published: (2025)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
by: Atad, Ido Andrew, et al.
Published: (2026)
by: Atad, Ido Andrew, et al.
Published: (2026)
Modeling Nonlinear Oscillator Networks Using Physics-Informed Hybrid Reservoir Computing
by: Shannon, Andrew, et al.
Published: (2024)
by: Shannon, Andrew, et al.
Published: (2024)
Newtonian and Lagrangian Neural Networks: A Comparison Towards Efficient Inverse Dynamics Identification
by: Trinh, Minh, et al.
Published: (2025)
by: Trinh, Minh, et al.
Published: (2025)
Towards Systematic Generalization for Power Grid Optimization Problems
by: Memon, Zeeshan, et al.
Published: (2026)
by: Memon, Zeeshan, et al.
Published: (2026)
Expressive Value Learning for Scalable Offline Reinforcement Learning
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
2Mamba2Furious: Linear in Complexity, Competitive in Accuracy
by: Mongaras, Gabriel, et al.
Published: (2026)
by: Mongaras, Gabriel, et al.
Published: (2026)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
by: Rozonoyer, Benjamin, et al.
Published: (2026)
by: Rozonoyer, Benjamin, et al.
Published: (2026)
Towards General Negotiation Strategies with End-to-End Reinforcement Learning
by: Renting, Bram M., et al.
Published: (2024)
by: Renting, Bram M., et al.
Published: (2024)
FluidWorld: Reaction-Diffusion Dynamics as a Predictive Substrate for World Models
by: Polly, Fabien
Published: (2026)
by: Polly, Fabien
Published: (2026)
I-GLIDE: Input Groups for Latent Health Indicators in Degradation Estimation
by: Thil, Lucas, et al.
Published: (2025)
by: Thil, Lucas, et al.
Published: (2025)
Towards A Flexible Accuracy-Oriented Deep Learning Module Inference Latency Prediction Framework for Adaptive Optimization Algorithms
by: Shen, Jingran, et al.
Published: (2023)
by: Shen, Jingran, et al.
Published: (2023)
Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis
by: Suetake, Yamato, et al.
Published: (2026)
by: Suetake, Yamato, et al.
Published: (2026)
Towards geological inference with process-based and deep generative modeling, part 1: training on fluvial deposits
by: Rongier, Guillaume, et al.
Published: (2025)
by: Rongier, Guillaume, et al.
Published: (2025)
How Reliable and Stable are Explanations of XAI Methods?
by: Ribeiro, José, et al.
Published: (2024)
by: Ribeiro, José, et al.
Published: (2024)
Towards geological inference with process-based and deep generative modeling, part 2: inversion of fluvial deposits and latent-space disentanglement
by: Rongier, Guillaume, et al.
Published: (2025)
by: Rongier, Guillaume, et al.
Published: (2025)
Universal Approximation of Continuous Functionals on Compact Subsets via Linear Measurements and Scalar Nonlinearities
by: Krylov, Andrey, et al.
Published: (2026)
by: Krylov, Andrey, et al.
Published: (2026)
How Many Ratings per Item are Necessary for Reliable Significance Testing?
by: Homan, Christopher, et al.
Published: (2024)
by: Homan, Christopher, et al.
Published: (2024)
Potential-Based Reward Shaping For Intrinsic Motivation
by: Forbes, Grant C., et al.
Published: (2024)
by: Forbes, Grant C., et al.
Published: (2024)
How to Boost Any Loss Function
by: Nock, Richard, et al.
Published: (2024)
by: Nock, Richard, et al.
Published: (2024)
Interpretable Multi-View Clustering
by: Jiang, Mudi, et al.
Published: (2024)
by: Jiang, Mudi, et al.
Published: (2024)
The Bayesian Confidence (BACON) Estimator for Deep Neural Networks
by: Kee, Patrick D., et al.
Published: (2024)
by: Kee, Patrick D., et al.
Published: (2024)
Pre-Ictal Seizure Prediction Using Personalized Deep Learning
by: Jaddu, Shriya, et al.
Published: (2024)
by: Jaddu, Shriya, et al.
Published: (2024)
xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar Memories
by: Kraus, Maurice, et al.
Published: (2024)
by: Kraus, Maurice, et al.
Published: (2024)
Securing Reliability: A Brief Overview on Enhancing In-Context Learning for Foundation Models
by: Huang, Yunpeng, et al.
Published: (2024)
by: Huang, Yunpeng, et al.
Published: (2024)
Representation learning with CGAN for casual inference
by: Weng, Zhaotian, et al.
Published: (2024)
by: Weng, Zhaotian, et al.
Published: (2024)
Boosting gets full Attention for Relational Learning
by: Guillame-Bert, Mathieu, et al.
Published: (2024)
by: Guillame-Bert, Mathieu, et al.
Published: (2024)
Data-Incremental Continual Offline Reinforcement Learning
by: Gai, Sibo, et al.
Published: (2024)
by: Gai, Sibo, et al.
Published: (2024)
Adaptive Epsilon Adversarial Training for Robust Gravitational Wave Parameter Estimation Using Normalizing Flows
by: Yang, Yiqian, et al.
Published: (2024)
by: Yang, Yiqian, et al.
Published: (2024)
Normalization Layer Per-Example Gradients are Sufficient to Predict Gradient Noise Scale in Transformers
by: Gray, Gavia, et al.
Published: (2024)
by: Gray, Gavia, et al.
Published: (2024)
Potential-Based Intrinsic Motivation: Preserving Optimality With Complex, Non-Markovian Shaping Rewards
by: Forbes, Grant C., et al.
Published: (2024)
by: Forbes, Grant C., et al.
Published: (2024)
CPT: Competence-progressive Training Strategy for Few-shot Node Classification
by: Yan, Qilong, et al.
Published: (2024)
by: Yan, Qilong, et al.
Published: (2024)
Learning Useful Representations of Recurrent Neural Network Weight Matrices
by: Herrmann, Vincent, et al.
Published: (2024)
by: Herrmann, Vincent, et al.
Published: (2024)
Similar Items
-
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025) -
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024) -
Learning Stochastic Nonlinear Dynamics with Embedded Latent Transfer Operators
by: Ke, Naichang, et al.
Published: (2025) -
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
by: Saxena, Udit
Published: (2025) -
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
by: Moghadasi, Mahdi Naser, et al.
Published: (2026)