Saved in:
| Main Authors: | Price, Ilan, Ball, Nicholas Daultry, Lam, Samuel C. H., Jones, Adam C., Tanner, Jared |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.16184 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Theory of Minimal Weight Perturbations in Deep Networks and its Applications for Low-Rank Activated Backdoor Attacks
by: Evans, Bethan, et al.
Published: (2026)
by: Evans, Bethan, et al.
Published: (2026)
How Controlling the Variance can Improve Training Stability of Sparsely Activated DNNs and CNNs
by: Dent, Emily, et al.
Published: (2026)
by: Dent, Emily, et al.
Published: (2026)
On the Hardness of Training Deep Neural Networks Discretely
by: Doron-Arad, Ilan
Published: (2024)
by: Doron-Arad, Ilan
Published: (2024)
Beyond IID weights: sparse and low-rank deep Neural Networks are also Gaussian Processes
by: Nait-Saada, Thiziri, et al.
Published: (2023)
by: Nait-Saada, Thiziri, et al.
Published: (2023)
SPADE: Sparsity-Guided Debugging for Deep Neural Networks
by: Moakhar, Arshia Soltani, et al.
Published: (2023)
by: Moakhar, Arshia Soltani, et al.
Published: (2023)
Chordal Sparsity for Lipschitz Constant Estimation of Deep Neural Networks
by: Xue, Anton, et al.
Published: (2022)
by: Xue, Anton, et al.
Published: (2022)
Effects of Initialization Biases on Deep Neural Network Training Dynamics
by: Pellegrino, Nicholas, et al.
Published: (2025)
by: Pellegrino, Nicholas, et al.
Published: (2025)
Investigating Sparsity in Recurrent Neural Networks
by: Darji, Harshil
Published: (2024)
by: Darji, Harshil
Published: (2024)
Why ReLU? A Bit-Model Dichotomy for Deep Network Training
by: Doron-Arad, Ilan, et al.
Published: (2026)
by: Doron-Arad, Ilan, et al.
Published: (2026)
Optimal Initialization in Depth: Lyapunov Initialization and Limit Theorems for Deep Leaky ReLU Networks
by: Kogler, Constantin, et al.
Published: (2026)
by: Kogler, Constantin, et al.
Published: (2026)
Mind the Gap: a Spectral Analysis of Rank Collapse and Signal Propagation in Attention Layers
by: Saada, Thiziri Nait, et al.
Published: (2024)
by: Saada, Thiziri Nait, et al.
Published: (2024)
Exploiting Subgradient Sparsity in Max-Plus Neural Networks
by: Enaieh, Ikhlas, et al.
Published: (2026)
by: Enaieh, Ikhlas, et al.
Published: (2026)
Online Optimisation of Machine Learning Collision Models to Accelerate Direct Molecular Simulation of Rarefied Gas Flows
by: Ball, Nicholas Daultry, et al.
Published: (2024)
by: Ball, Nicholas Daultry, et al.
Published: (2024)
Approximate Multiplier Induced Error Propagation in Deep Neural Networks
by: Alahakoon, A. M. H. H., et al.
Published: (2025)
by: Alahakoon, A. M. H. H., et al.
Published: (2025)
Optimal Condition for Initialization Variance in Deep Neural Networks: An SGD Dynamics Perspective
by: Horii, Hiroshi, et al.
Published: (2025)
by: Horii, Hiroshi, et al.
Published: (2025)
Optimized Weight Initialization on the Stiefel Manifold for Deep ReLU Neural Networks
by: Lee, Hyungu, et al.
Published: (2025)
by: Lee, Hyungu, et al.
Published: (2025)
Weight Initialization and Variance Dynamics in Deep Neural Networks and Large Language Models
by: Han, Yankun
Published: (2025)
by: Han, Yankun
Published: (2025)
Sparsity-Aware Communication for Distributed Graph Neural Network Training
by: Mukhodopadhyay, Ujjaini, et al.
Published: (2025)
by: Mukhodopadhyay, Ujjaini, et al.
Published: (2025)
Early Directional Convergence in Deep Homogeneous Neural Networks for Small Initializations
by: Kumar, Akshay, et al.
Published: (2024)
by: Kumar, Akshay, et al.
Published: (2024)
Chordal Sparsity for SDP-based Neural Network Verification
by: Xue, Anton, et al.
Published: (2022)
by: Xue, Anton, et al.
Published: (2022)
Parallel Algorithms for Exact Enumeration of Deep Neural Network Activation Regions
by: Drammis, Sabrina, et al.
Published: (2024)
by: Drammis, Sabrina, et al.
Published: (2024)
Exploring and Improving Initialization for Deep Graph Neural Networks: A Signal Propagation Perspective
by: Wang, Senmiao, et al.
Published: (2025)
by: Wang, Senmiao, et al.
Published: (2025)
Sparsity-Induced Global Matrix Autoregressive Model with Auxiliary Network Data
by: Wu, Sanyou, et al.
Published: (2025)
by: Wu, Sanyou, et al.
Published: (2025)
On Unbalanced Optimal Transport: Gradient Methods, Sparsity and Approximation Error
by: Nguyen, Quang Minh, et al.
Published: (2022)
by: Nguyen, Quang Minh, et al.
Published: (2022)
Model Merging by Output-Space Projection
by: Evans, Bethan, et al.
Published: (2026)
by: Evans, Bethan, et al.
Published: (2026)
Hamiltonian Monte Carlo on ReLU Neural Networks is Inefficient
by: Dinh, Vu C., et al.
Published: (2024)
by: Dinh, Vu C., et al.
Published: (2024)
Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning
by: Ma, Guozheng, et al.
Published: (2025)
by: Ma, Guozheng, et al.
Published: (2025)
Universal Properties of Activation Sparsity in Modern Large Language Models
by: Szatkowski, Filip, et al.
Published: (2025)
by: Szatkowski, Filip, et al.
Published: (2025)
Principal Components for Neural Network Initialization
by: Phan, Nhan, et al.
Published: (2025)
by: Phan, Nhan, et al.
Published: (2025)
SAUC: Sparsity-Aware Uncertainty Calibration for Spatiotemporal Prediction with Graph Neural Networks
by: Zhuang, Dingyi, et al.
Published: (2024)
by: Zhuang, Dingyi, et al.
Published: (2024)
DQA: An Efficient Method for Deep Quantization of Deep Neural Network Activations
by: Hu, Wenhao, et al.
Published: (2024)
by: Hu, Wenhao, et al.
Published: (2024)
Activation Bottleneck: Sigmoidal Neural Networks Cannot Forecast a Straight Line
by: Toller, Maximilian, et al.
Published: (2024)
by: Toller, Maximilian, et al.
Published: (2024)
Post-Training Statistical Calibration for Higher Activation Sparsity
by: Chua, Vui Seng, et al.
Published: (2024)
by: Chua, Vui Seng, et al.
Published: (2024)
Joint Training Across Multiple Activation Sparsity Regimes
by: Wang, Haotian
Published: (2026)
by: Wang, Haotian
Published: (2026)
Towards the Connection between Activation Sparsity and Flat Minima
by: Peng, Ze, et al.
Published: (2026)
by: Peng, Ze, et al.
Published: (2026)
From Activation to Initialization: Scaling Insights for Optimizing Neural Fields
by: Saratchandran, Hemanth, et al.
Published: (2024)
by: Saratchandran, Hemanth, et al.
Published: (2024)
Accelerating Linear Recurrent Neural Networks for the Edge with Unstructured Sparsity
by: Pierro, Alessandro, et al.
Published: (2025)
by: Pierro, Alessandro, et al.
Published: (2025)
Semiring Activation in Neural Networks
by: Smets, Bart M. N., et al.
Published: (2024)
by: Smets, Bart M. N., et al.
Published: (2024)
Neighbor-Sampling Based Momentum Stochastic Methods for Training Graph Neural Networks
by: Noel, Molly, et al.
Published: (2025)
by: Noel, Molly, et al.
Published: (2025)
A Proximal Operator for Inducing 2:4-Sparsity
by: Kübler, Jonas M, et al.
Published: (2025)
by: Kübler, Jonas M, et al.
Published: (2025)
Similar Items
-
Theory of Minimal Weight Perturbations in Deep Networks and its Applications for Low-Rank Activated Backdoor Attacks
by: Evans, Bethan, et al.
Published: (2026) -
How Controlling the Variance can Improve Training Stability of Sparsely Activated DNNs and CNNs
by: Dent, Emily, et al.
Published: (2026) -
On the Hardness of Training Deep Neural Networks Discretely
by: Doron-Arad, Ilan
Published: (2024) -
Beyond IID weights: sparse and low-rank deep Neural Networks are also Gaussian Processes
by: Nait-Saada, Thiziri, et al.
Published: (2023) -
SPADE: Sparsity-Guided Debugging for Deep Neural Networks
by: Moakhar, Arshia Soltani, et al.
Published: (2023)