An Adaptive and Stability-Promoting Layerwise Training Approach for Sparse Deep Neural Network Architecture
Fuente:
arXiv
Saved in:
| Main Authors: | Krishnanunni, C G, Bui-Thanh, Tan |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Momentum and Nonlinear Damping for Neural Network Training
by: Karoni, Aikaterini, et al.
Published: (2026)
by: Karoni, Aikaterini, et al.
Published: (2026)
A New Look at the Ensemble Kalman Filter for Inverse Problems: Duality, Non-Asymptotic Analysis and Convergence Acceleration
by: Krishnanunni, C G, et al.
Published: (2026)
by: Krishnanunni, C G, et al.
Published: (2026)
Multi-Objective Linear Ensembles for Robust and Sparse Training of Few-Bit Neural Networks
by: Bernardelli, Ambrogio Maria, et al.
Published: (2022)
by: Bernardelli, Ambrogio Maria, et al.
Published: (2022)
Optimization over Trained (and Sparse) Neural Networks: A Surrogate within a Surrogate
by: Pham, Hung, et al.
Published: (2025)
by: Pham, Hung, et al.
Published: (2025)
Convergence Analysis for Learning Orthonormal Deep Linear Neural Networks
by: Qin, Zhen, et al.
Published: (2023)
by: Qin, Zhen, et al.
Published: (2023)
Regularized Gradient Clipping Provably Trains Wide and Deep Neural Networks
by: Tucat, Matteo, et al.
Published: (2024)
by: Tucat, Matteo, et al.
Published: (2024)
Enhancing Stability of Physics-Informed Neural Network Training Through Saddle-Point Reformulation
by: Bylinkin, Dmitry, et al.
Published: (2025)
by: Bylinkin, Dmitry, et al.
Published: (2025)
A Recovery Guarantee for Sparse Neural Networks
by: Fridovich-Keil, Sara, et al.
Published: (2025)
by: Fridovich-Keil, Sara, et al.
Published: (2025)
A Unified Representation of Neural Networks Architectures
by: Prieur, Christophe, et al.
Published: (2025)
by: Prieur, Christophe, et al.
Published: (2025)
Regularized Adaptive Momentum Dual Averaging with an Efficient Inexact Subproblem Solver for Training Structured Neural Network
by: Huang, Zih-Syuan, et al.
Published: (2024)
by: Huang, Zih-Syuan, et al.
Published: (2024)
Sparse-ProxSkip: Accelerated Sparse-to-Sparse Training in Federated Learning
by: Meinhardt, Georg, et al.
Published: (2024)
by: Meinhardt, Georg, et al.
Published: (2024)
Tuning-Free Structured Sparse PCA via Deep Unfolding Networks
by: Chen, Long, et al.
Published: (2025)
by: Chen, Long, et al.
Published: (2025)
Verifying Properties of Binary Neural Networks Using Sparse Polynomial Optimization
by: Yang, Jianting, et al.
Published: (2024)
by: Yang, Jianting, et al.
Published: (2024)
Understanding Deep Representation Learning via Layerwise Feature Compression and Discrimination
by: Wang, Peng, et al.
Published: (2023)
by: Wang, Peng, et al.
Published: (2023)
Tight Robustness Certificates and Wasserstein Distributional Attacks for Deep Neural Networks
by: Le, Bach C., et al.
Published: (2025)
by: Le, Bach C., et al.
Published: (2025)
Relaxation-Informed Training of Neural Network Surrogate Models
by: Tsay, Calvin
Published: (2026)
by: Tsay, Calvin
Published: (2026)
Wide Neural Networks Trained with Weight Decay Provably Exhibit Neural Collapse
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
by: Cao, John, et al.
Published: (2025)
by: Cao, John, et al.
Published: (2025)
Compression-aware Training of Neural Networks using Frank-Wolfe
by: Zimmer, Max, et al.
Published: (2022)
by: Zimmer, Max, et al.
Published: (2022)
Deep Operator Neural Network Model Predictive Control
by: de Jong, Thomas Oliver, et al.
Published: (2025)
by: de Jong, Thomas Oliver, et al.
Published: (2025)
SGD with Partial Hessian for Deep Neural Networks Optimization
by: Sun, Ying, et al.
Published: (2024)
by: Sun, Ying, et al.
Published: (2024)
Improved Scalable Lipschitz Bounds for Deep Neural Networks
by: Syed, Usman, et al.
Published: (2025)
by: Syed, Usman, et al.
Published: (2025)
Optimization Over Trained Neural Networks: Taking a Relaxing Walk
by: Tong, Jiatai, et al.
Published: (2024)
by: Tong, Jiatai, et al.
Published: (2024)
Towards Guided Descent: Optimization Algorithms for Training Neural Networks At Scale
by: Nagwekar, Ansh
Published: (2025)
by: Nagwekar, Ansh
Published: (2025)
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
Neural Network Training Techniques Regularize Optimization Trajectory: An Empirical Study
by: Chen, Cheng, et al.
Published: (2020)
by: Chen, Cheng, et al.
Published: (2020)
Benchmarking Stochastic Approximation Algorithms for Fairness-Constrained Training of Deep Neural Networks
by: Kliachkin, Andrii, et al.
Published: (2025)
by: Kliachkin, Andrii, et al.
Published: (2025)
A Convexity-dependent Two-Phase Training Algorithm for Deep Neural Networks
by: Hrycej, Tomas, et al.
Published: (2025)
by: Hrycej, Tomas, et al.
Published: (2025)
Stability and Performance Analysis of Discrete-Time ReLU Recurrent Neural Networks
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
A Non-Monotone Preconditioned Trust-Region Method for Neural Network Training
by: Angino, Andrea, et al.
Published: (2026)
by: Angino, Andrea, et al.
Published: (2026)
Dual Natural Gradient Descent for Scalable Training of Physics-Informed Neural Networks
by: Jnini, Anas, et al.
Published: (2025)
by: Jnini, Anas, et al.
Published: (2025)
Neighbor-Sampling Based Momentum Stochastic Methods for Training Graph Neural Networks
by: Noel, Molly, et al.
Published: (2025)
by: Noel, Molly, et al.
Published: (2025)
Sparse Deep Learning Models with the $\ell_1$ Regularization
by: Shen, Lixin, et al.
Published: (2024)
by: Shen, Lixin, et al.
Published: (2024)
Early Directional Convergence in Deep Homogeneous Neural Networks for Small Initializations
by: Kumar, Akshay, et al.
Published: (2024)
by: Kumar, Akshay, et al.
Published: (2024)
Weight-Parameterization in Continuous Time Deep Neural Networks for Surrogate Modeling
by: Rosso, Haley, et al.
Published: (2025)
by: Rosso, Haley, et al.
Published: (2025)
Parameter-Adaptive Approximate MPC: Tuning Neural-Network Controllers without Retraining
by: Hose, Henrik, et al.
Published: (2024)
by: Hose, Henrik, et al.
Published: (2024)
Contraction-Guided Adaptive Partitioning for Reachability Analysis of Neural Network Controlled Systems
by: Harapanahalli, Akash, et al.
Published: (2023)
by: Harapanahalli, Akash, et al.
Published: (2023)
Does Weight Decay Enhance Training Stability?
by: Saether, Marius, et al.
Published: (2026)
by: Saether, Marius, et al.
Published: (2026)
Sparse Transformer Architectures via Regularized Wasserstein Proximal Operator with $L_1$ Prior
by: Han, Fuqun, et al.
Published: (2025)
by: Han, Fuqun, et al.
Published: (2025)
Mean-Field Limits for Two-Layer Neural Networks Trained with Consensus-Based Optimization
by: De Deyn, William, et al.
Published: (2025)
by: De Deyn, William, et al.
Published: (2025)
Similar Items
-
Adaptive Momentum and Nonlinear Damping for Neural Network Training
by: Karoni, Aikaterini, et al.
Published: (2026) -
A New Look at the Ensemble Kalman Filter for Inverse Problems: Duality, Non-Asymptotic Analysis and Convergence Acceleration
by: Krishnanunni, C G, et al.
Published: (2026) -
Multi-Objective Linear Ensembles for Robust and Sparse Training of Few-Bit Neural Networks
by: Bernardelli, Ambrogio Maria, et al.
Published: (2022) -
Optimization over Trained (and Sparse) Neural Networks: A Surrogate within a Surrogate
by: Pham, Hung, et al.
Published: (2025) -
Convergence Analysis for Learning Orthonormal Deep Linear Neural Networks
by: Qin, Zhen, et al.
Published: (2023)