LION-DG: Layer-Informed Initialization with Deep Gradient Protocols for Accelerated Neural Network Training
Fuente:
arXiv
Saved in:
| Main Author: | Kim, Hyunjun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometric Regularization in Mixture-of-Experts: The Disconnect Between Weights and Activations
by: Kim, Hyunjun
Published: (2026)
by: Kim, Hyunjun
Published: (2026)
Soft Deterministic Policy Gradient with Gaussian Smoothing
by: Na, Hyunjun, et al.
Published: (2026)
by: Na, Hyunjun, et al.
Published: (2026)
HOLOGRAPH: Active Causal Discovery via Sheaf-Theoretic Alignment of Large Language Model Priors
by: Kim, Hyunjun
Published: (2025)
by: Kim, Hyunjun
Published: (2025)
R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes
by: Na, Hyunjun, et al.
Published: (2026)
by: Na, Hyunjun, et al.
Published: (2026)
DeepDefense: Layer-Wise Gradient-Feature Alignment for Building Robust Neural Networks
by: Lin, Ci, et al.
Published: (2025)
by: Lin, Ci, et al.
Published: (2025)
GradINN: Gradient Informed Neural Network
by: Aglietti, Filippo, et al.
Published: (2024)
by: Aglietti, Filippo, et al.
Published: (2024)
Self-Abstraction Learning for Effective and Stable Training of Deep Neural Networks
by: Cho, Wonyong, et al.
Published: (2026)
by: Cho, Wonyong, et al.
Published: (2026)
Accelerating Storage-Based Training for Graph Neural Networks
by: Jang, Myung-Hwan, et al.
Published: (2026)
by: Jang, Myung-Hwan, et al.
Published: (2026)
Gradient-Free Training of Quantized Neural Networks
by: Cohen, Noa, et al.
Published: (2024)
by: Cohen, Noa, et al.
Published: (2024)
Enhancing Deep Learning with Optimized Gradient Descent: Bridging Numerical Methods and Neural Network Training
by: Ma, Yuhan, et al.
Published: (2024)
by: Ma, Yuhan, et al.
Published: (2024)
IDInit: A Universal and Stable Initialization Method for Neural Network Training
by: Pan, Yu, et al.
Published: (2025)
by: Pan, Yu, et al.
Published: (2025)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
by: Li, Kevin, et al.
Published: (2024)
by: Li, Kevin, et al.
Published: (2024)
Proximity-Informed Calibration for Deep Neural Networks
by: Xiong, Miao, et al.
Published: (2023)
by: Xiong, Miao, et al.
Published: (2023)
Layer Embedding Deep Fusion Graph Neural Network
by: Xu, Taihua, et al.
Published: (2026)
by: Xu, Taihua, et al.
Published: (2026)
Provable Acceleration of Nesterov's Accelerated Gradient Method over Heavy Ball Method in Training Over-Parameterized Neural Networks
by: Liu, Xin, et al.
Published: (2022)
by: Liu, Xin, et al.
Published: (2022)
NeuralGrok: Accelerate Grokking by Neural Gradient Transformation
by: Zhou, Xinyu, et al.
Published: (2025)
by: Zhou, Xinyu, et al.
Published: (2025)
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
by: Lee, Hyunwoo, et al.
Published: (2024)
by: Lee, Hyunwoo, et al.
Published: (2024)
Lightweight Geometric Adaptation for Training Physics-Informed Neural Networks
by: An, Kang, et al.
Published: (2026)
by: An, Kang, et al.
Published: (2026)
Improved Training of Physics-Informed Neural Networks with Model Ensembles
by: Haitsiukevich, Katsiaryna, et al.
Published: (2022)
by: Haitsiukevich, Katsiaryna, et al.
Published: (2022)
Q-Newton: Hybrid Quantum-Classical Scheduling for Accelerating Neural Network Training with Newton's Gradient Descent
by: Li, Pingzhi, et al.
Published: (2024)
by: Li, Pingzhi, et al.
Published: (2024)
Symbol Correctness in Deep Neural Networks Containing Symbolic Layers
by: Bembenek, Aaron, et al.
Published: (2024)
by: Bembenek, Aaron, et al.
Published: (2024)
Principal Components for Neural Network Initialization
by: Phan, Nhan, et al.
Published: (2025)
by: Phan, Nhan, et al.
Published: (2025)
DIM: Enforcing Domain-Informed Monotonicity in Deep Neural Networks
by: Salim, Joshua, et al.
Published: (2025)
by: Salim, Joshua, et al.
Published: (2025)
Communication-Efficient Federated Learning with Accelerated Client Gradient
by: Kim, Geeho, et al.
Published: (2022)
by: Kim, Geeho, et al.
Published: (2022)
Fast Training of Sinusoidal Neural Fields via Scaling Initialization
by: Yeom, Taesun, et al.
Published: (2024)
by: Yeom, Taesun, et al.
Published: (2024)
Reconstructing Deep Neural Networks: Unleashing the Optimization Potential of Natural Gradient Descent
by: Liu, Weihua, et al.
Published: (2024)
by: Liu, Weihua, et al.
Published: (2024)
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
Deep Neural Networks as Discrete Dynamical Systems: Implications for Physics-Informed Learning
by: Ganguly, Abhisek, et al.
Published: (2026)
by: Ganguly, Abhisek, et al.
Published: (2026)
Look Every Frame All at Once: Video-Ma$^2$mba for Efficient Long-form Video Understanding with Multi-Axis Gradient Checkpointing
by: Lee, Hosu, et al.
Published: (2024)
by: Lee, Hosu, et al.
Published: (2024)
Grokfast: Accelerated Grokking by Amplifying Slow Gradients
by: Lee, Jaerin, et al.
Published: (2024)
by: Lee, Jaerin, et al.
Published: (2024)
Gradient Routing: Masking Gradients to Localize Computation in Neural Networks
by: Cloud, Alex, et al.
Published: (2024)
by: Cloud, Alex, et al.
Published: (2024)
Axiomatization of Gradient Smoothing in Neural Networks
by: Zhou, Linjiang, et al.
Published: (2024)
by: Zhou, Linjiang, et al.
Published: (2024)
Gradient-Informed Temporal Sampling Improves Rollout Accuracy in PDE Surrogate Training
by: Wang, Wenshuo, et al.
Published: (2026)
by: Wang, Wenshuo, et al.
Published: (2026)
On-Device Training of Fully Quantized Deep Neural Networks on Cortex-M Microcontrollers
by: Deutel, Mark, et al.
Published: (2024)
by: Deutel, Mark, et al.
Published: (2024)
Native Fortran Implementation of TensorFlow-Trained Deep and Bayesian Neural Networks
by: Furlong, Aidan, et al.
Published: (2025)
by: Furlong, Aidan, et al.
Published: (2025)
BEND: Bagging Deep Learning Training Based on Efficient Neural Network Diffusion
by: Wei, Jia, et al.
Published: (2024)
by: Wei, Jia, et al.
Published: (2024)
Entropy-Aware Model Initialization for Effective Exploration in Deep Reinforcement Learning
by: Jang, Sooyoung, et al.
Published: (2021)
by: Jang, Sooyoung, et al.
Published: (2021)
Accelerating Training with Neuron Interaction and Nowcasting Networks
by: Knyazev, Boris, et al.
Published: (2024)
by: Knyazev, Boris, et al.
Published: (2024)
Gradient Inversion Attack on Graph Neural Networks
by: Sinha, Divya Anand, et al.
Published: (2024)
by: Sinha, Divya Anand, et al.
Published: (2024)
Complex Physics-Informed Neural Network
by: Si, Chenhao, et al.
Published: (2025)
by: Si, Chenhao, et al.
Published: (2025)
Similar Items
-
Geometric Regularization in Mixture-of-Experts: The Disconnect Between Weights and Activations
by: Kim, Hyunjun
Published: (2026) -
Soft Deterministic Policy Gradient with Gaussian Smoothing
by: Na, Hyunjun, et al.
Published: (2026) -
HOLOGRAPH: Active Causal Discovery via Sheaf-Theoretic Alignment of Large Language Model Priors
by: Kim, Hyunjun
Published: (2025) -
R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes
by: Na, Hyunjun, et al.
Published: (2026) -
DeepDefense: Layer-Wise Gradient-Feature Alignment for Building Robust Neural Networks
by: Lin, Ci, et al.
Published: (2025)