Layer-wise Adaptive Gradient Norm Penalizing Method for Efficient and Accurate Deep Learning
Fuente:
arXiv
Saved in:
| Main Author: | Lee, Sunwoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Layer-wise Update Aggregation with Recycling for Communication-Efficient Federated Learning
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
Asynchronous Sharpness-Aware Minimization For Fast and Accurate Deep Learning
by: Jo, Junhyuk, et al.
Published: (2025)
by: Jo, Junhyuk, et al.
Published: (2025)
On the Effect of Uncertainty on Layer-wise Inference Dynamics
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
Dynamic Rank Adjustment for Accurate and Efficient Neural Network Training
by: Shin, Hyuntak, et al.
Published: (2025)
by: Shin, Hyuntak, et al.
Published: (2025)
Biased Local SGD for Efficient Deep Learning on Heterogeneous Systems
by: Lim, Jihyun, et al.
Published: (2025)
by: Lim, Jihyun, et al.
Published: (2025)
FedLWS: Federated Learning with Adaptive Layer-wise Weight Shrinking
by: Shi, Changlong, et al.
Published: (2025)
by: Shi, Changlong, et al.
Published: (2025)
On The Concurrence of Layer-wise Preconditioning Methods and Provable Feature Learning
by: Zhang, Thomas T., et al.
Published: (2025)
by: Zhang, Thomas T., et al.
Published: (2025)
Deep Learning and Matrix Completion-aided IoT Network Localization in the Outlier Scenarios
by: Kim, Sunwoo
Published: (2025)
by: Kim, Sunwoo
Published: (2025)
Stochastic Layer-wise Learning: Scalable and Efficient Alternative to Backpropagation
by: Yin, Bojian, et al.
Published: (2025)
by: Yin, Bojian, et al.
Published: (2025)
DP-FedPGN: Finding Global Flat Minima for Differentially Private Federated Learning via Penalizing Gradient Norm
by: Liu, Junkang, et al.
Published: (2025)
by: Liu, Junkang, et al.
Published: (2025)
Resource-Efficient Federated Multimodal Learning via Layer-wise and Progressive Training
by: Tun, Ye Lin, et al.
Published: (2024)
by: Tun, Ye Lin, et al.
Published: (2024)
Differentially Private Block-wise Gradient Shuffle for Deep Learning
by: Zagardo, David
Published: (2024)
by: Zagardo, David
Published: (2024)
FedLAM: Low-latency Wireless Federated Learning via Layer-wise Adaptive Modulation
by: Qu, Linping, et al.
Published: (2025)
by: Qu, Linping, et al.
Published: (2025)
Element-wise Modulation of Random Matrices for Efficient Neural Layers
by: Szorc, Maksymilian
Published: (2025)
by: Szorc, Maksymilian
Published: (2025)
Geometric Layer-wise Approximation Rates for Deep Networks
by: Zhang, Shijun, et al.
Published: (2026)
by: Zhang, Shijun, et al.
Published: (2026)
GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning
by: Tian, Kaiyuan, et al.
Published: (2026)
by: Tian, Kaiyuan, et al.
Published: (2026)
Efficient, Accurate and Stable Gradients for Neural ODEs
by: McCallum, Sam, et al.
Published: (2024)
by: McCallum, Sam, et al.
Published: (2024)
Tight and Efficient Upper Bound on Spectral Norm of Convolutional Layers
by: Grishina, Ekaterina, et al.
Published: (2024)
by: Grishina, Ekaterina, et al.
Published: (2024)
Learning Accurate, Efficient, and Interpretable MLPs on Multiplex Graphs via Node-wise Multi-View Ensemble Distillation
by: Liu, Yunhui, et al.
Published: (2025)
by: Liu, Yunhui, et al.
Published: (2025)
Integrated Gradient Correlation: a Dataset-wise Attribution Method
by: Lelièvre, Pierre, et al.
Published: (2024)
by: Lelièvre, Pierre, et al.
Published: (2024)
Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Nuclear Norm Regularization for Deep Learning
by: Scarvelis, Christopher, et al.
Published: (2024)
by: Scarvelis, Christopher, et al.
Published: (2024)
Ghosted Layers: Unconstrained Activation Alignment for Recovering Layer-Pruned LLMs
by: Yun, Vincent-Daniel, et al.
Published: (2026)
by: Yun, Vincent-Daniel, et al.
Published: (2026)
GradientStabilizer:Fix the Norm, Not the Gradient
by: Huang, Tianjin, et al.
Published: (2025)
by: Huang, Tianjin, et al.
Published: (2025)
SPPCSO: Adaptive Penalized Estimation Method for High-Dimensional Correlated Data
by: Hu, Ying, et al.
Published: (2026)
by: Hu, Ying, et al.
Published: (2026)
Geometry and Dynamics of LayerNorm
by: Riechers, Paul M.
Published: (2024)
by: Riechers, Paul M.
Published: (2024)
Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
by: Chen, Chen, et al.
Published: (2026)
by: Chen, Chen, et al.
Published: (2026)
OUIDecay: Adaptive Layer-wise Weight Decay for CNNs Using Online Activation Patterns
by: Fernández-Hernández, Alberto, et al.
Published: (2026)
by: Fernández-Hernández, Alberto, et al.
Published: (2026)
AdaBet: Gradient-free Layer Selection for Efficient Training of Deep Neural Networks
by: Tenison, Irene, et al.
Published: (2025)
by: Tenison, Irene, et al.
Published: (2025)
Layer-wise Linear Mode Connectivity
by: Adilova, Linara, et al.
Published: (2023)
by: Adilova, Linara, et al.
Published: (2023)
Layer-wise Derivative Controlled Networks
by: Martnishn, Rowan, et al.
Published: (2026)
by: Martnishn, Rowan, et al.
Published: (2026)
An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms
by: Grünefeld, Nils, et al.
Published: (2026)
by: Grünefeld, Nils, et al.
Published: (2026)
GAS-Norm: Score-Driven Adaptive Normalization for Non-Stationary Time Series Forecasting in Deep Learning
by: Urettini, Edoardo, et al.
Published: (2024)
by: Urettini, Edoardo, et al.
Published: (2024)
Focusing Influence Mechanism for Multi-Agent Reinforcement Learning
by: Park, Yisak, et al.
Published: (2025)
by: Park, Yisak, et al.
Published: (2025)
Accurate, Efficient, and Explainable Deep Learning Approaches for Environmental Science Problems
by: Shi, Jimeng
Published: (2026)
by: Shi, Jimeng
Published: (2026)
Adaptive Gradient Methods at the Edge of Stability
by: Cohen, Jeremy M., et al.
Published: (2022)
by: Cohen, Jeremy M., et al.
Published: (2022)
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
by: Fang, Jiaxun, et al.
Published: (2025)
by: Fang, Jiaxun, et al.
Published: (2025)
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
Towards Layer-Wise Personalized Federated Learning: Adaptive Layer Disentanglement via Conflicting Gradients
by: Nguyen, Minh Duong, et al.
Published: (2024)
by: Nguyen, Minh Duong, et al.
Published: (2024)
Deep Q-Learning with Gradient Target Tracking
by: Park, Bum Geun, et al.
Published: (2025)
by: Park, Bum Geun, et al.
Published: (2025)
Similar Items
-
Layer-wise Update Aggregation with Recycling for Communication-Efficient Federated Learning
by: Kim, Jisoo, et al.
Published: (2025) -
Asynchronous Sharpness-Aware Minimization For Fast and Accurate Deep Learning
by: Jo, Junhyuk, et al.
Published: (2025) -
On the Effect of Uncertainty on Layer-wise Inference Dynamics
by: Kim, Sunwoo, et al.
Published: (2025) -
Dynamic Rank Adjustment for Accurate and Efficient Neural Network Training
by: Shin, Hyuntak, et al.
Published: (2025) -
Biased Local SGD for Efficient Deep Learning on Heterogeneous Systems
by: Lim, Jihyun, et al.
Published: (2025)