Saved in:
| Main Authors: | Lowell, Mark, Kastner, Catharine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.00127 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Adaptive and Stability-Promoting Layerwise Training Approach for Sparse Deep Neural Network Architecture
by: Krishnanunni, C G, et al.
Published: (2022)
by: Krishnanunni, C G, et al.
Published: (2022)
There is a Singularity in the Loss Landscape
by: Lowell, Mark
Published: (2022)
by: Lowell, Mark
Published: (2022)
Layerwise Progressive Freezing Enables STE-Free Training of Deep Binary Neural Networks
by: Smith, Evan Gibson, et al.
Published: (2026)
by: Smith, Evan Gibson, et al.
Published: (2026)
Understanding Gradient Descent through the Training Jacobian
by: Belrose, Nora, et al.
Published: (2024)
by: Belrose, Nora, et al.
Published: (2024)
Noise-Adaptive Layerwise Learning Rates: Accelerating Geometry-Aware Optimization for Deep Neural Network Training
by: Hao, Jie, et al.
Published: (2025)
by: Hao, Jie, et al.
Published: (2025)
Jacobian Regularization Stabilizes Long-Term Integration of Neural Differential Equations
by: Janvier, Maya, et al.
Published: (2026)
by: Janvier, Maya, et al.
Published: (2026)
Task-driven Layerwise Additive Activation Intervention
by: Nguyen, Hieu Trung, et al.
Published: (2025)
by: Nguyen, Hieu Trung, et al.
Published: (2025)
Identifiable Equivariant Networks are Layerwise Equivariant
by: Shahverdi, Vahid, et al.
Published: (2026)
by: Shahverdi, Vahid, et al.
Published: (2026)
On the Stability of the Jacobian Matrix in Deep Neural Networks
by: Dadoun, Benjamin, et al.
Published: (2025)
by: Dadoun, Benjamin, et al.
Published: (2025)
Training Implicit Networks for Image Deblurring using Jacobian-Free Backpropagation
by: Liu, Linghai, et al.
Published: (2024)
by: Liu, Linghai, et al.
Published: (2024)
An Infinite-Width Analysis on the Jacobian-Regularised Training of a Neural Network
by: Kim, Taeyoung, et al.
Published: (2023)
by: Kim, Taeyoung, et al.
Published: (2023)
BALI: Learning Neural Networks via Bayesian Layerwise Inference
by: Kurle, Richard, et al.
Published: (2024)
by: Kurle, Richard, et al.
Published: (2024)
Minimizing Layerwise Activation Norm Improves Generalization in Federated Learning
by: Yashwanth, M, et al.
Published: (2025)
by: Yashwanth, M, et al.
Published: (2025)
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
by: Dufort-Labbé, Simon, et al.
Published: (2026)
by: Dufort-Labbé, Simon, et al.
Published: (2026)
Layerwise Recall and the Geometry of Interwoven Knowledge in LLMs
by: Lei, Ge, et al.
Published: (2025)
by: Lei, Ge, et al.
Published: (2025)
Why Deep Jacobian Spectra Separate: Depth-Induced Scaling and Singular-Vector Alignment
by: Haas, Nathanaël, et al.
Published: (2026)
by: Haas, Nathanaël, et al.
Published: (2026)
Task Structure Reverses Layerwise State Encoding in Sequence Models
by: Jiang, Yuhang
Published: (2026)
by: Jiang, Yuhang
Published: (2026)
Robust Layerwise Scaling Rules by Proper Weight Decay Tuning
by: Fan, Zhiyuan, et al.
Published: (2025)
by: Fan, Zhiyuan, et al.
Published: (2025)
JacQuant: STE-Free Quantization-Aware Training via Learned Jacobian Surrogates
by: Yi, Kai, et al.
Published: (2026)
by: Yi, Kai, et al.
Published: (2026)
OWLed: Outlier-weighed Layerwise Pruning for Efficient Autonomous Driving Framework
by: Li, Jiaxi, et al.
Published: (2024)
by: Li, Jiaxi, et al.
Published: (2024)
Layerwise Proximal Replay: A Proximal Point Method for Online Continual Learning
by: Yoo, Jason, et al.
Published: (2024)
by: Yoo, Jason, et al.
Published: (2024)
CoopQ: Cooperative Game Inspired Layerwise Mixed Precision Quantization for LLMs
by: Zhao, Junchen, et al.
Published: (2025)
by: Zhao, Junchen, et al.
Published: (2025)
Outlier-weighed Layerwise Sampling for LLM Fine-tuning
by: Li, Pengxiang, et al.
Published: (2024)
by: Li, Pengxiang, et al.
Published: (2024)
Layerwise Change of Knowledge in Neural Networks
by: Cheng, Xu, et al.
Published: (2024)
by: Cheng, Xu, et al.
Published: (2024)
Edge of Stochastic Stability: Revisiting the Edge of Stability for SGD
by: Andreyev, Arseniy, et al.
Published: (2024)
by: Andreyev, Arseniy, et al.
Published: (2024)
The Origin of Edge of Stability
by: Litman, Elon
Published: (2026)
by: Litman, Elon
Published: (2026)
Jacobian-Enhanced Neural Networks
by: Berguin, Steven H.
Published: (2024)
by: Berguin, Steven H.
Published: (2024)
Jacobian Aligned Random Forests
by: Rauniyar, Sarwesh
Published: (2025)
by: Rauniyar, Sarwesh
Published: (2025)
Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification
by: Lutz, Patrick, et al.
Published: (2026)
by: Lutz, Patrick, et al.
Published: (2026)
Adaptive Layerwise Perturbation: Unifying Off-Policy Corrections for LLM RL
by: Ye, Chenlu, et al.
Published: (2026)
by: Ye, Chenlu, et al.
Published: (2026)
A Note on Non-Composability of Layerwise Approximate Verification for Neural Inference
by: Zamir, Or
Published: (2026)
by: Zamir, Or
Published: (2026)
Exploring Layerwise Adversarial Robustness Through the Lens of t-SNE
by: Valentim, Inês, et al.
Published: (2024)
by: Valentim, Inês, et al.
Published: (2024)
μP$^2$: Effective Sharpness Aware Minimization Requires Layerwise Perturbation Scaling
by: Haas, Moritz, et al.
Published: (2024)
by: Haas, Moritz, et al.
Published: (2024)
Offline Oracle-Efficient Learning for Contextual MDPs via Layerwise Exploration-Exploitation Tradeoff
by: Qian, Jian, et al.
Published: (2024)
by: Qian, Jian, et al.
Published: (2024)
From Compression to Expression: A Layerwise Analysis of In-Context Learning
by: Jiang, Jiachen, et al.
Published: (2025)
by: Jiang, Jiachen, et al.
Published: (2025)
EDGE-LLM: Enabling Efficient Large Language Model Adaptation on Edge Devices via Layerwise Unified Compression and Adaptive Layer Tuning and Voting
by: Yu, Zhongzhi, et al.
Published: (2024)
by: Yu, Zhongzhi, et al.
Published: (2024)
Maximum Redundancy Pruning: A Principle-Driven Layerwise Sparsity Allocation for LLMs
by: Gao, Chang, et al.
Published: (2025)
by: Gao, Chang, et al.
Published: (2025)
POLAR: Policy-based Layerwise Reinforcement Learning Method for Stealthy Backdoor Attacks in Federated Learning
by: Yu, Kuai, et al.
Published: (2025)
by: Yu, Kuai, et al.
Published: (2025)
Some Theoretical Results on Layerwise Effective Dimension Oscillations in Finite Width ReLU Networks
by: Makwana, Darshan
Published: (2025)
by: Makwana, Darshan
Published: (2025)
GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control
by: Xu, Haofeng, et al.
Published: (2026)
by: Xu, Haofeng, et al.
Published: (2026)
Similar Items
-
An Adaptive and Stability-Promoting Layerwise Training Approach for Sparse Deep Neural Network Architecture
by: Krishnanunni, C G, et al.
Published: (2022) -
There is a Singularity in the Loss Landscape
by: Lowell, Mark
Published: (2022) -
Layerwise Progressive Freezing Enables STE-Free Training of Deep Binary Neural Networks
by: Smith, Evan Gibson, et al.
Published: (2026) -
Understanding Gradient Descent through the Training Jacobian
by: Belrose, Nora, et al.
Published: (2024) -
Noise-Adaptive Layerwise Learning Rates: Accelerating Geometry-Aware Optimization for Deep Neural Network Training
by: Hao, Jie, et al.
Published: (2025)