Insights from Gradient Dynamics: Gradient Autoscaled Normalization
Fuente:
arXiv
Saved in:
| Main Author: | Yun, Vincent-Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sharpness-Aware Minimization with Z-Score Gradient Filtering
by: Yun, Vincent-Daniel
Published: (2025)
by: Yun, Vincent-Daniel
Published: (2025)
Normalized Conditional Mutual Information Surrogate Loss for Deep Neural Classifiers
by: Ye, Linfeng, et al.
Published: (2026)
by: Ye, Linfeng, et al.
Published: (2026)
Partial Information Decomposition via Normalizing Flows in Latent Gaussian Distributions
by: Zhao, Wenyuan, et al.
Published: (2025)
by: Zhao, Wenyuan, et al.
Published: (2025)
ASI: Accuracy-Stability Index for Evaluating Deep Learning Models
by: Dai, Wei, et al.
Published: (2023)
by: Dai, Wei, et al.
Published: (2023)
Information-Guided Diffusion Sampling for Dataset Distillation
by: Ye, Linfeng, et al.
Published: (2025)
by: Ye, Linfeng, et al.
Published: (2025)
I-Con: A Unifying Framework for Representation Learning
by: Alshammari, Shaden, et al.
Published: (2025)
by: Alshammari, Shaden, et al.
Published: (2025)
LVM4CSI: Enabling Direct Application of Pre-Trained Large Vision Models for Wireless Channel Tasks
by: Guo, Jiajia, et al.
Published: (2025)
by: Guo, Jiajia, et al.
Published: (2025)
Towards Efficient VLMs: Information-Theoretic Driven Compression via Adaptive Structural Pruning
by: Xu, Zhaoqi, et al.
Published: (2025)
by: Xu, Zhaoqi, et al.
Published: (2025)
SDQM: Synthetic Data Quality Metric for Object Detection Dataset Evaluation
by: Zenith, Ayush, et al.
Published: (2025)
by: Zenith, Ayush, et al.
Published: (2025)
Generalized Nested Latent Variable Models for Lossy Coding applied to Wind Turbine Scenarios
by: Pérez-Gonzalo, Raül, et al.
Published: (2024)
by: Pérez-Gonzalo, Raül, et al.
Published: (2024)
Noise Scheduling as Information-Guided Allocation in Diffusion Training
by: Raya, Gabriel, et al.
Published: (2026)
by: Raya, Gabriel, et al.
Published: (2026)
Forte : Finding Outliers with Representation Typicality Estimation
by: Ganguly, Debargha, et al.
Published: (2024)
by: Ganguly, Debargha, et al.
Published: (2024)
Understanding the Role of Equivariance in Self-supervised Learning
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
RPN 2: On Interdependence Function Learning Towards Unifying and Advancing CNN, RNN, GNN, and Transformer
by: Zhang, Jiawei
Published: (2024)
by: Zhang, Jiawei
Published: (2024)
Poly-View Contrastive Learning
by: Shidani, Amitis, et al.
Published: (2024)
by: Shidani, Amitis, et al.
Published: (2024)
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions
by: Bilmes, Jeff A., et al.
Published: (2026)
by: Bilmes, Jeff A., et al.
Published: (2026)
RPN: Reconciled Polynomial Network Towards Unifying PGMs, Kernel SVMs, MLP and KAN
by: Zhang, Jiawei
Published: (2024)
by: Zhang, Jiawei
Published: (2024)
AdjointDPM: Adjoint Sensitivity Method for Gradient Backpropagation of Diffusion Probabilistic Models
by: Pan, Jiachun, et al.
Published: (2023)
by: Pan, Jiachun, et al.
Published: (2023)
Stochastic Gradient Sampling for Enhancing Neural Networks Training
by: Yun, Juyoung
Published: (2023)
by: Yun, Juyoung
Published: (2023)
Why Does Stochastic Gradient Descent Slow Down in Low-Precision Training?
by: Yun, Vincent-Daniel
Published: (2025)
by: Yun, Vincent-Daniel
Published: (2025)
Fast Fourier Transform-Based Spectral and Temporal Gradient Filtering for Differential Privacy
by: Shin, Hyeju, et al.
Published: (2025)
by: Shin, Hyeju, et al.
Published: (2025)
Recursive KL Divergence Optimization: A Dynamic Framework for Representation Learning
by: Martin, Anthony D
Published: (2025)
by: Martin, Anthony D
Published: (2025)
Cooperative Meta-Learning with Gradient Augmentation
by: Shin, Jongyun, et al.
Published: (2024)
by: Shin, Jongyun, et al.
Published: (2024)
Learning to Explore for Stochastic Gradient MCMC
by: Kim, SeungHyun, et al.
Published: (2024)
by: Kim, SeungHyun, et al.
Published: (2024)
Perturbing the Gradient for Alleviating Meta Overfitting
by: Gogoi, Manas, et al.
Published: (2024)
by: Gogoi, Manas, et al.
Published: (2024)
Gradient Harmonization in Unsupervised Domain Adaptation
by: Huang, Fuxiang, et al.
Published: (2024)
by: Huang, Fuxiang, et al.
Published: (2024)
Manipulating Feature Visualizations with Gradient Slingshots
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
On Spectral Properties of Gradient-based Explanation Methods
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Natural Gradient Descent for Online Continual Learning
by: Khawand, Joe, et al.
Published: (2026)
by: Khawand, Joe, et al.
Published: (2026)
The Rate-Distortion-Perception-Classification Tradeoff: Joint Source Coding and Modulation via Inverse-Domain GANs
by: Fang, Junli, et al.
Published: (2023)
by: Fang, Junli, et al.
Published: (2023)
REG: Rectified Gradient Guidance for Conditional Diffusion Models
by: Gao, Zhengqi, et al.
Published: (2025)
by: Gao, Zhengqi, et al.
Published: (2025)
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
by: Zheng, Bowen, et al.
Published: (2026)
by: Zheng, Bowen, et al.
Published: (2026)
Diffusion State-Guided Projected Gradient for Inverse Problems
by: Zirvi, Rayhan, et al.
Published: (2024)
by: Zirvi, Rayhan, et al.
Published: (2024)
Gradient Correlation Subspace Learning against Catastrophic Forgetting
by: Dubnov, Tammuz, et al.
Published: (2024)
by: Dubnov, Tammuz, et al.
Published: (2024)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
by: Hartman, Max, et al.
Published: (2025)
by: Hartman, Max, et al.
Published: (2025)
Deep Generative Sampling in the Dual Divergence Space: A Data-efficient & Interpretative Approach for Generative AI
by: Garg, Sahil, et al.
Published: (2024)
by: Garg, Sahil, et al.
Published: (2024)
Visual Language Model based Cross-modal Semantic Communication Systems
by: Jiang, Feibo, et al.
Published: (2024)
by: Jiang, Feibo, et al.
Published: (2024)
Towards Formalizing Spuriousness of Biased Datasets Using Partial Information Decomposition
by: Halder, Barproda, et al.
Published: (2024)
by: Halder, Barproda, et al.
Published: (2024)
Learning from Synthetic Data via Provenance-Based Input Gradient Guidance
by: Nagano, Koshiro, et al.
Published: (2026)
by: Nagano, Koshiro, et al.
Published: (2026)
Similar Items
-
Sharpness-Aware Minimization with Z-Score Gradient Filtering
by: Yun, Vincent-Daniel
Published: (2025) -
Normalized Conditional Mutual Information Surrogate Loss for Deep Neural Classifiers
by: Ye, Linfeng, et al.
Published: (2026) -
Partial Information Decomposition via Normalizing Flows in Latent Gaussian Distributions
by: Zhao, Wenyuan, et al.
Published: (2025) -
ASI: Accuracy-Stability Index for Evaluating Deep Learning Models
by: Dai, Wei, et al.
Published: (2023) -
Information-Guided Diffusion Sampling for Dataset Distillation
by: Ye, Linfeng, et al.
Published: (2025)