Saved in:
| Main Authors: | Kiselev, Nikita, Grabovoy, Andrey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.14870 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unraveling the Hessian: A Key to Smooth Convergence in Loss Function Landscapes
by: Kiselev, Nikita, et al.
Published: (2024)
by: Kiselev, Nikita, et al.
Published: (2024)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025)
by: Petrov, Egor, et al.
Published: (2025)
A Scalable Measure of Loss Landscape Curvature for Analyzing the Training Dynamics of LLMs
by: Kalra, Dayal Singh, et al.
Published: (2026)
by: Kalra, Dayal Singh, et al.
Published: (2026)
On the Superlinear Relationship between SGD Noise Covariance and Loss Landscape Curvature
by: Zhang, Yikuan, et al.
Published: (2026)
by: Zhang, Yikuan, et al.
Published: (2026)
When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining
by: Karpukhin, Ivan, et al.
Published: (2026)
by: Karpukhin, Ivan, et al.
Published: (2026)
Curvature in the Looking-Glass: Optimal Methods to Exploit Curvature of Expectation in the Loss Landscape
by: Duersch, Jed A., et al.
Published: (2024)
by: Duersch, Jed A., et al.
Published: (2024)
Stable Coresets via Posterior Sampling: Aligning Induced and Full Loss Landscapes
by: Chang, Wei-Kai, et al.
Published: (2025)
by: Chang, Wei-Kai, et al.
Published: (2025)
N-Gram Induction Heads for In-Context RL: Improving Stability and Reducing Data Needs
by: Zisman, Ilya, et al.
Published: (2024)
by: Zisman, Ilya, et al.
Published: (2024)
Aligning Distributionally Robust Optimization with Practical Deep Learning Needs
by: Feoktistov, Dmitrii, et al.
Published: (2025)
by: Feoktistov, Dmitrii, et al.
Published: (2025)
Directions of Curvature as an Explanation for Loss of Plasticity
by: Lewandowski, Alex, et al.
Published: (2023)
by: Lewandowski, Alex, et al.
Published: (2023)
Sensitivity Analysis On Loss Landscape
by: Faroz, Salman
Published: (2024)
by: Faroz, Salman
Published: (2024)
Identifying Sparsely Active Circuits Through Local Loss Landscape Decomposition
by: Chrisman, Brianna, et al.
Published: (2025)
by: Chrisman, Brianna, et al.
Published: (2025)
Curvature Clues: Decoding Deep Learning Privacy with Input Loss Curvature
by: Ravikumar, Deepak, et al.
Published: (2024)
by: Ravikumar, Deepak, et al.
Published: (2024)
There is a Singularity in the Loss Landscape
by: Lowell, Mark
Published: (2022)
by: Lowell, Mark
Published: (2022)
From Memorization to Reasoning in the Spectrum of Loss Curvature
by: Merullo, Jack, et al.
Published: (2025)
by: Merullo, Jack, et al.
Published: (2025)
Paths and Ambient Spaces in Neural Loss Landscapes
by: Dold, Daniel, et al.
Published: (2025)
by: Dold, Daniel, et al.
Published: (2025)
Using Degeneracy in the Loss Landscape for Mechanistic Interpretability
by: Bushnaq, Lucius, et al.
Published: (2024)
by: Bushnaq, Lucius, et al.
Published: (2024)
LossLens: Diagnostics for Machine Learning through Loss Landscape Visual Analytics
by: Xie, Tiankai, et al.
Published: (2024)
by: Xie, Tiankai, et al.
Published: (2024)
Newton Losses: Using Curvature Information for Learning with Differentiable Algorithms
by: Petersen, Felix, et al.
Published: (2024)
by: Petersen, Felix, et al.
Published: (2024)
Landscaper: Understanding Loss Landscapes Through Multi-Dimensional Topological Analysis
by: Chen, Jiaqing, et al.
Published: (2026)
by: Chen, Jiaqing, et al.
Published: (2026)
Neural machine translation system for Lezgian, Russian and Azerbaijani languages
by: Asvarov, Alidar, et al.
Published: (2024)
by: Asvarov, Alidar, et al.
Published: (2024)
A Unified Noise-Curvature View of Loss of Trainability
by: Baveja, Gunbir Singh, et al.
Published: (2025)
by: Baveja, Gunbir Singh, et al.
Published: (2025)
Visualization and Analysis of the Loss Landscape in Graph Neural Networks
by: Moustafa, Samir, et al.
Published: (2025)
by: Moustafa, Samir, et al.
Published: (2025)
Probing Implicit Bias in Semi-gradient Q-learning: Visualizing the Effective Loss Landscapes via the Fokker--Planck Equation
by: Yin, Shuyu, et al.
Published: (2024)
by: Yin, Shuyu, et al.
Published: (2024)
Visualizing Loss Functions as Topological Landscape Profiles
by: Geniesse, Caleb, et al.
Published: (2024)
by: Geniesse, Caleb, et al.
Published: (2024)
On Measuring Localization of Shortcuts in Deep Networks
by: Tsoy, Nikita, et al.
Published: (2025)
by: Tsoy, Nikita, et al.
Published: (2025)
Efficient Sketches for Training Data Attribution and Studying the Loss Landscape
by: Schioppa, Andrea
Published: (2024)
by: Schioppa, Andrea
Published: (2024)
On the Hyperparameter Loss Landscapes of Machine Learning Models: An Exploratory Study
by: Huang, Mingyu, et al.
Published: (2023)
by: Huang, Mingyu, et al.
Published: (2023)
Visualizing, Rethinking, and Mining the Loss Landscape of Deep Neural Networks
by: Xu, Yichu, et al.
Published: (2024)
by: Xu, Yichu, et al.
Published: (2024)
Unveiling the Basin-Like Loss Landscape in Large Language Models
by: Chen, Huanran, et al.
Published: (2025)
by: Chen, Huanran, et al.
Published: (2025)
Gradient Aligned Regression via Pairwise Losses
by: Zhu, Dixian, et al.
Published: (2024)
by: Zhu, Dixian, et al.
Published: (2024)
Effective Structural Encodings via Local Curvature Profiles
by: Fesser, Lukas, et al.
Published: (2023)
by: Fesser, Lukas, et al.
Published: (2023)
A Minimal Bifurcation Model of Load Imbalance in a Softmax Mixture-of-Experts Router
by: Kiselev, O. M.
Published: (2026)
by: Kiselev, O. M.
Published: (2026)
Model Merging on Loss Landscape: A Geometry Perspective
by: Lu, Juanwu, et al.
Published: (2026)
by: Lu, Juanwu, et al.
Published: (2026)
Challenges in Training PINNs: A Loss Landscape Perspective
by: Rathore, Pratik, et al.
Published: (2024)
by: Rathore, Pratik, et al.
Published: (2024)
Evaluating Loss Landscapes from a Topology Perspective
by: Xie, Tiankai, et al.
Published: (2024)
by: Xie, Tiankai, et al.
Published: (2024)
Resource Constrained U-Net for Extraction of Retinal Vascular Trees
by: Kiselev, Georgiy
Published: (2024)
by: Kiselev, Georgiy
Published: (2024)
On the Loss Landscape Geometry of Regularized Deep Matrix Factorization: Uniqueness and Sharpness
by: Kamber, Anil, et al.
Published: (2026)
by: Kamber, Anil, et al.
Published: (2026)
Refined Risk Bounds for Unbounded Losses via Transductive Priors
by: Qian, Jian, et al.
Published: (2024)
by: Qian, Jian, et al.
Published: (2024)
Sharp Minima Can Generalize: A Loss Landscape Perspective On Data
by: Fan, Raymond, et al.
Published: (2025)
by: Fan, Raymond, et al.
Published: (2025)
Similar Items
-
Unraveling the Hessian: A Key to Smooth Convergence in Loss Function Landscapes
by: Kiselev, Nikita, et al.
Published: (2024) -
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025) -
A Scalable Measure of Loss Landscape Curvature for Analyzing the Training Dynamics of LLMs
by: Kalra, Dayal Singh, et al.
Published: (2026) -
On the Superlinear Relationship between SGD Noise Covariance and Loss Landscape Curvature
by: Zhang, Yikuan, et al.
Published: (2026) -
When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining
by: Karpukhin, Ivan, et al.
Published: (2026)