Saved in:
| Main Authors: | Kiselev, Nikita, Grabovoy, Andrey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.11995 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Curvature-Aligned Probing for Local Loss-Landscape Stabilization
by: Kiselev, Nikita, et al.
Published: (2026)
by: Kiselev, Nikita, et al.
Published: (2026)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025)
by: Petrov, Egor, et al.
Published: (2025)
Gradient-Normalized Smoothness for Optimization with Approximate Hessians
by: Semenov, Andrei, et al.
Published: (2025)
by: Semenov, Andrei, et al.
Published: (2025)
On Convergence of Incremental Gradient for Non-Convex Smooth Functions
by: Koloskova, Anastasia, et al.
Published: (2023)
by: Koloskova, Anastasia, et al.
Published: (2023)
The effects of Hessian eigenvalue spectral density type on the applicability of Hessian analysis to generalization capability assessment of neural networks
by: Gabdullin, Nikita
Published: (2025)
by: Gabdullin, Nikita
Published: (2025)
Practical Bayesian Inference for Speech SNNs: Uncertainty and Loss-Landscape Smoothing
by: Abdennadher, Yesmine, et al.
Published: (2026)
by: Abdennadher, Yesmine, et al.
Published: (2026)
Convergent Privacy Loss of Noisy-SGD without Convexity and Smoothness
by: Chien, Eli, et al.
Published: (2024)
by: Chien, Eli, et al.
Published: (2024)
Investigating generalization capabilities of neural networks by means of loss landscapes and Hessian analysis
by: Gabdullin, Nikita
Published: (2024)
by: Gabdullin, Nikita
Published: (2024)
Federated Optimization of Smooth Loss Functions
by: Jadbabaie, Ali, et al.
Published: (2022)
by: Jadbabaie, Ali, et al.
Published: (2022)
Asymptotic Smoothing of the Lipschitz Loss Landscape in Overparameterized One-Hidden-Layer ReLU Networks
by: Baturin, Saveliy
Published: (2026)
by: Baturin, Saveliy
Published: (2026)
Visualizing Loss Functions as Topological Landscape Profiles
by: Geniesse, Caleb, et al.
Published: (2024)
by: Geniesse, Caleb, et al.
Published: (2024)
The Expected Loss of Preconditioned Langevin Dynamics Reveals the Hessian Rank
by: Bar, Amitay, et al.
Published: (2024)
by: Bar, Amitay, et al.
Published: (2024)
Advancing Supervised Learning with the Wave Loss Function: A Robust and Smooth Approach
by: Akhtar, Mushir, et al.
Published: (2024)
by: Akhtar, Mushir, et al.
Published: (2024)
Sensitivity Analysis On Loss Landscape
by: Faroz, Salman
Published: (2024)
by: Faroz, Salman
Published: (2024)
HawkEye: Advancing Robust Regression with Bounded, Smooth, and Insensitive Loss Function
by: Akhtar, Mushir, et al.
Published: (2024)
by: Akhtar, Mushir, et al.
Published: (2024)
Designing a Robust, Bounded, and Smooth Loss Function for Improved Supervised Learning
by: Mahato, Soumi, et al.
Published: (2026)
by: Mahato, Soumi, et al.
Published: (2026)
Wolkowicz-Styan Upper Bound on the Hessian Eigenspectrum for Cross-Entropy Loss in Nonlinear Smooth Neural Networks
by: Omae, Yuto, et al.
Published: (2026)
by: Omae, Yuto, et al.
Published: (2026)
RoBoSS: A Robust, Bounded, Sparse, and Smooth Loss Function for Supervised Learning
by: Akhtar, Mushir, et al.
Published: (2023)
by: Akhtar, Mushir, et al.
Published: (2023)
Implicit score matching meets denoising score matching: improved rates of convergence and log-density Hessian estimation
by: Yakovlev, Konstantin, et al.
Published: (2025)
by: Yakovlev, Konstantin, et al.
Published: (2025)
Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured Prediction
by: Mohri, Mehryar, et al.
Published: (2026)
by: Mohri, Mehryar, et al.
Published: (2026)
Revisit, Extend, and Enhance Hessian-Free Influence Functions
by: Yang, Ziao, et al.
Published: (2024)
by: Yang, Ziao, et al.
Published: (2024)
Convergence Analysis of SGD under Expected Smoothness
by: Kawamoto, Yuta, et al.
Published: (2025)
by: Kawamoto, Yuta, et al.
Published: (2025)
Bayesian Influence Functions for Hessian-Free Data Attribution
by: Kreer, Philipp Alexander, et al.
Published: (2025)
by: Kreer, Philipp Alexander, et al.
Published: (2025)
LOTION: Smoothing the Optimization Landscape for Quantized Training
by: Kwun, Mujin, et al.
Published: (2025)
by: Kwun, Mujin, et al.
Published: (2025)
There is a Singularity in the Loss Landscape
by: Lowell, Mark
Published: (2022)
by: Lowell, Mark
Published: (2022)
Global Convergence of Natural Policy Gradient with Hessian-aided Momentum Variance Reduction
by: Feng, Jie, et al.
Published: (2024)
by: Feng, Jie, et al.
Published: (2024)
Improving Protein Optimization with Smoothed Fitness Landscapes
by: Kirjner, Andrew, et al.
Published: (2023)
by: Kirjner, Andrew, et al.
Published: (2023)
Using Degeneracy in the Loss Landscape for Mechanistic Interpretability
by: Bushnaq, Lucius, et al.
Published: (2024)
by: Bushnaq, Lucius, et al.
Published: (2024)
Paths and Ambient Spaces in Neural Loss Landscapes
by: Dold, Daniel, et al.
Published: (2025)
by: Dold, Daniel, et al.
Published: (2025)
A Scalable Measure of Loss Landscape Curvature for Analyzing the Training Dynamics of LLMs
by: Kalra, Dayal Singh, et al.
Published: (2026)
by: Kalra, Dayal Singh, et al.
Published: (2026)
A Universal Growth Rate for Learning with Smooth Surrogate Losses
by: Mao, Anqi, et al.
Published: (2024)
by: Mao, Anqi, et al.
Published: (2024)
Unified Interpretation of Smoothing Methods for Negative Sampling Loss Functions in Knowledge Graph Embedding
by: Feng, Xincan, et al.
Published: (2024)
by: Feng, Xincan, et al.
Published: (2024)
LossLens: Diagnostics for Machine Learning through Loss Landscape Visual Analytics
by: Xie, Tiankai, et al.
Published: (2024)
by: Xie, Tiankai, et al.
Published: (2024)
Unraveling the Key Components of OOD Generalization via Diversification
by: Benoit, Harold, et al.
Published: (2023)
by: Benoit, Harold, et al.
Published: (2023)
Landscaper: Understanding Loss Landscapes Through Multi-Dimensional Topological Analysis
by: Chen, Jiaqing, et al.
Published: (2026)
by: Chen, Jiaqing, et al.
Published: (2026)
MGDA Converges under Generalized Smoothness, Provably
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
Directional Smoothness and Gradient Methods: Convergence and Adaptivity
by: Mishkin, Aaron, et al.
Published: (2024)
by: Mishkin, Aaron, et al.
Published: (2024)
Smoothing the Landscape: Causal Structure Learning via Diffusion Denoising Objectives
by: Zhu, Hao, et al.
Published: (2026)
by: Zhu, Hao, et al.
Published: (2026)
LSAM: Asynchronous Distributed Training with Landscape-Smoothed Sharpness-Aware Minimization
by: Teng, Yunfei, et al.
Published: (2025)
by: Teng, Yunfei, et al.
Published: (2025)
Loss-Complexity Landscape and Model Structure Functions
by: Kolpakov, Alexander
Published: (2025)
by: Kolpakov, Alexander
Published: (2025)
Similar Items
-
Curvature-Aligned Probing for Local Loss-Landscape Stabilization
by: Kiselev, Nikita, et al.
Published: (2026) -
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
by: Petrov, Egor, et al.
Published: (2025) -
Gradient-Normalized Smoothness for Optimization with Approximate Hessians
by: Semenov, Andrei, et al.
Published: (2025) -
On Convergence of Incremental Gradient for Non-Convex Smooth Functions
by: Koloskova, Anastasia, et al.
Published: (2023) -
The effects of Hessian eigenvalue spectral density type on the applicability of Hessian analysis to generalization capability assessment of neural networks
by: Gabdullin, Nikita
Published: (2025)