Saved in:
| Main Authors: | Corlouer, Guillaume, Semler, Avi, Strang, Alexander, Oldenziel, Alexander Gietelink |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.06366 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Type-II Saddles and Probabilistic Stability of Stochastic Gradient Descent
by: Ziyin, Liu, et al.
Published: (2023)
by: Ziyin, Liu, et al.
Published: (2023)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
by: Bantzis, Ioannis, et al.
Published: (2025)
by: Bantzis, Ioannis, et al.
Published: (2025)
From Density Matrices to Phase Transitions in Deep Learning: Spectral Early Warnings and Interpretability
by: Hennick, Max, et al.
Published: (2026)
by: Hennick, Max, et al.
Published: (2026)
Transformers represent belief state geometry in their residual stream
by: Shai, Adam S., et al.
Published: (2024)
by: Shai, Adam S., et al.
Published: (2024)
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Saddle-to-Saddle Dynamics Explains A Simplicity Bias Across Neural Network Architectures
by: Zhang, Yedi, et al.
Published: (2025)
by: Zhang, Yedi, et al.
Published: (2025)
Neural Network-based High-index Saddle Dynamics Method for Searching Saddle Points and Solution Landscape
by: Liu, Yuankai, et al.
Published: (2024)
by: Liu, Yuankai, et al.
Published: (2024)
A Theory of Saddle Escape in Deep Nonlinear Networks
by: Rawal, Divit, et al.
Published: (2026)
by: Rawal, Divit, et al.
Published: (2026)
You Are What You Eat -- AI Alignment Requires Understanding How Data Shapes Structure and Generalisation
by: Lehalleur, Simon Pepin, et al.
Published: (2025)
by: Lehalleur, Simon Pepin, et al.
Published: (2025)
Mirror Descent Algorithms with Nearly Dimension-Independent Rates for Differentially-Private Stochastic Saddle-Point Problems
by: González, Tomás, et al.
Published: (2024)
by: González, Tomás, et al.
Published: (2024)
Exploring New Frontiers in Vertical Federated Learning: the Role of Saddle Point Reformulation
by: Beznosikov, Aleksandr, et al.
Published: (2026)
by: Beznosikov, Aleksandr, et al.
Published: (2026)
Plateaus, Optima, and Overfitting in Multi-Layer Perceptrons: A Saddle-Saddle-Attractor Scenario
by: Maleknia, Alex Alì, et al.
Published: (2026)
by: Maleknia, Alex Alì, et al.
Published: (2026)
Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points
by: Yamamoto, Naoya, et al.
Published: (2025)
by: Yamamoto, Naoya, et al.
Published: (2025)
Federated Composite Saddle Point Optimization
by: Bai, Site, et al.
Published: (2023)
by: Bai, Site, et al.
Published: (2023)
Variational Stochastic Gradient Descent for Deep Neural Networks
by: Chen, Haotian, et al.
Published: (2024)
by: Chen, Haotian, et al.
Published: (2024)
Dimension-Free Saddle-Point Escape in Muon
by: Long, Yanlin, et al.
Published: (2026)
by: Long, Yanlin, et al.
Published: (2026)
Regret Minimization via Saddle Point Optimization
by: Kirschner, Johannes, et al.
Published: (2024)
by: Kirschner, Johannes, et al.
Published: (2024)
Distributed Saddle-Point Problems: Lower Bounds, Near-Optimal and Robust Algorithms
by: Beznosikov, Aleksandr, et al.
Published: (2020)
by: Beznosikov, Aleksandr, et al.
Published: (2020)
Saddle Hierarchy in Dense Associative Memory
by: Thériault, Robin, et al.
Published: (2025)
by: Thériault, Robin, et al.
Published: (2025)
Private Algorithms for Stochastic Saddle Points and Variational Inequalities: Beyond Euclidean Geometry
by: Bassily, Raef, et al.
Published: (2024)
by: Bassily, Raef, et al.
Published: (2024)
Quantization Avoids Saddle Points in Distributed Optimization
by: Bo, Yanan, et al.
Published: (2024)
by: Bo, Yanan, et al.
Published: (2024)
Inertial Newton Algorithms Avoiding Strict Saddle Points
by: Castera, Camille
Published: (2021)
by: Castera, Camille
Published: (2021)
Proximal Point Method for Online Saddle Point Problem
by: Meng, Qing-xin, et al.
Published: (2024)
by: Meng, Qing-xin, et al.
Published: (2024)
Efficiently Escaping Saddle Points for Policy Optimization
by: Khorasani, Sadegh, et al.
Published: (2023)
by: Khorasani, Sadegh, et al.
Published: (2023)
Series of Hessian-Vector Products for Tractable Saddle-Free Newton Optimisation of Neural Networks
by: Oldewage, Elre T., et al.
Published: (2023)
by: Oldewage, Elre T., et al.
Published: (2023)
Geometry of Critical Sets and Existence of Saddle Branches for Two-layer Neural Networks
by: Zhang, Leyang, et al.
Published: (2024)
by: Zhang, Leyang, et al.
Published: (2024)
Directional Convergence Near Small Initializations and Saddles in Two-Homogeneous Neural Networks
by: Kumar, Akshay, et al.
Published: (2024)
by: Kumar, Akshay, et al.
Published: (2024)
Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?
by: Innocenti, Francesco, et al.
Published: (2024)
by: Innocenti, Francesco, et al.
Published: (2024)
Stochastic Gradient Descent for Gaussian Processes Done Right
by: Lin, Jihao Andreas, et al.
Published: (2023)
by: Lin, Jihao Andreas, et al.
Published: (2023)
Simultaneous Learning and Optimization via Misspecified Saddle Point Problems
by: Ahmadi, Mohammad Mahdi, et al.
Published: (2025)
by: Ahmadi, Mohammad Mahdi, et al.
Published: (2025)
Enhancing Stability of Physics-Informed Neural Network Training Through Saddle-Point Reformulation
by: Bylinkin, Dmitry, et al.
Published: (2025)
by: Bylinkin, Dmitry, et al.
Published: (2025)
Scaling Law for Stochastic Gradient Descent in Quadratically Parameterized Linear Regression
by: Ding, Shihong, et al.
Published: (2025)
by: Ding, Shihong, et al.
Published: (2025)
Stochastic Gradient Descent for Two-layer Neural Networks
by: Cao, Dinghao, et al.
Published: (2024)
by: Cao, Dinghao, et al.
Published: (2024)
Loss Landscape of Shallow ReLU-like Neural Networks: Stationary Points, Saddle Escape, and Network Embedding
by: Wu, Frank Zhengqing, et al.
Published: (2024)
by: Wu, Frank Zhengqing, et al.
Published: (2024)
Sampling from Gaussian Process Posteriors using Stochastic Gradient Descent
by: Lin, Jihao Andreas, et al.
Published: (2023)
by: Lin, Jihao Andreas, et al.
Published: (2023)
A Saddle Point Remedy: Power of Variable Elimination in Non-convex Optimization
by: Gan, Min, et al.
Published: (2025)
by: Gan, Min, et al.
Published: (2025)
On Linear Convergence in Smooth Convex-Concave Bilinearly-Coupled Saddle-Point Optimization: Lower Bounds and Optimal Algorithms
by: Kovalev, Dmitry, et al.
Published: (2024)
by: Kovalev, Dmitry, et al.
Published: (2024)
Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime
by: Ghane, Reza, et al.
Published: (2026)
by: Ghane, Reza, et al.
Published: (2026)
Stochastic Adaptive Gradient Descent Without Descent
by: Aujol, Jean-François, et al.
Published: (2025)
by: Aujol, Jean-François, et al.
Published: (2025)
Learning a Single Index Model from Anisotropic Data with vanilla Stochastic Gradient Descent
by: Braun, Guillaume, et al.
Published: (2025)
by: Braun, Guillaume, et al.
Published: (2025)
Similar Items
-
Type-II Saddles and Probabilistic Stability of Stochastic Gradient Descent
by: Ziyin, Liu, et al.
Published: (2023) -
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
by: Bantzis, Ioannis, et al.
Published: (2025) -
From Density Matrices to Phase Transitions in Deep Learning: Spectral Early Warnings and Interpretability
by: Hennick, Max, et al.
Published: (2026) -
Transformers represent belief state geometry in their residual stream
by: Shai, Adam S., et al.
Published: (2024) -
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026)