Gespeichert in:
| Hauptverfasser: | Long, Yanlin, Gu, Yufei, Xie, Zeke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.09331 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mano: Restriking Manifold Optimization for LLM Training
von: Gu, Yufei, et al.
Veröffentlicht: (2026)
von: Gu, Yufei, et al.
Veröffentlicht: (2026)
Efficiently Escaping Saddle Points for Policy Optimization
von: Khorasani, Sadegh, et al.
Veröffentlicht: (2023)
von: Khorasani, Sadegh, et al.
Veröffentlicht: (2023)
Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points
von: Yamamoto, Naoya, et al.
Veröffentlicht: (2025)
von: Yamamoto, Naoya, et al.
Veröffentlicht: (2025)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
von: Bantzis, Ioannis, et al.
Veröffentlicht: (2025)
von: Bantzis, Ioannis, et al.
Veröffentlicht: (2025)
Efficiently Escaping Saddle Points under Generalized Smoothness via Self-Bounding Regularity
von: Cao, Daniel Yiming, et al.
Veröffentlicht: (2025)
von: Cao, Daniel Yiming, et al.
Veröffentlicht: (2025)
Escaping Saddle Points for Nonsmooth Weakly Convex Functions via Perturbed Proximal Algorithms
von: Huang, Minhui, et al.
Veröffentlicht: (2021)
von: Huang, Minhui, et al.
Veröffentlicht: (2021)
Loss Landscape of Shallow ReLU-like Neural Networks: Stationary Points, Saddle Escape, and Network Embedding
von: Wu, Frank Zhengqing, et al.
Veröffentlicht: (2024)
von: Wu, Frank Zhengqing, et al.
Veröffentlicht: (2024)
Dimer-Enhanced Optimization: A First-Order Approach to Escaping Saddle Points in Neural Network Training
von: Hu, Yue, et al.
Veröffentlicht: (2025)
von: Hu, Yue, et al.
Veröffentlicht: (2025)
A Theory of Saddle Escape in Deep Nonlinear Networks
von: Rawal, Divit, et al.
Veröffentlicht: (2026)
von: Rawal, Divit, et al.
Veröffentlicht: (2026)
Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better
von: Zhao, Ji, et al.
Veröffentlicht: (2026)
von: Zhao, Ji, et al.
Veröffentlicht: (2026)
Regret Minimization via Saddle Point Optimization
von: Kirschner, Johannes, et al.
Veröffentlicht: (2024)
von: Kirschner, Johannes, et al.
Veröffentlicht: (2024)
Federated Composite Saddle Point Optimization
von: Bai, Site, et al.
Veröffentlicht: (2023)
von: Bai, Site, et al.
Veröffentlicht: (2023)
Proximal Point Method for Online Saddle Point Problem
von: Meng, Qing-xin, et al.
Veröffentlicht: (2024)
von: Meng, Qing-xin, et al.
Veröffentlicht: (2024)
Neural Network-based High-index Saddle Dynamics Method for Searching Saddle Points and Solution Landscape
von: Liu, Yuankai, et al.
Veröffentlicht: (2024)
von: Liu, Yuankai, et al.
Veröffentlicht: (2024)
Mirror Descent Algorithms with Nearly Dimension-Independent Rates for Differentially-Private Stochastic Saddle-Point Problems
von: González, Tomás, et al.
Veröffentlicht: (2024)
von: González, Tomás, et al.
Veröffentlicht: (2024)
Quantization Avoids Saddle Points in Distributed Optimization
von: Bo, Yanan, et al.
Veröffentlicht: (2024)
von: Bo, Yanan, et al.
Veröffentlicht: (2024)
Refining Multidimensional Video Reward Models via Disentangled Influence Functions
von: Wang, Muyao, et al.
Veröffentlicht: (2026)
von: Wang, Muyao, et al.
Veröffentlicht: (2026)
Inertial Newton Algorithms Avoiding Strict Saddle Points
von: Castera, Camille
Veröffentlicht: (2021)
von: Castera, Camille
Veröffentlicht: (2021)
Simultaneous Learning and Optimization via Misspecified Saddle Point Problems
von: Ahmadi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
von: Ahmadi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
A Saddle Point Remedy: Power of Variable Elimination in Non-convex Optimization
von: Gan, Min, et al.
Veröffentlicht: (2025)
von: Gan, Min, et al.
Veröffentlicht: (2025)
Stochastic Gradient Descent in the Saddle-to-Saddle Regime of Deep Linear Networks
von: Corlouer, Guillaume, et al.
Veröffentlicht: (2026)
von: Corlouer, Guillaume, et al.
Veröffentlicht: (2026)
Class-wise Activation Unravelling the Engima of Deep Double Descent
von: Gu, Yufei
Veröffentlicht: (2024)
von: Gu, Yufei
Veröffentlicht: (2024)
Accelerated Gradient Methods for Nonconvex Optimization: Escape Trajectories From Strict Saddle Points and Convergence to Local Minima
von: Dixit, Rishabh, et al.
Veröffentlicht: (2023)
von: Dixit, Rishabh, et al.
Veröffentlicht: (2023)
Weak-to-Strong Diffusion with Reflection
von: Bai, Lichen, et al.
Veröffentlicht: (2025)
von: Bai, Lichen, et al.
Veröffentlicht: (2025)
NuMuon: Nuclear-Norm-Constrained Muon for Compressible LLM Training
von: Dolatabadi, Hadi Mohaghegh, et al.
Veröffentlicht: (2026)
von: Dolatabadi, Hadi Mohaghegh, et al.
Veröffentlicht: (2026)
Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
Saddle-to-Saddle Dynamics Explains A Simplicity Bias Across Neural Network Architectures
von: Zhang, Yedi, et al.
Veröffentlicht: (2025)
von: Zhang, Yedi, et al.
Veröffentlicht: (2025)
Exploring New Frontiers in Vertical Federated Learning: the Role of Saddle Point Reformulation
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2026)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2026)
Online Min-Max Optimization: From Individual Regrets to Cumulative Saddle Points
von: Vyas, Abhijeet, et al.
Veröffentlicht: (2026)
von: Vyas, Abhijeet, et al.
Veröffentlicht: (2026)
Series of Hessian-Vector Products for Tractable Saddle-Free Newton Optimisation of Neural Networks
von: Oldewage, Elre T., et al.
Veröffentlicht: (2023)
von: Oldewage, Elre T., et al.
Veröffentlicht: (2023)
Bilevel Optimization over Saddle Points of Zero-Sum Markov Games
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
AdaMuon: Adaptive Muon Optimizer
von: Si, Chongjie, et al.
Veröffentlicht: (2025)
von: Si, Chongjie, et al.
Veröffentlicht: (2025)
Enhancing Stability of Physics-Informed Neural Network Training Through Saddle-Point Reformulation
von: Bylinkin, Dmitry, et al.
Veröffentlicht: (2025)
von: Bylinkin, Dmitry, et al.
Veröffentlicht: (2025)
Private Algorithms for Stochastic Saddle Points and Variational Inequalities: Beyond Euclidean Geometry
von: Bassily, Raef, et al.
Veröffentlicht: (2024)
von: Bassily, Raef, et al.
Veröffentlicht: (2024)
Channel Matters: Estimating Channel Influence for Multivariate Time Series
von: Wang, Muyao, et al.
Veröffentlicht: (2024)
von: Wang, Muyao, et al.
Veröffentlicht: (2024)
Multiphysics Bench: Benchmarking and Investigating Scientific Machine Learning for Multiphysics PDEs
von: Yang, Changfan, et al.
Veröffentlicht: (2025)
von: Yang, Changfan, et al.
Veröffentlicht: (2025)
Convergence, Sticking and Escape: Stochastic Dynamics Near Critical Points in SGD
von: Dudukalov, Dmitry, et al.
Veröffentlicht: (2025)
von: Dudukalov, Dmitry, et al.
Veröffentlicht: (2025)
From Saddle Points Toward Global Minima: A Newton-Type Method on Wasserstein Space
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2026)
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2026)
Saddle-Free Guidance: Improved On-Manifold Sampling without Labels or Additional Training
von: Yeats, Eric, et al.
Veröffentlicht: (2025)
von: Yeats, Eric, et al.
Veröffentlicht: (2025)
TrasMuon: Trust-Region Adaptive Scaling for Orthogonalized Momentum Optimizers
von: Cheng, Peng, et al.
Veröffentlicht: (2026)
von: Cheng, Peng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Mano: Restriking Manifold Optimization for LLM Training
von: Gu, Yufei, et al.
Veröffentlicht: (2026) -
Efficiently Escaping Saddle Points for Policy Optimization
von: Khorasani, Sadegh, et al.
Veröffentlicht: (2023) -
Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points
von: Yamamoto, Naoya, et al.
Veröffentlicht: (2025) -
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
von: Bantzis, Ioannis, et al.
Veröffentlicht: (2025) -
Efficiently Escaping Saddle Points under Generalized Smoothness via Self-Bounding Regularity
von: Cao, Daniel Yiming, et al.
Veröffentlicht: (2025)