Gespeichert in:
| Hauptverfasser: | Kim, Gyu Yeol, Oh, Min-hwan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.19156 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Newton-Muon Optimizer
von: Du, Zhehang, et al.
Veröffentlicht: (2026)
von: Du, Zhehang, et al.
Veröffentlicht: (2026)
ADAM Optimization with Adaptive Batch Selection
von: Kim, Gyu Yeol, et al.
Veröffentlicht: (2025)
von: Kim, Gyu Yeol, et al.
Veröffentlicht: (2025)
On the Convergence Analysis of Muon
von: Shen, Wei, et al.
Veröffentlicht: (2025)
von: Shen, Wei, et al.
Veröffentlicht: (2025)
Muon Does Not Converge on Convex Lipschitz Functions
von: Parshakova, Tetiana, et al.
Veröffentlicht: (2026)
von: Parshakova, Tetiana, et al.
Veröffentlicht: (2026)
Drop-Muon: Update Less, Converge Faster
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
Improved Convergence Rates of Muon Optimizer for Nonconvex Optimization
von: Nagashima, Shuntaro, et al.
Veröffentlicht: (2026)
von: Nagashima, Shuntaro, et al.
Veröffentlicht: (2026)
Gradient Regularized Newton Boosting Trees with Global Convergence
von: Zozoulenko, Nikita, et al.
Veröffentlicht: (2026)
von: Zozoulenko, Nikita, et al.
Veröffentlicht: (2026)
Incremental Gauss--Newton Methods with Superlinear Convergence Rates
von: Zhou, Zhiling, et al.
Veröffentlicht: (2024)
von: Zhou, Zhiling, et al.
Veröffentlicht: (2024)
Incremental Quasi-Newton Methods with Faster Superlinear Convergence Rates
von: Liu, Zhuanghua, et al.
Veröffentlicht: (2024)
von: Liu, Zhuanghua, et al.
Veröffentlicht: (2024)
Simple Stepsize for Quasi-Newton Methods with Global Convergence Guarantees
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
Unified Convergence Theory of Stochastic and Variance-Reduced Cubic Newton Methods
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2023)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2023)
Muon Converges under Heavy-Tailed Noise: Nonconvex Hölder-Smooth Empirical Risk Minimization
von: Iiduka, Hideaki
Veröffentlicht: (2026)
von: Iiduka, Hideaki
Veröffentlicht: (2026)
Phases of Muon: When Muon Eclipses SignSGD
von: Paquette, Elliot, et al.
Veröffentlicht: (2026)
von: Paquette, Elliot, et al.
Veröffentlicht: (2026)
Online Learning Guided Quasi-Newton Methods with Global Non-Asymptotic Convergence
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
A second-order method landing on the Stiefel manifold via Newton$\unicode{x2013}$Schulz iteration
von: Xiong, Xinhui, et al.
Veröffentlicht: (2026)
von: Xiong, Xinhui, et al.
Veröffentlicht: (2026)
MuonBP: Faster Muon via Block-Periodic Orthogonalization
von: Khaled, Ahmed, et al.
Veröffentlicht: (2025)
von: Khaled, Ahmed, et al.
Veröffentlicht: (2025)
LiMuon: Light and Fast Muon Optimizer for Large Models
von: Huang, Feihu, et al.
Veröffentlicht: (2025)
von: Huang, Feihu, et al.
Veröffentlicht: (2025)
Krylov Cubic Regularized Newton: A Subspace Second-Order Method with Dimension-Free Convergence Rate
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
Error Feedback for Muon and Friends
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
Sketch-and-Project Meets Newton Method: Global $\mathcal O(k^{-2})$ Convergence with Low-Rank Updates
von: Hanzely, Slavomír
Veröffentlicht: (2023)
von: Hanzely, Slavomír
Veröffentlicht: (2023)
Error whitening: Why Gauss-Newton outperforms Newton
von: McKay, Maricela Best, et al.
Veröffentlicht: (2026)
von: McKay, Maricela Best, et al.
Veröffentlicht: (2026)
Insights on Muon from Simple Quadratics
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
Muon is Provably Faster with Momentum Variance Reduction
von: Qian, Xun, et al.
Veröffentlicht: (2025)
von: Qian, Xun, et al.
Veröffentlicht: (2025)
Beyond the Ideal: Analyzing the Inexact Muon Update
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
Muon Optimizes Under Spectral Norm Constraints
von: Chen, Lizhang, et al.
Veröffentlicht: (2025)
von: Chen, Lizhang, et al.
Veröffentlicht: (2025)
On the Convergence of Black-Box Variational Inference
von: Kim, Kyurae, et al.
Veröffentlicht: (2023)
von: Kim, Kyurae, et al.
Veröffentlicht: (2023)
Lions and Muons: Optimization via Stochastic Frank-Wolfe
von: Sfyraki, Maria-Eleni, et al.
Veröffentlicht: (2025)
von: Sfyraki, Maria-Eleni, et al.
Veröffentlicht: (2025)
MiMuon: Mixed Muon Optimizer with Improved Generalization for Large Models
von: Huang, Feihu, et al.
Veröffentlicht: (2026)
von: Huang, Feihu, et al.
Veröffentlicht: (2026)
Muon in Associative Memory Learning: Training Dynamics and Scaling Laws
von: Li, Binghui, et al.
Veröffentlicht: (2026)
von: Li, Binghui, et al.
Veröffentlicht: (2026)
AdaGrad Meets Muon: Adaptive Stepsizes for Orthogonal Updates
von: Zhang, Minxin, et al.
Veröffentlicht: (2025)
von: Zhang, Minxin, et al.
Veröffentlicht: (2025)
Implicit Bias of Spectral Descent and Muon on Multiclass Separable Data
von: Fan, Chen, et al.
Veröffentlicht: (2025)
von: Fan, Chen, et al.
Veröffentlicht: (2025)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
von: Zhang, Fangzhao, et al.
Veröffentlicht: (2026)
von: Zhang, Fangzhao, et al.
Veröffentlicht: (2026)
Stochastic Newton Proximal Extragradient Method
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
Improving Stochastic Cubic Newton with Momentum
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2024)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2024)
FedMuon: Federated Learning with Bias-corrected LMO-based Optimization
von: Takezawa, Yuki, et al.
Veröffentlicht: (2025)
von: Takezawa, Yuki, et al.
Veröffentlicht: (2025)
Online Newton Method for Bandit Convex Optimisation
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
Incremental Gauss-Newton Descent for Machine Learning
von: Korbit, Mikalai, et al.
Veröffentlicht: (2024)
von: Korbit, Mikalai, et al.
Veröffentlicht: (2024)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
von: Tang, Xun, et al.
Veröffentlicht: (2024)
von: Tang, Xun, et al.
Veröffentlicht: (2024)
Sharpened Lazy Incremental Quasi-Newton Method
von: Lahoti, Aakash, et al.
Veröffentlicht: (2023)
von: Lahoti, Aakash, et al.
Veröffentlicht: (2023)
Efficient Graph Laplacian Estimation by Proximal Newton
von: Medvedovsky, Yakov, et al.
Veröffentlicht: (2023)
von: Medvedovsky, Yakov, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
The Newton-Muon Optimizer
von: Du, Zhehang, et al.
Veröffentlicht: (2026) -
ADAM Optimization with Adaptive Batch Selection
von: Kim, Gyu Yeol, et al.
Veröffentlicht: (2025) -
On the Convergence Analysis of Muon
von: Shen, Wei, et al.
Veröffentlicht: (2025) -
Muon Does Not Converge on Convex Lipschitz Functions
von: Parshakova, Tetiana, et al.
Veröffentlicht: (2026) -
Drop-Muon: Update Less, Converge Faster
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)