Error Bound Analysis for the Regularized Loss of Deep Linear Neural Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Po, Jiang, Rujun, Wang, Peng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Augmented Lagrangian Method for Training Recurrent Neural Networks
di: Wang, Yue, et al.
Pubblicazione: (2024)
di: Wang, Yue, et al.
Pubblicazione: (2024)
Inexact FPPA for the $\ell_0$ Sparse Regularization Problem
di: Fang, Ronglong, et al.
Pubblicazione: (2024)
di: Fang, Ronglong, et al.
Pubblicazione: (2024)
Riemannian Adaptive Regularized Newton Methods with Hölder Continuous Hessians
di: Zhang, Chenyu, et al.
Pubblicazione: (2023)
di: Zhang, Chenyu, et al.
Pubblicazione: (2023)
A Layer Separation Optimization Framework for Cross-Entropy Training in Deep Learning
di: Liu, Yaru, et al.
Pubblicazione: (2026)
di: Liu, Yaru, et al.
Pubblicazione: (2026)
Convergence of gradient descent for deep neural networks
di: Chatterjee, Sourav
Pubblicazione: (2022)
di: Chatterjee, Sourav
Pubblicazione: (2022)
Effectively Leveraging Momentum Terms in Stochastic Line Search Frameworks for Fast Optimization of Finite-Sum Problems
di: Lapucci, Matteo, et al.
Pubblicazione: (2024)
di: Lapucci, Matteo, et al.
Pubblicazione: (2024)
Convergence Conditions for Stochastic Line Search Based Optimization of Over-parametrized Models
di: Lapucci, Matteo, et al.
Pubblicazione: (2024)
di: Lapucci, Matteo, et al.
Pubblicazione: (2024)
Recent Advances in Non-convex Smoothness Conditions and Applicability to Deep Linear Neural Networks
di: Patel, Vivak, et al.
Pubblicazione: (2024)
di: Patel, Vivak, et al.
Pubblicazione: (2024)
Faster Adaptive Optimization via Expected Gradient Outer Product Reparameterization
di: DePavia, Adela, et al.
Pubblicazione: (2025)
di: DePavia, Adela, et al.
Pubblicazione: (2025)
Sample-wise Constrained Learning via a Sequential Penalty Approach with Applications in Image Processing
di: Lanzillotta, Francesca, et al.
Pubblicazione: (2026)
di: Lanzillotta, Francesca, et al.
Pubblicazione: (2026)
A Complete Loss Landscape Analysis of Regularized Deep Matrix Factorization
di: Chen, Po, et al.
Pubblicazione: (2025)
di: Chen, Po, et al.
Pubblicazione: (2025)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
di: Khanal, Kshitiz
Pubblicazione: (2025)
di: Khanal, Kshitiz
Pubblicazione: (2025)
Consensus-based optimization for closed-box adversarial attacks and a connection to evolution strategies
di: Roith, Tim, et al.
Pubblicazione: (2025)
di: Roith, Tim, et al.
Pubblicazione: (2025)
General Constrained Matrix Optimization
di: Garner, Casey, et al.
Pubblicazione: (2024)
di: Garner, Casey, et al.
Pubblicazione: (2024)
Spectrally Constrained Optimization
di: Garner, Casey, et al.
Pubblicazione: (2023)
di: Garner, Casey, et al.
Pubblicazione: (2023)
A Grover-compatible manifold optimization algorithm for quantum search
di: Lai, Zhijian, et al.
Pubblicazione: (2025)
di: Lai, Zhijian, et al.
Pubblicazione: (2025)
Representation and Regression Problems in Neural Networks: Relaxation, Generalization, and Numerics
di: Liu, Kang, et al.
Pubblicazione: (2024)
di: Liu, Kang, et al.
Pubblicazione: (2024)
SVD-Preconditioned Gradient Descent Method for Solving Nonlinear Least Squares Problems
di: Chang, Zhipeng, et al.
Pubblicazione: (2026)
di: Chang, Zhipeng, et al.
Pubblicazione: (2026)
Inexact Riemannian Gradient Descent Method for Nonconvex Optimization
di: Zhou, Juan, et al.
Pubblicazione: (2024)
di: Zhou, Juan, et al.
Pubblicazione: (2024)
On the role of semismoothness in nonsmooth numerical analysis: Theory
di: Gfrerer, H., et al.
Pubblicazione: (2024)
di: Gfrerer, H., et al.
Pubblicazione: (2024)
On the role of semismoothness in the implicit programming approach to selected nonsmooth optimization problems
di: Gfrerer, Helmut, et al.
Pubblicazione: (2024)
di: Gfrerer, Helmut, et al.
Pubblicazione: (2024)
Implicit augmented Lagrangian and generalized optimization
di: De Marchi, Alberto
Pubblicazione: (2023)
di: De Marchi, Alberto
Pubblicazione: (2023)
A Semismooth Newton Stochastic Proximal Point Algorithm with Variance Reduction
di: Milzarek, Andre, et al.
Pubblicazione: (2022)
di: Milzarek, Andre, et al.
Pubblicazione: (2022)
A Representation Optimization Dichotomy, Lie-Algebraic Policy Optimization
di: KC, Sooraj, et al.
Pubblicazione: (2026)
di: KC, Sooraj, et al.
Pubblicazione: (2026)
Solving nonconvex optimization problems via a second order dynamical system with unbounded damping
di: László, Szilárd Csaba
Pubblicazione: (2025)
di: László, Szilárd Csaba
Pubblicazione: (2025)
Binno: A 1st-order method for Bi-level Nonconvex Nonsmooth Optimization for Matrix Factorizations
di: Selicato, Laura, et al.
Pubblicazione: (2025)
di: Selicato, Laura, et al.
Pubblicazione: (2025)
Power Homotopy for Zeroth-Order Non-Convex Optimizations
di: Xu, Chen
Pubblicazione: (2025)
di: Xu, Chen
Pubblicazione: (2025)
Global Optimization with A Power-Transformed Objective and Gaussian Smoothing
di: Xu, Chen
Pubblicazione: (2024)
di: Xu, Chen
Pubblicazione: (2024)
Parameter-Free Accelerated Quasi-Newton Method for Nonconvex Optimization
di: Marumo, Naoki
Pubblicazione: (2025)
di: Marumo, Naoki
Pubblicazione: (2025)
Oracle complexities of augmented Lagrangian methods for nonsmooth manifold optimization
di: Deng, Kangkang, et al.
Pubblicazione: (2024)
di: Deng, Kangkang, et al.
Pubblicazione: (2024)
Preconditioned subgradient method for composite optimization: overparameterization and fast convergence
di: Díaz, Mateo, et al.
Pubblicazione: (2025)
di: Díaz, Mateo, et al.
Pubblicazione: (2025)
Subgradient Regularization: A Descent-Oriented Subgradient Method for Nonsmooth Optimization
di: Li, Hanyang, et al.
Pubblicazione: (2025)
di: Li, Hanyang, et al.
Pubblicazione: (2025)
Tight Error Bounds for the Sign-Constrained Stiefel Manifold
di: Chen, Xiaojun, et al.
Pubblicazione: (2022)
di: Chen, Xiaojun, et al.
Pubblicazione: (2022)
Local Convergence of Adaptively Regularized Tensor Methods
di: Welzel, Karl, et al.
Pubblicazione: (2025)
di: Welzel, Karl, et al.
Pubblicazione: (2025)
An Extended ADMM for 3-Block Nonconvex Nonseparable Problems with Applications
di: Liu, Zekun
Pubblicazione: (2024)
di: Liu, Zekun
Pubblicazione: (2024)
A Symplectic Discretization Based Proximal Point Algorithm for Convex Minimization
di: Yuan, Ya-xiang, et al.
Pubblicazione: (2024)
di: Yuan, Ya-xiang, et al.
Pubblicazione: (2024)
Convergence of Descent Optimization Algorithms under Polyak-Łojasiewicz-Kurdyka Conditions
di: Bento, G. C., et al.
Pubblicazione: (2024)
di: Bento, G. C., et al.
Pubblicazione: (2024)
On the local and global minimizers of the smooth stress function in Euclidean Distance Matrix problems
di: Song, Mengmeng, et al.
Pubblicazione: (2024)
di: Song, Mengmeng, et al.
Pubblicazione: (2024)
Riemannian Trust Region Methods for SC$^1$ Minimization
di: Zhang, Chenyu, et al.
Pubblicazione: (2023)
di: Zhang, Chenyu, et al.
Pubblicazione: (2023)
Exact Solutions for the NP-hard Wasserstein Barycenter Problem using a Doubly Nonnegative Relaxation and a Splitting Method
di: Jung, Woosuk L., et al.
Pubblicazione: (2023)
di: Jung, Woosuk L., et al.
Pubblicazione: (2023)
Documenti analoghi
-
An Augmented Lagrangian Method for Training Recurrent Neural Networks
di: Wang, Yue, et al.
Pubblicazione: (2024) -
Inexact FPPA for the $\ell_0$ Sparse Regularization Problem
di: Fang, Ronglong, et al.
Pubblicazione: (2024) -
Riemannian Adaptive Regularized Newton Methods with Hölder Continuous Hessians
di: Zhang, Chenyu, et al.
Pubblicazione: (2023) -
A Layer Separation Optimization Framework for Cross-Entropy Training in Deep Learning
di: Liu, Yaru, et al.
Pubblicazione: (2026) -
Convergence of gradient descent for deep neural networks
di: Chatterjee, Sourav
Pubblicazione: (2022)