ADMM Algorithms for Residual Network Training: Convergence Analysis and Parallel Implementation
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Jintao, Li, Yifei, Xing, Wenxun |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Natural Hypergradient Descent: Algorithm Design, Convergence Analysis, and Parallel Implementation
by: Kong, Deyi, et al.
Published: (2026)
by: Kong, Deyi, et al.
Published: (2026)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
by: Lin, Junan, et al.
Published: (2026)
by: Lin, Junan, et al.
Published: (2026)
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
by: Xu, Xianliang, et al.
Published: (2024)
by: Xu, Xianliang, et al.
Published: (2024)
Joint Cooperative and Non-Cooperative Localization in WSNs with Distributed Scaled Proximal ADMM Algorithms
by: Zhu, Qiaojia, et al.
Published: (2025)
by: Zhu, Qiaojia, et al.
Published: (2025)
The ADMM-PINNs Algorithmic Framework for Nonsmooth PDE-Constrained Optimization: A Deep Learning Approach
by: Song, Yongcun, et al.
Published: (2023)
by: Song, Yongcun, et al.
Published: (2023)
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections
by: Qin, Zhen, et al.
Published: (2025)
by: Qin, Zhen, et al.
Published: (2025)
Convergence Analysis of the Wasserstein Proximal Algorithm beyond Geodesic Convexity
by: Zhu, Shuailong, et al.
Published: (2025)
by: Zhu, Shuailong, et al.
Published: (2025)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
A General Continuous-Time Formulation of Stochastic ADMM and Its Variants
by: Li, Chris Junchi
Published: (2024)
by: Li, Chris Junchi
Published: (2024)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
by: Zhu, Feng, et al.
Published: (2026)
by: Zhu, Feng, et al.
Published: (2026)
Distributed Event-Based Learning via ADMM
by: Er, Guner Dilsad, et al.
Published: (2024)
by: Er, Guner Dilsad, et al.
Published: (2024)
Bilevel Optimization under Unbounded Smoothness: A New Algorithm and Convergence Analysis
by: Hao, Jie, et al.
Published: (2024)
by: Hao, Jie, et al.
Published: (2024)
Convergence Analysis of the PAGE Stochastic Algorithm for Weakly Convex Finite-Sum Optimization
by: Condat, Laurent, et al.
Published: (2025)
by: Condat, Laurent, et al.
Published: (2025)
Adaptive Batch Size Schedules for Distributed Training of Language Models with Data and Model Parallelism
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
Convergence Rate Analysis of LION
by: Dong, Yiming, et al.
Published: (2024)
by: Dong, Yiming, et al.
Published: (2024)
Convergence Analysis for Learning Orthonormal Deep Linear Neural Networks
by: Qin, Zhen, et al.
Published: (2023)
by: Qin, Zhen, et al.
Published: (2023)
Residual-Evasive Attacks on ADMM in Distributed Optimization
by: Bruckmeier, Sabrina, et al.
Published: (2025)
by: Bruckmeier, Sabrina, et al.
Published: (2025)
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
Convergence of SGD for Training Neural Networks with Sliced Wasserstein Losses
by: Tanguy, Eloi
Published: (2023)
by: Tanguy, Eloi
Published: (2023)
Learning to accelerate distributed ADMM using graph neural networks
by: Doerks, Henri, et al.
Published: (2025)
by: Doerks, Henri, et al.
Published: (2025)
Convergence and Complexity Guarantee for Inexact First-order Riemannian Optimization Algorithms
by: Li, Yuchen, et al.
Published: (2024)
by: Li, Yuchen, et al.
Published: (2024)
A Riemannian ADMM
by: Li, Jiaxiang, et al.
Published: (2022)
by: Li, Jiaxiang, et al.
Published: (2022)
Convergence of Gradient Descent for Recurrent Neural Networks: A Nonasymptotic Analysis
by: Cayci, Semih, et al.
Published: (2024)
by: Cayci, Semih, et al.
Published: (2024)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
by: Wang, Shengbo
Published: (2026)
by: Wang, Shengbo
Published: (2026)
HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization
by: Tran, Trinh, et al.
Published: (2026)
by: Tran, Trinh, et al.
Published: (2026)
A New Convergence Analysis of Plug-and-Play Proximal Gradient Descent Under Prior Mismatch
by: Xu, Guixian, et al.
Published: (2026)
by: Xu, Guixian, et al.
Published: (2026)
CoCoA Is ADMM: Unifying Two Paradigms in Distributed Optimization
by: Wu, Runxiong, et al.
Published: (2025)
by: Wu, Runxiong, et al.
Published: (2025)
Towards Guided Descent: Optimization Algorithms for Training Neural Networks At Scale
by: Nagwekar, Ansh
Published: (2025)
by: Nagwekar, Ansh
Published: (2025)
Convergence Rate of the Last Iterate of Stochastic Proximal Algorithms
by: Vaidyan, Kevin Kurian Thomas, et al.
Published: (2026)
by: Vaidyan, Kevin Kurian Thomas, et al.
Published: (2026)
ADMM-Based Training for Spiking Neural Networks
by: Perin, Giovanni, et al.
Published: (2025)
by: Perin, Giovanni, et al.
Published: (2025)
Unified Convergence Analysis for Adaptive Optimization with Moving Average Estimator
by: Guo, Zhishuai, et al.
Published: (2021)
by: Guo, Zhishuai, et al.
Published: (2021)
The Ball-Proximal (="Broximal") Point Method: a New Algorithm, Convergence Theory, and Applications
by: Gruntkowska, Kaja, et al.
Published: (2025)
by: Gruntkowska, Kaja, et al.
Published: (2025)
ADMM for Structured Fractional Minimization
by: Yuan, Ganzhao
Published: (2024)
by: Yuan, Ganzhao
Published: (2024)
Convergence Analysis of Stochastic Gradient Descent with MCMC Estimators
by: Li, Tianyou, et al.
Published: (2023)
by: Li, Tianyou, et al.
Published: (2023)
Adaptive Algorithms with Sharp Convergence Rates for Stochastic Hierarchical Optimization
by: Gong, Xiaochuan, et al.
Published: (2025)
by: Gong, Xiaochuan, et al.
Published: (2025)
A Systems-Theoretic View on the Convergence of Algorithms under Disturbances
by: Er, Guner Dilsad, et al.
Published: (2025)
by: Er, Guner Dilsad, et al.
Published: (2025)
Linear Convergence of the Frank-Wolfe Algorithm over Product Polytopes
by: Iommazzo, Gabriele, et al.
Published: (2025)
by: Iommazzo, Gabriele, et al.
Published: (2025)
Linearly Convergent Algorithms for Nonsmooth Problems with Unknown Smooth Pieces
by: Zhang, Zhe, et al.
Published: (2025)
by: Zhang, Zhe, et al.
Published: (2025)
Convergence of Stochastic Gradient Langevin Dynamics in the Lazy Training Regime
by: Oberweis, Noah, et al.
Published: (2025)
by: Oberweis, Noah, et al.
Published: (2025)
A Provably Convergent and Practical Algorithm for Gromov--Wasserstein Optimal Transport
by: Liang, Ling, et al.
Published: (2026)
by: Liang, Ling, et al.
Published: (2026)
Similar Items
-
Natural Hypergradient Descent: Algorithm Design, Convergence Analysis, and Parallel Implementation
by: Kong, Deyi, et al.
Published: (2026) -
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
by: Lin, Junan, et al.
Published: (2026) -
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
by: Xu, Xianliang, et al.
Published: (2024) -
Joint Cooperative and Non-Cooperative Localization in WSNs with Distributed Scaled Proximal ADMM Algorithms
by: Zhu, Qiaojia, et al.
Published: (2025) -
The ADMM-PINNs Algorithmic Framework for Nonsmooth PDE-Constrained Optimization: A Deep Learning Approach
by: Song, Yongcun, et al.
Published: (2023)