Convergence Analysis of Newton's Method for Neural Networks in the Overparameterized Limit
Fuente:
arXiv
Guardado en:
| Autores principales: | Riedl, Konstantin, Spiliopoulos, Konstantinos, Sirignano, Justin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Global Convergence of Adjoint-Optimized Neural PDEs
por: Riedl, Konstantin, et al.
Publicado: (2025)
por: Riedl, Konstantin, et al.
Publicado: (2025)
Convergence Analysis of Real-time Recurrent Learning (RTRL) for a class of Recurrent Neural Networks
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2025)
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2025)
Scaling Effects and Uncertainty Quantification in Neural Actor Critic Algorithms
por: Georgoudios, Nikos, et al.
Publicado: (2026)
por: Georgoudios, Nikos, et al.
Publicado: (2026)
Kernel Limit for a Class of Recurrent Neural Networks Trained on Ergodic Data Sequences
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2023)
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2023)
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2024)
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2024)
Transport map unadjusted Langevin algorithms: learning and discretizing perturbed samplers
por: Zhang, Benjamin J., et al.
Publicado: (2023)
por: Zhang, Benjamin J., et al.
Publicado: (2023)
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
por: Chemnitz, Dennis, et al.
Publicado: (2024)
por: Chemnitz, Dennis, et al.
Publicado: (2024)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
por: Adeoye, Adeyemi D., et al.
Publicado: (2024)
por: Adeoye, Adeyemi D., et al.
Publicado: (2024)
Convergence of SGD for Training Neural Networks with Sliced Wasserstein Losses
por: Tanguy, Eloi
Publicado: (2023)
por: Tanguy, Eloi
Publicado: (2023)
ResNets of All Shapes and Sizes: Convergence of Training Dynamics in the Large-scale Limit
por: Chaintron, Louis-Pierre, et al.
Publicado: (2026)
por: Chaintron, Louis-Pierre, et al.
Publicado: (2026)
Uniform-in-time quantitative fluctuations of large scale interacting particle systems
por: Bourguin, Solesne, et al.
Publicado: (2026)
por: Bourguin, Solesne, et al.
Publicado: (2026)
Quantitative fluctuation analysis of multiscale diffusion systems via Malliavin calculus
por: Bourguin, Solesne, et al.
Publicado: (2023)
por: Bourguin, Solesne, et al.
Publicado: (2023)
Central Limit Theorem for Bayesian Neural Network trained with Variational Inference
por: Descours, Arnaud, et al.
Publicado: (2024)
por: Descours, Arnaud, et al.
Publicado: (2024)
Structural and Convergence Analysis of Discrete-Time Denoising Diffusion Probabilistic Models
por: Nakano, Yumiharu
Publicado: (2024)
por: Nakano, Yumiharu
Publicado: (2024)
Convergence Analysis for General Probability Flow ODEs of Diffusion Models in Wasserstein Distances
por: Gao, Xuefeng, et al.
Publicado: (2024)
por: Gao, Xuefeng, et al.
Publicado: (2024)
Advancing Wasserstein Convergence Analysis of Score-Based Models: Insights from Discretization and Second-Order Acceleration
por: Yu, Yifeng, et al.
Publicado: (2025)
por: Yu, Yifeng, et al.
Publicado: (2025)
From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes
por: Chen, Zaiwei, et al.
Publicado: (2025)
por: Chen, Zaiwei, et al.
Publicado: (2025)
Partially Stochastic Infinitely Deep Bayesian Neural Networks
por: Calvo-Ordonez, Sergio, et al.
Publicado: (2024)
por: Calvo-Ordonez, Sergio, et al.
Publicado: (2024)
Exponential Convergence Guarantees for Iterative Markovian Fitting
por: Silveri, Marta Gentiloni, et al.
Publicado: (2025)
por: Silveri, Marta Gentiloni, et al.
Publicado: (2025)
Depth Degeneracy in Neural Networks: Vanishing Angles in Fully Connected ReLU Networks on Initialization
por: Jakub, Cameron, et al.
Publicado: (2023)
por: Jakub, Cameron, et al.
Publicado: (2023)
Symmetries in Overparametrized Neural Networks: A Mean-Field View
por: Maass, Javier, et al.
Publicado: (2024)
por: Maass, Javier, et al.
Publicado: (2024)
Random ReLU Neural Networks as Non-Gaussian Processes
por: Parhi, Rahul, et al.
Publicado: (2024)
por: Parhi, Rahul, et al.
Publicado: (2024)
Deep Learning for Computing Convergence Rates of Markov Chains
por: Qu, Yanlin, et al.
Publicado: (2024)
por: Qu, Yanlin, et al.
Publicado: (2024)
Toward Dynamic Stability Assessment of Power Grid Topologies using Graph Neural Networks
por: Nauck, Christian, et al.
Publicado: (2022)
por: Nauck, Christian, et al.
Publicado: (2022)
Stochastic Port-Hamiltonian Neural Networks: Universal Approximation with Passivity Guarantees
por: Di Persio, Luca, et al.
Publicado: (2026)
por: Di Persio, Luca, et al.
Publicado: (2026)
Universality in Deep Neural Networks: An approach via the Lindeberg exchange principle
por: Giovagnini, Filippo, et al.
Publicado: (2026)
por: Giovagnini, Filippo, et al.
Publicado: (2026)
Deep Neural Networks as Iterated Function Systems and a Generalization Bound
por: Vacher, Jonathan
Publicado: (2026)
por: Vacher, Jonathan
Publicado: (2026)
Implicit Compressibility of Overparametrized Neural Networks Trained with Heavy-Tailed SGD
por: Wan, Yijun, et al.
Publicado: (2023)
por: Wan, Yijun, et al.
Publicado: (2023)
Exact Gradients for Stochastic Spiking Neural Networks Driven by Rough Signals
por: Holberg, Christian, et al.
Publicado: (2024)
por: Holberg, Christian, et al.
Publicado: (2024)
Some Theoretical Limitations of t-SNE
por: Li, Rupert, et al.
Publicado: (2026)
por: Li, Rupert, et al.
Publicado: (2026)
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
por: Chen, Zaiwei, et al.
Publicado: (2026)
por: Chen, Zaiwei, et al.
Publicado: (2026)
Convergence, Sticking and Escape: Stochastic Dynamics Near Critical Points in SGD
por: Dudukalov, Dmitry, et al.
Publicado: (2025)
por: Dudukalov, Dmitry, et al.
Publicado: (2025)
Uniform-in-time bounds for a stochastic hybrid system with fast periodic sampling and small white-noise
por: Dhama, Shivam Singh, et al.
Publicado: (2024)
por: Dhama, Shivam Singh, et al.
Publicado: (2024)
On the Limits of Latent Reuse in Diffusion Models
por: Yu, Yifeng, et al.
Publicado: (2026)
por: Yu, Yifeng, et al.
Publicado: (2026)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
por: Srikant, R.
Publicado: (2024)
por: Srikant, R.
Publicado: (2024)
Wasserstein Convergence Guarantees for a General Class of Score-Based Generative Models
por: Gao, Xuefeng, et al.
Publicado: (2023)
por: Gao, Xuefeng, et al.
Publicado: (2023)
Neural Network Methods for Radiation Detectors and Imaging
por: Lin, S., et al.
Publicado: (2023)
por: Lin, S., et al.
Publicado: (2023)
Uniform Scaling Limits in AdamW-Trained Transformers
por: Gibson, William, et al.
Publicado: (2026)
por: Gibson, William, et al.
Publicado: (2026)
Limit Theorems for Stochastic Gradient Descent with Infinite Variance
por: Blanchet, Jose, et al.
Publicado: (2024)
por: Blanchet, Jose, et al.
Publicado: (2024)
Convergence of Unadjusted Langevin in High Dimensions: Delocalization of Bias
por: Chen, Yifan, et al.
Publicado: (2024)
por: Chen, Yifan, et al.
Publicado: (2024)
Ejemplares similares
-
Global Convergence of Adjoint-Optimized Neural PDEs
por: Riedl, Konstantin, et al.
Publicado: (2025) -
Convergence Analysis of Real-time Recurrent Learning (RTRL) for a class of Recurrent Neural Networks
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2025) -
Scaling Effects and Uncertainty Quantification in Neural Actor Critic Algorithms
por: Georgoudios, Nikos, et al.
Publicado: (2026) -
Kernel Limit for a Class of Recurrent Neural Networks Trained on Ergodic Data Sequences
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2023) -
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
por: Lam, Samuel Chun-Hei, et al.
Publicado: (2024)