Saved in:
| Main Authors: | Nenov, Rossen, Haider, Daniel, Balazs, Peter |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.00169 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Smoothing Methods for Automatic Differentiation Across Conditional Branches
by: Kreikemeyer, Justin N., et al.
Published: (2023)
by: Kreikemeyer, Justin N., et al.
Published: (2023)
Smoothing the Edges: Smooth Optimization for Sparse Regularization using Hadamard Overparametrization
by: Kolb, Chris, et al.
Published: (2023)
by: Kolb, Chris, et al.
Published: (2023)
Implicit Regularization in Perturbed Deep Matrix Factorization: Spectral Conditions and Stability
by: Wang, Jingzhe, et al.
Published: (2026)
by: Wang, Jingzhe, et al.
Published: (2026)
Physics-Informed Graph Neural Network for Dynamic Reconfiguration of Power Systems
by: Authier, Jules, et al.
Published: (2023)
by: Authier, Jules, et al.
Published: (2023)
Enhancing Stability of Physics-Informed Neural Network Training Through Saddle-Point Reformulation
by: Bylinkin, Dmitry, et al.
Published: (2025)
by: Bylinkin, Dmitry, et al.
Published: (2025)
Efficiently Escaping Saddle Points under Generalized Smoothness via Self-Bounding Regularity
by: Cao, Daniel Yiming, et al.
Published: (2025)
by: Cao, Daniel Yiming, et al.
Published: (2025)
Fixed-Point Automatic Differentiation of Forward--Backward Splitting Algorithms for Partly Smooth Functions
by: Mehmood, Sheheryar, et al.
Published: (2022)
by: Mehmood, Sheheryar, et al.
Published: (2022)
Almost Sure Convergence Analysis of Differentially Private Stochastic Gradient Methods
by: Mukherjee, Amartya, et al.
Published: (2025)
by: Mukherjee, Amartya, et al.
Published: (2025)
A full splitting algorithm for structured difference-of-convex programs
by: Bot, Radu Ioan, et al.
Published: (2025)
by: Bot, Radu Ioan, et al.
Published: (2025)
Fair Supervised Learning Through Constraints on Smooth Nonconvex Unfairness-Measure Surrogates
by: Khatti, Zahra, et al.
Published: (2025)
by: Khatti, Zahra, et al.
Published: (2025)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
by: Adeoye, Adeyemi D., et al.
Published: (2024)
by: Adeoye, Adeyemi D., et al.
Published: (2024)
Stability and Performance Analysis of Discrete-Time ReLU Recurrent Neural Networks
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
Stability Regularized Cross-Validation
by: Cory-Wright, Ryan, et al.
Published: (2025)
by: Cory-Wright, Ryan, et al.
Published: (2025)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
by: Beneventano, Pierfrancesco, et al.
Published: (2024)
by: Beneventano, Pierfrancesco, et al.
Published: (2024)
Regularized Gradient Clipping Provably Trains Wide and Deep Neural Networks
by: Tucat, Matteo, et al.
Published: (2024)
by: Tucat, Matteo, et al.
Published: (2024)
Neural Network Training Techniques Regularize Optimization Trajectory: An Empirical Study
by: Chen, Cheng, et al.
Published: (2020)
by: Chen, Cheng, et al.
Published: (2020)
Why Smooth Stability Assumptions Fail for ReLU Learning
by: Katende, Ronald
Published: (2025)
by: Katende, Ronald
Published: (2025)
On the Complexity of Finite-Sum Smooth Optimization under the Polyak-Łojasiewicz Condition
by: Bai, Yunyan, et al.
Published: (2024)
by: Bai, Yunyan, et al.
Published: (2024)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
by: Cao, John, et al.
Published: (2025)
by: Cao, John, et al.
Published: (2025)
A Study of Condition Numbers for First-Order Optimization
by: Guille-Escuret, Charles, et al.
Published: (2020)
by: Guille-Escuret, Charles, et al.
Published: (2020)
Towards Quantifying the Hessian Structure of Neural Networks
by: Dong, Zhaorui, et al.
Published: (2025)
by: Dong, Zhaorui, et al.
Published: (2025)
Stability of Data-Dependent Ridge-Regularization for Inverse Problems
by: Neumayer, Sebastian, et al.
Published: (2024)
by: Neumayer, Sebastian, et al.
Published: (2024)
Recent Advances in Non-convex Smoothness Conditions and Applicability to Deep Linear Neural Networks
by: Patel, Vivak, et al.
Published: (2024)
by: Patel, Vivak, et al.
Published: (2024)
Regularized Adaptive Momentum Dual Averaging with an Efficient Inexact Subproblem Solver for Training Structured Neural Network
by: Huang, Zih-Syuan, et al.
Published: (2024)
by: Huang, Zih-Syuan, et al.
Published: (2024)
Almost Surely $\sqrt{T}$ Regret for Adaptive LQR
by: Lu, Yiwen, et al.
Published: (2023)
by: Lu, Yiwen, et al.
Published: (2023)
Differentiation Through Black-Box Quadratic Programming Solvers
by: Magoon, Connor W., et al.
Published: (2024)
by: Magoon, Connor W., et al.
Published: (2024)
An Adaptive and Stability-Promoting Layerwise Training Approach for Sparse Deep Neural Network Architecture
by: Krishnanunni, C G, et al.
Published: (2022)
by: Krishnanunni, C G, et al.
Published: (2022)
ATE-SG: Alternate Through the Epochs Stochastic Gradient for Multi-Task Neural Networks
by: Bellavia, Stefania, et al.
Published: (2023)
by: Bellavia, Stefania, et al.
Published: (2023)
A Unified Lyapunov-IQC Framework for Uniform Stability of Smooth Quadratic First-Order Accelerated Optimizers
by: Li, Don, et al.
Published: (2026)
by: Li, Don, et al.
Published: (2026)
Towards Guided Descent: Optimization Algorithms for Training Neural Networks At Scale
by: Nagwekar, Ansh
Published: (2025)
by: Nagwekar, Ansh
Published: (2025)
On the Condition Number Dependency in Bilevel Optimization
by: Chen, Lesi, et al.
Published: (2025)
by: Chen, Lesi, et al.
Published: (2025)
A Unified Theory of Stochastic Proximal Point Methods without Smoothness
by: Richtárik, Peter, et al.
Published: (2024)
by: Richtárik, Peter, et al.
Published: (2024)
Wide Neural Networks Trained with Weight Decay Provably Exhibit Neural Collapse
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Kernel Learning with Adversarial Features: Numerical Efficiency and Adaptive Regularization
by: Ribeiro, Antônio H., et al.
Published: (2025)
by: Ribeiro, Antônio H., et al.
Published: (2025)
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024)
by: Schmitt-Förster, Peter, et al.
Published: (2024)
Towards Optimal Branching of Linear and Semidefinite Relaxations for Neural Network Robustness Certification
by: Anderson, Brendon G., et al.
Published: (2021)
by: Anderson, Brendon G., et al.
Published: (2021)
Towards Understanding Gradient Flow Dynamics of Homogeneous Neural Networks Beyond the Origin
by: Kumar, Akshay, et al.
Published: (2025)
by: Kumar, Akshay, et al.
Published: (2025)
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025)
by: Kassing, Sebastian, et al.
Published: (2025)
A Penalty Approach for Differentiation Through Black-Box Quadratic Programming Solvers
by: Linghu, Yuxuan, et al.
Published: (2026)
by: Linghu, Yuxuan, et al.
Published: (2026)
Dynamic Memory Based Adaptive Optimization
by: Szegedy, Balázs, et al.
Published: (2024)
by: Szegedy, Balázs, et al.
Published: (2024)
Similar Items
-
Smoothing Methods for Automatic Differentiation Across Conditional Branches
by: Kreikemeyer, Justin N., et al.
Published: (2023) -
Smoothing the Edges: Smooth Optimization for Sparse Regularization using Hadamard Overparametrization
by: Kolb, Chris, et al.
Published: (2023) -
Implicit Regularization in Perturbed Deep Matrix Factorization: Spectral Conditions and Stability
by: Wang, Jingzhe, et al.
Published: (2026) -
Physics-Informed Graph Neural Network for Dynamic Reconfiguration of Power Systems
by: Authier, Jules, et al.
Published: (2023) -
Enhancing Stability of Physics-Informed Neural Network Training Through Saddle-Point Reformulation
by: Bylinkin, Dmitry, et al.
Published: (2025)