Hidden Minima in Two-Layer ReLU Networks
Fuente:
arXiv
Saved in:
| Main Author: | Arjevani, Yossi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
Symmetry & Critical Points
by: Arjevani, Yossi
Published: (2024)
by: Arjevani, Yossi
Published: (2024)
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
by: Yang, Yahong, et al.
Published: (2023)
by: Yang, Yahong, et al.
Published: (2023)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
by: Kuelbs, Daniel, et al.
Published: (2024)
by: Kuelbs, Daniel, et al.
Published: (2024)
Symmetry & Critical Points for Symmetric Tensor Decomposition Problems
by: Arjevani, Yossi, et al.
Published: (2023)
by: Arjevani, Yossi, et al.
Published: (2023)
Computational Tradeoffs of Optimization-Based Bound Tightening in ReLU Networks
by: Badilla, Fabian, et al.
Published: (2023)
by: Badilla, Fabian, et al.
Published: (2023)
ReLU Networks for Model Predictive Control: Network Complexity and Performance Guarantees
by: Li, Xingchen, et al.
Published: (2026)
by: Li, Xingchen, et al.
Published: (2026)
Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks
by: Liao, Fangshuo, et al.
Published: (2023)
by: Liao, Fangshuo, et al.
Published: (2023)
Stability and Performance Analysis of Discrete-Time ReLU Recurrent Neural Networks
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
Why Smooth Stability Assumptions Fail for ReLU Learning
by: Katende, Ronald
Published: (2025)
by: Katende, Ronald
Published: (2025)
An analysis of optimization problems involving ReLU neural networks
by: Plate, Christoph, et al.
Published: (2025)
by: Plate, Christoph, et al.
Published: (2025)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
by: Kim, Sungyoon, et al.
Published: (2024)
by: Kim, Sungyoon, et al.
Published: (2024)
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
by: Lamperski, Andrew, et al.
Published: (2024)
by: Lamperski, Andrew, et al.
Published: (2024)
An Efficient Alternating Algorithm for ReLU-based Symmetric Matrix Decomposition
by: Wang, Qingsong
Published: (2025)
by: Wang, Qingsong
Published: (2025)
A Complete Set of Quadratic Constraints for Repeated ReLU and Generalizations
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
Neural Collapse under Gradient Flow on Shallow ReLU Networks for Orthogonally Separable Data
by: Min, Hancheng, et al.
Published: (2025)
by: Min, Hancheng, et al.
Published: (2025)
MIQCQP reformulation of the ReLU neural networks Lipschitz constant estimation problem
by: Sbihi, Mohammed, et al.
Published: (2024)
by: Sbihi, Mohammed, et al.
Published: (2024)
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
by: Lamperski, Andrew, et al.
Published: (2024)
by: Lamperski, Andrew, et al.
Published: (2024)
Path-conditioned training: a principled way to rescale ReLU neural networks
by: Lebeurrier, Arthur, et al.
Published: (2026)
by: Lebeurrier, Arthur, et al.
Published: (2026)
Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU
by: Liu, Jingzhou
Published: (2025)
by: Liu, Jingzhou
Published: (2025)
How Does the ReLU Activation Affect the Implicit Bias of Gradient Descent on High-dimensional Neural Network Regression?
by: Lai, Kuo-Wei, et al.
Published: (2026)
by: Lai, Kuo-Wei, et al.
Published: (2026)
Local Lipschitz Constant Computation of ReLU-FNNs: Upper Bound Computation with Exactness Verification
by: Ebihara, Yoshio, et al.
Published: (2023)
by: Ebihara, Yoshio, et al.
Published: (2023)
Geometry-induced Regularization in Deep ReLU Neural Networks
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
On bounds for norms of reparameterized ReLU artificial neural network parameters: sums of fractional powers of the Lipschitz norm control the network parameter vector
by: Jentzen, Arnulf, et al.
Published: (2022)
by: Jentzen, Arnulf, et al.
Published: (2022)
Constructive Universal Approximation and Finite Sample Memorization by Narrow Deep ReLU Networks
by: Hernández, Martín, et al.
Published: (2024)
by: Hernández, Martín, et al.
Published: (2024)
Sharpness of Minima in Deep Matrix Factorization
by: Kamber, Anil, et al.
Published: (2025)
by: Kamber, Anil, et al.
Published: (2025)
Non-asymptotic convergence analysis of the stochastic gradient Hamiltonian Monte Carlo algorithm with discontinuous stochastic gradient with applications to training of ReLU neural networks
by: Liang, Luxu, et al.
Published: (2024)
by: Liang, Luxu, et al.
Published: (2024)
ReLU Surrogates in Mixed-Integer MPC for Irrigation Scheduling
by: Agyeman, Bernard T., et al.
Published: (2024)
by: Agyeman, Bernard T., et al.
Published: (2024)
On the existence of optimal shallow feedforward networks with ReLU activation
by: Dereich, Steffen, et al.
Published: (2023)
by: Dereich, Steffen, et al.
Published: (2023)
On the ReLU Lagrangian Cuts for Stochastic Mixed Integer Programming
by: Deng, Haoyun, et al.
Published: (2024)
by: Deng, Haoyun, et al.
Published: (2024)
LMI hierarchies for stability analysis of ReLU feedback systems
by: Magron, Victor, et al.
Published: (2024)
by: Magron, Victor, et al.
Published: (2024)
Architecture independent generalization bounds for overparametrized deep ReLU networks
by: Bapu, Anandatheertha, et al.
Published: (2025)
by: Bapu, Anandatheertha, et al.
Published: (2025)
Gradient Descent Converges Linearly to Flatter Minima than Gradient Flow in Shallow Linear Networks
by: Beneventano, Pierfrancesco, et al.
Published: (2025)
by: Beneventano, Pierfrancesco, et al.
Published: (2025)
LoRA Training in the NTK Regime has No Spurious Local Minima
by: Jang, Uijeong, et al.
Published: (2024)
by: Jang, Uijeong, et al.
Published: (2024)
Normalization of ReLU Dual for Cut Generation in Stochastic Mixed-Integer Programs
by: Bansal, Akul, et al.
Published: (2026)
by: Bansal, Akul, et al.
Published: (2026)
Simplicity Bias of Two-Layer Networks beyond Linearly Separable Data
by: Tsoy, Nikita, et al.
Published: (2024)
by: Tsoy, Nikita, et al.
Published: (2024)
On the existence of minimizers in shallow residual ReLU neural network optimization landscapes
by: Dereich, Steffen, et al.
Published: (2023)
by: Dereich, Steffen, et al.
Published: (2023)
Gradient Descent with Polyak's Momentum Finds Flatter Minima via Large Catapults
by: Phunyaphibarn, Prin, et al.
Published: (2023)
by: Phunyaphibarn, Prin, et al.
Published: (2023)
Interpretable global minima of deep ReLU neural networks on sequentially separable data
by: Chen, Thomas, et al.
Published: (2024)
by: Chen, Thomas, et al.
Published: (2024)
Nonnegative Low-rank Matrix Recovery Can Have Spurious Local Minima
by: Zhang, Richard Y.
Published: (2025)
by: Zhang, Richard Y.
Published: (2025)
Similar Items
-
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024) -
Symmetry & Critical Points
by: Arjevani, Yossi
Published: (2024) -
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
by: Yang, Yahong, et al.
Published: (2023) -
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
by: Kuelbs, Daniel, et al.
Published: (2024) -
Symmetry & Critical Points for Symmetric Tensor Decomposition Problems
by: Arjevani, Yossi, et al.
Published: (2023)