Neural Collapse under Gradient Flow on Shallow ReLU Networks for Orthogonally Separable Data
Fuente:
arXiv
Saved in:
| Main Authors: | Min, Hancheng, Zhu, Zhihui, Vidal, René |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
by: Lamperski, Andrew, et al.
Published: (2024)
by: Lamperski, Andrew, et al.
Published: (2024)
Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU
by: Liu, Jingzhou
Published: (2025)
by: Liu, Jingzhou
Published: (2025)
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
by: Lamperski, Andrew, et al.
Published: (2024)
by: Lamperski, Andrew, et al.
Published: (2024)
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
Hidden Minima in Two-Layer ReLU Networks
by: Arjevani, Yossi
Published: (2023)
by: Arjevani, Yossi
Published: (2023)
Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks
by: Liao, Fangshuo, et al.
Published: (2023)
by: Liao, Fangshuo, et al.
Published: (2023)
Stability and Performance Analysis of Discrete-Time ReLU Recurrent Neural Networks
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
How Does the ReLU Activation Affect the Implicit Bias of Gradient Descent on High-dimensional Neural Network Regression?
by: Lai, Kuo-Wei, et al.
Published: (2026)
by: Lai, Kuo-Wei, et al.
Published: (2026)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
by: Kim, Sungyoon, et al.
Published: (2024)
by: Kim, Sungyoon, et al.
Published: (2024)
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
by: Yang, Yahong, et al.
Published: (2023)
by: Yang, Yahong, et al.
Published: (2023)
Computational Tradeoffs of Optimization-Based Bound Tightening in ReLU Networks
by: Badilla, Fabian, et al.
Published: (2023)
by: Badilla, Fabian, et al.
Published: (2023)
ReLU Networks for Model Predictive Control: Network Complexity and Performance Guarantees
by: Li, Xingchen, et al.
Published: (2026)
by: Li, Xingchen, et al.
Published: (2026)
Why Smooth Stability Assumptions Fail for ReLU Learning
by: Katende, Ronald
Published: (2025)
by: Katende, Ronald
Published: (2025)
An analysis of optimization problems involving ReLU neural networks
by: Plate, Christoph, et al.
Published: (2025)
by: Plate, Christoph, et al.
Published: (2025)
A Local Polyak-Lojasiewicz and Descent Lemma of Gradient Descent For Overparametrized Linear Models
by: Xu, Ziqing, et al.
Published: (2025)
by: Xu, Ziqing, et al.
Published: (2025)
An Efficient Alternating Algorithm for ReLU-based Symmetric Matrix Decomposition
by: Wang, Qingsong
Published: (2025)
by: Wang, Qingsong
Published: (2025)
A Complete Set of Quadratic Constraints for Repeated ReLU and Generalizations
by: Noori, Sahel Vahedi, et al.
Published: (2024)
by: Noori, Sahel Vahedi, et al.
Published: (2024)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
by: Kuelbs, Daniel, et al.
Published: (2024)
by: Kuelbs, Daniel, et al.
Published: (2024)
Early Neuron Alignment in Two-layer ReLU Networks with Small Initialization
by: Min, Hancheng, et al.
Published: (2023)
by: Min, Hancheng, et al.
Published: (2023)
MIQCQP reformulation of the ReLU neural networks Lipschitz constant estimation problem
by: Sbihi, Mohammed, et al.
Published: (2024)
by: Sbihi, Mohammed, et al.
Published: (2024)
Geometry-induced Regularization in Deep ReLU Neural Networks
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
Path-conditioned training: a principled way to rescale ReLU neural networks
by: Lebeurrier, Arthur, et al.
Published: (2026)
by: Lebeurrier, Arthur, et al.
Published: (2026)
Convergence Rates for Gradient Descent on the Edge of Stability in Overparametrised Least Squares
by: MacDonald, Lachlan Ewen, et al.
Published: (2025)
by: MacDonald, Lachlan Ewen, et al.
Published: (2025)
Local Lipschitz Constant Computation of ReLU-FNNs: Upper Bound Computation with Exactness Verification
by: Ebihara, Yoshio, et al.
Published: (2023)
by: Ebihara, Yoshio, et al.
Published: (2023)
Understanding Incremental Learning with Closed-form Solution to Gradient Flow on Overparamerterized Matrix Factorization
by: Min, Hancheng, et al.
Published: (2025)
by: Min, Hancheng, et al.
Published: (2025)
Gradient Descent Converges Linearly to Flatter Minima than Gradient Flow in Shallow Linear Networks
by: Beneventano, Pierfrancesco, et al.
Published: (2025)
by: Beneventano, Pierfrancesco, et al.
Published: (2025)
On bounds for norms of reparameterized ReLU artificial neural network parameters: sums of fractional powers of the Lipschitz norm control the network parameter vector
by: Jentzen, Arnulf, et al.
Published: (2022)
by: Jentzen, Arnulf, et al.
Published: (2022)
Convergence Analysis for Learning Orthonormal Deep Linear Neural Networks
by: Qin, Zhen, et al.
Published: (2023)
by: Qin, Zhen, et al.
Published: (2023)
Wasserstein Distributionally Robust Shallow Convex Neural Networks
by: Pallage, Julien, et al.
Published: (2024)
by: Pallage, Julien, et al.
Published: (2024)
Constructive Universal Approximation and Finite Sample Memorization by Narrow Deep ReLU Networks
by: Hernández, Martín, et al.
Published: (2024)
by: Hernández, Martín, et al.
Published: (2024)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections
by: Qin, Zhen, et al.
Published: (2025)
by: Qin, Zhen, et al.
Published: (2025)
Towards Understanding Gradient Flow Dynamics of Homogeneous Neural Networks Beyond the Origin
by: Kumar, Akshay, et al.
Published: (2025)
by: Kumar, Akshay, et al.
Published: (2025)
Implicit Bias of Mirror Flow on Separable Data
by: Pesme, Scott, et al.
Published: (2024)
by: Pesme, Scott, et al.
Published: (2024)
Wide Neural Networks Trained with Weight Decay Provably Exhibit Neural Collapse
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Non-asymptotic convergence analysis of the stochastic gradient Hamiltonian Monte Carlo algorithm with discontinuous stochastic gradient with applications to training of ReLU neural networks
by: Liang, Luxu, et al.
Published: (2024)
by: Liang, Luxu, et al.
Published: (2024)
Gradient descent provably escapes saddle points in the training of shallow ReLU networks
by: Cheridito, Patrick, et al.
Published: (2022)
by: Cheridito, Patrick, et al.
Published: (2022)
The Exploration of Neural Collapse under Imbalanced Data
by: Liu, Haixia
Published: (2024)
by: Liu, Haixia
Published: (2024)
Neural Collapse versus Low-rank Bias: Is Deep Neural Collapse Really Optimal?
by: Súkeník, Peter, et al.
Published: (2024)
by: Súkeník, Peter, et al.
Published: (2024)
ReLU Surrogates in Mixed-Integer MPC for Irrigation Scheduling
by: Agyeman, Bernard T., et al.
Published: (2024)
by: Agyeman, Bernard T., et al.
Published: (2024)
Similar Items
-
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
by: Lamperski, Andrew, et al.
Published: (2024) -
Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU
by: Liu, Jingzhou
Published: (2025) -
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
by: Lamperski, Andrew, et al.
Published: (2024) -
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024) -
Hidden Minima in Two-Layer ReLU Networks
by: Arjevani, Yossi
Published: (2023)