Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiao, Dan, Zhang, Kaiqi, Singh, Esha, Soudry, Daniel, Wang, Yu-Xiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
di: Qiao, Dan, et al.
Pubblicazione: (2025)
di: Qiao, Dan, et al.
Pubblicazione: (2025)
Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
di: Park, Junhyung, et al.
Pubblicazione: (2024)
di: Park, Junhyung, et al.
Pubblicazione: (2024)
Stable Minima of ReLU Neural Networks Suffer from the Curse of Dimensionality: The Neural Shattering Phenomenon
di: Liang, Tongtong, et al.
Pubblicazione: (2025)
di: Liang, Tongtong, et al.
Pubblicazione: (2025)
The Geometry of ReLU Networks through the ReLU Transition Graph
di: Dhayalkar, Sahil Rajesh
Pubblicazione: (2025)
di: Dhayalkar, Sahil Rajesh
Pubblicazione: (2025)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
di: Manik, Md Motaleb Hossen, et al.
Pubblicazione: (2025)
di: Manik, Md Motaleb Hossen, et al.
Pubblicazione: (2025)
The Resurrection of the ReLU
di: Horuz, Coşku Can, et al.
Pubblicazione: (2025)
di: Horuz, Coşku Can, et al.
Pubblicazione: (2025)
Expressive Power of ReLU and Step Networks under Floating-Point Operations
di: Park, Yeachan, et al.
Pubblicazione: (2024)
di: Park, Yeachan, et al.
Pubblicazione: (2024)
Pathwise Explanation of ReLU Neural Networks
di: Lim, Seongwoo, et al.
Pubblicazione: (2025)
di: Lim, Seongwoo, et al.
Pubblicazione: (2025)
Activation-Descent Regularization for Input Optimization of ReLU Networks
di: Yu, Hongzhan, et al.
Pubblicazione: (2024)
di: Yu, Hongzhan, et al.
Pubblicazione: (2024)
ReLU Networks for Exact Generation of Similar Graphs
di: Ghafoor, Mamoona, et al.
Pubblicazione: (2026)
di: Ghafoor, Mamoona, et al.
Pubblicazione: (2026)
Three Quantization Regimes for ReLU Networks
di: Ou, Weigutian, et al.
Pubblicazione: (2024)
di: Ou, Weigutian, et al.
Pubblicazione: (2024)
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
di: Schoots, Nandi, et al.
Pubblicazione: (2025)
di: Schoots, Nandi, et al.
Pubblicazione: (2025)
RePO: Understanding Preference Learning Through ReLU-Based Optimization
di: Wu, Junkang, et al.
Pubblicazione: (2025)
di: Wu, Junkang, et al.
Pubblicazione: (2025)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
di: Yazdannik, Saman, et al.
Pubblicazione: (2025)
di: Yazdannik, Saman, et al.
Pubblicazione: (2025)
Detecting Invariant Manifolds in ReLU-Based RNNs
di: Eisenmann, Lukas, et al.
Pubblicazione: (2025)
di: Eisenmann, Lukas, et al.
Pubblicazione: (2025)
Is ReLU Adversarially Robust?
di: Sooksatra, Korn, et al.
Pubblicazione: (2024)
di: Sooksatra, Korn, et al.
Pubblicazione: (2024)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
di: Dana, Léo, et al.
Pubblicazione: (2025)
di: Dana, Léo, et al.
Pubblicazione: (2025)
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models
di: Jha, Nandan Kumar, et al.
Pubblicazione: (2024)
di: Jha, Nandan Kumar, et al.
Pubblicazione: (2024)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
di: Ye, Longqing
Pubblicazione: (2025)
di: Ye, Longqing
Pubblicazione: (2025)
Hidden Minima in Two-Layer ReLU Networks
di: Arjevani, Yossi
Pubblicazione: (2023)
di: Arjevani, Yossi
Pubblicazione: (2023)
Noisy Interpolation Learning with Shallow Univariate ReLU Networks
di: Joshi, Nirmit, et al.
Pubblicazione: (2023)
di: Joshi, Nirmit, et al.
Pubblicazione: (2023)
Topological Signatures of ReLU Neural Network Activation Patterns
di: Bosca, Vicente, et al.
Pubblicazione: (2025)
di: Bosca, Vicente, et al.
Pubblicazione: (2025)
$λ$-GELU: Learning Gating Hardness for Controlled ReLU-ization in Deep Networks
di: Pérez-Corral, Cristian, et al.
Pubblicazione: (2026)
di: Pérez-Corral, Cristian, et al.
Pubblicazione: (2026)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
di: Zhang, Zhengyan, et al.
Pubblicazione: (2024)
di: Zhang, Zhengyan, et al.
Pubblicazione: (2024)
Deep ReLU Networks Have Surprisingly Simple Polytopes
di: Fan, Feng-Lei, et al.
Pubblicazione: (2023)
di: Fan, Feng-Lei, et al.
Pubblicazione: (2023)
Precise Verification of Transformers through ReLU-Catalyzed Abstraction Refinement
di: Liu, Hengjie, et al.
Pubblicazione: (2026)
di: Liu, Hengjie, et al.
Pubblicazione: (2026)
Uncovering Layer-Dependent Activation Sparsity Patterns in ReLU Transformers
di: Wild, Cody, et al.
Pubblicazione: (2024)
di: Wild, Cody, et al.
Pubblicazione: (2024)
A Lower Bound for the Number of Linear Regions of Ternary ReLU Regression Neural Networks
di: Nakahara, Yuta, et al.
Pubblicazione: (2025)
di: Nakahara, Yuta, et al.
Pubblicazione: (2025)
Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During Training
di: Milkert, Max, et al.
Pubblicazione: (2023)
di: Milkert, Max, et al.
Pubblicazione: (2023)
The Median is Easier than it Looks: Approximation with a Constant-Depth, Linear-Width ReLU Network
di: Dutta, Abhigyan, et al.
Pubblicazione: (2026)
di: Dutta, Abhigyan, et al.
Pubblicazione: (2026)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
di: Bantzis, Ioannis, et al.
Pubblicazione: (2025)
di: Bantzis, Ioannis, et al.
Pubblicazione: (2025)
Geometry-induced Regularization in Deep ReLU Neural Networks
di: Bona-Pellissier, Joachim, et al.
Pubblicazione: (2024)
di: Bona-Pellissier, Joachim, et al.
Pubblicazione: (2024)
From Tempered to Benign Overfitting in ReLU Neural Networks
di: Kornowski, Guy, et al.
Pubblicazione: (2023)
di: Kornowski, Guy, et al.
Pubblicazione: (2023)
The Cost of Robustness: Tighter Bounds on Parameter Complexity for Robust Memorization in ReLU Nets
di: Kim, Yujun, et al.
Pubblicazione: (2025)
di: Kim, Yujun, et al.
Pubblicazione: (2025)
Algebraic Approach to Ridge-Regularized Mean Squared Error Minimization in Minimal ReLU Neural Network
di: Fukasaku, Ryoya, et al.
Pubblicazione: (2025)
di: Fukasaku, Ryoya, et al.
Pubblicazione: (2025)
Uncertainty Quantification with Bayesian Higher Order ReLU KANs
di: Giroux, James, et al.
Pubblicazione: (2024)
di: Giroux, James, et al.
Pubblicazione: (2024)
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
On the Principles of ReLU Networks with One Hidden Layer
di: Huang, Changcun
Pubblicazione: (2024)
di: Huang, Changcun
Pubblicazione: (2024)
HiQ-Lip: A Hierarchical Quantum-Classical Method for Global Lipschitz Constant Estimation of ReLU Networks
di: He, Haoqi, et al.
Pubblicazione: (2025)
di: He, Haoqi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
di: Qiao, Dan, et al.
Pubblicazione: (2025) -
Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
di: Park, Junhyung, et al.
Pubblicazione: (2024) -
Stable Minima of ReLU Neural Networks Suffer from the Curse of Dimensionality: The Neural Shattering Phenomenon
di: Liang, Tongtong, et al.
Pubblicazione: (2025) -
The Geometry of ReLU Networks through the ReLU Transition Graph
di: Dhayalkar, Sahil Rajesh
Pubblicazione: (2025) -
N-ReLU: Zero-Mean Stochastic Extension of ReLU
di: Manik, Md Motaleb Hossen, et al.
Pubblicazione: (2025)