Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Junhyung, Bloebaum, Patrick, Kasiviswanathan, Shiva Prasad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Classical View on Benign Overfitting: The Role of Sample Size
von: Park, Junhyung, et al.
Veröffentlicht: (2025)
von: Park, Junhyung, et al.
Veröffentlicht: (2025)
Initialization Matters: On the Benign Overfitting of Two-Layer ReLU CNN with Fully Trainable Layers
von: Shang, Shuning, et al.
Veröffentlicht: (2024)
von: Shang, Shuning, et al.
Veröffentlicht: (2024)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
von: Qiao, Dan, et al.
Veröffentlicht: (2024)
von: Qiao, Dan, et al.
Veröffentlicht: (2024)
A Quantitative Characterization of Forgetting in Post-Training
von: Balasubramanian, Krishnakumar, et al.
Veröffentlicht: (2026)
von: Balasubramanian, Krishnakumar, et al.
Veröffentlicht: (2026)
From Tempered to Benign Overfitting in ReLU Neural Networks
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
von: Manik, Md Motaleb Hossen, et al.
Veröffentlicht: (2025)
von: Manik, Md Motaleb Hossen, et al.
Veröffentlicht: (2025)
The Geometry of ReLU Networks through the ReLU Transition Graph
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2025)
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2025)
The Resurrection of the ReLU
von: Horuz, Coşku Can, et al.
Veröffentlicht: (2025)
von: Horuz, Coşku Can, et al.
Veröffentlicht: (2025)
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
von: Qiao, Dan, et al.
Veröffentlicht: (2025)
von: Qiao, Dan, et al.
Veröffentlicht: (2025)
Expressive Power of ReLU and Step Networks under Floating-Point Operations
von: Park, Yeachan, et al.
Veröffentlicht: (2024)
von: Park, Yeachan, et al.
Veröffentlicht: (2024)
Pathwise Explanation of ReLU Neural Networks
von: Lim, Seongwoo, et al.
Veröffentlicht: (2025)
von: Lim, Seongwoo, et al.
Veröffentlicht: (2025)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
von: Ye, Longqing
Veröffentlicht: (2025)
von: Ye, Longqing
Veröffentlicht: (2025)
Training Large Language Models To Reason In Parallel With Global Forking Tokens
von: Jia, Sheng, et al.
Veröffentlicht: (2025)
von: Jia, Sheng, et al.
Veröffentlicht: (2025)
A Lower Bound for the Number of Linear Regions of Ternary ReLU Regression Neural Networks
von: Nakahara, Yuta, et al.
Veröffentlicht: (2025)
von: Nakahara, Yuta, et al.
Veröffentlicht: (2025)
Uncovering Layer-Dependent Activation Sparsity Patterns in ReLU Transformers
von: Wild, Cody, et al.
Veröffentlicht: (2024)
von: Wild, Cody, et al.
Veröffentlicht: (2024)
Three Quantization Regimes for ReLU Networks
von: Ou, Weigutian, et al.
Veröffentlicht: (2024)
von: Ou, Weigutian, et al.
Veröffentlicht: (2024)
Activation-Descent Regularization for Input Optimization of ReLU Networks
von: Yu, Hongzhan, et al.
Veröffentlicht: (2024)
von: Yu, Hongzhan, et al.
Veröffentlicht: (2024)
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
von: Schoots, Nandi, et al.
Veröffentlicht: (2025)
von: Schoots, Nandi, et al.
Veröffentlicht: (2025)
Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During Training
von: Milkert, Max, et al.
Veröffentlicht: (2023)
von: Milkert, Max, et al.
Veröffentlicht: (2023)
Benign Overfitting in Adversarial Training for Vision Transformers
von: Zhang, Jiaming, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2026)
ReLU Networks for Exact Generation of Similar Graphs
von: Ghafoor, Mamoona, et al.
Veröffentlicht: (2026)
von: Ghafoor, Mamoona, et al.
Veröffentlicht: (2026)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
von: Yazdannik, Saman, et al.
Veröffentlicht: (2025)
von: Yazdannik, Saman, et al.
Veröffentlicht: (2025)
Training a Two Layer ReLU Network Analytically
von: Barbu, Adrian
Veröffentlicht: (2023)
von: Barbu, Adrian
Veröffentlicht: (2023)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
von: Dana, Léo, et al.
Veröffentlicht: (2025)
von: Dana, Léo, et al.
Veröffentlicht: (2025)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
von: Harzli, Ouns El, et al.
Veröffentlicht: (2026)
von: Harzli, Ouns El, et al.
Veröffentlicht: (2026)
Is ReLU Adversarially Robust?
von: Sooksatra, Korn, et al.
Veröffentlicht: (2024)
von: Sooksatra, Korn, et al.
Veröffentlicht: (2024)
RePO: Understanding Preference Learning Through ReLU-Based Optimization
von: Wu, Junkang, et al.
Veröffentlicht: (2025)
von: Wu, Junkang, et al.
Veröffentlicht: (2025)
$λ$-GELU: Learning Gating Hardness for Controlled ReLU-ization in Deep Networks
von: Pérez-Corral, Cristian, et al.
Veröffentlicht: (2026)
von: Pérez-Corral, Cristian, et al.
Veröffentlicht: (2026)
Risk Phase Transitions in Spiked Regression: Alignment Driven Benign and Catastrophic Overfitting
von: Li, Jiping, et al.
Veröffentlicht: (2025)
von: Li, Jiping, et al.
Veröffentlicht: (2025)
Precise Verification of Transformers through ReLU-Catalyzed Abstraction Refinement
von: Liu, Hengjie, et al.
Veröffentlicht: (2026)
von: Liu, Hengjie, et al.
Veröffentlicht: (2026)
Detecting Invariant Manifolds in ReLU-Based RNNs
von: Eisenmann, Lukas, et al.
Veröffentlicht: (2025)
von: Eisenmann, Lukas, et al.
Veröffentlicht: (2025)
Topological Signatures of ReLU Neural Network Activation Patterns
von: Bosca, Vicente, et al.
Veröffentlicht: (2025)
von: Bosca, Vicente, et al.
Veröffentlicht: (2025)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models
von: Jha, Nandan Kumar, et al.
Veröffentlicht: (2024)
von: Jha, Nandan Kumar, et al.
Veröffentlicht: (2024)
Deep ReLU Networks Have Surprisingly Simple Polytopes
von: Fan, Feng-Lei, et al.
Veröffentlicht: (2023)
von: Fan, Feng-Lei, et al.
Veröffentlicht: (2023)
The Median is Easier than it Looks: Approximation with a Constant-Depth, Linear-Width ReLU Network
von: Dutta, Abhigyan, et al.
Veröffentlicht: (2026)
von: Dutta, Abhigyan, et al.
Veröffentlicht: (2026)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
von: Bantzis, Ioannis, et al.
Veröffentlicht: (2025)
von: Bantzis, Ioannis, et al.
Veröffentlicht: (2025)
Convex Formulations for Training Two-Layer ReLU Neural Networks
von: Prakhya, Karthik, et al.
Veröffentlicht: (2024)
von: Prakhya, Karthik, et al.
Veröffentlicht: (2024)
Directional Convergence, Benign Overfitting of Gradient Descent in leaky ReLU two-layer Neural Networks
von: Hashimoto, Ichiro
Veröffentlicht: (2025)
von: Hashimoto, Ichiro
Veröffentlicht: (2025)
On the Principles of ReLU Networks with One Hidden Layer
von: Huang, Changcun
Veröffentlicht: (2024)
von: Huang, Changcun
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Classical View on Benign Overfitting: The Role of Sample Size
von: Park, Junhyung, et al.
Veröffentlicht: (2025) -
Initialization Matters: On the Benign Overfitting of Two-Layer ReLU CNN with Fully Trainable Layers
von: Shang, Shuning, et al.
Veröffentlicht: (2024) -
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
von: Qiao, Dan, et al.
Veröffentlicht: (2024) -
A Quantitative Characterization of Forgetting in Post-Training
von: Balasubramanian, Krishnakumar, et al.
Veröffentlicht: (2026) -
From Tempered to Benign Overfitting in ReLU Neural Networks
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)