Why ReLU? A Bit-Model Dichotomy for Deep Network Training
Fuente:
arXiv
Saved in:
| Main Authors: | Doron-Arad, Ilan, Mossel, Elchanan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Realizable Regression and Applications for ReLU Networks
by: Doron-Arad, Ilan, et al.
Published: (2026)
by: Doron-Arad, Ilan, et al.
Published: (2026)
A Theory of Online Learning with Autoregressive Chain-of-Thought Reasoning
by: Doron-Arad, Ilan, et al.
Published: (2026)
by: Doron-Arad, Ilan, et al.
Published: (2026)
On the Hardness of Training Deep Neural Networks Discretely
by: Doron-Arad, Ilan
Published: (2024)
by: Doron-Arad, Ilan
Published: (2024)
Discrete Functional Geometry of ReLU Networks via ReLU Transition Graphs
by: Dhayalkar, Sahil Rajesh
Published: (2025)
by: Dhayalkar, Sahil Rajesh
Published: (2025)
The Geometry of ReLU Networks through the ReLU Transition Graph
by: Dhayalkar, Sahil Rajesh
Published: (2025)
by: Dhayalkar, Sahil Rajesh
Published: (2025)
A Mathematical Model for Curriculum Learning for Parities
by: Cornacchia, Elisabetta, et al.
Published: (2023)
by: Cornacchia, Elisabetta, et al.
Published: (2023)
SurvReLU: Inherently Interpretable Survival Analysis via Deep ReLU Networks
by: Sun, Xiaotong, et al.
Published: (2024)
by: Sun, Xiaotong, et al.
Published: (2024)
Training a Two Layer ReLU Network Analytically
by: Barbu, Adrian
Published: (2023)
by: Barbu, Adrian
Published: (2023)
On the Local Complexity of Linear Regions in Deep ReLU Networks
by: Patel, Niket, et al.
Published: (2024)
by: Patel, Niket, et al.
Published: (2024)
Implicit Hypersurface Approximation Capacity in Deep ReLU Networks
by: Vallin, Jonatan, et al.
Published: (2024)
by: Vallin, Jonatan, et al.
Published: (2024)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
The Resurrection of the ReLU
by: Horuz, Coşku Can, et al.
Published: (2025)
by: Horuz, Coşku Can, et al.
Published: (2025)
Deep Network Approximation: Beyond ReLU to Diverse Activation Functions
by: Zhang, Shijun, et al.
Published: (2023)
by: Zhang, Shijun, et al.
Published: (2023)
Why Smooth Stability Assumptions Fail for ReLU Learning
by: Katende, Ronald
Published: (2025)
by: Katende, Ronald
Published: (2025)
Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
by: Park, Junhyung, et al.
Published: (2024)
by: Park, Junhyung, et al.
Published: (2024)
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
Pathwise Explanation of ReLU Neural Networks
by: Lim, Seongwoo, et al.
Published: (2025)
by: Lim, Seongwoo, et al.
Published: (2025)
Online Learning of Neural Networks
by: Daniely, Amit, et al.
Published: (2025)
by: Daniely, Amit, et al.
Published: (2025)
Optimal Sets and Solution Paths of ReLU Networks
by: Mishkin, Aaron, et al.
Published: (2023)
by: Mishkin, Aaron, et al.
Published: (2023)
On Size-Independent Sample Complexity of ReLU Networks
by: Sellke, Mark
Published: (2023)
by: Sellke, Mark
Published: (2023)
Convexity in ReLU Neural Networks: beyond ICNNs?
by: Gagneux, Anne, et al.
Published: (2025)
by: Gagneux, Anne, et al.
Published: (2025)
Optimized Weight Initialization on the Stiefel Manifold for Deep ReLU Neural Networks
by: Lee, Hyungu, et al.
Published: (2025)
by: Lee, Hyungu, et al.
Published: (2025)
Sobolev Approximation of Deep ReLU Networks in Log-Barron Space
by: Song, Changhoon, et al.
Published: (2026)
by: Song, Changhoon, et al.
Published: (2026)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
by: Yazdannik, Saman, et al.
Published: (2025)
by: Yazdannik, Saman, et al.
Published: (2025)
Some Theoretical Limitations of t-SNE
by: Li, Rupert, et al.
Published: (2026)
by: Li, Rupert, et al.
Published: (2026)
Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
Component-based Sketching for Deep ReLU Nets
by: Wang, Di, et al.
Published: (2024)
by: Wang, Di, et al.
Published: (2024)
Stochastic Bandits with ReLU Neural Networks
by: Xu, Kan, et al.
Published: (2024)
by: Xu, Kan, et al.
Published: (2024)
On Space Folds of ReLU Neural Networks
by: Lewandowski, Michal, et al.
Published: (2025)
by: Lewandowski, Michal, et al.
Published: (2025)
Dense ReLU Neural Networks for Temporal-spatial Model
by: Padilla, Carlos Misael Madrid, et al.
Published: (2024)
by: Padilla, Carlos Misael Madrid, et al.
Published: (2024)
ReLU-KAN: New Kolmogorov-Arnold Networks that Only Need Matrix Addition, Dot Multiplication, and ReLU
by: Qiu, Qi, et al.
Published: (2024)
by: Qiu, Qi, et al.
Published: (2024)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
by: Ye, Longqing
Published: (2025)
by: Ye, Longqing
Published: (2025)
Topological Expressivity of ReLU Neural Networks
by: Ergen, Ekin, et al.
Published: (2023)
by: Ergen, Ekin, et al.
Published: (2023)
Three Quantization Regimes for ReLU Networks
by: Ou, Weigutian, et al.
Published: (2024)
by: Ou, Weigutian, et al.
Published: (2024)
Brownian ReLU(Br-ReLU): A New Activation Function for a Long-Short Term Memory (LSTM) Network
by: Awiakye-Marfo, George, et al.
Published: (2026)
by: Awiakye-Marfo, George, et al.
Published: (2026)
Complexity of Linear Regions in Self-supervised Deep ReLU Networks
by: Muthivhi, Mufhumudzi, et al.
Published: (2026)
by: Muthivhi, Mufhumudzi, et al.
Published: (2026)
Deep ReLU Networks Have Surprisingly Simple Polytopes
by: Fan, Feng-Lei, et al.
Published: (2023)
by: Fan, Feng-Lei, et al.
Published: (2023)
Convergence of Adam in Deep ReLU Networks via Directional Complexity and Kakeya Bounds
by: Sridhar, Anupama, et al.
Published: (2025)
by: Sridhar, Anupama, et al.
Published: (2025)
Hidden Minima in Two-Layer ReLU Networks
by: Arjevani, Yossi
Published: (2023)
by: Arjevani, Yossi
Published: (2023)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
by: Harzli, Ouns El, et al.
Published: (2026)
by: Harzli, Ouns El, et al.
Published: (2026)
Similar Items
-
Online Realizable Regression and Applications for ReLU Networks
by: Doron-Arad, Ilan, et al.
Published: (2026) -
A Theory of Online Learning with Autoregressive Chain-of-Thought Reasoning
by: Doron-Arad, Ilan, et al.
Published: (2026) -
On the Hardness of Training Deep Neural Networks Discretely
by: Doron-Arad, Ilan
Published: (2024) -
Discrete Functional Geometry of ReLU Networks via ReLU Transition Graphs
by: Dhayalkar, Sahil Rajesh
Published: (2025) -
The Geometry of ReLU Networks through the ReLU Transition Graph
by: Dhayalkar, Sahil Rajesh
Published: (2025)