Optimized Weight Initialization on the Stiefel Manifold for Deep ReLU Neural Networks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lee, Hyungu, Kim, Taehyeong, Choi, Hayoung |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Improved identification of breakpoints in piecewise regression and its applications
par: Kim, Taehyeong, et autres
Publié: (2024)
par: Kim, Taehyeong, et autres
Publié: (2024)
Pathwise Explanation of ReLU Neural Networks
par: Lim, Seongwoo, et autres
Publié: (2025)
par: Lim, Seongwoo, et autres
Publié: (2025)
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
par: Lee, Hyunwoo, et autres
Publié: (2024)
par: Lee, Hyunwoo, et autres
Publié: (2024)
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
par: Lee, Hyunwoo, et autres
Publié: (2025)
par: Lee, Hyunwoo, et autres
Publié: (2025)
Discrete Functional Geometry of ReLU Networks via ReLU Transition Graphs
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
Optimal Initialization in Depth: Lyapunov Initialization and Limit Theorems for Deep Leaky ReLU Networks
par: Kogler, Constantin, et autres
Publié: (2026)
par: Kogler, Constantin, et autres
Publié: (2026)
Depth Degeneracy in Neural Networks: Vanishing Angles in Fully Connected ReLU Networks on Initialization
par: Jakub, Cameron, et autres
Publié: (2023)
par: Jakub, Cameron, et autres
Publié: (2023)
Convexity in ReLU Neural Networks: beyond ICNNs?
par: Gagneux, Anne, et autres
Publié: (2025)
par: Gagneux, Anne, et autres
Publié: (2025)
On Space Folds of ReLU Neural Networks
par: Lewandowski, Michal, et autres
Publié: (2025)
par: Lewandowski, Michal, et autres
Publié: (2025)
Stochastic Bandits with ReLU Neural Networks
par: Xu, Kan, et autres
Publié: (2024)
par: Xu, Kan, et autres
Publié: (2024)
The Geometry of ReLU Networks through the ReLU Transition Graph
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
SurvReLU: Inherently Interpretable Survival Analysis via Deep ReLU Networks
par: Sun, Xiaotong, et autres
Publié: (2024)
par: Sun, Xiaotong, et autres
Publié: (2024)
Topological Expressivity of ReLU Neural Networks
par: Ergen, Ekin, et autres
Publié: (2023)
par: Ergen, Ekin, et autres
Publié: (2023)
Early Neuron Alignment in Two-layer ReLU Networks with Small Initialization
par: Min, Hancheng, et autres
Publié: (2023)
par: Min, Hancheng, et autres
Publié: (2023)
On the Local Complexity of Linear Regions in Deep ReLU Networks
par: Patel, Niket, et autres
Publié: (2024)
par: Patel, Niket, et autres
Publié: (2024)
Implicit Hypersurface Approximation Capacity in Deep ReLU Networks
par: Vallin, Jonatan, et autres
Publié: (2024)
par: Vallin, Jonatan, et autres
Publié: (2024)
Hamiltonian Monte Carlo on ReLU Neural Networks is Inefficient
par: Dinh, Vu C., et autres
Publié: (2024)
par: Dinh, Vu C., et autres
Publié: (2024)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
Detecting Invariant Manifolds in ReLU-Based RNNs
par: Eisenmann, Lukas, et autres
Publié: (2025)
par: Eisenmann, Lukas, et autres
Publié: (2025)
Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study
par: Liu, Hao, et autres
Publié: (2024)
par: Liu, Hao, et autres
Publié: (2024)
Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks
par: Liao, Fangshuo, et autres
Publié: (2023)
par: Liao, Fangshuo, et autres
Publié: (2023)
Geometry-induced Regularization in Deep ReLU Neural Networks
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
Random ReLU Neural Networks as Non-Gaussian Processes
par: Parhi, Rahul, et autres
Publié: (2024)
par: Parhi, Rahul, et autres
Publié: (2024)
The Resurrection of the ReLU
par: Horuz, Coşku Can, et autres
Publié: (2025)
par: Horuz, Coşku Can, et autres
Publié: (2025)
Complexity of Injectivity and Verification of ReLU Neural Networks
par: Froese, Vincent, et autres
Publié: (2024)
par: Froese, Vincent, et autres
Publié: (2024)
Neural Tangent Kernels and Fisher Information Matrices for Simple ReLU Networks with Random Hidden Weights
par: Takeuchi, Jun'ichi, et autres
Publié: (2025)
par: Takeuchi, Jun'ichi, et autres
Publié: (2025)
Deep Network Approximation: Beyond ReLU to Diverse Activation Functions
par: Zhang, Shijun, et autres
Publié: (2023)
par: Zhang, Shijun, et autres
Publié: (2023)
On the Expressiveness of Rational ReLU Neural Networks With Bounded Depth
par: Averkov, Gennadiy, et autres
Publié: (2025)
par: Averkov, Gennadiy, et autres
Publié: (2025)
Dense ReLU Neural Networks for Temporal-spatial Model
par: Padilla, Carlos Misael Madrid, et autres
Publié: (2024)
par: Padilla, Carlos Misael Madrid, et autres
Publié: (2024)
Neural Characteristic Activation Analysis and Geometric Parameterization for ReLU Networks
par: Chen, Wenlin, et autres
Publié: (2023)
par: Chen, Wenlin, et autres
Publié: (2023)
The Effects of Multi-Task Learning on ReLU Neural Network Functions
par: Nakhleh, Julia, et autres
Publié: (2024)
par: Nakhleh, Julia, et autres
Publié: (2024)
Activation-Descent Regularization for Input Optimization of ReLU Networks
par: Yu, Hongzhan, et autres
Publié: (2024)
par: Yu, Hongzhan, et autres
Publié: (2024)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
par: Kim, Sungyoon, et autres
Publié: (2024)
par: Kim, Sungyoon, et autres
Publié: (2024)
Geometry of Singular Foliations and Learning Manifolds in ReLU Networks via the Data Information Matrix
par: Tron, Eliot, et autres
Publié: (2024)
par: Tron, Eliot, et autres
Publié: (2024)
From Tempered to Benign Overfitting in ReLU Neural Networks
par: Kornowski, Guy, et autres
Publié: (2023)
par: Kornowski, Guy, et autres
Publié: (2023)
On the Depth of Monotone ReLU Neural Networks and ICNNs
par: Bakaev, Egor, et autres
Publié: (2025)
par: Bakaev, Egor, et autres
Publié: (2025)
Expressivity and Approximation Properties of Deep Neural Networks with ReLU$^k$ Activation
par: He, Juncai, et autres
Publié: (2023)
par: He, Juncai, et autres
Publié: (2023)
Weighted variation spaces and approximation by shallow ReLU networks
par: DeVore, Ronald, et autres
Publié: (2023)
par: DeVore, Ronald, et autres
Publié: (2023)
Memorization Capacity for Additive Fine-Tuning with Small ReLU Networks
par: Sohn, Jy-yong, et autres
Publié: (2024)
par: Sohn, Jy-yong, et autres
Publié: (2024)
ReLU integral probability metric and its applications
par: Park, Yuha, et autres
Publié: (2025)
par: Park, Yuha, et autres
Publié: (2025)
Documents similaires
-
Improved identification of breakpoints in piecewise regression and its applications
par: Kim, Taehyeong, et autres
Publié: (2024) -
Pathwise Explanation of ReLU Neural Networks
par: Lim, Seongwoo, et autres
Publié: (2025) -
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
par: Lee, Hyunwoo, et autres
Publié: (2024) -
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
par: Lee, Hyunwoo, et autres
Publié: (2025) -
Discrete Functional Geometry of ReLU Networks via ReLU Transition Graphs
par: Dhayalkar, Sahil Rajesh
Publié: (2025)