Path-conditioned training: a principled way to rescale ReLU neural networks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lebeurrier, Arthur, Vayer, Titouan, Gribonval, Rémi |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
An analysis of optimization problems involving ReLU neural networks
par: Plate, Christoph, et autres
Publié: (2025)
par: Plate, Christoph, et autres
Publié: (2025)
MIQCQP reformulation of the ReLU neural networks Lipschitz constant estimation problem
par: Sbihi, Mohammed, et autres
Publié: (2024)
par: Sbihi, Mohammed, et autres
Publié: (2024)
Hidden Minima in Two-Layer ReLU Networks
par: Arjevani, Yossi
Publié: (2023)
par: Arjevani, Yossi
Publié: (2023)
Implicit Differentiation for Hyperparameter Tuning the Weighted Graphical Lasso
par: Pouliquen, Can, et autres
Publié: (2023)
par: Pouliquen, Can, et autres
Publié: (2023)
On bounds for norms of reparameterized ReLU artificial neural network parameters: sums of fractional powers of the Lipschitz norm control the network parameter vector
par: Jentzen, Arnulf, et autres
Publié: (2022)
par: Jentzen, Arnulf, et autres
Publié: (2022)
Why Smooth Stability Assumptions Fail for ReLU Learning
par: Katende, Ronald
Publié: (2025)
par: Katende, Ronald
Publié: (2025)
Convex Formulations for Training Two-Layer ReLU Neural Networks
par: Prakhya, Karthik, et autres
Publié: (2024)
par: Prakhya, Karthik, et autres
Publié: (2024)
An Efficient Alternating Algorithm for ReLU-based Symmetric Matrix Decomposition
par: Wang, Qingsong
Publié: (2025)
par: Wang, Qingsong
Publié: (2025)
Computational Tradeoffs of Optimization-Based Bound Tightening in ReLU Networks
par: Badilla, Fabian, et autres
Publié: (2023)
par: Badilla, Fabian, et autres
Publié: (2023)
A Complete Set of Quadratic Constraints for Repeated ReLU and Generalizations
par: Noori, Sahel Vahedi, et autres
Publié: (2024)
par: Noori, Sahel Vahedi, et autres
Publié: (2024)
Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks
par: Liao, Fangshuo, et autres
Publié: (2023)
par: Liao, Fangshuo, et autres
Publié: (2023)
Non-asymptotic convergence analysis of the stochastic gradient Hamiltonian Monte Carlo algorithm with discontinuous stochastic gradient with applications to training of ReLU neural networks
par: Liang, Luxu, et autres
Publié: (2024)
par: Liang, Luxu, et autres
Publié: (2024)
ReLU Networks for Model Predictive Control: Network Complexity and Performance Guarantees
par: Li, Xingchen, et autres
Publié: (2026)
par: Li, Xingchen, et autres
Publié: (2026)
Stability and Performance Analysis of Discrete-Time ReLU Recurrent Neural Networks
par: Noori, Sahel Vahedi, et autres
Publié: (2024)
par: Noori, Sahel Vahedi, et autres
Publié: (2024)
Pruning for efficient deterministic global optimization over trained ReLU neural networks
par: Lastrucci, Giacomo, et autres
Publié: (2026)
par: Lastrucci, Giacomo, et autres
Publié: (2026)
A multilevel approach to accelerate the training of Transformers
par: Lauga, Guillaume, et autres
Publié: (2025)
par: Lauga, Guillaume, et autres
Publié: (2025)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
par: Kim, Sungyoon, et autres
Publié: (2024)
par: Kim, Sungyoon, et autres
Publié: (2024)
Abide by the Law and Follow the Flow: Conservation Laws for Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2023)
par: Marcotte, Sibylle, et autres
Publié: (2023)
Keep the Momentum: Conservation Laws beyond Euclidean Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2024)
par: Marcotte, Sibylle, et autres
Publié: (2024)
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
par: Lamperski, Andrew, et autres
Publié: (2024)
par: Lamperski, Andrew, et autres
Publié: (2024)
Local Lipschitz Constant Computation of ReLU-FNNs: Upper Bound Computation with Exactness Verification
par: Ebihara, Yoshio, et autres
Publié: (2023)
par: Ebihara, Yoshio, et autres
Publié: (2023)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
par: Kuelbs, Daniel, et autres
Publié: (2024)
par: Kuelbs, Daniel, et autres
Publié: (2024)
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
par: Yang, Yahong, et autres
Publié: (2023)
par: Yang, Yahong, et autres
Publié: (2023)
Neural Collapse under Gradient Flow on Shallow ReLU Networks for Orthogonally Separable Data
par: Min, Hancheng, et autres
Publié: (2025)
par: Min, Hancheng, et autres
Publié: (2025)
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
par: Lamperski, Andrew, et autres
Publié: (2024)
par: Lamperski, Andrew, et autres
Publié: (2024)
On the existence of minimizers in shallow residual ReLU neural network optimization landscapes
par: Dereich, Steffen, et autres
Publié: (2023)
par: Dereich, Steffen, et autres
Publié: (2023)
How Does the ReLU Activation Affect the Implicit Bias of Gradient Descent on High-dimensional Neural Network Regression?
par: Lai, Kuo-Wei, et autres
Publié: (2026)
par: Lai, Kuo-Wei, et autres
Publié: (2026)
Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU
par: Liu, Jingzhou
Publié: (2025)
par: Liu, Jingzhou
Publié: (2025)
Interpretable global minima of deep ReLU neural networks on sequentially separable data
par: Chen, Thomas, et autres
Publié: (2024)
par: Chen, Thomas, et autres
Publié: (2024)
Convexity in ReLU Neural Networks: beyond ICNNs?
par: Gagneux, Anne, et autres
Publié: (2025)
par: Gagneux, Anne, et autres
Publié: (2025)
A Rescaling-Invariant Lipschitz Bound Based on Path-Metrics for Modern ReLU Network Parameterizations
par: Gonon, Antoine, et autres
Publié: (2024)
par: Gonon, Antoine, et autres
Publié: (2024)
On the existence of optimal shallow feedforward networks with ReLU activation
par: Dereich, Steffen, et autres
Publié: (2023)
par: Dereich, Steffen, et autres
Publié: (2023)
Gradient descent provably escapes saddle points in the training of shallow ReLU networks
par: Cheridito, Patrick, et autres
Publié: (2022)
par: Cheridito, Patrick, et autres
Publié: (2022)
Architecture independent generalization bounds for overparametrized deep ReLU networks
par: Bapu, Anandatheertha, et autres
Publié: (2025)
par: Bapu, Anandatheertha, et autres
Publié: (2025)
Geometry-induced Regularization in Deep ReLU Neural Networks
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
Tuning the burn-in phase in training recurrent neural networks improves their performance
par: Schiller, Julian D., et autres
Publié: (2026)
par: Schiller, Julian D., et autres
Publié: (2026)
Tightening convex relaxations of trained neural networks: a unified approach for convex and S-shaped activations
par: Carrasco, Pablo, et autres
Publié: (2024)
par: Carrasco, Pablo, et autres
Publié: (2024)
ReLU Surrogates in Mixed-Integer MPC for Irrigation Scheduling
par: Agyeman, Bernard T., et autres
Publié: (2024)
par: Agyeman, Bernard T., et autres
Publié: (2024)
Convergence of stochastic gradient descent under a local Lojasiewicz condition for deep neural networks
par: An, Jing, et autres
Publié: (2023)
par: An, Jing, et autres
Publié: (2023)
On the ReLU Lagrangian Cuts for Stochastic Mixed Integer Programming
par: Deng, Haoyun, et autres
Publié: (2024)
par: Deng, Haoyun, et autres
Publié: (2024)
Documents similaires
-
An analysis of optimization problems involving ReLU neural networks
par: Plate, Christoph, et autres
Publié: (2025) -
MIQCQP reformulation of the ReLU neural networks Lipschitz constant estimation problem
par: Sbihi, Mohammed, et autres
Publié: (2024) -
Hidden Minima in Two-Layer ReLU Networks
par: Arjevani, Yossi
Publié: (2023) -
Implicit Differentiation for Hyperparameter Tuning the Weighted Graphical Lasso
par: Pouliquen, Can, et autres
Publié: (2023) -
On bounds for norms of reparameterized ReLU artificial neural network parameters: sums of fractional powers of the Lipschitz norm control the network parameter vector
par: Jentzen, Arnulf, et autres
Publié: (2022)