Activation-Descent Regularization for Input Optimization of ReLU Networks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yu, Hongzhan, Gao, Sicun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
N-ReLU: Zero-Mean Stochastic Extension of ReLU
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
The Geometry of ReLU Networks through the ReLU Transition Graph
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
The Resurrection of the ReLU
par: Horuz, Coşku Can, et autres
Publié: (2025)
par: Horuz, Coşku Can, et autres
Publié: (2025)
Pathwise Explanation of ReLU Neural Networks
par: Lim, Seongwoo, et autres
Publié: (2025)
par: Lim, Seongwoo, et autres
Publié: (2025)
RePO: Understanding Preference Learning Through ReLU-Based Optimization
par: Wu, Junkang, et autres
Publié: (2025)
par: Wu, Junkang, et autres
Publié: (2025)
Uncovering Layer-Dependent Activation Sparsity Patterns in ReLU Transformers
par: Wild, Cody, et autres
Publié: (2024)
par: Wild, Cody, et autres
Publié: (2024)
Topological Signatures of ReLU Neural Network Activation Patterns
par: Bosca, Vicente, et autres
Publié: (2025)
par: Bosca, Vicente, et autres
Publié: (2025)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
par: Zhang, Zhengyan, et autres
Publié: (2024)
par: Zhang, Zhengyan, et autres
Publié: (2024)
Three Quantization Regimes for ReLU Networks
par: Ou, Weigutian, et autres
Publié: (2024)
par: Ou, Weigutian, et autres
Publié: (2024)
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
par: Schoots, Nandi, et autres
Publié: (2025)
par: Schoots, Nandi, et autres
Publié: (2025)
Geometry-induced Regularization in Deep ReLU Neural Networks
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
par: Bona-Pellissier, Joachim, et autres
Publié: (2024)
ReLU Networks for Exact Generation of Similar Graphs
par: Ghafoor, Mamoona, et autres
Publié: (2026)
par: Ghafoor, Mamoona, et autres
Publié: (2026)
Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
par: Park, Junhyung, et autres
Publié: (2024)
par: Park, Junhyung, et autres
Publié: (2024)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
par: Yazdannik, Saman, et autres
Publié: (2025)
par: Yazdannik, Saman, et autres
Publié: (2025)
Algebraic Approach to Ridge-Regularized Mean Squared Error Minimization in Minimal ReLU Neural Network
par: Fukasaku, Ryoya, et autres
Publié: (2025)
par: Fukasaku, Ryoya, et autres
Publié: (2025)
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
par: Qiao, Dan, et autres
Publié: (2025)
par: Qiao, Dan, et autres
Publié: (2025)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
par: Dana, Léo, et autres
Publié: (2025)
par: Dana, Léo, et autres
Publié: (2025)
Expressive Power of ReLU and Step Networks under Floating-Point Operations
par: Park, Yeachan, et autres
Publié: (2024)
par: Park, Yeachan, et autres
Publié: (2024)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
par: Harzli, Ouns El, et autres
Publié: (2026)
par: Harzli, Ouns El, et autres
Publié: (2026)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
par: Ye, Longqing
Publié: (2025)
par: Ye, Longqing
Publié: (2025)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
par: Qiao, Dan, et autres
Publié: (2024)
par: Qiao, Dan, et autres
Publié: (2024)
Is ReLU Adversarially Robust?
par: Sooksatra, Korn, et autres
Publié: (2024)
par: Sooksatra, Korn, et autres
Publié: (2024)
$λ$-GELU: Learning Gating Hardness for Controlled ReLU-ization in Deep Networks
par: Pérez-Corral, Cristian, et autres
Publié: (2026)
par: Pérez-Corral, Cristian, et autres
Publié: (2026)
Precise Verification of Transformers through ReLU-Catalyzed Abstraction Refinement
par: Liu, Hengjie, et autres
Publié: (2026)
par: Liu, Hengjie, et autres
Publié: (2026)
Detecting Invariant Manifolds in ReLU-Based RNNs
par: Eisenmann, Lukas, et autres
Publié: (2025)
par: Eisenmann, Lukas, et autres
Publié: (2025)
A Lower Bound for the Number of Linear Regions of Ternary ReLU Regression Neural Networks
par: Nakahara, Yuta, et autres
Publié: (2025)
par: Nakahara, Yuta, et autres
Publié: (2025)
Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During Training
par: Milkert, Max, et autres
Publié: (2023)
par: Milkert, Max, et autres
Publié: (2023)
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models
par: Jha, Nandan Kumar, et autres
Publié: (2024)
par: Jha, Nandan Kumar, et autres
Publié: (2024)
Deep ReLU Networks Have Surprisingly Simple Polytopes
par: Fan, Feng-Lei, et autres
Publié: (2023)
par: Fan, Feng-Lei, et autres
Publié: (2023)
The Median is Easier than it Looks: Approximation with a Constant-Depth, Linear-Width ReLU Network
par: Dutta, Abhigyan, et autres
Publié: (2026)
par: Dutta, Abhigyan, et autres
Publié: (2026)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
par: Bantzis, Ioannis, et autres
Publié: (2025)
par: Bantzis, Ioannis, et autres
Publié: (2025)
The Cost of Robustness: Tighter Bounds on Parameter Complexity for Robust Memorization in ReLU Nets
par: Kim, Yujun, et autres
Publié: (2025)
par: Kim, Yujun, et autres
Publié: (2025)
Learning Quadruped Walking from Seconds of Demonstration
par: Zhang, Ruipeng, et autres
Publié: (2026)
par: Zhang, Ruipeng, et autres
Publié: (2026)
Uncertainty Quantification with Bayesian Higher Order ReLU KANs
par: Giroux, James, et autres
Publié: (2024)
par: Giroux, James, et autres
Publié: (2024)
Fractal Landscapes in Policy Optimization
par: Wang, Tao, et autres
Publié: (2023)
par: Wang, Tao, et autres
Publié: (2023)
Extremum-Seeking Action Selection for Accelerating Policy Optimization
par: Chang, Ya-Chien, et autres
Publié: (2024)
par: Chang, Ya-Chien, et autres
Publié: (2024)
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers
par: Liang, Yingyu, et autres
Publié: (2024)
par: Liang, Yingyu, et autres
Publié: (2024)
A Significantly Better Class of Activation Functions Than ReLU Like Activation Functions
par: Noel, Mathew Mithra, et autres
Publié: (2024)
par: Noel, Mathew Mithra, et autres
Publié: (2024)
Agnostic Learning of General ReLU Activation Using Gradient Descent
par: Awasthi, Pranjal, et autres
Publié: (2022)
par: Awasthi, Pranjal, et autres
Publié: (2022)
When Maximum Entropy Misleads Policy Optimization
par: Zhang, Ruipeng, et autres
Publié: (2025)
par: Zhang, Ruipeng, et autres
Publié: (2025)
Documents similaires
-
N-ReLU: Zero-Mean Stochastic Extension of ReLU
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025) -
The Geometry of ReLU Networks through the ReLU Transition Graph
par: Dhayalkar, Sahil Rajesh
Publié: (2025) -
The Resurrection of the ReLU
par: Horuz, Coşku Can, et autres
Publié: (2025) -
Pathwise Explanation of ReLU Neural Networks
par: Lim, Seongwoo, et autres
Publié: (2025) -
RePO: Understanding Preference Learning Through ReLU-Based Optimization
par: Wu, Junkang, et autres
Publié: (2025)