Expressive Power of ReLU and Step Networks under Floating-Point Operations
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Park, Yeachan, Hwang, Geonho, Lee, Wonyeol, Park, Sejun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Expressive Power of Floating-Point Neural Networks with Arbitrary Reduction Orders and Inexact Activation Implementations
par: Park, Yeachan, et autres
Publié: (2026)
par: Park, Yeachan, et autres
Publié: (2026)
On the Expressive Power of Floating-Point Transformers
par: Park, Sejun, et autres
Publié: (2026)
par: Park, Sejun, et autres
Publié: (2026)
On Expressive Power of Quantized Neural Networks under Fixed-Point Arithmetic
par: Park, Yeachan, et autres
Publié: (2024)
par: Park, Yeachan, et autres
Publié: (2024)
Floating-Point Neural Networks Are Provably Robust Universal Approximators
par: Hwang, Geonho, et autres
Publié: (2025)
par: Hwang, Geonho, et autres
Publié: (2025)
Floating-Point Networks with Automatic Differentiation Can Represent Almost All Floating-Point Functions and Their Gradients
par: Park, Sejun, et autres
Publié: (2026)
par: Park, Sejun, et autres
Publié: (2026)
Benign Overfitting for Regression with Trained Two-Layer ReLU Networks
par: Park, Junhyung, et autres
Publié: (2024)
par: Park, Junhyung, et autres
Publié: (2024)
Pathwise Explanation of ReLU Neural Networks
par: Lim, Seongwoo, et autres
Publié: (2025)
par: Lim, Seongwoo, et autres
Publié: (2025)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
par: Manik, Md Motaleb Hossen, et autres
Publié: (2025)
Intrinsic Task Symmetry Drives Generalization in Algorithmic Tasks
par: Hwang, Hyeonbin, et autres
Publié: (2026)
par: Hwang, Hyeonbin, et autres
Publié: (2026)
The Geometry of ReLU Networks through the ReLU Transition Graph
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
par: Dhayalkar, Sahil Rajesh
Publié: (2025)
The Resurrection of the ReLU
par: Horuz, Coşku Can, et autres
Publié: (2025)
par: Horuz, Coşku Can, et autres
Publié: (2025)
Minimum width for universal approximation using ReLU networks on compact domain
par: Kim, Namjun, et autres
Publié: (2023)
par: Kim, Namjun, et autres
Publié: (2023)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
par: Qiao, Dan, et autres
Publié: (2024)
par: Qiao, Dan, et autres
Publié: (2024)
Three Quantization Regimes for ReLU Networks
par: Ou, Weigutian, et autres
Publié: (2024)
par: Ou, Weigutian, et autres
Publié: (2024)
Activation-Descent Regularization for Input Optimization of ReLU Networks
par: Yu, Hongzhan, et autres
Publié: (2024)
par: Yu, Hongzhan, et autres
Publié: (2024)
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
par: Schoots, Nandi, et autres
Publié: (2025)
par: Schoots, Nandi, et autres
Publié: (2025)
Acceleration of Grokking in Learning Arithmetic Operations via Kolmogorov-Arnold Representation
par: Park, Yeachan, et autres
Publié: (2024)
par: Park, Yeachan, et autres
Publié: (2024)
ReLU Networks for Exact Generation of Similar Graphs
par: Ghafoor, Mamoona, et autres
Publié: (2026)
par: Ghafoor, Mamoona, et autres
Publié: (2026)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
par: Yazdannik, Saman, et autres
Publié: (2025)
par: Yazdannik, Saman, et autres
Publié: (2025)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
par: Dana, Léo, et autres
Publié: (2025)
par: Dana, Léo, et autres
Publié: (2025)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
par: Harzli, Ouns El, et autres
Publié: (2026)
par: Harzli, Ouns El, et autres
Publié: (2026)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
par: Ye, Longqing
Publié: (2025)
par: Ye, Longqing
Publié: (2025)
Is ReLU Adversarially Robust?
par: Sooksatra, Korn, et autres
Publié: (2024)
par: Sooksatra, Korn, et autres
Publié: (2024)
RePO: Understanding Preference Learning Through ReLU-Based Optimization
par: Wu, Junkang, et autres
Publié: (2025)
par: Wu, Junkang, et autres
Publié: (2025)
$λ$-GELU: Learning Gating Hardness for Controlled ReLU-ization in Deep Networks
par: Pérez-Corral, Cristian, et autres
Publié: (2026)
par: Pérez-Corral, Cristian, et autres
Publié: (2026)
IMPaCT GNN: Imposing invariance with Message Passing in Chronological split Temporal Graphs
par: Park, Sejun, et autres
Publié: (2024)
par: Park, Sejun, et autres
Publié: (2024)
Uncovering Layer-Dependent Activation Sparsity Patterns in ReLU Transformers
par: Wild, Cody, et autres
Publié: (2024)
par: Wild, Cody, et autres
Publié: (2024)
Precise Verification of Transformers through ReLU-Catalyzed Abstraction Refinement
par: Liu, Hengjie, et autres
Publié: (2026)
par: Liu, Hengjie, et autres
Publié: (2026)
Detecting Invariant Manifolds in ReLU-Based RNNs
par: Eisenmann, Lukas, et autres
Publié: (2025)
par: Eisenmann, Lukas, et autres
Publié: (2025)
Topological Signatures of ReLU Neural Network Activation Patterns
par: Bosca, Vicente, et autres
Publié: (2025)
par: Bosca, Vicente, et autres
Publié: (2025)
A Lower Bound for the Number of Linear Regions of Ternary ReLU Regression Neural Networks
par: Nakahara, Yuta, et autres
Publié: (2025)
par: Nakahara, Yuta, et autres
Publié: (2025)
Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During Training
par: Milkert, Max, et autres
Publié: (2023)
par: Milkert, Max, et autres
Publié: (2023)
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
par: Qiao, Dan, et autres
Publié: (2025)
par: Qiao, Dan, et autres
Publié: (2025)
Bridging the Gap Between Molecule and Textual Descriptions via Substructure-aware Alignment
par: Park, Hyuntae, et autres
Publié: (2025)
par: Park, Hyuntae, et autres
Publié: (2025)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
par: Zhang, Zhengyan, et autres
Publié: (2024)
par: Zhang, Zhengyan, et autres
Publié: (2024)
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models
par: Jha, Nandan Kumar, et autres
Publié: (2024)
par: Jha, Nandan Kumar, et autres
Publié: (2024)
Deep ReLU Networks Have Surprisingly Simple Polytopes
par: Fan, Feng-Lei, et autres
Publié: (2023)
par: Fan, Feng-Lei, et autres
Publié: (2023)
The Median is Easier than it Looks: Approximation with a Constant-Depth, Linear-Width ReLU Network
par: Dutta, Abhigyan, et autres
Publié: (2026)
par: Dutta, Abhigyan, et autres
Publié: (2026)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
par: Bantzis, Ioannis, et autres
Publié: (2025)
par: Bantzis, Ioannis, et autres
Publié: (2025)
Topological Expressivity of ReLU Neural Networks
par: Ergen, Ekin, et autres
Publié: (2023)
par: Ergen, Ekin, et autres
Publié: (2023)
Documents similaires
-
Expressive Power of Floating-Point Neural Networks with Arbitrary Reduction Orders and Inexact Activation Implementations
par: Park, Yeachan, et autres
Publié: (2026) -
On the Expressive Power of Floating-Point Transformers
par: Park, Sejun, et autres
Publié: (2026) -
On Expressive Power of Quantized Neural Networks under Fixed-Point Arithmetic
par: Park, Yeachan, et autres
Publié: (2024) -
Floating-Point Neural Networks Are Provably Robust Universal Approximators
par: Hwang, Geonho, et autres
Publié: (2025) -
Floating-Point Networks with Automatic Differentiation Can Represent Almost All Floating-Point Functions and Their Gradients
par: Park, Sejun, et autres
Publié: (2026)