TeLU Activation Function for Fast and Stable Deep Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Fernandez, Alfredo, Mali, Ankur |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stable and Robust Deep Learning By Hyperbolic Tangent Exponential Linear Unit (TeLU)
by: Fernandez, Alfredo, et al.
Published: (2024)
by: Fernandez, Alfredo, et al.
Published: (2024)
Integral Signatures of Activation Functions: A 9-Dimensional Taxonomy and Stability Theory for Deep Learning
by: Mali, Ankur, et al.
Published: (2025)
by: Mali, Ankur, et al.
Published: (2025)
Deep Network Approximation: Beyond ReLU to Diverse Activation Functions
by: Zhang, Shijun, et al.
Published: (2023)
by: Zhang, Shijun, et al.
Published: (2023)
Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning
by: Prada, Benjamin, et al.
Published: (2025)
by: Prada, Benjamin, et al.
Published: (2025)
PowLU: An Activation Function for Stable Pre-Training of LLMs
by: Jiang, Peijie, et al.
Published: (2026)
by: Jiang, Peijie, et al.
Published: (2026)
SwishReLU: A Unified Approach to Activation Functions for Enhanced Deep Neural Networks Performance
by: Rahman, Jamshaid Ul, et al.
Published: (2024)
by: Rahman, Jamshaid Ul, et al.
Published: (2024)
ReCA: A Parametric ReLU Composite Activation Function
by: Chidiac, John, et al.
Published: (2025)
by: Chidiac, John, et al.
Published: (2025)
Realizable Circuit Complexity: Embedding Computation in Space-Time
by: Prada, Benjamin, et al.
Published: (2025)
by: Prada, Benjamin, et al.
Published: (2025)
A Unified Framework for Continual Learning and Unlearning
by: Chatterjee, Romit, et al.
Published: (2024)
by: Chatterjee, Romit, et al.
Published: (2024)
Neuro-mimetic Task-free Unsupervised Online Learning with Continual Self-Organizing Maps
by: Vaidya, Hitesh, et al.
Published: (2024)
by: Vaidya, Hitesh, et al.
Published: (2024)
Stability Analysis of Various Symbolic Rule Extraction Methods from Recurrent Neural Network
by: Dave, Neisarg, et al.
Published: (2024)
by: Dave, Neisarg, et al.
Published: (2024)
Surprisal-Rényi Free Energy
by: Matsumoto, Shion, et al.
Published: (2026)
by: Matsumoto, Shion, et al.
Published: (2026)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
by: Zhang, Zhengyan, et al.
Published: (2024)
by: Zhang, Zhengyan, et al.
Published: (2024)
A Review of Neuroscience-Inspired Machine Learning
by: Ororbia, Alexander, et al.
Published: (2024)
by: Ororbia, Alexander, et al.
Published: (2024)
Brownian ReLU(Br-ReLU): A New Activation Function for a Long-Short Term Memory (LSTM) Network
by: Awiakye-Marfo, George, et al.
Published: (2026)
by: Awiakye-Marfo, George, et al.
Published: (2026)
Tight Stability, Convergence, and Robustness Bounds for Predictive Coding Networks
by: Mali, Ankur, et al.
Published: (2024)
by: Mali, Ankur, et al.
Published: (2024)
A Composite Activation Function for Learning Stable Binary Representations
by: Park, Seokhun, et al.
Published: (2026)
by: Park, Seokhun, et al.
Published: (2026)
Discrete Functional Geometry of ReLU Networks via ReLU Transition Graphs
by: Dhayalkar, Sahil Rajesh
Published: (2025)
by: Dhayalkar, Sahil Rajesh
Published: (2025)
Investigating Symbolic Capabilities of Large Language Models
by: Dave, Neisarg, et al.
Published: (2024)
by: Dave, Neisarg, et al.
Published: (2024)
The Effects of Multi-Task Learning on ReLU Neural Network Functions
by: Nakhleh, Julia, et al.
Published: (2024)
by: Nakhleh, Julia, et al.
Published: (2024)
SurvReLU: Inherently Interpretable Survival Analysis via Deep ReLU Networks
by: Sun, Xiaotong, et al.
Published: (2024)
by: Sun, Xiaotong, et al.
Published: (2024)
Agnostic Learning of Arbitrary ReLU Activation under Gaussian Marginals
by: Guo, Anxin, et al.
Published: (2024)
by: Guo, Anxin, et al.
Published: (2024)
Agnostic Learning of General ReLU Activation Using Gradient Descent
by: Awasthi, Pranjal, et al.
Published: (2022)
by: Awasthi, Pranjal, et al.
Published: (2022)
Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization
by: Amaefuna, Theophilus, et al.
Published: (2026)
by: Amaefuna, Theophilus, et al.
Published: (2026)
$λ$-GELU: Learning Gating Hardness for Controlled ReLU-ization in Deep Networks
by: Pérez-Corral, Cristian, et al.
Published: (2026)
by: Pérez-Corral, Cristian, et al.
Published: (2026)
Fast and Private Inference of Deep Neural Networks by Co-designing Activation Functions
by: Diaa, Abdulrahman, et al.
Published: (2023)
by: Diaa, Abdulrahman, et al.
Published: (2023)
Two-Dimensional Deep ReLU CNN Approximation for Korobov Functions: A Constructive Approach
by: Fang, Qin, et al.
Published: (2025)
by: Fang, Qin, et al.
Published: (2025)
Expressivity and Approximation Properties of Deep Neural Networks with ReLU$^k$ Activation
by: He, Juncai, et al.
Published: (2023)
by: He, Juncai, et al.
Published: (2023)
Neural Characteristic Activation Analysis and Geometric Parameterization for ReLU Networks
by: Chen, Wenlin, et al.
Published: (2023)
by: Chen, Wenlin, et al.
Published: (2023)
Zorro: A Flexible and Differentiable Parametric Family of Activation Functions That Extends ReLU and GELU
by: Roodschild, Matias, et al.
Published: (2024)
by: Roodschild, Matias, et al.
Published: (2024)
Unlearning or Concealment? A Critical Analysis and Evaluation Metrics for Unlearning in Diffusion Models
by: Sharma, Aakash Sen, et al.
Published: (2024)
by: Sharma, Aakash Sen, et al.
Published: (2024)
A Significantly Better Class of Activation Functions Than ReLU Like Activation Functions
by: Noel, Mathew Mithra, et al.
Published: (2024)
by: Noel, Mathew Mithra, et al.
Published: (2024)
Activation-Descent Regularization for Input Optimization of ReLU Networks
by: Yu, Hongzhan, et al.
Published: (2024)
by: Yu, Hongzhan, et al.
Published: (2024)
Linearization of ReLU Activation Function for Neural Network-Embedded Optimization: Optimal Day-Ahead Energy Scheduling
by: Zhao, Cunzhi, et al.
Published: (2023)
by: Zhao, Cunzhi, et al.
Published: (2023)
Combinations of Fast Activation and Trigonometric Functions in Kolmogorov-Arnold Networks
by: Ta, Hoang-Thang, et al.
Published: (2025)
by: Ta, Hoang-Thang, et al.
Published: (2025)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
by: Golowich, Noah, et al.
Published: (2024)
by: Golowich, Noah, et al.
Published: (2024)
On the Local Complexity of Linear Regions in Deep ReLU Networks
by: Patel, Niket, et al.
Published: (2024)
by: Patel, Niket, et al.
Published: (2024)
Implicit Hypersurface Approximation Capacity in Deep ReLU Networks
by: Vallin, Jonatan, et al.
Published: (2024)
by: Vallin, Jonatan, et al.
Published: (2024)
Beyond ReLU: How Activations Affect Neural Kernels and Random Wide Networks
by: Holzmüller, David, et al.
Published: (2025)
by: Holzmüller, David, et al.
Published: (2025)
ReLU Networks as Random Functions: Their Distribution in Probability Space
by: Chaudhari, Shreyas, et al.
Published: (2025)
by: Chaudhari, Shreyas, et al.
Published: (2025)
Similar Items
-
Stable and Robust Deep Learning By Hyperbolic Tangent Exponential Linear Unit (TeLU)
by: Fernandez, Alfredo, et al.
Published: (2024) -
Integral Signatures of Activation Functions: A 9-Dimensional Taxonomy and Stability Theory for Deep Learning
by: Mali, Ankur, et al.
Published: (2025) -
Deep Network Approximation: Beyond ReLU to Diverse Activation Functions
by: Zhang, Shijun, et al.
Published: (2023) -
Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning
by: Prada, Benjamin, et al.
Published: (2025) -
PowLU: An Activation Function for Stable Pre-Training of LLMs
by: Jiang, Peijie, et al.
Published: (2026)