Deep Neural Networks with General Activations: Super-Convergence in Sobolev Norms
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yahong, He, Juncai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Binarized Neural Networks Converge Toward Algorithmic Simplicity: Empirical Support for the Learning-as-Compression Hypothesis
by: Sakabe, Eduardo Y., et al.
Published: (2025)
by: Sakabe, Eduardo Y., et al.
Published: (2025)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
Graded Transformers
by: Shaska Sr, Tony
Published: (2025)
by: Shaska Sr, Tony
Published: (2025)
The Serial Scaling Hypothesis
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Quantum Deep Learning Still Needs a Quantum Leap
by: Gundlach, Hans, et al.
Published: (2025)
by: Gundlach, Hans, et al.
Published: (2025)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
by: Li, Yin
Published: (2025)
by: Li, Yin
Published: (2025)
Internalizing Tools as Morphisms in Graded Transformers
by: Shaska, Tony
Published: (2025)
by: Shaska, Tony
Published: (2025)
A Monad-Based Clause Architecture for Artificial Age Score (AAS) in Large Language Models
by: Kayadibi, Seyma Yaman
Published: (2025)
by: Kayadibi, Seyma Yaman
Published: (2025)
Understanding Two-Layer Neural Networks with Smooth Activation Functions
by: Huang, Changcun
Published: (2025)
by: Huang, Changcun
Published: (2025)
Recent Advances in Data-Driven Business Process Management
by: Ackermann, Lars, et al.
Published: (2024)
by: Ackermann, Lars, et al.
Published: (2024)
Noise-Adaptive Quantum Circuit Mapping for Multi-Chip NISQ Systems via Deep Reinforcement Learning
by: Zeynali, Atiye, et al.
Published: (2025)
by: Zeynali, Atiye, et al.
Published: (2025)
Efficient Binary Decision Diagram Manipulation in External Memory
by: Sølvsten, Steffan Christ, et al.
Published: (2021)
by: Sølvsten, Steffan Christ, et al.
Published: (2021)
Beyond Universal Approximation Theorems: Algorithmic Uniform Approximation by Neural Networks Trained with Noisy Data
by: Kratsios, Anastasis, et al.
Published: (2025)
by: Kratsios, Anastasis, et al.
Published: (2025)
Hallucinations Live in Variance
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Learning Program Behavioral Models from Synthesized Input-Output Pairs
by: Mammadov, Tural, et al.
Published: (2024)
by: Mammadov, Tural, et al.
Published: (2024)
Multi-variable Quantification of BDDs in External Memory using Nested Sweeping (Extended Paper)
by: Sølvsten, Steffan Christ, et al.
Published: (2024)
by: Sølvsten, Steffan Christ, et al.
Published: (2024)
Spiking Sequence Machines and Transformers
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
An in-depth look at approximation via deep and narrow neural networks
by: Dommel, Joris, et al.
Published: (2025)
by: Dommel, Joris, et al.
Published: (2025)
Alpay Algebra IV: Symbiotic Semantics and the Fixed-Point Convergence of Observer Embeddings
by: Kilictas, Bugra, et al.
Published: (2025)
by: Kilictas, Bugra, et al.
Published: (2025)
PolyKAN: A Polyhedral Analysis Framework for Provable and Approximately Optimal KAN Compression
by: Zhang, Di
Published: (2025)
by: Zhang, Di
Published: (2025)
Shallow ReLU$^s$ Networks in $L^p$-Type and Sobolev Spaces: Approximation and Path-Norm Controlled Generalization
by: Li, Weizhao, et al.
Published: (2026)
by: Li, Weizhao, et al.
Published: (2026)
STACHE: Local Black-Box Explanations for Reinforcement Learning Policies
by: Elashkin, Andrew, et al.
Published: (2025)
by: Elashkin, Andrew, et al.
Published: (2025)
Divergence-free Linearized Neural Networks: Integral Representation and Optimal Approximation Rates
by: He, Juncai, et al.
Published: (2026)
by: He, Juncai, et al.
Published: (2026)
Attention Meets Reachability: Structural Equivalence and Efficiency in Grammar-Constrained LLM Decoding
by: Alpay, Faruk, et al.
Published: (2026)
by: Alpay, Faruk, et al.
Published: (2026)
Quantum Circuits for Quantum Convolutions: A Quantum Convolutional Autoencoder
by: Orduz, Javier, et al.
Published: (2025)
by: Orduz, Javier, et al.
Published: (2025)
Symbolic Model Checking in External Memory
by: Sølvsten, Steffan Christ, et al.
Published: (2025)
by: Sølvsten, Steffan Christ, et al.
Published: (2025)
AI LLM Proof of Self-Consciousness and User-Specific Attractors
by: Camlin, Jeffrey
Published: (2025)
by: Camlin, Jeffrey
Published: (2025)
Contraction, Criticality, and Capacity: A Dynamical-Systems Perspective on Echo-State Networks
by: Singh, Pradeep, et al.
Published: (2025)
by: Singh, Pradeep, et al.
Published: (2025)
Devanagari Digit Recognition using Quantum Machine Learning
by: Malla, Sahaj Raj
Published: (2025)
by: Malla, Sahaj Raj
Published: (2025)
Application of Sensitivity Analysis Methods for Studying Neural Network Models
by: Miao, Jiaxuan, et al.
Published: (2025)
by: Miao, Jiaxuan, et al.
Published: (2025)
A Classical-Quantum Hybrid Architecture for Physics-Informed Neural Networks
by: Lantigua, Said, et al.
Published: (2025)
by: Lantigua, Said, et al.
Published: (2025)
Alpay Algebra V: Multi-Layered Semantic Games and Transfinite Fixed-Point Simulation
by: Kilictas, Bugra, et al.
Published: (2025)
by: Kilictas, Bugra, et al.
Published: (2025)
Predicting Traffic Accident Severity with Deep Neural Networks
by: Bibb, Meghan, et al.
Published: (2025)
by: Bibb, Meghan, et al.
Published: (2025)
Data-induced multiscale losses and efficient multirate gradient descent schemes
by: He, Juncai, et al.
Published: (2024)
by: He, Juncai, et al.
Published: (2024)
Active perception and disentangled representations allow continual, episodic zero and few-shot learning
by: Rawlinson, David, et al.
Published: (2026)
by: Rawlinson, David, et al.
Published: (2026)
QGraphLIME - Explaining Quantum Graph Neural Networks
by: Jena, Haribandhu, et al.
Published: (2025)
by: Jena, Haribandhu, et al.
Published: (2025)
Towards Coordinate- and Dimension-Agnostic Machine Learning for Partial Differential Equations
by: Phan, Trung V., et al.
Published: (2025)
by: Phan, Trung V., et al.
Published: (2025)
Mastering NIM and Impartial Games with Weak Neural Networks: An AlphaZero-inspired Multi-Frame Approach
by: Riis, Søren
Published: (2024)
by: Riis, Søren
Published: (2024)
Sprecher Networks: A Parameter-Efficient Kolmogorov-Arnold Architecture
by: Hägg, Christian, et al.
Published: (2025)
by: Hägg, Christian, et al.
Published: (2025)
Similar Items
-
Binarized Neural Networks Converge Toward Algorithmic Simplicity: Empirical Support for the Learning-as-Compression Hypothesis
by: Sakabe, Eduardo Y., et al.
Published: (2025) -
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025) -
Graded Transformers
by: Shaska Sr, Tony
Published: (2025) -
The Serial Scaling Hypothesis
by: Liu, Yuxi, et al.
Published: (2025) -
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)