Make Deep Networks Shallow Again
Fuente:
arXiv
Saved in:
| Main Authors: | Bermeitinger, Bernhard, Hrycej, Tomas, Handschuh, Siegfried |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Convexity-dependent Two-Phase Training Algorithm for Deep Neural Networks
by: Hrycej, Tomas, et al.
Published: (2025)
by: Hrycej, Tomas, et al.
Published: (2025)
Reducing the Transformer Architecture to a Minimum
by: Bermeitinger, Bernhard, et al.
Published: (2024)
by: Bermeitinger, Bernhard, et al.
Published: (2024)
Efficient Neural Network Training via Subset Pretraining
by: Spörer, Jan, et al.
Published: (2024)
by: Spörer, Jan, et al.
Published: (2024)
Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder
by: Wiegand, Götz-Henrik, et al.
Published: (2026)
by: Wiegand, Götz-Henrik, et al.
Published: (2026)
Born Again Neural Networks
by: Furlanello, Tommaso, et al.
Published: (2018)
by: Furlanello, Tommaso, et al.
Published: (2018)
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification
by: Tyagi, Kanishka, et al.
Published: (2024)
by: Tyagi, Kanishka, et al.
Published: (2024)
Deep Minds and Shallow Probes
by: Lee, Su Hyeong, et al.
Published: (2026)
by: Lee, Su Hyeong, et al.
Published: (2026)
Make Graph Neural Networks Great Again: A Generic Integration Paradigm of Topology-Free Patterns for Traffic Speed Prediction
by: Zhou, Yicheng, et al.
Published: (2024)
by: Zhou, Yicheng, et al.
Published: (2024)
Sparse Autoencoders, Again?
by: Lu, Yin, et al.
Published: (2025)
by: Lu, Yin, et al.
Published: (2025)
Exploring Deep-to-Shallow Transformable Neural Networks for Intelligent Embedded Systems
by: Luo, Xiangzhong, et al.
Published: (2025)
by: Luo, Xiangzhong, et al.
Published: (2025)
Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
by: Riabinin, Artem, et al.
Published: (2025)
by: Riabinin, Artem, et al.
Published: (2025)
Learning with Shallow Neural Networks on Cluster-Structured Features
by: Cornacchia, Elisabetta, et al.
Published: (2026)
by: Cornacchia, Elisabetta, et al.
Published: (2026)
Oops!... They Stole it Again: Attacks on Split Learning
by: Khan, Tanveer, et al.
Published: (2025)
by: Khan, Tanveer, et al.
Published: (2025)
Provable Learning of Random Hierarchy Models and Hierarchical Shallow-to-Deep Chaining
by: Ren, Yunwei, et al.
Published: (2026)
by: Ren, Yunwei, et al.
Published: (2026)
From Shallow to Deep: Pinning Semantic Intent via Causal GRPO
by: Zhou, Shuyi, et al.
Published: (2026)
by: Zhou, Shuyi, et al.
Published: (2026)
Neural Variance-aware Dueling Bandits with Deep Representation and Shallow Exploration
by: Oh, Youngmin, et al.
Published: (2025)
by: Oh, Youngmin, et al.
Published: (2025)
Time Scale Network: A Shallow Neural Network For Time Series Data
by: Meyer, Trevor, et al.
Published: (2023)
by: Meyer, Trevor, et al.
Published: (2023)
Noisy Interpolation Learning with Shallow Univariate ReLU Networks
by: Joshi, Nirmit, et al.
Published: (2023)
by: Joshi, Nirmit, et al.
Published: (2023)
Uniform-in-Time Weak Propagation-of-Chaos in Shallow Neural Networks
by: Glasgow, Margalit, et al.
Published: (2026)
by: Glasgow, Margalit, et al.
Published: (2026)
SDS-Net: Shallow-Deep Synergism-detection Network for infrared small target detection
by: Yue, Taoran, et al.
Published: (2025)
by: Yue, Taoran, et al.
Published: (2025)
Soft Contamination Means Benchmarks Test Shallow Generalization
by: Spiesberger, Ari, et al.
Published: (2026)
by: Spiesberger, Ari, et al.
Published: (2026)
Implicit Bias of Mirror Flow for Shallow Neural Networks in Univariate Regression
by: Liang, Shuang, et al.
Published: (2024)
by: Liang, Shuang, et al.
Published: (2024)
Finite Samples for Shallow Neural Networks
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
by: Yin, Bojian, et al.
Published: (2024)
by: Yin, Bojian, et al.
Published: (2024)
Condition Numbers and Eigenvalue Spectra of Shallow Networks on Spheres
by: Liu, Xinliang, et al.
Published: (2025)
by: Liu, Xinliang, et al.
Published: (2025)
Provable Privacy Attacks on Trained Shallow Neural Networks
by: Smorodinsky, Guy, et al.
Published: (2024)
by: Smorodinsky, Guy, et al.
Published: (2024)
Wasserstein Distributionally Robust Shallow Convex Neural Networks
by: Pallage, Julien, et al.
Published: (2024)
by: Pallage, Julien, et al.
Published: (2024)
Reduced Order Modeling with Shallow Recurrent Decoder Networks
by: Tomasetto, Matteo, et al.
Published: (2025)
by: Tomasetto, Matteo, et al.
Published: (2025)
Uncertainty Quantification in Graph Neural Networks with Shallow Ensembles
by: Vinchurkar, Tirtha, et al.
Published: (2025)
by: Vinchurkar, Tirtha, et al.
Published: (2025)
Beyond Unconstrained Features: Neural Collapse for Shallow Neural Networks with General Data
by: Hong, Wanli, et al.
Published: (2024)
by: Hong, Wanli, et al.
Published: (2024)
Deep Ridgelet Transform and Unified Universality Theorem for Deep and Shallow Joint-Group-Equivariant Machines
by: Sonoda, Sho, et al.
Published: (2024)
by: Sonoda, Sho, et al.
Published: (2024)
Why Shallow Networks Struggle to Approximate and Learn High Frequencies
by: Zhang, Shijun, et al.
Published: (2023)
by: Zhang, Shijun, et al.
Published: (2023)
Statistical Guarantees for Approximate Stationary Points of Shallow Neural Networks
by: Taheri, Mahsa, et al.
Published: (2022)
by: Taheri, Mahsa, et al.
Published: (2022)
Why and When Deep is Better than Shallow: Implementation-Agnostic State-Transition Model of Deep Learning
by: Sonoda, Sho, et al.
Published: (2025)
by: Sonoda, Sho, et al.
Published: (2025)
Surrogate models for nuclear fusion with parametric Shallow Recurrent Decoder Networks: applications to magnetohydrodynamics
by: Verso, M. Lo, et al.
Published: (2026)
by: Verso, M. Lo, et al.
Published: (2026)
Loss Landscape of Shallow ReLU-like Neural Networks: Stationary Points, Saddle Escape, and Network Embedding
by: Wu, Frank Zhengqing, et al.
Published: (2024)
by: Wu, Frank Zhengqing, et al.
Published: (2024)
Posterior Inference on Shallow Infinitely Wide Bayesian Neural Networks under Weights with Unbounded Variance
by: Loría, Jorge, et al.
Published: (2023)
by: Loría, Jorge, et al.
Published: (2023)
Resource-Efficient and Robust Inference of Deep and Bayesian Neural Networks on Embedded and Analog Computing Platforms
by: Klein, Bernhard
Published: (2025)
by: Klein, Bernhard
Published: (2025)
Application of parametric Shallow Recurrent Decoder Network to magnetohydrodynamic flows in liquid metal blankets of fusion reactors
by: Verso, M. Lo, et al.
Published: (2026)
by: Verso, M. Lo, et al.
Published: (2026)
Quotient Geometry, Effective Curvature, and Implicit Bias in Simple Shallow Neural Networks
by: Dong, Hang-Cheng, et al.
Published: (2026)
by: Dong, Hang-Cheng, et al.
Published: (2026)
Similar Items
-
A Convexity-dependent Two-Phase Training Algorithm for Deep Neural Networks
by: Hrycej, Tomas, et al.
Published: (2025) -
Reducing the Transformer Architecture to a Minimum
by: Bermeitinger, Bernhard, et al.
Published: (2024) -
Efficient Neural Network Training via Subset Pretraining
by: Spörer, Jan, et al.
Published: (2024) -
Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder
by: Wiegand, Götz-Henrik, et al.
Published: (2026) -
Born Again Neural Networks
by: Furlanello, Tommaso, et al.
Published: (2018)