Gespeichert in:
| Hauptverfasser: | Abeykoon, Chathurika S, Beknazaryan, Aleksandr, Sang, Hailin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.19351 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
von: Xu, Xianliang, et al.
Veröffentlicht: (2024)
von: Xu, Xianliang, et al.
Veröffentlicht: (2024)
Stochastic Gradient Descent for Two-layer Neural Networks
von: Cao, Dinghao, et al.
Veröffentlicht: (2024)
von: Cao, Dinghao, et al.
Veröffentlicht: (2024)
Generalization Guarantees of Gradient Descent for Multi-Layer Neural Networks
von: Wang, Puyu, et al.
Veröffentlicht: (2023)
von: Wang, Puyu, et al.
Veröffentlicht: (2023)
How Does Gradient Descent Learn Features -- A Local Analysis for Regularized Two-Layer Neural Networks
von: Zhou, Mo, et al.
Veröffentlicht: (2024)
von: Zhou, Mo, et al.
Veröffentlicht: (2024)
On the Lipschitz Constant of Deep Networks and Double Descent
von: Gamba, Matteo, et al.
Veröffentlicht: (2023)
von: Gamba, Matteo, et al.
Veröffentlicht: (2023)
Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent
von: Moniri, Behrad, et al.
Veröffentlicht: (2026)
von: Moniri, Behrad, et al.
Veröffentlicht: (2026)
Bayesian Double Descent
von: Polson, Nick, et al.
Veröffentlicht: (2025)
von: Polson, Nick, et al.
Veröffentlicht: (2025)
DSD$^2$: Can We Dodge Sparse Double Descent and Compress the Neural Network Worry-Free?
von: Quétu, Victor, et al.
Veröffentlicht: (2023)
von: Quétu, Victor, et al.
Veröffentlicht: (2023)
Manipulating Sparse Double Descent
von: Zhang, Ya Shi
Veröffentlicht: (2024)
von: Zhang, Ya Shi
Veröffentlicht: (2024)
The Benefits of Reusing Batches for Gradient Descent in Two-Layer Networks: Breaking the Curse of Information and Leap Exponents
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
Asymptotic Behavior of Multi--Task Learning: Implicit Regularization and Double Descent Effects
von: Alrashdi, Ayed M., et al.
Veröffentlicht: (2026)
von: Alrashdi, Ayed M., et al.
Veröffentlicht: (2026)
Rethinking Benign Overfitting in Two-Layer Neural Networks
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
Enhancing Deep Learning with Optimized Gradient Descent: Bridging Numerical Methods and Neural Network Training
von: Ma, Yuhan, et al.
Veröffentlicht: (2024)
von: Ma, Yuhan, et al.
Veröffentlicht: (2024)
KCNet: An Insect-Inspired Single-Hidden-Layer Neural Network with Randomized Binary Weights for Prediction and Classification Tasks
von: Hong, Jinyung, et al.
Veröffentlicht: (2021)
von: Hong, Jinyung, et al.
Veröffentlicht: (2021)
Two Speeds of Learning: A Representation-Readout Decomposition of Grokking and Double Descent
von: Chou, Chi-Ning, et al.
Veröffentlicht: (2026)
von: Chou, Chi-Ning, et al.
Veröffentlicht: (2026)
Decoupling Feature Extraction and Classification Layers for Calibrated Neural Networks
von: Jordahn, Mikkel, et al.
Veröffentlicht: (2024)
von: Jordahn, Mikkel, et al.
Veröffentlicht: (2024)
Dropout Drops Double Descent
von: Yang, Tian-Le, et al.
Veröffentlicht: (2023)
von: Yang, Tian-Le, et al.
Veröffentlicht: (2023)
Large Stepsize Gradient Descent for Non-Homogeneous Two-Layer Networks: Margin Improvement and Fast Optimization
von: Cai, Yuhang, et al.
Veröffentlicht: (2024)
von: Cai, Yuhang, et al.
Veröffentlicht: (2024)
Double Descent and Other Interpolation Phenomena in GANs
von: Luzi, Lorenzo, et al.
Veröffentlicht: (2021)
von: Luzi, Lorenzo, et al.
Veröffentlicht: (2021)
Hybrid Coordinate Descent for Efficient Neural Network Learning Using Line Search and Gradient Descent
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
Variational Stochastic Gradient Descent for Deep Neural Networks
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
Astrometric Binary Classification Via Artificial Neural Networks
von: Smith, Joe
Veröffentlicht: (2024)
von: Smith, Joe
Veröffentlicht: (2024)
Exploring Spiking Neural Networks for Binary Classification in Multivariate Time Series at the Edge
von: Ghawaly, James, et al.
Veröffentlicht: (2025)
von: Ghawaly, James, et al.
Veröffentlicht: (2025)
Preconditioned Norms: A Unified Framework for Steepest Descent, Quasi-Newton and Adaptive Methods
von: Veprikov, Andrey, et al.
Veröffentlicht: (2025)
von: Veprikov, Andrey, et al.
Veröffentlicht: (2025)
Towards Understanding Epoch-wise Double descent in Two-layer Linear Neural Networks
von: Olmin, Amanda, et al.
Veröffentlicht: (2024)
von: Olmin, Amanda, et al.
Veröffentlicht: (2024)
Generalization Bounds of Stochastic Gradient Descent in Homogeneous Neural Networks
von: Ma, Wenquan, et al.
Veröffentlicht: (2026)
von: Ma, Wenquan, et al.
Veröffentlicht: (2026)
On The Presence of Double-Descent in Deep Reinforcement Learning
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
Quantum Convolutional Neural Networks with Interaction Layers for Classification of Classical Data
von: Mahmud, Jishnu, et al.
Veröffentlicht: (2023)
von: Mahmud, Jishnu, et al.
Veröffentlicht: (2023)
Training Multi-Layer Binary Neural Networks With Local Binary Error Signals
von: Colombo, Luca, et al.
Veröffentlicht: (2024)
von: Colombo, Luca, et al.
Veröffentlicht: (2024)
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization
von: Boža, Vladimír, et al.
Veröffentlicht: (2024)
von: Boža, Vladimír, et al.
Veröffentlicht: (2024)
AutoGrid AI: Deep Reinforcement Learning Framework for Autonomous Microgrid Management
von: Guo, Kenny, et al.
Veröffentlicht: (2025)
von: Guo, Kenny, et al.
Veröffentlicht: (2025)
On the Theory of Continual Learning with Gradient Descent for Neural Networks
von: Taheri, Hossein, et al.
Veröffentlicht: (2025)
von: Taheri, Hossein, et al.
Veröffentlicht: (2025)
Convex Formulations for Training Two-Layer ReLU Neural Networks
von: Prakhya, Karthik, et al.
Veröffentlicht: (2024)
von: Prakhya, Karthik, et al.
Veröffentlicht: (2024)
Class-wise Activation Unravelling the Engima of Deep Double Descent
von: Gu, Yufei
Veröffentlicht: (2024)
von: Gu, Yufei
Veröffentlicht: (2024)
Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
von: Yu, Hao
Veröffentlicht: (2025)
von: Yu, Hao
Veröffentlicht: (2025)
Training Guarantees of Neural Network Classification Two-Sample Tests by Kernel Analysis
von: Khurana, Varun, et al.
Veröffentlicht: (2024)
von: Khurana, Varun, et al.
Veröffentlicht: (2024)
Provable Multi-Task Representation Learning by Two-Layer ReLU Neural Networks
von: Collins, Liam, et al.
Veröffentlicht: (2023)
von: Collins, Liam, et al.
Veröffentlicht: (2023)
How Two-Layer Neural Networks Learn, One (Giant) Step at a Time
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
Understanding the Benefits of SimCLR Pre-Training in Two-Layer Convolutional Neural Networks
von: Zhang, Han, et al.
Veröffentlicht: (2024)
von: Zhang, Han, et al.
Veröffentlicht: (2024)
Analyzing Neural Scaling Laws in Two-Layer Networks with Power-Law Data Spectra
von: Worschech, Roman, et al.
Veröffentlicht: (2024)
von: Worschech, Roman, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
von: Xu, Xianliang, et al.
Veröffentlicht: (2024) -
Stochastic Gradient Descent for Two-layer Neural Networks
von: Cao, Dinghao, et al.
Veröffentlicht: (2024) -
Generalization Guarantees of Gradient Descent for Multi-Layer Neural Networks
von: Wang, Puyu, et al.
Veröffentlicht: (2023) -
How Does Gradient Descent Learn Features -- A Local Analysis for Regularized Two-Layer Neural Networks
von: Zhou, Mo, et al.
Veröffentlicht: (2024) -
On the Lipschitz Constant of Deep Networks and Double Descent
von: Gamba, Matteo, et al.
Veröffentlicht: (2023)