Towards Understanding Epoch-wise Double descent in Two-layer Linear Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Olmin, Amanda, Lindsten, Fredrik |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the connection between Noise-Contrastive Estimation and Contrastive Divergence
by: Olmin, Amanda, et al.
Published: (2024)
by: Olmin, Amanda, et al.
Published: (2024)
Active Learning with Weak Supervision for Gaussian Processes
by: Olmin, Amanda, et al.
Published: (2022)
by: Olmin, Amanda, et al.
Published: (2022)
Probabilistic Weather Forecasting with Hierarchical Graph Neural Networks
by: Oskarsson, Joel, et al.
Published: (2024)
by: Oskarsson, Joel, et al.
Published: (2024)
Solving Linear-Gaussian Bayesian Inverse Problems with Decoupled Diffusion Sequential Monte Carlo
by: Kelvinius, Filip Ekström, et al.
Published: (2025)
by: Kelvinius, Filip Ekström, et al.
Published: (2025)
Discriminator Guidance for Autoregressive Diffusion Models
by: Kelvinius, Filip Ekström, et al.
Published: (2023)
by: Kelvinius, Filip Ekström, et al.
Published: (2023)
Prior Learning in Introspective VAEs
by: Athanasiadis, Ioannis, et al.
Published: (2024)
by: Athanasiadis, Ioannis, et al.
Published: (2024)
Particle-Guided Diffusion Models for Partial Differential Equations
by: Millard, Andrew, et al.
Published: (2026)
by: Millard, Andrew, et al.
Published: (2026)
Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting
by: Holmberg, Daniel, et al.
Published: (2026)
by: Holmberg, Daniel, et al.
Published: (2026)
Discriminative Ordering Through Ensemble Consensus
by: Ohl, Louis, et al.
Published: (2025)
by: Ohl, Louis, et al.
Published: (2025)
Fast and Scalable Score-Based Kernel Calibration Tests
by: Glaser, Pierre, et al.
Published: (2025)
by: Glaser, Pierre, et al.
Published: (2025)
ESS-Flow: Training-free guidance of flow-based models as inference in source space
by: Kalaivanan, Adhithyan, et al.
Published: (2025)
by: Kalaivanan, Adhithyan, et al.
Published: (2025)
CRPS-LAM: Regional ensemble weather forecasting from matching marginals
by: Larsson, Erik, et al.
Published: (2025)
by: Larsson, Erik, et al.
Published: (2025)
cryoSPHERE: Single-particle heterogeneous reconstruction from cryo EM
by: Ducrocq, Gabriel, et al.
Published: (2024)
by: Ducrocq, Gabriel, et al.
Published: (2024)
Elements of Sequential Monte Carlo
by: Naesseth, Christian A., et al.
Published: (2019)
by: Naesseth, Christian A., et al.
Published: (2019)
Continuous Ensemble Weather Forecasting with Diffusion models
by: Andrae, Martin, et al.
Published: (2024)
by: Andrae, Martin, et al.
Published: (2024)
Diffusion-LAM: Probabilistic Limited Area Weather Forecasting with Diffusion
by: Larsson, Erik, et al.
Published: (2025)
by: Larsson, Erik, et al.
Published: (2025)
Deep Exploration of Epoch-wise Double Descent in Noisy Data: Signal Separation, Large Activation, and Benign Overfitting
by: Kubo, Tomoki, et al.
Published: (2026)
by: Kubo, Tomoki, et al.
Published: (2026)
DINO as a von Mises-Fisher mixture model
by: Govindarajan, Hariprasath, et al.
Published: (2024)
by: Govindarajan, Hariprasath, et al.
Published: (2024)
On Partial Prototype Collapse in the DINO Family of Self-Supervised Methods
by: Govindarajan, Hariprasath, et al.
Published: (2024)
by: Govindarajan, Hariprasath, et al.
Published: (2024)
QUEST: A robust attention formulation using query-modulated spherical attention
by: Govindarajan, Hariprasath, et al.
Published: (2026)
by: Govindarajan, Hariprasath, et al.
Published: (2026)
DAISI: Data Assimilation with Inverse Sampling using Stochastic Interpolants
by: Andrae, Martin, et al.
Published: (2025)
by: Andrae, Martin, et al.
Published: (2025)
Stochastic Gradient Descent for Two-layer Neural Networks
by: Cao, Dinghao, et al.
Published: (2024)
by: Cao, Dinghao, et al.
Published: (2024)
Double descent in quantum kernel methods
by: Kempkes, Marie, et al.
Published: (2025)
by: Kempkes, Marie, et al.
Published: (2025)
The Double Descent Behavior in Two Layer Neural Network for Binary Classification
by: Abeykoon, Chathurika S, et al.
Published: (2025)
by: Abeykoon, Chathurika S, et al.
Published: (2025)
Learning with Norm Constrained, Over-parameterized, Two-layer Neural Networks
by: Liu, Fanghui, et al.
Published: (2024)
by: Liu, Fanghui, et al.
Published: (2024)
On the Learning Dynamics of Two-layer Linear Networks with Label Noise SGD
by: Zhang, Tongcheng, et al.
Published: (2026)
by: Zhang, Tongcheng, et al.
Published: (2026)
The Unreasonable Effectiveness Of Early Discarding After One Epoch In Neural Network Hyperparameter Optimization
by: Egele, Romain, et al.
Published: (2024)
by: Egele, Romain, et al.
Published: (2024)
Layer-wise Linear Mode Connectivity
by: Adilova, Linara, et al.
Published: (2023)
by: Adilova, Linara, et al.
Published: (2023)
Revisiting Likelihood-Based Out-of-Distribution Detection by Modeling Representations
by: Ding, Yifan, et al.
Published: (2025)
by: Ding, Yifan, et al.
Published: (2025)
Double descent for least-squares interpolation on contaminated data: A simulation study
by: Werner, Tino
Published: (2026)
by: Werner, Tino
Published: (2026)
Mean-field Analysis on Two-layer Neural Networks from a Kernel Perspective
by: Takakura, Shokichi, et al.
Published: (2024)
by: Takakura, Shokichi, et al.
Published: (2024)
Alternating Gradient Flows: A Theory of Feature Learning in Two-layer Neural Networks
by: Kunin, Daniel, et al.
Published: (2025)
by: Kunin, Daniel, et al.
Published: (2025)
Toward Understanding the Disagreement Problem in Neural Network Feature Attribution
by: Koenen, Niklas, et al.
Published: (2024)
by: Koenen, Niklas, et al.
Published: (2024)
Concept Learning in the Wild: Towards Algorithmic Understanding of Neural Networks
by: Shoham, Elad, et al.
Published: (2024)
by: Shoham, Elad, et al.
Published: (2024)
ATE-SG: Alternate Through the Epochs Stochastic Gradient for Multi-Task Neural Networks
by: Bellavia, Stefania, et al.
Published: (2023)
by: Bellavia, Stefania, et al.
Published: (2023)
Class-wise Activation Unravelling the Engima of Deep Double Descent
by: Gu, Yufei
Published: (2024)
by: Gu, Yufei
Published: (2024)
Fast Tensorization of Neural Networks via Slice-wise Feature Distillation
by: Hamreras, Safa, et al.
Published: (2026)
by: Hamreras, Safa, et al.
Published: (2026)
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization
by: Boža, Vladimír, et al.
Published: (2024)
by: Boža, Vladimír, et al.
Published: (2024)
Toward Understanding Convolutional Neural Networks from Volterra Convolution Perspective
by: Li, Tenghui, et al.
Published: (2021)
by: Li, Tenghui, et al.
Published: (2021)
Geometry of Critical Sets and Existence of Saddle Branches for Two-layer Neural Networks
by: Zhang, Leyang, et al.
Published: (2024)
by: Zhang, Leyang, et al.
Published: (2024)
Similar Items
-
On the connection between Noise-Contrastive Estimation and Contrastive Divergence
by: Olmin, Amanda, et al.
Published: (2024) -
Active Learning with Weak Supervision for Gaussian Processes
by: Olmin, Amanda, et al.
Published: (2022) -
Probabilistic Weather Forecasting with Hierarchical Graph Neural Networks
by: Oskarsson, Joel, et al.
Published: (2024) -
Solving Linear-Gaussian Bayesian Inverse Problems with Decoupled Diffusion Sequential Monte Carlo
by: Kelvinius, Filip Ekström, et al.
Published: (2025) -
Discriminator Guidance for Autoregressive Diffusion Models
by: Kelvinius, Filip Ekström, et al.
Published: (2023)