Deep neural networks with dependent weights: Gaussian Process mixture limit, heavy tails, sparsity and compressibility
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Hoil, Ayed, Fadhel, Jung, Paul, Lee, Juho, Yang, Hongseok, Caron, François |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Over-parameterised Shallow Neural Networks with Asymmetrical Node Scaling: Global Convergence Guarantees and Feature Learning
by: Caron, Francois, et al.
Published: (2023)
by: Caron, Francois, et al.
Published: (2023)
Generalization error bounds for two-layer neural networks with Lipschitz loss function
by: Nguwi, Jiang Yu, et al.
Published: (2026)
by: Nguwi, Jiang Yu, et al.
Published: (2026)
Non-asymptotic analysis of the performance of the penalized least trimmed squares in sparse models
by: Zuo, Yijun
Published: (2025)
by: Zuo, Yijun
Published: (2025)
Polynomial Chaos Surrogate Construction for Random Fields with Parametric Uncertainty
by: Mueller, Joy N., et al.
Published: (2023)
by: Mueller, Joy N., et al.
Published: (2023)
The Normal-Generalised Gamma-Pareto process: A novel pure-jump Lévy process with flexible tail and jump-activity properties
by: Ayed, Fadhel, et al.
Published: (2020)
by: Ayed, Fadhel, et al.
Published: (2020)
Asymptotic properties of the normalized discrete associated-kernel estimator for probability mass function
by: Esstafa, Youssef, et al.
Published: (2022)
by: Esstafa, Youssef, et al.
Published: (2022)
Projection predictive variable selection for discrete response families with finite support
by: Weber, Frank, et al.
Published: (2023)
by: Weber, Frank, et al.
Published: (2023)
The Predictive-Causal Gap: An Impossibility Theorem and Large-Scale Neural Evidence
by: Liu, Kejun
Published: (2026)
by: Liu, Kejun
Published: (2026)
On Finite Time Span Estimators of Parameters for Ornstein-Uhlenbeck Processes
by: Han, Jun S., et al.
Published: (2025)
by: Han, Jun S., et al.
Published: (2025)
Modeling Human Spatial Mobility Patterns with the Lévy Flight Cluster Model
by: Wolff, Malcolm, et al.
Published: (2025)
by: Wolff, Malcolm, et al.
Published: (2025)
Efficient reconstruction of multidimensional random field models with heterogeneous data using stochastic neural networks
by: Xia, Mingtao, et al.
Published: (2025)
by: Xia, Mingtao, et al.
Published: (2025)
Exchangeable Gaussian Processes for Staggered-Adoption Policy Evaluation
by: Gevorgyan, Hayk, et al.
Published: (2026)
by: Gevorgyan, Hayk, et al.
Published: (2026)
Multiple combined gamma kernel estimations for nonnegative data with Bayesian adaptive bandwidths
by: Somé, Sobom M., et al.
Published: (2022)
by: Somé, Sobom M., et al.
Published: (2022)
In almost all shallow analytic neural network optimization landscapes, efficient minimizers have strongly convex neighborhoods
by: Benning, Felix, et al.
Published: (2025)
by: Benning, Felix, et al.
Published: (2025)
Restricted Path Characteristic Function Determines the Law of Stochastic Processes
by: Li, Siran, et al.
Published: (2024)
by: Li, Siran, et al.
Published: (2024)
Poisson Regression in one Covariate on Massive Data
by: Reuter, Torsten, et al.
Published: (2024)
by: Reuter, Torsten, et al.
Published: (2024)
Sub-Riemannian Landmark Matching and its interpretation as residual neural networks
by: Jansson, Erik, et al.
Published: (2022)
by: Jansson, Erik, et al.
Published: (2022)
Conditional Density Estimation, Latent Variable Discovery and Optimal Transport
by: Yang, Hongkang, et al.
Published: (2019)
by: Yang, Hongkang, et al.
Published: (2019)
Beyond the Chinese Restaurant and Pitman-Yor processes: Statistical Models with Double Power-law Behavior
by: Ayed, Fadhel, et al.
Published: (2019)
by: Ayed, Fadhel, et al.
Published: (2019)
HRM-Agent: Training a recurrent reasoning model in dynamic environments using reinforcement learning
by: Dang, Long H, et al.
Published: (2025)
by: Dang, Long H, et al.
Published: (2025)
A monotonic MM-type algorithm for estimation of nonparametric finite mixture models with dependent marginals
by: Levine, Michael
Published: (2025)
by: Levine, Michael
Published: (2025)
Physical oceanography during Bjarni Saemundsson cruise B07/99
by: VEINS Members, et al.
Published: (2011)
by: VEINS Members, et al.
Published: (2011)
A Complete Symmetry Classification of Shallow ReLU Networks
by: Ramakrishnan, Pranavkrishnan
Published: (2026)
by: Ramakrishnan, Pranavkrishnan
Published: (2026)
A functional Hungarian construction for sums of independent random variables
by: Grama, Ion, et al.
Published: (2024)
by: Grama, Ion, et al.
Published: (2024)
Regularized least squares learning with heavy-tailed noise is minimax optimal
by: Mollenhauer, Mattes, et al.
Published: (2025)
by: Mollenhauer, Mattes, et al.
Published: (2025)
Sobol' Matrices For Multi-Output Models With Quantified Uncertainty
by: Milton, Robert A., et al.
Published: (2025)
by: Milton, Robert A., et al.
Published: (2025)
Hydrochemistry measured on water bottle samples during Bjarni Saemundsson cruise B07/99
by: VEINS Members, et al.
Published: (2011)
by: VEINS Members, et al.
Published: (2011)
Dimensionality-Aware Outlier Detection: Theoretical and Experimental Analysis
by: Anderberg, Alastair, et al.
Published: (2024)
by: Anderberg, Alastair, et al.
Published: (2024)
Bubble Lattices II: Combinatorics
by: McConville, Thomas, et al.
Published: (2022)
by: McConville, Thomas, et al.
Published: (2022)
Fractal and Regular Geometry of Deep Neural Networks
by: Di Lillo, Simmaco, et al.
Published: (2025)
by: Di Lillo, Simmaco, et al.
Published: (2025)
Integrating Attendance Tracking and Emotion Detection for Enhanced Student Engagement in Smart Classrooms
by: Ainebyona, Keith, et al.
Published: (2026)
by: Ainebyona, Keith, et al.
Published: (2026)
Autoencoders in Function Space
by: Bunker, Justin, et al.
Published: (2024)
by: Bunker, Justin, et al.
Published: (2024)
Preconditioned Conjugate Gradient methods for the estimation of General Linear Models
by: Foschi, Paolo
Published: (2025)
by: Foschi, Paolo
Published: (2025)
Density estimation for compositional data using nonparametric mixtures
by: Xie, Jiajin, et al.
Published: (2025)
by: Xie, Jiajin, et al.
Published: (2025)
Boosted generalized normal distributions: Integrating machine learning with operations knowledge
by: Gurlek, Ragip, et al.
Published: (2024)
by: Gurlek, Ragip, et al.
Published: (2024)
Statistical inference for Levy-driven graph supOU processes: From short- to long-memory in high-dimensional time series
by: Mehta, Shreya, et al.
Published: (2025)
by: Mehta, Shreya, et al.
Published: (2025)
Gaussian mixture models as a proxy for interacting language models
by: Wang, Edward L., et al.
Published: (2025)
by: Wang, Edward L., et al.
Published: (2025)
Error analysis for empirical risk minimization over clipped ReLU networks in solving linear Kolmogorov partial differential equations
by: Xiao, Jichang, et al.
Published: (2023)
by: Xiao, Jichang, et al.
Published: (2023)
Deconvolution of distribution functions without integral transforms
by: Kaiser, Henrik
Published: (2025)
by: Kaiser, Henrik
Published: (2025)
Universal Adaptive Environment Discovery
by: Matymov, Madi, et al.
Published: (2025)
by: Matymov, Madi, et al.
Published: (2025)
Similar Items
-
Over-parameterised Shallow Neural Networks with Asymmetrical Node Scaling: Global Convergence Guarantees and Feature Learning
by: Caron, Francois, et al.
Published: (2023) -
Generalization error bounds for two-layer neural networks with Lipschitz loss function
by: Nguwi, Jiang Yu, et al.
Published: (2026) -
Non-asymptotic analysis of the performance of the penalized least trimmed squares in sparse models
by: Zuo, Yijun
Published: (2025) -
Polynomial Chaos Surrogate Construction for Random Fields with Parametric Uncertainty
by: Mueller, Joy N., et al.
Published: (2023) -
The Normal-Generalised Gamma-Pareto process: A novel pure-jump Lévy process with flexible tail and jump-activity properties
by: Ayed, Fadhel, et al.
Published: (2020)