On efficiently computable functions, deep networks and sparse compositionality
Fuente:
arXiv
Salvato in:
| Autore principale: | Poggio, Tomaso |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Generalization Bounds for Neural Networks with Low Rank Layers
di: Pinto, Andrea, et al.
Pubblicazione: (2024)
di: Pinto, Andrea, et al.
Pubblicazione: (2024)
Ubiquity of Emergent Hebbian Dynamics in Regularized Learning
di: Koplow, David, et al.
Pubblicazione: (2025)
di: Koplow, David, et al.
Pubblicazione: (2025)
Hierarchical Reasoning Models: Perspectives and Misconceptions
di: Ge, Renee, et al.
Pubblicazione: (2025)
di: Ge, Renee, et al.
Pubblicazione: (2025)
Motif distribution and function of sparse deep neural networks
di: Zahn, Olivia T., et al.
Pubblicazione: (2024)
di: Zahn, Olivia T., et al.
Pubblicazione: (2024)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
di: Beneventano, Pierfrancesco, et al.
Pubblicazione: (2024)
di: Beneventano, Pierfrancesco, et al.
Pubblicazione: (2024)
pAI/MSc: ML Theory Research with Humans on the Loop
di: Abdelmoneum, Mahmoud, et al.
Pubblicazione: (2026)
di: Abdelmoneum, Mahmoud, et al.
Pubblicazione: (2026)
Does Feedback Alignment Work at Biological Timescales?
di: Bacvanski, Marc Gong, et al.
Pubblicazione: (2025)
di: Bacvanski, Marc Gong, et al.
Pubblicazione: (2025)
Topological Invariance and Breakdown in Learning
di: Yang, Yongyi, et al.
Pubblicazione: (2025)
di: Yang, Yongyi, et al.
Pubblicazione: (2025)
Learning Sparse Compositional Functions with Norm-Constrained Neural Networks
di: Huang, Shuo, et al.
Pubblicazione: (2026)
di: Huang, Shuo, et al.
Pubblicazione: (2026)
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network
di: Galanti, Tomer, et al.
Pubblicazione: (2022)
di: Galanti, Tomer, et al.
Pubblicazione: (2022)
Does Weight Decay Enhance Training Stability?
di: Saether, Marius, et al.
Pubblicazione: (2026)
di: Saether, Marius, et al.
Pubblicazione: (2026)
Iterative regularization in classification via hinge loss diagonal descent
di: Apidopoulos, Vassilis, et al.
Pubblicazione: (2022)
di: Apidopoulos, Vassilis, et al.
Pubblicazione: (2022)
Unraveling Syntax: How Language Models Learn Context-Free Grammars
di: Schulz, Laura Ying, et al.
Pubblicazione: (2025)
di: Schulz, Laura Ying, et al.
Pubblicazione: (2025)
Formation of Representations in Neural Networks
di: Ziyin, Liu, et al.
Pubblicazione: (2024)
di: Ziyin, Liu, et al.
Pubblicazione: (2024)
Too Sharp, Too Sure: When Calibration Follows Curvature
di: Morosini, Alessandro, et al.
Pubblicazione: (2026)
di: Morosini, Alessandro, et al.
Pubblicazione: (2026)
Heterosynaptic Circuits Are Universal Gradient Machines
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
Learning Multi-Index Models with Hyper-Kernel Ridge Regression
di: Huang, Shuo, et al.
Pubblicazione: (2025)
di: Huang, Shuo, et al.
Pubblicazione: (2025)
Parameter Symmetry Potentially Unifies Deep Learning Theory
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
Learning smooth functions in high dimensions: from sparse polynomials to deep neural networks
di: Adcock, Ben, et al.
Pubblicazione: (2024)
di: Adcock, Ben, et al.
Pubblicazione: (2024)
Position: A Theory of Deep Learning Must Include Compositional Sparsity
di: Danhofer, David A., et al.
Pubblicazione: (2025)
di: Danhofer, David A., et al.
Pubblicazione: (2025)
Same Error, Different Function: The Optimizer as an Implicit Prior in Financial Time Series
di: Cortesi, Federico Vittorio, et al.
Pubblicazione: (2026)
di: Cortesi, Federico Vittorio, et al.
Pubblicazione: (2026)
Momentum Further Constrains Sharpness at the Edge of Stochastic Stability
di: Andreyev, Arseniy, et al.
Pubblicazione: (2026)
di: Andreyev, Arseniy, et al.
Pubblicazione: (2026)
The Generalized Turing Test: A Foundation for Comparing Intelligence
di: Mitropolsky, Daniel, et al.
Pubblicazione: (2026)
di: Mitropolsky, Daniel, et al.
Pubblicazione: (2026)
Sparse deep neural networks for nonparametric estimation in high-dimensional sparse regression
di: Wu, Dongya, et al.
Pubblicazione: (2024)
di: Wu, Dongya, et al.
Pubblicazione: (2024)
Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias
di: Das, Mohua, et al.
Pubblicazione: (2026)
di: Das, Mohua, et al.
Pubblicazione: (2026)
Information Filtering Networks: Theoretical Foundations, Generative Methodologies, and Real-World Applications
di: Aste, Tomaso
Pubblicazione: (2025)
di: Aste, Tomaso
Pubblicazione: (2025)
Iteratively reweighted kernel machines efficiently learn sparse functions
di: Zhu, Libin, et al.
Pubblicazione: (2025)
di: Zhu, Libin, et al.
Pubblicazione: (2025)
Don't be lazy: CompleteP enables compute-efficient deep transformers
di: Dey, Nolan, et al.
Pubblicazione: (2025)
di: Dey, Nolan, et al.
Pubblicazione: (2025)
Dynamical system prediction from sparse observations using deep neural networks with Voronoi tessellation and physics constraint
di: Wang, Hanyang, et al.
Pubblicazione: (2024)
di: Wang, Hanyang, et al.
Pubblicazione: (2024)
A universal compression theory for lottery ticket hypothesis and neural scaling laws
di: Wang, Hong-Yi, et al.
Pubblicazione: (2025)
di: Wang, Hong-Yi, et al.
Pubblicazione: (2025)
Nonlinear functional regression by functional deep neural network with kernel embedding
di: Shi, Zhongjie, et al.
Pubblicazione: (2024)
di: Shi, Zhongjie, et al.
Pubblicazione: (2024)
Do deep neural networks utilize the weight space efficiently?
di: Koyun, Onur Can, et al.
Pubblicazione: (2024)
di: Koyun, Onur Can, et al.
Pubblicazione: (2024)
Training the Untrainable: Introducing Inductive Bias via Representational Alignment
di: Subramaniam, Vighnesh, et al.
Pubblicazione: (2024)
di: Subramaniam, Vighnesh, et al.
Pubblicazione: (2024)
Misclassification bounds for PAC-Bayesian sparse deep learning
di: Mai, The Tien
Pubblicazione: (2024)
di: Mai, The Tien
Pubblicazione: (2024)
Large-width functional asymptotics for deep Gaussian neural networks
di: Bracale, Daniele, et al.
Pubblicazione: (2021)
di: Bracale, Daniele, et al.
Pubblicazione: (2021)
Certified and accurate computation of function space norms of deep neural networks
di: Gründler, Johannes, et al.
Pubblicazione: (2026)
di: Gründler, Johannes, et al.
Pubblicazione: (2026)
From superposition to sparse codes: interpretable representations in neural networks
di: Klindt, David, et al.
Pubblicazione: (2025)
di: Klindt, David, et al.
Pubblicazione: (2025)
SLTrain: a sparse plus low-rank approach for parameter and memory efficient pretraining
di: Han, Andi, et al.
Pubblicazione: (2024)
di: Han, Andi, et al.
Pubblicazione: (2024)
Unraveling the Enigma of Double Descent: An In-depth Analysis through the Lens of Learned Feature Space
di: Gu, Yufei, et al.
Pubblicazione: (2023)
di: Gu, Yufei, et al.
Pubblicazione: (2023)
Exact discovery is polynomial for certain sparse causal Bayesian networks
di: Rios, Felix L., et al.
Pubblicazione: (2024)
di: Rios, Felix L., et al.
Pubblicazione: (2024)
Documenti analoghi
-
On Generalization Bounds for Neural Networks with Low Rank Layers
di: Pinto, Andrea, et al.
Pubblicazione: (2024) -
Ubiquity of Emergent Hebbian Dynamics in Regularized Learning
di: Koplow, David, et al.
Pubblicazione: (2025) -
Hierarchical Reasoning Models: Perspectives and Misconceptions
di: Ge, Renee, et al.
Pubblicazione: (2025) -
Motif distribution and function of sparse deep neural networks
di: Zahn, Olivia T., et al.
Pubblicazione: (2024) -
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
di: Beneventano, Pierfrancesco, et al.
Pubblicazione: (2024)