Learning Sparse Compositional Functions with Norm-Constrained Neural Networks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Shuo, Fiorito, Lorenzo, Rosasco, Lorenzo, Poggio, Tomaso |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Multi-Index Models with Hyper-Kernel Ridge Regression
von: Huang, Shuo, et al.
Veröffentlicht: (2025)
von: Huang, Shuo, et al.
Veröffentlicht: (2025)
Iterative regularization in classification via hinge loss diagonal descent
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
Learning functions, operators and dynamical systems with kernels
von: Rosasco, Lorenzo
Veröffentlicht: (2025)
von: Rosasco, Lorenzo
Veröffentlicht: (2025)
Towards a Learning Theory of Representation Alignment
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
On efficiently computable functions, deep networks and sparse compositionality
von: Poggio, Tomaso
Veröffentlicht: (2025)
von: Poggio, Tomaso
Veröffentlicht: (2025)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2024)
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2024)
On Generalization Bounds for Neural Networks with Low Rank Layers
von: Pinto, Andrea, et al.
Veröffentlicht: (2024)
von: Pinto, Andrea, et al.
Veröffentlicht: (2024)
The $φ$ Curve: The Shape of Generalization through the Lens of Norm-based Capacity Control
von: Wang, Yichen, et al.
Veröffentlicht: (2025)
von: Wang, Yichen, et al.
Veröffentlicht: (2025)
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network
von: Galanti, Tomer, et al.
Veröffentlicht: (2022)
von: Galanti, Tomer, et al.
Veröffentlicht: (2022)
Formation of Representations in Neural Networks
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
Ubiquity of Emergent Hebbian Dynamics in Regularized Learning
von: Koplow, David, et al.
Veröffentlicht: (2025)
von: Koplow, David, et al.
Veröffentlicht: (2025)
Position: A Theory of Deep Learning Must Include Compositional Sparsity
von: Danhofer, David A., et al.
Veröffentlicht: (2025)
von: Danhofer, David A., et al.
Veröffentlicht: (2025)
Topological Invariance and Breakdown in Learning
von: Yang, Yongyi, et al.
Veröffentlicht: (2025)
von: Yang, Yongyi, et al.
Veröffentlicht: (2025)
On the Sample Complexity of Learning for Blind Inverse Problems
von: Buskulic, Nathan, et al.
Veröffentlicht: (2025)
von: Buskulic, Nathan, et al.
Veröffentlicht: (2025)
Riemann Tensor Neural Networks: Learning Conservative Systems with Physics-Constrained Networks
von: Jnini, Anas, et al.
Veröffentlicht: (2025)
von: Jnini, Anas, et al.
Veröffentlicht: (2025)
Momentum Further Constrains Sharpness at the Edge of Stochastic Stability
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
Hierarchical Reasoning Models: Perspectives and Misconceptions
von: Ge, Renee, et al.
Veröffentlicht: (2025)
von: Ge, Renee, et al.
Veröffentlicht: (2025)
Neural reproducing kernel Banach spaces and representer theorems for deep networks
von: Bartolucci, Francesca, et al.
Veröffentlicht: (2024)
von: Bartolucci, Francesca, et al.
Veröffentlicht: (2024)
pAI/MSc: ML Theory Research with Humans on the Loop
von: Abdelmoneum, Mahmoud, et al.
Veröffentlicht: (2026)
von: Abdelmoneum, Mahmoud, et al.
Veröffentlicht: (2026)
Unraveling Syntax: How Language Models Learn Context-Free Grammars
von: Schulz, Laura Ying, et al.
Veröffentlicht: (2025)
von: Schulz, Laura Ying, et al.
Veröffentlicht: (2025)
Learning with Norm Constrained, Over-parameterized, Two-layer Neural Networks
von: Liu, Fanghui, et al.
Veröffentlicht: (2024)
von: Liu, Fanghui, et al.
Veröffentlicht: (2024)
Does Feedback Alignment Work at Biological Timescales?
von: Bacvanski, Marc Gong, et al.
Veröffentlicht: (2025)
von: Bacvanski, Marc Gong, et al.
Veröffentlicht: (2025)
A New Formulation for Zeroth-Order Optimization of Adversarial EXEmples in Malware Detection
von: Rando, Marco, et al.
Veröffentlicht: (2024)
von: Rando, Marco, et al.
Veröffentlicht: (2024)
Optimization Insights into Deep Diagonal Linear Networks
von: Labarrière, Hippolyte, et al.
Veröffentlicht: (2024)
von: Labarrière, Hippolyte, et al.
Veröffentlicht: (2024)
Same Error, Different Function: The Optimizer as an Implicit Prior in Financial Time Series
von: Cortesi, Federico Vittorio, et al.
Veröffentlicht: (2026)
von: Cortesi, Federico Vittorio, et al.
Veröffentlicht: (2026)
Parameter Symmetry Potentially Unifies Deep Learning Theory
von: Ziyin, Liu, et al.
Veröffentlicht: (2025)
von: Ziyin, Liu, et al.
Veröffentlicht: (2025)
Does Weight Decay Enhance Training Stability?
von: Saether, Marius, et al.
Veröffentlicht: (2026)
von: Saether, Marius, et al.
Veröffentlicht: (2026)
SGD for Variational Inference: Tackling Unbounded Variance via Preconditioning and Dynamic Batching
von: Labarrière, Hippolyte, et al.
Veröffentlicht: (2026)
von: Labarrière, Hippolyte, et al.
Veröffentlicht: (2026)
Stochastic Zeroth order Descent with Structured Directions
von: Rando, Marco, et al.
Veröffentlicht: (2022)
von: Rando, Marco, et al.
Veröffentlicht: (2022)
A Scalable Nystrom-Based Kernel Two-Sample Test with Permutations
von: Chatalic, Antoine, et al.
Veröffentlicht: (2025)
von: Chatalic, Antoine, et al.
Veröffentlicht: (2025)
The Nyström method for convex loss functions
von: Della Vecchia, Andrea, et al.
Veröffentlicht: (2020)
von: Della Vecchia, Andrea, et al.
Veröffentlicht: (2020)
Too Sharp, Too Sure: When Calibration Follows Curvature
von: Morosini, Alessandro, et al.
Veröffentlicht: (2026)
von: Morosini, Alessandro, et al.
Veröffentlicht: (2026)
Heterosynaptic Circuits Are Universal Gradient Machines
von: Ziyin, Liu, et al.
Veröffentlicht: (2025)
von: Ziyin, Liu, et al.
Veröffentlicht: (2025)
Efficient Numerical Integration in Reproducing Kernel Hilbert Spaces via Leverage Scores Sampling
von: Chatalic, Antoine, et al.
Veröffentlicht: (2023)
von: Chatalic, Antoine, et al.
Veröffentlicht: (2023)
Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias
von: Das, Mohua, et al.
Veröffentlicht: (2026)
von: Das, Mohua, et al.
Veröffentlicht: (2026)
Topological Neural Networks: Mitigating the Bottlenecks of Graph Neural Networks via Higher-Order Interactions
von: Giusti, Lorenzo
Veröffentlicht: (2024)
von: Giusti, Lorenzo
Veröffentlicht: (2024)
Key Design Choices in Source-Free Unsupervised Domain Adaptation: An In-depth Empirical Analysis
von: Maracani, Andrea, et al.
Veröffentlicht: (2024)
von: Maracani, Andrea, et al.
Veröffentlicht: (2024)
Implicit Bias of AdamW: $\ell_\infty$ Norm Constrained Optimization
von: Xie, Shuo, et al.
Veröffentlicht: (2024)
von: Xie, Shuo, et al.
Veröffentlicht: (2024)
Information Filtering Networks: Theoretical Foundations, Generative Methodologies, and Real-World Applications
von: Aste, Tomaso
Veröffentlicht: (2025)
von: Aste, Tomaso
Veröffentlicht: (2025)
Computational Efficiency under Covariate Shift in Kernel Ridge Regression
von: Della Vecchia, Andrea, et al.
Veröffentlicht: (2025)
von: Della Vecchia, Andrea, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Multi-Index Models with Hyper-Kernel Ridge Regression
von: Huang, Shuo, et al.
Veröffentlicht: (2025) -
Iterative regularization in classification via hinge loss diagonal descent
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022) -
Learning functions, operators and dynamical systems with kernels
von: Rosasco, Lorenzo
Veröffentlicht: (2025) -
Towards a Learning Theory of Representation Alignment
von: Insulla, Francesco, et al.
Veröffentlicht: (2025) -
On efficiently computable functions, deep networks and sparse compositionality
von: Poggio, Tomaso
Veröffentlicht: (2025)