On Generalization Bounds for Neural Networks with Low Rank Layers
Fuente:
arXiv
Saved in:
| Main Authors: | Pinto, Andrea, Rangamani, Akshay, Poggio, Tomaso |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Low Rank and Sparse Fourier Structure in Recurrent Networks Trained on Modular Addition
by: Rangamani, Akshay
Published: (2025)
by: Rangamani, Akshay
Published: (2025)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
by: Beneventano, Pierfrancesco, et al.
Published: (2024)
by: Beneventano, Pierfrancesco, et al.
Published: (2024)
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network
by: Galanti, Tomer, et al.
Published: (2022)
by: Galanti, Tomer, et al.
Published: (2022)
On efficiently computable functions, deep networks and sparse compositionality
by: Poggio, Tomaso
Published: (2025)
by: Poggio, Tomaso
Published: (2025)
Learning Sparse Compositional Functions with Norm-Constrained Neural Networks
by: Huang, Shuo, et al.
Published: (2026)
by: Huang, Shuo, et al.
Published: (2026)
Deep Neural Regression Collapse
by: Rangamani, Akshay, et al.
Published: (2026)
by: Rangamani, Akshay, et al.
Published: (2026)
Formation of Representations in Neural Networks
by: Ziyin, Liu, et al.
Published: (2024)
by: Ziyin, Liu, et al.
Published: (2024)
Ubiquity of Emergent Hebbian Dynamics in Regularized Learning
by: Koplow, David, et al.
Published: (2025)
by: Koplow, David, et al.
Published: (2025)
Hierarchical Reasoning Models: Perspectives and Misconceptions
by: Ge, Renee, et al.
Published: (2025)
by: Ge, Renee, et al.
Published: (2025)
Generalization Bounds for Rank-sparse Neural Networks
by: Ledent, Antoine, et al.
Published: (2025)
by: Ledent, Antoine, et al.
Published: (2025)
pAI/MSc: ML Theory Research with Humans on the Loop
by: Abdelmoneum, Mahmoud, et al.
Published: (2026)
by: Abdelmoneum, Mahmoud, et al.
Published: (2026)
Does Feedback Alignment Work at Biological Timescales?
by: Bacvanski, Marc Gong, et al.
Published: (2025)
by: Bacvanski, Marc Gong, et al.
Published: (2025)
Topological Invariance and Breakdown in Learning
by: Yang, Yongyi, et al.
Published: (2025)
by: Yang, Yongyi, et al.
Published: (2025)
Information Filtering Networks: Theoretical Foundations, Generative Methodologies, and Real-World Applications
by: Aste, Tomaso
Published: (2025)
by: Aste, Tomaso
Published: (2025)
An Illusion of Unlearning? Assessing Machine Unlearning Through Internal Representations
by: Gao, Yichen, et al.
Published: (2026)
by: Gao, Yichen, et al.
Published: (2026)
Does Weight Decay Enhance Training Stability?
by: Saether, Marius, et al.
Published: (2026)
by: Saether, Marius, et al.
Published: (2026)
Iterative regularization in classification via hinge loss diagonal descent
by: Apidopoulos, Vassilis, et al.
Published: (2022)
by: Apidopoulos, Vassilis, et al.
Published: (2022)
The Generalized Turing Test: A Foundation for Comparing Intelligence
by: Mitropolsky, Daniel, et al.
Published: (2026)
by: Mitropolsky, Daniel, et al.
Published: (2026)
Unraveling Syntax: How Language Models Learn Context-Free Grammars
by: Schulz, Laura Ying, et al.
Published: (2025)
by: Schulz, Laura Ying, et al.
Published: (2025)
Too Sharp, Too Sure: When Calibration Follows Curvature
by: Morosini, Alessandro, et al.
Published: (2026)
by: Morosini, Alessandro, et al.
Published: (2026)
Heterosynaptic Circuits Are Universal Gradient Machines
by: Ziyin, Liu, et al.
Published: (2025)
by: Ziyin, Liu, et al.
Published: (2025)
Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias
by: Das, Mohua, et al.
Published: (2026)
by: Das, Mohua, et al.
Published: (2026)
Learning Multi-Index Models with Hyper-Kernel Ridge Regression
by: Huang, Shuo, et al.
Published: (2025)
by: Huang, Shuo, et al.
Published: (2025)
Parameter Symmetry Potentially Unifies Deep Learning Theory
by: Ziyin, Liu, et al.
Published: (2025)
by: Ziyin, Liu, et al.
Published: (2025)
Position: A Theory of Deep Learning Must Include Compositional Sparsity
by: Danhofer, David A., et al.
Published: (2025)
by: Danhofer, David A., et al.
Published: (2025)
Norm-Bounded Low-Rank Adaptation
by: Wang, Ruigang, et al.
Published: (2025)
by: Wang, Ruigang, et al.
Published: (2025)
Same Error, Different Function: The Optimizer as an Implicit Prior in Financial Time Series
by: Cortesi, Federico Vittorio, et al.
Published: (2026)
by: Cortesi, Federico Vittorio, et al.
Published: (2026)
Momentum Further Constrains Sharpness at the Edge of Stochastic Stability
by: Andreyev, Arseniy, et al.
Published: (2026)
by: Andreyev, Arseniy, et al.
Published: (2026)
Low-Rank Prehab: Preparing Neural Networks for SVD Compression
by: Qin, Haoran, et al.
Published: (2025)
by: Qin, Haoran, et al.
Published: (2025)
Harnessing Orthogonality to Train Low-Rank Neural Networks
by: Coquelin, Daniel, et al.
Published: (2024)
by: Coquelin, Daniel, et al.
Published: (2024)
Low-Rank Tensor Decompositions for the Theory of Neural Networks
by: Borsoi, Ricardo, et al.
Published: (2025)
by: Borsoi, Ricardo, et al.
Published: (2025)
Low-Rank Matrix Approximation for Neural Network Compression
by: Cherukuri, Kalyan, et al.
Published: (2025)
by: Cherukuri, Kalyan, et al.
Published: (2025)
Personalizing Low-Rank Bayesian Neural Networks Via Federated Learning
by: Zhang, Boning, et al.
Published: (2024)
by: Zhang, Boning, et al.
Published: (2024)
Generalization and Risk Bounds for Recurrent Neural Networks
by: Cheng, Xuewei, et al.
Published: (2024)
by: Cheng, Xuewei, et al.
Published: (2024)
Theoretical Guarantees for Low-Rank Compression of Deep Neural Networks
by: Zhang, Shihao, et al.
Published: (2025)
by: Zhang, Shihao, et al.
Published: (2025)
Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks
by: Faller, Josua, et al.
Published: (2025)
by: Faller, Josua, et al.
Published: (2025)
Upper Bounds for Local Learning Coefficients of Three-Layer Neural Networks
by: Kurumadani, Yuki
Published: (2026)
by: Kurumadani, Yuki
Published: (2026)
Slicing Mutual Information Generalization Bounds for Neural Networks
by: Nadjahi, Kimia, et al.
Published: (2024)
by: Nadjahi, Kimia, et al.
Published: (2024)
Learning Neural Networks by Neuron Pursuit
by: Kumar, Akshay, et al.
Published: (2025)
by: Kumar, Akshay, et al.
Published: (2025)
Structure-Preserving Network Compression Via Low-Rank Induced Training Through Linear Layers Composition
by: Zhang, Xitong, et al.
Published: (2024)
by: Zhang, Xitong, et al.
Published: (2024)
Similar Items
-
Low Rank and Sparse Fourier Structure in Recurrent Networks Trained on Modular Addition
by: Rangamani, Akshay
Published: (2025) -
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
by: Beneventano, Pierfrancesco, et al.
Published: (2024) -
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network
by: Galanti, Tomer, et al.
Published: (2022) -
On efficiently computable functions, deep networks and sparse compositionality
by: Poggio, Tomaso
Published: (2025) -
Learning Sparse Compositional Functions with Norm-Constrained Neural Networks
by: Huang, Shuo, et al.
Published: (2026)