Shallow Neural Networks Learn Low-Degree Spherical Polynomials with Feature Learning by Learnable Channel Attention
Fuente:
arXiv
Guardado en:
| Autor principal: | Yang, Yingzhen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Over-parameterised Shallow Neural Networks with Asymmetrical Node Scaling: Global Convergence Guarantees and Feature Learning
por: Caron, Francois, et al.
Publicado: (2023)
por: Caron, Francois, et al.
Publicado: (2023)
Wasserstein Distributionally Robust Shallow Convex Neural Networks
por: Pallage, Julien, et al.
Publicado: (2024)
por: Pallage, Julien, et al.
Publicado: (2024)
Projective Proximal Gradient Descent for A Class of Nonconvex Nonsmooth Optimization Problems: Fast Convergence Without Kurdyka-Lojasiewicz (KL) Property
por: Yang, Yingzhen, et al.
Publicado: (2023)
por: Yang, Yingzhen, et al.
Publicado: (2023)
Locally Regularized Sparse Graph by Fast Proximal Gradient Descent
por: Sun, Dongfang, et al.
Publicado: (2024)
por: Sun, Dongfang, et al.
Publicado: (2024)
Verifying Properties of Binary Neural Networks Using Sparse Polynomial Optimization
por: Yang, Jianting, et al.
Publicado: (2024)
por: Yang, Jianting, et al.
Publicado: (2024)
Neural Collapse under Gradient Flow on Shallow ReLU Networks for Orthogonally Separable Data
por: Min, Hancheng, et al.
Publicado: (2025)
por: Min, Hancheng, et al.
Publicado: (2025)
Automated Proof of Polynomial Inequalities via Reinforcement Learning
por: Liu, Banglong, et al.
Publicado: (2025)
por: Liu, Banglong, et al.
Publicado: (2025)
Learning Neural Networks by Neuron Pursuit
por: Kumar, Akshay, et al.
Publicado: (2025)
por: Kumar, Akshay, et al.
Publicado: (2025)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
por: Kim, Sungyoon, et al.
Publicado: (2024)
por: Kim, Sungyoon, et al.
Publicado: (2024)
Fourier Learning Machines: Nonharmonic Fourier-Based Neural Networks for Scientific Machine Learning
por: Rubel, Mominul, et al.
Publicado: (2025)
por: Rubel, Mominul, et al.
Publicado: (2025)
Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU
por: Liu, Jingzhou
Publicado: (2025)
por: Liu, Jingzhou
Publicado: (2025)
Convergence Analysis for Learning Orthonormal Deep Linear Neural Networks
por: Qin, Zhen, et al.
Publicado: (2023)
por: Qin, Zhen, et al.
Publicado: (2023)
Approximation with Random Shallow ReLU Networks with Applications to Model Reference Adaptive Control
por: Lamperski, Andrew, et al.
Publicado: (2024)
por: Lamperski, Andrew, et al.
Publicado: (2024)
Global Convergence and Rich Feature Learning in $L$-Layer Infinite-Width Neural Networks under $μ$P Parametrization
por: Chen, Zixiang, et al.
Publicado: (2025)
por: Chen, Zixiang, et al.
Publicado: (2025)
Gradient Descent Converges Linearly to Flatter Minima than Gradient Flow in Shallow Linear Networks
por: Beneventano, Pierfrancesco, et al.
Publicado: (2025)
por: Beneventano, Pierfrancesco, et al.
Publicado: (2025)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
por: Beneventano, Pierfrancesco, et al.
Publicado: (2024)
por: Beneventano, Pierfrancesco, et al.
Publicado: (2024)
Learning of Linear Dynamical Systems as a Non-Commutative Polynomial Optimization Problem
por: Zhou, Quan, et al.
Publicado: (2020)
por: Zhou, Quan, et al.
Publicado: (2020)
Incremental Learning of Sparse Attention Patterns in Transformers
por: Yüksel, Oğuz Kaan, et al.
Publicado: (2026)
por: Yüksel, Oğuz Kaan, et al.
Publicado: (2026)
Distributed Learning over Arbitrary Topology: Linear Speed-Up with Polynomial Transient Time
por: You, Runze, et al.
Publicado: (2025)
por: You, Runze, et al.
Publicado: (2025)
Active Learning of Deep Neural Networks via Gradient-Free Cutting Planes
por: Zhang, Erica, et al.
Publicado: (2024)
por: Zhang, Erica, et al.
Publicado: (2024)
CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural Networks
por: Feng, Miria, et al.
Publicado: (2024)
por: Feng, Miria, et al.
Publicado: (2024)
Physics-Informed Neural Network Lyapunov Functions: PDE Characterization, Learning, and Verification
por: Liu, Jun, et al.
Publicado: (2023)
por: Liu, Jun, et al.
Publicado: (2023)
A Compositional Kernel Model for Feature Learning
por: Ruan, Feng, et al.
Publicado: (2025)
por: Ruan, Feng, et al.
Publicado: (2025)
Curse of Dimensionality in Neural Network Optimization
por: Na, Sanghoon, et al.
Publicado: (2025)
por: Na, Sanghoon, et al.
Publicado: (2025)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks (YANNs)
por: Braniff, Austin, et al.
Publicado: (2025)
por: Braniff, Austin, et al.
Publicado: (2025)
Learn2Aggregate: Supervised Generation of Chvátal-Gomory Cuts Using Graph Neural Networks
por: Deza, Arnaud, et al.
Publicado: (2024)
por: Deza, Arnaud, et al.
Publicado: (2024)
Learning-Augmented Decentralized Online Convex Optimization in Networks
por: Li, Pengfei, et al.
Publicado: (2023)
por: Li, Pengfei, et al.
Publicado: (2023)
Symmetric Linear Dynamical Systems are Learnable from Few Observations
por: Vu, Minh, et al.
Publicado: (2025)
por: Vu, Minh, et al.
Publicado: (2025)
A Nonlinear Separation Principle via Contraction Theory: Applications to Neural Networks, Control, and Learning
por: Gokhale, Anand, et al.
Publicado: (2026)
por: Gokhale, Anand, et al.
Publicado: (2026)
On The Concurrence of Layer-wise Preconditioning Methods and Provable Feature Learning
por: Zhang, Thomas T., et al.
Publicado: (2025)
por: Zhang, Thomas T., et al.
Publicado: (2025)
Learning Mixtures of Spherical Gaussians via Fourier Analysis
por: Chakraborty, Somnath, et al.
Publicado: (2020)
por: Chakraborty, Somnath, et al.
Publicado: (2020)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
por: Ye, Lintao, et al.
Publicado: (2024)
por: Ye, Lintao, et al.
Publicado: (2024)
Provable Exactness for Asymmetric Low-Rank SDP Learning
por: Hu, Enliang
Publicado: (2018)
por: Hu, Enliang
Publicado: (2018)
PAC Learnability of Scenario Decision-Making Algorithms: Necessary Conditions and Sufficient Conditions
por: Berger, Guillaume O., et al.
Publicado: (2025)
por: Berger, Guillaume O., et al.
Publicado: (2025)
Zero-Order Optimization for LLM Fine-Tuning via Learnable Direction Sampling
por: Parfenov, Valery, et al.
Publicado: (2026)
por: Parfenov, Valery, et al.
Publicado: (2026)
Inexact Column Generation for Bayesian Network Structure Learning via Difference-of-Submodular Optimization
por: Yang, Yiran, et al.
Publicado: (2025)
por: Yang, Yiran, et al.
Publicado: (2025)
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
por: Lamperski, Andrew, et al.
Publicado: (2024)
por: Lamperski, Andrew, et al.
Publicado: (2024)
Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons
por: Houédry, Pierre, et al.
Publicado: (2026)
por: Houédry, Pierre, et al.
Publicado: (2026)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
por: Kuelbs, Daniel, et al.
Publicado: (2024)
por: Kuelbs, Daniel, et al.
Publicado: (2024)
A Theory of Feature Learning in Kernel Models
por: Chen, Yunlu, et al.
Publicado: (2023)
por: Chen, Yunlu, et al.
Publicado: (2023)
Ejemplares similares
-
Over-parameterised Shallow Neural Networks with Asymmetrical Node Scaling: Global Convergence Guarantees and Feature Learning
por: Caron, Francois, et al.
Publicado: (2023) -
Wasserstein Distributionally Robust Shallow Convex Neural Networks
por: Pallage, Julien, et al.
Publicado: (2024) -
Projective Proximal Gradient Descent for A Class of Nonconvex Nonsmooth Optimization Problems: Fast Convergence Without Kurdyka-Lojasiewicz (KL) Property
por: Yang, Yingzhen, et al.
Publicado: (2023) -
Locally Regularized Sparse Graph by Fast Proximal Gradient Descent
por: Sun, Dongfang, et al.
Publicado: (2024) -
Verifying Properties of Binary Neural Networks Using Sparse Polynomial Optimization
por: Yang, Jianting, et al.
Publicado: (2024)