Provable Guarantees for Nonlinear Feature Learning in Three-Layer Neural Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Nichani, Eshaan, Damian, Alex, Lee, Jason D. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning Hierarchical Polynomials of Multiple Nonlinear Features with Three-Layer Networks
por: Fu, Hengyu, et al.
Publicado: (2024)
por: Fu, Hengyu, et al.
Publicado: (2024)
How Transformers Learn Causal Structure with Gradient Descent
por: Nichani, Eshaan, et al.
Publicado: (2024)
por: Nichani, Eshaan, et al.
Publicado: (2024)
Quantitative Bounds for Length Generalization in Transformers
por: Izzo, Zachary, et al.
Publicado: (2025)
por: Izzo, Zachary, et al.
Publicado: (2025)
On the Statistical Query Complexity of Learning Semiautomata: a Random Walk Approach
por: Giapitzakis, George, et al.
Publicado: (2025)
por: Giapitzakis, George, et al.
Publicado: (2025)
Understanding Factual Recall in Transformers via Associative Memories
por: Nichani, Eshaan, et al.
Publicado: (2024)
por: Nichani, Eshaan, et al.
Publicado: (2024)
Learning Compositional Functions with Transformers from Easy-to-Hard Data
por: Wang, Zixuan, et al.
Publicado: (2025)
por: Wang, Zixuan, et al.
Publicado: (2025)
Emergence and scaling laws in SGD learning of shallow neural networks
por: Ren, Yunwei, et al.
Publicado: (2025)
por: Ren, Yunwei, et al.
Publicado: (2025)
Fine-Tuning Dynamics of In-Context Factual Recall in Transformers
por: Huang, Ruomin, et al.
Publicado: (2026)
por: Huang, Ruomin, et al.
Publicado: (2026)
Sharp Capacity Scaling of Spectral Optimizers in Learning Associative Memory
por: Kim, Juno, et al.
Publicado: (2026)
por: Kim, Juno, et al.
Publicado: (2026)
Fine-Tuning Language Models with Just Forward Passes
por: Malladi, Sadhika, et al.
Publicado: (2023)
por: Malladi, Sadhika, et al.
Publicado: (2023)
Sharp Capacity Thresholds in Linear Associative Memory: From Winner-Take-All to Listwise Retrieval
por: Barnfield, Nicholas, et al.
Publicado: (2026)
por: Barnfield, Nicholas, et al.
Publicado: (2026)
The Generative Leap: Sharp Sample Complexity for Efficiently Learning Gaussian Multi-Index Models
por: Damian, Alex, et al.
Publicado: (2025)
por: Damian, Alex, et al.
Publicado: (2025)
Enumerating Safe Regions in Deep Neural Networks with Provable Probabilistic Guarantees
por: Marzari, Luca, et al.
Publicado: (2023)
por: Marzari, Luca, et al.
Publicado: (2023)
Provable Multi-Task Representation Learning by Two-Layer ReLU Neural Networks
por: Collins, Liam, et al.
Publicado: (2023)
por: Collins, Liam, et al.
Publicado: (2023)
Implicit Hypergraph Neural Networks: A Stable Framework for Higher-Order Relational Learning with Provable Guarantees
por: Li, Xiaoyu, et al.
Publicado: (2025)
por: Li, Xiaoyu, et al.
Publicado: (2025)
On The Concurrence of Layer-wise Preconditioning Methods and Provable Feature Learning
por: Zhang, Thomas T., et al.
Publicado: (2025)
por: Zhang, Thomas T., et al.
Publicado: (2025)
From Linear to Nonlinear: Provable Weak-to-Strong Generalization through Feature Learning
por: Oh, Junsoo, et al.
Publicado: (2025)
por: Oh, Junsoo, et al.
Publicado: (2025)
Improved high-dimensional estimation with Langevin dynamics and stochastic weight averaging
por: Wei, Stanley, et al.
Publicado: (2026)
por: Wei, Stanley, et al.
Publicado: (2026)
Generalization Guarantees of Gradient Descent for Multi-Layer Neural Networks
por: Wang, Puyu, et al.
Publicado: (2023)
por: Wang, Puyu, et al.
Publicado: (2023)
Over-parameterised Shallow Neural Networks with Asymmetrical Node Scaling: Global Convergence Guarantees and Feature Learning
por: Caron, Francois, et al.
Publicado: (2023)
por: Caron, Francois, et al.
Publicado: (2023)
Provable Learning of Random Hierarchy Models and Hierarchical Shallow-to-Deep Chaining
por: Ren, Yunwei, et al.
Publicado: (2026)
por: Ren, Yunwei, et al.
Publicado: (2026)
Provably-Stable Neural Network-Based Control of Nonlinear Systems
por: Li, Anran, et al.
Publicado: (2025)
por: Li, Anran, et al.
Publicado: (2025)
Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees
por: Chen, Yilei, et al.
Publicado: (2024)
por: Chen, Yilei, et al.
Publicado: (2024)
Provable In-Context Learning of Nonlinear Regression with Transformers
por: Li, Hongbo, et al.
Publicado: (2025)
por: Li, Hongbo, et al.
Publicado: (2025)
A Theory of Non-Linear Feature Learning with One Gradient Step in Two-Layer Neural Networks
por: Moniri, Behrad, et al.
Publicado: (2023)
por: Moniri, Behrad, et al.
Publicado: (2023)
Active Learning for Decision Trees with Provable Guarantees
por: Moakhar, Arshia Soltani, et al.
Publicado: (2026)
por: Moakhar, Arshia Soltani, et al.
Publicado: (2026)
Deciphering Raw Data in Neuro-Symbolic Learning with Provable Guarantees
por: Tao, Lue, et al.
Publicado: (2023)
por: Tao, Lue, et al.
Publicado: (2023)
Upper Bounds for Local Learning Coefficients of Three-Layer Neural Networks
por: Kurumadani, Yuki
Publicado: (2026)
por: Kurumadani, Yuki
Publicado: (2026)
Provable Reward-Agnostic Preference-Based Reinforcement Learning
por: Zhan, Wenhao, et al.
Publicado: (2023)
por: Zhan, Wenhao, et al.
Publicado: (2023)
Learning Neural Networks with Distribution Shift: Efficiently Certifiable Guarantees
por: Chandrasekaran, Gautam, et al.
Publicado: (2025)
por: Chandrasekaran, Gautam, et al.
Publicado: (2025)
A Guaranteed-Stable Neural Network Approach for Optimal Control of Nonlinear Systems
por: Li, Anran, et al.
Publicado: (2025)
por: Li, Anran, et al.
Publicado: (2025)
A Nonparametric Statistics Approach to Feature Selection in Deep Neural Networks with Theoretical Guarantees
por: Du, Junye, et al.
Publicado: (2025)
por: Du, Junye, et al.
Publicado: (2025)
Tuning Algorithmic and Architectural Hyperparameters in Graph-Based Semi-Supervised Learning with Provable Guarantees
por: Du, Ally Yalei, et al.
Publicado: (2025)
por: Du, Ally Yalei, et al.
Publicado: (2025)
PinnDE: Physics-Informed Neural Networks for Solving Differential Equations
por: Matthews, Jason, et al.
Publicado: (2024)
por: Matthews, Jason, et al.
Publicado: (2024)
Computational-Statistical Gaps in Gaussian Single-Index Models
por: Damian, Alex, et al.
Publicado: (2024)
por: Damian, Alex, et al.
Publicado: (2024)
Transformers Provably Learn Sparse Token Selection While Fully-Connected Nets Cannot
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Flow-Controlled Scheduling for LLM Inference with Provable Stability Guarantees
por: Dong, Zhuolun, et al.
Publicado: (2026)
por: Dong, Zhuolun, et al.
Publicado: (2026)
Non-Stationary Restless Multi-Armed Bandits with Provable Guarantee
por: Hung, Yu-Heng, et al.
Publicado: (2025)
por: Hung, Yu-Heng, et al.
Publicado: (2025)
Learning Guarantee of Reward Modeling Using Deep Neural Networks
por: Luo, Yuanhang, et al.
Publicado: (2025)
por: Luo, Yuanhang, et al.
Publicado: (2025)
Weak-to-Strong Generalization Even in Random Feature Networks, Provably
por: Medvedev, Marko, et al.
Publicado: (2025)
por: Medvedev, Marko, et al.
Publicado: (2025)
Ejemplares similares
-
Learning Hierarchical Polynomials of Multiple Nonlinear Features with Three-Layer Networks
por: Fu, Hengyu, et al.
Publicado: (2024) -
How Transformers Learn Causal Structure with Gradient Descent
por: Nichani, Eshaan, et al.
Publicado: (2024) -
Quantitative Bounds for Length Generalization in Transformers
por: Izzo, Zachary, et al.
Publicado: (2025) -
On the Statistical Query Complexity of Learning Semiautomata: a Random Walk Approach
por: Giapitzakis, George, et al.
Publicado: (2025) -
Understanding Factual Recall in Transformers via Associative Memories
por: Nichani, Eshaan, et al.
Publicado: (2024)