Learning with Norm Constrained, Over-parameterized, Two-layer Neural Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Fanghui, Dadi, Leello, Cevher, Volkan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Efficient Continual Finite-Sum Minimization
por: Mavrothalassitis, Ioannis, et al.
Publicado: (2024)
por: Mavrothalassitis, Ioannis, et al.
Publicado: (2024)
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
por: Davtyan, Aram, et al.
Publicado: (2026)
por: Davtyan, Aram, et al.
Publicado: (2026)
High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit Regularization
por: Chen, Yihang, et al.
Publicado: (2024)
por: Chen, Yihang, et al.
Publicado: (2024)
Training Deep Learning Models with Norm-Constrained LMOs
por: Pethick, Thomas, et al.
Publicado: (2025)
por: Pethick, Thomas, et al.
Publicado: (2025)
Generalization of Scaled Deep ResNets in the Mean-Field Regime
por: Chen, Yihang, et al.
Publicado: (2024)
por: Chen, Yihang, et al.
Publicado: (2024)
Polynomial Convergence of Bandit No-Regret Dynamics in Congestion Games
por: Dadi, Leello, et al.
Publicado: (2024)
por: Dadi, Leello, et al.
Publicado: (2024)
Truly No-Regret Learning in Constrained MDPs
por: Müller, Adrian, et al.
Publicado: (2024)
por: Müller, Adrian, et al.
Publicado: (2024)
Revisiting Character-level Adversarial Attacks for Language Models
por: Rocamora, Elias Abad, et al.
Publicado: (2024)
por: Rocamora, Elias Abad, et al.
Publicado: (2024)
Robust NAS under adversarial training: benchmark, theory, and beyond
por: Wu, Yongtao, et al.
Publicado: (2024)
por: Wu, Yongtao, et al.
Publicado: (2024)
Efficient Large Language Model Inference with Neural Block Linearization
por: Erdogan, Mete, et al.
Publicado: (2025)
por: Erdogan, Mete, et al.
Publicado: (2025)
Imitation Learning in Discounted Linear MDPs without exploration assumptions
por: Viano, Luca, et al.
Publicado: (2024)
por: Viano, Luca, et al.
Publicado: (2024)
IL-SOAR : Imitation Learning with Soft Optimistic Actor cRitic
por: Viel, Stefano, et al.
Publicado: (2025)
por: Viel, Stefano, et al.
Publicado: (2025)
Training Neural Networks at Any Scale
por: Pethick, Thomas, et al.
Publicado: (2025)
por: Pethick, Thomas, et al.
Publicado: (2025)
Efficient local linearity regularization to overcome catastrophic overfitting
por: Rocamora, Elias Abad, et al.
Publicado: (2024)
por: Rocamora, Elias Abad, et al.
Publicado: (2024)
Generalized Gradient Norm Clipping & Non-Euclidean $(L_0,L_1)$-Smoothness
por: Pethick, Thomas, et al.
Publicado: (2025)
por: Pethick, Thomas, et al.
Publicado: (2025)
Learning Sparse Compositional Functions with Norm-Constrained Neural Networks
por: Huang, Shuo, et al.
Publicado: (2026)
por: Huang, Shuo, et al.
Publicado: (2026)
On the Convergence of Federated Averaging under Partial Participation for Over-parameterized Neural Networks
por: Liu, Xin, et al.
Publicado: (2023)
por: Liu, Xin, et al.
Publicado: (2023)
SAMPa: Sharpness-aware Minimization Parallelized
por: Xie, Wanyun, et al.
Publicado: (2024)
por: Xie, Wanyun, et al.
Publicado: (2024)
MaD-Mix: Multi-Modal Data Mixtures via Latent Space Coupling for Vision-Language Model Training
por: Xie, Wanyun, et al.
Publicado: (2026)
por: Xie, Wanyun, et al.
Publicado: (2026)
Depth-induced NTK: Bridging Over-parameterized Neural Networks and Deep Neural Kernels
por: Tian, Yong-Ming, et al.
Publicado: (2025)
por: Tian, Yong-Ming, et al.
Publicado: (2025)
Matching Multiple Experts: On the Exploitability of Multi-Agent Imitation Learning
por: Bergerault, Antoine, et al.
Publicado: (2026)
por: Bergerault, Antoine, et al.
Publicado: (2026)
Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning
por: Xie, Wanyun, et al.
Publicado: (2025)
por: Xie, Wanyun, et al.
Publicado: (2025)
Stable Nonconvex-Nonconcave Training via Linear Interpolation
por: Pethick, Thomas, et al.
Publicado: (2023)
por: Pethick, Thomas, et al.
Publicado: (2023)
Learning Equilibria from Data: Provably Efficient Multi-Agent Imitation Learning
por: Freihaut, Till, et al.
Publicado: (2025)
por: Freihaut, Till, et al.
Publicado: (2025)
Convergence Analysis of Natural Gradient Descent for Over-parameterized Physics-Informed Neural Networks
por: Xu, Xianliang, et al.
Publicado: (2024)
por: Xu, Xianliang, et al.
Publicado: (2024)
Constrained Stochastic Spectral Preconditioning Converges for Nonconvex Objectives
por: Oikonomidis, Konstantinos, et al.
Publicado: (2026)
por: Oikonomidis, Konstantinos, et al.
Publicado: (2026)
Multilinear Operator Networks
por: Cheng, Yixin, et al.
Publicado: (2024)
por: Cheng, Yixin, et al.
Publicado: (2024)
MT-NAM: An Efficient and Adaptive Model for Epileptic Seizure Detection
por: Afzal, Arshia, et al.
Publicado: (2025)
por: Afzal, Arshia, et al.
Publicado: (2025)
Over-parameterization and Adversarial Robustness in Neural Networks: An Overview and Empirical Analysis
por: Gupta, Srishti, et al.
Publicado: (2024)
por: Gupta, Srishti, et al.
Publicado: (2024)
Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation
por: Sheebaelhamd, Ziyad, et al.
Publicado: (2026)
por: Sheebaelhamd, Ziyad, et al.
Publicado: (2026)
Optimistic Dual Averaging Unifies Modern Optimizers
por: Pethick, Thomas, et al.
Publicado: (2026)
por: Pethick, Thomas, et al.
Publicado: (2026)
Provably avoiding over-optimization in Direct Preference Optimization without knowing the data distribution
por: Barla, Adam, et al.
Publicado: (2026)
por: Barla, Adam, et al.
Publicado: (2026)
Easy Data Unlearning Bench
por: Rinberg, Roy, et al.
Publicado: (2026)
por: Rinberg, Roy, et al.
Publicado: (2026)
Adversarial Training for Defense Against Label Poisoning Attacks
por: Bal, Melis Ilayda, et al.
Publicado: (2025)
por: Bal, Melis Ilayda, et al.
Publicado: (2025)
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
por: Bal, Melis Ilayda, et al.
Publicado: (2025)
por: Bal, Melis Ilayda, et al.
Publicado: (2025)
The $φ$ Curve: The Shape of Generalization through the Lens of Norm-based Capacity Control
por: Wang, Yichen, et al.
Publicado: (2025)
por: Wang, Yichen, et al.
Publicado: (2025)
μP$^2$: Effective Sharpness Aware Minimization Requires Layerwise Perturbation Scaling
por: Haas, Moritz, et al.
Publicado: (2024)
por: Haas, Moritz, et al.
Publicado: (2024)
A Novel Spatiotemporal Coupling Graph Convolutional Network
por: Bi, Fanghui
Publicado: (2024)
por: Bi, Fanghui
Publicado: (2024)
Learning to Remove Cuts in Integer Linear Programming
por: Puigdemont, Pol, et al.
Publicado: (2024)
por: Puigdemont, Pol, et al.
Publicado: (2024)
Physics-Informed Design of Input Convex Neural Networks for Consistency Optimal Transport Flow Matching
por: Song, Fanghui, et al.
Publicado: (2025)
por: Song, Fanghui, et al.
Publicado: (2025)
Ejemplares similares
-
Efficient Continual Finite-Sum Minimization
por: Mavrothalassitis, Ioannis, et al.
Publicado: (2024) -
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
por: Davtyan, Aram, et al.
Publicado: (2026) -
High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit Regularization
por: Chen, Yihang, et al.
Publicado: (2024) -
Training Deep Learning Models with Norm-Constrained LMOs
por: Pethick, Thomas, et al.
Publicado: (2025) -
Generalization of Scaled Deep ResNets in the Mean-Field Regime
por: Chen, Yihang, et al.
Publicado: (2024)