Robust Implicit Regularization via Weight Normalization
Fuente:
arXiv
Guardado en:
| Autores principales: | Chou, Hung-Hsu, Rauhut, Holger, Ward, Rachel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
More is Less: Inducing Sparsity via Overparameterization
por: Chou, Hung-Hsu, et al.
Publicado: (2021)
por: Chou, Hung-Hsu, et al.
Publicado: (2021)
Implicit Regularization in Perturbed Deep Matrix Factorization: Spectral Conditions and Stability
por: Wang, Jingzhe, et al.
Publicado: (2026)
por: Wang, Jingzhe, et al.
Publicado: (2026)
Convergence of gradient flow for learning convolutional neural networks
por: Diederen, Jona-Maria, et al.
Publicado: (2026)
por: Diederen, Jona-Maria, et al.
Publicado: (2026)
Implicit Regularization Makes Overparameterized Asymmetric Matrix Sensing Robust to Perturbations
por: Wind, Johan S.
Publicado: (2023)
por: Wind, Johan S.
Publicado: (2023)
Improving Generalization and Convergence by Enhancing Implicit Regularization
por: Wang, Mingze, et al.
Publicado: (2024)
por: Wang, Mingze, et al.
Publicado: (2024)
Implicit Differentiation for Hyperparameter Tuning the Weighted Graphical Lasso
por: Pouliquen, Can, et al.
Publicado: (2023)
por: Pouliquen, Can, et al.
Publicado: (2023)
Convergence of Alternating Gradient Descent for Matrix Factorization
por: Ward, Rachel, et al.
Publicado: (2023)
por: Ward, Rachel, et al.
Publicado: (2023)
Understanding the Implicit Regularization of Gradient Descent in Over-parameterized Models
por: Ma, Jianhao, et al.
Publicado: (2025)
por: Ma, Jianhao, et al.
Publicado: (2025)
On Generalization and Regularization via Wasserstein Distributionally Robust Optimization
por: Wu, Qinyu, et al.
Publicado: (2022)
por: Wu, Qinyu, et al.
Publicado: (2022)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
por: Beneventano, Pierfrancesco, et al.
Publicado: (2024)
por: Beneventano, Pierfrancesco, et al.
Publicado: (2024)
Implicit Regularization for Tubal Tensor Factorizations via Gradient Descent
por: Karnik, Santhosh, et al.
Publicado: (2024)
por: Karnik, Santhosh, et al.
Publicado: (2024)
Regularization for Adversarial Robust Learning
por: Wang, Jie, et al.
Publicado: (2024)
por: Wang, Jie, et al.
Publicado: (2024)
Safety Beyond the Training Data: Robust Out-of-Distribution MPC via Conformalized System Level Synthesis
por: Srinivasan, Anutam, et al.
Publicado: (2026)
por: Srinivasan, Anutam, et al.
Publicado: (2026)
Offline Policy Learning with Weight Clipping and Heaviside Composite Optimization
por: Liu, Jingren, et al.
Publicado: (2026)
por: Liu, Jingren, et al.
Publicado: (2026)
Regularized Q-learning through Robust Averaging
por: Schmitt-Förster, Peter, et al.
Publicado: (2024)
por: Schmitt-Förster, Peter, et al.
Publicado: (2024)
Optimization and Generalization Guarantees for Weight Normalization
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
On the Benefits of Weight Normalization for Overparameterized Matrix Sensing
por: Wei, Yudong, et al.
Publicado: (2025)
por: Wei, Yudong, et al.
Publicado: (2025)
Implicit Regularization of Gradient Flow on One-Layer Softmax Attention
por: Sheen, Heejune, et al.
Publicado: (2024)
por: Sheen, Heejune, et al.
Publicado: (2024)
Provable Acceleration of Nesterov's Accelerated Gradient for Rectangular Matrix Factorization and Linear Neural Networks
por: Xu, Zhenghao, et al.
Publicado: (2024)
por: Xu, Zhenghao, et al.
Publicado: (2024)
Nested Stochastic Algorithm for Generalized Sinkhorn distance-Regularized Distributionally Robust Optimization
por: Yang, Yufeng, et al.
Publicado: (2025)
por: Yang, Yufeng, et al.
Publicado: (2025)
Adaptive Optimization via Momentum on Variance-Normalized Gradients
por: Patitucci, Francisco, et al.
Publicado: (2026)
por: Patitucci, Francisco, et al.
Publicado: (2026)
Scalable Solution of the Stochastic Multi-path Traveling Salesman Problem via Neural Networks
por: Chou, Xiaochen, et al.
Publicado: (2026)
por: Chou, Xiaochen, et al.
Publicado: (2026)
Deceptive Sequential Decision-Making via Regularized Policy Optimization
por: Kim, Yerin, et al.
Publicado: (2025)
por: Kim, Yerin, et al.
Publicado: (2025)
A Unifying View of Anchoring via Operator-Side Tikhonov Regularization
por: Chen, Zihao
Publicado: (2026)
por: Chen, Zihao
Publicado: (2026)
Worth Their Weight: Randomized and Regularized Block Kaczmarz Algorithms without Preprocessing
por: Goldshlager, Gil, et al.
Publicado: (2025)
por: Goldshlager, Gil, et al.
Publicado: (2025)
End-to-End Training of High-Dimensional Optimal Control with Implicit Hamiltonians via Jacobian-Free Backpropagation
por: Gelphman, Eric, et al.
Publicado: (2025)
por: Gelphman, Eric, et al.
Publicado: (2025)
Formal Safety Verification and Refinement for Generative Motion Planners via Certified Local Stabilization
por: Nath, Devesh, et al.
Publicado: (2025)
por: Nath, Devesh, et al.
Publicado: (2025)
The Rich and the Simple: On the Implicit Bias of Adam and SGD
por: Vasudeva, Bhavya, et al.
Publicado: (2025)
por: Vasudeva, Bhavya, et al.
Publicado: (2025)
Implicit Bias of Mirror Flow on Separable Data
por: Pesme, Scott, et al.
Publicado: (2024)
por: Pesme, Scott, et al.
Publicado: (2024)
Cauchy-Schwarz Regularizers
por: Taner, Sueda, et al.
Publicado: (2025)
por: Taner, Sueda, et al.
Publicado: (2025)
Efficiently Escaping Saddle Points under Generalized Smoothness via Self-Bounding Regularity
por: Cao, Daniel Yiming, et al.
Publicado: (2025)
por: Cao, Daniel Yiming, et al.
Publicado: (2025)
Sparse Transformer Architectures via Regularized Wasserstein Proximal Operator with $L_1$ Prior
por: Han, Fuqun, et al.
Publicado: (2025)
por: Han, Fuqun, et al.
Publicado: (2025)
Adjusted Shuffling SARAH: Advancing Complexity Analysis via Dynamic Gradient Weighting
por: Nguyen, Duc Toan, et al.
Publicado: (2025)
por: Nguyen, Duc Toan, et al.
Publicado: (2025)
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
por: Akhtiamov, Danil, et al.
Publicado: (2026)
por: Akhtiamov, Danil, et al.
Publicado: (2026)
Implicit Riemannian Optimism with Applications to Min-Max Problems
por: Roux, Christophe, et al.
Publicado: (2025)
por: Roux, Christophe, et al.
Publicado: (2025)
Implicit Bias and Fast Convergence Rates for Self-attention
por: Vasudeva, Bhavya, et al.
Publicado: (2024)
por: Vasudeva, Bhavya, et al.
Publicado: (2024)
Nonsmooth Implicit Differentiation: Deterministic and Stochastic Convergence Rates
por: Grazzi, Riccardo, et al.
Publicado: (2024)
por: Grazzi, Riccardo, et al.
Publicado: (2024)
Stability Regularized Cross-Validation
por: Cory-Wright, Ryan, et al.
Publicado: (2025)
por: Cory-Wright, Ryan, et al.
Publicado: (2025)
Cautious Weight Decay
por: Chen, Lizhang, et al.
Publicado: (2025)
por: Chen, Lizhang, et al.
Publicado: (2025)
Robust Gaussian Processes via Relevance Pursuit
por: Ament, Sebastian, et al.
Publicado: (2024)
por: Ament, Sebastian, et al.
Publicado: (2024)
Ejemplares similares
-
More is Less: Inducing Sparsity via Overparameterization
por: Chou, Hung-Hsu, et al.
Publicado: (2021) -
Implicit Regularization in Perturbed Deep Matrix Factorization: Spectral Conditions and Stability
por: Wang, Jingzhe, et al.
Publicado: (2026) -
Convergence of gradient flow for learning convolutional neural networks
por: Diederen, Jona-Maria, et al.
Publicado: (2026) -
Implicit Regularization Makes Overparameterized Asymmetric Matrix Sensing Robust to Perturbations
por: Wind, Johan S.
Publicado: (2023) -
Improving Generalization and Convergence by Enhancing Implicit Regularization
por: Wang, Mingze, et al.
Publicado: (2024)