Optimization for Neural Operators can Benefit from Width
Fuente:
arXiv
Salvato in:
| Autori principali: | Cisneros-Velarde, Pedro, Shrimali, Bhavesh, Banerjee, Arindam |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Optimization and Generalization Guarantees for Weight Normalization
di: Cisneros-Velarde, Pedro, et al.
Pubblicazione: (2024)
di: Cisneros-Velarde, Pedro, et al.
Pubblicazione: (2024)
On the Width Scaling of Neural Optimizers Under Matrix Operator Norms I: Row/Column Normalization and Hyperparameter Transfer
di: Xu, Ruihan, et al.
Pubblicazione: (2026)
di: Xu, Ruihan, et al.
Pubblicazione: (2026)
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
di: Yang, Yahong, et al.
Pubblicazione: (2023)
di: Yang, Yahong, et al.
Pubblicazione: (2023)
Fundamental Benefit of Alternating Updates in Minimax Optimization
di: Lee, Jaewook, et al.
Pubblicazione: (2024)
di: Lee, Jaewook, et al.
Pubblicazione: (2024)
Gearing Gaussian process modeling and sequential design towards stochastic simulators
di: Binois, Mickael, et al.
Pubblicazione: (2024)
di: Binois, Mickael, et al.
Pubblicazione: (2024)
Deep Operator Neural Network Model Predictive Control
di: de Jong, Thomas Oliver, et al.
Pubblicazione: (2025)
di: de Jong, Thomas Oliver, et al.
Pubblicazione: (2025)
Why Line Search when you can Plane Search? SO-Friendly Neural Networks allow Per-Iteration Optimization of Learning and Momentum Rates for Every Layer
di: Shea, Betty, et al.
Pubblicazione: (2024)
di: Shea, Betty, et al.
Pubblicazione: (2024)
Lotka-Sharpe Neural Operators for Control of Population PDEs
di: Krstic, Miroslav, et al.
Pubblicazione: (2026)
di: Krstic, Miroslav, et al.
Pubblicazione: (2026)
Closed-Loop Neural Operator-Based Observer of Traffic Density
di: Harting, Alice, et al.
Pubblicazione: (2025)
di: Harting, Alice, et al.
Pubblicazione: (2025)
RHYME-XT: A Neural Operator for Spatiotemporal Control Systems
di: Ruiter, Marijn, et al.
Pubblicazione: (2026)
di: Ruiter, Marijn, et al.
Pubblicazione: (2026)
Curse of Dimensionality in Neural Network Optimization
di: Na, Sanghoon, et al.
Pubblicazione: (2025)
di: Na, Sanghoon, et al.
Pubblicazione: (2025)
Attention is All You Need to Optimize Wind Farm Operations and Maintenance
di: Kazemian, Iman, et al.
Pubblicazione: (2024)
di: Kazemian, Iman, et al.
Pubblicazione: (2024)
Shape Derivative-Informed Neural Operators with Application to Risk-Averse Shape Optimization
di: Gong, Xindi, et al.
Pubblicazione: (2026)
di: Gong, Xindi, et al.
Pubblicazione: (2026)
Sampling-Horizon Neural Operator Predictors for Nonlinear Control under Delayed Inputs
di: Bhan, Luke, et al.
Pubblicazione: (2026)
di: Bhan, Luke, et al.
Pubblicazione: (2026)
Global Convergence and Rich Feature Learning in $L$-Layer Infinite-Width Neural Networks under $μ$P Parametrization
di: Chen, Zixiang, et al.
Pubblicazione: (2025)
di: Chen, Zixiang, et al.
Pubblicazione: (2025)
Momentum Benefits Non-IID Federated Learning Simply and Provably
di: Cheng, Ziheng, et al.
Pubblicazione: (2023)
di: Cheng, Ziheng, et al.
Pubblicazione: (2023)
FedSEA: Achieving Benefit of Parallelization in Federated Online Learning
di: Sahu, Harekrushna, et al.
Pubblicazione: (2026)
di: Sahu, Harekrushna, et al.
Pubblicazione: (2026)
Exploring the Potential of Bilevel Optimization for Calibrating Neural Networks
di: Sanguin, Gabriele, et al.
Pubblicazione: (2025)
di: Sanguin, Gabriele, et al.
Pubblicazione: (2025)
SGD with Partial Hessian for Deep Neural Networks Optimization
di: Sun, Ying, et al.
Pubblicazione: (2024)
di: Sun, Ying, et al.
Pubblicazione: (2024)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
di: Adeoye, Adeyemi D., et al.
Pubblicazione: (2024)
di: Adeoye, Adeyemi D., et al.
Pubblicazione: (2024)
Learning An Interpretable Risk Scoring System for Maximizing Decision Net Benefit
di: Chi, Wenhao, et al.
Pubblicazione: (2026)
di: Chi, Wenhao, et al.
Pubblicazione: (2026)
Optimality-Informed Neural Networks for Solving Parametric Optimization Problems
di: Hoffmann, Matthias K., et al.
Pubblicazione: (2025)
di: Hoffmann, Matthias K., et al.
Pubblicazione: (2025)
FSNet: Feasibility-Seeking Neural Network for Constrained Optimization with Guarantees
di: Nguyen, Hoang T., et al.
Pubblicazione: (2025)
di: Nguyen, Hoang T., et al.
Pubblicazione: (2025)
Training Neural ODEs Using Fully Discretized Simultaneous Optimization
di: Shapovalova, Mariia, et al.
Pubblicazione: (2025)
di: Shapovalova, Mariia, et al.
Pubblicazione: (2025)
Differentiable Convex Optimization Layers in Neural Architectures: Foundations and Perspectives
di: Katyal, Calder
Pubblicazione: (2024)
di: Katyal, Calder
Pubblicazione: (2024)
High-Dimensional Analysis of Gradient Flow for Extensive-Width Quadratic Neural Networks
di: Martin, Simon, et al.
Pubblicazione: (2026)
di: Martin, Simon, et al.
Pubblicazione: (2026)
Inverse Optimization for Routing Problems
di: Scroccaro, Pedro Zattoni, et al.
Pubblicazione: (2023)
di: Scroccaro, Pedro Zattoni, et al.
Pubblicazione: (2023)
Towards Guided Descent: Optimization Algorithms for Training Neural Networks At Scale
di: Nagwekar, Ansh
Pubblicazione: (2025)
di: Nagwekar, Ansh
Pubblicazione: (2025)
Scalable Mixed-Integer Optimization with Neural Constraints via Dual Decomposition
di: Zeng, Shuli, et al.
Pubblicazione: (2025)
di: Zeng, Shuli, et al.
Pubblicazione: (2025)
Optimization Over Trained Neural Networks: Taking a Relaxing Walk
di: Tong, Jiatai, et al.
Pubblicazione: (2024)
di: Tong, Jiatai, et al.
Pubblicazione: (2024)
Verifying Properties of Binary Neural Networks Using Sparse Polynomial Optimization
di: Yang, Jianting, et al.
Pubblicazione: (2024)
di: Yang, Jianting, et al.
Pubblicazione: (2024)
Neural Network Training Techniques Regularize Optimization Trajectory: An Empirical Study
di: Chen, Cheng, et al.
Pubblicazione: (2020)
di: Chen, Cheng, et al.
Pubblicazione: (2020)
Scalable Kernel Inverse Optimization
di: Long, Youyuan, et al.
Pubblicazione: (2024)
di: Long, Youyuan, et al.
Pubblicazione: (2024)
Less is More: Convergence Benefits of Fewer Data Weight Updates over Longer Horizon
di: Das, Rudrajit, et al.
Pubblicazione: (2026)
di: Das, Rudrajit, et al.
Pubblicazione: (2026)
Bayesian Optimization of a Lightweight and Accurate Neural Network for Aerodynamic Performance Prediction
di: Shihua, James M., et al.
Pubblicazione: (2025)
di: Shihua, James M., et al.
Pubblicazione: (2025)
Bayesian Optimization with Preference Exploration using a Monotonic Neural Network Ensemble
di: Wang, Hanyang, et al.
Pubblicazione: (2025)
di: Wang, Hanyang, et al.
Pubblicazione: (2025)
Analyzing Neural Network-Based Generative Diffusion Models through Convex Optimization
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
DisjunctiveNet: Neural Symbolic Learning via Differentiable Convexified Optimization Layers
di: Pal, Shraman, et al.
Pubblicazione: (2026)
di: Pal, Shraman, et al.
Pubblicazione: (2026)
Scale-Invariant Neural Network Optimization: Norm Geometry and Heavy-Tailed Noise
di: Zhang, Jiayu, et al.
Pubblicazione: (2026)
di: Zhang, Jiayu, et al.
Pubblicazione: (2026)
Learning in Inverse Optimization: Incenter Cost, Augmented Suboptimality Loss, and Algorithms
di: Scroccaro, Pedro Zattoni, et al.
Pubblicazione: (2023)
di: Scroccaro, Pedro Zattoni, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Optimization and Generalization Guarantees for Weight Normalization
di: Cisneros-Velarde, Pedro, et al.
Pubblicazione: (2024) -
On the Width Scaling of Neural Optimizers Under Matrix Operator Norms I: Row/Column Normalization and Hyperparameter Transfer
di: Xu, Ruihan, et al.
Pubblicazione: (2026) -
Homotopy Relaxation Training Algorithms for Infinite-Width Two-Layer ReLU Neural Networks
di: Yang, Yahong, et al.
Pubblicazione: (2023) -
Fundamental Benefit of Alternating Updates in Minimax Optimization
di: Lee, Jaewook, et al.
Pubblicazione: (2024) -
Gearing Gaussian process modeling and sequential design towards stochastic simulators
di: Binois, Mickael, et al.
Pubblicazione: (2024)