Gespeichert in:
| Hauptverfasser: | Pan, Jiangong, Wan, Wei, Zhang, Yuejin, Bao, Chenlong, Shi, Zuoqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2407.21346 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Pontryagin Maximum Principle for Training Convolutional Neural Networks
von: Hofmann, Sebastian, et al.
Veröffentlicht: (2025)
von: Hofmann, Sebastian, et al.
Veröffentlicht: (2025)
Expansive Natural Neural Gradient Flows for Energy Minimization
von: Dahmen, Wolfgang, et al.
Veröffentlicht: (2025)
von: Dahmen, Wolfgang, et al.
Veröffentlicht: (2025)
Recent Advances in Non-convex Smoothness Conditions and Applicability to Deep Linear Neural Networks
von: Patel, Vivak, et al.
Veröffentlicht: (2024)
von: Patel, Vivak, et al.
Veröffentlicht: (2024)
Self2Seg: Single-Image Self-Supervised Joint Segmentation and Denoising
von: Gruber, Nadja, et al.
Veröffentlicht: (2023)
von: Gruber, Nadja, et al.
Veröffentlicht: (2023)
Variational conditional normalizing flows for computing second-order mean field control problems
von: Zhao, Jiaxi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxi, et al.
Veröffentlicht: (2025)
Stochastic Mirror Descent for Convex Optimization with Consensus Constraints
von: Borovykh, Anastasia, et al.
Veröffentlicht: (2022)
von: Borovykh, Anastasia, et al.
Veröffentlicht: (2022)
Convergence to good non-optimal critical points in the training of neural networks: Gradient descent optimization with one random initialization overcomes all bad non-global local minima with high probability
von: Ibragimov, Shokhrukh, et al.
Veröffentlicht: (2022)
von: Ibragimov, Shokhrukh, et al.
Veröffentlicht: (2022)
Power Homotopy for Zeroth-Order Non-Convex Optimizations
von: Xu, Chen
Veröffentlicht: (2025)
von: Xu, Chen
Veröffentlicht: (2025)
Global Optimization with A Power-Transformed Objective and Gaussian Smoothing
von: Xu, Chen
Veröffentlicht: (2024)
von: Xu, Chen
Veröffentlicht: (2024)
Asymptotic stability properties and a priori bounds for Adam and other gradient descent optimization methods
von: Dereich, Steffen, et al.
Veröffentlicht: (2025)
von: Dereich, Steffen, et al.
Veröffentlicht: (2025)
Deep Predictor-Corrector Networks for Robust Parameter Estimation in Non-autonomous System with Discontinuous Inputs
von: Gu, Gyeongwan, et al.
Veröffentlicht: (2026)
von: Gu, Gyeongwan, et al.
Veröffentlicht: (2026)
Error Bound Analysis for the Regularized Loss of Deep Linear Neural Networks
von: Chen, Po, et al.
Veröffentlicht: (2025)
von: Chen, Po, et al.
Veröffentlicht: (2025)
Fixed-Point Neural Optimal Transport without Implicit Differentiation
von: Park, Yesom, et al.
Veröffentlicht: (2026)
von: Park, Yesom, et al.
Veröffentlicht: (2026)
A Layer Separation Optimization Framework for Cross-Entropy Training in Deep Learning
von: Liu, Yaru, et al.
Veröffentlicht: (2026)
von: Liu, Yaru, et al.
Veröffentlicht: (2026)
A First Step Towards Mesh-Free Probabilistic Shape Optimization
von: Schmidt, Stephan, et al.
Veröffentlicht: (2026)
von: Schmidt, Stephan, et al.
Veröffentlicht: (2026)
Learning where to learn: Training data distribution optimization for scientific machine learning
von: Guerra, Nicolas, et al.
Veröffentlicht: (2025)
von: Guerra, Nicolas, et al.
Veröffentlicht: (2025)
An Augmented Lagrangian Method for Training Recurrent Neural Networks
von: Wang, Yue, et al.
Veröffentlicht: (2024)
von: Wang, Yue, et al.
Veröffentlicht: (2024)
Progressive Power Homotopy for Non-convex Optimization
von: Xu, Chen
Veröffentlicht: (2026)
von: Xu, Chen
Veröffentlicht: (2026)
On a Generalization of Wasserstein Distance and the Beckmann Problem to Connection Graphs
von: Robertson, Sawyer, et al.
Veröffentlicht: (2023)
von: Robertson, Sawyer, et al.
Veröffentlicht: (2023)
Convergence of gradient descent for deep neural networks
von: Chatterjee, Sourav
Veröffentlicht: (2022)
von: Chatterjee, Sourav
Veröffentlicht: (2022)
Prox-PINNs: A Deep Learning Algorithmic Framework for Elliptic Variational Inequalities
von: Gao, Yu, et al.
Veröffentlicht: (2025)
von: Gao, Yu, et al.
Veröffentlicht: (2025)
Quantum circuit design from a retraction-based Riemannian optimization framework
von: Lai, Zhijian, et al.
Veröffentlicht: (2026)
von: Lai, Zhijian, et al.
Veröffentlicht: (2026)
An adaptive framework for first-order gradient methods
von: Hu, Xiaozhe, et al.
Veröffentlicht: (2026)
von: Hu, Xiaozhe, et al.
Veröffentlicht: (2026)
Polynomial algorithm for the disjoint bilinear programming problem with an acute-angled polytope for a disjoint subset
von: Lozovanu, Dmitrii
Veröffentlicht: (2025)
von: Lozovanu, Dmitrii
Veröffentlicht: (2025)
Sample-wise Constrained Learning via a Sequential Penalty Approach with Applications in Image Processing
von: Lanzillotta, Francesca, et al.
Veröffentlicht: (2026)
von: Lanzillotta, Francesca, et al.
Veröffentlicht: (2026)
Preconditioned subgradient method for composite optimization: overparameterization and fast convergence
von: Díaz, Mateo, et al.
Veröffentlicht: (2025)
von: Díaz, Mateo, et al.
Veröffentlicht: (2025)
SVD-Preconditioned Gradient Descent Method for Solving Nonlinear Least Squares Problems
von: Chang, Zhipeng, et al.
Veröffentlicht: (2026)
von: Chang, Zhipeng, et al.
Veröffentlicht: (2026)
Faster Adaptive Optimization via Expected Gradient Outer Product Reparameterization
von: DePavia, Adela, et al.
Veröffentlicht: (2025)
von: DePavia, Adela, et al.
Veröffentlicht: (2025)
On the boundedness of the sequence generated by minibatch stochastic gradient descent
von: Bauschke, Heinz H., et al.
Veröffentlicht: (2025)
von: Bauschke, Heinz H., et al.
Veröffentlicht: (2025)
An Operator Learning Approach to Nonsmooth Optimal Control of Nonlinear PDEs
von: Song, Yongcun, et al.
Veröffentlicht: (2024)
von: Song, Yongcun, et al.
Veröffentlicht: (2024)
To be or not to be stable, that is the question: understanding neural networks for inverse problems
von: Evangelista, Davide, et al.
Veröffentlicht: (2022)
von: Evangelista, Davide, et al.
Veröffentlicht: (2022)
Convergence, design and training of continuous-time dropout as a random batch method
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2025)
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2025)
PETScML: Second-order solvers for training regression problems in Scientific Machine Learning
von: Zampini, Stefano, et al.
Veröffentlicht: (2024)
von: Zampini, Stefano, et al.
Veröffentlicht: (2024)
Consensus-based optimization for closed-box adversarial attacks and a connection to evolution strategies
von: Roith, Tim, et al.
Veröffentlicht: (2025)
von: Roith, Tim, et al.
Veröffentlicht: (2025)
Objective Value Change and Shape-Based Accelerated Optimization for the Neural Network Approximation
von: Xie, Pengcheng, et al.
Veröffentlicht: (2025)
von: Xie, Pengcheng, et al.
Veröffentlicht: (2025)
Quantization Robustness of Monotone Operator Equilibrium Networks
von: Li, James, et al.
Veröffentlicht: (2026)
von: Li, James, et al.
Veröffentlicht: (2026)
Convergence of Momentum-Based Optimization Algorithms with Time-Varying Parameters
von: Vidyasagar, Mathukumalli
Veröffentlicht: (2025)
von: Vidyasagar, Mathukumalli
Veröffentlicht: (2025)
A lifted Bregman strategy for training unfolded proximal neural network Gaussian denoisers
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
A Unified Framework for Lifted Training and Inversion Approaches
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
Handbook of Convergence Theorems for (Stochastic) Gradient Methods
von: Garrigos, Guillaume, et al.
Veröffentlicht: (2023)
von: Garrigos, Guillaume, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
The Pontryagin Maximum Principle for Training Convolutional Neural Networks
von: Hofmann, Sebastian, et al.
Veröffentlicht: (2025) -
Expansive Natural Neural Gradient Flows for Energy Minimization
von: Dahmen, Wolfgang, et al.
Veröffentlicht: (2025) -
Recent Advances in Non-convex Smoothness Conditions and Applicability to Deep Linear Neural Networks
von: Patel, Vivak, et al.
Veröffentlicht: (2024) -
Self2Seg: Single-Image Self-Supervised Joint Segmentation and Denoising
von: Gruber, Nadja, et al.
Veröffentlicht: (2023) -
Variational conditional normalizing flows for computing second-order mean field control problems
von: Zhao, Jiaxi, et al.
Veröffentlicht: (2025)