A Convexity-dependent Two-Phase Training Algorithm for Deep Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Hrycej, Tomas, Bermeitinger, Bernhard, Pavone, Massimo, Wiegand, Götz-Henrik, Handschuh, Siegfried |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Make Deep Networks Shallow Again
by: Bermeitinger, Bernhard, et al.
Published: (2023)
by: Bermeitinger, Bernhard, et al.
Published: (2023)
Reducing the Transformer Architecture to a Minimum
by: Bermeitinger, Bernhard, et al.
Published: (2024)
by: Bermeitinger, Bernhard, et al.
Published: (2024)
Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder
by: Wiegand, Götz-Henrik, et al.
Published: (2026)
by: Wiegand, Götz-Henrik, et al.
Published: (2026)
Efficient Neural Network Training via Subset Pretraining
by: Spörer, Jan, et al.
Published: (2024)
by: Spörer, Jan, et al.
Published: (2024)
Convergence of Some Convex Message Passing Algorithms to a Fixed Point
by: Voracek, Vaclav, et al.
Published: (2024)
by: Voracek, Vaclav, et al.
Published: (2024)
Training Safe Neural Networks with Global SDP Bounds
by: Soletskyi, Roman, et al.
Published: (2024)
by: Soletskyi, Roman, et al.
Published: (2024)
BAGEL: Projection-Free Algorithm for Adversarially Constrained Online Convex Optimization
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Federated Distributionally Robust Optimization with Non-Convex Objectives: Algorithm and Analysis
by: Jiao, Yang, et al.
Published: (2023)
by: Jiao, Yang, et al.
Published: (2023)
Parametrizing Convex Sets Using Sublinear Neural Networks
by: Martinet, Eloi
Published: (2026)
by: Martinet, Eloi
Published: (2026)
An Improved Finite-time Analysis of Temporal Difference Learning with Deep Neural Networks
by: Ke, Zhifa, et al.
Published: (2024)
by: Ke, Zhifa, et al.
Published: (2024)
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
by: Pilanci, Mert
Published: (2023)
by: Pilanci, Mert
Published: (2023)
Training Infinitely Deep and Wide Transformers
by: Barboni, Raphaël, et al.
Published: (2026)
by: Barboni, Raphaël, et al.
Published: (2026)
Deep Learning for Two-Stage Robust Integer Optimization
by: Dumouchelle, Justin, et al.
Published: (2023)
by: Dumouchelle, Justin, et al.
Published: (2023)
Provable Acceleration of Nesterov's Accelerated Gradient Method over Heavy Ball Method in Training Over-Parameterized Neural Networks
by: Liu, Xin, et al.
Published: (2022)
by: Liu, Xin, et al.
Published: (2022)
Solving Integrated Process Planning and Scheduling Problem via Graph Neural Network Based Deep Reinforcement Learning
by: Li, Hongpei, et al.
Published: (2024)
by: Li, Hongpei, et al.
Published: (2024)
Convex Formulations for Training Two-Layer ReLU Neural Networks
by: Prakhya, Karthik, et al.
Published: (2024)
by: Prakhya, Karthik, et al.
Published: (2024)
Generating Likely Counterfactuals Using Sum-Product Networks
by: Nemecek, Jiri, et al.
Published: (2024)
by: Nemecek, Jiri, et al.
Published: (2024)
Online Submodular Maximization via Online Convex Optimization
by: Salem, Tareq Si, et al.
Published: (2023)
by: Salem, Tareq Si, et al.
Published: (2023)
Convex and Bilevel Optimization for Neuro-Symbolic Inference and Learning
by: Dickens, Charles, et al.
Published: (2024)
by: Dickens, Charles, et al.
Published: (2024)
A Library of Mirrors: Deep Neural Nets in Low Dimensions are Convex Lasso Models with Reflection Features
by: Zeger, Emi, et al.
Published: (2024)
by: Zeger, Emi, et al.
Published: (2024)
Unsupervised Training of Diffusion Models for Feasible Solution Generation in Neural Combinatorial Optimization
by: Hong, Seong-Hyun, et al.
Published: (2024)
by: Hong, Seong-Hyun, et al.
Published: (2024)
Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks
by: Kiyani, Elham, et al.
Published: (2025)
by: Kiyani, Elham, et al.
Published: (2025)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
by: Onorato, Gabriele
Published: (2024)
by: Onorato, Gabriele
Published: (2024)
Solving Functional Optimization with Deep Networks and Variational Principles
by: Kamtue, Kawisorn, et al.
Published: (2024)
by: Kamtue, Kawisorn, et al.
Published: (2024)
Hidden Convexity of Fair PCA and Fast Solver via Eigenvalue Optimization
by: Shen, Junhui, et al.
Published: (2025)
by: Shen, Junhui, et al.
Published: (2025)
Non-Smooth Weakly-Convex Finite-sum Coupled Compositional Optimization
by: Hu, Quanqi, et al.
Published: (2023)
by: Hu, Quanqi, et al.
Published: (2023)
Applications of 0-1 Neural Networks in Prescription and Prediction
by: Patil, Vrishabh, et al.
Published: (2024)
by: Patil, Vrishabh, et al.
Published: (2024)
Taming Binarized Neural Networks and Mixed-Integer Programs
by: Aspman, Johannes, et al.
Published: (2023)
by: Aspman, Johannes, et al.
Published: (2023)
Geometric Neural Operators (GNPs) for Data-Driven Deep Learning of Non-Euclidean Operators
by: Quackenbush, Blaine, et al.
Published: (2024)
by: Quackenbush, Blaine, et al.
Published: (2024)
Feed-Forward Neural Networks as a Mixed-Integer Program
by: Aftabi, Navid, et al.
Published: (2024)
by: Aftabi, Navid, et al.
Published: (2024)
Graph Neural Networks for the Offline Nanosatellite Task Scheduling Problem
by: Pacheco, Bruno Machado, et al.
Published: (2023)
by: Pacheco, Bruno Machado, et al.
Published: (2023)
PRISM: Distribution-free Adaptive Computation of Matrix Functions for Accelerating Neural Network Training
by: Yang, Shenghao, et al.
Published: (2026)
by: Yang, Shenghao, et al.
Published: (2026)
SMiLE: Provably Enforcing Global Relational Properties in Neural Networks
by: Francobaldi, Matteo, et al.
Published: (2025)
by: Francobaldi, Matteo, et al.
Published: (2025)
Ginger: An Efficient Curvature Approximation with Linear Complexity for General Neural Networks
by: Hao, Yongchang, et al.
Published: (2024)
by: Hao, Yongchang, et al.
Published: (2024)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
AdamZ: An Enhanced Optimisation Method for Neural Network Training
by: Zaznov, Ilia, et al.
Published: (2024)
by: Zaznov, Ilia, et al.
Published: (2024)
Learning the Riccati solution operator for time-varying LQR via Deep Operator Networks
by: Chen, Jun, et al.
Published: (2026)
by: Chen, Jun, et al.
Published: (2026)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
by: Cayci, Semih
Published: (2024)
by: Cayci, Semih
Published: (2024)
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial
by: Li, Haoyu, et al.
Published: (2026)
by: Li, Haoyu, et al.
Published: (2026)
Similar Items
-
Make Deep Networks Shallow Again
by: Bermeitinger, Bernhard, et al.
Published: (2023) -
Reducing the Transformer Architecture to a Minimum
by: Bermeitinger, Bernhard, et al.
Published: (2024) -
Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder
by: Wiegand, Götz-Henrik, et al.
Published: (2026) -
Efficient Neural Network Training via Subset Pretraining
by: Spörer, Jan, et al.
Published: (2024) -
Convergence of Some Convex Message Passing Algorithms to a Fixed Point
by: Voracek, Vaclav, et al.
Published: (2024)