Representation and Regression Problems in Neural Networks: Relaxation, Generalization, and Numerics
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Kang, Zuazua, Enrique |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Moments, Time-Inversion and Source Identification for the Heat Equation
by: Liu, Kang, et al.
Published: (2025)
by: Liu, Kang, et al.
Published: (2025)
Effectively Leveraging Momentum Terms in Stochastic Line Search Frameworks for Fast Optimization of Finite-Sum Problems
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
Convergence Conditions for Stochastic Line Search Based Optimization of Over-parametrized Models
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
Stable gradient-adjusted root mean square propagation on least squares problem
by: Li, Runze, et al.
Published: (2024)
by: Li, Runze, et al.
Published: (2024)
Two-level overlapping additive Schwarz preconditioner for training scientific machine learning applications
by: Lee, Youngkyu, et al.
Published: (2024)
by: Lee, Youngkyu, et al.
Published: (2024)
Enhancing training of physics-informed neural networks using domain-decomposition based preconditioning strategies
by: Kopaničáková, Alena, et al.
Published: (2023)
by: Kopaničáková, Alena, et al.
Published: (2023)
Sample-wise Constrained Learning via a Sequential Penalty Approach with Applications in Image Processing
by: Lanzillotta, Francesca, et al.
Published: (2026)
by: Lanzillotta, Francesca, et al.
Published: (2026)
An Augmented Lagrangian Method for Training Recurrent Neural Networks
by: Wang, Yue, et al.
Published: (2024)
by: Wang, Yue, et al.
Published: (2024)
Error Bound Analysis for the Regularized Loss of Deep Linear Neural Networks
by: Chen, Po, et al.
Published: (2025)
by: Chen, Po, et al.
Published: (2025)
Similarity-based fuzzy clustering scientific articles: potentials and challenges from mathematical and computational perspectives
by: Huong, Vu Thi, et al.
Published: (2025)
by: Huong, Vu Thi, et al.
Published: (2025)
Faster Adaptive Optimization via Expected Gradient Outer Product Reparameterization
by: DePavia, Adela, et al.
Published: (2025)
by: DePavia, Adela, et al.
Published: (2025)
Tracking the Median of Gradients with a Stochastic Proximal Point Method
by: Schaipp, Fabian, et al.
Published: (2024)
by: Schaipp, Fabian, et al.
Published: (2024)
BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization
by: Zhang, Hengrui, et al.
Published: (2026)
by: Zhang, Hengrui, et al.
Published: (2026)
A Two-Phase Adaptive Balanced Penalty Method for Controllable Pareto Front Learning under Split Feasibility Conditions
by: Hoang, Nguyen Viet, et al.
Published: (2026)
by: Hoang, Nguyen Viet, et al.
Published: (2026)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
Beyond Discreteness: Sample Complexity Analysis of Straight-Through Estimator for 1-bit Quantization
by: Jeong, Halyun, et al.
Published: (2025)
by: Jeong, Halyun, et al.
Published: (2025)
Supplementary Materials to Graph Convolutional Branch and Bound
by: Sciandra, Lorenzo, et al.
Published: (2024)
by: Sciandra, Lorenzo, et al.
Published: (2024)
Global Convergence of Sampling-Based Nonconvex Optimization through Diffusion-Style Smoothing
by: Yi, Zeji, et al.
Published: (2026)
by: Yi, Zeji, et al.
Published: (2026)
Global Optimization of Gaussian processes
by: Schweidtmann, Artur M., et al.
Published: (2020)
by: Schweidtmann, Artur M., et al.
Published: (2020)
Performance Estimation of second-order optimization methods on classes of univariate functions
by: Rubbens, Anne, et al.
Published: (2025)
by: Rubbens, Anne, et al.
Published: (2025)
Exponential convergence rates for momentum stochastic gradient descent in the overparametrized setting
by: Gess, Benjamin, et al.
Published: (2023)
by: Gess, Benjamin, et al.
Published: (2023)
Quantum-Inspired DRL Approach with LSTM and OU Noise for Cut Order Planning Optimization
by: Chrisnanto, Yulison Herry, et al.
Published: (2025)
by: Chrisnanto, Yulison Herry, et al.
Published: (2025)
Global Solutions to Non-Convex Functional Constrained Problems with Hidden Convexity
by: Fatkhullin, Ilyas, et al.
Published: (2025)
by: Fatkhullin, Ilyas, et al.
Published: (2025)
Deep Relaxation of Controlled Stochastic Gradient Descent via Singular Perturbations
by: Bardi, Martino, et al.
Published: (2022)
by: Bardi, Martino, et al.
Published: (2022)
Iso-Riemannian Optimization on Learned Data Manifolds
by: Diepeveen, Willem, et al.
Published: (2025)
by: Diepeveen, Willem, et al.
Published: (2025)
Policy Optimization over General State and Action Spaces
by: Ju, Caleb, et al.
Published: (2022)
by: Ju, Caleb, et al.
Published: (2022)
Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
by: Ueda, Ryota, et al.
Published: (2025)
by: Ueda, Ryota, et al.
Published: (2025)
Kurdyka-Łojasiewicz exponent via Hadamard parametrization
by: Ouyang, Wenqing, et al.
Published: (2024)
by: Ouyang, Wenqing, et al.
Published: (2024)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
Distributed Computing for Huge-Scale Aggregative Convex Programming
by: Tao, Luoyi
Published: (2026)
by: Tao, Luoyi
Published: (2026)
Stochastic versus Deterministic in Stochastic Gradient Descent
by: Li, Runze, et al.
Published: (2025)
by: Li, Runze, et al.
Published: (2025)
Kurdyka-Łojasiewicz exponent via square transformation
by: Ouyang, Wenqing
Published: (2025)
by: Ouyang, Wenqing
Published: (2025)
A KL-based Analysis Framework with Applications to Non-Descent Optimization Methods
by: Qiu, Junwen, et al.
Published: (2024)
by: Qiu, Junwen, et al.
Published: (2024)
Shuffling the Stochastic Mirror Descent via Dual Lipschitz Continuity and Kernel Conditioning
by: Qiu, Junwen, et al.
Published: (2026)
by: Qiu, Junwen, et al.
Published: (2026)
Optimization with Trained Machine Learning Models Embedded
by: Schweidtmann, Artur M., et al.
Published: (2022)
by: Schweidtmann, Artur M., et al.
Published: (2022)
PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming
by: Tang, Bo, et al.
Published: (2022)
by: Tang, Bo, et al.
Published: (2022)
Accelerating preconditioned ADMM via degenerate proximal point mappings
by: Sun, Defeng, et al.
Published: (2024)
by: Sun, Defeng, et al.
Published: (2024)
SUDA-Muon: Structural Design Principles and Boundaries for Fully Decentralized Muon
by: Zhang, Hengrui, et al.
Published: (2026)
by: Zhang, Hengrui, et al.
Published: (2026)
Stochastic First-Order Methods with Non-smooth and Non-Euclidean Proximal Terms for Nonconvex High-Dimensional Stochastic Optimization
by: Xie, Yue, et al.
Published: (2024)
by: Xie, Yue, et al.
Published: (2024)
How to beat a Bayesian adversary
by: Ding, Zihan, et al.
Published: (2024)
by: Ding, Zihan, et al.
Published: (2024)
Similar Items
-
Moments, Time-Inversion and Source Identification for the Heat Equation
by: Liu, Kang, et al.
Published: (2025) -
Effectively Leveraging Momentum Terms in Stochastic Line Search Frameworks for Fast Optimization of Finite-Sum Problems
by: Lapucci, Matteo, et al.
Published: (2024) -
Convergence Conditions for Stochastic Line Search Based Optimization of Over-parametrized Models
by: Lapucci, Matteo, et al.
Published: (2024) -
Stable gradient-adjusted root mean square propagation on least squares problem
by: Li, Runze, et al.
Published: (2024) -
Two-level overlapping additive Schwarz preconditioner for training scientific machine learning applications
by: Lee, Youngkyu, et al.
Published: (2024)