Convergence of SGD for Training Neural Networks with Sliced Wasserstein Losses
Fuente:
arXiv
Saved in:
| Main Author: | Tanguy, Eloi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Properties of Discrete Sliced Wasserstein Losses
by: Tanguy, Eloi, et al.
Published: (2023)
by: Tanguy, Eloi, et al.
Published: (2023)
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients
by: Bruno, Stefano, et al.
Published: (2025)
by: Bruno, Stefano, et al.
Published: (2025)
Neural Wasserstein Gradient Flows for Maximum Mean Discrepancies with Riesz Kernels
by: Altekrüger, Fabian, et al.
Published: (2023)
by: Altekrüger, Fabian, et al.
Published: (2023)
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
Global Convergence of SGD For Logistic Loss on Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2023)
by: Gopalani, Pulkit, et al.
Published: (2023)
Linear convergence of proximal descent schemes on the Wasserstein space
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
Large Deviation Upper Bounds and Improved MSE Rates of Nonlinear SGD: Heavy-tailed Noise and Power of Symmetry
by: Armacki, Aleksandar, et al.
Published: (2024)
by: Armacki, Aleksandar, et al.
Published: (2024)
Stochastic Inverse Problem: stability, regularization and Wasserstein gradient flow
by: Li, Qin, et al.
Published: (2024)
by: Li, Qin, et al.
Published: (2024)
Generalized Wasserstein Flow Matching: Transport Plans, Everywhere, All at Once
by: Piening, Moritz, et al.
Published: (2026)
by: Piening, Moritz, et al.
Published: (2026)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
by: Dus, Mathias
Published: (2026)
by: Dus, Mathias
Published: (2026)
Convergence rates for the Adam optimizer
by: Dereich, Steffen, et al.
Published: (2024)
by: Dereich, Steffen, et al.
Published: (2024)
A Generalization Result for Convergence in Learning-to-Optimize
by: Sucker, Michael, et al.
Published: (2024)
by: Sucker, Michael, et al.
Published: (2024)
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025)
by: Kassing, Sebastian, et al.
Published: (2025)
Statistical Inference for Linear Functionals of Online SGD in High-dimensional Linear Regression
by: Agrawalla, Bhavya, et al.
Published: (2023)
by: Agrawalla, Bhavya, et al.
Published: (2023)
Prelimit Coupling and Steady-State Convergence of Constant-stepsize Nonsmooth Contractive SA
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
High-dimensional scaling limits and fluctuations of online least-squares SGD with smooth covariance
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
A Piecewise Lyapunov Analysis of Sub-quadratic SGD: Applications to Robust and Quantile Regression
by: Zhang, Yixuan, et al.
Published: (2025)
by: Zhang, Yixuan, et al.
Published: (2025)
Global Convergence of SGD On Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2022)
by: Gopalani, Pulkit, et al.
Published: (2022)
Convergence of Actor-Critic Learning for Mean Field Games and Mean Field Control in Continuous Spaces
by: Fouque, Jean-Pierre, et al.
Published: (2025)
by: Fouque, Jean-Pierre, et al.
Published: (2025)
Convergence Error Analysis of Reflected Gradient Langevin Dynamics for Globally Optimizing Non-Convex Constrained Problems
by: Sato, Kanji, et al.
Published: (2022)
by: Sato, Kanji, et al.
Published: (2022)
The geometry of financial institutions -- Wasserstein clustering of financial data
by: Riess, Lorenz, et al.
Published: (2023)
by: Riess, Lorenz, et al.
Published: (2023)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022)
by: Neufeld, Ariel, et al.
Published: (2022)
Sliced Transport Plans
by: Tanguy, Eloi, et al.
Published: (2025)
by: Tanguy, Eloi, et al.
Published: (2025)
Neural Hilbert Ladders: Multi-Layer Neural Networks in Function Space
by: Chen, Zhengdao
Published: (2023)
by: Chen, Zhengdao
Published: (2023)
Wasserstein Contraction of Coordinate Ascent Variational Inference
by: Caprio, Rocco, et al.
Published: (2026)
by: Caprio, Rocco, et al.
Published: (2026)
Sliced Inner Product Gromov-Wasserstein Distances
by: Gong, Xiaoyun, et al.
Published: (2026)
by: Gong, Xiaoyun, et al.
Published: (2026)
Sliced Wasserstein Steering between Gaussian Measures
by: Ito, Kaito, et al.
Published: (2026)
by: Ito, Kaito, et al.
Published: (2026)
Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations
by: Huynh, Phuoc-Toan, et al.
Published: (2026)
by: Huynh, Phuoc-Toan, et al.
Published: (2026)
Convergence rate of Tsallis entropic regularized optimal transport
by: Suguro, Takeshi, et al.
Published: (2023)
by: Suguro, Takeshi, et al.
Published: (2023)
A Novel Sliced Fused Gromov-Wasserstein Distance
by: Piening, Moritz, et al.
Published: (2025)
by: Piening, Moritz, et al.
Published: (2025)
Convergence of linear programming hierarchies for Gibbs states of spin systems
by: Fawzi, Hamza, et al.
Published: (2025)
by: Fawzi, Hamza, et al.
Published: (2025)
SGD with Partial Hessian for Deep Neural Networks Optimization
by: Sun, Ying, et al.
Published: (2024)
by: Sun, Ying, et al.
Published: (2024)
Faster Convergence of Local SGD for Over-Parameterized Models
by: Qin, Tiancheng, et al.
Published: (2022)
by: Qin, Tiancheng, et al.
Published: (2022)
From PowerSGD to PowerSGD+: Low-Rank Gradient Compression for Distributed Optimization with Convergence Guarantees
by: Xie, Shengping, et al.
Published: (2025)
by: Xie, Shengping, et al.
Published: (2025)
Convergence of coordinate ascent variational inference for log-concave measures via optimal transport
by: Arnese, Manuel, et al.
Published: (2024)
by: Arnese, Manuel, et al.
Published: (2024)
Slicing Wasserstein Over Wasserstein Via Functional Optimal Transport
by: Piening, Moritz, et al.
Published: (2025)
by: Piening, Moritz, et al.
Published: (2025)
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
by: Attia, Amit, et al.
Published: (2025)
by: Attia, Amit, et al.
Published: (2025)
Optimization Trade-offs in Asynchronous Federated Learning: A Stochastic Networks Approach
by: Alahyane, Abdelkrim, et al.
Published: (2026)
by: Alahyane, Abdelkrim, et al.
Published: (2026)
Neural Brownian Motion
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Similar Items
-
Properties of Discrete Sliced Wasserstein Losses
by: Tanguy, Eloi, et al.
Published: (2023) -
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients
by: Bruno, Stefano, et al.
Published: (2025) -
Neural Wasserstein Gradient Flows for Maximum Mean Discrepancies with Riesz Kernels
by: Altekrüger, Fabian, et al.
Published: (2023) -
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
by: Lam, Samuel Chun-Hei, et al.
Published: (2024) -
Global Convergence of SGD For Logistic Loss on Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2023)