Global Dynamics of Heavy-Tailed SGDs in Nonconvex Loss Landscape: Characterization and Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Xingyu, Rhee, Chang-Han |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Large Deviations and Metastability Analysis for Heavy-Tailed Dynamical Systems
por: Wang, Xingyu, et al.
Publicado: (2023)
por: Wang, Xingyu, et al.
Publicado: (2023)
First-Exit Time Analysis for Truncated Heavy-Tailed Dynamical Systems
por: Wang, Xingyu, et al.
Publicado: (2026)
por: Wang, Xingyu, et al.
Publicado: (2026)
Implicit Compressibility of Overparametrized Neural Networks Trained with Heavy-Tailed SGD
por: Wan, Yijun, et al.
Publicado: (2023)
por: Wan, Yijun, et al.
Publicado: (2023)
Strongly Efficient Rare-Event Simulation for Regularly Varying Lévy Processes with Infinite Activities
por: Wang, Xingyu, et al.
Publicado: (2023)
por: Wang, Xingyu, et al.
Publicado: (2023)
Diffusion Models with Heavy-Tailed Targets: Score Estimation and Sampling Guarantees
por: Yu, Yifeng, et al.
Publicado: (2026)
por: Yu, Yifeng, et al.
Publicado: (2026)
Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise
por: Agrawal, Shubhada, et al.
Publicado: (2026)
por: Agrawal, Shubhada, et al.
Publicado: (2026)
Algorithmic Stability of Stochastic Gradient Descent with Momentum under Heavy-Tailed Noise
por: Dang, Thanh, et al.
Publicado: (2025)
por: Dang, Thanh, et al.
Publicado: (2025)
On the Nonconvexity of Push-Forward Constraints and Its Consequences in Machine Learning
por: de Lara, Lucas, et al.
Publicado: (2024)
por: de Lara, Lucas, et al.
Publicado: (2024)
Steady-State Behavior of Constant-Stepsize Stochastic Approximation: Gaussian Approximation and Tail Bounds
por: Wang, Zedong, et al.
Publicado: (2026)
por: Wang, Zedong, et al.
Publicado: (2026)
Tail Asymptotics of Cluster Sizes in Multivariate Heavy-Tailed Hawkes Processes
por: Blanchet, Jose, et al.
Publicado: (2025)
por: Blanchet, Jose, et al.
Publicado: (2025)
Stability and Generalization of Nonconvex Optimization with Heavy-Tailed Noise
por: Chen, Hongxu, et al.
Publicado: (2026)
por: Chen, Hongxu, et al.
Publicado: (2026)
Sharper Perturbed-Kullback-Leibler Exponential Tail Bounds for Beta and Dirichlet Distributions
por: Perrault, Pierre
Publicado: (2025)
por: Perrault, Pierre
Publicado: (2025)
Variational Tail Bounds for Norms of Random Vectors and Matrices
por: Bahmani, Sohail
Publicado: (2025)
por: Bahmani, Sohail
Publicado: (2025)
Near-Optimal Decentralized Stochastic Nonconvex Optimization with Heavy-Tailed Noise
por: Wang, Menglian, et al.
Publicado: (2026)
por: Wang, Menglian, et al.
Publicado: (2026)
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
por: Chemnitz, Dennis, et al.
Publicado: (2024)
por: Chemnitz, Dennis, et al.
Publicado: (2024)
Sharp Concentration Inequalities: Phase Transition and Mixing of Orlicz Tails with Variance
por: Shen, Yinan, et al.
Publicado: (2026)
por: Shen, Yinan, et al.
Publicado: (2026)
Sample Path Large Deviations for Multivariate Heavy-Tailed Hawkes Processes and Related Lévy Processes
por: Blanchet, Jose, et al.
Publicado: (2025)
por: Blanchet, Jose, et al.
Publicado: (2025)
A Method For Bounding Tail Probabilities
por: Zlatanov, Nikola
Publicado: (2024)
por: Zlatanov, Nikola
Publicado: (2024)
Variational Smoothing and Inference for SDEs from Sparse Data with Dynamic Neural Flows
por: Wang, Yu, et al.
Publicado: (2026)
por: Wang, Yu, et al.
Publicado: (2026)
Resolving Node Identifiability in Graph Neural Processes via Laplacian Spectral Encodings
por: Yan, Zimo, et al.
Publicado: (2025)
por: Yan, Zimo, et al.
Publicado: (2025)
Optimal Asynchronous Stochastic Nonconvex Optimization under Heavy-Tailed Noise
por: Wu, Yidong, et al.
Publicado: (2026)
por: Wu, Yidong, et al.
Publicado: (2026)
EVT-Based Rate-Preserving Distributional Robustness for Tail Risk Functionals
por: Deo, Anand
Publicado: (2025)
por: Deo, Anand
Publicado: (2025)
Properties of Discrete Sliced Wasserstein Losses
por: Tanguy, Eloi, et al.
Publicado: (2023)
por: Tanguy, Eloi, et al.
Publicado: (2023)
A General Reduction for High-Probability Analysis with General Light-Tailed Distributions
por: Attia, Amit, et al.
Publicado: (2024)
por: Attia, Amit, et al.
Publicado: (2024)
Tail-Aware Information-Theoretic Generalization for RLHF and SGLD
por: Zhang, Huiming, et al.
Publicado: (2026)
por: Zhang, Huiming, et al.
Publicado: (2026)
A Note on Asynchronous Challenges: Unveiling Formulaic Bias and Data Loss in the Hayashi-Yoshida Estimator
por: Georgiadis, Evangelos
Publicado: (2024)
por: Georgiadis, Evangelos
Publicado: (2024)
Convergence Error Analysis of Reflected Gradient Langevin Dynamics for Globally Optimizing Non-Convex Constrained Problems
por: Sato, Kanji, et al.
Publicado: (2022)
por: Sato, Kanji, et al.
Publicado: (2022)
Decentralized Proximal Stochastic Gradient Langevin Dynamics
por: Islam, Mohammad Rafiqul, et al.
Publicado: (2026)
por: Islam, Mohammad Rafiqul, et al.
Publicado: (2026)
Convergence of SGD for Training Neural Networks with Sliced Wasserstein Losses
por: Tanguy, Eloi
Publicado: (2023)
por: Tanguy, Eloi
Publicado: (2023)
Birth-death dynamics for sampling: Global convergence, approximations and their asymptotics
por: Lu, Yulong, et al.
Publicado: (2022)
por: Lu, Yulong, et al.
Publicado: (2022)
Static and Dynamic Approaches to Computing Barycenters of Probability Measures on Graphs
por: Gentile, David, et al.
Publicado: (2026)
por: Gentile, David, et al.
Publicado: (2026)
A Gaussian Comparison Theorem for Training Dynamics in Machine Learning
por: Panahi, Ashkan
Publicado: (2026)
por: Panahi, Ashkan
Publicado: (2026)
Tail Asymptotic of Heavy-Tail Risks with Elliptical Copula
por: Wang, Kai, et al.
Publicado: (2024)
por: Wang, Kai, et al.
Publicado: (2024)
Tail-Sensitive KL and Rényi Convergence of Unadjusted Hamiltonian Monte Carlo via One-Shot Couplings
por: Bou-Rabee, Nawaf, et al.
Publicado: (2026)
por: Bou-Rabee, Nawaf, et al.
Publicado: (2026)
Convergence, Sticking and Escape: Stochastic Dynamics Near Critical Points in SGD
por: Dudukalov, Dmitry, et al.
Publicado: (2025)
por: Dudukalov, Dmitry, et al.
Publicado: (2025)
Homogenization of Multi-agent Learning Dynamics in Finite-state Markov Games
por: Kerzreho, Yann
Publicado: (2025)
por: Kerzreho, Yann
Publicado: (2025)
Denoising distances beyond the volumetric barrier
por: Huang, Han, et al.
Publicado: (2026)
por: Huang, Han, et al.
Publicado: (2026)
Reconstructing the Geometry of Random Geometric Graphs
por: Huang, Han, et al.
Publicado: (2024)
por: Huang, Han, et al.
Publicado: (2024)
Large Deviation Upper Bounds and Improved MSE Rates of Nonlinear SGD: Heavy-tailed Noise and Power of Symmetry
por: Armacki, Aleksandar, et al.
Publicado: (2024)
por: Armacki, Aleksandar, et al.
Publicado: (2024)
ResNets of All Shapes and Sizes: Convergence of Training Dynamics in the Large-scale Limit
por: Chaintron, Louis-Pierre, et al.
Publicado: (2026)
por: Chaintron, Louis-Pierre, et al.
Publicado: (2026)
Ejemplares similares
-
Large Deviations and Metastability Analysis for Heavy-Tailed Dynamical Systems
por: Wang, Xingyu, et al.
Publicado: (2023) -
First-Exit Time Analysis for Truncated Heavy-Tailed Dynamical Systems
por: Wang, Xingyu, et al.
Publicado: (2026) -
Implicit Compressibility of Overparametrized Neural Networks Trained with Heavy-Tailed SGD
por: Wan, Yijun, et al.
Publicado: (2023) -
Strongly Efficient Rare-Event Simulation for Regularly Varying Lévy Processes with Infinite Activities
por: Wang, Xingyu, et al.
Publicado: (2023) -
Diffusion Models with Heavy-Tailed Targets: Score Estimation and Sampling Guarantees
por: Yu, Yifeng, et al.
Publicado: (2026)