Near-Optimal Streaming Heavy-Tailed Statistical Estimation with Clipped SGD
Fuente:
arXiv
Saved in:
| Main Authors: | Das, Aniket, Nagaraj, Dheeraj, Pal, Soumyabrata, Suggala, Arun, Varshney, Prateek |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
by: Vijayan, Sushant, et al.
Published: (2025)
by: Vijayan, Sushant, et al.
Published: (2025)
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
by: Kumar, Ramnath, et al.
Published: (2023)
by: Kumar, Ramnath, et al.
Published: (2023)
From Gradient Clipping to Normalization for Heavy Tailed SGD
by: Hübler, Florian, et al.
Published: (2024)
by: Hübler, Florian, et al.
Published: (2024)
Online Matrix Completion: A Collaborative Approach with Hott Items
by: Baby, Dheeraj, et al.
Published: (2024)
by: Baby, Dheeraj, et al.
Published: (2024)
Privacy of SGD under Gaussian or Heavy-Tailed Noise: Guarantees without Gradient Clipping
by: Şimşekli, Umut, et al.
Published: (2024)
by: Şimşekli, Umut, et al.
Published: (2024)
Revisiting Gradient Normalization and Clipping for Nonconvex SGD under Heavy-Tailed Noise: Necessity, Sufficiency, and Acceleration
by: Sun, Tao, et al.
Published: (2024)
by: Sun, Tao, et al.
Published: (2024)
Convergence of Clipped-SGD for Convex $(L_0,L_1)$-Smooth Optimization with Heavy-Tailed Noise
by: Chezhegov, Savelii, et al.
Published: (2025)
by: Chezhegov, Savelii, et al.
Published: (2025)
Heavy-Tail Phenomenon in Decentralized SGD
by: Gurbuzbalaban, Mert, et al.
Published: (2022)
by: Gurbuzbalaban, Mert, et al.
Published: (2022)
To Clip or not to Clip: the Dynamics of SGD with Gradient Clipping in High-Dimensions
by: Marshall, Noah, et al.
Published: (2024)
by: Marshall, Noah, et al.
Published: (2024)
Can SGD Handle Heavy-Tailed Noise?
by: Fatkhullin, Ilyas, et al.
Published: (2025)
by: Fatkhullin, Ilyas, et al.
Published: (2025)
Tight Long-Term Tail Decay of (Clipped) SGD in Non-Convex Optimization
by: Armacki, Aleksandar, et al.
Published: (2026)
by: Armacki, Aleksandar, et al.
Published: (2026)
Clip Body and Tail Separately: High Probability Guarantees for DPSGD with Heavy Tails
by: Sha, Haichao, et al.
Published: (2024)
by: Sha, Haichao, et al.
Published: (2024)
Nonconvex Stochastic Optimization under Heavy-Tailed Noises: Optimal Convergence without Gradient Clipping
by: Liu, Zijian, et al.
Published: (2024)
by: Liu, Zijian, et al.
Published: (2024)
Offline-to-online hyperparameter transfer for stochastic bandits
by: Sharma, Dravyansh, et al.
Published: (2025)
by: Sharma, Dravyansh, et al.
Published: (2025)
Robust Federated Learning Over the Air: Combating Heavy-Tailed Noise with Median Anchored Clipping
by: Li, Jiaxing, et al.
Published: (2024)
by: Li, Jiaxing, et al.
Published: (2024)
The Poisson Midpoint Method for Langevin Dynamics: Provably Efficient Discretization for Diffusion Models
by: Kandasamy, Saravanan, et al.
Published: (2024)
by: Kandasamy, Saravanan, et al.
Published: (2024)
Implicit Compressibility of Overparametrized Neural Networks Trained with Heavy-Tailed SGD
by: Wan, Yijun, et al.
Published: (2023)
by: Wan, Yijun, et al.
Published: (2023)
Efficient Private SCO for Heavy-Tailed Data via Averaged Clipping
by: Jin, Chenhan, et al.
Published: (2022)
by: Jin, Chenhan, et al.
Published: (2022)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
by: Patil, Gandharv, et al.
Published: (2022)
by: Patil, Gandharv, et al.
Published: (2022)
DC-SGD: Differentially Private SGD with Dynamic Clipping through Gradient Norm Distribution Estimation
by: Wei, Chengkun, et al.
Published: (2025)
by: Wei, Chengkun, et al.
Published: (2025)
Poisson Midpoint Method for Log Concave Sampling: Beyond the Strong Error Lower Bounds
by: Srinivasan, Rishikesh, et al.
Published: (2025)
by: Srinivasan, Rishikesh, et al.
Published: (2025)
Tractable Gaussian Phase Retrieval with Heavy Tails and Adversarial Corruption with Near-Linear Sample Complexity
by: Das, Santanu, et al.
Published: (2026)
by: Das, Santanu, et al.
Published: (2026)
Interleaved Gibbs Diffusion: Generating Discrete-Continuous Data with Implicit Constraints
by: Anil, Gautham Govind, et al.
Published: (2025)
by: Anil, Gautham Govind, et al.
Published: (2025)
GeoClip: Geometry-Aware Clipping for Differentially Private SGD
by: Gilani, Atefeh, et al.
Published: (2025)
by: Gilani, Atefeh, et al.
Published: (2025)
Near-Optimal Decentralized Stochastic Nonconvex Optimization with Heavy-Tailed Noise
by: Wang, Menglian, et al.
Published: (2026)
by: Wang, Menglian, et al.
Published: (2026)
On the Convergence of DP-SGD with Adaptive Clipping
by: Shulgin, Egor, et al.
Published: (2024)
by: Shulgin, Egor, et al.
Published: (2024)
Differentially Private Clipped-SGD: High-Probability Convergence with Arbitrary Clipping Level
by: Khah, Saleh Vatan, et al.
Published: (2025)
by: Khah, Saleh Vatan, et al.
Published: (2025)
Data-Driven Upper Confidence Bounds with Near-Optimal Regret for Heavy-Tailed Bandits
by: Tamás, Ambrus, et al.
Published: (2024)
by: Tamás, Ambrus, et al.
Published: (2024)
Glauber Generative Model: Discrete Diffusion Models via Binary Classification
by: Varma, Harshit, et al.
Published: (2024)
by: Varma, Harshit, et al.
Published: (2024)
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization
by: Varshney, Prateek, et al.
Published: (2024)
by: Varshney, Prateek, et al.
Published: (2024)
Clipping Improves Adam-Norm and AdaGrad-Norm when the Noise Is Heavy-Tailed
by: Chezhegov, Savelii, et al.
Published: (2024)
by: Chezhegov, Savelii, et al.
Published: (2024)
Clipped SGD Algorithms for Performative Prediction: Tight Bounds for Clipping Bias and Remedies
by: Li, Qiang, et al.
Published: (2024)
by: Li, Qiang, et al.
Published: (2024)
CDQuant: Greedy Coordinate Descent for Accurate LLM Quantization
by: Nair, Pranav Ajit, et al.
Published: (2024)
by: Nair, Pranav Ajit, et al.
Published: (2024)
Improving Generalization via Meta-Learning on Hard Samples
by: Jain, Nishant, et al.
Published: (2024)
by: Jain, Nishant, et al.
Published: (2024)
Second-order Optimization under Heavy-Tailed Noise: Hessian Clipping and Sample Complexity Limits
by: Sadiev, Abdurakhmon, et al.
Published: (2025)
by: Sadiev, Abdurakhmon, et al.
Published: (2025)
Dimension-free Score Matching and Time Bootstrapping for Diffusion Models
by: Kumar, Syamantak, et al.
Published: (2025)
by: Kumar, Syamantak, et al.
Published: (2025)
Sharp High-Probability Rates for Nonlinear SGD under Heavy-Tailed Noise via Symmetrization
by: Armacki, Aleksandar, et al.
Published: (2025)
by: Armacki, Aleksandar, et al.
Published: (2025)
DP-SGD Without Clipping: The Lipschitz Neural Network Way
by: Bethune, Louis, et al.
Published: (2023)
by: Bethune, Louis, et al.
Published: (2023)
Clipped Gradient Methods for Nonsmooth Convex Optimization under Heavy-Tailed Noise: A Refined Analysis
by: Liu, Zijian
Published: (2025)
by: Liu, Zijian
Published: (2025)
Second Order Methods for Bandit Optimization and Control
by: Suggala, Arun, et al.
Published: (2024)
by: Suggala, Arun, et al.
Published: (2024)
Similar Items
-
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
by: Vijayan, Sushant, et al.
Published: (2025) -
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
by: Kumar, Ramnath, et al.
Published: (2023) -
From Gradient Clipping to Normalization for Heavy Tailed SGD
by: Hübler, Florian, et al.
Published: (2024) -
Online Matrix Completion: A Collaborative Approach with Hott Items
by: Baby, Dheeraj, et al.
Published: (2024) -
Privacy of SGD under Gaussian or Heavy-Tailed Noise: Guarantees without Gradient Clipping
by: Şimşekli, Umut, et al.
Published: (2024)