Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sadiev, Abdurakhmon, Maranjyan, Artavazd, Ilin, Ivan, Richtárik, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ringmaster ASGD: The First Asynchronous SGD with Optimal Time Complexity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
Ringleader ASGD: The First Asynchronous SGD with Optimal Time Complexity under Data Heterogeneity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity
von: Mahran, Ammar, et al.
Veröffentlicht: (2026)
von: Mahran, Ammar, et al.
Veröffentlicht: (2026)
Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction
von: Tovmasyan, Zhirayr, et al.
Veröffentlicht: (2026)
von: Tovmasyan, Zhirayr, et al.
Veröffentlicht: (2026)
First Provably Optimal Asynchronous SGD for Homogeneous and Heterogeneous Data
von: Maranjyan, Artavazd
Veröffentlicht: (2026)
von: Maranjyan, Artavazd
Veröffentlicht: (2026)
GradSkip: Communication-Accelerated Local Gradient Methods with Better Computational Complexity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2022)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2022)
LoCoDL: Communication-Efficient Distributed Learning with Local Training and Compression
von: Condat, Laurent, et al.
Veröffentlicht: (2024)
von: Condat, Laurent, et al.
Veröffentlicht: (2024)
MindFlayer SGD: Efficient Parallel SGD in the Presence of Heterogeneous and Random Worker Compute Times
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging
von: Maziane, Yassine, et al.
Veröffentlicht: (2026)
von: Maziane, Yassine, et al.
Veröffentlicht: (2026)
ATA: Adaptive Task Allocation for Efficient Resource Management in Distributed Machine Learning
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025)
Differentially Private Random Block Coordinate Descent
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
Decentralized Personalized Federated Learning for Min-Max Problems
von: Borodich, Ekaterina, et al.
Veröffentlicht: (2021)
von: Borodich, Ekaterina, et al.
Veröffentlicht: (2021)
Towards a Better Theoretical Understanding of Independent Subnetwork Training
von: Shulgin, Egor, et al.
Veröffentlicht: (2023)
von: Shulgin, Egor, et al.
Veröffentlicht: (2023)
Better LMO-based Momentum Methods with Second-Order Information
von: Khirirat, Sarit, et al.
Veröffentlicht: (2025)
von: Khirirat, Sarit, et al.
Veröffentlicht: (2025)
On Biased Compression for Distributed Learning
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2020)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2020)
Birch SGD: A Tree Graph Framework for Local and Asynchronous SGD Methods
von: Tyurin, Alexander, et al.
Veröffentlicht: (2025)
von: Tyurin, Alexander, et al.
Veröffentlicht: (2025)
Smoothed Normalization for Efficient Distributed Private Optimization
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
Asynchronous Policy Gradient Aggregation for Efficient Distributed Reinforcement Learning
von: Tyurin, Alexander, et al.
Veröffentlicht: (2025)
von: Tyurin, Alexander, et al.
Veröffentlicht: (2025)
Decentralized Nonconvex Composite Federated Learning with Gradient Tracking and Momentum
von: Zhou, Yuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuan, et al.
Veröffentlicht: (2025)
A Bias-Correction Decentralized Stochastic Gradient Algorithm with Momentum Acceleration
von: Hu, Yuchen, et al.
Veröffentlicht: (2025)
von: Hu, Yuchen, et al.
Veröffentlicht: (2025)
MAST: Model-Agnostic Sparsified Training
von: Demidovich, Yury, et al.
Veröffentlicht: (2023)
von: Demidovich, Yury, et al.
Veröffentlicht: (2023)
Byzantine Robustness and Partial Participation Can Be Achieved at Once: Just Clip Gradient Differences
von: Malinovsky, Grigory, et al.
Veröffentlicht: (2023)
von: Malinovsky, Grigory, et al.
Veröffentlicht: (2023)
Correlated Quantization for Faster Nonconvex Distributed Optimization
von: Panferov, Andrei, et al.
Veröffentlicht: (2024)
von: Panferov, Andrei, et al.
Veröffentlicht: (2024)
Dynamic Regularized Sharpness Aware Minimization in Federated Learning: Approaching Global Consistency and Smooth Landscape
von: Sun, Yan, et al.
Veröffentlicht: (2023)
von: Sun, Yan, et al.
Veröffentlicht: (2023)
Local Methods with Adaptivity via Scaling
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2024)
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2024)
On Principled Local Optimization Methods for Federated Learning
von: Yuan, Honglin
Veröffentlicht: (2024)
von: Yuan, Honglin
Veröffentlicht: (2024)
A Privacy Preserving Randomized Gossip Algorithm via Controlled Noise Insertion
von: Hanzely, Filip, et al.
Veröffentlicht: (2019)
von: Hanzely, Filip, et al.
Veröffentlicht: (2019)
A Double Tracking Method for Optimization with Decentralized Generalized Orthogonality Constraints
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
A Penalty-Based Method for Communication-Efficient Decentralized Bilevel Programming
von: Nazari, Parvin, et al.
Veröffentlicht: (2022)
von: Nazari, Parvin, et al.
Veröffentlicht: (2022)
Accelerated Methods with Compressed Communications for Distributed Optimization Problems under Data Similarity
von: Bylinkin, Dmitry, et al.
Veröffentlicht: (2024)
von: Bylinkin, Dmitry, et al.
Veröffentlicht: (2024)
A Survey of Optimization Methods for Training DL Models: Theoretical Perspective on Convergence and Generalization
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
FADAS: Towards Federated Adaptive Asynchronous Optimization
von: Wang, Yujia, et al.
Veröffentlicht: (2024)
von: Wang, Yujia, et al.
Veröffentlicht: (2024)
A Hybrid Stochastic Gradient Tracking Method for Distributed Online Optimization Over Time-Varying Directed Networks
von: Shi, Xinli, et al.
Veröffentlicht: (2025)
von: Shi, Xinli, et al.
Veröffentlicht: (2025)
Optimizing Stochastic Gradient Push under Broadcast Communications
von: Nguyen, Tuan, et al.
Veröffentlicht: (2026)
von: Nguyen, Tuan, et al.
Veröffentlicht: (2026)
LoDAdaC: a unified local training-based decentralized framework with adaptive gradients and compressed communication
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
S$^3$LDBO: A Snapshot Single-Loop Algorithm for Decentralized Bilevel Optimization
von: Yin, Chao, et al.
Veröffentlicht: (2026)
von: Yin, Chao, et al.
Veröffentlicht: (2026)
FedCanon: Non-Convex Composite Federated Learning with Efficient Proximal Operation on Heterogeneous Data
von: Zhou, Yuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuan, et al.
Veröffentlicht: (2025)
Decentralized Directed Collaboration for Personalized Federated Learning
von: Liu, Yingqi, et al.
Veröffentlicht: (2024)
von: Liu, Yingqi, et al.
Veröffentlicht: (2024)
Stochastic Controlled Averaging for Federated Learning with Communication Compression
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
High-Performance Hybrid Algorithm for Minimum Sum-of-Squares Clustering of Infinitely Tall Data
von: Mussabayev, Ravil, et al.
Veröffentlicht: (2023)
von: Mussabayev, Ravil, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Ringmaster ASGD: The First Asynchronous SGD with Optimal Time Complexity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025) -
Ringleader ASGD: The First Asynchronous SGD with Optimal Time Complexity under Data Heterogeneity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2025) -
Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity
von: Mahran, Ammar, et al.
Veröffentlicht: (2026) -
Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction
von: Tovmasyan, Zhirayr, et al.
Veröffentlicht: (2026) -
First Provably Optimal Asynchronous SGD for Homogeneous and Heterogeneous Data
von: Maranjyan, Artavazd
Veröffentlicht: (2026)