A Novel Unified Parametric Assumption for Nonconvex Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Riabinin, Artem, Khaled, Ahmed, Richtárik, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
von: Riabinin, Artem, et al.
Veröffentlicht: (2025)
von: Riabinin, Artem, et al.
Veröffentlicht: (2025)
Correlated Quantization for Faster Nonconvex Distributed Optimization
von: Panferov, Andrei, et al.
Veröffentlicht: (2024)
von: Panferov, Andrei, et al.
Veröffentlicht: (2024)
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
von: Tyurin, Alexander, et al.
Veröffentlicht: (2022)
von: Tyurin, Alexander, et al.
Veröffentlicht: (2022)
Improving the Worst-Case Bidirectional Communication Complexity for Nonconvex Distributed Optimization under Function Similarity
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2024)
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2024)
Freya PAGE: First Optimal Time Complexity for Large-Scale Nonconvex Finite-Sum Optimization with Heterogeneous Asynchronous Computations
von: Tyurin, Alexander, et al.
Veröffentlicht: (2024)
von: Tyurin, Alexander, et al.
Veröffentlicht: (2024)
Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization
von: Xiao, Nachuan, et al.
Veröffentlicht: (2023)
von: Xiao, Nachuan, et al.
Veröffentlicht: (2023)
Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers
von: Riabinin, Artem, et al.
Veröffentlicht: (2026)
von: Riabinin, Artem, et al.
Veröffentlicht: (2026)
First Provable Guarantees for Practical Private FL: Beyond Restrictive Assumptions
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
von: Shulgin, Egor, et al.
Veröffentlicht: (2025)
Provable Acceleration for Diffusion Models under Minimal Assumptions
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Parametric Nonconvex Optimization via Convex Surrogates
von: Wang, Renzi, et al.
Veröffentlicht: (2026)
von: Wang, Renzi, et al.
Veröffentlicht: (2026)
A Unified Framework for Gradient Aggregation in Multi-Objective Optimization
von: Hu, Zeou, et al.
Veröffentlicht: (2026)
von: Hu, Zeou, et al.
Veröffentlicht: (2026)
Enhancing Stochastic Gradient Descent: A Unified Framework and Novel Acceleration Methods for Faster Convergence
von: Deng, Yichuan, et al.
Veröffentlicht: (2024)
von: Deng, Yichuan, et al.
Veröffentlicht: (2024)
A Single-Loop Gradient Descent and Perturbed Ascent Algorithm for Nonconvex Functional Constrained Optimization
von: Lu, Songtao
Veröffentlicht: (2022)
von: Lu, Songtao
Veröffentlicht: (2022)
Non-Euclidean Broximal Point Method: A Blueprint for Geometry-Aware Optimization
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2025)
Byzantine-Robust and Differentially Private Federated Optimization under Weaker Assumptions
von: Islamov, Rustem, et al.
Veröffentlicht: (2026)
von: Islamov, Rustem, et al.
Veröffentlicht: (2026)
A Unified Theory of Stochastic Proximal Point Methods without Smoothness
von: Richtárik, Peter, et al.
Veröffentlicht: (2024)
von: Richtárik, Peter, et al.
Veröffentlicht: (2024)
The Road Less Scheduled
von: Defazio, Aaron, et al.
Veröffentlicht: (2024)
von: Defazio, Aaron, et al.
Veröffentlicht: (2024)
Convergence Analysis of the PAGE Stochastic Algorithm for Weakly Convex Finite-Sum Optimization
von: Condat, Laurent, et al.
Veröffentlicht: (2025)
von: Condat, Laurent, et al.
Veröffentlicht: (2025)
MARINA-P: Superior Performance in Non-smooth Federated Optimization with Adaptive Stepsizes
von: Sokolov, Igor, et al.
Veröffentlicht: (2024)
von: Sokolov, Igor, et al.
Veröffentlicht: (2024)
$γ$-weakly $θ$-up-concavity: A Unified Framework for Non-Convex Optimization Beyond DR-Submodular and OSS Functions
von: Pedramfar, Mohammad, et al.
Veröffentlicht: (2026)
von: Pedramfar, Mohammad, et al.
Veröffentlicht: (2026)
Beyond Minimax Rates in Group Distributionally Robust Optimization via a Novel Notion of Sparsity
von: Nguyen, Quan, et al.
Veröffentlicht: (2024)
von: Nguyen, Quan, et al.
Veröffentlicht: (2024)
Robust Second-Order Nonconvex Optimization and Its Application to Low Rank Matrix Sensing
von: Li, Shuyao, et al.
Veröffentlicht: (2024)
von: Li, Shuyao, et al.
Veröffentlicht: (2024)
MAST: Model-Agnostic Sparsified Training
von: Demidovich, Yury, et al.
Veröffentlicht: (2023)
von: Demidovich, Yury, et al.
Veröffentlicht: (2023)
Byzantine Robustness and Partial Participation Can Be Achieved at Once: Just Clip Gradient Differences
von: Malinovsky, Grigory, et al.
Veröffentlicht: (2023)
von: Malinovsky, Grigory, et al.
Veröffentlicht: (2023)
Dynamic Memory Based Adaptive Optimization
von: Szegedy, Balázs, et al.
Veröffentlicht: (2024)
von: Szegedy, Balázs, et al.
Veröffentlicht: (2024)
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making
von: Asru, Avijit Saha, et al.
Veröffentlicht: (2025)
von: Asru, Avijit Saha, et al.
Veröffentlicht: (2025)
Unlocking FedNL: Self-Contained Compute-Optimized Implementation
von: Burlachenko, Konstantin, et al.
Veröffentlicht: (2024)
von: Burlachenko, Konstantin, et al.
Veröffentlicht: (2024)
Global Convergence and Rich Feature Learning in $L$-Layer Infinite-Width Neural Networks under $μ$P Parametrization
von: Chen, Zixiang, et al.
Veröffentlicht: (2025)
von: Chen, Zixiang, et al.
Veröffentlicht: (2025)
Local LMO: Constrained Gradient Optimization via a Local Linear Minimization Oracle
von: Richtárik, Peter, et al.
Veröffentlicht: (2026)
von: Richtárik, Peter, et al.
Veröffentlicht: (2026)
BiCoLoR: Communication-Efficient Optimization with Bidirectional Compression and Local Training
von: Condat, Laurent, et al.
Veröffentlicht: (2026)
von: Condat, Laurent, et al.
Veröffentlicht: (2026)
Second-order Optimization under Heavy-Tailed Noise: Hessian Clipping and Sample Complexity Limits
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
AdLoCo: adaptive batching significantly improves communications efficiency and convergence for Large Language Models
von: Kutuzov, Nikolay, et al.
Veröffentlicht: (2025)
von: Kutuzov, Nikolay, et al.
Veröffentlicht: (2025)
MetaOptimize: A Framework for Optimizing Step Sizes and Other Meta-parameters
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2024)
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2024)
EXAdam: The Power of Adaptive Cross-Moments
von: Adly, Ahmed M.
Veröffentlicht: (2024)
von: Adly, Ahmed M.
Veröffentlicht: (2024)
Min-Max Optimisation for Nonconvex-Nonconcave Functions Using a Random Zeroth-Order Extragradient Algorithm
von: Farzin, Amir Ali, et al.
Veröffentlicht: (2025)
von: Farzin, Amir Ali, et al.
Veröffentlicht: (2025)
From Large Language Models and Optimization to Decision Optimization CoPilot: A Research Manifesto
von: Wasserkrug, Segev, et al.
Veröffentlicht: (2024)
von: Wasserkrug, Segev, et al.
Veröffentlicht: (2024)
Unified Projection-Free Algorithms for Adversarial DR-Submodular Optimization
von: Pedramfar, Mohammad, et al.
Veröffentlicht: (2024)
von: Pedramfar, Mohammad, et al.
Veröffentlicht: (2024)
Tuning-Free Stochastic Optimization
von: Khaled, Ahmed, et al.
Veröffentlicht: (2024)
von: Khaled, Ahmed, et al.
Veröffentlicht: (2024)
A Minimalist Bayesian Framework for Stochastic Optimization
von: Wang, Kaizheng
Veröffentlicht: (2025)
von: Wang, Kaizheng
Veröffentlicht: (2025)
Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks
von: Kiyani, Elham, et al.
Veröffentlicht: (2025)
von: Kiyani, Elham, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
von: Riabinin, Artem, et al.
Veröffentlicht: (2025) -
Correlated Quantization for Faster Nonconvex Distributed Optimization
von: Panferov, Andrei, et al.
Veröffentlicht: (2024) -
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
von: Tyurin, Alexander, et al.
Veröffentlicht: (2022) -
Improving the Worst-Case Bidirectional Communication Complexity for Nonconvex Distributed Optimization under Function Similarity
von: Gruntkowska, Kaja, et al.
Veröffentlicht: (2024) -
Freya PAGE: First Optimal Time Complexity for Large-Scale Nonconvex Finite-Sum Optimization with Heterogeneous Asynchronous Computations
von: Tyurin, Alexander, et al.
Veröffentlicht: (2024)