From Logistic Regression to the Perceptron Algorithm: Exploring Gradient Descent with Large Step Sizes
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Tyurin, Alexander |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration
par: Tyurin, Alexander
Publié: (2025)
par: Tyurin, Alexander
Publié: (2025)
Gradient Descent on Logistic Regression with Non-Separable Data and Large Step Sizes
par: Meng, Si Yi, et autres
Publié: (2024)
par: Meng, Si Yi, et autres
Publié: (2024)
Gradient Descent on Logistic Regression: Do Large Step-Sizes Work with Data on the Sphere?
par: Meng, Si Yi, et autres
Publié: (2025)
par: Meng, Si Yi, et autres
Publié: (2025)
Large Stepsizes Accelerate Gradient Descent for Regularized Logistic Regression
par: Wu, Jingfeng, et autres
Publié: (2025)
par: Wu, Jingfeng, et autres
Publié: (2025)
Tight Bounds for Logistic Regression with Large Stepsize Gradient Descent in Low Dimension
par: Crawshaw, Michael, et autres
Publié: (2026)
par: Crawshaw, Michael, et autres
Publié: (2026)
Benefits of Early Stopping in Gradient Descent for Overparameterized Logistic Regression
par: Wu, Jingfeng, et autres
Publié: (2025)
par: Wu, Jingfeng, et autres
Publié: (2025)
Minimax Optimal Convergence of Gradient Descent in Logistic Regression via Large and Adaptive Stepsizes
par: Zhang, Ruiqi, et autres
Publié: (2025)
par: Zhang, Ruiqi, et autres
Publié: (2025)
Gradient Descent with Large Step Sizes: Chaos and Fractal Convergence Region
par: Liang, Shuang, et autres
Publié: (2025)
par: Liang, Shuang, et autres
Publié: (2025)
Exponential Convergence of (Stochastic) Gradient Descent for Separable Logistic Regression
par: Kale, Sacchit, et autres
Publié: (2026)
par: Kale, Sacchit, et autres
Publié: (2026)
Near-Optimal Convergence of Accelerated Gradient Methods under Generalized and $(L_0, L_1)$-Smoothness
par: Tyurin, Alexander
Publié: (2025)
par: Tyurin, Alexander
Publié: (2025)
Transformers Efficiently Perform In-Context Logistic Regression via Normalized Gradient Descent
par: Zhang, Chenyang, et autres
Publié: (2026)
par: Zhang, Chenyang, et autres
Publié: (2026)
Adaptive Step Sizes for Preconditioned Stochastic Gradient Descent
par: Köhne, Frederik, et autres
Publié: (2023)
par: Köhne, Frederik, et autres
Publié: (2023)
Large Stepsize Gradient Descent for Logistic Loss: Non-Monotonicity of the Loss Improves Optimization Efficiency
par: Wu, Jingfeng, et autres
Publié: (2024)
par: Wu, Jingfeng, et autres
Publié: (2024)
Exploring the Optimized Value of Each Hyperparameter in Various Gradient Descent Algorithms
par: Chen, Abel C. H.
Publié: (2022)
par: Chen, Abel C. H.
Publié: (2022)
Stochastic Gradient Descent for Nonparametric Additive Regression
par: Chen, Xin, et autres
Publié: (2024)
par: Chen, Xin, et autres
Publié: (2024)
Learning Curves of Stochastic Gradient Descent in Kernel Regression
par: Zhang, Haihan, et autres
Publié: (2025)
par: Zhang, Haihan, et autres
Publié: (2025)
Step by Step: Adaptive Gradient Descent for Training L-Lipschitz Neural Networks
par: Sung, Kyle, et autres
Publié: (2025)
par: Sung, Kyle, et autres
Publié: (2025)
Proving the Limited Scalability of Centralized Distributed Optimization via a New Lower Bound Construction
par: Tyurin, Alexander
Publié: (2025)
par: Tyurin, Alexander
Publié: (2025)
Gradient Descent Algorithm Survey
par: Fucheng, Deng, et autres
Publié: (2025)
par: Fucheng, Deng, et autres
Publié: (2025)
Asynchronous Policy Gradient Aggregation for Efficient Distributed Reinforcement Learning
par: Tyurin, Alexander, et autres
Publié: (2025)
par: Tyurin, Alexander, et autres
Publié: (2025)
Local Steps Speed Up Local GD for Heterogeneous Distributed Logistic Regression
par: Crawshaw, Michael, et autres
Publié: (2025)
par: Crawshaw, Michael, et autres
Publié: (2025)
Toward a Unified Theory of Gradient Descent under Generalized Smoothness
par: Tyurin, Alexander
Publié: (2024)
par: Tyurin, Alexander
Publié: (2024)
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization
par: Abuduweili, Abulikemu, et autres
Publié: (2024)
par: Abuduweili, Abulikemu, et autres
Publié: (2024)
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
par: Tyurin, Alexander, et autres
Publié: (2022)
par: Tyurin, Alexander, et autres
Publié: (2022)
Freya PAGE: First Optimal Time Complexity for Large-Scale Nonconvex Finite-Sum Optimization with Heterogeneous Asynchronous Computations
par: Tyurin, Alexander, et autres
Publié: (2024)
par: Tyurin, Alexander, et autres
Publié: (2024)
Trained Mamba Emulates Online Gradient Descent in In-Context Linear Regression
par: Jiang, Jiarui, et autres
Publié: (2025)
par: Jiang, Jiarui, et autres
Publié: (2025)
Scaling Law for Stochastic Gradient Descent in Quadratically Parameterized Linear Regression
par: Ding, Shihong, et autres
Publié: (2025)
par: Ding, Shihong, et autres
Publié: (2025)
Increasing Batch Size Improves Convergence of Stochastic Gradient Descent with Momentum
par: Kamo, Keisuke, et autres
Publié: (2025)
par: Kamo, Keisuke, et autres
Publié: (2025)
Privacy-Preserving Logistic Regression Training with A Faster Gradient Variant
par: Chiang, John
Publié: (2022)
par: Chiang, John
Publié: (2022)
Enhancing Policy Gradient with the Polyak Step-Size Adaption
par: Li, Yunxiang, et autres
Publié: (2024)
par: Li, Yunxiang, et autres
Publié: (2024)
Provably Faster Gradient Descent via Long Steps
par: Grimmer, Benjamin
Publié: (2023)
par: Grimmer, Benjamin
Publié: (2023)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
par: Lei, Yunwen, et autres
Publié: (2026)
par: Lei, Yunwen, et autres
Publié: (2026)
Relationship between Batch Size and Number of Steps Needed for Nonconvex Optimization of Stochastic Gradient Descent using Armijo Line Search
par: Tsukada, Yuki, et autres
Publié: (2023)
par: Tsukada, Yuki, et autres
Publié: (2023)
Complex Equation Learner: Rational Symbolic Regression with Gradient Descent in Complex Domain
par: Garmaev, Sergei, et autres
Publié: (2026)
par: Garmaev, Sergei, et autres
Publié: (2026)
FIRAL: An Active Learning Algorithm for Multinomial Logistic Regression
par: Chen, Youguang, et autres
Publié: (2024)
par: Chen, Youguang, et autres
Publié: (2024)
PMGDA: A Preference-based Multiple Gradient Descent Algorithm
par: Zhang, Xiaoyuan, et autres
Publié: (2024)
par: Zhang, Xiaoyuan, et autres
Publié: (2024)
Data-Driven Logistic Regression Ensembles With Applications in Genomics
par: Christidis, Anthony-Alexander, et autres
Publié: (2021)
par: Christidis, Anthony-Alexander, et autres
Publié: (2021)
A Provably Accurate Randomized Sampling Algorithm for Logistic Regression
par: Chowdhury, Agniva, et autres
Publié: (2024)
par: Chowdhury, Agniva, et autres
Publié: (2024)
Faster Convergence of Riemannian Stochastic Gradient Descent with Increasing Batch Size
par: Oowada, Kanata, et autres
Publié: (2025)
par: Oowada, Kanata, et autres
Publié: (2025)
Effectiveness of Distributed Gradient Descent with Local Steps for Overparameterized Models
par: Zhu, Heng, et autres
Publié: (2024)
par: Zhu, Heng, et autres
Publié: (2024)
Documents similaires
-
Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration
par: Tyurin, Alexander
Publié: (2025) -
Gradient Descent on Logistic Regression with Non-Separable Data and Large Step Sizes
par: Meng, Si Yi, et autres
Publié: (2024) -
Gradient Descent on Logistic Regression: Do Large Step-Sizes Work with Data on the Sphere?
par: Meng, Si Yi, et autres
Publié: (2025) -
Large Stepsizes Accelerate Gradient Descent for Regularized Logistic Regression
par: Wu, Jingfeng, et autres
Publié: (2025) -
Tight Bounds for Logistic Regression with Large Stepsize Gradient Descent in Low Dimension
par: Crawshaw, Michael, et autres
Publié: (2026)