Saved in:
| Main Authors: | Bui, Nghia, Wang, Lijing |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.27987 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing the Macro and Micro Effects of Random Seeds on Fine-Tuning Large Language Models
by: Bui, Nghia, et al.
Published: (2025)
by: Bui, Nghia, et al.
Published: (2025)
Continual Fine-Tuning with Provably Accurate and Parameter-Free Task Retrieval
by: Le, Hang Thi-Thuy, et al.
Published: (2026)
by: Le, Hang Thi-Thuy, et al.
Published: (2026)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Multi-kernel Correntropy-based Orientation Estimation of IMUs: Gradient Descent Methods
by: Li, Shilei, et al.
Published: (2023)
by: Li, Shilei, et al.
Published: (2023)
Gradient Descent with Provably Tuned Learning-rate Schedules
by: Sharma, Dravyansh
Published: (2025)
by: Sharma, Dravyansh
Published: (2025)
Antibody: Strengthening Defense Against Harmful Fine-Tuning for Large Language Models via Attenuating Harmful Gradient Influence
by: Nguyen, Quoc Minh, et al.
Published: (2026)
by: Nguyen, Quoc Minh, et al.
Published: (2026)
Unraveling the Gradient Descent Dynamics of Transformers
by: Song, Bingqing, et al.
Published: (2024)
by: Song, Bingqing, et al.
Published: (2024)
Generalization Guarantees on Data-Driven Tuning of Gradient Descent with Langevin Updates
by: Goyal, Saumya, et al.
Published: (2026)
by: Goyal, Saumya, et al.
Published: (2026)
Prompt Tuning for Natural Language to SQL with Embedding Fine-Tuning and RAG
by: Jang, Jisoo, et al.
Published: (2025)
by: Jang, Jisoo, et al.
Published: (2025)
Optimal Rates for Generalization of Gradient Descent for Deep ReLU Classification
by: Li, Yuanfan, et al.
Published: (2025)
by: Li, Yuanfan, et al.
Published: (2025)
Multiclass Loss Geometry Matters for Generalization of Gradient Descent in Separable Classification
by: Schliserman, Matan, et al.
Published: (2025)
by: Schliserman, Matan, et al.
Published: (2025)
Auditing Information Disclosure During LLM-Scale Gradient Descent Using Gradient Uniqueness
by: Abdelghafar, Sleem, et al.
Published: (2025)
by: Abdelghafar, Sleem, et al.
Published: (2025)
Convergence and Implicit Bias of Gradient Descent on Continual Linear Classification
by: Jung, Hyunji, et al.
Published: (2025)
by: Jung, Hyunji, et al.
Published: (2025)
Occam Gradient Descent
by: Kausik, B. N.
Published: (2024)
by: Kausik, B. N.
Published: (2024)
Gradient Descent Algorithm Survey
by: Fucheng, Deng, et al.
Published: (2025)
by: Fucheng, Deng, et al.
Published: (2025)
Scaling Law for Stochastic Gradient Descent in Quadratically Parameterized Linear Regression
by: Ding, Shihong, et al.
Published: (2025)
by: Ding, Shihong, et al.
Published: (2025)
Scaling Laws for Gradient Descent and Sign Descent for Linear Bigram Models under Zipf's Law
by: Kunstner, Frederik, et al.
Published: (2025)
by: Kunstner, Frederik, et al.
Published: (2025)
Modified Loss of Momentum Gradient Descent: Fine-Grained Analysis
by: Cattaneo, Matias D., et al.
Published: (2025)
by: Cattaneo, Matias D., et al.
Published: (2025)
Stochastic Adaptive Gradient Descent Without Descent
by: Aujol, Jean-François, et al.
Published: (2025)
by: Aujol, Jean-François, et al.
Published: (2025)
Save It All: Enabling Full Parameter Tuning for Federated Large Language Models via Cycle Block Gradient Descent
by: Wang, Lin, et al.
Published: (2024)
by: Wang, Lin, et al.
Published: (2024)
Stacking as Accelerated Gradient Descent
by: Agarwal, Naman, et al.
Published: (2024)
by: Agarwal, Naman, et al.
Published: (2024)
Grams: Gradient Descent with Adaptive Momentum Scaling
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Corner Gradient Descent
by: Yarotsky, Dmitry
Published: (2025)
by: Yarotsky, Dmitry
Published: (2025)
Adjacent Leader Decentralized Stochastic Gradient Descent
by: He, Haoze, et al.
Published: (2024)
by: He, Haoze, et al.
Published: (2024)
Turning Stale Gradients into Stable Gradients: Coherent Coordinate Descent with Implicit Landscape Smoothing for Lightweight Zeroth-Order Optimization
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
Curl Descent: Non-Gradient Learning Dynamics with Sign-Diverse Plasticity
by: Ninou, Hugo, et al.
Published: (2025)
by: Ninou, Hugo, et al.
Published: (2025)
Dynamic Decoupling of Placid Terminal Attractor-based Gradient Descent Algorithm
by: Zhao, Jinwei, et al.
Published: (2024)
by: Zhao, Jinwei, et al.
Published: (2024)
Sparse Gradient Compression for Fine-Tuning Large Language Models
by: Yang, David H., et al.
Published: (2025)
by: Yang, David H., et al.
Published: (2025)
Light-FMP: Lightweight Feature and Model Pruning for Enhanced Deep Recommender Systems
by: Bui, Nghia, et al.
Published: (2026)
by: Bui, Nghia, et al.
Published: (2026)
Rethinking Cross-Modal Fine-Tuning: Optimizing the Interaction Between Feature Alignment and Target Fitting
by: Tran, Trong Khiem, et al.
Published: (2026)
by: Tran, Trong Khiem, et al.
Published: (2026)
Parameter-Efficient Fine-Tuning of DINOv2 for Large-Scale Font Classification
by: Chen, Daniel, et al.
Published: (2026)
by: Chen, Daniel, et al.
Published: (2026)
Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent
by: Hennick, Max, et al.
Published: (2025)
by: Hennick, Max, et al.
Published: (2025)
Training Instabilities Induce Flatness Bias in Gradient Descent
by: Wang, Lawrence, et al.
Published: (2025)
by: Wang, Lawrence, et al.
Published: (2025)
Distributed Gradient Descent for Functional Learning
by: Yu, Zhan, et al.
Published: (2023)
by: Yu, Zhan, et al.
Published: (2023)
Robust Gradient Descent for Phase Retrieval
by: Buna, Alex, et al.
Published: (2024)
by: Buna, Alex, et al.
Published: (2024)
On the Generalization of Stochastic Gradient Descent with Momentum
by: Ramezani-Kebrya, Ali, et al.
Published: (2018)
by: Ramezani-Kebrya, Ali, et al.
Published: (2018)
First-ish Order Methods: Hessian-aware Scalings of Gradient Descent
by: Smee, Oscar, et al.
Published: (2025)
by: Smee, Oscar, et al.
Published: (2025)
Limited Memory Online Gradient Descent for Kernelized Pairwise Learning with Dynamic Averaging
by: AlQuabeh, Hilal, et al.
Published: (2024)
by: AlQuabeh, Hilal, et al.
Published: (2024)
The Implicit Bias of Gradient Descent on Separable Multiclass Data
by: Ravi, Hrithik, et al.
Published: (2024)
by: Ravi, Hrithik, et al.
Published: (2024)
VFOG: Variance-Reduced Fast Optimistic Gradient Methods for a Class of Nonmonotone Generalized Equations
by: Tran-Dinh, Quoc, et al.
Published: (2025)
by: Tran-Dinh, Quoc, et al.
Published: (2025)
Similar Items
-
Assessing the Macro and Micro Effects of Random Seeds on Fine-Tuning Large Language Models
by: Bui, Nghia, et al.
Published: (2025) -
Continual Fine-Tuning with Provably Accurate and Parameter-Free Task Retrieval
by: Le, Hang Thi-Thuy, et al.
Published: (2026) -
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026) -
Multi-kernel Correntropy-based Orientation Estimation of IMUs: Gradient Descent Methods
by: Li, Shilei, et al.
Published: (2023) -
Gradient Descent with Provably Tuned Learning-rate Schedules
by: Sharma, Dravyansh
Published: (2025)