Gespeichert in:
| Hauptverfasser: | Bu, Zhiqi, Xu, Shiyun, Mao, Jialin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.07145 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding Forgetting in LLM Supervised Fine-Tuning and Preference Learning -- A Convex Optimization Perspective
von: Fernando, Heshan, et al.
Veröffentlicht: (2024)
von: Fernando, Heshan, et al.
Veröffentlicht: (2024)
Gradient descent with generalized Newton's method
von: Bu, Zhiqi, et al.
Veröffentlicht: (2024)
von: Bu, Zhiqi, et al.
Veröffentlicht: (2024)
Variational Learning is Effective for Large Deep Networks
von: Shen, Yuesong, et al.
Veröffentlicht: (2024)
von: Shen, Yuesong, et al.
Veröffentlicht: (2024)
Grokking or Glitching? How Low-Precision Drives Slingshot Loss Spikes
von: Hanqing, Liu, et al.
Veröffentlicht: (2026)
von: Hanqing, Liu, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning: A Convex Optimization Approach
von: Gattami, Ather
Veröffentlicht: (2024)
von: Gattami, Ather
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback
von: Wang, Yikai, et al.
Veröffentlicht: (2026)
von: Wang, Yikai, et al.
Veröffentlicht: (2026)
The Surprising Agreement Between Convex Optimization Theory and Learning-Rate Scheduling for Large Model Training
von: Schaipp, Fabian, et al.
Veröffentlicht: (2025)
von: Schaipp, Fabian, et al.
Veröffentlicht: (2025)
ACING: Actor-Critic for Instruction Learning in Black-Box LLMs
von: Kharrat, Salma, et al.
Veröffentlicht: (2024)
von: Kharrat, Salma, et al.
Veröffentlicht: (2024)
Muon in Associative Memory Learning: Training Dynamics and Scaling Laws
von: Li, Binghui, et al.
Veröffentlicht: (2026)
von: Li, Binghui, et al.
Veröffentlicht: (2026)
Reinforcement Learning from Human Feedback with Active Queries
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
In-Context Learning with Representations: Contextual Generalization of Trained Transformers
von: Yang, Tong, et al.
Veröffentlicht: (2024)
von: Yang, Tong, et al.
Veröffentlicht: (2024)
CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural Networks
von: Feng, Miria, et al.
Veröffentlicht: (2024)
von: Feng, Miria, et al.
Veröffentlicht: (2024)
When and How Unlabeled Data Provably Improve In-Context Learning
von: Li, Yingcong, et al.
Veröffentlicht: (2025)
von: Li, Yingcong, et al.
Veröffentlicht: (2025)
Online Learning on Hidden-Convex Losses via Algorithmic Equivalence: Optimal Regret, Geometric Barrier, and Bandit Feedback
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
Gating is Weighting: Understanding Gated Linear Attention through In-context Learning
von: Li, Yingcong, et al.
Veröffentlicht: (2025)
von: Li, Yingcong, et al.
Veröffentlicht: (2025)
AutoGD: Automatic Learning Rate Selection for Gradient Descent
von: Surjanovic, Nikola, et al.
Veröffentlicht: (2025)
von: Surjanovic, Nikola, et al.
Veröffentlicht: (2025)
Optimal Rates for Robust Stochastic Convex Optimization
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
Learning to optimize: A tutorial for continuous and mixed-integer optimization
von: Chen, Xiaohan, et al.
Veröffentlicht: (2024)
von: Chen, Xiaohan, et al.
Veröffentlicht: (2024)
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training
von: Liu, Hong, et al.
Veröffentlicht: (2023)
von: Liu, Hong, et al.
Veröffentlicht: (2023)
A Unified Understanding of Offline Data Selection and Online Self-refining Generation for Post-training LLMs
von: Xiao, Quan, et al.
Veröffentlicht: (2025)
von: Xiao, Quan, et al.
Veröffentlicht: (2025)
AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent
von: Surjanovic, Nikola, et al.
Veröffentlicht: (2025)
von: Surjanovic, Nikola, et al.
Veröffentlicht: (2025)
CLASP: An online learning algorithm for Convex Losses And Squared Penalties
von: Ferreira, Ricardo N., et al.
Veröffentlicht: (2026)
von: Ferreira, Ricardo N., et al.
Veröffentlicht: (2026)
Private Federated Learning Without a Trusted Server: Optimal Algorithms for Convex Losses
von: Lowy, Andrew, et al.
Veröffentlicht: (2021)
von: Lowy, Andrew, et al.
Veröffentlicht: (2021)
Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
von: Yu, Dingzhi, et al.
Veröffentlicht: (2026)
von: Yu, Dingzhi, et al.
Veröffentlicht: (2026)
Adapprox: Adaptive Approximation in Adam Optimization via Randomized Low-Rank Matrices
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2024)
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2024)
COS-DPO: Conditioned One-Shot Multi-Objective Fine-Tuning Framework
von: Ren, Yinuo, et al.
Veröffentlicht: (2024)
von: Ren, Yinuo, et al.
Veröffentlicht: (2024)
Distributional Surgery for Language Model Activations
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
FOCUS: First Order Concentrated Updating Scheme
von: Liu, Yizhou, et al.
Veröffentlicht: (2025)
von: Liu, Yizhou, et al.
Veröffentlicht: (2025)
Heavy-Tailed Class Imbalance and Why Adam Outperforms Gradient Descent on Language Models
von: Kunstner, Frederik, et al.
Veröffentlicht: (2024)
von: Kunstner, Frederik, et al.
Veröffentlicht: (2024)
SUMO: Subspace-Aware Moment-Orthogonalization for Accelerating Memory-Efficient LLM Training
von: Refael, Yehonathan, et al.
Veröffentlicht: (2025)
von: Refael, Yehonathan, et al.
Veröffentlicht: (2025)
Accelerated Rates between Stochastic and Adversarial Online Convex Optimization
von: Sachs, Sarah, et al.
Veröffentlicht: (2023)
von: Sachs, Sarah, et al.
Veröffentlicht: (2023)
Interpreting Adaptive Gradient Methods by Parameter Scaling for Learning-Rate-Free Optimization
von: Suh, Min-Kook, et al.
Veröffentlicht: (2024)
von: Suh, Min-Kook, et al.
Veröffentlicht: (2024)
BOOOM: Loss-Function-Agnostic Black-Box Optimization over Orthonormal Manifolds for Machine Learning and Statistical Inference
von: Kim, Beomchang, et al.
Veröffentlicht: (2026)
von: Kim, Beomchang, et al.
Veröffentlicht: (2026)
Operator Splitting for Learning to Predict Equilibria in Convex Games
von: McKenzie, Daniel, et al.
Veröffentlicht: (2021)
von: McKenzie, Daniel, et al.
Veröffentlicht: (2021)
First-Order Sparse Convex Optimization: Better Rates with Sparse Updates
von: Garber, Dan
Veröffentlicht: (2025)
von: Garber, Dan
Veröffentlicht: (2025)
Learning Algorithm Hyperparameters for Fast Parametric Convex Optimization
von: Sambharya, Rajiv, et al.
Veröffentlicht: (2024)
von: Sambharya, Rajiv, et al.
Veröffentlicht: (2024)
Learning-Augmented Decentralized Online Convex Optimization in Networks
von: Li, Pengfei, et al.
Veröffentlicht: (2023)
von: Li, Pengfei, et al.
Veröffentlicht: (2023)
Universal Architectures for the Learning of Polyhedral Norms and Convex Regularizers
von: Unser, Michael, et al.
Veröffentlicht: (2025)
von: Unser, Michael, et al.
Veröffentlicht: (2025)
Online (Non-)Convex Learning via Tempered Optimism
von: Haddouche, Maxime, et al.
Veröffentlicht: (2023)
von: Haddouche, Maxime, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Understanding Forgetting in LLM Supervised Fine-Tuning and Preference Learning -- A Convex Optimization Perspective
von: Fernando, Heshan, et al.
Veröffentlicht: (2024) -
Gradient descent with generalized Newton's method
von: Bu, Zhiqi, et al.
Veröffentlicht: (2024) -
Variational Learning is Effective for Large Deep Networks
von: Shen, Yuesong, et al.
Veröffentlicht: (2024) -
Grokking or Glitching? How Low-Precision Drives Slingshot Loss Spikes
von: Hanqing, Liu, et al.
Veröffentlicht: (2026) -
Deep Reinforcement Learning: A Convex Optimization Approach
von: Gattami, Ather
Veröffentlicht: (2024)