Saved in:
| Main Authors: | Seung, Hyunseok, Lee, Jaewoo, Ko, Hyunsuk |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.08464 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NysAct: A Scalable Preconditioned Gradient Descent using Nystrom Approximation
by: Seung, Hyunseok, et al.
Published: (2025)
by: Seung, Hyunseok, et al.
Published: (2025)
Low-Rank Curvature for Zeroth-Order Optimization in LLM Fine-Tuning
by: Seung, Hyunseok, et al.
Published: (2025)
by: Seung, Hyunseok, et al.
Published: (2025)
An Adaptive Method Stabilizing Activations for Enhanced Generalization
by: Seung, Hyunseok, et al.
Published: (2025)
by: Seung, Hyunseok, et al.
Published: (2025)
Sign Gradient Descent-based Neuronal Dynamics: ANN-to-SNN Conversion Beyond ReLU Network
by: Oh, Hyunseok, et al.
Published: (2024)
by: Oh, Hyunseok, et al.
Published: (2024)
Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation
by: Hwang, Jisung, et al.
Published: (2026)
by: Hwang, Jisung, et al.
Published: (2026)
On the Nystrom Approximation for Preconditioning in Kernel Machines
by: Abedsoltan, Amirhesam, et al.
Published: (2023)
by: Abedsoltan, Amirhesam, et al.
Published: (2023)
A More Accurate Approximation of Activation Function with Few Spikes Neurons
by: Jeong, Dayena, et al.
Published: (2024)
by: Jeong, Dayena, et al.
Published: (2024)
Preconditioned DeltaNet: Curvature-aware Sequence Modeling for Linear Recurrences
by: Tumma, Neehal, et al.
Published: (2026)
by: Tumma, Neehal, et al.
Published: (2026)
Mousse: Rectifying the Geometry of Muon with Curvature-Aware Preconditioning
by: Zhang, Yechen, et al.
Published: (2026)
by: Zhang, Yechen, et al.
Published: (2026)
A Stein Gradient Descent Approach for Doubly Intractable Distributions
by: Lee, Heesang, et al.
Published: (2024)
by: Lee, Heesang, et al.
Published: (2024)
Stochastic Gradient Methods with Preconditioned Updates
by: Sadiev, Abdurakhmon, et al.
Published: (2022)
by: Sadiev, Abdurakhmon, et al.
Published: (2022)
Papez: Resource-Efficient Speech Separation with Auditory Working Memory
by: Oh, Hyunseok, et al.
Published: (2024)
by: Oh, Hyunseok, et al.
Published: (2024)
PROMISE: Preconditioned Stochastic Optimization Methods by Incorporating Scalable Curvature Estimates
by: Frangella, Zachary, et al.
Published: (2023)
by: Frangella, Zachary, et al.
Published: (2023)
Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime
by: Ghane, Reza, et al.
Published: (2026)
by: Ghane, Reza, et al.
Published: (2026)
Mirror and Preconditioned Gradient Descent in Wasserstein Space
by: Bonet, Clément, et al.
Published: (2024)
by: Bonet, Clément, et al.
Published: (2024)
Preconditioning for Accelerated Gradient Descent Optimization and Regularization
by: Ye, Qiang
Published: (2024)
by: Ye, Qiang
Published: (2024)
Efficient Curvature-Aware Hypergradient Approximation for Bilevel Optimization
by: Dong, Youran, et al.
Published: (2025)
by: Dong, Youran, et al.
Published: (2025)
Efficient Low-Tubal-Rank Tensor Estimation via Alternating Preconditioned Gradient Descent
by: Liu, Zhiyu, et al.
Published: (2025)
by: Liu, Zhiyu, et al.
Published: (2025)
Diffusion Fine-Tuning via Reparameterized Policy Gradient of the Soft Q-Function
by: Kang, Hyeongyu, et al.
Published: (2025)
by: Kang, Hyeongyu, et al.
Published: (2025)
Adaptive Step Sizes for Preconditioned Stochastic Gradient Descent
by: Köhne, Frederik, et al.
Published: (2023)
by: Köhne, Frederik, et al.
Published: (2023)
Mitigating Quantization Errors Due to Activation Spikes in GLU-Based LLMs
by: Yang, Jaewoo, et al.
Published: (2024)
by: Yang, Jaewoo, et al.
Published: (2024)
DyKAF: Dynamical Kronecker Approximation of the Fisher Information Matrix for Gradient Preconditioning
by: Yudin, Nikolay, et al.
Published: (2025)
by: Yudin, Nikolay, et al.
Published: (2025)
Scalable Mean-Field Variational Inference via Preconditioned Primal-Dual Optimization
by: Lyu, Jinhua, et al.
Published: (2026)
by: Lyu, Jinhua, et al.
Published: (2026)
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
by: Lee, Hyunseok, et al.
Published: (2024)
by: Lee, Hyunseok, et al.
Published: (2024)
Geometric Embedding Alignment via Curvature Matching in Transfer Learning
by: Ko, Sung Moon, et al.
Published: (2025)
by: Ko, Sung Moon, et al.
Published: (2025)
Preconditioned Gradient Descent for Over-Parameterized Nonconvex Matrix Factorization
by: Zhang, Gavin, et al.
Published: (2025)
by: Zhang, Gavin, et al.
Published: (2025)
Efficient Over-parameterized Matrix Sensing from Noisy Measurements via Alternating Preconditioned Gradient Descent
by: Liu, Zhiyu, et al.
Published: (2025)
by: Liu, Zhiyu, et al.
Published: (2025)
Deep Learning-Enhanced Preconditioning for Efficient Conjugate Gradient Solvers in Large-Scale PDE Systems
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
AGD: an Auto-switchable Optimizer using Stepwise Gradient Difference for Preconditioning Matrix
by: Yue, Yun, et al.
Published: (2023)
by: Yue, Yun, et al.
Published: (2023)
Gradient Flows for Sampling: Mean-Field Models, Gaussian Approximations and Affine Invariance
by: Chen, Yifan, et al.
Published: (2023)
by: Chen, Yifan, et al.
Published: (2023)
Ginger: An Efficient Curvature Approximation with Linear Complexity for General Neural Networks
by: Hao, Yongchang, et al.
Published: (2024)
by: Hao, Yongchang, et al.
Published: (2024)
Inverting the Leverage Score Gradient: An Efficient Approximate Newton Method
by: Li, Chenyang, et al.
Published: (2024)
by: Li, Chenyang, et al.
Published: (2024)
Kronecker-factored Approximate Curvature (KFAC) From Scratch
by: Dangel, Felix, et al.
Published: (2025)
by: Dangel, Felix, et al.
Published: (2025)
How many classifiers do we need?
by: Kim, Hyunsuk, et al.
Published: (2024)
by: Kim, Hyunsuk, et al.
Published: (2024)
ConditionNET: Learning Preconditions and Effects for Execution Monitoring
by: Sliwowski, Daniel, et al.
Published: (2025)
by: Sliwowski, Daniel, et al.
Published: (2025)
Leveraging Operator Learning to Accelerate Convergence of the Preconditioned Conjugate Gradient Method
by: Kopaničáková, Alena, et al.
Published: (2025)
by: Kopaničáková, Alena, et al.
Published: (2025)
Beyond Correctness: Learning Robust Reasoning via Transfer
by: Lee, Hyunseok, et al.
Published: (2026)
by: Lee, Hyunseok, et al.
Published: (2026)
Efficient Search for Customized Activation Functions with Gradient Descent
by: Strack, Lukas, et al.
Published: (2024)
by: Strack, Lukas, et al.
Published: (2024)
Posterior Approximation using Stochastic Gradient Ascent with Adaptive Stepsize
by: Lim, Kart-Leong, et al.
Published: (2024)
by: Lim, Kart-Leong, et al.
Published: (2024)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
by: Song, Chihyeon, et al.
Published: (2025)
by: Song, Chihyeon, et al.
Published: (2025)
Similar Items
-
NysAct: A Scalable Preconditioned Gradient Descent using Nystrom Approximation
by: Seung, Hyunseok, et al.
Published: (2025) -
Low-Rank Curvature for Zeroth-Order Optimization in LLM Fine-Tuning
by: Seung, Hyunseok, et al.
Published: (2025) -
An Adaptive Method Stabilizing Activations for Enhanced Generalization
by: Seung, Hyunseok, et al.
Published: (2025) -
Sign Gradient Descent-based Neuronal Dynamics: ANN-to-SNN Conversion Beyond ReLU Network
by: Oh, Hyunseok, et al.
Published: (2024) -
Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation
by: Hwang, Jisung, et al.
Published: (2026)