Saved in:
| Main Authors: | Lu, Yanqing, Wang, Letao, Liu, Jinbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.22938 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding and Improving Shampoo and SOAP via Kullback-Leibler Minimization
by: Lin, Wu, et al.
Published: (2025)
by: Lin, Wu, et al.
Published: (2025)
Zipfian Whitening
by: Yokoi, Sho, et al.
Published: (2024)
by: Yokoi, Sho, et al.
Published: (2024)
Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes
by: Kamal, Sadia, et al.
Published: (2025)
by: Kamal, Sadia, et al.
Published: (2025)
Whitening Not Recommended for Classification Tasks in LLMs
by: Forooghi, Ali, et al.
Published: (2024)
by: Forooghi, Ali, et al.
Published: (2024)
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)
by: Vyas, Nikhil, et al.
Published: (2024)
What Really Matters in Matrix-Whitening Optimizers?
by: Frans, Kevin, et al.
Published: (2025)
by: Frans, Kevin, et al.
Published: (2025)
Reparametrizing Shampoo and SOAP for Subspace Basis Updates and BFloat16 Storage
by: Milligan, Alan, et al.
Published: (2026)
by: Milligan, Alan, et al.
Published: (2026)
Whitening Spherical Gaussian Mixtures in the Large-Dimensional Regime
by: Boudjemaa, Mohammed Racim Moussa, et al.
Published: (2025)
by: Boudjemaa, Mohammed Racim Moussa, et al.
Published: (2025)
Data Whitening Improves Sparse Autoencoder Learning
by: Saraswatula, Ashwin, et al.
Published: (2025)
by: Saraswatula, Ashwin, et al.
Published: (2025)
A Stable Whitening Optimizer for Efficient Neural Network Training
by: Frans, Kevin, et al.
Published: (2025)
by: Frans, Kevin, et al.
Published: (2025)
Conditionally Whitened Generative Models for Probabilistic Time Series Forecasting
by: Yang, Yanfeng, et al.
Published: (2025)
by: Yang, Yanfeng, et al.
Published: (2025)
Improving Relative Representations with Learned Anchors and Whitened Inner Products
by: Svendsen, Oscar Thorsted, et al.
Published: (2026)
by: Svendsen, Oscar Thorsted, et al.
Published: (2026)
SOAP-RL: Sequential Option Advantage Propagation for Reinforcement Learning in POMDP Environments
by: Ishida, Shu, et al.
Published: (2024)
by: Ishida, Shu, et al.
Published: (2024)
SWAN: SGD with Normalization and Whitening Enables Stateless LLM Training
by: Ma, Chao, et al.
Published: (2024)
by: Ma, Chao, et al.
Published: (2024)
Pro-KLShampoo: Projected KL-Shampoo with Whitening Recovered by Orthogonalization
by: Sun, Ruotong, et al.
Published: (2026)
by: Sun, Ruotong, et al.
Published: (2026)
Exploring Robustness in Doctor-Patient Conversation Summarization: An Analysis of Out-of-Domain SOAP Notes
by: Chen, Yu-Wen, et al.
Published: (2024)
by: Chen, Yu-Wen, et al.
Published: (2024)
Training-Free Policy Violation Detection via Activation-Space Whitening in LLMs
by: Rachmil, Oren, et al.
Published: (2025)
by: Rachmil, Oren, et al.
Published: (2025)
Out-of-Distribution Radar Detection with Complex VAEs: Theory, Whitening, and ANMF Fusion
by: Rouzoumka, Yadang Alexis, et al.
Published: (2026)
by: Rouzoumka, Yadang Alexis, et al.
Published: (2026)
IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
An Integrated Fusion Framework for Ensemble Learning Leveraging Gradient Boosting and Fuzzy Rule-Based Models
by: Li, Jinbo, et al.
Published: (2025)
by: Li, Jinbo, et al.
Published: (2025)
Controllable Feature Whitening for Hyperparameter-Free Bias Mitigation
by: Cho, Yooshin, et al.
Published: (2025)
by: Cho, Yooshin, et al.
Published: (2025)
Reservoir Subspace Injection for Online ICA under Top-n Whitening
by: Xiao, Wenjun, et al.
Published: (2026)
by: Xiao, Wenjun, et al.
Published: (2026)
Towards Scalable SOAP Note Generation: A Weakly Supervised Multimodal Framework
by: Kamal, Sadia, et al.
Published: (2025)
by: Kamal, Sadia, et al.
Published: (2025)
Grouped Satisficing Paths in Pure Strategy Games: a Topological Perspective
by: Fu, Yanqing, et al.
Published: (2025)
by: Fu, Yanqing, et al.
Published: (2025)
A Bootstrap Perspective on Stochastic Gradient Descent
by: Lan, Hongjian, et al.
Published: (2025)
by: Lan, Hongjian, et al.
Published: (2025)
GradPower: Powering Gradients for Faster Language Model Pre-Training
by: Wang, Jinbo, et al.
Published: (2025)
by: Wang, Jinbo, et al.
Published: (2025)
Principled Understanding of Generalization for Generative Transformer Models in Arithmetic Reasoning Tasks
by: Xu, Xingcheng, et al.
Published: (2024)
by: Xu, Xingcheng, et al.
Published: (2024)
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation
by: Faisal, Faizan
Published: (2026)
by: Faisal, Faizan
Published: (2026)
It Ain't That Bad: Understanding the Mysterious Performance Drop in OOD Generalization for Generative Transformer Models
by: Xu, Xingcheng, et al.
Published: (2023)
by: Xu, Xingcheng, et al.
Published: (2023)
Gradient Guidance for Diffusion Models: An Optimization Perspective
by: Guo, Yingqing, et al.
Published: (2024)
by: Guo, Yingqing, et al.
Published: (2024)
Understanding Robust Overfitting from the Feature Generalization Perspective
by: Yu, Chaojian, et al.
Published: (2023)
by: Yu, Chaojian, et al.
Published: (2023)
Whitened Score Diffusion: A Structured Prior for Imaging Inverse Problems
by: Alido, Jeffrey, et al.
Published: (2025)
by: Alido, Jeffrey, et al.
Published: (2025)
Understanding Private Learning From Feature Perspective
by: Ding, Meng, et al.
Published: (2025)
by: Ding, Meng, et al.
Published: (2025)
How to Defend Against Large-scale Model Poisoning Attacks in Federated Learning: A Vertical Solution
by: Wang, Jinbo, et al.
Published: (2024)
by: Wang, Jinbo, et al.
Published: (2024)
Gradient Shaping Beyond Clipping: A Functional Perspective on Update Magnitude Control
by: You, Haochen, et al.
Published: (2025)
by: You, Haochen, et al.
Published: (2025)
A Theoretical Understanding of Gradient Bias in Meta-Reinforcement Learning
by: Feng, Xidong, et al.
Published: (2021)
by: Feng, Xidong, et al.
Published: (2021)
Balance Reward and Safety Optimization for Safe Reinforcement Learning: A Perspective of Gradient Manipulation
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
Toward a Better Understanding of Fourier Neural Operators from a Spectral Perspective
by: Qin, Shaoxiang, et al.
Published: (2024)
by: Qin, Shaoxiang, et al.
Published: (2024)
Understanding Stochastic Natural Gradient Variational Inference
by: Wu, Kaiwen, et al.
Published: (2024)
by: Wu, Kaiwen, et al.
Published: (2024)
Understanding Gradient Descent through the Training Jacobian
by: Belrose, Nora, et al.
Published: (2024)
by: Belrose, Nora, et al.
Published: (2024)
Similar Items
-
Understanding and Improving Shampoo and SOAP via Kullback-Leibler Minimization
by: Lin, Wu, et al.
Published: (2025) -
Zipfian Whitening
by: Yokoi, Sho, et al.
Published: (2024) -
Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes
by: Kamal, Sadia, et al.
Published: (2025) -
Whitening Not Recommended for Classification Tasks in LLMs
by: Forooghi, Ali, et al.
Published: (2024) -
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)