PolarAdamW: Disentangling Spectral Control and Schur Gauge-Equivariance in Matrix Optimisation
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Zhang, Haozhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chain-of-Thought and Compressed Looped Transformers: A Memory-Budget Separation
von: Zhang, Haozhou
Veröffentlicht: (2026)
von: Zhang, Haozhou
Veröffentlicht: (2026)
HomeAdam: Adam and AdamW Algorithms Sometimes Go Home to Obtain Better Provable Generalization
von: Huang, Feihu, et al.
Veröffentlicht: (2026)
von: Huang, Feihu, et al.
Veröffentlicht: (2026)
Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers
von: Ranganath, Aditya
Veröffentlicht: (2026)
von: Ranganath, Aditya
Veröffentlicht: (2026)
Gauge-Equivariant Graph Networks via Self-Interference Cancellation
von: Choi, Yoonhyuk, et al.
Veröffentlicht: (2025)
von: Choi, Yoonhyuk, et al.
Veröffentlicht: (2025)
Learning Disentangled Equivariant Representation for Explicitly Controllable 3D Molecule Generation
von: Liu, Haoran, et al.
Veröffentlicht: (2024)
von: Liu, Haoran, et al.
Veröffentlicht: (2024)
Incorporating Arbitrary Matrix Group Equivariance into KANs
von: Hu, Lexiang, et al.
Veröffentlicht: (2024)
von: Hu, Lexiang, et al.
Veröffentlicht: (2024)
Gauge-Equivariant Graph Neural Networks for Lattice Gauge Theories
von: Rayat, Ali, et al.
Veröffentlicht: (2026)
von: Rayat, Ali, et al.
Veröffentlicht: (2026)
Monomial Matrix Group Equivariant Neural Functional Networks
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2024)
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2024)
Disentanglement-Based Equivariant Learning for Compositional VQA
von: Du, Zhou, et al.
Veröffentlicht: (2026)
von: Du, Zhou, et al.
Veröffentlicht: (2026)
Novel Dynamic Batch-Sensitive Adam Optimiser for Vehicular Accident Injury Severity Prediction
von: Kyei, Daniel Asare, et al.
Veröffentlicht: (2026)
von: Kyei, Daniel Asare, et al.
Veröffentlicht: (2026)
Uniform Scaling Limits in AdamW-Trained Transformers
von: Gibson, William, et al.
Veröffentlicht: (2026)
von: Gibson, William, et al.
Veröffentlicht: (2026)
Equivariant Machine Learning on Graphs with Nonlinear Spectral Filters
von: Lin, Ya-Wei Eileen, et al.
Veröffentlicht: (2024)
von: Lin, Ya-Wei Eileen, et al.
Veröffentlicht: (2024)
Beyond First-Order: Training LLMs with Stochastic Conjugate Subgradients and AdamW
von: Zhang, Di, et al.
Veröffentlicht: (2025)
von: Zhang, Di, et al.
Veröffentlicht: (2025)
AdamZ: An Enhanced Optimisation Method for Neural Network Training
von: Zaznov, Ilia, et al.
Veröffentlicht: (2024)
von: Zaznov, Ilia, et al.
Veröffentlicht: (2024)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
von: Zhang, Fangzhao, et al.
Veröffentlicht: (2026)
von: Zhang, Fangzhao, et al.
Veröffentlicht: (2026)
Equivariant Matrix Function Neural Networks
von: Batatia, Ilyes, et al.
Veröffentlicht: (2023)
von: Batatia, Ilyes, et al.
Veröffentlicht: (2023)
Learning Chern Numbers of Topological Insulators with Gauge Equivariant Neural Networks
von: Huang, Longde, et al.
Veröffentlicht: (2025)
von: Huang, Longde, et al.
Veröffentlicht: (2025)
APOLLO: SGD-like Memory, AdamW-level Performance
von: Zhu, Hanqing, et al.
Veröffentlicht: (2024)
von: Zhu, Hanqing, et al.
Veröffentlicht: (2024)
Generalizable Equivariant Diffusion Models for Non-Abelian Lattice Gauge Theory
von: Aarts, Gert, et al.
Veröffentlicht: (2026)
von: Aarts, Gert, et al.
Veröffentlicht: (2026)
Gauge-Equivariant Intrinsic Neural Operators for Geometry-Consistent Learning of Elliptic PDE Maps
von: Cheng, Pengcheng
Veröffentlicht: (2026)
von: Cheng, Pengcheng
Veröffentlicht: (2026)
Diagnosing Spectral Ceilings in Equivariant Neural Force Fields
von: Kim, Hyunmog
Veröffentlicht: (2026)
von: Kim, Hyunmog
Veröffentlicht: (2026)
GIST: Gauge-Invariant Spectral Transformers for Scalable Graph Neural Operators
von: Rigotti, Mattia, et al.
Veröffentlicht: (2026)
von: Rigotti, Mattia, et al.
Veröffentlicht: (2026)
Implicit Bias of AdamW: $\ell_\infty$ Norm Constrained Optimization
von: Xie, Shuo, et al.
Veröffentlicht: (2024)
von: Xie, Shuo, et al.
Veröffentlicht: (2024)
Multiple Invertible and Partial-Equivariant Function for Latent Vector Transformation to Enhance Disentanglement in VAEs
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2025)
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2025)
DP-AdamW: Investigating Decoupled Weight Decay and Bias Correction in Private Deep Learning
von: Chooi, Jay, et al.
Veröffentlicht: (2025)
von: Chooi, Jay, et al.
Veröffentlicht: (2025)
G-RepsNet: A Fast and General Construction of Equivariant Networks for Arbitrary Matrix Groups
von: Basu, Sourya, et al.
Veröffentlicht: (2024)
von: Basu, Sourya, et al.
Veröffentlicht: (2024)
Compact Matrix Quantum Group Equivariant Neural Networks
von: Pearce-Crump, Edward
Veröffentlicht: (2023)
von: Pearce-Crump, Edward
Veröffentlicht: (2023)
The Implicit Bias of Adam on Separable Data
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
CaAdam: Improving Adam optimizer using connection aware methods
von: Genet, Remi, et al.
Veröffentlicht: (2024)
von: Genet, Remi, et al.
Veröffentlicht: (2024)
Matrix Completion via Residual Spectral Matching
von: Chen, Ziyuan, et al.
Veröffentlicht: (2024)
von: Chen, Ziyuan, et al.
Veröffentlicht: (2024)
DP-FedAdamW: An Efficient Optimizer for Differentially Private Federated Large Models
von: Liu, Jin, et al.
Veröffentlicht: (2026)
von: Liu, Jin, et al.
Veröffentlicht: (2026)
How to set AdamW's weight decay as you scale model and dataset size
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2024)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2024)
A Machine Learning Approach Towards Runtime Optimisation of Matrix Multiplication
von: Xia, Yufan, et al.
Veröffentlicht: (2026)
von: Xia, Yufan, et al.
Veröffentlicht: (2026)
FreDN: Spectral Disentanglement for Time Series Forecasting via Learnable Frequency Decomposition
von: An, Zhongde, et al.
Veröffentlicht: (2025)
von: An, Zhongde, et al.
Veröffentlicht: (2025)
Disentangling Shared and Target-Enriched Topics via Background-Contrastive Non-negative Matrix Factorization
von: Li, Yixuan, et al.
Veröffentlicht: (2026)
von: Li, Yixuan, et al.
Veröffentlicht: (2026)
Disentangled (Un)Controllable Features
von: Kooi, Jacob E., et al.
Veröffentlicht: (2022)
von: Kooi, Jacob E., et al.
Veröffentlicht: (2022)
First-Passage Prediction of Grokking Delay: ACalibrated Law under AdamW with Causal Validation
von: Khanh, Truong Xuan, et al.
Veröffentlicht: (2026)
von: Khanh, Truong Xuan, et al.
Veröffentlicht: (2026)
Controlled Generation with Equivariant Variational Flow Matching
von: Eijkelboom, Floor, et al.
Veröffentlicht: (2025)
von: Eijkelboom, Floor, et al.
Veröffentlicht: (2025)
On Suppressing Range of Adaptive Stepsizes of Adam to Improve Generalisation Performance
von: Zhang, Guoqiang
Veröffentlicht: (2023)
von: Zhang, Guoqiang
Veröffentlicht: (2023)
Ähnliche Einträge
-
Chain-of-Thought and Compressed Looped Transformers: A Memory-Budget Separation
von: Zhang, Haozhou
Veröffentlicht: (2026) -
HomeAdam: Adam and AdamW Algorithms Sometimes Go Home to Obtain Better Provable Generalization
von: Huang, Feihu, et al.
Veröffentlicht: (2026) -
Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers
von: Ranganath, Aditya
Veröffentlicht: (2026) -
Gauge-Equivariant Graph Networks via Self-Interference Cancellation
von: Choi, Yoonhyuk, et al.
Veröffentlicht: (2025) -
Learning Disentangled Equivariant Representation for Explicitly Controllable 3D Molecule Generation
von: Liu, Haoran, et al.
Veröffentlicht: (2024)