Nora: Normalized Orthogonal Row Alignment for Scalable Matrix Optimizer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Jinghui, Zou, Jiaxuan, Wang, Shuo, Liu, Yong, Nie, Feiping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Riemannian Optimization on Relaxed Indicator Matrix Manifold
von: Yuan, Jinghui, et al.
Veröffentlicht: (2025)
von: Yuan, Jinghui, et al.
Veröffentlicht: (2025)
RMNP: Row-Momentum Normalized Preconditioning for Scalable Matrix-Based Optimization
von: Deng, Shenyang, et al.
Veröffentlicht: (2026)
von: Deng, Shenyang, et al.
Veröffentlicht: (2026)
A Margin-Maximizing Fine-Grained Ensemble Method
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
Dual-Bounded Nonlinear Optimal Transport for Size Constrained Min Cut Clustering
von: Xie, Fangyuan, et al.
Veröffentlicht: (2025)
von: Xie, Fangyuan, et al.
Veröffentlicht: (2025)
Achieving More with Less: A Tensor-Optimization-Powered Ensemble Method
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
Decoupled Orthogonal Dynamics: Regularization for Deep Network Optimizers
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
MOMA: Masked Orthogonal Matrix Alignment for Zero-Additional-Parameter Model Merging
von: Kong, Fanshuang, et al.
Veröffentlicht: (2024)
von: Kong, Fanshuang, et al.
Veröffentlicht: (2024)
Multi-class Support Vector Machine with Maximizing Minimum Margin
von: Hao, Zhezheng, et al.
Veröffentlicht: (2023)
von: Hao, Zhezheng, et al.
Veröffentlicht: (2023)
Doubly Stochastic Adaptive Neighbors Clustering via the Marcus Mapping
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)
Adaptive Fuzzy C-Means with Graph Embedding
von: Chen, Qiang, et al.
Veröffentlicht: (2024)
von: Chen, Qiang, et al.
Veröffentlicht: (2024)
PRformer: Pyramidal Recurrent Transformer for Multivariate Time Series Forecasting
von: Yu, Yongbo, et al.
Veröffentlicht: (2024)
von: Yu, Yongbo, et al.
Veröffentlicht: (2024)
On the Width Scaling of Neural Optimizers Under Matrix Operator Norms I: Row/Column Normalization and Hyperparameter Transfer
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
Fast Semi-supervised Learning on Large Graphs: An Improved Green-function Method
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
Clustering Based on Density Propagation and Subcluster Merging
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
A Greedy Strategy for Graph Cut
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
Kaczmarz Linear Attention
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
COMPOT: Calibration-Optimized Matrix Procrustes Orthogonalization for Transformers Compression
von: Makhov, Denis, et al.
Veröffentlicht: (2026)
von: Makhov, Denis, et al.
Veröffentlicht: (2026)
Towards Federated Clustering: A Client-wise Private Graph Aggregation Framework
von: He, Guanxiong, et al.
Veröffentlicht: (2025)
von: He, Guanxiong, et al.
Veröffentlicht: (2025)
DP-Muon: Differentially Private Optimization via Matrix-Orthogonalized Momentum
von: Kim, Jihwan, et al.
Veröffentlicht: (2026)
von: Kim, Jihwan, et al.
Veröffentlicht: (2026)
Robust Capped lp-Norm Support Vector Ordinal Regression
von: Xiang, Haorui, et al.
Veröffentlicht: (2024)
von: Xiang, Haorui, et al.
Veröffentlicht: (2024)
Cross-attention Secretly Performs Orthogonal Alignment in Recommendation Models
von: Lee, Hyunin, et al.
Veröffentlicht: (2025)
von: Lee, Hyunin, et al.
Veröffentlicht: (2025)
Muown: Row-Norm Control for Muon Optimization
von: Lion, Kai, et al.
Veröffentlicht: (2026)
von: Lion, Kai, et al.
Veröffentlicht: (2026)
Simple Multigraph Convolution Networks
von: Wu, Danyang, et al.
Veröffentlicht: (2024)
von: Wu, Danyang, et al.
Veröffentlicht: (2024)
FedMuon: Accelerating Federated Learning with Matrix Orthogonalization
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
Capabilities and Fundamental Limits of Latent Chain-of-Thought
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
UniDiff: A Unified Diffusion Framework for Multimodal Time Series Forecasting
von: Zhang, Da, et al.
Veröffentlicht: (2025)
von: Zhang, Da, et al.
Veröffentlicht: (2025)
FusAD: Time-Frequency Fusion with Adaptive Denoising for General Time Series Analysis
von: Zhang, Da, et al.
Veröffentlicht: (2025)
von: Zhang, Da, et al.
Veröffentlicht: (2025)
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
von: Zixian, Wang
Veröffentlicht: (2026)
von: Zixian, Wang
Veröffentlicht: (2026)
Group Orthogonalized Policy Optimization:Group Policy Optimization as Orthogonal Projection in Hilbert Space
von: Zixian, Wang
Veröffentlicht: (2026)
von: Zixian, Wang
Veröffentlicht: (2026)
FAIM: Frequency-Aware Interactive Mamba for Time Series Classification
von: Zhang, Da, et al.
Veröffentlicht: (2025)
von: Zhang, Da, et al.
Veröffentlicht: (2025)
PITE: Multi-Prototype Alignment for Individual Treatment Effect Estimation
von: Cao, Fuyuan, et al.
Veröffentlicht: (2025)
von: Cao, Fuyuan, et al.
Veröffentlicht: (2025)
OrthAlign: Orthogonal Subspace Decomposition for Non-Interfering Multi-Objective Alignment
von: Lin, Liang, et al.
Veröffentlicht: (2025)
von: Lin, Liang, et al.
Veröffentlicht: (2025)
Variational Inference, Entropy, and Orthogonality: A Unified Theory of Mixture-of-Experts
von: Su, Ye, et al.
Veröffentlicht: (2026)
von: Su, Ye, et al.
Veröffentlicht: (2026)
Sparse Orthogonal Parameters Tuning for Continual Learning
von: Ning, Kun-Peng, et al.
Veröffentlicht: (2024)
von: Ning, Kun-Peng, et al.
Veröffentlicht: (2024)
Row-Stochastic Matrices Can Provably Outperform Doubly Stochastic Matrices in Decentralized Learning
von: Liu, Bing, et al.
Veröffentlicht: (2025)
von: Liu, Bing, et al.
Veröffentlicht: (2025)
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Safety Alignment as Continual Learning: Mitigating the Alignment Tax via Orthogonal Gradient Projection
von: Sun, Guanglong, et al.
Veröffentlicht: (2026)
von: Sun, Guanglong, et al.
Veröffentlicht: (2026)
Memory-Guided Trust-Region Bayesian Optimization (MG-TuRBO) for High Dimensions
von: Saroj, Abhilasha, et al.
Veröffentlicht: (2026)
von: Saroj, Abhilasha, et al.
Veröffentlicht: (2026)
Hyperspherical Normalization for Scalable Deep Reinforcement Learning
von: Lee, Hojoon, et al.
Veröffentlicht: (2025)
von: Lee, Hojoon, et al.
Veröffentlicht: (2025)
Effective Frontiers: A Unification of Neural Scaling Laws
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Zou, Jiaxuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Riemannian Optimization on Relaxed Indicator Matrix Manifold
von: Yuan, Jinghui, et al.
Veröffentlicht: (2025) -
RMNP: Row-Momentum Normalized Preconditioning for Scalable Matrix-Based Optimization
von: Deng, Shenyang, et al.
Veröffentlicht: (2026) -
A Margin-Maximizing Fine-Grained Ensemble Method
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024) -
Dual-Bounded Nonlinear Optimal Transport for Size Constrained Min Cut Clustering
von: Xie, Fangyuan, et al.
Veröffentlicht: (2025) -
Achieving More with Less: A Tensor-Optimization-Powered Ensemble Method
von: Yuan, Jinghui, et al.
Veröffentlicht: (2024)