Saved in:
| Main Authors: | Han, Andi, Li, Jiaxiang, Huang, Wei, Hong, Mingyi, Takeda, Akiko, Jawanpuria, Pratik, Mishra, Bamdev |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.02214 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Framework for Bilevel Optimization on Riemannian Manifolds
by: Han, Andi, et al.
Published: (2024)
by: Han, Andi, et al.
Published: (2024)
Riemannian coordinate descent algorithms on matrix manifolds
by: Han, Andi, et al.
Published: (2024)
by: Han, Andi, et al.
Published: (2024)
Generalized infinite dimensional Alpha-Procrustes based geometries
by: Goomanee, Salvish, et al.
Published: (2025)
by: Goomanee, Salvish, et al.
Published: (2025)
Riemannian Optimization for Hadamard Products of Low-Rank Matrices
by: Jawanpuria, Pratik, et al.
Published: (2026)
by: Jawanpuria, Pratik, et al.
Published: (2026)
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
by: Mishra, Neel, et al.
Published: (2024)
by: Mishra, Neel, et al.
Published: (2024)
LOFT: Low-Rank Orthogonal Fine-Tuning via Task-Aware Support Selection
by: Zhao, Lanxin, et al.
Published: (2026)
by: Zhao, Lanxin, et al.
Published: (2026)
A Riemannian Approach to Ground Metric Learning for Optimal Transport
by: Jawanpuria, Pratik, et al.
Published: (2024)
by: Jawanpuria, Pratik, et al.
Published: (2024)
Riemannian Federated Learning via Averaging Gradient Streams
by: Huang, Zhenwei, et al.
Published: (2024)
by: Huang, Zhenwei, et al.
Published: (2024)
Federated Learning on Riemannian Manifolds with Differential Privacy
by: Huang, Zhenwei, et al.
Published: (2024)
by: Huang, Zhenwei, et al.
Published: (2024)
Intrinsic Muon: Spectral Optimization on Riemannian Matrix Manifolds
by: Li, Yibang, et al.
Published: (2026)
by: Li, Yibang, et al.
Published: (2026)
Submodular Framework for Structured-Sparse Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2024)
by: Manupriya, Piyushi, et al.
Published: (2024)
UniPROT: Uniform Prototype Selection via Partial Optimal Transport with Submodular Guarantees
by: Chanda, Prateek, et al.
Published: (2026)
by: Chanda, Prateek, et al.
Published: (2026)
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
by: Glentis, Athanasios, et al.
Published: (2025)
by: Glentis, Athanasios, et al.
Published: (2025)
Nyström Approximation on Manifolds
by: Nie, Hantao, et al.
Published: (2026)
by: Nie, Hantao, et al.
Published: (2026)
Efficient Optimization with Orthogonality Constraint: a Randomized Riemannian Submanifold Method
by: Han, Andi, et al.
Published: (2025)
by: Han, Andi, et al.
Published: (2025)
Scalable Parameter and Memory Efficient Pretraining for LLM: Recent Algorithmic Advances and Benchmarking
by: Glentis, Athanasios, et al.
Published: (2025)
by: Glentis, Athanasios, et al.
Published: (2025)
Modified K-means Algorithm with Local Optimality Guarantees
by: Li, Mingyi, et al.
Published: (2025)
by: Li, Mingyi, et al.
Published: (2025)
FedSEA: Achieving Benefit of Parallelization in Federated Online Learning
by: Sahu, Harekrushna, et al.
Published: (2026)
by: Sahu, Harekrushna, et al.
Published: (2026)
MMD-Regularized Unbalanced Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2020)
by: Manupriya, Piyushi, et al.
Published: (2020)
On the Role of Label Noise in the Feature Learning Process
by: Han, Andi, et al.
Published: (2025)
by: Han, Andi, et al.
Published: (2025)
A Framework for Quantifying How Pre-Training and Context Benefit In-Context Learning
by: Song, Bingqing, et al.
Published: (2025)
by: Song, Bingqing, et al.
Published: (2025)
AdaFish: Fast low-rank parameter-efficient fine-tuning by using second-order information
by: Hu, Jiang, et al.
Published: (2024)
by: Hu, Jiang, et al.
Published: (2024)
Do pretrained Transformers Learn In-Context by Gradient Descent?
by: Shen, Lingfeng, et al.
Published: (2023)
by: Shen, Lingfeng, et al.
Published: (2023)
Beyond IID weights: sparse and low-rank deep Neural Networks are also Gaussian Processes
by: Nait-Saada, Thiziri, et al.
Published: (2023)
by: Nait-Saada, Thiziri, et al.
Published: (2023)
PCA recovery thresholds in low-rank matrix inference with sparse noise
by: Adomaityte, Urte, et al.
Published: (2025)
by: Adomaityte, Urte, et al.
Published: (2025)
On the Feature Learning in Diffusion Models
by: Han, Andi, et al.
Published: (2024)
by: Han, Andi, et al.
Published: (2024)
Robust Least-Squares Optimization for Data-Driven Predictive Control: A Geometric Approach
by: Bharadwaj, Shreyas, et al.
Published: (2025)
by: Bharadwaj, Shreyas, et al.
Published: (2025)
Efficient transformer adaptation for analog in-memory computing via low-rank adapters
by: Li, Chen, et al.
Published: (2024)
by: Li, Chen, et al.
Published: (2024)
Adaptive regularization parameter selection for high-dimensional inverse problems: A Bayesian approach with Tucker low-rank constraints
by: Yang, Qing-Mei, et al.
Published: (2026)
by: Yang, Qing-Mei, et al.
Published: (2026)
Harnessing small projectors and multiple views for efficient vision pretraining
by: Agrawal, Kumar Krishna, et al.
Published: (2023)
by: Agrawal, Kumar Krishna, et al.
Published: (2023)
The effectiveness of MAE pre-pretraining for billion-scale pretraining
by: Singh, Mannat, et al.
Published: (2023)
by: Singh, Mannat, et al.
Published: (2023)
Universal priors: solving empirical Bayes via Bayesian inference and pretraining
by: Cannella, Nick, et al.
Published: (2026)
by: Cannella, Nick, et al.
Published: (2026)
Storing overlapping associative memories on latent manifolds in low-rank spiking networks
by: Podlaski, William F., et al.
Published: (2024)
by: Podlaski, William F., et al.
Published: (2024)
Evolution Strategies for Deep RL pretraining
by: Martínez, Adrian, et al.
Published: (2026)
by: Martínez, Adrian, et al.
Published: (2026)
Synthetic continued pretraining
by: Yang, Zitong, et al.
Published: (2024)
by: Yang, Zitong, et al.
Published: (2024)
Convergence Error Analysis of Reflected Gradient Langevin Dynamics for Globally Optimizing Non-Convex Constrained Problems
by: Sato, Kanji, et al.
Published: (2022)
by: Sato, Kanji, et al.
Published: (2022)
Physics-informed waveform inversion using pretrained wavefield neural operators
by: Huang, Xinquan, et al.
Published: (2025)
by: Huang, Xinquan, et al.
Published: (2025)
On efficiently computable functions, deep networks and sparse compositionality
by: Poggio, Tomaso
Published: (2025)
by: Poggio, Tomaso
Published: (2025)
MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency
by: Dufour, Nicolas, et al.
Published: (2025)
by: Dufour, Nicolas, et al.
Published: (2025)
Enhancing pretraining efficiency for medical image segmentation via transferability metrics
by: Hidy, Gábor, et al.
Published: (2024)
by: Hidy, Gábor, et al.
Published: (2024)
Similar Items
-
A Framework for Bilevel Optimization on Riemannian Manifolds
by: Han, Andi, et al.
Published: (2024) -
Riemannian coordinate descent algorithms on matrix manifolds
by: Han, Andi, et al.
Published: (2024) -
Generalized infinite dimensional Alpha-Procrustes based geometries
by: Goomanee, Salvish, et al.
Published: (2025) -
Riemannian Optimization for Hadamard Products of Low-Rank Matrices
by: Jawanpuria, Pratik, et al.
Published: (2026) -
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
by: Mishra, Neel, et al.
Published: (2024)