Breaking the Frozen Subspace: Importance Sampling for Low-Rank Optimization in LLM Pretraining
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Haochen, Yin, Junze, Wang, Guanchu, Liu, Zirui, Yang, Lin F., Zhang, Tianyi, Shrivastava, Anshumali, Braverman, Vladimir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoVE: Compressed Vocabulary Expansion Makes Better LLM-based Recommender Systems
von: Zhang, Haochen, et al.
Veröffentlicht: (2025)
von: Zhang, Haochen, et al.
Veröffentlicht: (2025)
LeanQuant: Accurate and Scalable Large Language Model Quantization with Loss-error-aware Grid
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
Support Basis: Fast Attention Beyond Bounded Entries
von: Aliakbarpour, Maryam, et al.
Veröffentlicht: (2025)
von: Aliakbarpour, Maryam, et al.
Veröffentlicht: (2025)
KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
IDentity with Locality: An ideal hash for gene sequence search
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
Learning Scalable Structural Representations for Link Prediction with Bloom Signatures
von: Zhang, Tianyi, et al.
Veröffentlicht: (2023)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2023)
NoMAD-Attention: Efficient LLM Inference on CPUs Through Multiply-add-free Attention
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
High-Dimensional Robust Mean Estimation with Untrusted Batches
von: Aliakbarpour, Maryam, et al.
Veröffentlicht: (2026)
von: Aliakbarpour, Maryam, et al.
Veröffentlicht: (2026)
Sketch to Adapt: Fine-Tunable Sketches for Efficient LLM Adaptation
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2025)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2025)
70% Size, 100% Accuracy: Lossless LLM Compression for Efficient GPU Inference via Dynamic-Length Float (DFloat11)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
To Compress or Not? Pushing the Frontier of Lossless GenAI Model Weights Compression with Exponent Concentration
von: Yang, Zeyu, et al.
Veröffentlicht: (2025)
von: Yang, Zeyu, et al.
Veröffentlicht: (2025)
Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2026)
von: Luo, Feng, et al.
Veröffentlicht: (2026)
Efficient Alternating Minimization with Applications to Weighted Low Rank Approximation
von: Song, Zhao, et al.
Veröffentlicht: (2023)
von: Song, Zhao, et al.
Veröffentlicht: (2023)
Low Rank Matrix Completion via Robust Alternating Minimization in Nearly Linear Time
von: Gu, Yuzhou, et al.
Veröffentlicht: (2023)
von: Gu, Yuzhou, et al.
Veröffentlicht: (2023)
A Dynamic Low-Rank Fast Gaussian Transform
von: Huang, Baihe, et al.
Veröffentlicht: (2022)
von: Huang, Baihe, et al.
Veröffentlicht: (2022)
Low-Rank Compression of Pretrained Models via Randomized Subspace Iteration
von: Pourkamali-Anaraki, Farhad
Veröffentlicht: (2026)
von: Pourkamali-Anaraki, Farhad
Veröffentlicht: (2026)
Superintelligent Retrieval Agent: The Next Frontier of Information Retrieval
von: Yang, Zeyu, et al.
Veröffentlicht: (2026)
von: Yang, Zeyu, et al.
Veröffentlicht: (2026)
SRLoRA: Subspace Recomposition in Low-Rank Adaptation via Importance-Based Fusion and Reinitialization
von: Yang, Haodong, et al.
Veröffentlicht: (2025)
von: Yang, Haodong, et al.
Veröffentlicht: (2025)
Personalizing Low-Rank Bayesian Neural Networks Via Federated Learning
von: Zhang, Boning, et al.
Veröffentlicht: (2024)
von: Zhang, Boning, et al.
Veröffentlicht: (2024)
Scout Before You Attend: Sketch-and-Walk Sparse Attention for Efficient LLM Inference
von: Le, Hoang Anh Duy, et al.
Veröffentlicht: (2026)
von: Le, Hoang Anh Duy, et al.
Veröffentlicht: (2026)
DTS: Enhancing Large Reasoning Models via Decoding Tree Sketching
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
GIPO: Gaussian Importance Sampling Policy Optimization
von: Lu, Chengxuan, et al.
Veröffentlicht: (2026)
von: Lu, Chengxuan, et al.
Veröffentlicht: (2026)
Stabilizing Native Low-Rank LLM Pretraining
von: Janson, Paul, et al.
Veröffentlicht: (2026)
von: Janson, Paul, et al.
Veröffentlicht: (2026)
Regularizing Subspace Redundancy of Low-Rank Adaptation
von: Zhu, Yue, et al.
Veröffentlicht: (2025)
von: Zhu, Yue, et al.
Veröffentlicht: (2025)
From Low Rank Gradient Subspace Stabilization to Low-Rank Weights: Observations, Theories, and Applications
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
Lotus: Efficient LLM Training by Randomized Low-Rank Gradient Projection with Adaptive Subspace Switching
von: Miao, Tianhao, et al.
Veröffentlicht: (2026)
von: Miao, Tianhao, et al.
Veröffentlicht: (2026)
Assessing and Enhancing Large Language Models in Rare Disease Question-answering
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
Scalable Importance Sampling in High Dimensions with Low-Rank Mixture Proposals
von: Kruse, Liam A., et al.
Veröffentlicht: (2025)
von: Kruse, Liam A., et al.
Veröffentlicht: (2025)
CoRA: Optimizing Low-Rank Adaptation with Common Subspace of Large Language Models
von: Xiao, Xiaojun, et al.
Veröffentlicht: (2024)
von: Xiao, Xiaojun, et al.
Veröffentlicht: (2024)
Mixture-of-Subspaces in Low-Rank Adaptation
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
REFRAG: Rethinking RAG based Decoding
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
Borrowed Geometry: Cross-Distribution Head-Importance Fingerprints of Frozen Pretrained Gemma 4 31B
von: Bektursun, Abay
Veröffentlicht: (2026)
von: Bektursun, Abay
Veröffentlicht: (2026)
Breaking the Blocks: Continuous Low-Rank Decomposed Scaling for Unified LLM Quantization and Adaptation
von: Tang, Pingzhi, et al.
Veröffentlicht: (2026)
von: Tang, Pingzhi, et al.
Veröffentlicht: (2026)
Self-ensemble: Mitigating Confidence Mis-calibration for Large Language Models
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
ESPO: Entropy Importance Sampling Policy Optimization
von: Sheng, Yuepeng, et al.
Veröffentlicht: (2025)
von: Sheng, Yuepeng, et al.
Veröffentlicht: (2025)
CARAMEL: A Succinct Read-Only Lookup Table via Compressed Static Functions
von: Coleman, Benjamin, et al.
Veröffentlicht: (2023)
von: Coleman, Benjamin, et al.
Veröffentlicht: (2023)
Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective
von: Zhang, Yuheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2026)
PeriodicLoRA: Breaking the Low-Rank Bottleneck in LoRA Optimization
von: Meng, Xiangdi, et al.
Veröffentlicht: (2024)
von: Meng, Xiangdi, et al.
Veröffentlicht: (2024)
Low-Rank Robust Subspace Tensor Clustering for Metro Passenger Flow Modeling
von: Hu, Jiuyun, et al.
Veröffentlicht: (2024)
von: Hu, Jiuyun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CoVE: Compressed Vocabulary Expansion Makes Better LLM-based Recommender Systems
von: Zhang, Haochen, et al.
Veröffentlicht: (2025) -
LeanQuant: Accurate and Scalable Large Language Model Quantization with Loss-error-aware Grid
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024) -
Support Basis: Fast Attention Beyond Bounded Entries
von: Aliakbarpour, Maryam, et al.
Veröffentlicht: (2025) -
KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024) -
IDentity with Locality: An ideal hash for gene sequence search
von: Desai, Aditya, et al.
Veröffentlicht: (2024)