Low-Rank Compression of Language Models via Differentiable Rank Selection
Fuente:
arXiv
Salvato in:
| Autori principali: | Sundrani, Sidhant, Tudisco, Francesco, Minervini, Pasquale |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Low-Rank Adversarial PGD Attack
di: Savostianova, Dayana, et al.
Pubblicazione: (2024)
di: Savostianova, Dayana, et al.
Pubblicazione: (2024)
Lillama: Large Language Models Compression via Low-Rank Feature Distillation
di: Sy, Yaya, et al.
Pubblicazione: (2024)
di: Sy, Yaya, et al.
Pubblicazione: (2024)
Compressing Large Language Models using Low Rank and Low Precision Decomposition
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
Memory-Efficient Fine-Tuning via Low-Rank Activation Compression
di: Shi, Jiang-Xin, et al.
Pubblicazione: (2025)
di: Shi, Jiang-Xin, et al.
Pubblicazione: (2025)
Lossless Model Compression via Joint Low-Rank Factorization Optimization
di: Zhang, Boyang, et al.
Pubblicazione: (2024)
di: Zhang, Boyang, et al.
Pubblicazione: (2024)
Low-Rank Compression of Pretrained Models via Randomized Subspace Iteration
di: Pourkamali-Anaraki, Farhad
Pubblicazione: (2026)
di: Pourkamali-Anaraki, Farhad
Pubblicazione: (2026)
CALR: Corrective Adaptive Low-Rank Decomposition for Efficient Large Language Model Layer Compression
di: Kautsar, Muchammad Daniyal, et al.
Pubblicazione: (2025)
di: Kautsar, Muchammad Daniyal, et al.
Pubblicazione: (2025)
Hierarchical Sparse Plus Low Rank Compression of LLM
di: Kumar, Pawan, et al.
Pubblicazione: (2025)
di: Kumar, Pawan, et al.
Pubblicazione: (2025)
Palu: Compressing KV-Cache with Low-Rank Projection
di: Chang, Chi-Chih, et al.
Pubblicazione: (2024)
di: Chang, Chi-Chih, et al.
Pubblicazione: (2024)
Large Language Model Compression with Global Rank and Sparsity Optimization
di: Zhou, Changhai, et al.
Pubblicazione: (2025)
di: Zhou, Changhai, et al.
Pubblicazione: (2025)
SynDARin: Synthesising Datasets for Automated Reasoning in Low-Resource Languages
di: Ghazaryan, Gayane, et al.
Pubblicazione: (2024)
di: Ghazaryan, Gayane, et al.
Pubblicazione: (2024)
D-QRELO: Training- and Data-Free Delta Compression for Large Language Models via Quantization and Residual Low-Rank Approximation
di: Li, Junlin, et al.
Pubblicazione: (2026)
di: Li, Junlin, et al.
Pubblicazione: (2026)
FFT-based Dynamic Subspace Selection for Low-Rank Adaptive Optimization of Large Language Models
di: Modoranu, Ionut-Vlad, et al.
Pubblicazione: (2025)
di: Modoranu, Ionut-Vlad, et al.
Pubblicazione: (2025)
DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights
di: Mikaelyan, Liana, et al.
Pubblicazione: (2025)
di: Mikaelyan, Liana, et al.
Pubblicazione: (2025)
MGAA: Multi-Granular Adaptive Allocation fof Low-Rank Compression of LLMs
di: Li, Guangyan, et al.
Pubblicazione: (2025)
di: Li, Guangyan, et al.
Pubblicazione: (2025)
ReCalKV: Low-Rank KV Cache Compression via Head Reordering and Offline Calibration
di: Yan, Xianglong, et al.
Pubblicazione: (2025)
di: Yan, Xianglong, et al.
Pubblicazione: (2025)
EliteKV: Scalable KV Cache Compression via RoPE Frequency Selection and Joint Low-Rank Projection
di: Zhou, Yuhao, et al.
Pubblicazione: (2025)
di: Zhou, Yuhao, et al.
Pubblicazione: (2025)
Compressible Dynamics in Deep Overparameterized Low-Rank Learning & Adaptation
di: Yaras, Can, et al.
Pubblicazione: (2024)
di: Yaras, Can, et al.
Pubblicazione: (2024)
TLoRA: Tri-Matrix Low-Rank Adaptation of Large Language Models
di: Islam, Tanvir
Pubblicazione: (2025)
di: Islam, Tanvir
Pubblicazione: (2025)
Efficient Low Rank Attention for Long-Context Inference in Large Language Models
di: Li, Tenghui, et al.
Pubblicazione: (2025)
di: Li, Tenghui, et al.
Pubblicazione: (2025)
Model Stealing for Any Low-Rank Language Model
di: Liu, Allen, et al.
Pubblicazione: (2024)
di: Liu, Allen, et al.
Pubblicazione: (2024)
Dynamic Model Selection for Trajectory Prediction via Pairwise Ranking and Meta-Features
di: Bowen, Lu
Pubblicazione: (2025)
di: Bowen, Lu
Pubblicazione: (2025)
Global Low-Rank, Local Full-Rank: The Holographic Encoding of Learned Algorithms
di: Xu, Yongzhong
Pubblicazione: (2026)
di: Xu, Yongzhong
Pubblicazione: (2026)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
di: Muñoz, J. Pablo, et al.
Pubblicazione: (2025)
di: Muñoz, J. Pablo, et al.
Pubblicazione: (2025)
Eigen Attention: Attention in Low-Rank Space for KV Cache Compression
di: Saxena, Utkarsh, et al.
Pubblicazione: (2024)
di: Saxena, Utkarsh, et al.
Pubblicazione: (2024)
ELAS: Efficient Pre-Training of Low-Rank Large Language Models via 2:4 Activation Sparsity
di: Li, Jiaxi, et al.
Pubblicazione: (2026)
di: Li, Jiaxi, et al.
Pubblicazione: (2026)
Sequences of Logits Reveal the Low Rank Structure of Language Models
di: Golowich, Noah, et al.
Pubblicazione: (2025)
di: Golowich, Noah, et al.
Pubblicazione: (2025)
Multiple Choice Learning of Low-Rank Adapters for Language Modeling
di: Letzelter, Victor, et al.
Pubblicazione: (2025)
di: Letzelter, Victor, et al.
Pubblicazione: (2025)
LoLCATs: On Low-Rank Linearizing of Large Language Models
di: Zhang, Michael, et al.
Pubblicazione: (2024)
di: Zhang, Michael, et al.
Pubblicazione: (2024)
Sparsity-Aware Low-Rank Representation for Efficient Fine-Tuning of Large Language Models
di: Zhang, Longteng, et al.
Pubblicazione: (2026)
di: Zhang, Longteng, et al.
Pubblicazione: (2026)
Towards Symmetric Low-Rank Adapters
di: Panoutsos, Tales, et al.
Pubblicazione: (2025)
di: Panoutsos, Tales, et al.
Pubblicazione: (2025)
The Primacy of Magnitude in Low-Rank Adaptation
di: Zhang, Zicheng, et al.
Pubblicazione: (2025)
di: Zhang, Zicheng, et al.
Pubblicazione: (2025)
Mixture-of-Subspaces in Low-Rank Adaptation
di: Wu, Taiqiang, et al.
Pubblicazione: (2024)
di: Wu, Taiqiang, et al.
Pubblicazione: (2024)
Provably Learning from Modern Language Models via Low Logit Rank
di: Golowich, Noah, et al.
Pubblicazione: (2025)
di: Golowich, Noah, et al.
Pubblicazione: (2025)
OjaKV: Context-Aware Online Low-Rank KV Cache Compression
di: Zhu, Yuxuan, et al.
Pubblicazione: (2025)
di: Zhu, Yuxuan, et al.
Pubblicazione: (2025)
TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models
di: Mu, Lin, et al.
Pubblicazione: (2026)
di: Mu, Lin, et al.
Pubblicazione: (2026)
Polynomial Expansion Rank Adaptation: Enhancing Low-Rank Fine-Tuning with High-Order Interactions
di: Zhang, Wenhao, et al.
Pubblicazione: (2026)
di: Zhang, Wenhao, et al.
Pubblicazione: (2026)
Training-Free Bayesianization for Low-Rank Adapters of Large Language Models
di: Shi, Haizhou, et al.
Pubblicazione: (2024)
di: Shi, Haizhou, et al.
Pubblicazione: (2024)
MoR: Mixture of Ranks for Low-Rank Adaptation Tuning
di: Tang, Chuanyu, et al.
Pubblicazione: (2024)
di: Tang, Chuanyu, et al.
Pubblicazione: (2024)
LANCE: Low Rank Activation Compression for Efficient On-Device Continual Learning
di: Apolinario, Marco Paul E., et al.
Pubblicazione: (2025)
di: Apolinario, Marco Paul E., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Low-Rank Adversarial PGD Attack
di: Savostianova, Dayana, et al.
Pubblicazione: (2024) -
Lillama: Large Language Models Compression via Low-Rank Feature Distillation
di: Sy, Yaya, et al.
Pubblicazione: (2024) -
Compressing Large Language Models using Low Rank and Low Precision Decomposition
di: Saha, Rajarshi, et al.
Pubblicazione: (2024) -
Memory-Efficient Fine-Tuning via Low-Rank Activation Compression
di: Shi, Jiang-Xin, et al.
Pubblicazione: (2025) -
Lossless Model Compression via Joint Low-Rank Factorization Optimization
di: Zhang, Boyang, et al.
Pubblicazione: (2024)