3BASiL: An Algorithmic Framework for Sparse plus Low-Rank Compression of LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Makni, Mehdi, Meng, Xiang, Mazumder, Rahul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HASSLE-free: A unified Framework for Sparse plus Low-Rank Matrix Decomposition for LLMs
by: Makni, Mehdi, et al.
Published: (2025)
by: Makni, Mehdi, et al.
Published: (2025)
TSENOR: Highly-Efficient Algorithm for Finding Transposable N:M Sparse Masks
by: Meng, Xiang, et al.
Published: (2025)
by: Meng, Xiang, et al.
Published: (2025)
ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models
by: Lucas, Ryan, et al.
Published: (2026)
by: Lucas, Ryan, et al.
Published: (2026)
An Optimization Framework for Differentially Private Sparse Fine-Tuning
by: Makni, Mehdi, et al.
Published: (2025)
by: Makni, Mehdi, et al.
Published: (2025)
ALPS: Improved Optimization for Highly Sparse One-Shot Pruning for Large Language Models
by: Meng, Xiang, et al.
Published: (2024)
by: Meng, Xiang, et al.
Published: (2024)
Sparse NMF with Archetypal Regularization: Computational and Robustness Properties
by: Behdin, Kayhan, et al.
Published: (2021)
by: Behdin, Kayhan, et al.
Published: (2021)
MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models
by: Afriat, Gabriel, et al.
Published: (2026)
by: Afriat, Gabriel, et al.
Published: (2026)
FAST: An Optimization Framework for Fast Additive Segmentation in Transparent ML
by: Liu, Brian, et al.
Published: (2024)
by: Liu, Brian, et al.
Published: (2024)
Computation of Least Trimmed Squares: A Branch-and-Bound framework with Hyperplane Arrangement Enhancements
by: Meng, Xiang, et al.
Published: (2026)
by: Meng, Xiang, et al.
Published: (2026)
FALCON: FLOP-Aware Combinatorial Optimization for Neural Network Pruning
by: Meng, Xiang, et al.
Published: (2024)
by: Meng, Xiang, et al.
Published: (2024)
Preserving Deep Representations In One-Shot Pruning: A Hessian-Free Second-Order Optimization Framework
by: Lucas, Ryan, et al.
Published: (2024)
by: Lucas, Ryan, et al.
Published: (2024)
Sparse Gaussian Graphical Models with Discrete Optimization: Computational and Statistical Perspectives
by: Behdin, Kayhan, et al.
Published: (2023)
by: Behdin, Kayhan, et al.
Published: (2023)
Hierarchical Sparse Plus Low Rank Compression of LLM
by: Kumar, Pawan, et al.
Published: (2025)
by: Kumar, Pawan, et al.
Published: (2025)
LoLA: Low-Rank Linear Attention With Sparse Caching
by: McDermott, Luke, et al.
Published: (2025)
by: McDermott, Luke, et al.
Published: (2025)
Randomization Can Reduce Both Bias and Variance: A Case Study in Random Forests
by: Liu, Brian, et al.
Published: (2024)
by: Liu, Brian, et al.
Published: (2024)
Theoretical Compression Bounds for Wide Multilayer Perceptrons
by: Cheairi, Houssam El, et al.
Published: (2025)
by: Cheairi, Houssam El, et al.
Published: (2025)
MOSS: Multi-Objective Optimization for Stable Rule Sets
by: Liu, Brian, et al.
Published: (2025)
by: Liu, Brian, et al.
Published: (2025)
Learning on Transformers is Provable Low-Rank and Sparse: A One-layer Analysis
by: Li, Hongkang, et al.
Published: (2024)
by: Li, Hongkang, et al.
Published: (2024)
DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights
by: Mikaelyan, Liana, et al.
Published: (2025)
by: Mikaelyan, Liana, et al.
Published: (2025)
MGAA: Multi-Granular Adaptive Allocation fof Low-Rank Compression of LLMs
by: Li, Guangyan, et al.
Published: (2025)
by: Li, Guangyan, et al.
Published: (2025)
Low-Rank Adapting Models for Sparse Autoencoders
by: Chen, Matthew, et al.
Published: (2025)
by: Chen, Matthew, et al.
Published: (2025)
Scalable Variational Bayesian Fine-Tuning of LLMs via Orthogonalized Low-Rank Adapters
by: Xiang, Haotian, et al.
Published: (2026)
by: Xiang, Haotian, et al.
Published: (2026)
Low-Rank Correction for Quantized LLMs
by: Scetbon, Meyer, et al.
Published: (2024)
by: Scetbon, Meyer, et al.
Published: (2024)
Extracting Interpretable Models from Tree Ensembles: Computational and Statistical Perspectives
by: Liu, Brian, et al.
Published: (2025)
by: Liu, Brian, et al.
Published: (2025)
End-to-end Feature Selection Approach for Learning Skinny Trees
by: Ibrahim, Shibal, et al.
Published: (2023)
by: Ibrahim, Shibal, et al.
Published: (2023)
SLoPe: Double-Pruned Sparse Plus Lazy Low-Rank Adapter Pretraining of LLMs
by: Mozaffari, Mohammad, et al.
Published: (2024)
by: Mozaffari, Mohammad, et al.
Published: (2024)
LoRC: Low-Rank Compression for LLMs KV Cache with a Progressive Compression Strategy
by: Zhang, Rongzhi, et al.
Published: (2024)
by: Zhang, Rongzhi, et al.
Published: (2024)
Dynamic Low-Rank Sparse Adaptation for Large Language Models
by: Huang, Weizhong, et al.
Published: (2025)
by: Huang, Weizhong, et al.
Published: (2025)
Low-Rank Compression of Language Models via Differentiable Rank Selection
by: Sundrani, Sidhant, et al.
Published: (2025)
by: Sundrani, Sidhant, et al.
Published: (2025)
Low-Rank Prehab: Preparing Neural Networks for SVD Compression
by: Qin, Haoran, et al.
Published: (2025)
by: Qin, Haoran, et al.
Published: (2025)
A GPU-accelerated Nonlinear Branch-and-Bound Framework for Sparse Linear Models
by: Meng, Xiang, et al.
Published: (2026)
by: Meng, Xiang, et al.
Published: (2026)
OSSCAR: One-Shot Structured Pruning in Vision and Language Models with Combinatorial Optimization
by: Meng, Xiang, et al.
Published: (2024)
by: Meng, Xiang, et al.
Published: (2024)
Differentially Private High-dimensional Variable Selection via Integer Programming
by: Prastakos, Petros, et al.
Published: (2025)
by: Prastakos, Petros, et al.
Published: (2025)
Activation Compression in LLMs: Theoretical Analysis and Efficient Algorithm
by: Wei, Wen-Da, et al.
Published: (2026)
by: Wei, Wen-Da, et al.
Published: (2026)
FairLRF: Achieving Fairness through Sparse Low Rank Factorization
by: Guo, Yuanbo, et al.
Published: (2025)
by: Guo, Yuanbo, et al.
Published: (2025)
Low-Rank Matrix Approximation for Neural Network Compression
by: Cherukuri, Kalyan, et al.
Published: (2025)
by: Cherukuri, Kalyan, et al.
Published: (2025)
Palu: Compressing KV-Cache with Low-Rank Projection
by: Chang, Chi-Chih, et al.
Published: (2024)
by: Chang, Chi-Chih, et al.
Published: (2024)
Clustering-Based Low-Rank Matrix Approximation for Medical Image Compression
by: Hamlomo, Sisipho, et al.
Published: (2025)
by: Hamlomo, Sisipho, et al.
Published: (2025)
Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition
by: He, Zhengfu, et al.
Published: (2025)
by: He, Zhengfu, et al.
Published: (2025)
Zero Sum SVD: Balancing Loss Sensitivity for Low Rank LLM Compression
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
Similar Items
-
HASSLE-free: A unified Framework for Sparse plus Low-Rank Matrix Decomposition for LLMs
by: Makni, Mehdi, et al.
Published: (2025) -
TSENOR: Highly-Efficient Algorithm for Finding Transposable N:M Sparse Masks
by: Meng, Xiang, et al.
Published: (2025) -
ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models
by: Lucas, Ryan, et al.
Published: (2026) -
An Optimization Framework for Differentially Private Sparse Fine-Tuning
by: Makni, Mehdi, et al.
Published: (2025) -
ALPS: Improved Optimization for Highly Sparse One-Shot Pruning for Large Language Models
by: Meng, Xiang, et al.
Published: (2024)