WSVD: Weighted Low-Rank Approximation for Fast and Efficient Execution of Low-Precision Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Haiyu, Wang, Yutong, Jiang, Jack, Zhang, Sai Qian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
QSVD: Efficient Low-rank Approximation for Unified Query-Key-Value Weight Compression in Low-Precision Vision-Language Models
by: Wang, Yutong, et al.
Published: (2025)
by: Wang, Yutong, et al.
Published: (2025)
LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models
by: Wang, Haiyu, et al.
Published: (2026)
by: Wang, Haiyu, et al.
Published: (2026)
MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation
by: Ghiasvand, Sajjad, et al.
Published: (2026)
by: Ghiasvand, Sajjad, et al.
Published: (2026)
ALoRE: Efficient Visual Adaptation via Aggregating Low Rank Experts
by: Du, Sinan, et al.
Published: (2024)
by: Du, Sinan, et al.
Published: (2024)
RoPeSLR: 3D RoPE-driven Sparse-LowRank Attention for Efficient Diffusion Transformers
by: Liu, Yuxi, et al.
Published: (2026)
by: Liu, Yuxi, et al.
Published: (2026)
Weight Copy and Low-Rank Adaptation for Few-Shot Distillation of Vision Transformers
by: Grigore, Diana-Nicoleta, et al.
Published: (2024)
by: Grigore, Diana-Nicoleta, et al.
Published: (2024)
Fast Learnings of Coupled Nonnegative Tensor Decomposition Using Optimal Gradient and Low-rank Approximation
by: Wang, Xiulin, et al.
Published: (2023)
by: Wang, Xiulin, et al.
Published: (2023)
DL-QAT: Weight-Decomposed Low-Rank Quantization-Aware Training for Large Language Models
by: Ke, Wenjin, et al.
Published: (2025)
by: Ke, Wenjin, et al.
Published: (2025)
RankFeat&RankWeight: Rank-1 Feature/Weight Removal for Out-of-distribution Detection
by: Song, Yue, et al.
Published: (2023)
by: Song, Yue, et al.
Published: (2023)
Robust PCA Based on Adaptive Weighted Least Squares and Low-Rank Matrix Factorization
by: Li, Kexin, et al.
Published: (2024)
by: Li, Kexin, et al.
Published: (2024)
Parameter Efficient Continual Learning with Dynamic Low-Rank Adaptation
by: Bhat, Prashant Shivaram, et al.
Published: (2025)
by: Bhat, Prashant Shivaram, et al.
Published: (2025)
RankSEG-RMA: An Efficient Segmentation Algorithm via Reciprocal Moment Approximation
by: Wang, Zixun, et al.
Published: (2025)
by: Wang, Zixun, et al.
Published: (2025)
CLoRA: Parameter-Efficient Continual Learning with Low-Rank Adaptation
by: Muralidhara, Shishir, et al.
Published: (2025)
by: Muralidhara, Shishir, et al.
Published: (2025)
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks
by: Ding, Yuhe, et al.
Published: (2024)
by: Ding, Yuhe, et al.
Published: (2024)
A Systematic Review of Low-Rank and Local Low-Rank Matrix Approximation in Big Data Medical Imaging
by: Hamlomo, Sisipho, et al.
Published: (2024)
by: Hamlomo, Sisipho, et al.
Published: (2024)
Learning Low-Rank Feature for Thorax Disease Classification
by: Goel, Rajeev, et al.
Published: (2024)
by: Goel, Rajeev, et al.
Published: (2024)
LoRAPrune: Structured Pruning Meets Low-Rank Parameter-Efficient Fine-Tuning
by: Zhang, Mingyang, et al.
Published: (2023)
by: Zhang, Mingyang, et al.
Published: (2023)
Low Rank Support Quaternion Matrix Machine
by: Chen, Wang, et al.
Published: (2025)
by: Chen, Wang, et al.
Published: (2025)
DoRAN: Stabilizing Weight-Decomposed Low-Rank Adaptation via Noise Injection and Auxiliary Networks
by: Diep, Nghiem T., et al.
Published: (2025)
by: Diep, Nghiem T., et al.
Published: (2025)
FastVLM: Efficient Vision Encoding for Vision Language Models
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
CAPA: Contribution-Aware Pruning and FFN Approximation for Efficient Large Vision-Language Models
by: Jha, Samyak, et al.
Published: (2026)
by: Jha, Samyak, et al.
Published: (2026)
HLQ: Fast and Efficient Backpropagation via Hadamard Low-rank Quantization
by: Kim, Seonggon, et al.
Published: (2024)
by: Kim, Seonggon, et al.
Published: (2024)
Rethinking Low-Rank Adaptation in Vision: Exploring Head-Level Responsiveness across Diverse Tasks
by: Zhong, Yibo, et al.
Published: (2024)
by: Zhong, Yibo, et al.
Published: (2024)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
by: Zhu, Lianghui, et al.
Published: (2024)
by: Zhu, Lianghui, et al.
Published: (2024)
SVDQuant: Absorbing Outliers by Low-Rank Components for 4-Bit Diffusion Models
by: Li, Muyang, et al.
Published: (2024)
by: Li, Muyang, et al.
Published: (2024)
Efficient and Long-Tailed Generalization for Pre-trained Vision-Language Model
by: Shi, Jiang-Xin, et al.
Published: (2024)
by: Shi, Jiang-Xin, et al.
Published: (2024)
Low-Rank Similarity Mining for Multimodal Dataset Distillation
by: Xu, Yue, et al.
Published: (2024)
by: Xu, Yue, et al.
Published: (2024)
Continual Low-Rank Scaled Dot-product Attention
by: Picón, Ginés Carreto, et al.
Published: (2024)
by: Picón, Ginés Carreto, et al.
Published: (2024)
Rank Matters: Understanding and Defending Model Inversion Attacks via Low-Rank Feature Filtering
by: Yu, Hongyao, et al.
Published: (2024)
by: Yu, Hongyao, et al.
Published: (2024)
OTLRM: Orthogonal Learning-based Low-Rank Metric for Multi-Dimensional Inverse Problems
by: Wang, Xiangming, et al.
Published: (2024)
by: Wang, Xiangming, et al.
Published: (2024)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
Fast-Slow Efficient Training for Multimodal Large Language Models via Visual Token Pruning
by: Zhang, Dingkun, et al.
Published: (2026)
by: Zhang, Dingkun, et al.
Published: (2026)
MLoRQ: Bridging Low-Rank and Quantization for Transformer Compression
by: Gordon, Ofir, et al.
Published: (2025)
by: Gordon, Ofir, et al.
Published: (2025)
Replay-Free Continual Low-Rank Adaptation with Dynamic Memory
by: Chen, Huancheng, et al.
Published: (2024)
by: Chen, Huancheng, et al.
Published: (2024)
FREE: Fast and Robust Vision Language Models with Early Exits
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
Privacy-Preserving Low-Rank Adaptation against Membership Inference Attacks for Latent Diffusion Models
by: Luo, Zihao, et al.
Published: (2024)
by: Luo, Zihao, et al.
Published: (2024)
Extreme Model Compression with Structured Sparsity at Low Precision
by: Liu, Dan, et al.
Published: (2025)
by: Liu, Dan, et al.
Published: (2025)
Self-Supervised Weight Templates for Scalable Vision Model Initialization
by: Xie, Yucheng, et al.
Published: (2026)
by: Xie, Yucheng, et al.
Published: (2026)
VCR: A Task for Pixel-Level Complex Reasoning in Vision Language Models via Restoring Occluded Text
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
by: Lu, Ziqi, et al.
Published: (2024)
by: Lu, Ziqi, et al.
Published: (2024)
Similar Items
-
QSVD: Efficient Low-rank Approximation for Unified Query-Key-Value Weight Compression in Low-Precision Vision-Language Models
by: Wang, Yutong, et al.
Published: (2025) -
LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models
by: Wang, Haiyu, et al.
Published: (2026) -
MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation
by: Ghiasvand, Sajjad, et al.
Published: (2026) -
ALoRE: Efficient Visual Adaptation via Aggregating Low Rank Experts
by: Du, Sinan, et al.
Published: (2024) -
RoPeSLR: 3D RoPE-driven Sparse-LowRank Attention for Efficient Diffusion Transformers
by: Liu, Yuxi, et al.
Published: (2026)