CALR: Corrective Adaptive Low-Rank Decomposition for Efficient Large Language Model Layer Compression
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kautsar, Muchammad Daniyal, Hariono, Afra Majida, Widyawan, Alfarozi, Syukron Abu Ishaq, Woraratpanya, Kuntpong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Attention vs LSTM: Improving Word-level BISINDO Recognition
von: Kautsar, Muchammad Daniyal, et al.
Veröffentlicht: (2024)
von: Kautsar, Muchammad Daniyal, et al.
Veröffentlicht: (2024)
The Application of Artificial Neural Network Model to Predicting the Acid Mine Drainage from Long-Term Lab Scale Kinetic Test
von: Abfertiawan, Muhammad Sonny, et al.
Veröffentlicht: (2024)
von: Abfertiawan, Muhammad Sonny, et al.
Veröffentlicht: (2024)
Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports
von: Singh, Punit Kumar, et al.
Veröffentlicht: (2025)
von: Singh, Punit Kumar, et al.
Veröffentlicht: (2025)
Compressing Large Language Models using Low Rank and Low Precision Decomposition
von: Saha, Rajarshi, et al.
Veröffentlicht: (2024)
von: Saha, Rajarshi, et al.
Veröffentlicht: (2024)
ARA: Adaptive Rank Allocation for Efficient Large Language Model SVD Compression
von: Xv, Lin, et al.
Veröffentlicht: (2025)
von: Xv, Lin, et al.
Veröffentlicht: (2025)
SoLA: Leveraging Soft Activation Sparsity and Low-Rank Decomposition for Large Language Model Compression
von: Huang, Xinhao, et al.
Veröffentlicht: (2026)
von: Huang, Xinhao, et al.
Veröffentlicht: (2026)
Adaptive Feature-based Low-Rank Compression of Large Language Models via Bayesian Optimization
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
Mission im kolonialen Umfeld - Deutsche protestantische Missionsgesellschaften in Deutsch-Ostafrika
von: Hamilton, Majida,
Veröffentlicht: (2016)
von: Hamilton, Majida,
Veröffentlicht: (2016)
Mission im kolonialen Umfeld
von: Hamilton, Majida
Veröffentlicht: (2020)
von: Hamilton, Majida
Veröffentlicht: (2020)
Prompt Injection as an Emerging Threat: Evaluating the Resilience of Large Language Models
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
A Hybrid Machine Learning Framework for Systematic Trading in Cryptocurrency and FX Markets
von: Izzuddin, Muchammad Fikri
Veröffentlicht: (2025)
von: Izzuddin, Muchammad Fikri
Veröffentlicht: (2025)
Trustworthiness Calibration Framework for Phishing Email Detection Using Large Language Models
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
FlexRank: Nested Low-Rank Knowledge Decomposition for Adaptive Model Deployment
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2026)
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2026)
A Systematic Literature Review of Fleksibilitas Dan Budaya Organisasi Penting Bagi UKM Di Pasar Yang Kompetitif
von: Yani Dwi Restanti, et al.
Veröffentlicht: (2025)
von: Yani Dwi Restanti, et al.
Veröffentlicht: (2025)
Importance-Guided Basis Selection for Low-Rank Decomposition of Large Language Models
von: Asante, Daniel Agyei, et al.
Veröffentlicht: (2026)
von: Asante, Daniel Agyei, et al.
Veröffentlicht: (2026)
Compressed BC-LISTA via Low-Rank Convolutional Decomposition
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
Convolutional Neural Network Compression Based on Low-Rank Decomposition
von: He, Yaping, et al.
Veröffentlicht: (2024)
von: He, Yaping, et al.
Veröffentlicht: (2024)
Lillama: Large Language Models Compression via Low-Rank Feature Distillation
von: Sy, Yaya, et al.
Veröffentlicht: (2024)
von: Sy, Yaya, et al.
Veröffentlicht: (2024)
Basis Selection: Low-Rank Decomposition of Pretrained Large Language Models for Target Applications
von: Li, Yang, et al.
Veröffentlicht: (2024)
von: Li, Yang, et al.
Veröffentlicht: (2024)
Low-Rank Compression of Language Models via Differentiable Rank Selection
von: Sundrani, Sidhant, et al.
Veröffentlicht: (2025)
von: Sundrani, Sidhant, et al.
Veröffentlicht: (2025)
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition
von: Xu, Mingxue, et al.
Veröffentlicht: (2023)
von: Xu, Mingxue, et al.
Veröffentlicht: (2023)
SkipCat: Rank-Maximized Low-Rank Compression of Large Language Models via Shared Projection and Block Skipping
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
Interpretable Bayesian Tensor Network Kernel Machines with Automatic Rank and Feature Selection
von: Kilic, Afra, et al.
Veröffentlicht: (2025)
von: Kilic, Afra, et al.
Veröffentlicht: (2025)
Reinterpretasi Konsep Pendidikan Tauhid Muhammad bin Abdul Wahhab dan Relevansinya dengan Kurikulum Merdeka PAI di Indonesia
von: abdbul qodir, Muchammad Idham Cholid
Veröffentlicht: (2025)
von: abdbul qodir, Muchammad Idham Cholid
Veröffentlicht: (2025)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
von: Zhang, Yuxuan
Veröffentlicht: (2025)
von: Zhang, Yuxuan
Veröffentlicht: (2025)
Evaluating Vision-Language and Large Language Models for Automated Student Assessment in Indonesian Classrooms
von: Aisyah, Nurul, et al.
Veröffentlicht: (2025)
von: Aisyah, Nurul, et al.
Veröffentlicht: (2025)
Dynamic Rank Reinforcement Learning for Adaptive Low-Rank Multi-Head Self Attention in Large Language Models
von: Erden, Caner
Veröffentlicht: (2025)
von: Erden, Caner
Veröffentlicht: (2025)
Efficient Low Rank Attention for Long-Context Inference in Large Language Models
von: Li, Tenghui, et al.
Veröffentlicht: (2025)
von: Li, Tenghui, et al.
Veröffentlicht: (2025)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models
von: Wang, Haiyu, et al.
Veröffentlicht: (2026)
von: Wang, Haiyu, et al.
Veröffentlicht: (2026)
1+1>2: A Synergistic Sparse and Low-Rank Compression Method for Large Language Models
von: Zong, Zeliang, et al.
Veröffentlicht: (2025)
von: Zong, Zeliang, et al.
Veröffentlicht: (2025)
Large Language Model Compression with Global Rank and Sparsity Optimization
von: Zhou, Changhai, et al.
Veröffentlicht: (2025)
von: Zhou, Changhai, et al.
Veröffentlicht: (2025)
Streamlining Redundant Layers to Compress Large Language Models
von: Chen, Xiaodong, et al.
Veröffentlicht: (2024)
von: Chen, Xiaodong, et al.
Veröffentlicht: (2024)
Adaptive Regularized Low-Rank Tensor Decomposition for Hyperspectral Image Denoising and Destriping
von: Li, Dongyi, et al.
Veröffentlicht: (2024)
von: Li, Dongyi, et al.
Veröffentlicht: (2024)
Dynamic Adaptive Rank Space Exploration for Efficient Sentiment Analysis with Large Language Models
von: Ding, Hongcheng, et al.
Veröffentlicht: (2024)
von: Ding, Hongcheng, et al.
Veröffentlicht: (2024)
Low-Rank Knowledge Decomposition for Medical Foundation Models
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
MoDeGPT: Modular Decomposition for Large Language Model Compression
von: Lin, Chi-Heng, et al.
Veröffentlicht: (2024)
von: Lin, Chi-Heng, et al.
Veröffentlicht: (2024)
Large Language Model Compression via the Nested Activation-Aware Decomposition
von: Lu, Jun, et al.
Veröffentlicht: (2025)
von: Lu, Jun, et al.
Veröffentlicht: (2025)
FFT-based Dynamic Subspace Selection for Low-Rank Adaptive Optimization of Large Language Models
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2025)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Attention vs LSTM: Improving Word-level BISINDO Recognition
von: Kautsar, Muchammad Daniyal, et al.
Veröffentlicht: (2024) -
The Application of Artificial Neural Network Model to Predicting the Acid Mine Drainage from Long-Term Lab Scale Kinetic Test
von: Abfertiawan, Muhammad Sonny, et al.
Veröffentlicht: (2024) -
Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports
von: Singh, Punit Kumar, et al.
Veröffentlicht: (2025) -
Compressing Large Language Models using Low Rank and Low Precision Decomposition
von: Saha, Rajarshi, et al.
Veröffentlicht: (2024) -
ARA: Adaptive Rank Allocation for Efficient Large Language Model SVD Compression
von: Xv, Lin, et al.
Veröffentlicht: (2025)