Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Zhang, Jia-Chen, Xiong, Yu-Jie, Xia, Chun-Ming, Zhu, Dong-Hai, Qiu, Xi-He
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866929739156946944
author Zhang, Jia-Chen
Xiong, Yu-Jie
Xia, Chun-Ming
Zhu, Dong-Hai
Qiu, Xi-He
author_facet Zhang, Jia-Chen
Xiong, Yu-Jie
Xia, Chun-Ming
Zhu, Dong-Hai
Qiu, Xi-He
contents Large language model (LLM) is considered a milestone towards achieving Artificial General Intelligence (AGI). With its advanced emergent capabilities, it adapt to a wide range of specific applications. Fine-tuning LLMs for various downstream tasks has become a new paradigm. Low-Rank Adaptation (LoRA) is well-known for its parameter efficiency. It can reduce the number of parameters needed to fine-tune LLMs by several orders of magnitude. However, LoRA-based approaches encounter a significant limitation due to the bottleneck imposed by rank one decomposition. As the parameters count in LLMs increase, even rank one decomposition might surpass the number of parameters truly necessary for handling more downstream tasks. In this paper, we propose a new method for Parameter-Efficient Fine-Tuning (PEFT) via deconvolution in subspace, dubbed as DCFT. We innovatively use deconvolution to complete details and enhance knowledge in subspace incremental matrices, and dynamically control parameters by adjusting the kernel size, unconstrained by rank-one decomposition. Extensive experiments are conducted to validate the effectiveness of DCFT. Results show that compared to LoRA, DCFT achieve an 8$\times$ reduction in parameters, and still achieves highly impressive performance. Our code is available here: https://github.com/Godz-z/DCFT.
format Preprint
id arxiv_https___arxiv_org_abs_2503_01419
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace
Zhang, Jia-Chen
Xiong, Yu-Jie
Xia, Chun-Ming
Zhu, Dong-Hai
Qiu, Xi-He
Computation and Language
Artificial Intelligence
Large language model (LLM) is considered a milestone towards achieving Artificial General Intelligence (AGI). With its advanced emergent capabilities, it adapt to a wide range of specific applications. Fine-tuning LLMs for various downstream tasks has become a new paradigm. Low-Rank Adaptation (LoRA) is well-known for its parameter efficiency. It can reduce the number of parameters needed to fine-tune LLMs by several orders of magnitude. However, LoRA-based approaches encounter a significant limitation due to the bottleneck imposed by rank one decomposition. As the parameters count in LLMs increase, even rank one decomposition might surpass the number of parameters truly necessary for handling more downstream tasks. In this paper, we propose a new method for Parameter-Efficient Fine-Tuning (PEFT) via deconvolution in subspace, dubbed as DCFT. We innovatively use deconvolution to complete details and enhance knowledge in subspace incremental matrices, and dynamically control parameters by adjusting the kernel size, unconstrained by rank-one decomposition. Extensive experiments are conducted to validate the effectiveness of DCFT. Results show that compared to LoRA, DCFT achieve an 8$\times$ reduction in parameters, and still achieves highly impressive performance. Our code is available here: https://github.com/Godz-z/DCFT.
title Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2503.01419