Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Li, Changqun, Ding, Chaofan, Luan, Kexin, Di, Xinhan
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910743776985088
author Li, Changqun
Ding, Chaofan
Luan, Kexin
Di, Xinhan
author_facet Li, Changqun
Ding, Chaofan
Luan, Kexin
Di, Xinhan
contents Fine-tuning pre-trained large language models in a parameter-efficient manner is widely studied for its effectiveness and efficiency. LoRA is one of the most widely used methods, which assumes that the optimization process is essentially low dimensional. Although LoRA has demonstrated commendable performance, there remains a significant performance gap between LoRA and full fine-tuning when learning new tasks. In this work, we propose Low-Rank Adaptation with Task-Relevant Feature Enhancement(LoRATRF) for enhancing task-relevant features from the perspective of editing neural network representations. To prioritize task-relevant features, a task-aware filter that selectively extracts valuable knowledge from hidden representations for the target or current task is designed. As the experiments on a vareity of datasets including NLU, commonsense reasoning and mathematical reasoning tasks demonstrates, our method reduces 33.71% parameters and achieves better performance on a variety of datasets in comparison with SOTA low-rank methods.
format Preprint
id arxiv_https___arxiv_org_abs_2412_09827
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models
Li, Changqun
Ding, Chaofan
Luan, Kexin
Di, Xinhan
Computation and Language
Computer Vision and Pattern Recognition
Fine-tuning pre-trained large language models in a parameter-efficient manner is widely studied for its effectiveness and efficiency. LoRA is one of the most widely used methods, which assumes that the optimization process is essentially low dimensional. Although LoRA has demonstrated commendable performance, there remains a significant performance gap between LoRA and full fine-tuning when learning new tasks. In this work, we propose Low-Rank Adaptation with Task-Relevant Feature Enhancement(LoRATRF) for enhancing task-relevant features from the perspective of editing neural network representations. To prioritize task-relevant features, a task-aware filter that selectively extracts valuable knowledge from hidden representations for the target or current task is designed. As the experiments on a vareity of datasets including NLU, commonsense reasoning and mathematical reasoning tasks demonstrates, our method reduces 33.71% parameters and achieves better performance on a variety of datasets in comparison with SOTA low-rank methods.
title Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models
topic Computation and Language
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2412.09827