DoRA: Enhancing Parameter-Efficient Fine-Tuning with Dynamic Rank Distribution

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Mao, Yulong, Huang, Kaiyu, Guan, Changhao, Bao, Ganglin, Mo, Fengran, Xu, Jinan
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916301166870528
author Mao, Yulong
Huang, Kaiyu
Guan, Changhao
Bao, Ganglin
Mo, Fengran
Xu, Jinan
author_facet Mao, Yulong
Huang, Kaiyu
Guan, Changhao
Bao, Ganglin
Mo, Fengran
Xu, Jinan
contents Fine-tuning large-scale pre-trained models is inherently a resource-intensive task. While it can enhance the capabilities of the model, it also incurs substantial computational costs, posing challenges to the practical application of downstream tasks. Existing parameter-efficient fine-tuning (PEFT) methods such as Low-Rank Adaptation (LoRA) rely on a bypass framework that ignores the differential parameter budget requirements across weight matrices, which may lead to suboptimal fine-tuning outcomes. To address this issue, we introduce the Dynamic Low-Rank Adaptation (DoRA) method. DoRA decomposes high-rank LoRA layers into structured single-rank components, allowing for dynamic pruning of parameter budget based on their importance to specific tasks during training, which makes the most of the limited parameter budget. Experimental results demonstrate that DoRA can achieve competitive performance compared with LoRA and full model fine-tuning, and outperform various strong baselines with the same storage parameter budget. Our code is available at https://github.com/MIkumikumi0116/DoRA
format Preprint
id arxiv_https___arxiv_org_abs_2405_17357
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle DoRA: Enhancing Parameter-Efficient Fine-Tuning with Dynamic Rank Distribution
Mao, Yulong
Huang, Kaiyu
Guan, Changhao
Bao, Ganglin
Mo, Fengran
Xu, Jinan
Computation and Language
Fine-tuning large-scale pre-trained models is inherently a resource-intensive task. While it can enhance the capabilities of the model, it also incurs substantial computational costs, posing challenges to the practical application of downstream tasks. Existing parameter-efficient fine-tuning (PEFT) methods such as Low-Rank Adaptation (LoRA) rely on a bypass framework that ignores the differential parameter budget requirements across weight matrices, which may lead to suboptimal fine-tuning outcomes. To address this issue, we introduce the Dynamic Low-Rank Adaptation (DoRA) method. DoRA decomposes high-rank LoRA layers into structured single-rank components, allowing for dynamic pruning of parameter budget based on their importance to specific tasks during training, which makes the most of the limited parameter budget. Experimental results demonstrate that DoRA can achieve competitive performance compared with LoRA and full model fine-tuning, and outperform various strong baselines with the same storage parameter budget. Our code is available at https://github.com/MIkumikumi0116/DoRA
title DoRA: Enhancing Parameter-Efficient Fine-Tuning with Dynamic Rank Distribution
topic Computation and Language
url https://arxiv.org/abs/2405.17357