Multi-Task Dense Prediction via Mixture of Low-Rank Experts

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yang, Yuqi, Jiang, Peng-Tao, Hou, Qibin, Zhang, Hao, Chen, Jinwei, Li, Bo
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913363697598464
author Yang, Yuqi
Jiang, Peng-Tao
Hou, Qibin
Zhang, Hao
Chen, Jinwei
Li, Bo
author_facet Yang, Yuqi
Jiang, Peng-Tao
Hou, Qibin
Zhang, Hao
Chen, Jinwei
Li, Bo
contents Previous multi-task dense prediction methods based on the Mixture of Experts (MoE) have received great performance but they neglect the importance of explicitly modeling the global relations among all tasks. In this paper, we present a novel decoder-focused method for multi-task dense prediction, called Mixture-of-Low-Rank-Experts (MLoRE). To model the global task relationships, MLoRE adds a generic convolution path to the original MoE structure, where each task feature can go through this path for explicit parameter sharing. Furthermore, to control the parameters and computational cost brought by the increase in the number of experts, we take inspiration from LoRA and propose to leverage the low-rank format of a vanilla convolution in the expert network. Since the low-rank experts have fewer parameters and can be dynamically parameterized into the generic convolution, the parameters and computational cost do not change much with the increase of experts. Benefiting from this design, we increase the number of experts and its reception field to enlarge the representation capacity, facilitating multiple dense tasks learning in a unified network. Extensive experiments on the PASCAL-Context and NYUD-v2 benchmarks show that our MLoRE achieves superior performance compared to previous state-of-the-art methods on all metrics. Our code is available at https://github.com/YuqiYang213/MLoRE.
format Preprint
id arxiv_https___arxiv_org_abs_2403_17749
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Multi-Task Dense Prediction via Mixture of Low-Rank Experts
Yang, Yuqi
Jiang, Peng-Tao
Hou, Qibin
Zhang, Hao
Chen, Jinwei
Li, Bo
Computer Vision and Pattern Recognition
Previous multi-task dense prediction methods based on the Mixture of Experts (MoE) have received great performance but they neglect the importance of explicitly modeling the global relations among all tasks. In this paper, we present a novel decoder-focused method for multi-task dense prediction, called Mixture-of-Low-Rank-Experts (MLoRE). To model the global task relationships, MLoRE adds a generic convolution path to the original MoE structure, where each task feature can go through this path for explicit parameter sharing. Furthermore, to control the parameters and computational cost brought by the increase in the number of experts, we take inspiration from LoRA and propose to leverage the low-rank format of a vanilla convolution in the expert network. Since the low-rank experts have fewer parameters and can be dynamically parameterized into the generic convolution, the parameters and computational cost do not change much with the increase of experts. Benefiting from this design, we increase the number of experts and its reception field to enlarge the representation capacity, facilitating multiple dense tasks learning in a unified network. Extensive experiments on the PASCAL-Context and NYUD-v2 benchmarks show that our MLoRE achieves superior performance compared to previous state-of-the-art methods on all metrics. Our code is available at https://github.com/YuqiYang213/MLoRE.
title Multi-Task Dense Prediction via Mixture of Low-Rank Experts
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2403.17749