KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Tran, Quyen, Phan, Hoang, Tran, Lam, Than, Khoat, Tran, Toan, Phung, Dinh, Le, Trung
Format: Preprint
Publié: 2023
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866912128218169344
author Tran, Quyen
Phan, Hoang
Tran, Lam
Than, Khoat
Tran, Toan
Phung, Dinh
Le, Trung
author_facet Tran, Quyen
Phan, Hoang
Tran, Lam
Than, Khoat
Tran, Toan
Phung, Dinh
Le, Trung
contents Drawing inspiration from prompt tuning techniques applied to Large Language Models, recent methods based on pre-trained ViT networks have achieved remarkable results in the field of Continual Learning. Specifically, these approaches propose to maintain a set of prompts and allocate a subset of them to learn each task using a key-query matching strategy. However, they may encounter limitations when lacking control over the correlations between old task queries and keys of future tasks, the shift of features in the latent space, and the relative separation of latent vectors learned in independent tasks. In this work, we introduce a novel key-query learning strategy based on orthogonal projection, inspired by model-agnostic meta-learning, to enhance prompt matching efficiency and address the challenge of shifting features. Furthermore, we introduce a One-Versus-All (OVA) prototype-based component that enhances the classification head distinction. Experimental results on benchmark datasets demonstrate that our method empowers the model to achieve results surpassing those of current state-of-the-art approaches by a large margin of up to 20%.
format Preprint
id arxiv_https___arxiv_org_abs_2311_15414
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All
Tran, Quyen
Phan, Hoang
Tran, Lam
Than, Khoat
Tran, Toan
Phung, Dinh
Le, Trung
Machine Learning
Computer Vision and Pattern Recognition
Drawing inspiration from prompt tuning techniques applied to Large Language Models, recent methods based on pre-trained ViT networks have achieved remarkable results in the field of Continual Learning. Specifically, these approaches propose to maintain a set of prompts and allocate a subset of them to learn each task using a key-query matching strategy. However, they may encounter limitations when lacking control over the correlations between old task queries and keys of future tasks, the shift of features in the latent space, and the relative separation of latent vectors learned in independent tasks. In this work, we introduce a novel key-query learning strategy based on orthogonal projection, inspired by model-agnostic meta-learning, to enhance prompt matching efficiency and address the challenge of shifting features. Furthermore, we introduce a One-Versus-All (OVA) prototype-based component that enhances the classification head distinction. Experimental results on benchmark datasets demonstrate that our method empowers the model to achieve results surpassing those of current state-of-the-art approaches by a large margin of up to 20%.
title KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All
topic Machine Learning
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2311.15414