Saved in:
Bibliographic Details
Main Authors: Liu, Lihui, Kim, Jinha, Bansal, Vidit
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2404.08701
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916204335071232
author Liu, Lihui
Kim, Jinha
Bansal, Vidit
author_facet Liu, Lihui
Kim, Jinha
Bansal, Vidit
contents Recent advancements in contrastive learning have revolutionized self-supervised representation learning and achieved state-of-the-art performance on benchmark tasks. While most existing methods focus on applying contrastive learning to input data modalities such as images, natural language sentences, or networks, they overlook the potential of utilizing outputs from previously trained encoders. In this paper, we introduce SIMSKIP, a novel contrastive learning framework that specifically refines input embeddings for downstream tasks. Unlike traditional unsupervised learning approaches, SIMSKIP takes advantage of the output embeddings of encoder models as its input. Through theoretical analysis, we provide evidence that applying SIMSKIP does not result in larger upper bounds on downstream task errors than those of the original embeddings, which serve as SIMSKIP's input. Experimental results on various open datasets demonstrate that the embeddings produced by SIMSKIP improve performance on downstream tasks.
format Preprint
id arxiv_https___arxiv_org_abs_2404_08701
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Can Contrastive Learning Refine Embeddings
Liu, Lihui
Kim, Jinha
Bansal, Vidit
Machine Learning
Recent advancements in contrastive learning have revolutionized self-supervised representation learning and achieved state-of-the-art performance on benchmark tasks. While most existing methods focus on applying contrastive learning to input data modalities such as images, natural language sentences, or networks, they overlook the potential of utilizing outputs from previously trained encoders. In this paper, we introduce SIMSKIP, a novel contrastive learning framework that specifically refines input embeddings for downstream tasks. Unlike traditional unsupervised learning approaches, SIMSKIP takes advantage of the output embeddings of encoder models as its input. Through theoretical analysis, we provide evidence that applying SIMSKIP does not result in larger upper bounds on downstream task errors than those of the original embeddings, which serve as SIMSKIP's input. Experimental results on various open datasets demonstrate that the embeddings produced by SIMSKIP improve performance on downstream tasks.
title Can Contrastive Learning Refine Embeddings
topic Machine Learning
url https://arxiv.org/abs/2404.08701