A Systematic Comparison of Contextualized Word Embeddings for Lexical Semantic Change

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Periti, Francesco, Tahmasebi, Nina
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929567189434368
author Periti, Francesco
Tahmasebi, Nina
author_facet Periti, Francesco
Tahmasebi, Nina
contents Contextualized embeddings are the preferred tool for modeling Lexical Semantic Change (LSC). Current evaluations typically focus on a specific task known as Graded Change Detection (GCD). However, performance comparison across work are often misleading due to their reliance on diverse settings. In this paper, we evaluate state-of-the-art models and approaches for GCD under equal conditions. We further break the LSC problem into Word-in-Context (WiC) and Word Sense Induction (WSI) tasks, and compare models across these different levels. Our evaluation is performed across different languages on eight available benchmarks for LSC, and shows that (i) APD outperforms other approaches for GCD; (ii) XL-LEXEME outperforms other contextualized models for WiC, WSI, and GCD, while being comparable to GPT-4; (iii) there is a clear need for improving the modeling of word meanings, as well as focus on how, when, and why these meanings change, rather than solely focusing on the extent of semantic change.
format Preprint
id arxiv_https___arxiv_org_abs_2402_12011
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle A Systematic Comparison of Contextualized Word Embeddings for Lexical Semantic Change
Periti, Francesco
Tahmasebi, Nina
Computation and Language
Contextualized embeddings are the preferred tool for modeling Lexical Semantic Change (LSC). Current evaluations typically focus on a specific task known as Graded Change Detection (GCD). However, performance comparison across work are often misleading due to their reliance on diverse settings. In this paper, we evaluate state-of-the-art models and approaches for GCD under equal conditions. We further break the LSC problem into Word-in-Context (WiC) and Word Sense Induction (WSI) tasks, and compare models across these different levels. Our evaluation is performed across different languages on eight available benchmarks for LSC, and shows that (i) APD outperforms other approaches for GCD; (ii) XL-LEXEME outperforms other contextualized models for WiC, WSI, and GCD, while being comparable to GPT-4; (iii) there is a clear need for improving the modeling of word meanings, as well as focus on how, when, and why these meanings change, rather than solely focusing on the extent of semantic change.
title A Systematic Comparison of Contextualized Word Embeddings for Lexical Semantic Change
topic Computation and Language
url https://arxiv.org/abs/2402.12011