Assesing the Feature-Driven Nature of Similarity-based Sorting of Verbs

Fuente: Redalyc
Enregistré dans:
Détails bibliographiques
Auteur principal: Pinar Öztürk
Format: Artículo científico
Langue:en
Publié: Instituto Politécnico Nacional 2011
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1876467590830751744
author Pinar Öztürk
author_facet Pinar Öztürk
contents Assesing the Feature-Driven Nature of Similarity-based Sorting of Verbs Pinar Öztürk Mila Vulchanova Christian Tumyr Liliana Martinez David Kabath Computación similarity verb sorting Verb features The paper presents a computational analysis of the results from a sorting task with motion verbs in Norwegian. The sorting behavior of humans rests on the features they use when they compare two or more words. We investigate what these features are and how differential each feature may be in sorting. The key rationale for our method of analysis is the assumption that a sorting task rests on a similarity assessment process. The main idea is that a set of features underlies this similarity judgment, and similarity between two verbs amounts to the sum of the weighted similarity between the given set of features. The computational methodology used to investigate the features is as follows. Based on the frequency of co-occurrence of verbs in the human generated cluster, weights of a given set of features are computed using linear regression. The weights are used, in turn, to compute a similarity matrix between the verbs. This matrix is used as an input for the agglomerative hierarchical clustering. If the selected/projected set of features aligns with the features the participants used when sorting verbs in groups, then the clusters we obtain using this computational method would align with the clusters generated by humans. Otherwise, the method proceeds with modifying the feature set and repeating the process. Features promoting clusters that align with human-generated clusters are evaluated by a set of human experts and the results show that the method manages to identify the appropriate feature sets. This method can be applied in analyzing a variety of data ranging from experimental free production data, to linguistic data from controlled experiments in the assessment of semantic relations and hierarchies within languages and across languages. 2011 artículo científico 1870-9044 https://www.redalyc.org/articulo.oa?id=402640456002 en http://www.redalyc.org/revista.oa?id=4026 Polibits application/pdf Instituto Politécnico Nacional Polibits (México) Vol.43
format Artículo científico
id redalyc_402640456002
institution Redalyc
language en
publishDate 2011
publisher Instituto Politécnico Nacional
spellingShingle Assesing the Feature-Driven Nature of Similarity-based Sorting of Verbs
Pinar Öztürk
Computación
similarity
verb sorting
Verb features
Assesing the Feature-Driven Nature of Similarity-based Sorting of Verbs Pinar Öztürk Mila Vulchanova Christian Tumyr Liliana Martinez David Kabath Computación similarity verb sorting Verb features The paper presents a computational analysis of the results from a sorting task with motion verbs in Norwegian. The sorting behavior of humans rests on the features they use when they compare two or more words. We investigate what these features are and how differential each feature may be in sorting. The key rationale for our method of analysis is the assumption that a sorting task rests on a similarity assessment process. The main idea is that a set of features underlies this similarity judgment, and similarity between two verbs amounts to the sum of the weighted similarity between the given set of features. The computational methodology used to investigate the features is as follows. Based on the frequency of co-occurrence of verbs in the human generated cluster, weights of a given set of features are computed using linear regression. The weights are used, in turn, to compute a similarity matrix between the verbs. This matrix is used as an input for the agglomerative hierarchical clustering. If the selected/projected set of features aligns with the features the participants used when sorting verbs in groups, then the clusters we obtain using this computational method would align with the clusters generated by humans. Otherwise, the method proceeds with modifying the feature set and repeating the process. Features promoting clusters that align with human-generated clusters are evaluated by a set of human experts and the results show that the method manages to identify the appropriate feature sets. This method can be applied in analyzing a variety of data ranging from experimental free production data, to linguistic data from controlled experiments in the assessment of semantic relations and hierarchies within languages and across languages. 2011 artículo científico 1870-9044 https://www.redalyc.org/articulo.oa?id=402640456002 en http://www.redalyc.org/revista.oa?id=4026 Polibits application/pdf Instituto Politécnico Nacional Polibits (México) Vol.43
title Assesing the Feature-Driven Nature of Similarity-based Sorting of Verbs
topic Computación
similarity
verb sorting
Verb features
url https://www.redalyc.org/articulo.oa?id=402640456002