Phonological Neighbourhood Density in Dutch Verbs: From Classification to Corpus Annotation

Fuente: Zenodo
Saved in:
Bibliographic Details
Main Authors: Zimianiti, Eleni, Kievit, Rogier, Rowland, Caroline, Donnelly, Seamus
Format: Recurso digital
Language:Dutch
Published: Zenodo 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866902324211875840
author Zimianiti, Eleni
Kievit, Rogier
Rowland, Caroline
Donnelly, Seamus
author_facet Zimianiti, Eleni
Kievit, Rogier
Rowland, Caroline
Donnelly, Seamus
contents <p><span lang="EN-US">Phonological Neighbourhood Density (PND) quantifies how many phonologically similar verbs undergo the same type of stem alternation across tenses. In Dutch, for instance, <em>blijken, lijken</em>, and <em>bijten</em> form a neighbourhood through the transformation to <em>bleek, leek</em>, and <em>beet</em> in the simple past tense. PND has been widely studied in psycholinguistics, where denser neighbourhoods are linked to stronger analogical effects in lexical processing, morphological generalization, and language acquisition.</span></p> <p><span lang="EN-US">A detailed PND classification of 1,452 Dutch verbs has been developed based on systematic vowel stem alternations between the simple present, simple past, and present perfect. Two types measures of PND are used: the PND-form measure counts how many verbs share the same phonological transformation from base to target tense (e.g., <em>h<strong>e</strong>lpen–h<strong>ie</strong>lp, w<strong>e</strong>rpen–w<strong>ie</strong>rp</em>), while the PND-class measure captures broader morphophonological generalizations, grouping verbs that follow the same stem alternation patterns across multiple tenses, e.g., simple present: <em>l<strong>o</strong>pen</em>, simple past: <em>l<strong>ie</strong>p</em>, present perfect: <em>gel<strong>o</strong>pen</em>. </span></p> <p><span lang="EN-US">Building on this resource, we propose creating a phonological annotation layer for Dutch corpora within the CLARIN infrastructure (e.g., CGN, OpenSoNaR). Using standardized formats such as TEI or FoLiA, this layer will assign structured phonological categories to verb tokens and include relevant metadata. The resulting resource will support corpus-based research on language acquisition, phonology, morphology, and their interface with syntax and semantics, while also contributing to language technology applications and CLARIN’s mission to provide open, high-quality linguistic resources.</span></p>
format Recurso digital
id zenodo_https___doi_org_10_5281_zenodo_17339578
institution Zenodo
language nld
publishDate 2025
publisher Zenodo
record_format zenodo
spellingShingle Phonological Neighbourhood Density in Dutch Verbs: From Classification to Corpus Annotation
Zimianiti, Eleni
Kievit, Rogier
Rowland, Caroline
Donnelly, Seamus
Phonological Neighborhood Density
Corpus Annotation
Language Resources
<p><span lang="EN-US">Phonological Neighbourhood Density (PND) quantifies how many phonologically similar verbs undergo the same type of stem alternation across tenses. In Dutch, for instance, <em>blijken, lijken</em>, and <em>bijten</em> form a neighbourhood through the transformation to <em>bleek, leek</em>, and <em>beet</em> in the simple past tense. PND has been widely studied in psycholinguistics, where denser neighbourhoods are linked to stronger analogical effects in lexical processing, morphological generalization, and language acquisition.</span></p> <p><span lang="EN-US">A detailed PND classification of 1,452 Dutch verbs has been developed based on systematic vowel stem alternations between the simple present, simple past, and present perfect. Two types measures of PND are used: the PND-form measure counts how many verbs share the same phonological transformation from base to target tense (e.g., <em>h<strong>e</strong>lpen–h<strong>ie</strong>lp, w<strong>e</strong>rpen–w<strong>ie</strong>rp</em>), while the PND-class measure captures broader morphophonological generalizations, grouping verbs that follow the same stem alternation patterns across multiple tenses, e.g., simple present: <em>l<strong>o</strong>pen</em>, simple past: <em>l<strong>ie</strong>p</em>, present perfect: <em>gel<strong>o</strong>pen</em>. </span></p> <p><span lang="EN-US">Building on this resource, we propose creating a phonological annotation layer for Dutch corpora within the CLARIN infrastructure (e.g., CGN, OpenSoNaR). Using standardized formats such as TEI or FoLiA, this layer will assign structured phonological categories to verb tokens and include relevant metadata. The resulting resource will support corpus-based research on language acquisition, phonology, morphology, and their interface with syntax and semantics, while also contributing to language technology applications and CLARIN’s mission to provide open, high-quality linguistic resources.</span></p>
title Phonological Neighbourhood Density in Dutch Verbs: From Classification to Corpus Annotation
topic Phonological Neighborhood Density
Corpus Annotation
Language Resources
url https://doi.org/10.5281/zenodo.17339578