Phonetic Segmentation of the UCLA Phonetics Lab Archive

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chodroff, Eleanor, Pažon, Blaž, Baker, Annie, Moran, Steven
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917624831541248
author Chodroff, Eleanor
Pažon, Blaž
Baker, Annie
Moran, Steven
author_facet Chodroff, Eleanor
Pažon, Blaž
Baker, Annie
Moran, Steven
contents Research in speech technologies and comparative linguistics depends on access to diverse and accessible speech data. The UCLA Phonetics Lab Archive is one of the earliest multilingual speech corpora, with long-form audio recordings and phonetic transcriptions for 314 languages (Ladefoged et al., 2009). Recently, 95 of these languages were time-aligned with word-level phonetic transcriptions (Li et al., 2021). Here we present VoxAngeles, a corpus of audited phonetic transcriptions and phone-level alignments of the UCLA Phonetics Lab Archive, which uses the 95-language CMU re-release as our starting point. VoxAngeles also includes word- and phone-level segmentations from the original UCLA corpus, as well as phonetic measurements of word and phone durations, vowel formants, and vowel f0. This corpus enhances the usability of the original data, particularly for quantitative phonetic typology, as demonstrated through a case study of vowel intrinsic f0. We also discuss the utility of the VoxAngeles corpus for general research and pedagogy in crosslinguistic phonetics, as well as for low-resource and multilingual speech technologies. VoxAngeles is free to download and use under a CC-BY-NC 4.0 license.
format Preprint
id arxiv_https___arxiv_org_abs_2403_19509
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Phonetic Segmentation of the UCLA Phonetics Lab Archive
Chodroff, Eleanor
Pažon, Blaž
Baker, Annie
Moran, Steven
Computation and Language
Sound
Audio and Speech Processing
Research in speech technologies and comparative linguistics depends on access to diverse and accessible speech data. The UCLA Phonetics Lab Archive is one of the earliest multilingual speech corpora, with long-form audio recordings and phonetic transcriptions for 314 languages (Ladefoged et al., 2009). Recently, 95 of these languages were time-aligned with word-level phonetic transcriptions (Li et al., 2021). Here we present VoxAngeles, a corpus of audited phonetic transcriptions and phone-level alignments of the UCLA Phonetics Lab Archive, which uses the 95-language CMU re-release as our starting point. VoxAngeles also includes word- and phone-level segmentations from the original UCLA corpus, as well as phonetic measurements of word and phone durations, vowel formants, and vowel f0. This corpus enhances the usability of the original data, particularly for quantitative phonetic typology, as demonstrated through a case study of vowel intrinsic f0. We also discuss the utility of the VoxAngeles corpus for general research and pedagogy in crosslinguistic phonetics, as well as for low-resource and multilingual speech technologies. VoxAngeles is free to download and use under a CC-BY-NC 4.0 license.
title Phonetic Segmentation of the UCLA Phonetics Lab Archive
topic Computation and Language
Sound
Audio and Speech Processing
url https://arxiv.org/abs/2403.19509