Exploring the anatomy of articulation rate in spontaneous English speech: relationships between utterance length effects and social factors
Fuente:
arXiv
Guardado en:
| Autores principales: | Tanner, James, Sonderegger, Morgan, Stuart-Smith, Jane, Kendall, Tyler, Mielke, Jeff, Dodsworth, Robin, Thomas, Erik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automatic classification of stop realisation with wav2vec2.0
por: Tanner, James, et al.
Publicado: (2025)
por: Tanner, James, et al.
Publicado: (2025)
On the relationship between speech and hearing
por: Umesh, Srinivasan, et al.
Publicado: (2024)
por: Umesh, Srinivasan, et al.
Publicado: (2024)
Improving speaker verification robustness with synthetic emotional utterances
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024)
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024)
Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech
por: Jiang, Pan-Pan, et al.
Publicado: (2024)
por: Jiang, Pan-Pan, et al.
Publicado: (2024)
Dementia classification from spontaneous speech using wrapper-based feature selection
por: Niemelä, Marko, et al.
Publicado: (2025)
por: Niemelä, Marko, et al.
Publicado: (2025)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
por: Tabatabaee, Saba, et al.
Publicado: (2026)
por: Tabatabaee, Saba, et al.
Publicado: (2026)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
Direct Punjabi to English speech translation using discrete units
por: Kaur, Prabhjot, et al.
Publicado: (2024)
por: Kaur, Prabhjot, et al.
Publicado: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
por: Deng, Qingkun, et al.
Publicado: (2024)
por: Deng, Qingkun, et al.
Publicado: (2024)
Cross-utterance ASR Rescoring with Graph-based Label Propagation
por: Tankasala, Srinath, et al.
Publicado: (2023)
por: Tankasala, Srinath, et al.
Publicado: (2023)
SponTTS: modeling and transferring spontaneous style for TTS
por: Li, Hanzhao, et al.
Publicado: (2023)
por: Li, Hanzhao, et al.
Publicado: (2023)
Probing mental health information in speech foundation models
por: de Gennes, Marc, et al.
Publicado: (2024)
por: de Gennes, Marc, et al.
Publicado: (2024)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
por: Li, Junjie, et al.
Publicado: (2024)
por: Li, Junjie, et al.
Publicado: (2024)
Distilling a speech and music encoder with task arithmetic
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2025)
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2025)
WhisperFlow: speech foundation models in real time
por: Wang, Rongxiang, et al.
Publicado: (2024)
por: Wang, Rongxiang, et al.
Publicado: (2024)
Omni-directional attention mechanism based on Mamba for speech separation
por: Xue, Ke, et al.
Publicado: (2026)
por: Xue, Ke, et al.
Publicado: (2026)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
por: Ohnaka, Hien, et al.
Publicado: (2024)
por: Ohnaka, Hien, et al.
Publicado: (2024)
BFA: Real-time Multilingual Text-to-speech Forced Alignment
por: Rehman, Abdul, et al.
Publicado: (2025)
por: Rehman, Abdul, et al.
Publicado: (2025)
SPGM: Prioritizing Local Features for enhanced speech separation performance
por: Yip, Jia Qi, et al.
Publicado: (2023)
por: Yip, Jia Qi, et al.
Publicado: (2023)
Inter-channel Conv-TasNet for multichannel speech enhancement
por: Lee, Dongheon, et al.
Publicado: (2021)
por: Lee, Dongheon, et al.
Publicado: (2021)
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
por: Yanir, Efrayim, et al.
Publicado: (2025)
por: Yanir, Efrayim, et al.
Publicado: (2025)
Graph-based multi-Feature fusion method for speech emotion recognition
por: Liu, Xueyu, et al.
Publicado: (2024)
por: Liu, Xueyu, et al.
Publicado: (2024)
Expressive paragraph text-to-speech synthesis with multi-step variational autoencoder
por: Li, Xuyuan, et al.
Publicado: (2023)
por: Li, Xuyuan, et al.
Publicado: (2023)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
por: Pascu, Octavian, et al.
Publicado: (2023)
por: Pascu, Octavian, et al.
Publicado: (2023)
A lightweight and robust method for blind wideband-to-fullband extension of speech
por: Büthe, Jan, et al.
Publicado: (2024)
por: Büthe, Jan, et al.
Publicado: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
por: Chen, Xingyu, et al.
Publicado: (2024)
por: Chen, Xingyu, et al.
Publicado: (2024)
FreeCodec: A disentangled neural speech codec with fewer tokens
por: Zheng, Youqiang, et al.
Publicado: (2024)
por: Zheng, Youqiang, et al.
Publicado: (2024)
Prosodic Parameter Manipulation in TTS generated speech for Controlled Speech Generation
por: Chary, Podakanti Satyajith
Publicado: (2024)
por: Chary, Podakanti Satyajith
Publicado: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
por: Chen, Shihao, et al.
Publicado: (2024)
por: Chen, Shihao, et al.
Publicado: (2024)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
por: Das, Sneha, et al.
Publicado: (2020)
por: Das, Sneha, et al.
Publicado: (2020)
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
por: Biberger, Thomas, et al.
Publicado: (2021)
por: Biberger, Thomas, et al.
Publicado: (2021)
Building speech corpus with diverse voice characteristics for its prompt-based representation
por: Watanabe, Aya, et al.
Publicado: (2024)
por: Watanabe, Aya, et al.
Publicado: (2024)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
por: Kammoun, Sofiene, et al.
Publicado: (2025)
por: Kammoun, Sofiene, et al.
Publicado: (2025)
DQR-TTS: Semi-supervised Text-to-speech Synthesis with Dynamic Quantized Representation
por: Wang, Jianzong, et al.
Publicado: (2023)
por: Wang, Jianzong, et al.
Publicado: (2023)
Language model integration based on memory control for sequence to sequence speech recognition
por: Cho, Jaejin, et al.
Publicado: (2018)
por: Cho, Jaejin, et al.
Publicado: (2018)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
por: Kumar, Anurag, et al.
Publicado: (2024)
por: Kumar, Anurag, et al.
Publicado: (2024)
Phoneme-based speech recognition driven by large language models and sampling marginalization
por: Ma, Te, et al.
Publicado: (2025)
por: Ma, Te, et al.
Publicado: (2025)
Single-channel speech enhancement by using psychoacoustical model inspired fusion framework
por: Samui, Suman
Publicado: (2022)
por: Samui, Suman
Publicado: (2022)
Multichannel blind speech source separation with a disjoint constraint source model
por: Wang, Jianyu, et al.
Publicado: (2024)
por: Wang, Jianyu, et al.
Publicado: (2024)
Non-verbal information in spontaneous speech -- towards a new framework of analysis
por: Biron, Tirza, et al.
Publicado: (2024)
por: Biron, Tirza, et al.
Publicado: (2024)
Ejemplares similares
-
Automatic classification of stop realisation with wav2vec2.0
por: Tanner, James, et al.
Publicado: (2025) -
On the relationship between speech and hearing
por: Umesh, Srinivasan, et al.
Publicado: (2024) -
Improving speaker verification robustness with synthetic emotional utterances
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024) -
Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech
por: Jiang, Pan-Pan, et al.
Publicado: (2024) -
Dementia classification from spontaneous speech using wrapper-based feature selection
por: Niemelä, Marko, et al.
Publicado: (2025)