Towards robust paralinguistic assessment for real-world mobile health (mHealth) monitoring: an initial study of reverberation effects on speech
Fuente:
arXiv
Guardado en:
| Autores principales: | Dineley, Judith, Carr, Ewan, Matcham, Faith, Downs, Johnny, Dobson, Richard, Quatieri, Thomas F, Cummins, Nicholas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A methodological framework and exemplar protocol for the collection and analysis of repeated speech samples
por: Cummins, Nicholas, et al.
Publicado: (2024)
por: Cummins, Nicholas, et al.
Publicado: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Cross-lingual Alzheimer's Disease detection based on paralinguistic and pre-trained features
por: Chen, Xuchu, et al.
Publicado: (2023)
por: Chen, Xuchu, et al.
Publicado: (2023)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
por: Tabatabaee, Saba, et al.
Publicado: (2026)
por: Tabatabaee, Saba, et al.
Publicado: (2026)
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
por: Dang, Shaoxiang, et al.
Publicado: (2024)
por: Dang, Shaoxiang, et al.
Publicado: (2024)
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech
por: Pahar, Madhurananda, et al.
Publicado: (2025)
por: Pahar, Madhurananda, et al.
Publicado: (2025)
A lightweight and robust method for blind wideband-to-fullband extension of speech
por: Büthe, Jan, et al.
Publicado: (2024)
por: Büthe, Jan, et al.
Publicado: (2024)
Probing mental health information in speech foundation models
por: de Gennes, Marc, et al.
Publicado: (2024)
por: de Gennes, Marc, et al.
Publicado: (2024)
Paraformer-v2: An improved non-autoregressive transformer for noise-robust speech recognition
por: An, Keyu, et al.
Publicado: (2024)
por: An, Keyu, et al.
Publicado: (2024)
Perceptual implications of simplifying geometrical acoustics models for Ambisonics-based binaural reverberation
por: Martin, Vincent, et al.
Publicado: (2024)
por: Martin, Vincent, et al.
Publicado: (2024)
Effects of auditory distance cues and reverberation on spatial perception and listening strategies
por: Missoni, Fulvio, et al.
Publicado: (2025)
por: Missoni, Fulvio, et al.
Publicado: (2025)
Online neural fusion of distortionless differential beamformers for robust speech enhancement
por: Qian, Yuanhang, et al.
Publicado: (2025)
por: Qian, Yuanhang, et al.
Publicado: (2025)
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
por: Fan, Junyi, et al.
Publicado: (2025)
por: Fan, Junyi, et al.
Publicado: (2025)
Evaluating pretrained speech embedding systems for dysarthria detection across heterogenous datasets
por: Wihlborg, Lovisa, et al.
Publicado: (2025)
por: Wihlborg, Lovisa, et al.
Publicado: (2025)
Noise-robust zero-shot text-to-speech synthesis conditioned on self-supervised speech-representation model with adapters
por: Fujita, Kenichi, et al.
Publicado: (2024)
por: Fujita, Kenichi, et al.
Publicado: (2024)
Systematic evaluation of commercially available pain‐management mHealth apps for chronic pain in the United Kingdom
por: Rebecca P. Harding, et al.
Publicado: (2026)
por: Rebecca P. Harding, et al.
Publicado: (2026)
Neighbors and relatives: How do speech embeddings reflect linguistic connections across the world?
por: Törö, Tuukka, et al.
Publicado: (2025)
por: Törö, Tuukka, et al.
Publicado: (2025)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
por: Huang, Ziling, et al.
Publicado: (2025)
por: Huang, Ziling, et al.
Publicado: (2025)
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
por: Saon, George, et al.
Publicado: (2025)
por: Saon, George, et al.
Publicado: (2025)
SLM-S2ST: A multimodal language model for direct speech-to-speech translation
por: Hu, Yuxuan, et al.
Publicado: (2025)
por: Hu, Yuxuan, et al.
Publicado: (2025)
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions
por: Mack, Wolfgang, et al.
Publicado: (2025)
por: Mack, Wolfgang, et al.
Publicado: (2025)
Speech-preserving active noise control: a deep learning approach in reverberant environments
por: Dai, Shuning
Publicado: (2026)
por: Dai, Shuning
Publicado: (2026)
Prominence-aware automatic speech recognition for conversational speech
por: Linke, Julian, et al.
Publicado: (2025)
por: Linke, Julian, et al.
Publicado: (2025)
Predicting speech intelligibility in older adults for speech enhancement using the Gammachirp Envelope Similarity Index, GESI
por: Yamamoto, Ayako, et al.
Publicado: (2025)
por: Yamamoto, Ayako, et al.
Publicado: (2025)
Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
por: Shirahata, Yuma, et al.
Publicado: (2024)
por: Shirahata, Yuma, et al.
Publicado: (2024)
Good practices for evaluation of synthesized speech
por: Cooper, Erica, et al.
Publicado: (2025)
por: Cooper, Erica, et al.
Publicado: (2025)
On the relationship between speech and hearing
por: Umesh, Srinivasan, et al.
Publicado: (2024)
por: Umesh, Srinivasan, et al.
Publicado: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
por: Gowda, Harshavardhana T., et al.
Publicado: (2025)
por: Gowda, Harshavardhana T., et al.
Publicado: (2025)
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
por: Lee, Dongheon, et al.
Publicado: (2026)
por: Lee, Dongheon, et al.
Publicado: (2026)
Improving child speech recognition with augmented child-like speech
por: Zhang, Yuanyuan, et al.
Publicado: (2024)
por: Zhang, Yuanyuan, et al.
Publicado: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
por: Deng, Qingkun, et al.
Publicado: (2024)
por: Deng, Qingkun, et al.
Publicado: (2024)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
por: Maisonneuve, Malo, et al.
Publicado: (2024)
por: Maisonneuve, Malo, et al.
Publicado: (2024)
Non-invasive electromyographic speech neuroprosthesis: a geometric perspective
por: Gowda, Harshavardhana T., et al.
Publicado: (2025)
por: Gowda, Harshavardhana T., et al.
Publicado: (2025)
FINALLY: fast and universal speech enhancement with studio-like quality
por: Babaev, Nicholas, et al.
Publicado: (2024)
por: Babaev, Nicholas, et al.
Publicado: (2024)
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
por: Huo, Robin, et al.
Publicado: (2025)
por: Huo, Robin, et al.
Publicado: (2025)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
por: Kim, Yunsik, et al.
Publicado: (2025)
por: Kim, Yunsik, et al.
Publicado: (2025)
Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation
por: Ma, Jianbo, et al.
Publicado: (2026)
por: Ma, Jianbo, et al.
Publicado: (2026)
Transcribe, Align and Segment: Creating speech datasets for low-resource languages
por: Sereda, Taras
Publicado: (2024)
por: Sereda, Taras
Publicado: (2024)
Ejemplares similares
-
A methodological framework and exemplar protocol for the collection and analysis of repeated speech samples
por: Cummins, Nicholas, et al.
Publicado: (2024) -
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024) -
Cross-lingual Alzheimer's Disease detection based on paralinguistic and pre-trained features
por: Chen, Xuchu, et al.
Publicado: (2023) -
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024) -
Towards noise-robust speech inversion through multi-task learning with speech enhancement
por: Tabatabaee, Saba, et al.
Publicado: (2026)