Towards explainable reference-free speech intelligibility evaluation of people with pathological speech
Fuente:
arXiv
Salvato in:
| Autori principali: | Halpern, Bence Mark, Tienkamp, Thomas, Abur, Defne, Toda, Tomoki |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PathBench: Speech Intelligibility Benchmark for Automatic Pathological Speech Assessment
di: Halpern, Bence Mark, et al.
Pubblicazione: (2026)
di: Halpern, Bence Mark, et al.
Pubblicazione: (2026)
Reference-free automatic speech severity evaluation using acoustic unit language modelling
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
XPPG-PCA: Reference-free automatic speech severity evaluation with principal components
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025)
Quantifying the effect of speech pathology on automatic and human speaker verification
di: Halpern, Bence Mark, et al.
Pubblicazione: (2024)
di: Halpern, Bence Mark, et al.
Pubblicazione: (2024)
Multi-speaker Text-to-speech Training with Speaker Anonymized Data
di: Huang, Wen-Chin, et al.
Pubblicazione: (2024)
di: Huang, Wen-Chin, et al.
Pubblicazione: (2024)
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
di: Biberger, Thomas, et al.
Pubblicazione: (2021)
di: Biberger, Thomas, et al.
Pubblicazione: (2021)
Selfsupervised learning for pathological speech detection
di: Sheikh, Shakeel Ahmad
Pubblicazione: (2024)
di: Sheikh, Shakeel Ahmad
Pubblicazione: (2024)
asr_eval: Algorithms and tools for multi-reference and streaming speech recognition evaluation
di: Sedukhin, Oleg, et al.
Pubblicazione: (2026)
di: Sedukhin, Oleg, et al.
Pubblicazione: (2026)
Ensemble of classifiers for speech evaluation
di: Belokrylov, G., et al.
Pubblicazione: (2024)
di: Belokrylov, G., et al.
Pubblicazione: (2024)
Covertly improving intelligibility with data-driven adaptations of speech timing
di: Tuttösí, Paige, et al.
Pubblicazione: (2026)
di: Tuttösí, Paige, et al.
Pubblicazione: (2026)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
di: Tabatabaee, Saba, et al.
Pubblicazione: (2026)
di: Tabatabaee, Saba, et al.
Pubblicazione: (2026)
Towards the Synthesis of Non-speech Vocalizations
di: Hoq, Enjamamul, et al.
Pubblicazione: (2024)
di: Hoq, Enjamamul, et al.
Pubblicazione: (2024)
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2025)
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2025)
Learning to assess subjective impressions from speech
di: Kondo, Yuto, et al.
Pubblicazione: (2025)
di: Kondo, Yuto, et al.
Pubblicazione: (2025)
Investigation of perceptual music similarity focusing on each instrumental part
di: Hashizume, Yuka, et al.
Pubblicazione: (2025)
di: Hashizume, Yuka, et al.
Pubblicazione: (2025)
Improving child speech recognition with augmented child-like speech
di: Zhang, Yuanyuan, et al.
Pubblicazione: (2024)
di: Zhang, Yuanyuan, et al.
Pubblicazione: (2024)
Automated evaluation of children's speech fluency for low-resource languages
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
di: Irino, Toshio, et al.
Pubblicazione: (2025)
di: Irino, Toshio, et al.
Pubblicazione: (2025)
Selective Classifier-free Guidance for Zero-shot Text-to-speech
di: Zheng, John, et al.
Pubblicazione: (2025)
di: Zheng, John, et al.
Pubblicazione: (2025)
Target matching based generative model for speech enhancement
di: Wang, Taihui, et al.
Pubblicazione: (2025)
di: Wang, Taihui, et al.
Pubblicazione: (2025)
On the relationship between speech and hearing
di: Umesh, Srinivasan, et al.
Pubblicazione: (2024)
di: Umesh, Srinivasan, et al.
Pubblicazione: (2024)
Wearable intelligent throat enables natural speech in stroke patients with dysarthria
di: Tang, Chenyu, et al.
Pubblicazione: (2024)
di: Tang, Chenyu, et al.
Pubblicazione: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
di: Ducorroy, Alexandre, et al.
Pubblicazione: (2025)
di: Ducorroy, Alexandre, et al.
Pubblicazione: (2025)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
di: Fernández-Díaz, Miguel, et al.
Pubblicazione: (2024)
di: Fernández-Díaz, Miguel, et al.
Pubblicazione: (2024)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
di: Kim, Yunsik, et al.
Pubblicazione: (2025)
di: Kim, Yunsik, et al.
Pubblicazione: (2025)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
QHARMA-GAN: Quasi-Harmonic Neural Vocoder based on Autoregressive Moving Average Model
di: Chen, Shaowen, et al.
Pubblicazione: (2025)
di: Chen, Shaowen, et al.
Pubblicazione: (2025)
Investigating training objective for flow matching-based speech enhancement
di: Yang, Liusha, et al.
Pubblicazione: (2025)
di: Yang, Liusha, et al.
Pubblicazione: (2025)
Serial-OE: Anomalous sound detection based on serial method with outlier exposure capable of using small amounts of anomalous data for training
di: Kuroyanagi, Ibuki, et al.
Pubblicazione: (2025)
di: Kuroyanagi, Ibuki, et al.
Pubblicazione: (2025)
Eigenvoice Synthesis based on Model Editing for Speaker Generation
di: Murata, Masato, et al.
Pubblicazione: (2025)
di: Murata, Masato, et al.
Pubblicazione: (2025)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
di: Deng, Qingkun, et al.
Pubblicazione: (2024)
di: Deng, Qingkun, et al.
Pubblicazione: (2024)
Gen-SER: When the generative model meets speech emotion recognition
di: Wang, Taihui, et al.
Pubblicazione: (2026)
di: Wang, Taihui, et al.
Pubblicazione: (2026)
Translating speech with just images
di: Oneata, Dan, et al.
Pubblicazione: (2024)
di: Oneata, Dan, et al.
Pubblicazione: (2024)
Data-driven grapheme-to-phoneme representations for a lexicon-free text-to-speech
di: Garg, Abhinav, et al.
Pubblicazione: (2024)
di: Garg, Abhinav, et al.
Pubblicazione: (2024)
AS-70: A Mandarin stuttered speech dataset for automatic speech recognition and stuttering event detection
di: Gong, Rong, et al.
Pubblicazione: (2024)
di: Gong, Rong, et al.
Pubblicazione: (2024)
Introduction to speech recognition
di: Dauphin, Gabriel
Pubblicazione: (2024)
di: Dauphin, Gabriel
Pubblicazione: (2024)
Improved Architecture for High-resolution Piano Transcription to Efficiently Capture Acoustic Characteristics of Music Signals
di: Mi, Jinyi, et al.
Pubblicazione: (2024)
di: Mi, Jinyi, et al.
Pubblicazione: (2024)
Improvements of Discriminative Feature Space Training for Anomalous Sound Detection in Unlabeled Conditions
di: Fujimura, Takuya, et al.
Pubblicazione: (2024)
di: Fujimura, Takuya, et al.
Pubblicazione: (2024)
Multi-channel multi-speaker transformer for speech recognition
di: Yifan, Guo, et al.
Pubblicazione: (2026)
di: Yifan, Guo, et al.
Pubblicazione: (2026)
Documenti analoghi
-
PathBench: Speech Intelligibility Benchmark for Automatic Pathological Speech Assessment
di: Halpern, Bence Mark, et al.
Pubblicazione: (2026) -
Reference-free automatic speech severity evaluation using acoustic unit language modelling
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025) -
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025) -
XPPG-PCA: Reference-free automatic speech severity evaluation with principal components
di: Halpern, Bence Mark, et al.
Pubblicazione: (2025) -
Quantifying the effect of speech pathology on automatic and human speaker verification
di: Halpern, Bence Mark, et al.
Pubblicazione: (2024)