Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Chin-Jou, Yeo, Eunjung, Choi, Kwanghee, Pérez-Toro, Paula Andrea, Someki, Masao, Das, Rohan Kumar, Yue, Zhengjun, Orozco-Arroyave, Juan Rafael, Nöth, Elmar, Mortensen, David R. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Voice Biomarker Analysis and Automated Severity Classification of Dysarthric Speech in a Multilingual Context
por: Yeo, Eunjung
Publicado: (2024)
por: Yeo, Eunjung
Publicado: (2024)
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
por: Hernandez, Abner, et al.
Publicado: (2026)
por: Hernandez, Abner, et al.
Publicado: (2026)
Applications of Artificial Intelligence for Cross-language Intelligibility Assessment of Dysarthric Speech
por: Yeo, Eunjung, et al.
Publicado: (2025)
por: Yeo, Eunjung, et al.
Publicado: (2025)
An Empirical Recipe for Universal Phone Recognition
por: Bharadwaj, Shikhar, et al.
Publicado: (2026)
por: Bharadwaj, Shikhar, et al.
Publicado: (2026)
Multilingual Dysarthric Speech Assessment Using Universal Phone Recognition and Language-Specific Phonemic Contrast Modeling
por: Yeo, Eunjung, et al.
Publicado: (2026)
por: Yeo, Eunjung, et al.
Publicado: (2026)
POWSM: A Phonetic Open Whisper-Style Speech Foundation Model
por: Li, Chin-Jou, et al.
Publicado: (2025)
por: Li, Chin-Jou, et al.
Publicado: (2025)
On-device Streaming Discrete Speech Units
por: Choi, Kwanghee, et al.
Publicado: (2025)
por: Choi, Kwanghee, et al.
Publicado: (2025)
Leveraging Allophony in Self-Supervised Speech Models for Atypical Pronunciation Assessment
por: Choi, Kwanghee, et al.
Publicado: (2025)
por: Choi, Kwanghee, et al.
Publicado: (2025)
Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech
por: Hajal, Karl El, et al.
Publicado: (2025)
por: Hajal, Karl El, et al.
Publicado: (2025)
Unsupervised Rhythm and Voice Conversion of Dysarthric to Healthy Speech for ASR
por: Hajal, Karl El, et al.
Publicado: (2025)
por: Hajal, Karl El, et al.
Publicado: (2025)
Personalized Fine-Tuning with Controllable Synthetic Speech from LLM-Generated Transcripts for Dysarthric Speech Recognition
por: Wagner, Dominik, et al.
Publicado: (2025)
por: Wagner, Dominik, et al.
Publicado: (2025)
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
por: Choi, Kwanghee, et al.
Publicado: (2026)
por: Choi, Kwanghee, et al.
Publicado: (2026)
Self-Supervised Speech Models Encode Phonetic Context via Position-dependent Orthogonal Subspaces
por: Choi, Kwanghee, et al.
Publicado: (2026)
por: Choi, Kwanghee, et al.
Publicado: (2026)
Self-supervised ASR Models and Features For Dysarthric and Elderly Speech Recognition
por: Hu, Shujie, et al.
Publicado: (2024)
por: Hu, Shujie, et al.
Publicado: (2024)
Probing Whisper for Dysarthric Speech in Detection and Assessment
por: Yue, Zhengjun, et al.
Publicado: (2025)
por: Yue, Zhengjun, et al.
Publicado: (2025)
Finding My Voice: Generative Reconstruction of Disordered Speech for Automated Clinical Evaluation
por: Rosero, Karen, et al.
Publicado: (2025)
por: Rosero, Karen, et al.
Publicado: (2025)
Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches
por: Aboeitta, Ahmed, et al.
Publicado: (2025)
por: Aboeitta, Ahmed, et al.
Publicado: (2025)
Prosodic ABX: A Language-Agnostic Method for Measuring Prosodic Contrast in Speech Representations
por: Sun, Haitong, et al.
Publicado: (2026)
por: Sun, Haitong, et al.
Publicado: (2026)
Objective and Subjective Evaluation of Diffusion-Based Speech Enhancement for Dysarthric Speech
por: de Groot, Dimme, et al.
Publicado: (2025)
por: de Groot, Dimme, et al.
Publicado: (2025)
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
por: Zheng, Xiuwen, et al.
Publicado: (2026)
por: Zheng, Xiuwen, et al.
Publicado: (2026)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
por: Leung, Wing-Zin, et al.
Publicado: (2024)
por: Leung, Wing-Zin, et al.
Publicado: (2024)
PRiSM: Benchmarking Phone Realization in Speech Models
por: Bharadwaj, Shikhar, et al.
Publicado: (2026)
por: Bharadwaj, Shikhar, et al.
Publicado: (2026)
Enhancing Pre-trained ASR System Fine-tuning for Dysarthric Speech Recognition using Adversarial Data Augmentation
por: Wang, Huimeng, et al.
Publicado: (2024)
por: Wang, Huimeng, et al.
Publicado: (2024)
Exploring Generative Error Correction for Dysarthric Speech Recognition
por: La Quatra, Moreno, et al.
Publicado: (2025)
por: La Quatra, Moreno, et al.
Publicado: (2025)
Voice Cloning for Dysarthric Speech Synthesis: Addressing Data Scarcity in Speech-Language Pathology
por: Moell, Birger, et al.
Publicado: (2025)
por: Moell, Birger, et al.
Publicado: (2025)
Improved Dysarthric Speech to Text Conversion via TTS Personalization
por: Mihajlik, Péter, et al.
Publicado: (2025)
por: Mihajlik, Péter, et al.
Publicado: (2025)
Personalized Adversarial Data Augmentation for Dysarthric and Elderly Speech Recognition
por: Jin, Zengrui, et al.
Publicado: (2022)
por: Jin, Zengrui, et al.
Publicado: (2022)
Phone-purity Guided Discrete Tokens for Dysarthric Speech Recognition
por: Wang, Huimeng, et al.
Publicado: (2025)
por: Wang, Huimeng, et al.
Publicado: (2025)
Robust Cross-Etiology and Speaker-Independent Dysarthric Speech Recognition
por: Singh, Satwinder, et al.
Publicado: (2025)
por: Singh, Satwinder, et al.
Publicado: (2025)
Robust Dysarthric Speech Recognition with GAN Enhancement and LLM Correction
por: Yibo He, et al.
Publicado: (2025)
por: Yibo He, et al.
Publicado: (2025)
Inappropriate Pause Detection In Dysarthric Speech Using Large-Scale Speech Recognition
por: Lee, Jeehyun, et al.
Publicado: (2024)
por: Lee, Jeehyun, et al.
Publicado: (2024)
A Speech-to-Video Synthesis Approach Using Spatio-Temporal Diffusion for Vocal Tract MRI
por: Pérez-Toro, Paula Andrea, et al.
Publicado: (2025)
por: Pérez-Toro, Paula Andrea, et al.
Publicado: (2025)
Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching
por: Das, Shoutrik, et al.
Publicado: (2025)
por: Das, Shoutrik, et al.
Publicado: (2025)
Speech Recognition-based Feature Extraction for Enhanced Automatic Severity Classification in Dysarthric Speech
por: Choi, Yerin, et al.
Publicado: (2024)
por: Choi, Yerin, et al.
Publicado: (2024)
Recognizing Every Voice: Towards Inclusive ASR for Rural Bhojpuri Women
por: Joshi, Sakshi, et al.
Publicado: (2025)
por: Joshi, Sakshi, et al.
Publicado: (2025)
Regularized Federated Learning for Privacy-Preserving Dysarthric and Elderly Speech Recognition
por: Zhong, Tao, et al.
Publicado: (2025)
por: Zhong, Tao, et al.
Publicado: (2025)
Variational Auto-Encoder Based Variability Encoding for Dysarthric Speech Recognition
por: Xie, Xurong, et al.
Publicado: (2022)
por: Xie, Xurong, et al.
Publicado: (2022)
Large Language Models for Dysfluency Detection in Stuttered Speech
por: Wagner, Dominik, et al.
Publicado: (2024)
por: Wagner, Dominik, et al.
Publicado: (2024)
Segment-Level Vectorized Beam Search Based on Partially Autoregressive Inference
por: Someki, Masao, et al.
Publicado: (2023)
por: Someki, Masao, et al.
Publicado: (2023)
Zero-Shot Recognition of Dysarthric Speech Using Commercial Automatic Speech Recognition and Multimodal Large Language Models
por: Alsayegh, Ali, et al.
Publicado: (2025)
por: Alsayegh, Ali, et al.
Publicado: (2025)
Ejemplares similares
-
Voice Biomarker Analysis and Automated Severity Classification of Dysarthric Speech in a Multilingual Context
por: Yeo, Eunjung
Publicado: (2024) -
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
por: Hernandez, Abner, et al.
Publicado: (2026) -
Applications of Artificial Intelligence for Cross-language Intelligibility Assessment of Dysarthric Speech
por: Yeo, Eunjung, et al.
Publicado: (2025) -
An Empirical Recipe for Universal Phone Recognition
por: Bharadwaj, Shikhar, et al.
Publicado: (2026) -
Multilingual Dysarthric Speech Assessment Using Universal Phone Recognition and Language-Specific Phonemic Contrast Modeling
por: Yeo, Eunjung, et al.
Publicado: (2026)