Improving Automatic Speech Recognition for Speakers Treated for Oral Cancer using Data Augmentation and LLM Error Correction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Folkertsma, Hidde, Tienkamp, Thomas, de Visscher, Sebastiaan, Witjes, Max, van Son, Rob, Guo, Jiapan, Halpern, Bence Mark |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Quantifying the effect of speech pathology on automatic and human speaker verification
par: Halpern, Bence Mark, et autres
Publié: (2024)
par: Halpern, Bence Mark, et autres
Publié: (2024)
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
par: Halpern, Bence Mark, et autres
Publié: (2025)
par: Halpern, Bence Mark, et autres
Publié: (2025)
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
par: Aronowitz, Hagai, et autres
Publié: (2026)
par: Aronowitz, Hagai, et autres
Publié: (2026)
Retrieval Augmented Correction of Named Entity Speech Recognition Errors
par: Pusateri, Ernest, et autres
Publié: (2024)
par: Pusateri, Ernest, et autres
Publié: (2024)
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems
par: Robatian, Amin, et autres
Publié: (2025)
par: Robatian, Amin, et autres
Publié: (2025)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
par: Guo, Jiaxin, et autres
Publié: (2024)
par: Guo, Jiaxin, et autres
Publié: (2024)
Diarization-Aware Multi-Speaker Automatic Speech Recognition via Large Language Models
par: Lin, Yuke, et autres
Publié: (2025)
par: Lin, Yuke, et autres
Publié: (2025)
Using Songs to Improve Kazakh Automatic Speech Recognition
par: Yeshpanov, Rustem
Publié: (2026)
par: Yeshpanov, Rustem
Publié: (2026)
The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge
par: Lin, Yuke, et autres
Publié: (2025)
par: Lin, Yuke, et autres
Publié: (2025)
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
par: Polok, Alexander, et autres
Publié: (2024)
par: Polok, Alexander, et autres
Publié: (2024)
A Comprehensive Investigation on Speaker Augmentation for Speaker Recognition
par: Zhou, Zhenyu, et autres
Publié: (2024)
par: Zhou, Zhenyu, et autres
Publié: (2024)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
par: Bondaruk, Łukasz, et autres
Publié: (2024)
par: Bondaruk, Łukasz, et autres
Publié: (2024)
Error Correction by Paying Attention to Both Acoustic and Confidence References for Automatic Speech Recognition
par: Shu, Yuchun, et autres
Publié: (2024)
par: Shu, Yuchun, et autres
Publié: (2024)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
par: Leung, Wing-Zin, et autres
Publié: (2024)
par: Leung, Wing-Zin, et autres
Publié: (2024)
Voice Conversion Augmentation for Speaker Recognition on Defective Datasets
par: Tao, Ruijie, et autres
Publié: (2024)
par: Tao, Ruijie, et autres
Publié: (2024)
Improving Automatic Speech Recognition with Decoder-Centric Regularisation in Encoder-Decoder Models
par: Polok, Alexander, et autres
Publié: (2024)
par: Polok, Alexander, et autres
Publié: (2024)
Speaker-Aware Simulation Improves Conversational Speech Recognition
par: Gedeon, Máté, et autres
Publié: (2026)
par: Gedeon, Máté, et autres
Publié: (2026)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
par: Nespoli, Francesco, et autres
Publié: (2024)
par: Nespoli, Francesco, et autres
Publié: (2024)
AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition
par: Dai, Yuhang, et autres
Publié: (2025)
par: Dai, Yuhang, et autres
Publié: (2025)
Exploring Generative Error Correction for Dysarthric Speech Recognition
par: La Quatra, Moreno, et autres
Publié: (2025)
par: La Quatra, Moreno, et autres
Publié: (2025)
Serialized Speech Information Guidance with Overlapped Encoding Separation for Multi-Speaker Automatic Speech Recognition
par: Shi, Hao, et autres
Publié: (2024)
par: Shi, Hao, et autres
Publié: (2024)
Phonetic Richness for Improved Automatic Speaker Verification
par: Klein, Nicholas, et autres
Publié: (2024)
par: Klein, Nicholas, et autres
Publié: (2024)
Multi-stage Large Language Model Correction for Speech Recognition
par: Pu, Jie, et autres
Publié: (2023)
par: Pu, Jie, et autres
Publié: (2023)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
par: Bhattacharjee, Susmita, et autres
Publié: (2025)
par: Bhattacharjee, Susmita, et autres
Publié: (2025)
Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition
par: Wang, Shih-heng, et autres
Publié: (2024)
par: Wang, Shih-heng, et autres
Publié: (2024)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
par: Ghosh, Sreyan, et autres
Publié: (2024)
par: Ghosh, Sreyan, et autres
Publié: (2024)
Beyond Manual Transcripts: The Potential of Automated Speech Recognition Errors in Improving Alzheimer's Disease Detection
par: Liu, Yin-Long, et autres
Publié: (2025)
par: Liu, Yin-Long, et autres
Publié: (2025)
MOPSA: Mixture of Prompt-Experts Based Speaker Adaptation for Elderly Speech Recognition
par: Deng, Chengxi, et autres
Publié: (2025)
par: Deng, Chengxi, et autres
Publié: (2025)
Unsupervised Online Continual Learning for Automatic Speech Recognition
par: Eeckt, Steven Vander, et autres
Publié: (2024)
par: Eeckt, Steven Vander, et autres
Publié: (2024)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
par: Min, Do June, et autres
Publié: (2024)
par: Min, Do June, et autres
Publié: (2024)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
par: Tian, Jingguang, et autres
Publié: (2024)
par: Tian, Jingguang, et autres
Publié: (2024)
Disentangled Representation Learning for Environment-agnostic Speaker Recognition
par: Nam, KiHyun, et autres
Publié: (2024)
par: Nam, KiHyun, et autres
Publié: (2024)
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
par: Ravenscroft, William, et autres
Publié: (2024)
par: Ravenscroft, William, et autres
Publié: (2024)
Benchmarking Japanese Speech Recognition on ASR-LLM Setups with Multi-Pass Augmented Generative Error Correction
par: Ko, Yuka, et autres
Publié: (2024)
par: Ko, Yuka, et autres
Publié: (2024)
Automatic Speech Recognition System-Independent Word Error Rate Estimation
par: Park, Chanho, et autres
Publié: (2024)
par: Park, Chanho, et autres
Publié: (2024)
Hallucinations in Neural Automatic Speech Recognition: Identifying Errors and Hallucinatory Models
par: Frieske, Rita, et autres
Publié: (2024)
par: Frieske, Rita, et autres
Publié: (2024)
Joint vs Sequential Speaker-Role Detection and Automatic Speech Recognition for Air-traffic Control
par: Blatt, Alexander, et autres
Publié: (2024)
par: Blatt, Alexander, et autres
Publié: (2024)
Non-Intrusive Automatic Speech Recognition Refinement: A Survey
par: Peyghan, Mohammad Reza, et autres
Publié: (2025)
par: Peyghan, Mohammad Reza, et autres
Publié: (2025)
Study on Inter and Intra Speaker Variability in Speaker Recognition
par: Okhotnikov, Anton, et autres
Publié: (2024)
par: Okhotnikov, Anton, et autres
Publié: (2024)
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
par: Jacobs, Christiaan, et autres
Publié: (2025)
par: Jacobs, Christiaan, et autres
Publié: (2025)
Documents similaires
-
Quantifying the effect of speech pathology on automatic and human speaker verification
par: Halpern, Bence Mark, et autres
Publié: (2024) -
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
par: Halpern, Bence Mark, et autres
Publié: (2025) -
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
par: Aronowitz, Hagai, et autres
Publié: (2026) -
Retrieval Augmented Correction of Named Entity Speech Recognition Errors
par: Pusateri, Ernest, et autres
Publié: (2024) -
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems
par: Robatian, Amin, et autres
Publié: (2025)