AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Owodunni, Abraham Toluwase, Yadavalli, Aditya, Emezue, Chris Chinenye, Olatunji, Tobi, Mbataku, Clinton C |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Performant ASR Models for Medical Entities in Accented Speech
von: Afonja, Tejumade, et al.
Veröffentlicht: (2024)
von: Afonja, Tejumade, et al.
Veröffentlicht: (2024)
Effects of Speaker Count, Duration, and Accent Diversity on Zero-Shot Accent Robustness in Low-Resource ASR
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
AccentBox: Towards High-Fidelity Zero-Shot Accent Generation
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2024)
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2024)
DITTO: Data-efficient and Fair Targeted Subset Selection for ASR Accent Adaptation
von: Kothawade, Suraj, et al.
Veröffentlicht: (2021)
von: Kothawade, Suraj, et al.
Veröffentlicht: (2021)
MacST: Multi-Accent Speech Synthesis via Text Transliteration for Accent Conversion
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
Multimodal Consistency-Guided Reference-Free Data Selection for ASR Accent Adaptation
von: Lei, Ligong, et al.
Veröffentlicht: (2026)
von: Lei, Ligong, et al.
Veröffentlicht: (2026)
LID Models are Actually Accent Classifiers: Implications and Solutions for LID on Accented Speech
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
CodecMOS-Accent: A MOS Benchmark of Resynthesized and TTS Speech from Neural Codecs Across English Accents
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2026)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2026)
CosyAccent: Duration-Controllable Accent Normalization Using Source-Synthesis Training Data
von: Bai, Qibing, et al.
Veröffentlicht: (2026)
von: Bai, Qibing, et al.
Veröffentlicht: (2026)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
SpeechAccentLLM: A Unified Framework for Foreign Accent Conversion and Text to Speech
von: Cheng, Zhuangfei, et al.
Veröffentlicht: (2025)
von: Cheng, Zhuangfei, et al.
Veröffentlicht: (2025)
Accent-VITS:accent transfer for end-to-end TTS
von: Ma, Linhan, et al.
Veröffentlicht: (2023)
von: Ma, Linhan, et al.
Veröffentlicht: (2023)
Advancing African-Accented Speech Recognition: Epistemic Uncertainty-Driven Data Selection for Generalizable ASR Models
von: Dossou, Bonaventure F. P.
Veröffentlicht: (2023)
von: Dossou, Bonaventure F. P.
Veröffentlicht: (2023)
Convert and Speak: Zero-shot Accent Conversion with Minimum Supervision
von: Jia, Zhijun, et al.
Veröffentlicht: (2024)
von: Jia, Zhijun, et al.
Veröffentlicht: (2024)
Study on the Fairness of Speaker Verification Systems on Underrepresented Accents in English
von: Estevez, Mariel, et al.
Veröffentlicht: (2022)
von: Estevez, Mariel, et al.
Veröffentlicht: (2022)
Pairwise Evaluation of Accent Similarity in Speech Synthesis
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2025)
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2025)
Controllable Accent Normalization via Discrete Diffusion
von: Bai, Qibing, et al.
Veröffentlicht: (2026)
von: Bai, Qibing, et al.
Veröffentlicht: (2026)
Optimizing Multilingual Text-To-Speech with Accents & Emotions
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
von: Poncelet, Jakob, et al.
Veröffentlicht: (2023)
von: Poncelet, Jakob, et al.
Veröffentlicht: (2023)
Accent Normalization Using Self-Supervised Discrete Tokens with Non-Parallel Data
von: Bai, Qibing, et al.
Veröffentlicht: (2025)
von: Bai, Qibing, et al.
Veröffentlicht: (2025)
Rethinking Discrete Speech Representation Tokens for Accent Generation
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2026)
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2026)
GLOBE: A High-quality English Corpus with Global Accents for Zero-shot Speaker Adaptive Text-to-Speech
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data
von: Onda, Kentaro, et al.
Veröffentlicht: (2025)
von: Onda, Kentaro, et al.
Veröffentlicht: (2025)
Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder
von: Melechovsky, Jan, et al.
Veröffentlicht: (2022)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2022)
Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2024)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
von: Sasu, David, et al.
Veröffentlicht: (2025)
von: Sasu, David, et al.
Veröffentlicht: (2025)
Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
On the Relationship between Accent Strength and Articulatory Features
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
Empowering Communication: Speech Technology for Indian and Western Accents through AI-powered Speech Synthesis
von: R, Vinotha, et al.
Veröffentlicht: (2024)
von: R, Vinotha, et al.
Veröffentlicht: (2024)
Adapting Automatic Speech Recognition for Accented Air Traffic Control Communications
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
Non-autoregressive real-time Accent Conversion model with voice cloning
von: Nechaev, Vladimir, et al.
Veröffentlicht: (2024)
von: Nechaev, Vladimir, et al.
Veröffentlicht: (2024)
Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition
von: Xie, Xurong, et al.
Veröffentlicht: (2022)
von: Xie, Xurong, et al.
Veröffentlicht: (2022)
Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora
von: Onda, Kentaro, et al.
Veröffentlicht: (2025)
von: Onda, Kentaro, et al.
Veröffentlicht: (2025)
MMGER: Multi-modal and Multi-granularity Generative Error Correction with LLM for Joint Accent and Speech Recognition
von: Mu, Bingshen, et al.
Veröffentlicht: (2024)
von: Mu, Bingshen, et al.
Veröffentlicht: (2024)
Accent-Invariant Automatic Speech Recognition via Saliency-Driven Spectrogram Masking
von: Sameti, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Sameti, Mohammad Hossein, et al.
Veröffentlicht: (2025)
Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
ACES: Accent Subspaces for Coupling, Explanations, and Stress-Testing in Automatic Speech Recognition
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Performant ASR Models for Medical Entities in Accented Speech
von: Afonja, Tejumade, et al.
Veröffentlicht: (2024) -
Effects of Speaker Count, Duration, and Accent Diversity on Zero-Shot Accent Robustness in Low-Resource ASR
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025) -
AccentBox: Towards High-Fidelity Zero-Shot Accent Generation
von: Zhong, Jinzuomu, et al.
Veröffentlicht: (2024) -
DITTO: Data-efficient and Fair Targeted Subset Selection for ASR Accent Adaptation
von: Kothawade, Suraj, et al.
Veröffentlicht: (2021) -
MacST: Multi-Accent Speech Synthesis via Text Transliteration for Accent Conversion
von: Inoue, Sho, et al.
Veröffentlicht: (2024)