1000 African Voices: Advancing inclusive multi-speaker multi-accent speech synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Ogun, Sewade, Owodunni, Abraham T., Olatunji, Tobi, Alese, Eniola, Oladimeji, Babatunde, Afonja, Tejumade, Olaleye, Kayode, Etori, Naome A., Adewumi, Tosin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performant ASR Models for Medical Entities in Accented Speech
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Sometin Beta Pass Notin (SBPN): Improving Multilingual ASR for Nigerian Languages via Knowledge Distillation
by: Ogun, Sewade
Published: (2026)
by: Ogun, Sewade
Published: (2026)
End-to-end multi-channel speaker extraction and binaural speech synthesis
by: Chi, Cheng, et al.
Published: (2024)
by: Chi, Cheng, et al.
Published: (2024)
An Exhaustive Evaluation of TTS- and VC-based Data Augmentation for ASR
by: Ogun, Sewade, et al.
Published: (2025)
by: Ogun, Sewade, et al.
Published: (2025)
AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents
by: Owodunni, Abraham Toluwase, et al.
Published: (2024)
by: Owodunni, Abraham Toluwase, et al.
Published: (2024)
RideKE: Leveraging Low-Resource, User-Generated Twitter Content for Sentiment and Emotion Detection in Kenyan Code-Switched Dataset
by: Etori, Naome A., et al.
Published: (2025)
by: Etori, Naome A., et al.
Published: (2025)
The NaijaVoices Dataset: Cultivating Large-Scale, High-Quality, Culturally-Rich Speech Data for African Languages
by: Emezue, Chris, et al.
Published: (2025)
by: Emezue, Chris, et al.
Published: (2025)
Multi-channel multi-speaker transformer for speech recognition
by: Yifan, Guo, et al.
Published: (2026)
by: Yifan, Guo, et al.
Published: (2026)
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
by: Huang, Ziling, et al.
Published: (2025)
by: Huang, Ziling, et al.
Published: (2025)
Status of fisheries and aquaculture development in Ekiti State, Nigeria
by: Adewumi, A.A., et al.
Published: (2009)
by: Adewumi, A.A., et al.
Published: (2009)
Toward Accessible Mobile Money: A Voice-Driven, Biometrically Secured USSD Automation Framework for Visually Impaired Users
by: Ajayi, Sunday, et al.
Published: (2026)
by: Ajayi, Sunday, et al.
Published: (2026)
Prompting Towards Alleviating Code-Switched Data Scarcity in Under-Resourced Languages with GPT as a Pivot
by: Terblanche, Michelle, et al.
Published: (2024)
by: Terblanche, Michelle, et al.
Published: (2024)
Self-Improving Tabular Language Models via Iterative Reward-Guided Post-Training
by: Long, Yunbo, et al.
Published: (2026)
by: Long, Yunbo, et al.
Published: (2026)
LAG-MMLU: Benchmarking Frontier LLM Understanding in Latvian and Giriama
by: Etori, Naome A., et al.
Published: (2025)
by: Etori, Naome A., et al.
Published: (2025)
A Novel Preprocessing-Driven Approach to Remaining Useful Life (RUL) Prediction Using Temporal Convolutional Networks (TCN)
by: Imbert, Florent, et al.
Published: (2026)
by: Imbert, Florent, et al.
Published: (2026)
Target speaker anonymization in multi-speaker recordings
by: Tomashenko, Natalia, et al.
Published: (2025)
by: Tomashenko, Natalia, et al.
Published: (2025)
Towards Biologically Plausible and Private Gene Expression Data Generation
by: Chen, Dingfan, et al.
Published: (2024)
by: Chen, Dingfan, et al.
Published: (2024)
Top advances of the year: Reappraisal of systemic versus local therapy for hepatocellular carcinoma
by: Mosunmoluwa Oyenuga, et al.
Published: (2025)
by: Mosunmoluwa Oyenuga, et al.
Published: (2025)
Language translation, and change of accent for speech-to-speech task using diffusion model
by: Mishra, Abhishek, et al.
Published: (2025)
by: Mishra, Abhishek, et al.
Published: (2025)
Continually Adding New Languages to Multilingual Language Models
by: Owodunni, Abraham Toluwase, et al.
Published: (2025)
by: Owodunni, Abraham Toluwase, et al.
Published: (2025)
Essentials of the kinetic theory of multi-agent systems
by: Loy, Nadia, et al.
Published: (2025)
by: Loy, Nadia, et al.
Published: (2025)
TEXT2TASTE: A Versatile Egocentric Vision System for Intelligent Reading Assistance Using Large Language Model
by: Mucha, Wiktor, et al.
Published: (2024)
by: Mucha, Wiktor, et al.
Published: (2024)
Afrispeech-Dialog: A Benchmark Dataset for Spontaneous English Conversations in Healthcare and Beyond
by: Sanni, Mardhiyah, et al.
Published: (2025)
by: Sanni, Mardhiyah, et al.
Published: (2025)
The Epistemic Harms of Botched Apologies for Past Wrongs
by: Abraham Tobi
Published: (2025)
by: Abraham Tobi
Published: (2025)
Stigma, self‐styling and ‘forced accents’ among English L2 speakers in Spain
by: Eva Codó, et al.
Published: (2025)
by: Eva Codó, et al.
Published: (2025)
Expressive paragraph text-to-speech synthesis with multi-step variational autoencoder
by: Li, Xuyuan, et al.
Published: (2023)
by: Li, Xuyuan, et al.
Published: (2023)
New experiments on speaker diarization for unsupervised speaking style voice building for speech synthesis
by: Beatriz Martínez-González
Published: (2014)
by: Beatriz Martínez-González
Published: (2014)
Instruction Makes a Difference
by: Adewumi, Tosin, et al.
Published: (2024)
by: Adewumi, Tosin, et al.
Published: (2024)
Trends and Challenges in Authorship Analysis: A Review of ML, DL, and LLM Approaches
by: Habib, Nudrat, et al.
Published: (2025)
by: Habib, Nudrat, et al.
Published: (2025)
Generative AI and Teachers -- For Us or Against Us? A Case Study
by: Pettersson, Jenny, et al.
Published: (2024)
by: Pettersson, Jenny, et al.
Published: (2024)
On the Limitations of Large Language Models (LLMs): False Attribution
by: Adewumi, Tosin, et al.
Published: (2024)
by: Adewumi, Tosin, et al.
Published: (2024)
BatteryPass-12K: The First Dataset for the Novel Digital Battery Passport Conformance Task
by: Adewumi, Tosin, et al.
Published: (2026)
by: Adewumi, Tosin, et al.
Published: (2026)
AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition
by: Awobade, Busayo, et al.
Published: (2026)
by: Awobade, Busayo, et al.
Published: (2026)
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Acoustic and perceptual differences between standard and accented speech and their voice clones
by: Yang, Tianle, et al.
Published: (2026)
by: Yang, Tianle, et al.
Published: (2026)
Preschoolers benefit from sentential context in familiar‐ and unfamiliar‐accented speech
by: Naz Deniz Atik, et al.
Published: (2024)
by: Naz Deniz Atik, et al.
Published: (2024)
Efficient training strategies for natural sounding speech synthesis and speaker adaptation based on FastPitch
by: Răgman, Teodora, et al.
Published: (2024)
by: Răgman, Teodora, et al.
Published: (2024)
FLEXITOKENS: Flexible Tokenization for Evolving Language Models
by: Owodunni, Abraham Toluwase, et al.
Published: (2025)
by: Owodunni, Abraham Toluwase, et al.
Published: (2025)
SPGISpeech 2.0: Transcribed multi-speaker financial audio for speaker-tagged transcription
by: Grossman, Raymond, et al.
Published: (2025)
by: Grossman, Raymond, et al.
Published: (2025)
Similar Items
-
Performant ASR Models for Medical Entities in Accented Speech
by: Afonja, Tejumade, et al.
Published: (2024) -
Sometin Beta Pass Notin (SBPN): Improving Multilingual ASR for Nigerian Languages via Knowledge Distillation
by: Ogun, Sewade
Published: (2026) -
End-to-end multi-channel speaker extraction and binaural speech synthesis
by: Chi, Cheng, et al.
Published: (2024) -
An Exhaustive Evaluation of TTS- and VC-based Data Augmentation for ASR
by: Ogun, Sewade, et al.
Published: (2025) -
AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents
by: Owodunni, Abraham Toluwase, et al.
Published: (2024)