AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Awobade, Busayo, Ashungafac, Gabrial Zencha, Olatunji, Tobi |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR
par: Ashungafac, Gabrial Zencha, et autres
Publié: (2025)
par: Ashungafac, Gabrial Zencha, et autres
Publié: (2025)
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
par: Guda, Blessed, et autres
Publié: (2024)
par: Guda, Blessed, et autres
Publié: (2024)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
par: Zhang, Hezhao, et autres
Publié: (2026)
par: Zhang, Hezhao, et autres
Publié: (2026)
Quantifying and Mitigating Selection Bias in LLMs: A Transferable LoRA Fine-Tuning and Efficient Majority Voting Approach
par: Guda, Blessed, et autres
Publié: (2025)
par: Guda, Blessed, et autres
Publié: (2025)
AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents
par: Owodunni, Abraham Toluwase, et autres
Publié: (2024)
par: Owodunni, Abraham Toluwase, et autres
Publié: (2024)
Performant ASR Models for Medical Entities in Accented Speech
par: Afonja, Tejumade, et autres
Publié: (2024)
par: Afonja, Tejumade, et autres
Publié: (2024)
Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild
par: Durmus, Berkin, et autres
Publié: (2026)
par: Durmus, Berkin, et autres
Publié: (2026)
VoxRole: A Comprehensive Benchmark for Evaluating Speech-Based Role-Playing Agents
par: Wu, Weihao, et autres
Publié: (2025)
par: Wu, Weihao, et autres
Publié: (2025)
Benchmarking Automatic Speech Recognition Models for African Languages
par: Nahabwe, Alvin, et autres
Publié: (2025)
par: Nahabwe, Alvin, et autres
Publié: (2025)
AfriHuBERT: A self-supervised speech representation model for African languages
par: Alabi, Jesujoba O., et autres
Publié: (2024)
par: Alabi, Jesujoba O., et autres
Publié: (2024)
The Multicultural Medical Assistant: Can LLMs Improve Medical ASR Errors Across Borders?
par: Adedeji, Ayo, et autres
Publié: (2025)
par: Adedeji, Ayo, et autres
Publié: (2025)
WearVox: An Egocentric Multichannel Voice Assistant Benchmark for Wearables
par: Lin, Zhaojiang, et autres
Publié: (2025)
par: Lin, Zhaojiang, et autres
Publié: (2025)
What Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language Models
par: Awobade, Busayo, et autres
Publié: (2024)
par: Awobade, Busayo, et autres
Publié: (2024)
VoxVietnam: a Large-Scale Multi-Genre Dataset for Vietnamese Speaker Recognition
par: Vu, Hoang Long, et autres
Publié: (2024)
par: Vu, Hoang Long, et autres
Publié: (2024)
Retrieval-Augmented Speech Recognition Approach for Domain Challenges
par: Shen, Peng, et autres
Publié: (2025)
par: Shen, Peng, et autres
Publié: (2025)
BlasBench: An Open Benchmark for Irish Speech Recognition
par: Raj, Jyoutir, et autres
Publié: (2026)
par: Raj, Jyoutir, et autres
Publié: (2026)
VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models
par: Cui, Wenqian, et autres
Publié: (2025)
par: Cui, Wenqian, et autres
Publié: (2025)
Cocktail-Party Audio-Visual Speech Recognition
par: Nguyen, Thai-Binh, et autres
Publié: (2025)
par: Nguyen, Thai-Binh, et autres
Publié: (2025)
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions
par: Imam, Sukairaj Hafiz, et autres
Publié: (2025)
par: Imam, Sukairaj Hafiz, et autres
Publié: (2025)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
par: Zhu, Han, et autres
Publié: (2024)
par: Zhu, Han, et autres
Publié: (2024)
Speech Robust Bench: A Robustness Benchmark For Speech Recognition
par: Shah, Muhammad A., et autres
Publié: (2024)
par: Shah, Muhammad A., et autres
Publié: (2024)
BERSting at the Screams: A Benchmark for Distanced, Emotional and Shouted Speech Recognition
par: Tuttösí, Paige, et autres
Publié: (2025)
par: Tuttösí, Paige, et autres
Publié: (2025)
ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark
par: Wang, He, et autres
Publié: (2025)
par: Wang, He, et autres
Publié: (2025)
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
par: Mei, Yuxiang, et autres
Publié: (2026)
par: Mei, Yuxiang, et autres
Publié: (2026)
WildScore: Benchmarking MLLMs in-the-Wild Symbolic Music Reasoning
par: Mundada, Gagan, et autres
Publié: (2025)
par: Mundada, Gagan, et autres
Publié: (2025)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
par: Rai, Anand, et autres
Publié: (2025)
par: Rai, Anand, et autres
Publié: (2025)
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
par: Ong, Michael, et autres
Publié: (2024)
par: Ong, Michael, et autres
Publié: (2024)
Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark
par: Turetzky, Arnon, et autres
Publié: (2026)
par: Turetzky, Arnon, et autres
Publié: (2026)
PRiSM: Benchmarking Phone Realization in Speech Models
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
VoxHakka: A Dialectally Diverse Multi-speaker Text-to-Speech System for Taiwanese Hakka
par: Chen, Li-Wei, et autres
Publié: (2024)
par: Chen, Li-Wei, et autres
Publié: (2024)
Universal Robust Speech Adaptation for Cross-Domain Speech Recognition and Enhancement
par: Wang, Chien-Chun, et autres
Publié: (2026)
par: Wang, Chien-Chun, et autres
Publié: (2026)
Towards Orthographically-Informed Evaluation of Speech Recognition Systems for Indian Languages
par: Bhogale, Kaushal Santosh, et autres
Publié: (2026)
par: Bhogale, Kaushal Santosh, et autres
Publié: (2026)
Explainable Transformer-CNN Fusion for Noise-Robust Speech Emotion Recognition
par: Chakrabarty, Sudip, et autres
Publié: (2025)
par: Chakrabarty, Sudip, et autres
Publié: (2025)
Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India
par: Bhogale, Kaushal, et autres
Publié: (2026)
par: Bhogale, Kaushal, et autres
Publié: (2026)
Advancing African-Accented Speech Recognition: Epistemic Uncertainty-Driven Data Selection for Generalizable ASR Models
par: Dossou, Bonaventure F. P.
Publié: (2023)
par: Dossou, Bonaventure F. P.
Publié: (2023)
WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition
par: Ramezani, Erfan, et autres
Publié: (2026)
par: Ramezani, Erfan, et autres
Publié: (2026)
PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
par: Piskala, Deepak Babu
Publié: (2025)
par: Piskala, Deepak Babu
Publié: (2025)
PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech
par: Menta, Venkata Pushpak Teja
Publié: (2026)
par: Menta, Venkata Pushpak Teja
Publié: (2026)
Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs
par: Quang, Trung Nguyen, et autres
Publié: (2026)
par: Quang, Trung Nguyen, et autres
Publié: (2026)
Learning More with Less: Self-Supervised Approaches for Low-Resource Speech Emotion Recognition
par: Gong, Ziwei, et autres
Publié: (2025)
par: Gong, Ziwei, et autres
Publié: (2025)
Documents similaires
-
AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR
par: Ashungafac, Gabrial Zencha, et autres
Publié: (2025) -
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
par: Guda, Blessed, et autres
Publié: (2024) -
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
par: Zhang, Hezhao, et autres
Publié: (2026) -
Quantifying and Mitigating Selection Bias in LLMs: A Transferable LoRA Fine-Tuning and Efficient Majority Voting Approach
par: Guda, Blessed, et autres
Publié: (2025) -
AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents
par: Owodunni, Abraham Toluwase, et autres
Publié: (2024)