SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Du, Jiayu, Li, Jinpeng, Chen, Guoguo, Zhang, Wei-Qiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dolphin: A Large-Scale Automatic Speech Recognition Model for Eastern Languages
von: Meng, Yangyang, et al.
Veröffentlicht: (2025)
von: Meng, Yangyang, et al.
Veröffentlicht: (2025)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
von: Wang, Yujin, et al.
Veröffentlicht: (2022)
von: Wang, Yujin, et al.
Veröffentlicht: (2022)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
von: Abu, Turi, et al.
Veröffentlicht: (2025)
von: Abu, Turi, et al.
Veröffentlicht: (2025)
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text
von: Li, Jinpeng, et al.
Veröffentlicht: (2024)
von: Li, Jinpeng, et al.
Veröffentlicht: (2024)
Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual and Long-Form Speech Recognition Evaluation
von: Srivastav, Vaibhav, et al.
Veröffentlicht: (2025)
von: Srivastav, Vaibhav, et al.
Veröffentlicht: (2025)
The THUEE System Description for the IARPA OpenASR21 Challenge
von: Zhao, Jing, et al.
Veröffentlicht: (2022)
von: Zhao, Jing, et al.
Veröffentlicht: (2022)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
A Large Dataset of Spontaneous Speech with the Accent Spoken in São Paulo for Automatic Speech Recognition Evaluation
von: Lima, Rodrigo, et al.
Veröffentlicht: (2024)
von: Lima, Rodrigo, et al.
Veröffentlicht: (2024)
PhoWhisper: Automatic Speech Recognition for Vietnamese
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Hindi
von: Saha, Anish, et al.
Veröffentlicht: (2024)
von: Saha, Anish, et al.
Veröffentlicht: (2024)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary
von: Sudo, Yui, et al.
Veröffentlicht: (2025)
von: Sudo, Yui, et al.
Veröffentlicht: (2025)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition with BERT and CTC Transformers: A Review
von: Djeffal, Noussaiba, et al.
Veröffentlicht: (2024)
von: Djeffal, Noussaiba, et al.
Veröffentlicht: (2024)
Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2025)
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2025)
GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
Automatic Screening for Children with Speech Disorder using Automatic Speech Recognition: Opportunities and Challenges
von: Liu, Dancheng, et al.
Veröffentlicht: (2024)
von: Liu, Dancheng, et al.
Veröffentlicht: (2024)
Multi-Level Embedding Conformer Framework for Bengali Automatic Speech Recognition
von: Sakib, Md. Nazmus, et al.
Veröffentlicht: (2025)
von: Sakib, Md. Nazmus, et al.
Veröffentlicht: (2025)
CAMÕES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese
von: Carvalho, Carlos, et al.
Veröffentlicht: (2025)
von: Carvalho, Carlos, et al.
Veröffentlicht: (2025)
Stateful Conformer with Cache-based Inference for Streaming Automatic Speech Recognition
von: Noroozi, Vahid, et al.
Veröffentlicht: (2023)
von: Noroozi, Vahid, et al.
Veröffentlicht: (2023)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
von: Zhu, Han, et al.
Veröffentlicht: (2024)
von: Zhu, Han, et al.
Veröffentlicht: (2024)
How to Evaluate Automatic Speech Recognition: Comparing Different Performance and Bias Measures
von: Patel, Tanvina, et al.
Veröffentlicht: (2025)
von: Patel, Tanvina, et al.
Veröffentlicht: (2025)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
von: Min, Do June, et al.
Veröffentlicht: (2024)
von: Min, Do June, et al.
Veröffentlicht: (2024)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
PAC: Pronunciation-Aware Contextualized Large Language Model-based Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2025)
von: Fu, Li, et al.
Veröffentlicht: (2025)
Streaming Speech-to-Confusion Network Speech Recognition
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition for Biomedical Data in Bengali Language
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
Benchmarking Automatic Speech Recognition Models for African Languages
von: Nahabwe, Alvin, et al.
Veröffentlicht: (2025)
von: Nahabwe, Alvin, et al.
Veröffentlicht: (2025)
Exploring Gender Disparities in Automatic Speech Recognition Technology
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
Evaluation of LLMs in Speech is Often Flawed: Test Set Contamination in Large Language Models for Speech Recognition
von: Tseng, Yuan, et al.
Veröffentlicht: (2025)
von: Tseng, Yuan, et al.
Veröffentlicht: (2025)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
von: Hu, Ke, et al.
Veröffentlicht: (2025)
von: Hu, Ke, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition System-Independent Word Error Rate Estimation
von: Park, Chanho, et al.
Veröffentlicht: (2024)
von: Park, Chanho, et al.
Veröffentlicht: (2024)
CIF-T: A Novel CIF-based Transducer Architecture for Automatic Speech Recognition
von: Zhang, Tian-Hao, et al.
Veröffentlicht: (2023)
von: Zhang, Tian-Hao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Dolphin: A Large-Scale Automatic Speech Recognition Model for Eastern Languages
von: Meng, Yangyang, et al.
Veröffentlicht: (2025) -
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
von: Wang, Yujin, et al.
Veröffentlicht: (2022) -
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
von: Abu, Turi, et al.
Veröffentlicht: (2025) -
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text
von: Li, Jinpeng, et al.
Veröffentlicht: (2024) -
Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual and Long-Form Speech Recognition Evaluation
von: Srivastav, Vaibhav, et al.
Veröffentlicht: (2025)