Investigating Transcription Normalization in the Faetar ASR Benchmark
Fuente:
arXiv
Guardado en:
| Autores principales: | Peckham, Leo, Ong, Michael, Nagy, Naomi, Dunbar, Ewan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
por: Ong, Michael, et al.
Publicado: (2024)
por: Ong, Michael, et al.
Publicado: (2024)
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
por: Huo, Robin, et al.
Publicado: (2025)
por: Huo, Robin, et al.
Publicado: (2025)
Quantifying the Role of Textual Predictability in Automatic Speech Recognition
por: Robertson, Sean, et al.
Publicado: (2024)
por: Robertson, Sean, et al.
Publicado: (2024)
Bigger is not Always Better: The Effect of Context Size on Speech Pre-Training
por: Robertson, Sean, et al.
Publicado: (2023)
por: Robertson, Sean, et al.
Publicado: (2023)
Distinguishing Repetition Disfluency from Morphological Reduplication in Bangla ASR Transcripts: A Novel Corpus and Benchmarking Analysis
por: Arpa, Zaara Zabeen, et al.
Publicado: (2025)
por: Arpa, Zaara Zabeen, et al.
Publicado: (2025)
Beyond Transcription: Mechanistic Interpretability in ASR
por: Glazer, Neta, et al.
Publicado: (2025)
por: Glazer, Neta, et al.
Publicado: (2025)
Joint Beamforming and Speaker-Attributed ASR for Real Distant-Microphone Meeting Transcription
por: Cui, Can, et al.
Publicado: (2024)
por: Cui, Can, et al.
Publicado: (2024)
Recording for Eyes, Not Echoing to Ears: Contextualized Spoken-to-Written Conversion of ASR Transcripts
por: Liu, Jiaqing, et al.
Publicado: (2024)
por: Liu, Jiaqing, et al.
Publicado: (2024)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
por: Poli, Maxime, et al.
Publicado: (2026)
por: Poli, Maxime, et al.
Publicado: (2026)
Refining Transcripts With TV Subtitles by Prompt-Based Weakly Supervised Training of ASR
por: Zhao, Xinnian, et al.
Publicado: (2025)
por: Zhao, Xinnian, et al.
Publicado: (2025)
Zero Resource Code-switched Speech Benchmark Using Speech Utterance Pairs For Multiple Spoken Languages
por: Huang, Kuan-Po, et al.
Publicado: (2023)
por: Huang, Kuan-Po, et al.
Publicado: (2023)
Mind the Gap: Entity-Preserved Context-Aware ASR Structured Transcriptions
por: Altinok, Duygu
Publicado: (2025)
por: Altinok, Duygu
Publicado: (2025)
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
por: Pattnayak, Priyaranjan
Publicado: (2026)
por: Pattnayak, Priyaranjan
Publicado: (2026)
Investigating the Emergent Audio Classification Ability of ASR Foundation Models
por: Ma, Rao, et al.
Publicado: (2023)
por: Ma, Rao, et al.
Publicado: (2023)
The Sound of Healthcare: Improving Medical Transcription ASR Accuracy with Large Language Models
por: Adedeji, Ayo, et al.
Publicado: (2024)
por: Adedeji, Ayo, et al.
Publicado: (2024)
'The Order in the Horse's Heart': A Case Study in LLM-Assisted Stylometry for the Discovery of Biblical Allusion in Modern Literary Fiction
por: Cameron, Ewan
Publicado: (2026)
por: Cameron, Ewan
Publicado: (2026)
ASR-EC Benchmark: Evaluating Large Language Models on Chinese ASR Error Correction
por: Wei, Victor Junqiu, et al.
Publicado: (2024)
por: Wei, Victor Junqiu, et al.
Publicado: (2024)
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
por: Liang, Siyu, et al.
Publicado: (2025)
por: Liang, Siyu, et al.
Publicado: (2025)
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
por: Javed, Tahir, et al.
Publicado: (2024)
por: Javed, Tahir, et al.
Publicado: (2024)
Alzheimer Disease Classification through ASR-based Transcriptions: Exploring the Impact of Punctuation and Pauses
por: Gómez-Zaragozá, Lucía, et al.
Publicado: (2023)
por: Gómez-Zaragozá, Lucía, et al.
Publicado: (2023)
AdaCS: Adaptive Normalization for Enhanced Code-Switching ASR
por: Chu, The Chuong, et al.
Publicado: (2025)
por: Chu, The Chuong, et al.
Publicado: (2025)
WER We Stand: Benchmarking Urdu ASR Models
por: Arif, Samee, et al.
Publicado: (2024)
por: Arif, Samee, et al.
Publicado: (2024)
A Benchmark of French ASR Systems Based on Error Severity
por: Tholly, Antoine, et al.
Publicado: (2025)
por: Tholly, Antoine, et al.
Publicado: (2025)
Speech Emotion Recognition with ASR Transcripts: A Comprehensive Study on Word Error Rate and Fusion Techniques
por: Li, Yuanchao, et al.
Publicado: (2024)
por: Li, Yuanchao, et al.
Publicado: (2024)
Dysarthria Normalization via Local Lie Group Transformations for Robust ASR
por: Osipov, Mikhail
Publicado: (2025)
por: Osipov, Mikhail
Publicado: (2025)
PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
por: Piskala, Deepak Babu
Publicado: (2025)
por: Piskala, Deepak Babu
Publicado: (2025)
From Text Segmentation to Smart Chaptering: A Novel Benchmark for Structuring Video Transcriptions
por: Retkowski, Fabian, et al.
Publicado: (2024)
por: Retkowski, Fabian, et al.
Publicado: (2024)
(Not) Understanding Latin Poetic Style with Deep Learning
por: Nagy, Ben
Publicado: (2024)
por: Nagy, Ben
Publicado: (2024)
AppTek Call-Center Dialogues: A Multi-Accent Long-Form Benchmark for English ASR
por: Beck, Eugen, et al.
Publicado: (2026)
por: Beck, Eugen, et al.
Publicado: (2026)
Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation
por: Rahman, Hanif
Publicado: (2026)
por: Rahman, Hanif
Publicado: (2026)
Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German
por: Abdoli, Sajjad, et al.
Publicado: (2026)
por: Abdoli, Sajjad, et al.
Publicado: (2026)
Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry
por: Kumar, Sujeet, et al.
Publicado: (2025)
por: Kumar, Sujeet, et al.
Publicado: (2025)
ASR Benchmarking: Need for a More Representative Conversational Dataset
por: Maheshwari, Gaurav, et al.
Publicado: (2024)
por: Maheshwari, Gaurav, et al.
Publicado: (2024)
Benchmarks as Microscopes: A Call for Model Metrology
por: Saxon, Michael, et al.
Publicado: (2024)
por: Saxon, Michael, et al.
Publicado: (2024)
End-to-end Joint Punctuated and Normalized ASR with a Limited Amount of Punctuated Training Data
por: Cui, Can, et al.
Publicado: (2023)
por: Cui, Can, et al.
Publicado: (2023)
On the Robust Approximation of ASR Metrics
por: Waheed, Abdul, et al.
Publicado: (2025)
por: Waheed, Abdul, et al.
Publicado: (2025)
AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR
por: Ashungafac, Gabrial Zencha, et al.
Publicado: (2025)
por: Ashungafac, Gabrial Zencha, et al.
Publicado: (2025)
Kunnafonidilaw ka Cadeau: an ASR dataset of present-day Bambara
por: Diarra, Yacouba, et al.
Publicado: (2025)
por: Diarra, Yacouba, et al.
Publicado: (2025)
CantoASR: Prosody-Aware ASR-LALM Collaboration for Low-Resource Cantonese
por: Chen, Dazhong, et al.
Publicado: (2025)
por: Chen, Dazhong, et al.
Publicado: (2025)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
por: Rai, Anand, et al.
Publicado: (2025)
por: Rai, Anand, et al.
Publicado: (2025)
Ejemplares similares
-
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
por: Ong, Michael, et al.
Publicado: (2024) -
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
por: Huo, Robin, et al.
Publicado: (2025) -
Quantifying the Role of Textual Predictability in Automatic Speech Recognition
por: Robertson, Sean, et al.
Publicado: (2024) -
Bigger is not Always Better: The Effect of Context Size on Speech Pre-Training
por: Robertson, Sean, et al.
Publicado: (2023) -
Distinguishing Repetition Disfluency from Morphological Reduplication in Bangla ASR Transcripts: A Novel Corpus and Benchmarking Analysis
por: Arpa, Zaara Zabeen, et al.
Publicado: (2025)