Cost Analysis of Human-corrected Transcription for Predominately Oral Languages
Fuente:
arXiv
Guardado en:
| Autores principales: | Diarra, Yacouba, Coulibaly, Nouhoum Souleymane, Leventhal, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Kunnafonidilaw ka Cadeau: an ASR dataset of present-day Bambara
por: Diarra, Yacouba, et al.
Publicado: (2025)
por: Diarra, Yacouba, et al.
Publicado: (2025)
The Serendipity of Claude AI: Case of the 13 Low-Resource National Languages of Mali
por: Dembele, Alou, et al.
Publicado: (2025)
por: Dembele, Alou, et al.
Publicado: (2025)
Dealing with the Hard Facts of Low-Resource African NLP
por: Diarra, Yacouba, et al.
Publicado: (2025)
por: Diarra, Yacouba, et al.
Publicado: (2025)
Listen, Attend, Understand: a Regularization Technique for Stable E2E Speech Translation Training on High Variance labels
por: Diarra, Yacouba, et al.
Publicado: (2026)
por: Diarra, Yacouba, et al.
Publicado: (2026)
Generative Artificial Intelligence, Musical Heritage and the Construction of Peace Narratives: A Case Study in Mali
por: Coulibaly, Nouhoum, et al.
Publicado: (2026)
por: Coulibaly, Nouhoum, et al.
Publicado: (2026)
Where Are We At with Automatic Speech Recognition for the Bambara Language?
por: Diallo, Seydou, et al.
Publicado: (2026)
por: Diallo, Seydou, et al.
Publicado: (2026)
Evaluating the trade effect of developing regional trade agreements : a semi-parametric approach / Souleymane Coulibaly
por: Coulibaly, Souleymane
Publicado: (2007)
por: Coulibaly, Souleymane
Publicado: (2007)
Position: Thematic Analysis of Unstructured Clinical Transcripts with Large Language Models
por: Yi, Seungjun, et al.
Publicado: (2025)
por: Yi, Seungjun, et al.
Publicado: (2025)
R2T: Rule-Encoded Loss Functions for Low-Resource Sequence Tagging
por: Keita, Mamadou K., et al.
Publicado: (2025)
por: Keita, Mamadou K., et al.
Publicado: (2025)
Antioxidant and antiamylase activities of leaf and root extracts of Ziziphus mauritiana Lam
por: Bandjougou Diarra, et al.
Publicado: (2024)
por: Bandjougou Diarra, et al.
Publicado: (2024)
Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis
por: Shakur, Ameer Hamza, et al.
Publicado: (2024)
por: Shakur, Ameer Hamza, et al.
Publicado: (2024)
TWeddit : A Dataset of Triggering Stories Predominantly Shared by Women on Reddit
por: Bandela, Shirlene Rose, et al.
Publicado: (2026)
por: Bandela, Shirlene Rose, et al.
Publicado: (2026)
Geometric Latent Reasoning Induces Shorter Generations in LLMs
por: Kumar, Shashi, et al.
Publicado: (2026)
por: Kumar, Shashi, et al.
Publicado: (2026)
Measuring the Effect of Transcription Noise on Downstream Language Understanding Tasks
por: Shapira, Ori, et al.
Publicado: (2025)
por: Shapira, Ori, et al.
Publicado: (2025)
Searching for Best Practices in Medical Transcription with Large Language Model
por: Li, Jiafeng, et al.
Publicado: (2024)
por: Li, Jiafeng, et al.
Publicado: (2024)
Investigating Transcription Normalization in the Faetar ASR Benchmark
por: Peckham, Leo, et al.
Publicado: (2025)
por: Peckham, Leo, et al.
Publicado: (2025)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
por: Fanconi, Claudio, et al.
Publicado: (2025)
por: Fanconi, Claudio, et al.
Publicado: (2025)
An Application of Large Language Models to Coding Negotiation Transcripts
por: Friedman, Ray, et al.
Publicado: (2024)
por: Friedman, Ray, et al.
Publicado: (2024)
TALENT: Table VQA via Augmented Language-Enhanced Natural-text Transcription
por: Yutong, Guo, et al.
Publicado: (2025)
por: Yutong, Guo, et al.
Publicado: (2025)
Self-correction is Not An Innate Capability in Language Models
por: Liu, Guangliang, et al.
Publicado: (2024)
por: Liu, Guangliang, et al.
Publicado: (2024)
Beyond Transcripts: Iterative Peer-Editing with Audio Unlocks High-Quality Human Summaries of Conversational Speech
por: Chaparala, Kaavya, et al.
Publicado: (2026)
por: Chaparala, Kaavya, et al.
Publicado: (2026)
IPA Transcription of Bengali Texts
por: Fatema, Kanij, et al.
Publicado: (2024)
por: Fatema, Kanij, et al.
Publicado: (2024)
Preserving Privacy, Increasing Accessibility, and Reducing Cost: An On-Device Artificial Intelligence Model for Medical Transcription and Note Generation
por: Thomas, Johnson, et al.
Publicado: (2025)
por: Thomas, Johnson, et al.
Publicado: (2025)
Fotheidil: an Automatic Transcription System for the Irish Language
por: Lonergan, Liam, et al.
Publicado: (2024)
por: Lonergan, Liam, et al.
Publicado: (2024)
Large Language Models for Oral History Understanding with Text Classification and Sentiment Analysis
por: Cherukuri, Komala Subramanyam, et al.
Publicado: (2025)
por: Cherukuri, Komala Subramanyam, et al.
Publicado: (2025)
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
por: Nguyen, Minh, et al.
Publicado: (2024)
por: Nguyen, Minh, et al.
Publicado: (2024)
Profiling Patient Transcript Using Large Language Model Reasoning Augmentation for Alzheimer's Disease Detection
por: Chen, Chin-Po, et al.
Publicado: (2024)
por: Chen, Chin-Po, et al.
Publicado: (2024)
Self-Debias: Self-correcting for Debiasing Large Language Models
por: Feng, Xuan, et al.
Publicado: (2026)
por: Feng, Xuan, et al.
Publicado: (2026)
Lyrics Transcription for Humans: A Readability-Aware Benchmark
por: Cífka, Ondřej, et al.
Publicado: (2024)
por: Cífka, Ondřej, et al.
Publicado: (2024)
Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization?
por: Sharma, Roshan, et al.
Publicado: (2024)
por: Sharma, Roshan, et al.
Publicado: (2024)
Leveraging Large Language Models for Predictive Analysis of Human Misery
por: Seal, Bishanka, et al.
Publicado: (2025)
por: Seal, Bishanka, et al.
Publicado: (2025)
Cost-effective Instruction Learning for Pathology Vision and Language Analysis
por: Chen, Kaitao, et al.
Publicado: (2024)
por: Chen, Kaitao, et al.
Publicado: (2024)
SBAAM! Eliminating Transcript Dependency in Automatic Subtitling
por: Gaido, Marco, et al.
Publicado: (2024)
por: Gaido, Marco, et al.
Publicado: (2024)
WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data
por: Zhang, Ziheng, et al.
Publicado: (2026)
por: Zhang, Ziheng, et al.
Publicado: (2026)
Small Language Model Can Self-correct
por: Han, Haixia, et al.
Publicado: (2024)
por: Han, Haixia, et al.
Publicado: (2024)
Oral to Web: Digitizing 'Zero Resource'Languages of Bangladesh
por: Rashid, Mohammad Mamun Or
Publicado: (2026)
por: Rashid, Mohammad Mamun Or
Publicado: (2026)
Distinguishing Repetition Disfluency from Morphological Reduplication in Bangla ASR Transcripts: A Novel Corpus and Benchmarking Analysis
por: Arpa, Zaara Zabeen, et al.
Publicado: (2025)
por: Arpa, Zaara Zabeen, et al.
Publicado: (2025)
Human Variability vs. Machine Consistency: A Linguistic Analysis of Texts Generated by Humans and Large Language Models
por: Zanotto, Sergio E., et al.
Publicado: (2024)
por: Zanotto, Sergio E., et al.
Publicado: (2024)
Developing a Guideline for the Labovian-Structural Analysis of Oral Narratives in Japanese
por: Watahiki, Amane, et al.
Publicado: (2026)
por: Watahiki, Amane, et al.
Publicado: (2026)
Extracting Biomedical Entities from Noisy Audio Transcripts
por: Ebadi, Nima, et al.
Publicado: (2024)
por: Ebadi, Nima, et al.
Publicado: (2024)
Ejemplares similares
-
Kunnafonidilaw ka Cadeau: an ASR dataset of present-day Bambara
por: Diarra, Yacouba, et al.
Publicado: (2025) -
The Serendipity of Claude AI: Case of the 13 Low-Resource National Languages of Mali
por: Dembele, Alou, et al.
Publicado: (2025) -
Dealing with the Hard Facts of Low-Resource African NLP
por: Diarra, Yacouba, et al.
Publicado: (2025) -
Listen, Attend, Understand: a Regularization Technique for Stable E2E Speech Translation Training on High Variance labels
por: Diarra, Yacouba, et al.
Publicado: (2026) -
Generative Artificial Intelligence, Musical Heritage and the Construction of Peace Narratives: A Case Study in Mali
por: Coulibaly, Nouhoum, et al.
Publicado: (2026)