HATS: An Open data set Integrating Human Perception Applied to the Evaluation of Automatic Speech Recognition Metrics
Fuente:
arXiv
Guardado en:
| Autores principales: | Roux, Thibault Bañeras, Wottawa, Jane, Rouvier, Mickael, Merlin, Teva, Dufour, Richard |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
A Benchmark of French ASR Systems Based on Error Severity
por: Tholly, Antoine, et al.
Publicado: (2025)
por: Tholly, Antoine, et al.
Publicado: (2025)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
por: Labrak, Yanis, et al.
Publicado: (2023)
por: Labrak, Yanis, et al.
Publicado: (2023)
Probing the Information Encoded in Neural-based Acoustic Models of Automatic Speech Recognition Systems
por: Raymondaud, Quentin, et al.
Publicado: (2024)
por: Raymondaud, Quentin, et al.
Publicado: (2024)
An Empirical Analysis of Discrete Unit Representations in Speech Language Modeling Pre-training
por: Labrak, Yanis, et al.
Publicado: (2025)
por: Labrak, Yanis, et al.
Publicado: (2025)
Zero-Shot End-To-End Spoken Question Answering In Medical Domain
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
How Important Is Tokenization in French Medical Masked Language Models?
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
Asymmetric and trial-dependent modeling: the contribution of LIA to SdSV Challenge Task 2
por: Bousquet, Pierre-Michel, et al.
Publicado: (2024)
por: Bousquet, Pierre-Michel, et al.
Publicado: (2024)
SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
por: Du, Jiayu, et al.
Publicado: (2024)
por: Du, Jiayu, et al.
Publicado: (2024)
MSP-Podcast SER Challenge 2024: L'antenne du Ventoux Multimodal Self-Supervised Learning for Speech Emotion Recognition
por: Duret, Jarod, et al.
Publicado: (2024)
por: Duret, Jarod, et al.
Publicado: (2024)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
por: Sperber, Matthias, et al.
Publicado: (2024)
por: Sperber, Matthias, et al.
Publicado: (2024)
Identifying Reliable Evaluation Metrics for Scientific Text Revision
por: Jourdan, Léane, et al.
Publicado: (2025)
por: Jourdan, Léane, et al.
Publicado: (2025)
Speech-Aware Long Context Pruning and Integration for Contextualized Automatic Speech Recognition
por: Rong, Yiming, et al.
Publicado: (2025)
por: Rong, Yiming, et al.
Publicado: (2025)
Responsible Benchmarking of Fairness for Automatic Speech Recognition
por: Herron, Felix, et al.
Publicado: (2026)
por: Herron, Felix, et al.
Publicado: (2026)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
por: Luu, Nam, et al.
Publicado: (2025)
por: Luu, Nam, et al.
Publicado: (2025)
Closing the Speech-Text Gap with Limited Audio for Effective Domain Adaptation in LLM-Based ASR
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Developing an Automatic Pronunciation Scorer: Aligning Speech Evaluation Models and Applied Linguistics Constructs
por: Danwei Cai, et al.
Publicado: (2025)
por: Danwei Cai, et al.
Publicado: (2025)
Categorize Early, Integrate Late: Divergent Processing Strategies in Automatic Speech Recognition
por: Roll, Nathan, et al.
Publicado: (2026)
por: Roll, Nathan, et al.
Publicado: (2026)
Automatic Speech Recognition for the Ika Language
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
Late Fusion and Multi-Level Fission Amplify Cross-Modal Transfer in Text-Speech LMs
por: Cuervo, Santiago, et al.
Publicado: (2025)
por: Cuervo, Santiago, et al.
Publicado: (2025)
HATS: High-Accuracy Triple-Set Watermarking for Large Language Models
por: Hu, Zhiqing, et al.
Publicado: (2025)
por: Hu, Zhiqing, et al.
Publicado: (2025)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
por: Abu, Turi, et al.
Publicado: (2025)
por: Abu, Turi, et al.
Publicado: (2025)
DrBenchmark: A Large Language Understanding Evaluation Benchmark for French Biomedical Domain
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
por: Ryan, Michael J., et al.
Publicado: (2025)
por: Ryan, Michael J., et al.
Publicado: (2025)
The Role of Natural Language Processing Tasks in Automatic Literary Character Network Construction
por: Amalvy, Arthur, et al.
Publicado: (2024)
por: Amalvy, Arthur, et al.
Publicado: (2024)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
por: Park, ChaeHun, et al.
Publicado: (2024)
por: Park, ChaeHun, et al.
Publicado: (2024)
Open Automatic Speech Recognition Models for Classical and Modern Standard Arabic
por: Grigoryan, Lilit, et al.
Publicado: (2025)
por: Grigoryan, Lilit, et al.
Publicado: (2025)
Vietnamese Automatic Speech Recognition: A Revisit
por: Vu, Thi, et al.
Publicado: (2026)
por: Vu, Thi, et al.
Publicado: (2026)
The Role of Global and Local Context in Named Entity Recognition
por: Amalvy, Arthur, et al.
Publicado: (2023)
por: Amalvy, Arthur, et al.
Publicado: (2023)
OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary
por: Sudo, Yui, et al.
Publicado: (2025)
por: Sudo, Yui, et al.
Publicado: (2025)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
por: Adila, Aulia, et al.
Publicado: (2024)
por: Adila, Aulia, et al.
Publicado: (2024)
A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain
por: Obaidah, Qusai Abo, et al.
Publicado: (2024)
por: Obaidah, Qusai Abo, et al.
Publicado: (2024)
Learning to Rank Context for Named Entity Recognition Using a Synthetic Dataset
por: Amalvy, Arthur, et al.
Publicado: (2023)
por: Amalvy, Arthur, et al.
Publicado: (2023)
WST: Weakly Supervised Transducer for Automatic Speech Recognition
por: Gao, Dongji, et al.
Publicado: (2025)
por: Gao, Dongji, et al.
Publicado: (2025)
Stuttering-Aware Automatic Speech Recognition for Indonesian Language
por: Muhammad, Fadhil, et al.
Publicado: (2026)
por: Muhammad, Fadhil, et al.
Publicado: (2026)
Ejemplares similares
-
A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026) -
Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026) -
A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026) -
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026) -
A Benchmark of French ASR Systems Based on Error Severity
por: Tholly, Antoine, et al.
Publicado: (2025)