What is lost in Normalization? Exploring Pitfalls in Multilingual ASR Model Evaluations
Fuente:
arXiv
Saved in:
| Main Authors: | Manohar, Kavya, Pillai, Leena G, Sherly, Elizabeth |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Humor Mechanics: Advancing Humor Generation with Multistep Reasoning
by: Tikhonov, Alexey, et al.
Published: (2024)
by: Tikhonov, Alexey, et al.
Published: (2024)
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
by: Anam, Rizal Khoirul
Published: (2025)
by: Anam, Rizal Khoirul
Published: (2025)
A Collaborative Content Moderation Framework for Toxicity Detection based on Conformalized Estimates of Annotation Disagreement
by: Villate-Castillo, Guillermo, et al.
Published: (2024)
by: Villate-Castillo, Guillermo, et al.
Published: (2024)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
by: Otal, Hakan T., et al.
Published: (2024)
by: Otal, Hakan T., et al.
Published: (2024)
Exploring the Structure of AI-Induced Language Change in Scientific English
by: Galpin, Riley, et al.
Published: (2025)
by: Galpin, Riley, et al.
Published: (2025)
Reviewriter: AI-Generated Instructions For Peer Review Writing
by: Su, Xiaotian, et al.
Published: (2025)
by: Su, Xiaotian, et al.
Published: (2025)
Rethinking the Multilingual Reasoning Gap with Layer Swap
by: Lasbordes, Maxence, et al.
Published: (2026)
by: Lasbordes, Maxence, et al.
Published: (2026)
Low-Resource Neural Machine Translation Using Recurrent Neural Networks and Transfer Learning: A Case Study on English-to-Igbo
by: Ekle, Ocheme Anthony, et al.
Published: (2025)
by: Ekle, Ocheme Anthony, et al.
Published: (2025)
Designing Synthetic Discussion Generation Systems: A Case Study for Online Facilitation
by: Tsirmpas, Dimitris, et al.
Published: (2025)
by: Tsirmpas, Dimitris, et al.
Published: (2025)
Quantitative Assessment of Intersectional Empathetic Bias and Understanding
by: Formanek, Vojtech, et al.
Published: (2024)
by: Formanek, Vojtech, et al.
Published: (2024)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
by: Zhang, Xue
Published: (2025)
by: Zhang, Xue
Published: (2025)
Large Language Models are Inconsistent and Biased Evaluators
by: Stureborg, Rickard, et al.
Published: (2024)
by: Stureborg, Rickard, et al.
Published: (2024)
CognitiveArm: Enabling Real-Time EEG-Controlled Prosthetic Arm Using Embodied Machine Learning
by: Basit, Abdul, et al.
Published: (2025)
by: Basit, Abdul, et al.
Published: (2025)
An Epidemiological Knowledge Graph extracted from the World Health Organization's Disease Outbreak News
by: Consoli, Sergio, et al.
Published: (2025)
by: Consoli, Sergio, et al.
Published: (2025)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
by: Wang, Yihao, et al.
Published: (2026)
by: Wang, Yihao, et al.
Published: (2026)
Towards Conversational AI for Human-Machine Collaborative MLOps
by: Fatouros, George, et al.
Published: (2025)
by: Fatouros, George, et al.
Published: (2025)
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
by: Sun, Mingrui, et al.
Published: (2026)
by: Sun, Mingrui, et al.
Published: (2026)
Value Lens: Using Large Language Models to Understand Human Values
by: Fernández, Eduardo de la Cruz, et al.
Published: (2025)
by: Fernández, Eduardo de la Cruz, et al.
Published: (2025)
Pitfalls in Evaluating Interpretability Agents
by: Haklay, Tal, et al.
Published: (2026)
by: Haklay, Tal, et al.
Published: (2026)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
Extracting Abstraction Dimensions by Identifying Syntax Pattern from Texts
by: Zhou, Jian, et al.
Published: (2025)
by: Zhou, Jian, et al.
Published: (2025)
Unveiling factors influencing judgment variation in Sentiment Analysis with Natural Language Processing and Statistics
by: Kellert, Olga, et al.
Published: (2024)
by: Kellert, Olga, et al.
Published: (2024)
Exploring State Tracking Capabilities of Large Language Models
by: Rezaee, Kiamehr, et al.
Published: (2025)
by: Rezaee, Kiamehr, et al.
Published: (2025)
Exploring Italian sentence embeddings properties through multi-tasking
by: Nastase, Vivi, et al.
Published: (2024)
by: Nastase, Vivi, et al.
Published: (2024)
Exploring syntactic information in sentence embeddings through multilingual subject-verb agreement
by: Nastase, Vivi, et al.
Published: (2024)
by: Nastase, Vivi, et al.
Published: (2024)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
by: Sun, Jingyi, et al.
Published: (2024)
by: Sun, Jingyi, et al.
Published: (2024)
Named entity recognition for Serbian legal documents: Design, methodology and dataset development
by: Kalušev, Vladimir, et al.
Published: (2025)
by: Kalušev, Vladimir, et al.
Published: (2025)
Optimizing What We Trust: Reliability-Guided QUBO Selection of Multi-Agent Weak Framing Signals for Arabic Sentiment Prediction
by: Alkhalifa, Rabab
Published: (2026)
by: Alkhalifa, Rabab
Published: (2026)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
by: Tu, Songjun, et al.
Published: (2026)
by: Tu, Songjun, et al.
Published: (2026)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
by: Moradbeiki, Pardis, et al.
Published: (2024)
by: Moradbeiki, Pardis, et al.
Published: (2024)
Evaluating Pixel Language Models on Non-Standardized Languages
by: Muñoz-Ortiz, Alberto, et al.
Published: (2024)
by: Muñoz-Ortiz, Alberto, et al.
Published: (2024)
The Pursuit of Empathy: Evaluating Small Language Models for PTSD Dialogue Support
by: BN, Suhas, et al.
Published: (2025)
by: BN, Suhas, et al.
Published: (2025)
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis
by: Petrov, Nikolay B, et al.
Published: (2024)
by: Petrov, Nikolay B, et al.
Published: (2024)
Tailoring Vaccine Messaging with Common-Ground Opinions
by: Stureborg, Rickard, et al.
Published: (2024)
by: Stureborg, Rickard, et al.
Published: (2024)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
by: Gómez-Rodríguez, Carlos, et al.
Published: (2024)
by: Gómez-Rodríguez, Carlos, et al.
Published: (2024)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
by: Bayram, M. Ali, et al.
Published: (2024)
by: Bayram, M. Ali, et al.
Published: (2024)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
by: Pan, Leyi, et al.
Published: (2025)
by: Pan, Leyi, et al.
Published: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
by: Weigang, Li, et al.
Published: (2025)
by: Weigang, Li, et al.
Published: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
by: Yao, Ben, et al.
Published: (2025)
by: Yao, Ben, et al.
Published: (2025)
Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data
by: Borisov, Vadim
Published: (2026)
by: Borisov, Vadim
Published: (2026)
Similar Items
-
Humor Mechanics: Advancing Humor Generation with Multistep Reasoning
by: Tikhonov, Alexey, et al.
Published: (2024) -
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
by: Anam, Rizal Khoirul
Published: (2025) -
A Collaborative Content Moderation Framework for Toxicity Detection based on Conformalized Estimates of Annotation Disagreement
by: Villate-Castillo, Guillermo, et al.
Published: (2024) -
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
by: Otal, Hakan T., et al.
Published: (2024) -
Exploring the Structure of AI-Induced Language Change in Scientific English
by: Galpin, Riley, et al.
Published: (2025)