Audio-Based Crowd-Sourced Evaluation of Machine Translation Quality
Fuente:
arXiv
Saved in:
| Main Authors: | Haq, Sami Ul, Castilho, Sheila, Graham, Yvette |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Aware Monolingual Human Evaluation of Machine Translation
by: Picinini, Silvio, et al.
Published: (2025)
by: Picinini, Silvio, et al.
Published: (2025)
Questionnaires for Everyone: Streamlining Cross-Cultural Questionnaire Adaptation with GPT-Based Translation Quality Evaluation
by: Haavisto, Otso, et al.
Published: (2024)
by: Haavisto, Otso, et al.
Published: (2024)
Introducing Quality Estimation to Machine Translation Post-editing Workflow: An Empirical Study on Its Usefulness
by: Liu, Siqi, et al.
Published: (2025)
by: Liu, Siqi, et al.
Published: (2025)
Machine Translation in the Wild: User Reaction to Xiaohongshu's Built-In Translation Feature
by: He, Sui
Published: (2026)
by: He, Sui
Published: (2026)
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)
by: Zouhar, Vilém, et al.
Published: (2026)
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows
by: Balashov, Yuri, et al.
Published: (2026)
by: Balashov, Yuri, et al.
Published: (2026)
Emojinize: Enriching Any Text with Emoji Translations
by: Klein, Lars Henning, et al.
Published: (2024)
by: Klein, Lars Henning, et al.
Published: (2024)
Human-LLM Hybrid Text Answer Aggregation for Crowd Annotations
by: Li, Jiyi
Published: (2024)
by: Li, Jiyi
Published: (2024)
AudioInsight: Detecting Social Contexts Relevant to Social Anxiety from Speech
by: Reddy, Varun, et al.
Published: (2024)
by: Reddy, Varun, et al.
Published: (2024)
Media of Langue: The Interface for Exploring Word Translation Network/Space
by: Muramoto, Goki, et al.
Published: (2023)
by: Muramoto, Goki, et al.
Published: (2023)
Lost Before Translation: Social Information Transmission and Survival in AI-AI Communication
by: Ghafouri, Bijean, et al.
Published: (2026)
by: Ghafouri, Bijean, et al.
Published: (2026)
Telephone Surveys Meet Conversational AI: Evaluating a LLM-Based Telephone Survey System at Scale
by: Lang, Max M., et al.
Published: (2025)
by: Lang, Max M., et al.
Published: (2025)
Prompting ChatGPT for Translation: A Comparative Analysis of Translation Brief and Persona Prompts
by: He, Sui
Published: (2024)
by: He, Sui
Published: (2024)
From Scratch to Fine-Tuned: A Comparative Study of Transformer Training Strategies for Legal Machine Translation
by: Barman, Amit, et al.
Published: (2025)
by: Barman, Amit, et al.
Published: (2025)
Efficient Machine Translation Corpus Generation: Integrating Human-in-the-Loop Post-Editing with Large Language Models
by: Yuksel, Kamer Ali, et al.
Published: (2025)
by: Yuksel, Kamer Ali, et al.
Published: (2025)
Open-Source Large Language Models as Multilingual Crowdworkers: Synthesizing Open-Domain Dialogues in Several Languages With No Examples in Targets and No Machine Translation
by: Njifenjou, Ahmed, et al.
Published: (2025)
by: Njifenjou, Ahmed, et al.
Published: (2025)
Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
Multi-Intent Recognition in Dialogue Understanding: A Comparison Between Smaller Open-Source LLMs
by: Ahmad, Adnan, et al.
Published: (2025)
by: Ahmad, Adnan, et al.
Published: (2025)
Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations
by: He, Keyu, et al.
Published: (2025)
by: He, Keyu, et al.
Published: (2025)
QE4PE: Word-level Quality Estimation for Human Post-Editing
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
A Comparative Study on Annotation Quality of Crowdsourcing and LLM via Label Aggregation
by: Li, Jiyi
Published: (2024)
by: Li, Jiyi
Published: (2024)
SwissADT: An Audio Description Translation System for Swiss Languages
by: Fischer, Lukas, et al.
Published: (2024)
by: Fischer, Lukas, et al.
Published: (2024)
Language Modelling Approaches to Adaptive Machine Translation
by: Moslem, Yasmin
Published: (2024)
by: Moslem, Yasmin
Published: (2024)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
by: Vo, Truong, et al.
Published: (2025)
by: Vo, Truong, et al.
Published: (2025)
Deep Learning and Machine Learning -- Natural Language Processing: From Theory to Application
by: Chen, Keyu, et al.
Published: (2024)
by: Chen, Keyu, et al.
Published: (2024)
Thematic Analysis with Open-Source Generative AI and Machine Learning: A New Method for Inductive Qualitative Codebook Development
by: Katz, Andrew, et al.
Published: (2024)
by: Katz, Andrew, et al.
Published: (2024)
AutoTAMP: Autoregressive Task and Motion Planning with LLMs as Translators and Checkers
by: Chen, Yongchao, et al.
Published: (2023)
by: Chen, Yongchao, et al.
Published: (2023)
Machine Learning for Enhancing Deliberation in Online Political Discussions and Participatory Processes: A Survey
by: Behrendt, Maike, et al.
Published: (2025)
by: Behrendt, Maike, et al.
Published: (2025)
Long-context Reference-based MT Quality Estimation
by: Haq, Sami Ul, et al.
Published: (2025)
by: Haq, Sami Ul, et al.
Published: (2025)
Combine Virtual Reality and Machine-Learning to Identify the Presence of Dyslexia: A Cross-Linguistic Approach
by: Materazzini, Michele, et al.
Published: (2025)
by: Materazzini, Michele, et al.
Published: (2025)
Robots in the Middle: Evaluating LLMs in Dispute Resolution
by: Tan, Jinzhe, et al.
Published: (2024)
by: Tan, Jinzhe, et al.
Published: (2024)
Conversations Gone Awry, But Then? Evaluating Conversational Forecasting Models
by: Tran, Son Quoc, et al.
Published: (2025)
by: Tran, Son Quoc, et al.
Published: (2025)
Designing and Evaluating Chain-of-Hints for Scientific Question Answering
by: Jangra, Anubhav, et al.
Published: (2025)
by: Jangra, Anubhav, et al.
Published: (2025)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
From Text to Self: Users' Perceptions of Potential of AI on Interpersonal Communication and Self
by: Fu, Yue, et al.
Published: (2023)
by: Fu, Yue, et al.
Published: (2023)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
DICE: A Framework for Dimensional and Contextual Evaluation of Language Models
by: Shrivastava, Aryan, et al.
Published: (2025)
by: Shrivastava, Aryan, et al.
Published: (2025)
Evaluation of a Sign Language Avatar on Comprehensibility, User Experience \& Acceptability
by: Wasserroth, Fenya, et al.
Published: (2025)
by: Wasserroth, Fenya, et al.
Published: (2025)
IP-Dialog: Evaluating Implicit Personalization in Dialogue Systems with Synthetic Data
by: Peng, Bo, et al.
Published: (2025)
by: Peng, Bo, et al.
Published: (2025)
Similar Items
-
Context-Aware Monolingual Human Evaluation of Machine Translation
by: Picinini, Silvio, et al.
Published: (2025) -
Questionnaires for Everyone: Streamlining Cross-Cultural Questionnaire Adaptation with GPT-Based Translation Quality Evaluation
by: Haavisto, Otso, et al.
Published: (2024) -
Introducing Quality Estimation to Machine Translation Post-editing Workflow: An Empirical Study on Its Usefulness
by: Liu, Siqi, et al.
Published: (2025) -
Machine Translation in the Wild: User Reaction to Xiaohongshu's Built-In Translation Feature
by: He, Sui
Published: (2026) -
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)