Simpson's Paradox and the Accuracy-Fluency Tradeoff in Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Lim, Zheng Wei, Vylomova, Ekaterina, Cohn, Trevor, Kemp, Charles |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OpenWHO: A Document-Level Parallel Corpus for Health Translation in Low-Resource Languages
by: Merx, Raphaël, et al.
Published: (2025)
by: Merx, Raphaël, et al.
Published: (2025)
TULUN: Transparent and Adaptable Low-resource Machine Translation
by: Merx, Raphaël, et al.
Published: (2025)
by: Merx, Raphaël, et al.
Published: (2025)
Mufu: Multilingual Fused Learning for Low-Resource Translation with LLM
by: Lim, Zheng Wei, et al.
Published: (2024)
by: Lim, Zheng Wei, et al.
Published: (2024)
Language-Specific Latent Process Hinders Cross-Lingual Performance
by: Lim, Zheng Wei, et al.
Published: (2025)
by: Lim, Zheng Wei, et al.
Published: (2025)
A Multidimensional Framework for Evaluating Lexical Semantic Change with Social Science Applications
by: Baes, Naomi, et al.
Published: (2024)
by: Baes, Naomi, et al.
Published: (2024)
Vavanagi: a Community-run Platform for Documentation of the Hula Language in Papua New Guinea
by: Olewale, Bri, et al.
Published: (2026)
by: Olewale, Bri, et al.
Published: (2026)
A Computational Approach to Identifying Cultural Keywords Across Languages
by: Zheng Wei Lim, et al.
Published: (2024)
by: Zheng Wei Lim, et al.
Published: (2024)
Generating bilingual example sentences with large language models as lexicography assistants
by: Merx, Raphael, et al.
Published: (2024)
by: Merx, Raphael, et al.
Published: (2024)
Low-resource Machine Translation: what for? who for? An observational study on a dedicated Tetun language translation service
by: Merx, Raphael, et al.
Published: (2024)
by: Merx, Raphael, et al.
Published: (2024)
Low-Resource Machine Translation through Retrieval-Augmented LLM Prompting: A Study on the Mambai Language
by: Merx, Raphaël, et al.
Published: (2024)
by: Merx, Raphaël, et al.
Published: (2024)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
by: Shayegh, Behzad, et al.
Published: (2025)
by: Shayegh, Behzad, et al.
Published: (2025)
You Cannot Feed Two Birds with One Score: the Accuracy-Naturalness Tradeoff in Translation
by: Flamich, Gergely, et al.
Published: (2025)
by: Flamich, Gergely, et al.
Published: (2025)
Fluency and Faithfulness in Human and Machine Literary Translation
by: Griebel, Sarah, et al.
Published: (2026)
by: Griebel, Sarah, et al.
Published: (2026)
LSC-Eval: A General Framework to Evaluate Methods for Assessing Dimensions of Lexical Semantic Change Using LLM-Generated Synthetic Data
by: Baes, Naomi, et al.
Published: (2025)
by: Baes, Naomi, et al.
Published: (2025)
A Joint Multitask Model for Morpho-Syntactic Parsing
by: Inostroza, Demian, et al.
Published: (2025)
by: Inostroza, Demian, et al.
Published: (2025)
Paradigm Completion for Derivational Morphology
by: Cotterell, Ryan, et al.
Published: (2017)
by: Cotterell, Ryan, et al.
Published: (2017)
Backdoor Attack on Multilingual Machine Translation
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
by: Aji, Alham Fikri, et al.
Published: (2025)
by: Aji, Alham Fikri, et al.
Published: (2025)
Log-Likelihood, Simpson's Paradox, and the Detection of Machine-Generated Text
by: Kempton, Tom, et al.
Published: (2026)
by: Kempton, Tom, et al.
Published: (2026)
CommonMorph: Participatory Morphological Documentation Platform
by: Mahmudi, Aso, et al.
Published: (2026)
by: Mahmudi, Aso, et al.
Published: (2026)
Can a Neural Model Guide Fieldwork? A Case Study on Morphological Data Collection
by: Mahmudi, Aso, et al.
Published: (2024)
by: Mahmudi, Aso, et al.
Published: (2024)
Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations
by: Castilho, Sheila, et al.
Published: (2025)
by: Castilho, Sheila, et al.
Published: (2025)
Pre-training Cross-lingual Open Domain Question Answering with Large-scale Synthetic Supervision
by: Jiang, Fan, et al.
Published: (2024)
by: Jiang, Fan, et al.
Published: (2024)
Chasing Progress, Not Perfection: Revisiting Strategies for End-to-End LLM Plan Generation
by: Huang, Sukai, et al.
Published: (2024)
by: Huang, Sukai, et al.
Published: (2024)
Don't Ignore the Tail: Decoupling top-K Probabilities for Efficient Language Model Distillation
by: Dasgupta, Sayantan, et al.
Published: (2026)
by: Dasgupta, Sayantan, et al.
Published: (2026)
Large Reasoning Models Struggle to Transfer Parametric Knowledge Across Scripts
by: Bandarkar, Lucas, et al.
Published: (2026)
by: Bandarkar, Lucas, et al.
Published: (2026)
Few-Shot Multilingual Open-Domain QA from 5 Examples
by: Jiang, Fan, et al.
Published: (2025)
by: Jiang, Fan, et al.
Published: (2025)
Quantity Convergence, Quality Divergence: Disentangling Fluency and Accuracy in L2 Mandarin Prosody
by: Shi, Yuqi, et al.
Published: (2026)
by: Shi, Yuqi, et al.
Published: (2026)
Franken-Adapter: Cross-Lingual Adaptation of LLMs by Embedding Surgery
by: Jiang, Fan, et al.
Published: (2025)
by: Jiang, Fan, et al.
Published: (2025)
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
Knowledge Localization in Mixture-of-Experts LLMs Using Cross-Lingual Inconsistency
by: Bandarkar, Lucas, et al.
Published: (2026)
by: Bandarkar, Lucas, et al.
Published: (2026)
Learning Robust Negation Text Representations
by: Truong, Thinh Hung, et al.
Published: (2025)
by: Truong, Thinh Hung, et al.
Published: (2025)
Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models
by: Srivastava, Gaurav, et al.
Published: (2025)
by: Srivastava, Gaurav, et al.
Published: (2025)
Zero-Shot Performance Prediction for Probabilistic Scaling Laws
by: Schram, Viktoria, et al.
Published: (2025)
by: Schram, Viktoria, et al.
Published: (2025)
Multi-EuP: The Multilingual European Parliament Dataset for Analysis of Bias in Information Retrieval
by: Yang, Jinrui, et al.
Published: (2023)
by: Yang, Jinrui, et al.
Published: (2023)
Machine Translation Meta Evaluation through Translation Accuracy Challenge Sets
by: Moghe, Nikita, et al.
Published: (2024)
by: Moghe, Nikita, et al.
Published: (2024)
The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models
by: Chen, Yanjun, et al.
Published: (2024)
by: Chen, Yanjun, et al.
Published: (2024)
Revisiting subword tokenization: A case study on affixal negation in large language models
by: Truong, Thinh Hung, et al.
Published: (2024)
by: Truong, Thinh Hung, et al.
Published: (2024)
The Fine-Tuning Paradox: Boosting Translation Quality Without Sacrificing LLM Abilities
by: Stap, David, et al.
Published: (2024)
by: Stap, David, et al.
Published: (2024)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
Similar Items
-
OpenWHO: A Document-Level Parallel Corpus for Health Translation in Low-Resource Languages
by: Merx, Raphaël, et al.
Published: (2025) -
TULUN: Transparent and Adaptable Low-resource Machine Translation
by: Merx, Raphaël, et al.
Published: (2025) -
Mufu: Multilingual Fused Learning for Low-Resource Translation with LLM
by: Lim, Zheng Wei, et al.
Published: (2024) -
Language-Specific Latent Process Hinders Cross-Lingual Performance
by: Lim, Zheng Wei, et al.
Published: (2025) -
A Multidimensional Framework for Evaluating Lexical Semantic Change with Social Science Applications
by: Baes, Naomi, et al.
Published: (2024)