Creativity Bias: How Machine Evaluation Struggles with Creativity in Literary Translations
Fuente:
arXiv
Saved in:
| Main Authors: | Gerrits, Kyo, van Noord, Rik, Arenas, Ana Guerberof |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To MT or not to MT: An eye-tracking study on the reception by Dutch readers of different translation and creativity levels
by: Gerrits, Kyo, et al.
Published: (2025)
by: Gerrits, Kyo, et al.
Published: (2025)
Optimising ChatGPT for creativity in literary translation: A case study from English into Dutch, Chinese, Catalan and Spanish
by: Du, Shuxiang, et al.
Published: (2025)
by: Du, Shuxiang, et al.
Published: (2025)
Towards Tailored Recovery of Lexical Diversity in Literary Machine Translation
by: Ploeger, Esther, et al.
Published: (2024)
by: Ploeger, Esther, et al.
Published: (2024)
What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered Study
by: Savoldi, Beatrice, et al.
Published: (2024)
by: Savoldi, Beatrice, et al.
Published: (2024)
Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation
by: Lai, Huiyuan, et al.
Published: (2024)
by: Lai, Huiyuan, et al.
Published: (2024)
Evaluating the Creativity of LLMs in Persian Literary Text Generation
by: Tourajmehr, Armin, et al.
Published: (2025)
by: Tourajmehr, Armin, et al.
Published: (2025)
Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation
by: Zhang, Ran, et al.
Published: (2026)
by: Zhang, Ran, et al.
Published: (2026)
Are Character-level Translations Worth the Wait? Comparing ByT5 and mT5 for Machine Translation
by: Edman, Lukas, et al.
Published: (2023)
by: Edman, Lukas, et al.
Published: (2023)
LiteraryTaste: A Preference Dataset for Creative Writing Personalization
by: Chung, John Joon Young, et al.
Published: (2025)
by: Chung, John Joon Young, et al.
Published: (2025)
How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs
by: Zhang, Ran, et al.
Published: (2024)
by: Zhang, Ran, et al.
Published: (2024)
Rethinking Creativity Evaluation: A Critical Analysis of Existing Creativity Evaluations
by: Lu, Li-Chun, et al.
Published: (2025)
by: Lu, Li-Chun, et al.
Published: (2025)
CreativEval: Evaluating Creativity of LLM-Based Hardware Code Generation
by: DeLorenzo, Matthew, et al.
Published: (2024)
by: DeLorenzo, Matthew, et al.
Published: (2024)
Playing With AI: How Do State-Of-The-Art Large Language Models Perform in the 1977 Text-Based Adventure Game Zork?
by: Gerrits, Berry
Published: (2026)
by: Gerrits, Berry
Published: (2026)
CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges
by: Wang, Zi-Han, et al.
Published: (2026)
by: Wang, Zi-Han, et al.
Published: (2026)
Beyond Divergent Creativity: A Human-Based Evaluation of Creativity in Large Language Models
by: Nakajima, Kumiko, et al.
Published: (2026)
by: Nakajima, Kumiko, et al.
Published: (2026)
Fluency and Faithfulness in Human and Machine Literary Translation
by: Griebel, Sarah, et al.
Published: (2026)
by: Griebel, Sarah, et al.
Published: (2026)
PMB5: Gaining More Insight into Neural Semantic Parsing with Challenging Benchmarks
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
What Shapes a Creative Machine Mind? Comprehensively Benchmarking Creativity in Foundation Models
by: He, Zicong, et al.
Published: (2025)
by: He, Zicong, et al.
Published: (2025)
How to Evaluate Coreference in Literary Texts?
by: Duron-Tejedor, Ana-Isabel, et al.
Published: (2023)
by: Duron-Tejedor, Ana-Isabel, et al.
Published: (2023)
CreativityPrism: A Holistic Evaluation Framework for Large Language Model Creativity
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
Uncertainty Quantification for Evaluating Machine Translation Bias
by: Staliūnaitė, Ieva Raminta, et al.
Published: (2025)
by: Staliūnaitė, Ieva Raminta, et al.
Published: (2025)
QE4PE: Word-level Quality Estimation for Human Post-Editing
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
Creative and Context-Aware Translation of East Asian Idioms with GPT-4
by: Tang, Kenan, et al.
Published: (2024)
by: Tang, Kenan, et al.
Published: (2024)
Creative Preference Optimization
by: Ismayilzada, Mete, et al.
Published: (2025)
by: Ismayilzada, Mete, et al.
Published: (2025)
Hallucination or Creativity: How to Evaluate AI-Generated Scientific Stories?
by: Argese, Alex, et al.
Published: (2026)
by: Argese, Alex, et al.
Published: (2026)
Creativity in AI: Progresses and Challenges
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Multiple References with Meaningful Variations Improve Literary Machine Translation
by: Wu, Si, et al.
Published: (2024)
by: Wu, Si, et al.
Published: (2024)
IDEAFix: Evaluation Framework for Creative Defixation Prompting in LLMs
by: Carichon, F., et al.
Published: (2026)
by: Carichon, F., et al.
Published: (2026)
Language Models on a Diet: Cost-Efficient Development of Encoders for Closely-Related Languages via Additional Pretraining
by: Ljubešić, Nikola, et al.
Published: (2024)
by: Ljubešić, Nikola, et al.
Published: (2024)
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
by: Qian, Cheng, et al.
Published: (2026)
by: Qian, Cheng, et al.
Published: (2026)
CAP: Evaluation of Persuasive and Creative Image Generation
by: Aghazadeh, Aysan, et al.
Published: (2024)
by: Aghazadeh, Aysan, et al.
Published: (2024)
Fine-Tuned Machine Translation Metrics Struggle in Unseen Domains
by: Zouhar, Vilém, et al.
Published: (2024)
by: Zouhar, Vilém, et al.
Published: (2024)
Redefining <Creative> in Dictionary: Towards an Enhanced Semantic Understanding of Creative Generation
by: Feng, Fu, et al.
Published: (2024)
by: Feng, Fu, et al.
Published: (2024)
Designing and Evaluating Dialogue LLMs for Co-Creative Improvised Theatre
by: Branch, Boyd, et al.
Published: (2024)
by: Branch, Boyd, et al.
Published: (2024)
SimulBench: Evaluating Language Models with Creative Simulation Tasks
by: Jia, Qi, et al.
Published: (2024)
by: Jia, Qi, et al.
Published: (2024)
How Creative Are Large Language Models in Generating Molecules?
by: Tao, Wen, et al.
Published: (2026)
by: Tao, Wen, et al.
Published: (2026)
Gender Inflected or Bias Inflicted: On Using Grammatical Gender Cues for Bias Evaluation in Machine Translation
by: Singh, Pushpdeep
Published: (2023)
by: Singh, Pushpdeep
Published: (2023)
The Reader is the Metric: How Textual Features and Reader Profiles Explain Conflicting Evaluations of AI Creative Writing
by: Marco, Guillermo, et al.
Published: (2025)
by: Marco, Guillermo, et al.
Published: (2025)
Steering Large Language Models to Evaluate and Amplify Creativity
by: Olson, Matthew Lyle, et al.
Published: (2024)
by: Olson, Matthew Lyle, et al.
Published: (2024)
Similar Items
-
To MT or not to MT: An eye-tracking study on the reception by Dutch readers of different translation and creativity levels
by: Gerrits, Kyo, et al.
Published: (2025) -
Optimising ChatGPT for creativity in literary translation: A case study from English into Dutch, Chinese, Catalan and Spanish
by: Du, Shuxiang, et al.
Published: (2025) -
Towards Tailored Recovery of Lexical Diversity in Literary Machine Translation
by: Ploeger, Esther, et al.
Published: (2024) -
What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered Study
by: Savoldi, Beatrice, et al.
Published: (2024) -
Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation
by: Lai, Huiyuan, et al.
Published: (2024)