Uncertainty Quantification for Evaluating Machine Translation Bias
Fuente:
arXiv
Saved in:
| Main Authors: | Staliūnaitė, Ieva Raminta, Cheng, Julius, Vlachos, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
by: Staliūnaitė, Ieva Raminta, et al.
Published: (2026)
by: Staliūnaitė, Ieva Raminta, et al.
Published: (2026)
A Bayesian Optimization Approach to Machine Translation Reranking
by: Cheng, Julius, et al.
Published: (2024)
by: Cheng, Julius, et al.
Published: (2024)
Gender Inflected or Bias Inflicted: On Using Grammatical Gender Cues for Bias Evaluation in Machine Translation
by: Singh, Pushpdeep
Published: (2023)
by: Singh, Pushpdeep
Published: (2023)
Creativity Bias: How Machine Evaluation Struggles with Creativity in Literary Translations
by: Gerrits, Kyo, et al.
Published: (2026)
by: Gerrits, Kyo, et al.
Published: (2026)
TSVer: A Benchmark for Fact Verification Against Time-Series Evidence
by: Strong, Marek, et al.
Published: (2025)
by: Strong, Marek, et al.
Published: (2025)
Capturing Symmetry and Antisymmetry in Language Models through Symmetry-Aware Training Objectives
by: Yuan, Zhangdie, et al.
Published: (2025)
by: Yuan, Zhangdie, et al.
Published: (2025)
TabVer: Tabular Fact Verification with Natural Logic
by: Aly, Rami, et al.
Published: (2024)
by: Aly, Rami, et al.
Published: (2024)
Are All Spanish Doctors Male? Evaluating Gender Bias in German Machine Translation
by: Kappl, Michelle
Published: (2025)
by: Kappl, Michelle
Published: (2025)
Causal Estimation of Tokenisation Bias
by: Lesci, Pietro, et al.
Published: (2025)
by: Lesci, Pietro, et al.
Published: (2025)
Gender Bias in English-to-Greek Machine Translation
by: Gkovedarou, Eleni, et al.
Published: (2025)
by: Gkovedarou, Eleni, et al.
Published: (2025)
Good, but not always Fair: An Evaluation of Gender Bias for three commercial Machine Translation Systems
by: Piazzolla, Silvia Alma, et al.
Published: (2023)
by: Piazzolla, Silvia Alma, et al.
Published: (2023)
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
by: Santilli, Andrea, et al.
Published: (2025)
by: Santilli, Andrea, et al.
Published: (2025)
FairTranslate: An English-French Dataset for Gender Bias Evaluation in Machine Translation by Overcoming Gender Binarity
by: Jourdan, Fanny, et al.
Published: (2025)
by: Jourdan, Fanny, et al.
Published: (2025)
GAMBIT+: A Challenge Set for Evaluating Gender Bias in Machine Translation Quality Estimation Metrics
by: Filandrianos, Giorgos, et al.
Published: (2025)
by: Filandrianos, Giorgos, et al.
Published: (2025)
AnchorAL: Computationally Efficient Active Learning for Large and Imbalanced Datasets
by: Lesci, Pietro, et al.
Published: (2024)
by: Lesci, Pietro, et al.
Published: (2024)
MAQA: Evaluating Uncertainty Quantification in LLMs Regarding Data Uncertainty
by: Yang, Yongjin, et al.
Published: (2024)
by: Yang, Yongjin, et al.
Published: (2024)
Beyond BLEU: A Semantic Evaluation Method for Code Translation
by: Näumann, Julius, et al.
Published: (2026)
by: Näumann, Julius, et al.
Published: (2026)
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
by: Goloburda, Maiya, et al.
Published: (2026)
by: Goloburda, Maiya, et al.
Published: (2026)
PRobELM: Plausibility Ranking Evaluation for Language Models
by: Yuan, Zhangdie, et al.
Published: (2024)
by: Yuan, Zhangdie, et al.
Published: (2024)
Improving Zero-shot Sentence Decontextualisation with Content Selection and Planning
by: Deng, Zhenyun, et al.
Published: (2025)
by: Deng, Zhenyun, et al.
Published: (2025)
Do Language Models Update their Forecasts with New Information?
by: Yuan, Zhangdie, et al.
Published: (2025)
by: Yuan, Zhangdie, et al.
Published: (2025)
Document-level Claim Extraction and Decontextualisation for Fact-Checking
by: Deng, Zhenyun, et al.
Published: (2024)
by: Deng, Zhenyun, et al.
Published: (2024)
The effect of diversity on group decision-making
by: Karadzhov, Georgi, et al.
Published: (2024)
by: Karadzhov, Georgi, et al.
Published: (2024)
Zero-Shot Fact Verification via Natural Logic and Large Language Models
by: Strong, Marek, et al.
Published: (2024)
by: Strong, Marek, et al.
Published: (2024)
Do We Need Language-Specific Fact-Checking Models? The Case of Chinese
by: Zhang, Caiqi, et al.
Published: (2024)
by: Zhang, Caiqi, et al.
Published: (2024)
Automated Focused Feedback Generation for Scientific Writing Assistance
by: Chamoun, Eric, et al.
Published: (2024)
by: Chamoun, Eric, et al.
Published: (2024)
Investigating Markers and Drivers of Gender Bias in Machine Translations
by: Barclay, Peter J, et al.
Published: (2024)
by: Barclay, Peter J, et al.
Published: (2024)
Bias in News Summarization: Measures, Pitfalls and Corpora
by: Steen, Julius, et al.
Published: (2023)
by: Steen, Julius, et al.
Published: (2023)
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking
by: Akhtar, Mubashara, et al.
Published: (2024)
by: Akhtar, Mubashara, et al.
Published: (2024)
Does Context Help Mitigate Gender Bias in Neural Machine Translation?
by: Gete, Harritxu, et al.
Published: (2024)
by: Gete, Harritxu, et al.
Published: (2024)
COPU: Conformal Prediction for Uncertainty Quantification in Natural Language Generation
by: Wang, Sean, et al.
Published: (2025)
by: Wang, Sean, et al.
Published: (2025)
Gender Bias in Machine Translation and The Era of Large Language Models
by: Vanmassenhove, Eva
Published: (2024)
by: Vanmassenhove, Eva
Published: (2024)
Agentic Uncertainty Quantification
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models
by: Zhou, Kevin, et al.
Published: (2025)
by: Zhou, Kevin, et al.
Published: (2025)
Uncertainty Quantification for LLM Function-Calling
by: Ye, Zihuiwen, et al.
Published: (2026)
by: Ye, Zihuiwen, et al.
Published: (2026)
Benchmarking LLMs via Uncertainty Quantification
by: Ye, Fanghua, et al.
Published: (2024)
by: Ye, Fanghua, et al.
Published: (2024)
FOReCAst: The Future Outcome Reasoning and Confidence Assessment Benchmark
by: Yuan, Zhangdie, et al.
Published: (2025)
by: Yuan, Zhangdie, et al.
Published: (2025)
Machine Translation Meta Evaluation through Translation Accuracy Challenge Sets
by: Moghe, Nikita, et al.
Published: (2024)
by: Moghe, Nikita, et al.
Published: (2024)
Assumed Identities: Quantifying Gender Bias in Machine Translation of Gender-Ambiguous Occupational Terms
by: Mastromichalakis, Orfeas Menis, et al.
Published: (2025)
by: Mastromichalakis, Orfeas Menis, et al.
Published: (2025)
Evaluating Structural Generalization in Neural Machine Translation
by: Kumon, Ryoma, et al.
Published: (2024)
by: Kumon, Ryoma, et al.
Published: (2024)
Similar Items
-
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
by: Staliūnaitė, Ieva Raminta, et al.
Published: (2026) -
A Bayesian Optimization Approach to Machine Translation Reranking
by: Cheng, Julius, et al.
Published: (2024) -
Gender Inflected or Bias Inflicted: On Using Grammatical Gender Cues for Bias Evaluation in Machine Translation
by: Singh, Pushpdeep
Published: (2023) -
Creativity Bias: How Machine Evaluation Struggles with Creativity in Literary Translations
by: Gerrits, Kyo, et al.
Published: (2026) -
TSVer: A Benchmark for Fact Verification Against Time-Series Evidence
by: Strong, Marek, et al.
Published: (2025)