Adding Chocolate to Mint: Mitigating Metric Interference in Machine Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Pombal, José, Guerreiro, Nuno M., Rei, Ricardo, Martins, André F. T. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zero-shot Benchmarking: A Framework for Flexible and Scalable Automatic Evaluation of Language Models
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
Tower+: Bridging Generality and Translation Specialization in Multilingual LLMs
por: Rei, Ricardo, et al.
Publicado: (2025)
por: Rei, Ricardo, et al.
Publicado: (2025)
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models
por: Pombal, José, et al.
Publicado: (2026)
por: Pombal, José, et al.
Publicado: (2026)
Translate Smart, not Hard: Cascaded Translation Systems with Quality-Aware Deferral
por: Farinhas, António, et al.
Publicado: (2025)
por: Farinhas, António, et al.
Publicado: (2025)
MindEval: Benchmarking Language Models on Multi-turn Mental Health Support
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
M-Prometheus: A Suite of Open Multilingual LLM Judges
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
Analyzing Context Contributions in LLM-based Machine Translation
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
Modeling User Preferences with Automatic Metrics: Creating a High-Quality Preference Dataset for Machine Translation
por: Agrawal, Sweta, et al.
Publicado: (2024)
por: Agrawal, Sweta, et al.
Publicado: (2024)
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors
por: Treviso, Marcos, et al.
Publicado: (2024)
por: Treviso, Marcos, et al.
Publicado: (2024)
EuroLLM-9B: Technical Report
por: Martins, Pedro Henrique, et al.
Publicado: (2025)
por: Martins, Pedro Henrique, et al.
Publicado: (2025)
EuroLLM-22B: Technical Report
por: Ramos, Miguel Moura, et al.
Publicado: (2026)
por: Ramos, Miguel Moura, et al.
Publicado: (2026)
LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation
por: Magdy, Samar M., et al.
Publicado: (2026)
por: Magdy, Samar M., et al.
Publicado: (2026)
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
por: Perrella, Stefano, et al.
Publicado: (2024)
por: Perrella, Stefano, et al.
Publicado: (2024)
Can Automatic Metrics Assess High-Quality Translations?
por: Agrawal, Sweta, et al.
Publicado: (2024)
por: Agrawal, Sweta, et al.
Publicado: (2024)
An Analysis on Automated Metrics for Evaluating Japanese-English Chat Translation
por: Rusli, Andre, et al.
Publicado: (2024)
por: Rusli, Andre, et al.
Publicado: (2024)
Fine-Tuned Machine Translation Metrics Struggle in Unseen Domains
por: Zouhar, Vilém, et al.
Publicado: (2024)
por: Zouhar, Vilém, et al.
Publicado: (2024)
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
por: Perrella, Stefano, et al.
Publicado: (2024)
por: Perrella, Stefano, et al.
Publicado: (2024)
Textual Similarity as a Key Metric in Machine Translation Quality Estimation
por: Sun, Kun, et al.
Publicado: (2024)
por: Sun, Kun, et al.
Publicado: (2024)
Fine-Grained and Multi-Dimensional Metrics for Document-Level Machine Translation
por: Sun, Yirong, et al.
Publicado: (2024)
por: Sun, Yirong, et al.
Publicado: (2024)
Tower: An Open Multilingual Large Language Model for Translation-Related Tasks
por: Alves, Duarte M., et al.
Publicado: (2024)
por: Alves, Duarte M., et al.
Publicado: (2024)
Mitigating Stylistic Biases of Machine Translation Systems via Monolingual Corpora Only
por: Gao, Xuanqi, et al.
Publicado: (2025)
por: Gao, Xuanqi, et al.
Publicado: (2025)
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation
por: Prahallad, Lavanya, et al.
Publicado: (2024)
por: Prahallad, Lavanya, et al.
Publicado: (2024)
An Interdisciplinary Approach to Human-Centered Machine Translation
por: Carpuat, Marine, et al.
Publicado: (2025)
por: Carpuat, Marine, et al.
Publicado: (2025)
Trainable Reference-Based Evaluation Metric for Identifying Quality of English-Gujarati Machine Translation System
por: Joshi, Nisheeth, et al.
Publicado: (2025)
por: Joshi, Nisheeth, et al.
Publicado: (2025)
Lottery Ticket Adaptation: Mitigating Destructive Interference in LLMs
por: Panda, Ashwinee, et al.
Publicado: (2024)
por: Panda, Ashwinee, et al.
Publicado: (2024)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
por: Anugraha, David, et al.
Publicado: (2024)
por: Anugraha, David, et al.
Publicado: (2024)
MindGuard: Guardrail Classifiers for Multi-Turn Mental Health Support
por: Farinhas, António, et al.
Publicado: (2026)
por: Farinhas, António, et al.
Publicado: (2026)
EuroBERT: Scaling Multilingual Encoders for European Languages
por: Boizard, Nicolas, et al.
Publicado: (2025)
por: Boizard, Nicolas, et al.
Publicado: (2025)
Mitigating Metric Bias in Minimum Bayes Risk Decoding
por: Kovacs, Geza, et al.
Publicado: (2024)
por: Kovacs, Geza, et al.
Publicado: (2024)
A Context-aware Framework for Translation-mediated Conversations
por: Pombal, José, et al.
Publicado: (2024)
por: Pombal, José, et al.
Publicado: (2024)
Sociotechnical Effects of Machine Translation
por: Moorkens, Joss, et al.
Publicado: (2025)
por: Moorkens, Joss, et al.
Publicado: (2025)
What's Holding Back Latent Visual Reasoning?
por: Viveiros, André G., et al.
Publicado: (2026)
por: Viveiros, André G., et al.
Publicado: (2026)
Automatic Evaluation Metrics for Document-level Translation: Overview, Challenges and Trends
por: GUO, Jiaxin, et al.
Publicado: (2025)
por: GUO, Jiaxin, et al.
Publicado: (2025)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
por: Cettolo, Mauro, et al.
Publicado: (2025)
por: Cettolo, Mauro, et al.
Publicado: (2025)
Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation
por: Guan, Yiwen, et al.
Publicado: (2025)
por: Guan, Yiwen, et al.
Publicado: (2025)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
por: Moon, Hyeonseok, et al.
Publicado: (2024)
por: Moon, Hyeonseok, et al.
Publicado: (2024)
Generating Gender Alternatives in Machine Translation
por: Garg, Sarthak, et al.
Publicado: (2024)
por: Garg, Sarthak, et al.
Publicado: (2024)
Word Alignment as Preference for Machine Translation
por: Wu, Qiyu, et al.
Publicado: (2024)
por: Wu, Qiyu, et al.
Publicado: (2024)
Interplay of Machine Translation, Diacritics, and Diacritization
por: Chen, Wei-Rui, et al.
Publicado: (2024)
por: Chen, Wei-Rui, et al.
Publicado: (2024)
Glancing Future for Simultaneous Machine Translation
por: Guo, Shoutao, et al.
Publicado: (2023)
por: Guo, Shoutao, et al.
Publicado: (2023)
Ejemplares similares
-
Zero-shot Benchmarking: A Framework for Flexible and Scalable Automatic Evaluation of Language Models
por: Pombal, José, et al.
Publicado: (2025) -
Tower+: Bridging Generality and Translation Specialization in Multilingual LLMs
por: Rei, Ricardo, et al.
Publicado: (2025) -
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models
por: Pombal, José, et al.
Publicado: (2026) -
Translate Smart, not Hard: Cascaded Translation Systems with Quality-Aware Deferral
por: Farinhas, António, et al.
Publicado: (2025) -
MindEval: Benchmarking Language Models on Multi-turn Mental Health Support
por: Pombal, José, et al.
Publicado: (2025)