Salvato in:
| Autori principali: | Perrella, Stefano, Agostinho, Eric Morales, Zaragoza, Hugo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.19921 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
Has Machine Translation Evaluation Achieved Human Parity? The Human Reference and the Limits of Progress
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation
di: Lyu, Boxuan, et al.
Pubblicazione: (2025)
di: Lyu, Boxuan, et al.
Pubblicazione: (2025)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
di: Shayegh, Behzad, et al.
Pubblicazione: (2025)
di: Shayegh, Behzad, et al.
Pubblicazione: (2025)
Fine-Grained and Multi-Dimensional Metrics for Document-Level Machine Translation
di: Sun, Yirong, et al.
Pubblicazione: (2024)
di: Sun, Yirong, et al.
Pubblicazione: (2024)
Evaluation of Machine Translation Based on Semantic Dependencies and Keywords
di: Yuan, Kewei, et al.
Pubblicazione: (2024)
di: Yuan, Kewei, et al.
Pubblicazione: (2024)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
Should I Share this Translation? Evaluating Quality Feedback for User Reliance on Machine Translation
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
Estimating Machine Translation Difficulty
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
di: Kreutzer, Julia, et al.
Pubblicazione: (2025)
di: Kreutzer, Julia, et al.
Pubblicazione: (2025)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
di: Anugraha, David, et al.
Pubblicazione: (2024)
di: Anugraha, David, et al.
Pubblicazione: (2024)
FairTranslate: An English-French Dataset for Gender Bias Evaluation in Machine Translation by Overcoming Gender Binarity
di: Jourdan, Fanny, et al.
Pubblicazione: (2025)
di: Jourdan, Fanny, et al.
Pubblicazione: (2025)
Align-then-Slide: A complete evaluation framework for Ultra-Long Document-Level Machine Translation
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
POMP: Probability-driven Meta-graph Prompter for LLMs in Low-resource Unsupervised Neural Machine Translation
di: Pan, Shilong, et al.
Pubblicazione: (2024)
di: Pan, Shilong, et al.
Pubblicazione: (2024)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
di: Polák, Peter, et al.
Pubblicazione: (2025)
di: Polák, Peter, et al.
Pubblicazione: (2025)
Meta-Judging with Large Language Models: Concepts, Methods, and Challenges
di: Silva, Hugo, et al.
Pubblicazione: (2026)
di: Silva, Hugo, et al.
Pubblicazione: (2026)
Evaluating the Meta- and Object-Level Reasoning of Large Language Models for Question Answering
di: Ferguson, Nick, et al.
Pubblicazione: (2025)
di: Ferguson, Nick, et al.
Pubblicazione: (2025)
Enhancing Document-Level Machine Translation via Filtered Synthetic Corpora and Two-Stage LLM Adaptation
di: Kim, Ireh, et al.
Pubblicazione: (2026)
di: Kim, Ireh, et al.
Pubblicazione: (2026)
M-MAD: Multidimensional Multi-Agent Debate for Advanced Machine Translation Evaluation
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
Sentiment Analysis Across Languages: Evaluation Before and After Machine Translation to English
di: Kathunia, Aekansh, et al.
Pubblicazione: (2024)
di: Kathunia, Aekansh, et al.
Pubblicazione: (2024)
Sociotechnical Effects of Machine Translation
di: Moorkens, Joss, et al.
Pubblicazione: (2025)
di: Moorkens, Joss, et al.
Pubblicazione: (2025)
SLIDE: Reference-free Evaluation for Machine Translation using a Sliding Document Window
di: Raunak, Vikas, et al.
Pubblicazione: (2023)
di: Raunak, Vikas, et al.
Pubblicazione: (2023)
Lost in the Source Language: How Large Language Models Evaluate the Quality of Machine Translation
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
di: Powers, Maximus, et al.
Pubblicazione: (2024)
di: Powers, Maximus, et al.
Pubblicazione: (2024)
Scaling Bidirectional Spans and Span Violations in Attention Mechanism
di: Kim, Jongwook, et al.
Pubblicazione: (2025)
di: Kim, Jongwook, et al.
Pubblicazione: (2025)
Convergences and Divergences between Automatic Assessment and Human Evaluation: Insights from Comparing ChatGPT-Generated Translation and Neural Machine Translation
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation
di: Guan, Yiwen, et al.
Pubblicazione: (2025)
di: Guan, Yiwen, et al.
Pubblicazione: (2025)
Generating Gender Alternatives in Machine Translation
di: Garg, Sarthak, et al.
Pubblicazione: (2024)
di: Garg, Sarthak, et al.
Pubblicazione: (2024)
Word Alignment as Preference for Machine Translation
di: Wu, Qiyu, et al.
Pubblicazione: (2024)
di: Wu, Qiyu, et al.
Pubblicazione: (2024)
Interplay of Machine Translation, Diacritics, and Diacritization
di: Chen, Wei-Rui, et al.
Pubblicazione: (2024)
di: Chen, Wei-Rui, et al.
Pubblicazione: (2024)
Glancing Future for Simultaneous Machine Translation
di: Guo, Shoutao, et al.
Pubblicazione: (2023)
di: Guo, Shoutao, et al.
Pubblicazione: (2023)
Trainable Reference-Based Evaluation Metric for Identifying Quality of English-Gujarati Machine Translation System
di: Joshi, Nisheeth, et al.
Pubblicazione: (2025)
di: Joshi, Nisheeth, et al.
Pubblicazione: (2025)
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels
di: Yan, Jianhao, et al.
Pubblicazione: (2024)
di: Yan, Jianhao, et al.
Pubblicazione: (2024)
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
di: Pires, Ramon, et al.
Pubblicazione: (2026)
di: Pires, Ramon, et al.
Pubblicazione: (2026)
DelTA: An Online Document-Level Translation Agent Based on Multi-Level Memory
di: Wang, Yutong, et al.
Pubblicazione: (2024)
di: Wang, Yutong, et al.
Pubblicazione: (2024)
Memory Reviving, Continuing Learning and Beyond: Evaluation of Pre-trained Encoders and Decoders for Multimodal Machine Translation
di: Yu, Zhuang, et al.
Pubblicazione: (2025)
di: Yu, Zhuang, et al.
Pubblicazione: (2025)
Overestimation in LLM Evaluation: A Controlled Large-Scale Study on Data Contamination's Impact on Machine Translation
di: Kocyigit, Muhammed Yusuf, et al.
Pubblicazione: (2025)
di: Kocyigit, Muhammed Yusuf, et al.
Pubblicazione: (2025)
Retrieval-Augmented Machine Translation with Unstructured Knowledge
di: Wang, Jiaan, et al.
Pubblicazione: (2024)
di: Wang, Jiaan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
di: Perrella, Stefano, et al.
Pubblicazione: (2024) -
Has Machine Translation Evaluation Achieved Human Parity? The Human Reference and the Limits of Progress
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025) -
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
di: Perrella, Stefano, et al.
Pubblicazione: (2024) -
Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation
di: Lyu, Boxuan, et al.
Pubblicazione: (2025) -
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
di: Shayegh, Behzad, et al.
Pubblicazione: (2025)