Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Perrella, Stefano, Proietti, Lorenzo, Cabot, Pere-Lluís Huguet, Barba, Edoardo, Navigli, Roberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
von: Perrella, Stefano, et al.
Veröffentlicht: (2024)
von: Perrella, Stefano, et al.
Veröffentlicht: (2024)
Has Machine Translation Evaluation Achieved Human Parity? The Human Reference and the Limits of Progress
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
ReLiK: Retrieve and LinK, Fast and Accurate Entity Linking and Relation Extraction on an Academic Budget
von: Orlando, Riccardo, et al.
Veröffentlicht: (2024)
von: Orlando, Riccardo, et al.
Veröffentlicht: (2024)
BOOKCOREF: Coreference Resolution at Book Scale
von: Martinelli, Giuliano, et al.
Veröffentlicht: (2025)
von: Martinelli, Giuliano, et al.
Veröffentlicht: (2025)
Estimating Machine Translation Difficulty
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends
von: Martinelli, Giuliano, et al.
Veröffentlicht: (2024)
von: Martinelli, Giuliano, et al.
Veröffentlicht: (2024)
Span-Level Machine Translation Meta-Evaluation
von: Perrella, Stefano, et al.
Veröffentlicht: (2026)
von: Perrella, Stefano, et al.
Veröffentlicht: (2026)
Interpretable Coreference Resolution Evaluation Using Explicit Semantics
von: Gatti, Bruno, et al.
Veröffentlicht: (2026)
von: Gatti, Bruno, et al.
Veröffentlicht: (2026)
Word Sense Linking: Disambiguating Outside the Sandbox
von: Bejgu, Andrei Stefan, et al.
Veröffentlicht: (2024)
von: Bejgu, Andrei Stefan, et al.
Veröffentlicht: (2024)
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
von: Moroni, Luca, et al.
Veröffentlicht: (2025)
von: Moroni, Luca, et al.
Veröffentlicht: (2025)
LiteraryQA: Towards Effective Evaluation of Long-document Narrative QA
von: Bonomo, Tommaso, et al.
Veröffentlicht: (2025)
von: Bonomo, Tommaso, et al.
Veröffentlicht: (2025)
AutoML-guided Fusion of Entity and LLM-based Representations for Document Classification
von: Koloski, Boshko, et al.
Veröffentlicht: (2024)
von: Koloski, Boshko, et al.
Veröffentlicht: (2024)
Trainable Reference-Based Evaluation Metric for Identifying Quality of English-Gujarati Machine Translation System
von: Joshi, Nisheeth, et al.
Veröffentlicht: (2025)
von: Joshi, Nisheeth, et al.
Veröffentlicht: (2025)
Fine-Tuned Machine Translation Metrics Struggle in Unseen Domains
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
Adding Chocolate to Mint: Mitigating Metric Interference in Machine Translation
von: Pombal, José, et al.
Veröffentlicht: (2025)
von: Pombal, José, et al.
Veröffentlicht: (2025)
LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation
von: Magdy, Samar M., et al.
Veröffentlicht: (2026)
von: Magdy, Samar M., et al.
Veröffentlicht: (2026)
Do Large Language Models Understand Word Senses?
von: Meconi, Domenico, et al.
Veröffentlicht: (2025)
von: Meconi, Domenico, et al.
Veröffentlicht: (2025)
Textual Similarity as a Key Metric in Machine Translation Quality Estimation
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
Fine-Grained and Multi-Dimensional Metrics for Document-Level Machine Translation
von: Sun, Yirong, et al.
Veröffentlicht: (2024)
von: Sun, Yirong, et al.
Veröffentlicht: (2024)
An Analysis on Automated Metrics for Evaluating Japanese-English Chat Translation
von: Rusli, Andre, et al.
Veröffentlicht: (2024)
von: Rusli, Andre, et al.
Veröffentlicht: (2024)
Memory Reviving, Continuing Learning and Beyond: Evaluation of Pre-trained Encoders and Decoders for Multimodal Machine Translation
von: Yu, Zhuang, et al.
Veröffentlicht: (2025)
von: Yu, Zhuang, et al.
Veröffentlicht: (2025)
PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2026)
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2026)
Automatic Evaluation Metrics for Document-level Translation: Overview, Challenges and Trends
von: GUO, Jiaxin, et al.
Veröffentlicht: (2025)
von: GUO, Jiaxin, et al.
Veröffentlicht: (2025)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
Emergent Communication Pretraining for Few-Shot Machine Translation
von: Li, Yaoyiran, et al.
Veröffentlicht: (2020)
von: Li, Yaoyiran, et al.
Veröffentlicht: (2020)
Vision-Grounded Machine Interpreting: Improving the Translation Process through Visual Cues
von: Fantinuoli, Claudio
Veröffentlicht: (2025)
von: Fantinuoli, Claudio
Veröffentlicht: (2025)
Evaluation of Machine Translation Based on Semantic Dependencies and Keywords
von: Yuan, Kewei, et al.
Veröffentlicht: (2024)
von: Yuan, Kewei, et al.
Veröffentlicht: (2024)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Should I Share this Translation? Evaluating Quality Feedback for User Reliance on Machine Translation
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
Beyond the Sentence: A Survey on Context-Aware Machine Translation with Large Language Models
von: Appicharla, Ramakrishna, et al.
Veröffentlicht: (2025)
von: Appicharla, Ramakrishna, et al.
Veröffentlicht: (2025)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
von: Polák, Peter, et al.
Veröffentlicht: (2025)
von: Polák, Peter, et al.
Veröffentlicht: (2025)
Beyond Metrics: A Critical Analysis of the Variability in Large Language Model Evaluation Frameworks
von: Pimentel, Marco AF, et al.
Veröffentlicht: (2024)
von: Pimentel, Marco AF, et al.
Veröffentlicht: (2024)
FairTranslate: An English-French Dataset for Gender Bias Evaluation in Machine Translation by Overcoming Gender Binarity
von: Jourdan, Fanny, et al.
Veröffentlicht: (2025)
von: Jourdan, Fanny, et al.
Veröffentlicht: (2025)
Visualizing Uncertainty in Translation Tasks: An Evaluation of LLM Performance and Confidence Metrics
von: Park, Jin Hyun, et al.
Veröffentlicht: (2025)
von: Park, Jin Hyun, et al.
Veröffentlicht: (2025)
Cross-Platform Evaluation of Reasoning Capabilities in Foundation Models
von: de Curtò, J., et al.
Veröffentlicht: (2025)
von: de Curtò, J., et al.
Veröffentlicht: (2025)
Beyond Semantics: Measuring Fine-Grained Emotion Preservation in Small Language Model-Based Machine Translation
von: Wisniewski, Dawid, et al.
Veröffentlicht: (2026)
von: Wisniewski, Dawid, et al.
Veröffentlicht: (2026)
M-MAD: Multidimensional Multi-Agent Debate for Advanced Machine Translation Evaluation
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
Sentiment Analysis Across Languages: Evaluation Before and After Machine Translation to English
von: Kathunia, Aekansh, et al.
Veröffentlicht: (2024)
von: Kathunia, Aekansh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
von: Perrella, Stefano, et al.
Veröffentlicht: (2024) -
Has Machine Translation Evaluation Achieved Human Parity? The Human Reference and the Limits of Progress
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025) -
ReLiK: Retrieve and LinK, Fast and Accurate Entity Linking and Relation Extraction on an Academic Budget
von: Orlando, Riccardo, et al.
Veröffentlicht: (2024) -
BOOKCOREF: Coreference Resolution at Book Scale
von: Martinelli, Giuliano, et al.
Veröffentlicht: (2025) -
Estimating Machine Translation Difficulty
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)