Evaluating Automatic Metrics with Incremental Machine Translation Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Guojun, Cohen, Shay B., Sennrich, Rico |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Machine Translation Models are Zero-Shot Detectors of Translation Direction
di: Wastl, Michelle, et al.
Pubblicazione: (2024)
di: Wastl, Michelle, et al.
Pubblicazione: (2024)
Mitigating Hallucinations and Off-target Machine Translation with Source-Contrastive and Language-Contrastive Decoding
di: Sennrich, Rico, et al.
Pubblicazione: (2023)
di: Sennrich, Rico, et al.
Pubblicazione: (2023)
Investigating Multi-Pivot Ensembling with Massively Multilingual Machine Translation Models
di: Mohammadshahi, Alireza, et al.
Pubblicazione: (2023)
di: Mohammadshahi, Alireza, et al.
Pubblicazione: (2023)
Machine Translation Meta Evaluation through Translation Accuracy Challenge Sets
di: Moghe, Nikita, et al.
Pubblicazione: (2024)
di: Moghe, Nikita, et al.
Pubblicazione: (2024)
An Analysis of BPE Vocabulary Trimming in Neural Machine Translation
di: Cognetta, Marco, et al.
Pubblicazione: (2024)
di: Cognetta, Marco, et al.
Pubblicazione: (2024)
Source-primed Multi-turn Conversation Helps Large Language Models Translate Documents
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
Linear-time Minimum Bayes Risk Decoding with Reference Aggregation
di: Vamvas, Jannis, et al.
Pubblicazione: (2024)
di: Vamvas, Jannis, et al.
Pubblicazione: (2024)
Measuring the Effect of Disfluency in Multilingual Knowledge Probing Benchmarks
di: Semenov, Kirill, et al.
Pubblicazione: (2025)
di: Semenov, Kirill, et al.
Pubblicazione: (2025)
Quality and Quantity of Machine Translation References for Automatic Metrics
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
Modular Adaptation of Multilingual Encoders to Written Swiss German Dialect
di: Vamvas, Jannis, et al.
Pubblicazione: (2024)
di: Vamvas, Jannis, et al.
Pubblicazione: (2024)
Examining Multilingual Embedding Models Cross-Lingually Through LLM-Generated Adversarial Examples
di: Michail, Andrianos, et al.
Pubblicazione: (2025)
di: Michail, Andrianos, et al.
Pubblicazione: (2025)
SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents
di: Wastl, Michelle, et al.
Pubblicazione: (2025)
di: Wastl, Michelle, et al.
Pubblicazione: (2025)
Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?
di: Kew, Tannon, et al.
Pubblicazione: (2023)
di: Kew, Tannon, et al.
Pubblicazione: (2023)
SwissBERT: The Multilingual Language Model for Switzerland
di: Vamvas, Jannis, et al.
Pubblicazione: (2023)
di: Vamvas, Jannis, et al.
Pubblicazione: (2023)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
di: Wang, Shun, et al.
Pubblicazione: (2024)
di: Wang, Shun, et al.
Pubblicazione: (2024)
Towards Explainable Evaluation Metrics for Machine Translation
di: Leiter, Christoph, et al.
Pubblicazione: (2023)
di: Leiter, Christoph, et al.
Pubblicazione: (2023)
20min-XD: A Comparable Corpus of Swiss News Articles
di: Wastl, Michelle, et al.
Pubblicazione: (2025)
di: Wastl, Michelle, et al.
Pubblicazione: (2025)
Translation Asymmetry in LLMs as a Data Augmentation Factor: A Case Study for 6 Romansh Language Varieties
di: Vamvas, Jannis, et al.
Pubblicazione: (2026)
di: Vamvas, Jannis, et al.
Pubblicazione: (2026)
PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation
di: Proietti, Lorenzo, et al.
Pubblicazione: (2026)
di: Proietti, Lorenzo, et al.
Pubblicazione: (2026)
AskQE: Question Answering as Automatic Evaluation for Machine Translation
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
Extending Automatic Machine Translation Evaluation to Book-Length Documents
di: Wang, Kuang-Da, et al.
Pubblicazione: (2025)
di: Wang, Kuang-Da, et al.
Pubblicazione: (2025)
Automatic Evaluation Metrics for Document-level Translation: Overview, Challenges and Trends
di: GUO, Jiaxin, et al.
Pubblicazione: (2025)
di: GUO, Jiaxin, et al.
Pubblicazione: (2025)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
Can Automatic Metrics Assess High-Quality Translations?
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
Leveraging In-Context Learning for Political Bias Testing of LLMs
di: Haller, Patrick, et al.
Pubblicazione: (2025)
di: Haller, Patrick, et al.
Pubblicazione: (2025)
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
The Mediomatix Corpus: Parallel Data for Romansh Language Varieties via Comparable Schoolbooks
di: Hopton, Zachary, et al.
Pubblicazione: (2025)
di: Hopton, Zachary, et al.
Pubblicazione: (2025)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
Modeling User Preferences with Automatic Metrics: Creating a High-Quality Preference Dataset for Machine Translation
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
QueST: Incentivizing LLMs to Generate Difficult Problems
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
Can Large Language Model Summarizers Adapt to Diverse Scientific Communication Goals?
di: Fonseca, Marcio, et al.
Pubblicazione: (2024)
di: Fonseca, Marcio, et al.
Pubblicazione: (2024)
Can Large Language Models Follow Concept Annotation Guidelines? A Case Study on Scientific and Financial Domains
di: Fonseca, Marcio, et al.
Pubblicazione: (2023)
di: Fonseca, Marcio, et al.
Pubblicazione: (2023)
Trainable Reference-Based Evaluation Metric for Identifying Quality of English-Gujarati Machine Translation System
di: Joshi, Nisheeth, et al.
Pubblicazione: (2025)
di: Joshi, Nisheeth, et al.
Pubblicazione: (2025)
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language
di: Sennrich, Kilian, et al.
Pubblicazione: (2025)
di: Sennrich, Kilian, et al.
Pubblicazione: (2025)
Automatically Generating Chinese Homophone Words to Probe Machine Translation Estimation Systems
di: Qian, Shenbin, et al.
Pubblicazione: (2025)
di: Qian, Shenbin, et al.
Pubblicazione: (2025)
Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
Robust Native Language Identification through Agentic Decomposition
di: Uluslu, Ahmet Yavuz, et al.
Pubblicazione: (2025)
di: Uluslu, Ahmet Yavuz, et al.
Pubblicazione: (2025)
An Automatic Quality Metric for Evaluating Simultaneous Interpretation
di: Makinae, Mana, et al.
Pubblicazione: (2024)
di: Makinae, Mana, et al.
Pubblicazione: (2024)
LeanReasoner: Boosting Complex Logical Reasoning with Lean
di: Jiang, Dongwei, et al.
Pubblicazione: (2024)
di: Jiang, Dongwei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Machine Translation Models are Zero-Shot Detectors of Translation Direction
di: Wastl, Michelle, et al.
Pubblicazione: (2024) -
Mitigating Hallucinations and Off-target Machine Translation with Source-Contrastive and Language-Contrastive Decoding
di: Sennrich, Rico, et al.
Pubblicazione: (2023) -
Investigating Multi-Pivot Ensembling with Massively Multilingual Machine Translation Models
di: Mohammadshahi, Alireza, et al.
Pubblicazione: (2023) -
Machine Translation Meta Evaluation through Translation Accuracy Challenge Sets
di: Moghe, Nikita, et al.
Pubblicazione: (2024) -
An Analysis of BPE Vocabulary Trimming in Neural Machine Translation
di: Cognetta, Marco, et al.
Pubblicazione: (2024)