Dynamic Meta-Metrics: Source-Sentence Conditioned Weighting for MT Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Luke, Vasselli, Justin, Khan, Aditya, Ng, York Hay, Lee, En-Shiun Annie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Less is More: The Effectiveness of Compact Typological Language Representations
von: Ng, York Hay, et al.
Veröffentlicht: (2025)
von: Ng, York Hay, et al.
Veröffentlicht: (2025)
\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding
von: Min, Junghyun, et al.
Veröffentlicht: (2025)
von: Min, Junghyun, et al.
Veröffentlicht: (2025)
Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+
von: Shipton, Mason, et al.
Veröffentlicht: (2025)
von: Shipton, Mason, et al.
Veröffentlicht: (2025)
Improving Explainability of Sentence-level Metrics via Edit-level Attribution for Grammatical Error Correction
von: Goto, Takumi, et al.
Veröffentlicht: (2024)
von: Goto, Takumi, et al.
Veröffentlicht: (2024)
Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+
von: Ng, York Hay, et al.
Veröffentlicht: (2025)
von: Ng, York Hay, et al.
Veröffentlicht: (2025)
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems
von: Vasselli, Justin, et al.
Veröffentlicht: (2025)
von: Vasselli, Justin, et al.
Veröffentlicht: (2025)
OasisSimp: An Open-source Asian-English Sentence Simplification Dataset
von: Liu, Hannah, et al.
Veröffentlicht: (2026)
von: Liu, Hannah, et al.
Veröffentlicht: (2026)
TranslationCorrect: A Unified Framework for Machine Translation Post-Editing with Predictive Error Assistance
von: Wasti, Syed Mekael, et al.
Veröffentlicht: (2025)
von: Wasti, Syed Mekael, et al.
Veröffentlicht: (2025)
Beyond Vanilla Fine-Tuning: Leveraging Multistage, Multilingual, and Domain-Specific Methods for Low-Resource Machine Translation
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2025)
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2025)
BridG MT: Enhancing LLMs' Machine Translation Capabilities with Sentence Bridging and Gradual MT
von: Choi, Seung-Woo, et al.
Veröffentlicht: (2024)
von: Choi, Seung-Woo, et al.
Veröffentlicht: (2024)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
URIEL+: Enhancing Linguistic Inclusion and Usability in a Typological and Multilingual Knowledge Base
von: Khan, Aditya, et al.
Veröffentlicht: (2024)
von: Khan, Aditya, et al.
Veröffentlicht: (2024)
Unlocking Parameter-Efficient Fine-Tuning for Low-Resource Language Translation
von: Su, Tong, et al.
Veröffentlicht: (2024)
von: Su, Tong, et al.
Veröffentlicht: (2024)
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
Enhancing Taiwanese Hokkien Dual Translation by Exploring and Standardizing of Four Writing Systems
von: Lu, Bo-Han, et al.
Veröffentlicht: (2024)
von: Lu, Bo-Han, et al.
Veröffentlicht: (2024)
Multilingual Dialogue Generation and Localization with Dialogue Act Scripting
von: Vasselli, Justin, et al.
Veröffentlicht: (2025)
von: Vasselli, Justin, et al.
Veröffentlicht: (2025)
Rethinking what Matters: Effective and Robust Multilingual Realignment for Low-Resource Languages
von: Nguyen, Quang Phuoc, et al.
Veröffentlicht: (2025)
von: Nguyen, Quang Phuoc, et al.
Veröffentlicht: (2025)
SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects
von: Adelani, David Ifeoluwa, et al.
Veröffentlicht: (2023)
von: Adelani, David Ifeoluwa, et al.
Veröffentlicht: (2023)
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation
von: Barančíková, Petra, et al.
Veröffentlicht: (2025)
von: Barančíková, Petra, et al.
Veröffentlicht: (2025)
A Reproducibility Study on Quantifying Language Similarity: The Impact of Missing Values in the URIEL Knowledge Base
von: Toossi, Hasti, et al.
Veröffentlicht: (2024)
von: Toossi, Hasti, et al.
Veröffentlicht: (2024)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAG
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
AlignFreeze: Navigating the Impact of Realignment on the Layers of Multilingual Models Across Diverse Languages
von: Bakos, Steve, et al.
Veröffentlicht: (2025)
von: Bakos, Steve, et al.
Veröffentlicht: (2025)
Findings of the BEA 2025 Shared Task on Pedagogical Ability Assessment of AI-powered Tutors
von: Kochmar, Ekaterina, et al.
Veröffentlicht: (2025)
von: Kochmar, Ekaterina, et al.
Veröffentlicht: (2025)
SiniticMTError: A Machine Translation Dataset with Error Annotations for Sinitic Languages
von: Liu, Hannah, et al.
Veröffentlicht: (2025)
von: Liu, Hannah, et al.
Veröffentlicht: (2025)
Computational Sentence-level Metrics Predicting Human Sentence Comprehension
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
What am I missing here?: Evaluating Large Language Models for Masked Sentence Prediction
von: Wyatt, Charlie, et al.
Veröffentlicht: (2025)
von: Wyatt, Charlie, et al.
Veröffentlicht: (2025)
Label Confidence Weighted Learning for Target-level Sentence Simplification
von: Qiu, Xinying, et al.
Veröffentlicht: (2024)
von: Qiu, Xinying, et al.
Veröffentlicht: (2024)
FUSE : A Ridge and Random Forest-Based Metric for Evaluating MT in Indigenous Languages
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Contextual Metric Meta-Evaluation by Measuring Local Metric Accuracy
von: Deviyani, Athiya, et al.
Veröffentlicht: (2025)
von: Deviyani, Athiya, et al.
Veröffentlicht: (2025)
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
xCOMET-lite: Bridging the Gap Between Efficiency and Quality in Learned MT Evaluation Metrics
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
Low-Cost Generation and Evaluation of Dictionary Example Sentences
von: Cai, Bill, et al.
Veröffentlicht: (2024)
von: Cai, Bill, et al.
Veröffentlicht: (2024)
CoAM: Corpus of All-Type Multiword Expressions
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
Hyper-CL: Conditioning Sentence Representations with Hypernetworks
von: Yoo, Young Hyun, et al.
Veröffentlicht: (2024)
von: Yoo, Young Hyun, et al.
Veröffentlicht: (2024)
Expect the unexpected: Harnessing Sentence Completion for Sarcasm Detection
von: Joshi, Aditya, et al.
Veröffentlicht: (2017)
von: Joshi, Aditya, et al.
Veröffentlicht: (2017)
ImpScore: A Learnable Metric For Quantifying The Implicitness Level of Sentence
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
Exploiting Domain-Specific Parallel Data on Multilingual Language Models for Low-resource Language Translation
von: Ranathungaa, Surangika, et al.
Veröffentlicht: (2024)
von: Ranathungaa, Surangika, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Less is More: The Effectiveness of Compact Typological Language Representations
von: Ng, York Hay, et al.
Veröffentlicht: (2025) -
\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding
von: Min, Junghyun, et al.
Veröffentlicht: (2025) -
Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+
von: Shipton, Mason, et al.
Veröffentlicht: (2025) -
Improving Explainability of Sentence-level Metrics via Edit-level Attribution for Grammatical Error Correction
von: Goto, Takumi, et al.
Veröffentlicht: (2024) -
Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+
von: Ng, York Hay, et al.
Veröffentlicht: (2025)