MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
Fuente:
arXiv
Guardado en:
| Autores principales: | Moosa, Ibraheem Muhammad, Zhang, Rui, Yin, Wenpeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Low-resource neural machine translation with morphological modeling
por: Nzeyimana, Antoine
Publicado: (2024)
por: Nzeyimana, Antoine
Publicado: (2024)
Omnilingual MT: Machine Translation for 1,600 Languages
por: Omnilingual MT Team, et al.
Publicado: (2026)
por: Omnilingual MT Team, et al.
Publicado: (2026)
LLMs Are Not Scorers: Rethinking MT Evaluation with Generation-Based Methods
por: Cui, Hyang
Publicado: (2025)
por: Cui, Hyang
Publicado: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models
por: Fuad, Kazi Ahmed Asif, et al.
Publicado: (2024)
por: Fuad, Kazi Ahmed Asif, et al.
Publicado: (2024)
When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models
por: Bae, Ji Ho
Publicado: (2026)
por: Bae, Ji Ho
Publicado: (2026)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
por: Stewart, Ian, et al.
Publicado: (2024)
por: Stewart, Ian, et al.
Publicado: (2024)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
por: Evelo, Bart, et al.
Publicado: (2026)
por: Evelo, Bart, et al.
Publicado: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025)
por: Peters, Sydney, et al.
Publicado: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
Strategy Adaptation in Large Language Model Werewolf Agents
por: Nakamori, Fuya, et al.
Publicado: (2025)
por: Nakamori, Fuya, et al.
Publicado: (2025)
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
por: Ahmad, Sarfraz, et al.
Publicado: (2025)
por: Ahmad, Sarfraz, et al.
Publicado: (2025)
EduGuardBench: A Holistic Benchmark for Evaluating the Pedagogical Fidelity and Adversarial Safety of LLMs as Simulated Teachers
por: Jiang, Yilin, et al.
Publicado: (2025)
por: Jiang, Yilin, et al.
Publicado: (2025)
Large Language Model (LLM) Bias Index -- LLMBI
por: Oketunji, Abiodun Finbarrs, et al.
Publicado: (2023)
por: Oketunji, Abiodun Finbarrs, et al.
Publicado: (2023)
HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agents
por: Jiang, Yilin, et al.
Publicado: (2026)
por: Jiang, Yilin, et al.
Publicado: (2026)
Yes-MT's Submission to the Low-Resource Indic Language Translation Shared Task in WMT 2024
por: Bhaskar, Yash, et al.
Publicado: (2025)
por: Bhaskar, Yash, et al.
Publicado: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
por: Smădu, Răzvan-Alexandru, et al.
Publicado: (2025)
por: Smădu, Răzvan-Alexandru, et al.
Publicado: (2025)
Evaluating an evidence-guided reinforcement learning framework in aligning light-parameter large language models with decision-making cognition in psychiatric clinical reasoning
por: Lin, Xinxin, et al.
Publicado: (2026)
por: Lin, Xinxin, et al.
Publicado: (2026)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
por: Wang, Xintao, et al.
Publicado: (2026)
por: Wang, Xintao, et al.
Publicado: (2026)
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
por: Qi, Jinhu, et al.
Publicado: (2024)
por: Qi, Jinhu, et al.
Publicado: (2024)
CauESC: A Causal Aware Model for Emotional Support Conversation
por: Chen, Wei, et al.
Publicado: (2024)
por: Chen, Wei, et al.
Publicado: (2024)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
por: Liu, Han, et al.
Publicado: (2026)
por: Liu, Han, et al.
Publicado: (2026)
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
por: Xiaohui, Han, et al.
Publicado: (2025)
por: Xiaohui, Han, et al.
Publicado: (2025)
Graphemic Normalization of the Perso-Arabic Script
por: Doctor, Raiomond, et al.
Publicado: (2022)
por: Doctor, Raiomond, et al.
Publicado: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
por: Gutkin, Alexander, et al.
Publicado: (2023)
por: Gutkin, Alexander, et al.
Publicado: (2023)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
por: Liu, Han, et al.
Publicado: (2026)
por: Liu, Han, et al.
Publicado: (2026)
SemEval-2026 Task 3: Dimensional Aspect-Based Sentiment Analysis (DimABSA)
por: Yu, Liang-Chih, et al.
Publicado: (2026)
por: Yu, Liang-Chih, et al.
Publicado: (2026)
DimABSA: Building Multilingual and Multidomain Datasets for Dimensional Aspect-Based Sentiment Analysis
por: Lee, Lung-Hao, et al.
Publicado: (2026)
por: Lee, Lung-Hao, et al.
Publicado: (2026)
LinkNER: Linking Local Named Entity Recognition Models to Large Language Models using Uncertainty
por: Zhang, Zhen, et al.
Publicado: (2024)
por: Zhang, Zhen, et al.
Publicado: (2024)
DimStance: Multilingual Datasets for Dimensional Stance Analysis
por: Becker, Jonas, et al.
Publicado: (2026)
por: Becker, Jonas, et al.
Publicado: (2026)
AsyncTLS: Efficient Generative LLM Inference with Asynchronous Two-level Sparse Attention
por: Hu, Yuxuan, et al.
Publicado: (2026)
por: Hu, Yuxuan, et al.
Publicado: (2026)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
por: Wang, Renxi, et al.
Publicado: (2024)
por: Wang, Renxi, et al.
Publicado: (2024)
Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
por: Ma, Longxuan, et al.
Publicado: (2024)
por: Ma, Longxuan, et al.
Publicado: (2024)
Schema as Parameterized Tools for Universal Information Extraction
por: Liang, Sheng, et al.
Publicado: (2025)
por: Liang, Sheng, et al.
Publicado: (2025)
Effective and Efficient Schema-aware Information Extraction Using On-Device Large Language Models
por: Wen, Zhihao, et al.
Publicado: (2025)
por: Wen, Zhihao, et al.
Publicado: (2025)
Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
por: Ma, Longxuan, et al.
Publicado: (2024)
por: Ma, Longxuan, et al.
Publicado: (2024)
Towards Greater Leverage: Scaling Laws for Efficient Mixture-of-Experts Language Models
por: Tian, Changxin, et al.
Publicado: (2025)
por: Tian, Changxin, et al.
Publicado: (2025)
I run as fast as a rabbit, can you? A Multilingual Simile Dialogue Dataset
por: Ma, Longxuan, et al.
Publicado: (2023)
por: Ma, Longxuan, et al.
Publicado: (2023)
Ejemplares similares
-
Low-resource neural machine translation with morphological modeling
por: Nzeyimana, Antoine
Publicado: (2024) -
Omnilingual MT: Machine Translation for 1,600 Languages
por: Omnilingual MT Team, et al.
Publicado: (2026) -
LLMs Are Not Scorers: Rethinking MT Evaluation with Generation-Based Methods
por: Cui, Hyang
Publicado: (2025) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025) -
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models
por: Fuad, Kazi Ahmed Asif, et al.
Publicado: (2024)