TranslateGemma Technical Report
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Finkelstein, Mara, Caswell, Isaac, Domhan, Tobias, Peter, Jan-Thorsten, Juraska, Juraj, Riley, Parker, Deutsch, Daniel, Kovacs, Geza, Dilanni, Cole, Cherry, Colin, Briakou, Eleftheria, Nielsen, Elizabeth, Luo, Jiaming, Black, Kat, Mullins, Ryan, Agrawal, Sweta, Xu, Wenda, Kats, Erin, Jaskiewicz, Stephane, Freitag, Markus, Vilar, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Jack of All Trades to Master of One: Specializing LLM-based Autoraters to a Test Set
von: Finkelstein, Mara, et al.
Veröffentlicht: (2024)
von: Finkelstein, Mara, et al.
Veröffentlicht: (2024)
MetricX-25 and GemSpanEval: Google Translate Submissions to the WMT25 Evaluation Shared Task
von: Juraska, Juraj, et al.
Veröffentlicht: (2025)
von: Juraska, Juraj, et al.
Veröffentlicht: (2025)
Generating Difficult-to-Translate Texts
von: Zouhar, Vilém, et al.
Veröffentlicht: (2025)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2025)
MetricX-24: The Google Submission to the WMT 2024 Metrics Shared Task
von: Juraska, Juraj, et al.
Veröffentlicht: (2024)
von: Juraska, Juraj, et al.
Veröffentlicht: (2024)
WMT24++: Expanding the Language Coverage of WMT24 to 55 Languages & Dialects
von: Deutsch, Daniel, et al.
Veröffentlicht: (2025)
von: Deutsch, Daniel, et al.
Veröffentlicht: (2025)
MQM Re-Annotation: A Technique for Collaborative Evaluation of Machine Translation
von: Riley, Parker, et al.
Veröffentlicht: (2025)
von: Riley, Parker, et al.
Veröffentlicht: (2025)
Translating Step-by-Step: Decomposing the Translation Process for Improved Translation Quality of Long-Form Texts
von: Briakou, Eleftheria, et al.
Veröffentlicht: (2024)
von: Briakou, Eleftheria, et al.
Veröffentlicht: (2024)
On the Implications of Verbose LLM Outputs: A Case Study in Translation Evaluation
von: Briakou, Eleftheria, et al.
Veröffentlicht: (2024)
von: Briakou, Eleftheria, et al.
Veröffentlicht: (2024)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
von: Shayegh, Behzad, et al.
Veröffentlicht: (2025)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2025)
Overestimation in LLM Evaluation: A Controlled Large-Scale Study on Data Contamination's Impact on Machine Translation
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2025)
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2025)
Mitigating Metric Bias in Minimum Bayes Risk Decoding
von: Kovacs, Geza, et al.
Veröffentlicht: (2024)
von: Kovacs, Geza, et al.
Veröffentlicht: (2024)
LLMRefine: Pinpointing and Refining Large Language Models via Fine-Grained Actionable Feedback
von: Xu, Wenda, et al.
Veröffentlicht: (2023)
von: Xu, Wenda, et al.
Veröffentlicht: (2023)
Rethinking Cross-lingual Alignment: Balancing Transfer and Cultural Erasure in Multilingual LLMs
von: Han, HyoJung, et al.
Veröffentlicht: (2025)
von: Han, HyoJung, et al.
Veröffentlicht: (2025)
Leveraging Domain Knowledge at Inference Time for LLM Translation: Retrieval versus Generation
von: Li, Bryan, et al.
Veröffentlicht: (2025)
von: Li, Bryan, et al.
Veröffentlicht: (2025)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
Introducing the NewsPaLM MBR and QE Dataset: LLM-Generated High-Quality Parallel Data Outperforms Traditional Web-Crawled Data
von: Finkelstein, Mara, et al.
Veröffentlicht: (2024)
von: Finkelstein, Mara, et al.
Veröffentlicht: (2024)
Don't Throw Away Data: Better Sequence Knowledge Distillation
von: Wang, Jun, et al.
Veröffentlicht: (2024)
von: Wang, Jun, et al.
Veröffentlicht: (2024)
Distribution-Calibrated Inference time compute for Thinking LLM-as-a-Judge
von: Dadkhahi, Hamid, et al.
Veröffentlicht: (2025)
von: Dadkhahi, Hamid, et al.
Veröffentlicht: (2025)
Mind the Gap... or Not? How Translation Errors and Evaluation Details Skew Multilingual Results
von: Peter, Jan-Thorsten, et al.
Veröffentlicht: (2025)
von: Peter, Jan-Thorsten, et al.
Veröffentlicht: (2025)
Pulsation-driven helium transport as a potential source of the Blazhko effect
von: Kovacs, Geza
Veröffentlicht: (2026)
von: Kovacs, Geza
Veröffentlicht: (2026)
Digging Deeper for RR Lyrae Stars with Low Modulation Amplitudes
von: Kovacs, Geza
Veröffentlicht: (2025)
von: Kovacs, Geza
Veröffentlicht: (2025)
Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion Algorithms
von: Trabelsi, Firas, et al.
Veröffentlicht: (2024)
von: Trabelsi, Firas, et al.
Veröffentlicht: (2024)
Enhancing Human Evaluation in Machine Translation with Comparative Judgment
von: Song, Yixiao, et al.
Veröffentlicht: (2025)
von: Song, Yixiao, et al.
Veröffentlicht: (2025)
Quality-Aware Translation Models: Efficient Generation and Quality Estimation in a Single Model
von: Tomani, Christian, et al.
Veröffentlicht: (2023)
von: Tomani, Christian, et al.
Veröffentlicht: (2023)
SSA-COMET: Do LLMs Outperform Learned Metrics in Evaluating MT for Under-Resourced African Languages?
von: Li, Senyu, et al.
Veröffentlicht: (2025)
von: Li, Senyu, et al.
Veröffentlicht: (2025)
ShieldGemma: Generative AI Content Moderation Based on Gemma
von: Zeng, Wenjun, et al.
Veröffentlicht: (2024)
von: Zeng, Wenjun, et al.
Veröffentlicht: (2024)
Quark mass corrections in di-Higgs production amplitude at high-energy
von: Jaskiewicz, Sebastian
Veröffentlicht: (2025)
von: Jaskiewicz, Sebastian
Veröffentlicht: (2025)
Lightcone expansion beyond leading power
von: Jaskiewicz, Sebastian
Veröffentlicht: (2024)
von: Jaskiewicz, Sebastian
Veröffentlicht: (2024)
The last stage of development: The restructuring and plasticity of the cortex during adolescence especially at puberty
von: Janice M. Juraska
Veröffentlicht: (2024)
von: Janice M. Juraska
Veröffentlicht: (2024)
Same evaluation, more tokens: On the effect of input length for machine translation evaluation using Large Language Models
von: Domhan, Tobias, et al.
Veröffentlicht: (2025)
von: Domhan, Tobias, et al.
Veröffentlicht: (2025)
Finding Replicable Human Evaluations via Stable Ranking Probability
von: Riley, Parker, et al.
Veröffentlicht: (2024)
von: Riley, Parker, et al.
Veröffentlicht: (2024)
T5Gemma 2: Seeing, Reading, and Understanding Longer
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
Secondary eclipses of two brown dwarfs in the K2 fields: detection by multiple dataset merging
von: Kovacs, Geza, et al.
Veröffentlicht: (2025)
von: Kovacs, Geza, et al.
Veröffentlicht: (2025)
Beyond Human-Only: Evaluating Human-Machine Collaboration for Collecting High-Quality Translation Data
von: Liu, Zhongtao, et al.
Veröffentlicht: (2024)
von: Liu, Zhongtao, et al.
Veröffentlicht: (2024)
Stochastic dynamic programming under recursive Epstein-Zin preferences
von: Jaśkiewicz, Anna, et al.
Veröffentlicht: (2024)
von: Jaśkiewicz, Anna, et al.
Veröffentlicht: (2024)
Time-consistency in the mean-variance problem: A new perspective
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
Markov Decision Processes with Risk-Sensitive Criteria: An Overview
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
Gemma 3 Technical Report
von: Gemma Team, et al.
Veröffentlicht: (2025)
von: Gemma Team, et al.
Veröffentlicht: (2025)
Urgent Archives
von: Caswell, Michelle
Veröffentlicht: (2025)
von: Caswell, Michelle
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Jack of All Trades to Master of One: Specializing LLM-based Autoraters to a Test Set
von: Finkelstein, Mara, et al.
Veröffentlicht: (2024) -
MetricX-25 and GemSpanEval: Google Translate Submissions to the WMT25 Evaluation Shared Task
von: Juraska, Juraj, et al.
Veröffentlicht: (2025) -
Generating Difficult-to-Translate Texts
von: Zouhar, Vilém, et al.
Veröffentlicht: (2025) -
MetricX-24: The Google Submission to the WMT 2024 Metrics Shared Task
von: Juraska, Juraj, et al.
Veröffentlicht: (2024) -
WMT24++: Expanding the Language Coverage of WMT24 to 55 Languages & Dialects
von: Deutsch, Daniel, et al.
Veröffentlicht: (2025)