Are Character-level Translations Worth the Wait? Comparing ByT5 and mT5 for Machine Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Edman, Lukas, Sarti, Gabriele, Toral, Antonio, van Noord, Gertjan, Bisazza, Arianna |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
by: Dang, Thao Anh, et al.
Published: (2024)
by: Dang, Thao Anh, et al.
Published: (2024)
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
by: Sarti, Gabriele, et al.
Published: (2023)
by: Sarti, Gabriele, et al.
Published: (2023)
Steering Large Language Models for Machine Translation Personalization
by: Scalena, Daniel, et al.
Published: (2025)
by: Scalena, Daniel, et al.
Published: (2025)
Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation
by: Lai, Huiyuan, et al.
Published: (2024)
by: Lai, Huiyuan, et al.
Published: (2024)
Towards Tailored Recovery of Lexical Diversity in Literary Machine Translation
by: Ploeger, Esther, et al.
Published: (2024)
by: Ploeger, Esther, et al.
Published: (2024)
Assessing the Impact of Typological Features on Multilingual Machine Translation in the Age of Large Language Models
by: Hirak, Vitalii, et al.
Published: (2026)
by: Hirak, Vitalii, et al.
Published: (2026)
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
by: Sarti, Gabriele, et al.
Published: (2024)
by: Sarti, Gabriele, et al.
Published: (2024)
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation
by: Qi, Jirui, et al.
Published: (2024)
by: Qi, Jirui, et al.
Published: (2024)
Efficacy of ByT5 in Multilingual Translation of Biblical Texts for Underrepresented Languages
by: Aars, Corinne, et al.
Published: (2024)
by: Aars, Corinne, et al.
Published: (2024)
A Primer on the Inner Workings of Transformer-based Language Models
by: Ferrando, Javier, et al.
Published: (2024)
by: Ferrando, Javier, et al.
Published: (2024)
Creativity Bias: How Machine Evaluation Struggles with Creativity in Literary Translations
by: Gerrits, Kyo, et al.
Published: (2026)
by: Gerrits, Kyo, et al.
Published: (2026)
QE4PE: Word-level Quality Estimation for Human Post-Editing
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
Towards Santali Linguistic Inclusion: Building the First Santali-to-English Translation Model using mT5 Transformer and Data Augmentation
by: Billah, Syed Mohammed Mostaque, et al.
Published: (2024)
by: Billah, Syed Mohammed Mostaque, et al.
Published: (2024)
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
by: Sarti, Gabriele, et al.
Published: (2022)
by: Sarti, Gabriele, et al.
Published: (2022)
MBBQ: A Dataset for Cross-Lingual Comparison of Stereotypes in Generative LLMs
by: Neplenbroek, Vera, et al.
Published: (2024)
by: Neplenbroek, Vera, et al.
Published: (2024)
Reading Between the Prompts: How Stereotypes Shape LLM's Implicit Personalization
by: Neplenbroek, Vera, et al.
Published: (2025)
by: Neplenbroek, Vera, et al.
Published: (2025)
Cross-Lingual Transfer of Debiasing and Detoxification in Multilingual LLMs: An Extensive Investigation
by: Neplenbroek, Vera, et al.
Published: (2024)
by: Neplenbroek, Vera, et al.
Published: (2024)
Simulating the Emergence of Differential Case Marking with Communicating Neural-Network Agents
by: Lian, Yuchen, et al.
Published: (2025)
by: Lian, Yuchen, et al.
Published: (2025)
NeLLCom-X: A Comprehensive Neural-Agent Framework to Simulate Language Learning and Group Communication
by: Lian, Yuchen, et al.
Published: (2024)
by: Lian, Yuchen, et al.
Published: (2024)
Mask and You Shall Receive: Optimizing Masked Language Modeling For Pretraining BabyLMs
by: Edman, Lukas, et al.
Published: (2025)
by: Edman, Lukas, et al.
Published: (2025)
Improving Language and Modality Transfer in Translation by Character-level Modeling
by: Tsiamas, Ioannis, et al.
Published: (2025)
by: Tsiamas, Ioannis, et al.
Published: (2025)
On the Consistency of Multilingual Context Utilization in Retrieval-Augmented Generation
by: Qi, Jirui, et al.
Published: (2025)
by: Qi, Jirui, et al.
Published: (2025)
PMB5: Gaining More Insight into Neural Semantic Parsing with Challenging Benchmarks
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
An Empirical Study on Chinese Character Decomposition in Multiword Expression-Aware Neural Machine Translation
by: Han, Lifeng, et al.
Published: (2025)
by: Han, Lifeng, et al.
Published: (2025)
Medical mT5: An Open-Source Multilingual Text-to-Text LLM for The Medical Domain
by: García-Ferrero, Iker, et al.
Published: (2024)
by: García-Ferrero, Iker, et al.
Published: (2024)
Enhancing Human Evaluation in Machine Translation with Comparative Judgment
by: Song, Yixiao, et al.
Published: (2025)
by: Song, Yixiao, et al.
Published: (2025)
Comparative Evaluation of Machine Translation Systems on Images with Text
by: Puchol, Blai, et al.
Published: (2026)
by: Puchol, Blai, et al.
Published: (2026)
Cross-Lingual Consistency of Factual Knowledge in Multilingual Language Models
by: Qi, Jirui, et al.
Published: (2023)
by: Qi, Jirui, et al.
Published: (2023)
Is Child-Directed Language Optimized for Word Learning? A Computational Study of Verb Meaning Acquisition
by: Padovani, Francesca, et al.
Published: (2026)
by: Padovani, Francesca, et al.
Published: (2026)
Child-Directed Language Does Not Consistently Boost Syntax Learning in Language Models
by: Padovani, Francesca, et al.
Published: (2025)
by: Padovani, Francesca, et al.
Published: (2025)
TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs
by: Başar, Ezgi, et al.
Published: (2025)
by: Başar, Ezgi, et al.
Published: (2025)
MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs
by: Jumelet, Jaap, et al.
Published: (2025)
by: Jumelet, Jaap, et al.
Published: (2025)
CUTE: Measuring LLMs' Understanding of Their Tokens
by: Edman, Lukas, et al.
Published: (2024)
by: Edman, Lukas, et al.
Published: (2024)
EXECUTE: A Multilingual Benchmark for LLM Token Understanding
by: Edman, Lukas, et al.
Published: (2025)
by: Edman, Lukas, et al.
Published: (2025)
Do Language Models Care About Text Quality? Evaluating Web-Crawled Corpora Across 11 Languages
by: van Noord, Rik, et al.
Published: (2024)
by: van Noord, Rik, et al.
Published: (2024)
One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks
by: Nehrdich, Sebastian, et al.
Published: (2024)
by: Nehrdich, Sebastian, et al.
Published: (2024)
Investigating Length Issues in Document-level Machine Translation
by: Peng, Ziqian, et al.
Published: (2024)
by: Peng, Ziqian, et al.
Published: (2024)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
by: Wastl, Michelle, et al.
Published: (2024)
by: Wastl, Michelle, et al.
Published: (2024)
The Impact of Syntactic and Semantic Proximity on Machine Translation with Back-Translation
by: Guerin, Nicolas, et al.
Published: (2024)
by: Guerin, Nicolas, et al.
Published: (2024)
Similar Items
-
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
by: Sarti, Gabriele, et al.
Published: (2025) -
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
by: Dang, Thao Anh, et al.
Published: (2024) -
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
by: Sarti, Gabriele, et al.
Published: (2023) -
Steering Large Language Models for Machine Translation Personalization
by: Scalena, Daniel, et al.
Published: (2025) -
Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation
by: Lai, Huiyuan, et al.
Published: (2024)