How Effective are State Space Models for Machine Translation?
Fuente:
arXiv
Guardado en:
| Autores principales: | Pitorro, Hugo, Vasylenko, Pavlo, Treviso, Marcos, Martins, André F. T. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Long-Context Generalization with Sparse Attention
por: Vasylenko, Pavlo, et al.
Publicado: (2025)
por: Vasylenko, Pavlo, et al.
Publicado: (2025)
LaTIM: Measuring Latent Token-to-Token Interactions in Mamba Models
por: Pitorro, Hugo, et al.
Publicado: (2025)
por: Pitorro, Hugo, et al.
Publicado: (2025)
AdaSplash-2: Faster Differentiable Sparse Attention
por: Gonçalves, Nuno, et al.
Publicado: (2026)
por: Gonçalves, Nuno, et al.
Publicado: (2026)
AdaSplash: Adaptive Sparse Flash Attention
por: Gonçalves, Nuno, et al.
Publicado: (2025)
por: Gonçalves, Nuno, et al.
Publicado: (2025)
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors
por: Treviso, Marcos, et al.
Publicado: (2024)
por: Treviso, Marcos, et al.
Publicado: (2024)
Aligning Neural Machine Translation Models: Human Feedback in Training and Inference
por: Ramos, Miguel Moura, et al.
Publicado: (2023)
por: Ramos, Miguel Moura, et al.
Publicado: (2023)
Multilingual Contextualization of Large Language Models for Document-Level Machine Translation
por: Ramos, Miguel Moura, et al.
Publicado: (2025)
por: Ramos, Miguel Moura, et al.
Publicado: (2025)
Analyzing Context Contributions in LLM-based Machine Translation
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
EntmaxKV: Support-Aware Decoding for Entmax Attention
por: Duarte, Gonçalo, et al.
Publicado: (2026)
por: Duarte, Gonçalo, et al.
Publicado: (2026)
Watching the Watchers: Exposing Gender Disparities in Machine Translation Quality Estimation
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
por: Zaranis, Emmanouil, et al.
Publicado: (2024)
Did Translation Models Get More Robust Without Anyone Even Noticing?
por: Peters, Ben, et al.
Publicado: (2024)
por: Peters, Ben, et al.
Publicado: (2024)
Adding Chocolate to Mint: Mitigating Metric Interference in Machine Translation
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
Self-Modifying State Modeling for Simultaneous Machine Translation
por: Yu, Donglei, et al.
Publicado: (2024)
por: Yu, Donglei, et al.
Publicado: (2024)
Fine-Grained Reward Optimization for Machine Translation using Error Severity Mappings
por: Ramos, Miguel Moura, et al.
Publicado: (2024)
por: Ramos, Miguel Moura, et al.
Publicado: (2024)
DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention
por: Huang, Yuxiang, et al.
Publicado: (2026)
por: Huang, Yuxiang, et al.
Publicado: (2026)
Different Speech Translation Models Encode and Translate Speaker Gender Differently
por: Fucci, Dennis, et al.
Publicado: (2025)
por: Fucci, Dennis, et al.
Publicado: (2025)
Can Automatic Metrics Assess High-Quality Translations?
por: Agrawal, Sweta, et al.
Publicado: (2024)
por: Agrawal, Sweta, et al.
Publicado: (2024)
QUEST: Quality-Aware Metropolis-Hastings Sampling for Machine Translation
por: Faria, Gonçalo R. A., et al.
Publicado: (2024)
por: Faria, Gonçalo R. A., et al.
Publicado: (2024)
Is Context Helpful for Chat Translation Evaluation?
por: Agrawal, Sweta, et al.
Publicado: (2024)
por: Agrawal, Sweta, et al.
Publicado: (2024)
How Well Do Large Reasoning Models Translate? A Comprehensive Evaluation for Multi-Domain Machine Translation
por: Ye, Yongshi, et al.
Publicado: (2025)
por: Ye, Yongshi, et al.
Publicado: (2025)
How Important is `Perfect' English for Machine Translation Prompts?
por: Schmidtová, Patrícia, et al.
Publicado: (2025)
por: Schmidtová, Patrícia, et al.
Publicado: (2025)
Toward Machine Translation Literacy: How Lay Users Perceive and Rely on Imperfect Translations
por: Xiao, Yimin, et al.
Publicado: (2025)
por: Xiao, Yimin, et al.
Publicado: (2025)
Modeling User Preferences with Automatic Metrics: Creating a High-Quality Preference Dataset for Machine Translation
por: Agrawal, Sweta, et al.
Publicado: (2024)
por: Agrawal, Sweta, et al.
Publicado: (2024)
A Context-aware Framework for Translation-mediated Conversations
por: Pombal, José, et al.
Publicado: (2024)
por: Pombal, José, et al.
Publicado: (2024)
Effective Self-Mining of In-Context Examples for Unsupervised Machine Translation with LLMs
por: Mekki, Abdellah El, et al.
Publicado: (2024)
por: Mekki, Abdellah El, et al.
Publicado: (2024)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
por: Belouadi, Jonas, et al.
Publicado: (2022)
por: Belouadi, Jonas, et al.
Publicado: (2022)
Span-Level Machine Translation Meta-Evaluation
por: Perrella, Stefano, et al.
Publicado: (2026)
por: Perrella, Stefano, et al.
Publicado: (2026)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
por: Wastl, Michelle, et al.
Publicado: (2024)
por: Wastl, Michelle, et al.
Publicado: (2024)
Creativity Bias: How Machine Evaluation Struggles with Creativity in Literary Translations
por: Gerrits, Kyo, et al.
Publicado: (2026)
por: Gerrits, Kyo, et al.
Publicado: (2026)
Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation
por: Qu, Zhi, et al.
Publicado: (2025)
por: Qu, Zhi, et al.
Publicado: (2025)
Are Large Language Models State-of-the-art Quality Estimators for Machine Translation of User-generated Content?
por: Qian, Shenbin, et al.
Publicado: (2024)
por: Qian, Shenbin, et al.
Publicado: (2024)
Lost in the Source Language: How Large Language Models Evaluate the Quality of Machine Translation
por: Huang, Xu, et al.
Publicado: (2024)
por: Huang, Xu, et al.
Publicado: (2024)
Large Language Models "Ad Referendum": How Good Are They at Machine Translation in the Legal Domain?
por: Briva-Iglesias, Vicent, et al.
Publicado: (2024)
por: Briva-Iglesias, Vicent, et al.
Publicado: (2024)
The Effectiveness of Morphology-aware Segmentation in Low-Resource Neural Machine Translation
por: Sälevä, Jonne, et al.
Publicado: (2021)
por: Sälevä, Jonne, et al.
Publicado: (2021)
Simultaneous Machine Translation with Large Language Models
por: Wang, Minghan, et al.
Publicado: (2023)
por: Wang, Minghan, et al.
Publicado: (2023)
Translate Smart, not Hard: Cascaded Translation Systems with Quality-Aware Deferral
por: Farinhas, António, et al.
Publicado: (2025)
por: Farinhas, António, et al.
Publicado: (2025)
From TOWER to SPIRE: Adding the Speech Modality to a Translation-Specialist LLM
por: Ambilduke, Kshitij, et al.
Publicado: (2025)
por: Ambilduke, Kshitij, et al.
Publicado: (2025)
An Interdisciplinary Approach to Human-Centered Machine Translation
por: Carpuat, Marine, et al.
Publicado: (2025)
por: Carpuat, Marine, et al.
Publicado: (2025)
Tower: An Open Multilingual Large Language Model for Translation-Related Tasks
por: Alves, Duarte M., et al.
Publicado: (2024)
por: Alves, Duarte M., et al.
Publicado: (2024)
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models
por: Xu, Haoran, et al.
Publicado: (2023)
por: Xu, Haoran, et al.
Publicado: (2023)
Ejemplares similares
-
Long-Context Generalization with Sparse Attention
por: Vasylenko, Pavlo, et al.
Publicado: (2025) -
LaTIM: Measuring Latent Token-to-Token Interactions in Mamba Models
por: Pitorro, Hugo, et al.
Publicado: (2025) -
AdaSplash-2: Faster Differentiable Sparse Attention
por: Gonçalves, Nuno, et al.
Publicado: (2026) -
AdaSplash: Adaptive Sparse Flash Attention
por: Gonçalves, Nuno, et al.
Publicado: (2025) -
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors
por: Treviso, Marcos, et al.
Publicado: (2024)