Neuron Specialization: Leveraging intrinsic task modularity for multilingual machine translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tan, Shaomu, Wu, Di, Monz, Christof |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling
por: Tan, Shaomu, et al.
Publicado: (2025)
por: Tan, Shaomu, et al.
Publicado: (2025)
How Far Can 100 Samples Go? Unlocking Overall Zero-Shot Multilingual Translation via Tiny Multi-Parallel Data
por: Wu, Di, et al.
Publicado: (2024)
por: Wu, Di, et al.
Publicado: (2024)
Remedy-R: Generative Reasoning for Machine Translation Evaluation without Error Annotations
por: Tan, Shaomu, et al.
Publicado: (2025)
por: Tan, Shaomu, et al.
Publicado: (2025)
Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
por: Wu, Di, et al.
Publicado: (2023)
por: Wu, Di, et al.
Publicado: (2023)
Please Translate Again: Two Simple Experiments on Whether Human-Like Reasoning Helps Translation
por: Wu, Di, et al.
Publicado: (2025)
por: Wu, Di, et al.
Publicado: (2025)
Calibrating Translation Decoding with Quality Estimation on LLMs
por: Wu, Di, et al.
Publicado: (2025)
por: Wu, Di, et al.
Publicado: (2025)
How to Learn in a Noisy World? Self-Correcting the Real-World Data Noise in Machine Translation
por: Meng, Yan, et al.
Publicado: (2024)
por: Meng, Yan, et al.
Publicado: (2024)
The Effect of Language Diversity When Fine-Tuning Large Language Models for Translation
por: Stap, David, et al.
Publicado: (2025)
por: Stap, David, et al.
Publicado: (2025)
Is It a Free Lunch for Removing Outliers during Pretraining?
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
Can LLMs Really Learn to Translate a Low-Resource Language from One Grammar Book?
por: Aycock, Seth, et al.
Publicado: (2024)
por: Aycock, Seth, et al.
Publicado: (2024)
What Does Neuro Mean to Cardio? Investigating the Role of Clinical Specialty Data in Medical LLMs
por: Yan, Xinlan, et al.
Publicado: (2025)
por: Yan, Xinlan, et al.
Publicado: (2025)
Analyzing the Evaluation of Cross-Lingual Knowledge Transfer in Multilingual Language Models
por: Rajaee, Sara, et al.
Publicado: (2024)
por: Rajaee, Sara, et al.
Publicado: (2024)
Disentangling the Roles of Target-Side Transfer and Regularization in Multilingual Machine Translation
por: Meng, Yan, et al.
Publicado: (2024)
por: Meng, Yan, et al.
Publicado: (2024)
3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
Investigating Test-Time Scaling with Reranking for Machine Translation
por: Tan, Shaomu, et al.
Publicado: (2025)
por: Tan, Shaomu, et al.
Publicado: (2025)
Do Language Models Reason Across Languages?
por: Meng, Yan, et al.
Publicado: (2026)
por: Meng, Yan, et al.
Publicado: (2026)
IKUN for WMT24 General MT Task: LLMs Are here for Multilingual Machine Translation
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
When Contextual Inference Fails: Cancelability in Interactive Instruction Following
por: Bila, Natalia, et al.
Publicado: (2026)
por: Bila, Natalia, et al.
Publicado: (2026)
Communicating with Speakers and Listeners of Different Pragmatic Levels
por: Naszadi, Kata, et al.
Publicado: (2024)
por: Naszadi, Kata, et al.
Publicado: (2024)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
por: Choenni, Rochelle, et al.
Publicado: (2024)
por: Choenni, Rochelle, et al.
Publicado: (2024)
ApiQ: Finetuning of 2-Bit Quantized Large Language Model
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
por: Rajaee, Sara, et al.
Publicado: (2025)
por: Rajaee, Sara, et al.
Publicado: (2025)
The Fine-Tuning Paradox: Boosting Translation Quality Without Sacrificing LLM Abilities
por: Stap, David, et al.
Publicado: (2024)
por: Stap, David, et al.
Publicado: (2024)
On the Limits of Model Merging for Multilinguality in Pre-Training
por: Aycock, Seth, et al.
Publicado: (2026)
por: Aycock, Seth, et al.
Publicado: (2026)
SENSE models: an open source solution for multilingual and multimodal semantic-based tasks
por: Mdhaffar, Salima, et al.
Publicado: (2025)
por: Mdhaffar, Salima, et al.
Publicado: (2025)
Are AI agents the new machine translation frontier? Challenges and opportunities of single- and multi-agent systems for multilingual digital communication
por: Briva-Iglesias, Vicent
Publicado: (2025)
por: Briva-Iglesias, Vicent
Publicado: (2025)
Optimizing example selection for retrieval-augmented machine translation with translation memories
por: Bouthors, Maxime, et al.
Publicado: (2024)
por: Bouthors, Maxime, et al.
Publicado: (2024)
Escaping the sentence-level paradigm in machine translation
por: Post, Matt, et al.
Publicado: (2023)
por: Post, Matt, et al.
Publicado: (2023)
Self-Hinting Language Models Enhance Reinforcement Learning
por: Liao, Baohao, et al.
Publicado: (2026)
por: Liao, Baohao, et al.
Publicado: (2026)
What Does LLM Refinement Actually Improve? A Systematic Study on Document-Level Literary Translation
por: Tan, Shaomu, et al.
Publicado: (2026)
por: Tan, Shaomu, et al.
Publicado: (2026)
Lost in translation: using global fact-checks to measure multilingual misinformation prevalence, spread, and evolution
por: Quelle, Dorian, et al.
Publicado: (2023)
por: Quelle, Dorian, et al.
Publicado: (2023)
The first open machine translation system for the Chechen language
por: Umishov, Abu-Viskhan A., et al.
Publicado: (2025)
por: Umishov, Abu-Viskhan A., et al.
Publicado: (2025)
Contextual effects of sentiment deployment in human and machine translation
por: Comstock, Lindy, et al.
Publicado: (2025)
por: Comstock, Lindy, et al.
Publicado: (2025)
The SIFo Benchmark: Investigating the Sequential Instruction Following Ability of Large Language Models
por: Chen, Xinyi, et al.
Publicado: (2024)
por: Chen, Xinyi, et al.
Publicado: (2024)
Neural machine translation system for Lezgian, Russian and Azerbaijani languages
por: Asvarov, Alidar, et al.
Publicado: (2024)
por: Asvarov, Alidar, et al.
Publicado: (2024)
Disentangling meaning from language in LLM-based machine translation
por: Lasnier, Théo, et al.
Publicado: (2026)
por: Lasnier, Théo, et al.
Publicado: (2026)
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning
por: Liao, Baohao, et al.
Publicado: (2025)
por: Liao, Baohao, et al.
Publicado: (2025)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
por: Troshin, Sergey, et al.
Publicado: (2025)
por: Troshin, Sergey, et al.
Publicado: (2025)
Can professional translators identify machine-generated text?
por: Farrell, Michael
Publicado: (2026)
por: Farrell, Michael
Publicado: (2026)
Language translation, and change of accent for speech-to-speech task using diffusion model
por: Mishra, Abhishek, et al.
Publicado: (2025)
por: Mishra, Abhishek, et al.
Publicado: (2025)
Ejemplares similares
-
Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling
por: Tan, Shaomu, et al.
Publicado: (2025) -
How Far Can 100 Samples Go? Unlocking Overall Zero-Shot Multilingual Translation via Tiny Multi-Parallel Data
por: Wu, Di, et al.
Publicado: (2024) -
Remedy-R: Generative Reasoning for Machine Translation Evaluation without Error Annotations
por: Tan, Shaomu, et al.
Publicado: (2025) -
Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
por: Wu, Di, et al.
Publicado: (2023) -
Please Translate Again: Two Simple Experiments on Whether Human-Like Reasoning Helps Translation
por: Wu, Di, et al.
Publicado: (2025)