Toucan: Many-to-Many Translation for 150 African Language Pairs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Elmadany, AbdelRahim, Adebara, Ife, Abdul-Mageed, Muhammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cheetah: Natural Language Generation for 517 African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2024)
von: Adebara, Ife, et al.
Veröffentlicht: (2024)
Where Are We? Evaluating LLM Performance on African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
AfroScope: A Framework for Studying the Linguistic Landscape of Africa
von: Kwon, Sang Yun, et al.
Veröffentlicht: (2026)
von: Kwon, Sang Yun, et al.
Veröffentlicht: (2026)
Interplay of Machine Translation, Diacritics, and Diacritization
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2024)
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2024)
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
von: Sullivan, Peter, et al.
Veröffentlicht: (2026)
von: Sullivan, Peter, et al.
Veröffentlicht: (2026)
WojoodNER 2024: The Second Arabic Named Entity Recognition Shared Task
von: Jarrar, Mustafa, et al.
Veröffentlicht: (2024)
von: Jarrar, Mustafa, et al.
Veröffentlicht: (2024)
Voice of a Continent: Mapping Africa's Speech Technology Frontier
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
Fumbling in Babel: An Investigation into ChatGPT's Language Identification Ability
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2023)
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2023)
NADI 2024: The Fifth Nuanced Arabic Dialect Identification Shared Task
von: Abdul-Mageed, Muhammad, et al.
Veröffentlicht: (2024)
von: Abdul-Mageed, Muhammad, et al.
Veröffentlicht: (2024)
NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
Effective Self-Mining of In-Context Examples for Unsupervised Machine Translation with LLMs
von: Mekki, Abdellah El, et al.
Veröffentlicht: (2024)
von: Mekki, Abdellah El, et al.
Veröffentlicht: (2024)
Towards Boosting Many-to-Many Multilingual Machine Translation with Large Language Models
von: Gao, Pengzhi, et al.
Veröffentlicht: (2024)
von: Gao, Pengzhi, et al.
Veröffentlicht: (2024)
MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages
von: Du, Yexing, et al.
Veröffentlicht: (2025)
von: Du, Yexing, et al.
Veröffentlicht: (2025)
Making LLMs Better Many-to-Many Speech-to-Text Translators with Curriculum Learning
von: Du, Yexing, et al.
Veröffentlicht: (2024)
von: Du, Yexing, et al.
Veröffentlicht: (2024)
EnAnchored-X2X: English-Anchored Optimization for Many-to-Many Translation
von: Yang, Sen, et al.
Veröffentlicht: (2025)
von: Yang, Sen, et al.
Veröffentlicht: (2025)
Dallah: A Dialect-Aware Multimodal Large Language Model for Arabic
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2024)
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2024)
Translation in the Hands of Many:Centering Lay Users in Machine Translation Interactions
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
An Empirical Study of Many-to-Many Summarization with Large Language Models
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
von: Wang, Jiaan, et al.
Veröffentlicht: (2025)
To Distill or Not to Distill? On the Robustness of Robust Knowledge Distillation
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
Arabic Automatic Story Generation with Large Language Models
von: El-Shangiti, Ahmed Oumar, et al.
Veröffentlicht: (2024)
von: El-Shangiti, Ahmed Oumar, et al.
Veröffentlicht: (2024)
On Barriers to Archival Audio Processing
von: Sullivan, Peter, et al.
Veröffentlicht: (2025)
von: Sullivan, Peter, et al.
Veröffentlicht: (2025)
An Empirical Study of Many-Shot In-Context Learning for Machine Translation of Low-Resource Languages
von: Lu, Yinhan, et al.
Veröffentlicht: (2026)
von: Lu, Yinhan, et al.
Veröffentlicht: (2026)
Gazelle: An Instruction Dataset for Arabic Writing Assistance
von: Magdy, Samar M., et al.
Veröffentlicht: (2024)
von: Magdy, Samar M., et al.
Veröffentlicht: (2024)
Zero-Shot Context-Aware ASR for Diverse Arabic Varieties
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
Towards Zero-Shot Text-To-Speech for Arabic Dialects
von: Doan, Khai Duy, et al.
Veröffentlicht: (2024)
von: Doan, Khai Duy, et al.
Veröffentlicht: (2024)
On Many-Shot In-Context Learning for Long-Context Evaluation
von: Zou, Kaijian, et al.
Veröffentlicht: (2024)
von: Zou, Kaijian, et al.
Veröffentlicht: (2024)
LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation
von: Magdy, Samar M., et al.
Veröffentlicht: (2026)
von: Magdy, Samar M., et al.
Veröffentlicht: (2026)
Soft Language Identification for Language-Agnostic Many-to-One End-to-End Speech Translation
von: Wang, Peidong, et al.
Veröffentlicht: (2024)
von: Wang, Peidong, et al.
Veröffentlicht: (2024)
uDistil-Whisper: Label-Free Data Filtering for Knowledge Distillation in Low-Data Regimes
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions
von: Wu, Minghao, et al.
Veröffentlicht: (2023)
von: Wu, Minghao, et al.
Veröffentlicht: (2023)
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
LLM Performance Predictors are good initializers for Architecture Search
von: Jawahar, Ganesh, et al.
Veröffentlicht: (2023)
von: Jawahar, Ganesh, et al.
Veröffentlicht: (2023)
Many-to-English Machine Translation Tools, Data, and Pretrained Models
von: Gowda, Thamme, et al.
Veröffentlicht: (2021)
von: Gowda, Thamme, et al.
Veröffentlicht: (2021)
Many-Turn Jailbreaking
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
von: Yang, Xianjun, et al.
Veröffentlicht: (2025)
Peacock: A Family of Arabic Multimodal Large Language Models and Benchmarks
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2024)
von: Alwajih, Fakhraddin, et al.
Veröffentlicht: (2024)
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
Beyond Many-Shot Translation: Scaling In-Context Demonstrations For Low-Resource Machine Translation
von: Salim, Luis Frentzen, et al.
Veröffentlicht: (2026)
von: Salim, Luis Frentzen, et al.
Veröffentlicht: (2026)
DetoxLLM: A Framework for Detoxification with Explanations
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Cheetah: Natural Language Generation for 517 African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2024) -
Where Are We? Evaluating LLM Performance on African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2025) -
AfroScope: A Framework for Studying the Linguistic Landscape of Africa
von: Kwon, Sang Yun, et al.
Veröffentlicht: (2026) -
Interplay of Machine Translation, Diacritics, and Diacritization
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2024) -
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
von: Sullivan, Peter, et al.
Veröffentlicht: (2026)