Effective Self-Mining of In-Context Examples for Unsupervised Machine Translation with LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Mekki, Abdellah El, Abdul-Mageed, Muhammad |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
di: Naeem, Numaan, et al.
Pubblicazione: (2025)
di: Naeem, Numaan, et al.
Pubblicazione: (2025)
LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation
di: Magdy, Samar M., et al.
Pubblicazione: (2026)
di: Magdy, Samar M., et al.
Pubblicazione: (2026)
NileChat: Towards Linguistically Diverse and Culturally Aware LLMs for Local Communities
di: Mekki, Abdellah El, et al.
Pubblicazione: (2025)
di: Mekki, Abdellah El, et al.
Pubblicazione: (2025)
PalmX 2025: The First Shared Task on Benchmarking LLMs on Arabic and Islamic Culture
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2025)
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2025)
Swan and ArabicMTEB: Dialect-Aware, Arabic-Centric, Cross-Lingual, and Cross-Cultural Embedding Models and Benchmarks
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
Interplay of Machine Translation, Diacritics, and Diacritization
di: Chen, Wei-Rui, et al.
Pubblicazione: (2024)
di: Chen, Wei-Rui, et al.
Pubblicazione: (2024)
Toucan: Many-to-Many Translation for 150 African Language Pairs
di: Elmadany, AbdelRahim, et al.
Pubblicazione: (2024)
di: Elmadany, AbdelRahim, et al.
Pubblicazione: (2024)
Zero-Shot Context-Aware ASR for Diverse Arabic Varieties
di: Talafha, Bashar, et al.
Pubblicazione: (2025)
di: Talafha, Bashar, et al.
Pubblicazione: (2025)
Distilling Text Style Transfer With Self-Explanation From LLMs
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
Arabic Automatic Story Generation with Large Language Models
di: El-Shangiti, Ahmed Oumar, et al.
Pubblicazione: (2024)
di: El-Shangiti, Ahmed Oumar, et al.
Pubblicazione: (2024)
To Distill or Not to Distill? On the Robustness of Robust Knowledge Distillation
di: Waheed, Abdul, et al.
Pubblicazione: (2024)
di: Waheed, Abdul, et al.
Pubblicazione: (2024)
On Barriers to Archival Audio Processing
di: Sullivan, Peter, et al.
Pubblicazione: (2025)
di: Sullivan, Peter, et al.
Pubblicazione: (2025)
Ensemble Self-Training for Unsupervised Machine Translation
di: Aharon, Ido, et al.
Pubblicazione: (2026)
di: Aharon, Ido, et al.
Pubblicazione: (2026)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
Cheetah: Natural Language Generation for 517 African Languages
di: Adebara, Ife, et al.
Pubblicazione: (2024)
di: Adebara, Ife, et al.
Pubblicazione: (2024)
Desert Camels and Oil Sheikhs: Arab-Centric Red Teaming of Frontier LLMs
di: Saeed, Muhammed, et al.
Pubblicazione: (2024)
di: Saeed, Muhammed, et al.
Pubblicazione: (2024)
Dallah: A Dialect-Aware Multimodal Large Language Model for Arabic
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2024)
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2024)
Alexandria: A Multi-Domain Dialectal Arabic Machine Translation Dataset for Culturally Inclusive and Linguistically Diverse LLMs
di: Mekki, Abdellah El, et al.
Pubblicazione: (2026)
di: Mekki, Abdellah El, et al.
Pubblicazione: (2026)
In-Context Example Selection via Similarity Search Improves Low-Resource Machine Translation
di: Zebaze, Armel, et al.
Pubblicazione: (2024)
di: Zebaze, Armel, et al.
Pubblicazione: (2024)
Towards Zero-Shot Text-To-Speech for Arabic Dialects
di: Doan, Khai Duy, et al.
Pubblicazione: (2024)
di: Doan, Khai Duy, et al.
Pubblicazione: (2024)
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
Attention Mechanism and Context Modeling System for Text Mining Machine Translation
di: Zhang, Yuwei, et al.
Pubblicazione: (2024)
di: Zhang, Yuwei, et al.
Pubblicazione: (2024)
AfroScope: A Framework for Studying the Linguistic Landscape of Africa
di: Kwon, Sang Yun, et al.
Pubblicazione: (2026)
di: Kwon, Sang Yun, et al.
Pubblicazione: (2026)
Exploring In-context Example Generation for Machine Translation
di: Lee, Dohyun, et al.
Pubblicazione: (2025)
di: Lee, Dohyun, et al.
Pubblicazione: (2025)
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer
di: Zhang, Xiang, et al.
Pubblicazione: (2024)
di: Zhang, Xiang, et al.
Pubblicazione: (2024)
uDistil-Whisper: Label-Free Data Filtering for Knowledge Distillation in Low-Data Regimes
di: Waheed, Abdul, et al.
Pubblicazione: (2024)
di: Waheed, Abdul, et al.
Pubblicazione: (2024)
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions
di: Wu, Minghao, et al.
Pubblicazione: (2023)
di: Wu, Minghao, et al.
Pubblicazione: (2023)
Self-Augmented In-Context Learning for Unsupervised Word Translation
di: Li, Yaoyiran, et al.
Pubblicazione: (2024)
di: Li, Yaoyiran, et al.
Pubblicazione: (2024)
Guiding In-Context Learning of LLMs through Quality Estimation for Machine Translation
di: Sharami, Javad Pourmostafa Roshan, et al.
Pubblicazione: (2024)
di: Sharami, Javad Pourmostafa Roshan, et al.
Pubblicazione: (2024)
LLM Performance Predictors are good initializers for Architecture Search
di: Jawahar, Ganesh, et al.
Pubblicazione: (2023)
di: Jawahar, Ganesh, et al.
Pubblicazione: (2023)
Effective In-Context Example Selection through Data Compression
di: Sun, Zhongxiang, et al.
Pubblicazione: (2024)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2024)
Learning to Search Effective Example Sequences for In-Context Learning
di: Gao, Xiang, et al.
Pubblicazione: (2025)
di: Gao, Xiang, et al.
Pubblicazione: (2025)
Peacock: A Family of Arabic Multimodal Large Language Models and Benchmarks
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2024)
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2024)
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
di: Sullivan, Peter, et al.
Pubblicazione: (2026)
di: Sullivan, Peter, et al.
Pubblicazione: (2026)
DetoxLLM: A Framework for Detoxification with Explanations
di: Khondaker, Md Tawkat Islam, et al.
Pubblicazione: (2024)
di: Khondaker, Md Tawkat Islam, et al.
Pubblicazione: (2024)
Revisiting Context Choices for Context-aware Machine Translation
di: Rikters, Matīss, et al.
Pubblicazione: (2021)
di: Rikters, Matīss, et al.
Pubblicazione: (2021)
Benchmarking LLMs for Mimicking Child-Caregiver Language in Interaction
di: Liu, Jing, et al.
Pubblicazione: (2024)
di: Liu, Jing, et al.
Pubblicazione: (2024)
SCOI: Syntax-augmented Coverage-based In-context Example Selection for Machine Translation
di: Tang, Chenming, et al.
Pubblicazione: (2024)
di: Tang, Chenming, et al.
Pubblicazione: (2024)
POMP: Probability-driven Meta-graph Prompter for LLMs in Low-resource Unsupervised Neural Machine Translation
di: Pan, Shilong, et al.
Pubblicazione: (2024)
di: Pan, Shilong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
di: Naeem, Numaan, et al.
Pubblicazione: (2025) -
LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation
di: Magdy, Samar M., et al.
Pubblicazione: (2026) -
NileChat: Towards Linguistically Diverse and Culturally Aware LLMs for Local Communities
di: Mekki, Abdellah El, et al.
Pubblicazione: (2025) -
PalmX 2025: The First Shared Task on Benchmarking LLMs on Arabic and Islamic Culture
di: Alwajih, Fakhraddin, et al.
Pubblicazione: (2025) -
Swan and ArabicMTEB: Dialect-Aware, Arabic-Centric, Cross-Lingual, and Cross-Cultural Embedding Models and Benchmarks
di: Bhatia, Gagan, et al.
Pubblicazione: (2024)