Exploring Performance Variations in Finetuned Translators of Ultra-Low Resource Languages: Do Linguistic Differences Matter?
Fuente:
arXiv
Salvato in:
| Autori principali: | Gonçalves, Isabel, Cavalin, Paulo, Pinhanez, Claudio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Non-Determinism of Small LLMs: Evidence of Low Answer Consistency in Repetition Trials of Standard Multiple-Choice Benchmarks
di: Pinhanez, Claudio, et al.
Pubblicazione: (2025)
di: Pinhanez, Claudio, et al.
Pubblicazione: (2025)
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
di: Cavalin, Paulo, et al.
Pubblicazione: (2025)
di: Cavalin, Paulo, et al.
Pubblicazione: (2025)
Sentence-level Aggregation of Lexical Metrics Correlates Stronger with Human Judgements than Corpus-level Aggregation
di: Cavalin, Paulo, et al.
Pubblicazione: (2024)
di: Cavalin, Paulo, et al.
Pubblicazione: (2024)
Harnessing the Power of Artificial Intelligence to Vitalize Endangered Indigenous Languages: Technologies and Experiences
di: Pinhanez, Claudio, et al.
Pubblicazione: (2024)
di: Pinhanez, Claudio, et al.
Pubblicazione: (2024)
CAT: A Metric-Driven Framework for Analyzing the Consistency-Accuracy Relation of LLMs under Controlled Input Variations
di: Cavalin, Paulo, et al.
Pubblicazione: (2025)
di: Cavalin, Paulo, et al.
Pubblicazione: (2025)
Can Linguistically Related Languages Guide LLM Translation in Low-Resource Settings?
di: Ramasethu, Aishwarya, et al.
Pubblicazione: (2026)
di: Ramasethu, Aishwarya, et al.
Pubblicazione: (2026)
Transcending Language Boundaries: Harnessing LLMs for Low-Resource Language Translation
di: Shu, Peng, et al.
Pubblicazione: (2024)
di: Shu, Peng, et al.
Pubblicazione: (2024)
Harnessing Linguistic Dissimilarity for Language Generalization on Unseen Low-Resource Varieties
di: Kim, Jinju, et al.
Pubblicazione: (2026)
di: Kim, Jinju, et al.
Pubblicazione: (2026)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
di: Somayajula, Sai Ashish, et al.
Pubblicazione: (2024)
di: Somayajula, Sai Ashish, et al.
Pubblicazione: (2024)
LSR: Linguistic Safety Robustness Benchmark for Low-Resource West African Languages
di: Faruna, Godwin Abuh
Pubblicazione: (2026)
di: Faruna, Godwin Abuh
Pubblicazione: (2026)
On Instruction-Finetuning Neural Machine Translation Models
di: Raunak, Vikas, et al.
Pubblicazione: (2024)
di: Raunak, Vikas, et al.
Pubblicazione: (2024)
TravelBench : Exploring LLM Performance in Low-Resource Domains
di: Billa, Srinivas, et al.
Pubblicazione: (2025)
di: Billa, Srinivas, et al.
Pubblicazione: (2025)
LLMs for Translation: Historical, Low-Resourced Languages and Contemporary AI Models
di: Tekgurler, Merve
Pubblicazione: (2025)
di: Tekgurler, Merve
Pubblicazione: (2025)
Machine Translation Advancements of Low-Resource Indian Languages by Transfer Learning
di: Wei, Bin, et al.
Pubblicazione: (2024)
di: Wei, Bin, et al.
Pubblicazione: (2024)
Democratizing LLMs for Low-Resource Languages by Leveraging their English Dominant Abilities with Linguistically-Diverse Prompts
di: Nguyen, Xuan-Phi, et al.
Pubblicazione: (2023)
di: Nguyen, Xuan-Phi, et al.
Pubblicazione: (2023)
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
di: Truong, Sang T., et al.
Pubblicazione: (2024)
di: Truong, Sang T., et al.
Pubblicazione: (2024)
Rethinking what Matters: Effective and Robust Multilingual Realignment for Low-Resource Languages
di: Nguyen, Quang Phuoc, et al.
Pubblicazione: (2025)
di: Nguyen, Quang Phuoc, et al.
Pubblicazione: (2025)
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
di: Wang, Yu, et al.
Pubblicazione: (2026)
di: Wang, Yu, et al.
Pubblicazione: (2026)
SPRING Lab IITM's submission to Low Resource Indic Language Translation Shared Task
di: Sayed, Hamees, et al.
Pubblicazione: (2024)
di: Sayed, Hamees, et al.
Pubblicazione: (2024)
Whisper Finetuning on Nepali Language
di: Rijal, Sanjay, et al.
Pubblicazione: (2024)
di: Rijal, Sanjay, et al.
Pubblicazione: (2024)
Context Volume Drives Performance: Tackling Domain Shift in Extremely Low-Resource Translation via RAG
di: Setiawan, David Samuel, et al.
Pubblicazione: (2026)
di: Setiawan, David Samuel, et al.
Pubblicazione: (2026)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
di: Gao, Yufei, et al.
Pubblicazione: (2025)
di: Gao, Yufei, et al.
Pubblicazione: (2025)
Enhancing Low-Resource Minority Language Translation with LLMs and Retrieval-Augmented Generation for Cultural Nuances
di: Chang, Chen-Chi, et al.
Pubblicazione: (2025)
di: Chang, Chen-Chi, et al.
Pubblicazione: (2025)
OpenWHO: A Document-Level Parallel Corpus for Health Translation in Low-Resource Languages
di: Merx, Raphaël, et al.
Pubblicazione: (2025)
di: Merx, Raphaël, et al.
Pubblicazione: (2025)
Narrow Finetuning Leaves Clearly Readable Traces in Activation Differences
di: Minder, Julian, et al.
Pubblicazione: (2025)
di: Minder, Julian, et al.
Pubblicazione: (2025)
VLURes: Benchmarking VLM Visual and Linguistic Understanding in Low-Resource Languages
di: Atuhurra, Jesse, et al.
Pubblicazione: (2025)
di: Atuhurra, Jesse, et al.
Pubblicazione: (2025)
Leveraging Large Language Models to Geolocate Linguistic Variations in Social Media Posts
di: Savarro, Davide, et al.
Pubblicazione: (2024)
di: Savarro, Davide, et al.
Pubblicazione: (2024)
Improving Multilingual Instruction Finetuning via Linguistically Natural and Diverse Datasets
di: Indurthi, Sathish Reddy, et al.
Pubblicazione: (2024)
di: Indurthi, Sathish Reddy, et al.
Pubblicazione: (2024)
Can Large Language Models Code Like a Linguist?: A Case Study in Low Resource Sound Law Induction
di: Naik, Atharva, et al.
Pubblicazione: (2024)
di: Naik, Atharva, et al.
Pubblicazione: (2024)
Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen
di: Zainaldin, James L., et al.
Pubblicazione: (2026)
di: Zainaldin, James L., et al.
Pubblicazione: (2026)
Bridging the Linguistic Divide: A Survey on Leveraging Large Language Models for Machine Translation
di: Gain, Baban, et al.
Pubblicazione: (2025)
di: Gain, Baban, et al.
Pubblicazione: (2025)
MyCulture: Exploring Malaysia's Diverse Culture under Low-Resource Language Constraints
di: Hew, Zhong Ken, et al.
Pubblicazione: (2025)
di: Hew, Zhong Ken, et al.
Pubblicazione: (2025)
Exploring Linguistic Properties of Monolingual BERTs with Typological Classification among Languages
di: Ruzzetti, Elena Sofia, et al.
Pubblicazione: (2023)
di: Ruzzetti, Elena Sofia, et al.
Pubblicazione: (2023)
TALL -- A Trainable Architecture for Enhancing LLM Performance in Low-Resource Languages
di: Ofer, Moshe, et al.
Pubblicazione: (2025)
di: Ofer, Moshe, et al.
Pubblicazione: (2025)
Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese
di: Yakhni, Silvana, et al.
Pubblicazione: (2025)
di: Yakhni, Silvana, et al.
Pubblicazione: (2025)
The Effectiveness of Morphology-aware Segmentation in Low-Resource Neural Machine Translation
di: Sälevä, Jonne, et al.
Pubblicazione: (2021)
di: Sälevä, Jonne, et al.
Pubblicazione: (2021)
From LLM to NMT: Advancing Low-Resource Machine Translation with Claude
di: Enis, Maxim, et al.
Pubblicazione: (2024)
di: Enis, Maxim, et al.
Pubblicazione: (2024)
Enhancing Neural Machine Translation of Low-Resource Languages: Corpus Development, Human Evaluation and Explainable AI Architectures
di: Lankford, Séamus
Pubblicazione: (2024)
di: Lankford, Séamus
Pubblicazione: (2024)
Risk-Averse Finetuning of Large Language Models
di: Chaudhary, Sapana, et al.
Pubblicazione: (2025)
di: Chaudhary, Sapana, et al.
Pubblicazione: (2025)
LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages
di: Lu, Yinquan, et al.
Pubblicazione: (2024)
di: Lu, Yinquan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
The Non-Determinism of Small LLMs: Evidence of Low Answer Consistency in Repetition Trials of Standard Multiple-Choice Benchmarks
di: Pinhanez, Claudio, et al.
Pubblicazione: (2025) -
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
di: Cavalin, Paulo, et al.
Pubblicazione: (2025) -
Sentence-level Aggregation of Lexical Metrics Correlates Stronger with Human Judgements than Corpus-level Aggregation
di: Cavalin, Paulo, et al.
Pubblicazione: (2024) -
Harnessing the Power of Artificial Intelligence to Vitalize Endangered Indigenous Languages: Technologies and Experiences
di: Pinhanez, Claudio, et al.
Pubblicazione: (2024) -
CAT: A Metric-Driven Framework for Analyzing the Consistency-Accuracy Relation of LLMs under Controlled Input Variations
di: Cavalin, Paulo, et al.
Pubblicazione: (2025)