Low-resource neural machine translation with morphological modeling
Fuente:
arXiv
Salvato in:
| Autore principale: | Nzeyimana, Antoine |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
di: Nzeyimana, Antoine, et al.
Pubblicazione: (2025)
di: Nzeyimana, Antoine, et al.
Pubblicazione: (2025)
MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
di: Moosa, Ibraheem Muhammad, et al.
Pubblicazione: (2024)
di: Moosa, Ibraheem Muhammad, et al.
Pubblicazione: (2024)
Mitigating Translationese in Low-resource Languages: The Storyboard Approach
di: Kuwanto, Garry, et al.
Pubblicazione: (2024)
di: Kuwanto, Garry, et al.
Pubblicazione: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
QFS-Composer: Query-focused summarization pipeline for less resourced languages
di: Đuranović, Vuk, et al.
Pubblicazione: (2026)
di: Đuranović, Vuk, et al.
Pubblicazione: (2026)
Learning and communication pressures in neural networks: Lessons from emergent communication
di: Galke, Lukas, et al.
Pubblicazione: (2024)
di: Galke, Lukas, et al.
Pubblicazione: (2024)
Graphemic Normalization of the Perso-Arabic Script
di: Doctor, Raiomond, et al.
Pubblicazione: (2022)
di: Doctor, Raiomond, et al.
Pubblicazione: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
di: Gutkin, Alexander, et al.
Pubblicazione: (2023)
di: Gutkin, Alexander, et al.
Pubblicazione: (2023)
A Benchmark of French ASR Systems Based on Error Severity
di: Tholly, Antoine, et al.
Pubblicazione: (2025)
di: Tholly, Antoine, et al.
Pubblicazione: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
di: Schneider, Felix, et al.
Pubblicazione: (2026)
di: Schneider, Felix, et al.
Pubblicazione: (2026)
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
di: Galke, Lukas, et al.
Pubblicazione: (2023)
di: Galke, Lukas, et al.
Pubblicazione: (2023)
The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
Low-Resource Court Judgment Summarization for Common Law Systems
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
di: Wanjawa, Barack Wamkaya, et al.
Pubblicazione: (2025)
di: Wanjawa, Barack Wamkaya, et al.
Pubblicazione: (2025)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
di: Gaim, Fitsum, et al.
Pubblicazione: (2025)
di: Gaim, Fitsum, et al.
Pubblicazione: (2025)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
Reflective Translation: Improving Low-Resource Machine Translation via Structured Self-Reflection
di: Cheng, Nicholas
Pubblicazione: (2026)
di: Cheng, Nicholas
Pubblicazione: (2026)
MALT: Mechanistic Ablation of Lossy Translation in LLMs for a Low-Resource Language: Urdu
di: Bajwa, Taaha Saleem
Pubblicazione: (2025)
di: Bajwa, Taaha Saleem
Pubblicazione: (2025)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
di: Dhasmana, Akriti, et al.
Pubblicazione: (2026)
di: Dhasmana, Akriti, et al.
Pubblicazione: (2026)
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
di: Akavarapu, V. S. D. S. Mahesh, et al.
Pubblicazione: (2026)
di: Akavarapu, V. S. D. S. Mahesh, et al.
Pubblicazione: (2026)
SITA: Learning Speaker-Invariant and Tone-Aware Speech Representations for Low-Resource Tonal Languages
di: Xu, Tianyi, et al.
Pubblicazione: (2026)
di: Xu, Tianyi, et al.
Pubblicazione: (2026)
Learning the meanings of function words from grounded language using a visual question answering model
di: Portelance, Eva, et al.
Pubblicazione: (2023)
di: Portelance, Eva, et al.
Pubblicazione: (2023)
Heidelberg-Boston @ SIGTYP 2024 Shared Task: Enhancing Low-Resource Language Analysis With Character-Aware Hierarchical Transformers
di: Riemenschneider, Frederick, et al.
Pubblicazione: (2024)
di: Riemenschneider, Frederick, et al.
Pubblicazione: (2024)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
di: Stewart, Ian, et al.
Pubblicazione: (2024)
di: Stewart, Ian, et al.
Pubblicazione: (2024)
Are formal and functional linguistic mechanisms dissociated in language models?
di: Hanna, Michael, et al.
Pubblicazione: (2025)
di: Hanna, Michael, et al.
Pubblicazione: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data
di: de Mello, Guilherme Lamartine, et al.
Pubblicazione: (2024)
di: de Mello, Guilherme Lamartine, et al.
Pubblicazione: (2024)
Towards interpretable models for language proficiency assessment: Predicting the CEFR level of Estonian learner texts
di: Allkivi, Kais
Pubblicazione: (2026)
di: Allkivi, Kais
Pubblicazione: (2026)
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
di: Thamma, Abishek, et al.
Pubblicazione: (2025)
di: Thamma, Abishek, et al.
Pubblicazione: (2025)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
di: Guo, Willis, et al.
Pubblicazione: (2024)
di: Guo, Willis, et al.
Pubblicazione: (2024)
LLMs and the Human Condition
di: Wallis, Peter
Pubblicazione: (2024)
di: Wallis, Peter
Pubblicazione: (2024)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
di: Dai, Song, et al.
Pubblicazione: (2025)
di: Dai, Song, et al.
Pubblicazione: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
On the Influence of Discourse Relations in Persuasive Texts
di: Turk, Nawar, et al.
Pubblicazione: (2025)
di: Turk, Nawar, et al.
Pubblicazione: (2025)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
di: Tong, Jingqi, et al.
Pubblicazione: (2025)
di: Tong, Jingqi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
di: Nzeyimana, Antoine, et al.
Pubblicazione: (2025) -
MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
di: Moosa, Ibraheem Muhammad, et al.
Pubblicazione: (2024) -
Mitigating Translationese in Low-resource Languages: The Storyboard Approach
di: Kuwanto, Garry, et al.
Pubblicazione: (2024) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
di: Ashuach, Tomer, et al.
Pubblicazione: (2025) -
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)