A Reproducibility Study on Quantifying Language Similarity: The Impact of Missing Values in the URIEL Knowledge Base
Fuente:
arXiv
Saved in:
| Main Authors: | Toossi, Hasti, Huai, Guo Qing, Liu, Jinyu, Khiu, Eric, Doğruöz, A. Seza, Lee, En-Shiun Annie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Predicting Machine Translation Performance on Low-Resource Languages: The Role of Domain Similarity
by: Khiu, Eric, et al.
Published: (2024)
by: Khiu, Eric, et al.
Published: (2024)
URIEL+: Enhancing Linguistic Inclusion and Usability in a Typological and Multilingual Knowledge Base
by: Khan, Aditya, et al.
Published: (2024)
by: Khan, Aditya, et al.
Published: (2024)
Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+
by: Shipton, Mason, et al.
Published: (2025)
by: Shipton, Mason, et al.
Published: (2025)
Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+
by: Ng, York Hay, et al.
Published: (2025)
by: Ng, York Hay, et al.
Published: (2025)
Readability Measures and Automatic Text Simplification: In the Search of a Construct
by: Cardon, Rémi, et al.
Published: (2025)
by: Cardon, Rémi, et al.
Published: (2025)
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
A Typology of Synthetic Datasets for Dialogue Processing in Clinical Contexts
by: Bedrick, Steven, et al.
Published: (2025)
by: Bedrick, Steven, et al.
Published: (2025)
Leveraging LLMs for Translating and Classifying Mental Health Data
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
Less is More: The Effectiveness of Compact Typological Language Representations
by: Ng, York Hay, et al.
Published: (2025)
by: Ng, York Hay, et al.
Published: (2025)
Comparing LLM prompting with Cross-lingual transfer performance on Indigenous and Low-resource Brazilian Languages
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
Single- vs. Dual-Prompt Dialogue Generation with LLMs for Job Interviews in Human Resources
by: De Baer, Joachim, et al.
Published: (2025)
by: De Baer, Joachim, et al.
Published: (2025)
Who is bragging more online? A large scale analysis of bragging in social media
by: Jin, Mali, et al.
Published: (2024)
by: Jin, Mali, et al.
Published: (2024)
Does Generative AI speak Nigerian-Pidgin?: Issues about Representativeness and Bias for Multilingualism in LLMs
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
AlignFreeze: Navigating the Impact of Realignment on the Layers of Multilingual Models Across Diverse Languages
by: Bakos, Steve, et al.
Published: (2025)
by: Bakos, Steve, et al.
Published: (2025)
Unlocking Parameter-Efficient Fine-Tuning for Low-Resource Language Translation
by: Su, Tong, et al.
Published: (2024)
by: Su, Tong, et al.
Published: (2024)
Rethinking what Matters: Effective and Robust Multilingual Realignment for Low-Resource Languages
by: Nguyen, Quang Phuoc, et al.
Published: (2025)
by: Nguyen, Quang Phuoc, et al.
Published: (2025)
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
by: Anugraha, David, et al.
Published: (2024)
by: Anugraha, David, et al.
Published: (2024)
TranslationCorrect: A Unified Framework for Machine Translation Post-Editing with Predictive Error Assistance
by: Wasti, Syed Mekael, et al.
Published: (2025)
by: Wasti, Syed Mekael, et al.
Published: (2025)
Beyond Vanilla Fine-Tuning: Leveraging Multistage, Multilingual, and Domain-Specific Methods for Low-Resource Machine Translation
by: Thillainathan, Sarubi, et al.
Published: (2025)
by: Thillainathan, Sarubi, et al.
Published: (2025)
Dynamic Meta-Metrics: Source-Sentence Conditioned Weighting for MT Evaluation
by: Zhang, Luke, et al.
Published: (2026)
by: Zhang, Luke, et al.
Published: (2026)
AKGNet: Attribute Knowledge-Guided Unsupervised Lung-Infected Area Segmentation
by: En, Qing, et al.
Published: (2024)
by: En, Qing, et al.
Published: (2024)
M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAG
by: Anugraha, David, et al.
Published: (2025)
by: Anugraha, David, et al.
Published: (2025)
\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding
by: Min, Junghyun, et al.
Published: (2025)
by: Min, Junghyun, et al.
Published: (2025)
Enhancing Taiwanese Hokkien Dual Translation by Exploring and Standardizing of Four Writing Systems
by: Lu, Bo-Han, et al.
Published: (2024)
by: Lu, Bo-Han, et al.
Published: (2024)
Quantifying Uncertainty in Natural Language Explanations of Large Language Models for Question Answering
by: Li, Yangyi, et al.
Published: (2025)
by: Li, Yangyi, et al.
Published: (2025)
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
by: Kim, Eunsu, et al.
Published: (2025)
by: Kim, Eunsu, et al.
Published: (2025)
Impact of Language Guidance: A Reproducibility Study
by: Puniani, Cherish, et al.
Published: (2025)
by: Puniani, Cherish, et al.
Published: (2025)
The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility
by: Pape, David, et al.
Published: (2026)
by: Pape, David, et al.
Published: (2026)
SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects
by: Adelani, David Ifeoluwa, et al.
Published: (2023)
by: Adelani, David Ifeoluwa, et al.
Published: (2023)
Similarity-Based Assessment of Computational Reproducibility in Jupyter Notebooks
by: Hossain, A S M Shahadat, et al.
Published: (2025)
by: Hossain, A S M Shahadat, et al.
Published: (2025)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
by: Anugraha, David, et al.
Published: (2025)
by: Anugraha, David, et al.
Published: (2025)
Overview of MWE history, challenges, and horizons: standing at the 20th anniversary of the MWE workshop series via MWE-UD2024
by: Han, Lifeng, et al.
Published: (2024)
by: Han, Lifeng, et al.
Published: (2024)
Exploiting Domain-Specific Parallel Data on Multilingual Language Models for Low-resource Language Translation
by: Ranathungaa, Surangika, et al.
Published: (2024)
by: Ranathungaa, Surangika, et al.
Published: (2024)
Enhancing Mental Health Counseling Support in Bangladesh using Culturally-Grounded Knowledge
by: Hasan, Md Arid, et al.
Published: (2026)
by: Hasan, Md Arid, et al.
Published: (2026)
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph
by: Tamir, Azwad, et al.
Published: (2024)
by: Tamir, Azwad, et al.
Published: (2024)
Retrieving Implicit and Explicit Emotional Events Using Large Language Models
by: Hu, Guimin, et al.
Published: (2024)
by: Hu, Guimin, et al.
Published: (2024)
Impact of Missing Values on the Efficiency of Dispersion Control Charts
by: Faisal Maqbool Zahid, et al.
Published: (2026)
by: Faisal Maqbool Zahid, et al.
Published: (2026)
Machine Learning Based Missing Values Imputation in Categorical Datasets
by: Ishaq, Muhammad, et al.
Published: (2023)
by: Ishaq, Muhammad, et al.
Published: (2023)
Beyond Pairwise Similarity: Quantifying and Characterizing Linguistic Similarity between Groups of Languages by MDL
by: Andrea K. Fischer
Published: (2017)
by: Andrea K. Fischer
Published: (2017)
Cross-model Mutual Learning for Exemplar-based Medical Image Segmentation
by: En, Qing, et al.
Published: (2024)
by: En, Qing, et al.
Published: (2024)
Similar Items
-
Predicting Machine Translation Performance on Low-Resource Languages: The Role of Domain Similarity
by: Khiu, Eric, et al.
Published: (2024) -
URIEL+: Enhancing Linguistic Inclusion and Usability in a Typological and Multilingual Knowledge Base
by: Khan, Aditya, et al.
Published: (2024) -
Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+
by: Shipton, Mason, et al.
Published: (2025) -
Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+
by: Ng, York Hay, et al.
Published: (2025) -
Readability Measures and Automatic Text Simplification: In the Search of a Construct
by: Cardon, Rémi, et al.
Published: (2025)