A Principled Framework for Evaluating on Typologically Diverse Languages
Fuente:
arXiv
Salvato in:
| Autori principali: | Ploeger, Esther, Poelman, Wessel, Høeg-Petersen, Andreas Holck, Schlichtkrull, Anders, de Lhoneux, Miryam, Bjerva, Johannes |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What is "Typological Diversity" in NLP?
di: Ploeger, Esther, et al.
Pubblicazione: (2024)
di: Ploeger, Esther, et al.
Pubblicazione: (2024)
The Roles of English in Evaluating Multilingual Language Models
di: Poelman, Wessel, et al.
Pubblicazione: (2024)
di: Poelman, Wessel, et al.
Pubblicazione: (2024)
Form and Meaning in Intrinsic Multilingual Evaluations
di: Poelman, Wessel, et al.
Pubblicazione: (2026)
di: Poelman, Wessel, et al.
Pubblicazione: (2026)
QQ: A Toolkit for Language Identifiers and Metadata
di: Poelman, Wessel, et al.
Pubblicazione: (2026)
di: Poelman, Wessel, et al.
Pubblicazione: (2026)
How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
Multilingual Gradient Word-Order Typology from Universal Dependencies
di: Baylor, Emi, et al.
Pubblicazione: (2024)
di: Baylor, Emi, et al.
Pubblicazione: (2024)
On the Interplay between Positional Encodings, Morphological Complexity, and Word Order Flexibility
di: Tatariya, Kushal, et al.
Pubblicazione: (2025)
di: Tatariya, Kushal, et al.
Pubblicazione: (2025)
Confounding Factors in Relating Model Performance to Morphology
di: Poelman, Wessel, et al.
Pubblicazione: (2025)
di: Poelman, Wessel, et al.
Pubblicazione: (2025)
Typologically Informed Parameter Aggregation
di: Accou, Stef, et al.
Pubblicazione: (2026)
di: Accou, Stef, et al.
Pubblicazione: (2026)
Sociolinguistically Informed Interpretability: A Case Study on Hinglish Emotion Classification
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
We Need to Measure Data Diversity in NLP -- Better and Broader
di: Nguyen, Dong, et al.
Pubblicazione: (2025)
di: Nguyen, Dong, et al.
Pubblicazione: (2025)
Type and Complexity Signals in Multilingual Question Representations
di: Kokot, Robin, et al.
Pubblicazione: (2025)
di: Kokot, Robin, et al.
Pubblicazione: (2025)
Against All Odds: Overcoming Typology, Script, and Language Confusion in Multilingual Embedding Inversion Attacks
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
Recipe for Zero-shot POS Tagging: Is It Useful in Realistic Scenarios?
di: Vandenbulcke, Zeno, et al.
Pubblicazione: (2024)
di: Vandenbulcke, Zeno, et al.
Pubblicazione: (2024)
Pixology: Probing the Linguistic and Visual Capabilities of Pixel-based Language Models
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
Patterns of Persistence and Diffusibility across the World's Languages
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
Linguistically Grounded Analysis of Language Models using Shapley Head Values
di: Fekete, Marcell, et al.
Pubblicazione: (2024)
di: Fekete, Marcell, et al.
Pubblicazione: (2024)
Towards Tailored Recovery of Lexical Diversity in Literary Machine Translation
di: Ploeger, Esther, et al.
Pubblicazione: (2024)
di: Ploeger, Esther, et al.
Pubblicazione: (2024)
Engineering Conversational Search Systems: A Review of Applications, Architectures, and Functional Components
di: Schneider, Phillip, et al.
Pubblicazione: (2024)
di: Schneider, Phillip, et al.
Pubblicazione: (2024)
When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs
di: Fekete, Marcell, et al.
Pubblicazione: (2026)
di: Fekete, Marcell, et al.
Pubblicazione: (2026)
Generating Media Background Checks for Automated Source Critical Reasoning
di: Schlichtkrull, Michael
Pubblicazione: (2024)
di: Schlichtkrull, Michael
Pubblicazione: (2024)
Trans-Tokenization and Cross-lingual Vocabulary Transfers: Language Adaptation of LLMs for Low-Resource NLP
di: Remy, François, et al.
Pubblicazione: (2024)
di: Remy, François, et al.
Pubblicazione: (2024)
CreoleVal: Multilingual Multitask Benchmarks for Creoles
di: Lent, Heather, et al.
Pubblicazione: (2023)
di: Lent, Heather, et al.
Pubblicazione: (2023)
Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality
di: Zhang, Mike, et al.
Pubblicazione: (2025)
di: Zhang, Mike, et al.
Pubblicazione: (2025)
Attacks by Content: Automated Fact-checking is an AI Security Issue
di: Schlichtkrull, Michael
Pubblicazione: (2025)
di: Schlichtkrull, Michael
Pubblicazione: (2025)
MultiHal: Multilingual Dataset for Knowledge-Graph Grounded Evaluation of LLM Hallucinations
di: Lavrinovics, Ernests, et al.
Pubblicazione: (2025)
di: Lavrinovics, Ernests, et al.
Pubblicazione: (2025)
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
Text Embedding Inversion Security for Multilingual Language Models
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
di: Chen, Yiyi, et al.
Pubblicazione: (2024)
Document-level Claim Extraction and Decontextualisation for Fact-Checking
di: Deng, Zhenyun, et al.
Pubblicazione: (2024)
di: Deng, Zhenyun, et al.
Pubblicazione: (2024)
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking
di: Akhtar, Mubashara, et al.
Pubblicazione: (2024)
di: Akhtar, Mubashara, et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Actionable Course Evaluation Student Feedback to Lecturers
di: Zhang, Mike, et al.
Pubblicazione: (2024)
di: Zhang, Mike, et al.
Pubblicazione: (2024)
Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation
di: Lai, Huiyuan, et al.
Pubblicazione: (2024)
di: Lai, Huiyuan, et al.
Pubblicazione: (2024)
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
di: Luo, Xiaoyu, et al.
Pubblicazione: (2025)
di: Luo, Xiaoyu, et al.
Pubblicazione: (2025)
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages
di: Onyame, Eric, et al.
Pubblicazione: (2026)
di: Onyame, Eric, et al.
Pubblicazione: (2026)
Efficient Shield Synthesis via State-Space Transformation
di: Brorholt, Asger Horn, et al.
Pubblicazione: (2024)
di: Brorholt, Asger Horn, et al.
Pubblicazione: (2024)
Social Good or Scientific Curiosity? Uncovering the Research Framing Behind NLP Artefacts
di: Chamoun, Eric, et al.
Pubblicazione: (2025)
di: Chamoun, Eric, et al.
Pubblicazione: (2025)
Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages
di: Brinkmann, Jannik, et al.
Pubblicazione: (2025)
di: Brinkmann, Jannik, et al.
Pubblicazione: (2025)
AVerImaTeC: A Dataset for Automatic Verification of Image-Text Claims with Evidence from the Web
di: Cao, Rui, et al.
Pubblicazione: (2025)
di: Cao, Rui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
What is "Typological Diversity" in NLP?
di: Ploeger, Esther, et al.
Pubblicazione: (2024) -
The Roles of English in Evaluating Multilingual Language Models
di: Poelman, Wessel, et al.
Pubblicazione: (2024) -
Form and Meaning in Intrinsic Multilingual Evaluations
di: Poelman, Wessel, et al.
Pubblicazione: (2026) -
QQ: A Toolkit for Language Identifiers and Metadata
di: Poelman, Wessel, et al.
Pubblicazione: (2026) -
How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP
di: Tatariya, Kushal, et al.
Pubblicazione: (2024)