The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shani, Chen, Reif, Yuval, Roll, Nathan, Jurafsky, Dan, Shutova, Ekaterina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
PolyPrompt: Automating Knowledge Extraction from Multilingual Language Models with Dynamic Prompt Generation
von: Roll, Nathan
Veröffentlicht: (2025)
von: Roll, Nathan
Veröffentlicht: (2025)
Entanglement as Memory: Mechanistic Interpretability of Quantum Language Models
von: Roll, Nathan
Veröffentlicht: (2026)
von: Roll, Nathan
Veröffentlicht: (2026)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2023)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2023)
Categorize Early, Integrate Late: Divergent Processing Strategies in Automatic Speech Recognition
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
Beyond Performance: Quantifying and Mitigating Label Bias in LLMs
von: Reif, Yuval, et al.
Veröffentlicht: (2024)
von: Reif, Yuval, et al.
Veröffentlicht: (2024)
In-Context Learning Boosts Speech Recognition via Human-like Adaptation to Speakers and Language Varieties
von: Roll, Nathan, et al.
Veröffentlicht: (2025)
von: Roll, Nathan, et al.
Veröffentlicht: (2025)
Induction Heads as an Essential Mechanism for Pattern Matching in In-context Learning
von: Crosbie, Joy, et al.
Veröffentlicht: (2024)
von: Crosbie, Joy, et al.
Veröffentlicht: (2024)
Self-Alignment: Improving Alignment of Cultural Values in LLMs via In-Context Learning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Beyond Tokens: Concept-Level Training Objectives for LLMs
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Cross-modal Information Flow in Multimodal Large Language Models
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
Yesterday's News: Benchmarking Multi-Dimensional Out-of-Distribution Generalization of Misinformation Detection Models
von: Verhoeven, Ivo, et al.
Veröffentlicht: (2024)
von: Verhoeven, Ivo, et al.
Veröffentlicht: (2024)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
Quantifying Language Disparities in Multilingual Large Language Models
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
Cooking Up Creativity: Enhancing LLM Creativity through Structured Recombination
von: Mizrahi, Moran, et al.
Veröffentlicht: (2025)
von: Mizrahi, Moran, et al.
Veröffentlicht: (2025)
Artificial Aphasias in Lesioned Language Models
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
Beyond Words: Exploring Cultural Value Sensitivity in Multimodal Models
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
A framework for annotating and modelling intentions behind metaphor use
von: Michelli, Gianluca, et al.
Veröffentlicht: (2024)
von: Michelli, Gianluca, et al.
Veröffentlicht: (2024)
Density Matrices for Metaphor Understanding
von: Owers, Jay, et al.
Veröffentlicht: (2024)
von: Owers, Jay, et al.
Veröffentlicht: (2024)
A Shared Geometry of Difficulty in Multilingual Language Models
von: Civelli, Stefano, et al.
Veröffentlicht: (2026)
von: Civelli, Stefano, et al.
Veröffentlicht: (2026)
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
von: Shani, Chen, et al.
Veröffentlicht: (2025)
von: Shani, Chen, et al.
Veröffentlicht: (2025)
CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition
von: Bartelds, Martijn, et al.
Veröffentlicht: (2025)
von: Bartelds, Martijn, et al.
Veröffentlicht: (2025)
Metaphor Understanding Challenge Dataset for LLMs
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
Grounding Gaps in Language Model Generations
von: Shaikh, Omar, et al.
Veröffentlicht: (2023)
von: Shaikh, Omar, et al.
Veröffentlicht: (2023)
Are LLMs classical or nonmonotonic reasoners? Lessons from generics
von: Leidinger, Alina, et al.
Veröffentlicht: (2024)
von: Leidinger, Alina, et al.
Veröffentlicht: (2024)
Learning New Tasks from a Few Examples with Soft-Label Prototypes
von: Singh, Avyav Kumar, et al.
Veröffentlicht: (2022)
von: Singh, Avyav Kumar, et al.
Veröffentlicht: (2022)
Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic
von: Reif, Yuval, et al.
Veröffentlicht: (2025)
von: Reif, Yuval, et al.
Veröffentlicht: (2025)
Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models
von: Kaplan, Guy, et al.
Veröffentlicht: (2025)
von: Kaplan, Guy, et al.
Veröffentlicht: (2025)
Difficulty-Controllable Multiple-Choice Question Generation Using Large Language Models and Direct Preference Optimization
von: Tomikawa, Yuto, et al.
Veröffentlicht: (2025)
von: Tomikawa, Yuto, et al.
Veröffentlicht: (2025)
Exploring Representational Disparities Between Multilingual and Bilingual Translation Models
von: Verma, Neha, et al.
Veröffentlicht: (2023)
von: Verma, Neha, et al.
Veröffentlicht: (2023)
ML-SUPERB 2.0: Benchmarking Multilingual Speech Models Across Modeling Constraints, Languages, and Datasets
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
A (More) Realistic Evaluation Setup for Generalisation of Community Models on Malicious Content Detection
von: Verhoeven, Ivo, et al.
Veröffentlicht: (2024)
von: Verhoeven, Ivo, et al.
Veröffentlicht: (2024)
The Script Tax: Measuring Tokenization-Driven Efficiency and Latency Disparities in Multilingual Language Models
von: Dixit, Aradhya, et al.
Veröffentlicht: (2026)
von: Dixit, Aradhya, et al.
Veröffentlicht: (2026)
Can Model Uncertainty Function as a Proxy for Multiple-Choice Question Item Difficulty?
von: Zotos, Leonidas, et al.
Veröffentlicht: (2024)
von: Zotos, Leonidas, et al.
Veröffentlicht: (2024)
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
von: de la Fuente, Antón, et al.
Veröffentlicht: (2024)
von: de la Fuente, Antón, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
von: Zhang, Christine, et al.
Veröffentlicht: (2026) -
PolyPrompt: Automating Knowledge Extraction from Multilingual Language Models with Dynamic Prompt Generation
von: Roll, Nathan
Veröffentlicht: (2025) -
Entanglement as Memory: Mechanistic Interpretability of Quantum Language Models
von: Roll, Nathan
Veröffentlicht: (2026) -
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2025) -
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)