Understanding the effects of language-specific class imbalance in multilingual fine-tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Jung, Vincent, van der Plas, Lonneke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can language models learn analogical reasoning? Investigating training objectives and comparisons to human performance
by: Petersen, Molly R., et al.
Published: (2023)
by: Petersen, Molly R., et al.
Published: (2023)
From 124 Million Tokens to 1,021 Neologisms: A Large-Scale Pipeline for Automatic Neologism Detection
by: Rossini, Diego, et al.
Published: (2026)
by: Rossini, Diego, et al.
Published: (2026)
Binary Token-Level Classification with DeBERTa for All-Type MWE Identification: A Lightweight Approach with Linguistic Enhancement
by: Rossini, Diego, et al.
Published: (2026)
by: Rossini, Diego, et al.
Published: (2026)
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Modelling Analogies and Analogical Reasoning: Connecting Cognitive Science Theory and NLP Research
by: Petersen, Molly R, et al.
Published: (2025)
by: Petersen, Molly R, et al.
Published: (2025)
Creativity in AI: Progresses and Challenges
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge
by: Ismayilzada, Mete, et al.
Published: (2026)
by: Ismayilzada, Mete, et al.
Published: (2026)
A dual task learning approach to fine-tune a multilingual semantic speech encoder for Spoken Language Understanding
by: Laperrière, Gaëlle, et al.
Published: (2024)
by: Laperrière, Gaëlle, et al.
Published: (2024)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
Creative Preference Optimization
by: Ismayilzada, Mete, et al.
Published: (2025)
by: Ismayilzada, Mete, et al.
Published: (2025)
Comparison between parameter-efficient techniques and full fine-tuning: A case study on multilingual news article classification
by: Razuvayevskaya, Olesya, et al.
Published: (2023)
by: Razuvayevskaya, Olesya, et al.
Published: (2023)
Fine-tuning multilingual language models in Twitter/X sentiment analysis: a study on Eastern-European V4 languages
by: Filip, Tomáš, et al.
Published: (2024)
by: Filip, Tomáš, et al.
Published: (2024)
LLMs that Understand Processes: Instruction-tuning for Semantics-Aware Process Mining
by: Pyrih, Vira, et al.
Published: (2025)
by: Pyrih, Vira, et al.
Published: (2025)
Evaluating Morphological Compositional Generalization in Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Effective vocabulary expanding of multilingual language models for extremely low-resource languages
by: Zheng, Jianyu
Published: (2026)
by: Zheng, Jianyu
Published: (2026)
Deep literature reviews: an application of fine-tuned language models to migration research
by: Iacus, Stefano M., et al.
Published: (2025)
by: Iacus, Stefano M., et al.
Published: (2025)
The representation landscape of few-shot learning and fine-tuning in large language models
by: Doimo, Diego, et al.
Published: (2024)
by: Doimo, Diego, et al.
Published: (2024)
How does fine-tuning improve sensorimotor representations in large language models?
by: Wu, Minghua, et al.
Published: (2026)
by: Wu, Minghua, et al.
Published: (2026)
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
by: Han, Yifu, et al.
Published: (2025)
by: Han, Yifu, et al.
Published: (2025)
MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Multi-task retriever fine-tuning for domain-specific and efficient RAG
by: Béchard, Patrice, et al.
Published: (2025)
by: Béchard, Patrice, et al.
Published: (2025)
EuroGEST: Investigating gender stereotypes in multilingual language models
by: Rowe, Jacqueline, et al.
Published: (2025)
by: Rowe, Jacqueline, et al.
Published: (2025)
Automated scoring of the Ambiguous Intentions Hostility Questionnaire using fine-tuned large language models
by: Lyu, Y., et al.
Published: (2025)
by: Lyu, Y., et al.
Published: (2025)
SRS-Stories: Vocabulary-constrained multilingual story generation for language learning
by: Kamzela, Wiktor, et al.
Published: (2025)
by: Kamzela, Wiktor, et al.
Published: (2025)
Large Language Models Align with the Human Brain during Creative Thinking
by: Ismayilzada, Mete, et al.
Published: (2026)
by: Ismayilzada, Mete, et al.
Published: (2026)
Understanding the role of FFNs in driving multilingual behaviour in LLMs
by: Bhattacharya, Sunit, et al.
Published: (2024)
by: Bhattacharya, Sunit, et al.
Published: (2024)
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
by: Choenni, Rochelle, et al.
Published: (2023)
by: Choenni, Rochelle, et al.
Published: (2023)
Information availability in different languages and various technological constraints related to multilinguism on the Internet
by: Khosla, Sonal, et al.
Published: (2025)
by: Khosla, Sonal, et al.
Published: (2025)
Complexity-aware fine-tuning
by: Goncharov, Andrey, et al.
Published: (2025)
by: Goncharov, Andrey, et al.
Published: (2025)
Evolution of meta's llama models and parameter-efficient fine-tuning of large language models: a survey
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
Metaphor identification using large language models: A comparison of RAG, prompt engineering, and fine-tuning
by: Fuoli, Matteo, et al.
Published: (2025)
by: Fuoli, Matteo, et al.
Published: (2025)
Immunization against harmful fine-tuning attacks
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
Temporal fine-tuning for early risk detection
by: Thompson, Horacio, et al.
Published: (2025)
by: Thompson, Horacio, et al.
Published: (2025)
Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead
by: Briva-Iglesias, Vicent
Published: (2026)
by: Briva-Iglesias, Vicent
Published: (2026)
One ruler to measure them all: Benchmarking multilingual long-context language models
by: Kim, Yekyung, et al.
Published: (2025)
by: Kim, Yekyung, et al.
Published: (2025)
Loose and Tight: Creative Formation but Rigid Use of Nominal Compounds in Conspiracist Texts
by: Alessandro Miani, et al.
Published: (2024)
by: Alessandro Miani, et al.
Published: (2024)
Labeling supervised fine-tuning data with the scaling law
by: Kong, Huanjun
Published: (2024)
by: Kong, Huanjun
Published: (2024)
MultiLoKo: a multilingual local knowledge benchmark for LLMs spanning 31 languages
by: Hupkes, Dieuwke, et al.
Published: (2025)
by: Hupkes, Dieuwke, et al.
Published: (2025)
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation
by: Chirkova, Nadezhda, et al.
Published: (2023)
by: Chirkova, Nadezhda, et al.
Published: (2023)
Symbol tuning improves in-context learning in language models
by: Wei, Jerry, et al.
Published: (2023)
by: Wei, Jerry, et al.
Published: (2023)
Similar Items
-
Can language models learn analogical reasoning? Investigating training objectives and comparisons to human performance
by: Petersen, Molly R., et al.
Published: (2023) -
From 124 Million Tokens to 1,021 Neologisms: A Large-Scale Pipeline for Automatic Neologism Detection
by: Rossini, Diego, et al.
Published: (2026) -
Binary Token-Level Classification with DeBERTa for All-Type MWE Identification: A Lightweight Approach with Linguistic Enhancement
by: Rossini, Diego, et al.
Published: (2026) -
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024) -
Modelling Analogies and Analogical Reasoning: Connecting Cognitive Science Theory and NLP Research
by: Petersen, Molly R, et al.
Published: (2025)