What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty
Fuente:
arXiv
Saved in:
| Main Authors: | Martins, Jonas Mayer, Huang, Zhuojing, Herygers, Aaricia, Beinborn, Lisa |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vocabulary shapes cross-lingual variation of word-order learnability in language models
by: Martins, Jonas Mayer, et al.
Published: (2026)
by: Martins, Jonas Mayer, et al.
Published: (2026)
Deep learning models for representing out-of-vocabulary words
by: Lochter, Johannes V., et al.
Published: (2020)
by: Lochter, Johannes V., et al.
Published: (2020)
Once Upon a Time: Interactive Learning for Storytelling with Small Language Models
by: Martins, Jonas Mayer, et al.
Published: (2025)
by: Martins, Jonas Mayer, et al.
Published: (2025)
Round and Round We Go! What makes Rotary Positional Encodings useful?
by: Barbero, Federico, et al.
Published: (2024)
by: Barbero, Federico, et al.
Published: (2024)
Exploiting the English Vocabulary Profile for L2 word-level vocabulary assessment with LLMs
by: Bannò, Stefano, et al.
Published: (2025)
by: Bannò, Stefano, et al.
Published: (2025)
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
OmniDraft: A Cross-vocabulary, Online Adaptive Drafter for On-device Speculative Decoding
by: Ramakrishnan, Ramchalam Kinattinkara, et al.
Published: (2025)
by: Ramakrishnan, Ramchalam Kinattinkara, et al.
Published: (2025)
Multi-word Tokenization for Sequence Compression
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
CroissantLLM: A Truly Bilingual French-English Language Model
by: Faysse, Manuel, et al.
Published: (2024)
by: Faysse, Manuel, et al.
Published: (2024)
Do Multilingual LLMs Think In English?
by: Schut, Lisa, et al.
Published: (2025)
by: Schut, Lisa, et al.
Published: (2025)
What Large Language Models Know and What People Think They Know
by: Steyvers, Mark, et al.
Published: (2024)
by: Steyvers, Mark, et al.
Published: (2024)
Critical biblical studies via word frequency analysis: unveiling text authorship
by: Faigenbaum-Golovin, Shira, et al.
Published: (2024)
by: Faigenbaum-Golovin, Shira, et al.
Published: (2024)
Why is prompting hard? Understanding prompts on binary sequence predictors
by: Wenliang, Li Kevin, et al.
Published: (2025)
by: Wenliang, Li Kevin, et al.
Published: (2025)
A meta-analysis on the performance of machine-learning based language models for sentiment analysis
by: Rohde, Elena, et al.
Published: (2025)
by: Rohde, Elena, et al.
Published: (2025)
What's the Difference? Supporting Users in Identifying the Effects of Prompt and Model Changes Through Token Patterns
by: Hedderich, Michael A., et al.
Published: (2025)
by: Hedderich, Michael A., et al.
Published: (2025)
Assessing the validity of new paradigmatic complexity measures as criterial features for proficiency in L2 writings in English
by: Mallart, Cyriel, et al.
Published: (2025)
by: Mallart, Cyriel, et al.
Published: (2025)
Efficient Parallel Samplers for Recurrent-Depth Models and Their Connection to Diffusion Language Models
by: Geiping, Jonas, et al.
Published: (2025)
by: Geiping, Jonas, et al.
Published: (2025)
Do "English" Named Entity Recognizers Work Well on Global Englishes?
by: Shan, Alexander, et al.
Published: (2024)
by: Shan, Alexander, et al.
Published: (2024)
Learning from Sufficient Rationales: Analysing the Relationship Between Explanation Faithfulness and Token-level Regularisation Strategies
by: Kamp, Jonathan, et al.
Published: (2025)
by: Kamp, Jonathan, et al.
Published: (2025)
The Role of Syntactic Span Preferences in Post-Hoc Explanation Disagreement
by: Kamp, Jonathan, et al.
Published: (2024)
by: Kamp, Jonathan, et al.
Published: (2024)
What Evidence Do Language Models Find Convincing?
by: Wan, Alexander, et al.
Published: (2024)
by: Wan, Alexander, et al.
Published: (2024)
What is Wrong with Perplexity for Long-context Language Modeling?
by: Fang, Lizhe, et al.
Published: (2024)
by: Fang, Lizhe, et al.
Published: (2024)
Better To Ask in English? Evaluating Factual Accuracy of Multilingual LLMs in English and Low-Resource Languages
by: Rohera, Pritika, et al.
Published: (2025)
by: Rohera, Pritika, et al.
Published: (2025)
CueBuddy: helping non-native English speakers navigate English-centric STEM education
by: Gupta, Pranav
Published: (2025)
by: Gupta, Pranav
Published: (2025)
Evaluating Machine Translation Models for English-Hindi Language Pairs: A Comparative Analysis
by: Shetty, Ahan Prasannakumar
Published: (2025)
by: Shetty, Ahan Prasannakumar
Published: (2025)
What Will My Model Forget? Forecasting Forgotten Examples in Language Model Refinement
by: Jin, Xisen, et al.
Published: (2024)
by: Jin, Xisen, et al.
Published: (2024)
From communities to interpretable network and word embedding: an unified approach
by: Prouteau, Thibault, et al.
Published: (2024)
by: Prouteau, Thibault, et al.
Published: (2024)
Traditional Readability Formulas Compared for English
by: Lee, Bruce W., et al.
Published: (2023)
by: Lee, Bruce W., et al.
Published: (2023)
Few-shot learning for automated content analysis: Efficient coding of arguments and claims in the debate on arms deliveries to Ukraine
by: Rieger, Jonas, et al.
Published: (2023)
by: Rieger, Jonas, et al.
Published: (2023)
Towards Understanding What Code Language Models Learned
by: Ahmed, Toufique, et al.
Published: (2023)
by: Ahmed, Toufique, et al.
Published: (2023)
What Do Language Models Learn in Context? The Structured Task Hypothesis
by: Li, Jiaoda, et al.
Published: (2024)
by: Li, Jiaoda, et al.
Published: (2024)
SoK: Machine Learning for Misinformation Detection
by: Xiao, Madelyne, et al.
Published: (2023)
by: Xiao, Madelyne, et al.
Published: (2023)
Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs
by: Su, Guinan, et al.
Published: (2026)
by: Su, Guinan, et al.
Published: (2026)
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
by: Dang, Quy-Anh, et al.
Published: (2025)
by: Dang, Quy-Anh, et al.
Published: (2025)
Derailing Non-Answers via Logit Suppression at Output Subspace Boundaries in RLHF-Aligned Language Models
by: Dam, Harvey, et al.
Published: (2025)
by: Dam, Harvey, et al.
Published: (2025)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
by: Kimera, Richard, et al.
Published: (2024)
by: Kimera, Richard, et al.
Published: (2024)
ParaScopes: What do Language Models Activations Encode About Future Text?
by: Pochinkov, Nicky, et al.
Published: (2025)
by: Pochinkov, Nicky, et al.
Published: (2025)
What Matters in LLM-generated Data: Diversity and Its Effect on Model Fine-Tuning
by: Zhu, Yuchang, et al.
Published: (2025)
by: Zhu, Yuchang, et al.
Published: (2025)
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
by: MacPhail, Dorothea, et al.
Published: (2024)
by: MacPhail, Dorothea, et al.
Published: (2024)
What to align in multimodal contrastive learning?
by: Dufumier, Benoit, et al.
Published: (2024)
by: Dufumier, Benoit, et al.
Published: (2024)
Similar Items
-
Vocabulary shapes cross-lingual variation of word-order learnability in language models
by: Martins, Jonas Mayer, et al.
Published: (2026) -
Deep learning models for representing out-of-vocabulary words
by: Lochter, Johannes V., et al.
Published: (2020) -
Once Upon a Time: Interactive Learning for Storytelling with Small Language Models
by: Martins, Jonas Mayer, et al.
Published: (2025) -
Round and Round We Go! What makes Rotary Positional Encodings useful?
by: Barbero, Federico, et al.
Published: (2024) -
Exploiting the English Vocabulary Profile for L2 word-level vocabulary assessment with LLMs
by: Bannò, Stefano, et al.
Published: (2025)