Modeling the human lexicon under temperature variations: linguistic factors, diversity and typicality in LLM word associations
Fuente:
arXiv
Saved in:
| Main Authors: | Rodriguez, Maria Andueza, Candito, Marie, Huyghe, Richard |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Auxiliary Tasks to Boost Biaffine Semantic Dependency Parsing
by: Candito, Marie
Published: (2024)
by: Candito, Marie
Published: (2024)
In the LLM era, Word Sense Induction remains unsolved
by: Mosolova, Anna, et al.
Published: (2026)
by: Mosolova, Anna, et al.
Published: (2026)
Probing structural constraints of negation in Pretrained Language Models
by: Kletz, David, et al.
Published: (2024)
by: Kletz, David, et al.
Published: (2024)
Learning to vary: Teaching LMs to reproduce human linguistic variability in next-word prediction
by: Groot, Tobias, et al.
Published: (2025)
by: Groot, Tobias, et al.
Published: (2025)
Target word activity detector: An approach to obtain ASR word boundaries without lexicon
by: Sivasankaran, Sunit, et al.
Published: (2024)
by: Sivasankaran, Sunit, et al.
Published: (2024)
Injecting Wiktionary to improve token-level contextual representations using contrastive learning
by: Mosolova, Anna, et al.
Published: (2024)
by: Mosolova, Anna, et al.
Published: (2024)
The Self-Contained Negation Test Set
by: Kletz, David, et al.
Published: (2024)
by: Kletz, David, et al.
Published: (2024)
Towards a resource for multilingual lexicons: an MT assisted and human-in-the-loop multilingual parallel corpus with multi-word expression annotation
by: Han, Lifeng, et al.
Published: (2020)
by: Han, Lifeng, et al.
Published: (2020)
How desirable is alignment between LLMs and linguistically diverse human users?
by: Knoeferle, Pia, et al.
Published: (2025)
by: Knoeferle, Pia, et al.
Published: (2025)
Understanding the effects of word-level linguistic annotations in under-resourced neural machine translation
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
Playing with words: Comparing the vocabulary and lexical diversity of ChatGPT and humans
by: Reviriego, Pedro, et al.
Published: (2023)
by: Reviriego, Pedro, et al.
Published: (2023)
French parsing enhanced with a word clustering method based on a syntactic lexicon
by: Sigogne, Anthony, et al.
Published: (2026)
by: Sigogne, Anthony, et al.
Published: (2026)
Advancements and limitations of LLMs in replicating human color-word associations
by: Fukushima, Makoto, et al.
Published: (2024)
by: Fukushima, Makoto, et al.
Published: (2024)
Quantification and object perception in Multimodal Large Language Models and human linguistic cognition
by: Montero, Raquel, et al.
Published: (2025)
by: Montero, Raquel, et al.
Published: (2025)
A multimodal multiplex of the mental lexicon for multilingual individuals
by: Huynh, Maria, et al.
Published: (2025)
by: Huynh, Maria, et al.
Published: (2025)
A word association network methodology for evaluating implicit biases in LLMs compared to humans
by: Abramski, Katherine, et al.
Published: (2025)
by: Abramski, Katherine, et al.
Published: (2025)
Evolution of the lexicon: a probabilistic point of view
by: Serva, Maurizio
Published: (2025)
by: Serva, Maurizio
Published: (2025)
Identifying the sources of ideological bias in GPT models through linguistic variation in output
by: Walker, Christina, et al.
Published: (2024)
by: Walker, Christina, et al.
Published: (2024)
Does GPT-4 surpass human performance in linguistic pragmatics?
by: Bojic, Ljubisa, et al.
Published: (2023)
by: Bojic, Ljubisa, et al.
Published: (2023)
Towards a theory of morphology-driven marking in the lexicon: The case of the state
by: Idrissi, Mohamed El
Published: (2026)
by: Idrissi, Mohamed El
Published: (2026)
A lexicon obtained and validated by a data-driven approach for organic residues valorization in emerging and developing countries
by: Rakotomalala, Christiane, et al.
Published: (2024)
by: Rakotomalala, Christiane, et al.
Published: (2024)
Prompt and circumstance: A word-by-word LLM prompting approach to interlinear glossing for low-resource languages
by: Elsner, Micha, et al.
Published: (2025)
by: Elsner, Micha, et al.
Published: (2025)
Inducing lexicons of in-group language with socio-temporal context
by: de Kock, Christine
Published: (2024)
by: de Kock, Christine
Published: (2024)
The truth is no diaper: Human and AI-generated associations to emotional words
by: Vintar, Špela, et al.
Published: (2025)
by: Vintar, Špela, et al.
Published: (2025)
Revisiting speech segmentation and lexicon learning with better features
by: Kamper, Herman, et al.
Published: (2024)
by: Kamper, Herman, et al.
Published: (2024)
Swap distance minimization beyond entropy minimization in word order variation
by: Franco-Sánchez, Víctor, et al.
Published: (2024)
by: Franco-Sánchez, Víctor, et al.
Published: (2024)
How communicatively optimal are exact numeral systems? Once more on lexicon size and morphosyntactic complexity
by: Cathcart, Chundra, et al.
Published: (2026)
by: Cathcart, Chundra, et al.
Published: (2026)
BenCSSmark: Making the Social Sciences Count in LLM Research
by: Chatelain, Arnault, et al.
Published: (2026)
by: Chatelain, Arnault, et al.
Published: (2026)
Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations
by: Baroni, Marco, et al.
Published: (2026)
by: Baroni, Marco, et al.
Published: (2026)
Investigating the structure of emotions by analyzing similarity and association of emotion words
by: Iwaki, Fumitaka, et al.
Published: (2026)
by: Iwaki, Fumitaka, et al.
Published: (2026)
Syntax, semantics, and the lexicon
Published: (2025)
Published: (2025)
Effects of diversity incentives on sample diversity and downstream model performance in LLM-based text augmentation
by: Cegin, Jan, et al.
Published: (2024)
by: Cegin, Jan, et al.
Published: (2024)
Is linguistically-motivated data augmentation worth it?
by: Groshan, Ray, et al.
Published: (2025)
by: Groshan, Ray, et al.
Published: (2025)
Is it the end of (generative) linguistics as we know it?
by: Chesi, Cristiano
Published: (2024)
by: Chesi, Cristiano
Published: (2024)
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
What is a word?
by: Murphy, Elliot
Published: (2024)
by: Murphy, Elliot
Published: (2024)
Fin-Bias: Comprehensive Evaluation for LLM Decision-Making under human bias in Finance Domain
by: Hu, Xiaoyu, et al.
Published: (2026)
by: Hu, Xiaoyu, et al.
Published: (2026)
Vocabulary shapes cross-lingual variation of word-order learnability in language models
by: Martins, Jonas Mayer, et al.
Published: (2026)
by: Martins, Jonas Mayer, et al.
Published: (2026)
Do LLMs produce texts with "human-like" lexical diversity?
by: Kendro, Kelly, et al.
Published: (2025)
by: Kendro, Kelly, et al.
Published: (2025)
Unsupervised lexicon learning from speech is limited by representations rather than clustering
by: Slabbert, Danel, et al.
Published: (2025)
by: Slabbert, Danel, et al.
Published: (2025)
Similar Items
-
Auxiliary Tasks to Boost Biaffine Semantic Dependency Parsing
by: Candito, Marie
Published: (2024) -
In the LLM era, Word Sense Induction remains unsolved
by: Mosolova, Anna, et al.
Published: (2026) -
Probing structural constraints of negation in Pretrained Language Models
by: Kletz, David, et al.
Published: (2024) -
Learning to vary: Teaching LMs to reproduce human linguistic variability in next-word prediction
by: Groot, Tobias, et al.
Published: (2025) -
Target word activity detector: An approach to obtain ASR word boundaries without lexicon
by: Sivasankaran, Sunit, et al.
Published: (2024)