Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Timkey, William, Dillon, Brian, Linzen, Tal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deconstructing sentence disambiguation by joint latent modeling of reading paradigms: LLM surprisal is not enough
von: Paape, Dario, et al.
Veröffentlicht: (2026)
von: Paape, Dario, et al.
Veröffentlicht: (2026)
Manipulating language models' training data to study syntactic constraint learning: the case of English passivization
von: Leong, Cara Su-Yi, et al.
Veröffentlicht: (2024)
von: Leong, Cara Su-Yi, et al.
Veröffentlicht: (2024)
Do Language Models' Words Refer?
von: Mandelkern, Matthew, et al.
Veröffentlicht: (2023)
von: Mandelkern, Matthew, et al.
Veröffentlicht: (2023)
SPAWNing Structural Priming Predictions from a Cognitively Motivated Parser
von: Prasad, Grusha, et al.
Veröffentlicht: (2024)
von: Prasad, Grusha, et al.
Veröffentlicht: (2024)
Entailment Semantics Can Be Extracted from an Ideal Language Model
von: Merrill, William, et al.
Veröffentlicht: (2022)
von: Merrill, William, et al.
Veröffentlicht: (2022)
Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction
von: Petty, Jackson, et al.
Veröffentlicht: (2026)
von: Petty, Jackson, et al.
Veröffentlicht: (2026)
Can You Learn Semantics Through Next-Word Prediction? The Case of Entailment
von: Merrill, William, et al.
Veröffentlicht: (2024)
von: Merrill, William, et al.
Veröffentlicht: (2024)
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2024)
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2024)
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax
von: Mueller, Aaron, et al.
Veröffentlicht: (2023)
von: Mueller, Aaron, et al.
Veröffentlicht: (2023)
Multilingual Prompting for Improving LLM Generation Diversity
von: Wang, Qihan, et al.
Veröffentlicht: (2025)
von: Wang, Qihan, et al.
Veröffentlicht: (2025)
How Does Code Pretraining Affect Language Model Task Performance?
von: Petty, Jackson, et al.
Veröffentlicht: (2024)
von: Petty, Jackson, et al.
Veröffentlicht: (2024)
Post-training makes large language models less human-like
von: Binz, Marcel, et al.
Veröffentlicht: (2026)
von: Binz, Marcel, et al.
Veröffentlicht: (2026)
Decomposition of surprisal: Unified computational model of ERP components in language processing
von: Li, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Li, Jiaxuan, et al.
Veröffentlicht: (2024)
Emergence of Linear Truth Encodings in Language Models
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2025)
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2025)
RELIC: Evaluating Complex Reasoning via the Recognition of Languages In-Context
von: Petty, Jackson, et al.
Veröffentlicht: (2025)
von: Petty, Jackson, et al.
Veröffentlicht: (2025)
Language Models Struggle to Use Representations Learned In-Context
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
Tiny language models
von: Gross, Ronit D., et al.
Veröffentlicht: (2025)
von: Gross, Ronit D., et al.
Veröffentlicht: (2025)
Rapid Word Learning Through Meta In-Context Learning
von: Wang, Wentao, et al.
Veröffentlicht: (2025)
von: Wang, Wentao, et al.
Veröffentlicht: (2025)
Why Are Parsing Actions for Understanding Message Hierarchies Not Random?
von: Kato, Daichi, et al.
Veröffentlicht: (2025)
von: Kato, Daichi, et al.
Veröffentlicht: (2025)
Why do language models perform worse for morphologically complex languages?
von: Arnett, Catherine, et al.
Veröffentlicht: (2024)
von: Arnett, Catherine, et al.
Veröffentlicht: (2024)
Why transformers are obviously good models of language
von: Hill, Felix
Veröffentlicht: (2024)
von: Hill, Felix
Veröffentlicht: (2024)
The Impact of Depth on Compositional Generalization in Transformer Language Models
von: Petty, Jackson, et al.
Veröffentlicht: (2023)
von: Petty, Jackson, et al.
Veröffentlicht: (2023)
Subword models struggle with word learning, but surprisal hides it
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2025)
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2025)
Do large language models resemble humans in language use?
von: Cai, Zhenguang G., et al.
Veröffentlicht: (2023)
von: Cai, Zhenguang G., et al.
Veröffentlicht: (2023)
Studies with impossible languages falsify LMs as models of human language
von: Bowers, Jeffrey S., et al.
Veröffentlicht: (2025)
von: Bowers, Jeffrey S., et al.
Veröffentlicht: (2025)
Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time
von: Hu, Michael Y., et al.
Veröffentlicht: (2026)
von: Hu, Michael Y., et al.
Veröffentlicht: (2026)
Hypothesis Testing for Quantifying LLM-Human Misalignment in Multiple Choice Settings
von: Hong, Harbin, et al.
Veröffentlicht: (2025)
von: Hong, Harbin, et al.
Veröffentlicht: (2025)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
Multilingual large language models leak human stereotypes across language boundaries
von: Cao, Yang Trista, et al.
Veröffentlicht: (2023)
von: Cao, Yang Trista, et al.
Veröffentlicht: (2023)
Aligning language models with human preferences
von: Korbak, Tomasz
Veröffentlicht: (2024)
von: Korbak, Tomasz
Veröffentlicht: (2024)
Training language models to be warm and empathetic makes them less reliable and more sycophantic
von: Ibrahim, Lujain, et al.
Veröffentlicht: (2025)
von: Ibrahim, Lujain, et al.
Veröffentlicht: (2025)
Prefix Parsing is Just Parsing
von: Pasti, Clemente, et al.
Veröffentlicht: (2026)
von: Pasti, Clemente, et al.
Veröffentlicht: (2026)
Are they human? Detecting large language models by probing human memory constraints
von: Schug, Simon, et al.
Veröffentlicht: (2026)
von: Schug, Simon, et al.
Veröffentlicht: (2026)
A Systematic Comparison of Syllogistic Reasoning in Humans and Language Models
von: Eisape, Tiwalayo, et al.
Veröffentlicht: (2023)
von: Eisape, Tiwalayo, et al.
Veröffentlicht: (2023)
Testing the Deliteralization Hypothesis in Human and Machine Translation
von: Marmonier, Malik, et al.
Veröffentlicht: (2026)
von: Marmonier, Malik, et al.
Veröffentlicht: (2026)
Temperature-scaling surprisal estimates improve fit to human reading times -- but does it do so for the "right reasons"?
von: Liu, Tong, et al.
Veröffentlicht: (2023)
von: Liu, Tong, et al.
Veröffentlicht: (2023)
A surprisal oracle for when every layer counts
von: Hong, Xudong, et al.
Veröffentlicht: (2024)
von: Hong, Xudong, et al.
Veröffentlicht: (2024)
Exploring the Knowledge Mismatch Hypothesis: Hallucination Propensity in Small Models Fine-tuned on Data from Larger Models
von: Wee, Phil, et al.
Veröffentlicht: (2024)
von: Wee, Phil, et al.
Veröffentlicht: (2024)
Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models
von: Occhipinti, Daniela, et al.
Veröffentlicht: (2024)
von: Occhipinti, Daniela, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Deconstructing sentence disambiguation by joint latent modeling of reading paradigms: LLM surprisal is not enough
von: Paape, Dario, et al.
Veröffentlicht: (2026) -
Manipulating language models' training data to study syntactic constraint learning: the case of English passivization
von: Leong, Cara Su-Yi, et al.
Veröffentlicht: (2024) -
Do Language Models' Words Refer?
von: Mandelkern, Matthew, et al.
Veröffentlicht: (2023) -
SPAWNing Structural Priming Predictions from a Cognitively Motivated Parser
von: Prasad, Grusha, et al.
Veröffentlicht: (2024) -
Entailment Semantics Can Be Extracted from an Ideal Language Model
von: Merrill, William, et al.
Veröffentlicht: (2022)