A suite of LMs comprehend puzzle statements as well as humans
Fuente:
arXiv
Saved in:
| Main Authors: | Goldberg, Adele E, Rakshit, Supantho, Hu, Jennifer, Mahowald, Kyle |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
by: Rakshit, Supantho, et al.
Published: (2025)
by: Rakshit, Supantho, et al.
Published: (2025)
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
by: Weissweiler, Leonie, et al.
Published: (2025)
by: Weissweiler, Leonie, et al.
Published: (2025)
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies
by: Cuneo, Nicole, et al.
Published: (2025)
by: Cuneo, Nicole, et al.
Published: (2025)
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
by: Boguraev, Sasha, et al.
Published: (2026)
by: Boguraev, Sasha, et al.
Published: (2026)
Language models align with human judgments on key grammatical constructions
by: Hu, Jennifer, et al.
Published: (2024)
by: Hu, Jennifer, et al.
Published: (2024)
Language Models Fail to Introspect About Their Knowledge of Language
by: Song, Siyuan, et al.
Published: (2025)
by: Song, Siyuan, et al.
Published: (2025)
Privileged Self-Access Matters for Introspection in AI
by: Song, Siyuan, et al.
Published: (2025)
by: Song, Siyuan, et al.
Published: (2025)
How Linguistics Learned to Stop Worrying and Love the Language Models
by: Futrell, Richard, et al.
Published: (2025)
by: Futrell, Richard, et al.
Published: (2025)
You Can't Fight in Here! This is BBS!
by: Futrell, Richard, et al.
Published: (2026)
by: Futrell, Richard, et al.
Published: (2026)
Language Models Learn Rare Phenomena from Less Rare Phenomena: The Case of the Missing AANNs
by: Misra, Kanishka, et al.
Published: (2024)
by: Misra, Kanishka, et al.
Published: (2024)
For Generated Text, Is NLI-Neutral Text the Best Text?
by: Mersinias, Michail, et al.
Published: (2023)
by: Mersinias, Michail, et al.
Published: (2023)
Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs
by: Lederman, Harvey, et al.
Published: (2024)
by: Lederman, Harvey, et al.
Published: (2024)
Emergent Introspection in AI is Content-Agnostic
by: Lederman, Harvey, et al.
Published: (2026)
by: Lederman, Harvey, et al.
Published: (2026)
The Counterexample Game: Iterated Conceptual Analysis and Repair in Language Models
by: Drucker, Daniel, et al.
Published: (2026)
by: Drucker, Daniel, et al.
Published: (2026)
France or Spain or Germany or France: A Neural Account of Non-Redundant Redundant Disjunctions
by: Boguraev, Sasha, et al.
Published: (2026)
by: Boguraev, Sasha, et al.
Published: (2026)
Participle-Prepended Nominals Have Lower Entropy Than Nominals Appended After the Participle
by: Denlinger, Kristie, et al.
Published: (2024)
by: Denlinger, Kristie, et al.
Published: (2024)
Convergence and Divergence of Language Models under Different Random Seeds
by: Fehlauer, Finlay, et al.
Published: (2025)
by: Fehlauer, Finlay, et al.
Published: (2025)
Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions
by: Boguraev, Sasha, et al.
Published: (2025)
by: Boguraev, Sasha, et al.
Published: (2025)
What Can String Probability Tell Us About Grammaticality?
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
by: Misra, Kanishka, et al.
Published: (2024)
by: Misra, Kanishka, et al.
Published: (2024)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
by: Rozner, Josh, et al.
Published: (2021)
by: Rozner, Josh, et al.
Published: (2021)
Studies with impossible languages falsify LMs as models of human language
by: Bowers, Jeffrey S., et al.
Published: (2025)
by: Bowers, Jeffrey S., et al.
Published: (2025)
Constructions are Revealed in Word Distributions
by: Rozner, Joshua, et al.
Published: (2025)
by: Rozner, Joshua, et al.
Published: (2025)
On Language Models' Sensitivity to Suspicious Coincidences
by: Padmanabhan, Sriram, et al.
Published: (2025)
by: Padmanabhan, Sriram, et al.
Published: (2025)
Both Direct and Indirect Evidence Contribute to Dative Alternation Preferences in Language Models
by: Yao, Qing, et al.
Published: (2025)
by: Yao, Qing, et al.
Published: (2025)
semantic-features: A User-Friendly Tool for Studying Contextual Word Embeddings in Interpretable Semantic Spaces
by: Ranganathan, Jwalanthi, et al.
Published: (2025)
by: Ranganathan, Jwalanthi, et al.
Published: (2025)
Lil-Bevo: Explorations of Strategies for Training Language Models in More Humanlike Ways
by: Govindarajan, Venkata S, et al.
Published: (2023)
by: Govindarajan, Venkata S, et al.
Published: (2023)
Models Can and Should Embrace the Communicative Nature of Human-Generated Math
by: Boguraev, Sasha, et al.
Published: (2024)
by: Boguraev, Sasha, et al.
Published: (2024)
Learning to vary: Teaching LMs to reproduce human linguistic variability in next-word prediction
by: Groot, Tobias, et al.
Published: (2025)
by: Groot, Tobias, et al.
Published: (2025)
Counterfactual Probing for the Influence of Affect and Specificity on Intergroup Bias
by: Govindarajan, Venkata S, et al.
Published: (2023)
by: Govindarajan, Venkata S, et al.
Published: (2023)
Multilingual Large Language Models do not comprehend all natural languages to equal degrees
by: Moskvina, Natalia, et al.
Published: (2026)
by: Moskvina, Natalia, et al.
Published: (2026)
Do they mean 'us'? Interpreting Referring Expressions in Intergroup Bias
by: Govindarajan, Venkata S, et al.
Published: (2024)
by: Govindarajan, Venkata S, et al.
Published: (2024)
Are BabyLMs Second Language Learners?
by: Edman, Lukas, et al.
Published: (2024)
by: Edman, Lukas, et al.
Published: (2024)
Learning Unacceptability: Repeated Exposure to Acceptable Sentences Improves Adult Learners’ Recognition of Unacceptable Sentences
by: Karina Tachihara, et al.
Published: (2024)
by: Karina Tachihara, et al.
Published: (2024)
Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
by: Prashanth, USVSN Sai, et al.
Published: (2024)
by: Prashanth, USVSN Sai, et al.
Published: (2024)
How to Make LMs Strong Node Classifiers?
by: Xu, Zhe, et al.
Published: (2024)
by: Xu, Zhe, et al.
Published: (2024)
Compositional preference models for aligning LMs
by: Go, Dongyoung, et al.
Published: (2023)
by: Go, Dongyoung, et al.
Published: (2023)
Beyond Pattern Recognition: Probing Mental Representations of LMs
by: Miller, Moritz, et al.
Published: (2025)
by: Miller, Moritz, et al.
Published: (2025)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
by: Alnuhait, Deema, et al.
Published: (2024)
by: Alnuhait, Deema, et al.
Published: (2024)
Crafting In-context Examples according to LMs' Parametric Knowledge
by: Lee, Yoonsang, et al.
Published: (2023)
by: Lee, Yoonsang, et al.
Published: (2023)
Similar Items
-
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
by: Rakshit, Supantho, et al.
Published: (2025) -
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
by: Weissweiler, Leonie, et al.
Published: (2025) -
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies
by: Cuneo, Nicole, et al.
Published: (2025) -
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
by: Boguraev, Sasha, et al.
Published: (2026) -
Language models align with human judgments on key grammatical constructions
by: Hu, Jennifer, et al.
Published: (2024)