A suite of LMs comprehend puzzle statements as well as humans
Fuente:
arXiv
Guardado en:
| Autores principales: | Goldberg, Adele E, Rakshit, Supantho, Hu, Jennifer, Mahowald, Kyle |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
por: Rakshit, Supantho, et al.
Publicado: (2025)
por: Rakshit, Supantho, et al.
Publicado: (2025)
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
por: Weissweiler, Leonie, et al.
Publicado: (2025)
por: Weissweiler, Leonie, et al.
Publicado: (2025)
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies
por: Cuneo, Nicole, et al.
Publicado: (2025)
por: Cuneo, Nicole, et al.
Publicado: (2025)
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
por: Boguraev, Sasha, et al.
Publicado: (2026)
por: Boguraev, Sasha, et al.
Publicado: (2026)
Language models align with human judgments on key grammatical constructions
por: Hu, Jennifer, et al.
Publicado: (2024)
por: Hu, Jennifer, et al.
Publicado: (2024)
Language Models Fail to Introspect About Their Knowledge of Language
por: Song, Siyuan, et al.
Publicado: (2025)
por: Song, Siyuan, et al.
Publicado: (2025)
Privileged Self-Access Matters for Introspection in AI
por: Song, Siyuan, et al.
Publicado: (2025)
por: Song, Siyuan, et al.
Publicado: (2025)
How Linguistics Learned to Stop Worrying and Love the Language Models
por: Futrell, Richard, et al.
Publicado: (2025)
por: Futrell, Richard, et al.
Publicado: (2025)
You Can't Fight in Here! This is BBS!
por: Futrell, Richard, et al.
Publicado: (2026)
por: Futrell, Richard, et al.
Publicado: (2026)
Language Models Learn Rare Phenomena from Less Rare Phenomena: The Case of the Missing AANNs
por: Misra, Kanishka, et al.
Publicado: (2024)
por: Misra, Kanishka, et al.
Publicado: (2024)
For Generated Text, Is NLI-Neutral Text the Best Text?
por: Mersinias, Michail, et al.
Publicado: (2023)
por: Mersinias, Michail, et al.
Publicado: (2023)
Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs
por: Lederman, Harvey, et al.
Publicado: (2024)
por: Lederman, Harvey, et al.
Publicado: (2024)
Emergent Introspection in AI is Content-Agnostic
por: Lederman, Harvey, et al.
Publicado: (2026)
por: Lederman, Harvey, et al.
Publicado: (2026)
The Counterexample Game: Iterated Conceptual Analysis and Repair in Language Models
por: Drucker, Daniel, et al.
Publicado: (2026)
por: Drucker, Daniel, et al.
Publicado: (2026)
France or Spain or Germany or France: A Neural Account of Non-Redundant Redundant Disjunctions
por: Boguraev, Sasha, et al.
Publicado: (2026)
por: Boguraev, Sasha, et al.
Publicado: (2026)
Participle-Prepended Nominals Have Lower Entropy Than Nominals Appended After the Participle
por: Denlinger, Kristie, et al.
Publicado: (2024)
por: Denlinger, Kristie, et al.
Publicado: (2024)
Convergence and Divergence of Language Models under Different Random Seeds
por: Fehlauer, Finlay, et al.
Publicado: (2025)
por: Fehlauer, Finlay, et al.
Publicado: (2025)
Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions
por: Boguraev, Sasha, et al.
Publicado: (2025)
por: Boguraev, Sasha, et al.
Publicado: (2025)
What Can String Probability Tell Us About Grammaticality?
por: Hu, Jennifer, et al.
Publicado: (2025)
por: Hu, Jennifer, et al.
Publicado: (2025)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
por: Misra, Kanishka, et al.
Publicado: (2024)
por: Misra, Kanishka, et al.
Publicado: (2024)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
por: Rozner, Josh, et al.
Publicado: (2021)
por: Rozner, Josh, et al.
Publicado: (2021)
Studies with impossible languages falsify LMs as models of human language
por: Bowers, Jeffrey S., et al.
Publicado: (2025)
por: Bowers, Jeffrey S., et al.
Publicado: (2025)
Constructions are Revealed in Word Distributions
por: Rozner, Joshua, et al.
Publicado: (2025)
por: Rozner, Joshua, et al.
Publicado: (2025)
On Language Models' Sensitivity to Suspicious Coincidences
por: Padmanabhan, Sriram, et al.
Publicado: (2025)
por: Padmanabhan, Sriram, et al.
Publicado: (2025)
Both Direct and Indirect Evidence Contribute to Dative Alternation Preferences in Language Models
por: Yao, Qing, et al.
Publicado: (2025)
por: Yao, Qing, et al.
Publicado: (2025)
semantic-features: A User-Friendly Tool for Studying Contextual Word Embeddings in Interpretable Semantic Spaces
por: Ranganathan, Jwalanthi, et al.
Publicado: (2025)
por: Ranganathan, Jwalanthi, et al.
Publicado: (2025)
Lil-Bevo: Explorations of Strategies for Training Language Models in More Humanlike Ways
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
Learning to vary: Teaching LMs to reproduce human linguistic variability in next-word prediction
por: Groot, Tobias, et al.
Publicado: (2025)
por: Groot, Tobias, et al.
Publicado: (2025)
Models Can and Should Embrace the Communicative Nature of Human-Generated Math
por: Boguraev, Sasha, et al.
Publicado: (2024)
por: Boguraev, Sasha, et al.
Publicado: (2024)
Counterfactual Probing for the Influence of Affect and Specificity on Intergroup Bias
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
Do they mean 'us'? Interpreting Referring Expressions in Intergroup Bias
por: Govindarajan, Venkata S, et al.
Publicado: (2024)
por: Govindarajan, Venkata S, et al.
Publicado: (2024)
Are BabyLMs Second Language Learners?
por: Edman, Lukas, et al.
Publicado: (2024)
por: Edman, Lukas, et al.
Publicado: (2024)
Multilingual Large Language Models do not comprehend all natural languages to equal degrees
por: Moskvina, Natalia, et al.
Publicado: (2026)
por: Moskvina, Natalia, et al.
Publicado: (2026)
Learning Unacceptability: Repeated Exposure to Acceptable Sentences Improves Adult Learners’ Recognition of Unacceptable Sentences
por: Karina Tachihara, et al.
Publicado: (2024)
por: Karina Tachihara, et al.
Publicado: (2024)
How to Make LMs Strong Node Classifiers?
por: Xu, Zhe, et al.
Publicado: (2024)
por: Xu, Zhe, et al.
Publicado: (2024)
Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
por: Prashanth, USVSN Sai, et al.
Publicado: (2024)
por: Prashanth, USVSN Sai, et al.
Publicado: (2024)
Compositional preference models for aligning LMs
por: Go, Dongyoung, et al.
Publicado: (2023)
por: Go, Dongyoung, et al.
Publicado: (2023)
Beyond Pattern Recognition: Probing Mental Representations of LMs
por: Miller, Moritz, et al.
Publicado: (2025)
por: Miller, Moritz, et al.
Publicado: (2025)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
por: Alnuhait, Deema, et al.
Publicado: (2024)
por: Alnuhait, Deema, et al.
Publicado: (2024)
Crafting In-context Examples according to LMs' Parametric Knowledge
por: Lee, Yoonsang, et al.
Publicado: (2023)
por: Lee, Yoonsang, et al.
Publicado: (2023)
Ejemplares similares
-
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
por: Rakshit, Supantho, et al.
Publicado: (2025) -
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
por: Weissweiler, Leonie, et al.
Publicado: (2025) -
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies
por: Cuneo, Nicole, et al.
Publicado: (2025) -
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
por: Boguraev, Sasha, et al.
Publicado: (2026) -
Language models align with human judgments on key grammatical constructions
por: Hu, Jennifer, et al.
Publicado: (2024)