Looking forward: Linguistic theory and methods
Fuente:
arXiv
Saved in:
| Main Authors: | Mansfield, John, Wilcox, Ethan Gotlieb |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Predicting the Emergence of Induction Heads in Language Model Pretraining
by: Aoyama, Tatsuya, et al.
Published: (2025)
by: Aoyama, Tatsuya, et al.
Published: (2025)
Information-Theoretic Storage Cost in Sentence Comprehension
by: Kajikawa, Kohei, et al.
Published: (2026)
by: Kajikawa, Kohei, et al.
Published: (2026)
Function Words as Statistical Cues for Language Learning
by: Yang, Xiulin, et al.
Published: (2026)
by: Yang, Xiulin, et al.
Published: (2026)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
A Unified Assessment of the Poverty of the Stimulus Argument for Neural Language Models
by: Yang, Xiulin, et al.
Published: (2026)
by: Yang, Xiulin, et al.
Published: (2026)
On the Role of Context in Reading Time Prediction
by: Opedal, Andreas, et al.
Published: (2024)
by: Opedal, Andreas, et al.
Published: (2024)
Modeling Bottom-up Information Quality during Language Processing
by: Ding, Cui, et al.
Published: (2025)
by: Ding, Cui, et al.
Published: (2025)
Testing the Predictions of Surprisal Theory in 11 Languages
by: Wilcox, Ethan Gotlieb, et al.
Published: (2023)
by: Wilcox, Ethan Gotlieb, et al.
Published: (2023)
What Can String Probability Tell Us About Grammaticality?
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Using Information Theory to Characterize Prosodic Typology: The Case of Tone, Pitch-Accent and Stress-Accent
by: Wilcox, Ethan Gotlieb, et al.
Published: (2025)
by: Wilcox, Ethan Gotlieb, et al.
Published: (2025)
Reverse-Engineering the Reader
by: Kiegeland, Samuel, et al.
Published: (2024)
by: Kiegeland, Samuel, et al.
Published: (2024)
Language Models Grow Less Humanlike beyond Phase Transition
by: Aoyama, Tatsuya, et al.
Published: (2025)
by: Aoyama, Tatsuya, et al.
Published: (2025)
Linguistic Diversification and Rates of Change: Insights From a Diverse Sample of Sociolinguistic Studies
by: John Mansfield
Published: (2025)
by: John Mansfield
Published: (2025)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
by: Hu, Michael Y., et al.
Published: (2024)
by: Hu, Michael Y., et al.
Published: (2024)
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
by: Choshen, Leshem, et al.
Published: (2026)
by: Choshen, Leshem, et al.
Published: (2026)
Anything Goes? A Crosslinguistic Study of (Im)possible Language Learning in LMs
by: Yang, Xiulin, et al.
Published: (2025)
by: Yang, Xiulin, et al.
Published: (2025)
Unpacking Let Alone: Human-Scale Models Generalize to a Rare Construction in Form but not Meaning
by: Scivetti, Wesley, et al.
Published: (2025)
by: Scivetti, Wesley, et al.
Published: (2025)
What Do Prosody and Text Convey? Characterizing How Meaningful Information is Distributed Across Multiple Channels
by: Yadavalli, Aditya, et al.
Published: (2025)
by: Yadavalli, Aditya, et al.
Published: (2025)
On the Efficacy of Sampling Adapters
by: Meister, Clara, et al.
Published: (2023)
by: Meister, Clara, et al.
Published: (2023)
Visual Merit or Linguistic Crutch? A Close Look at DeepSeek-OCR
by: Liang, Yunhao, et al.
Published: (2026)
by: Liang, Yunhao, et al.
Published: (2026)
Surprise! Uniform Information Density Isn't the Whole Story: Predicting Surprisal Contours in Long-form Discourse
by: Tsipidi, Eleftheria, et al.
Published: (2024)
by: Tsipidi, Eleftheria, et al.
Published: (2024)
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions
by: Scivetti, Wesley, et al.
Published: (2026)
by: Scivetti, Wesley, et al.
Published: (2026)
Looking Inward: Language Models Can Learn About Themselves by Introspection
by: Binder, Felix J, et al.
Published: (2024)
by: Binder, Felix J, et al.
Published: (2024)
The Harmonic Structure of Information Contours
by: Tsipidi, Eleftheria, et al.
Published: (2025)
by: Tsipidi, Eleftheria, et al.
Published: (2025)
modeLing: A Novel Dataset for Testing Linguistic Reasoning in Language Models
by: Chi, Nathan A., et al.
Published: (2024)
by: Chi, Nathan A., et al.
Published: (2024)
LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages
by: Bean, Andrew M., et al.
Published: (2024)
by: Bean, Andrew M., et al.
Published: (2024)
Almost Clinical: Linguistic properties of synthetic electronic health records
by: Sharoff, Serge, et al.
Published: (2026)
by: Sharoff, Serge, et al.
Published: (2026)
Agent-Driven Corpus Linguistics: A Framework for Autonomous Linguistic Discovery
by: Yu, Jia, et al.
Published: (2026)
by: Yu, Jia, et al.
Published: (2026)
Linguistic Minimal Pairs Elicit Linguistic Similarity in Large Language Models
by: Zhou, Xinyu, et al.
Published: (2024)
by: Zhou, Xinyu, et al.
Published: (2024)
Hire a Linguist!: Learning Endangered Languages with In-Context Linguistic Descriptions
by: Zhang, Kexun, et al.
Published: (2024)
by: Zhang, Kexun, et al.
Published: (2024)
Logical Computational Linguistics
by: Morrill, Glyn V., et al.
Published: (2026)
by: Morrill, Glyn V., et al.
Published: (2026)
Quantum-Like Contextuality in Large Language Models
by: Lo, Kin Ian, et al.
Published: (2024)
by: Lo, Kin Ian, et al.
Published: (2024)
Developments in Sheaf-Theoretic Models of Natural Language Ambiguities
by: Lo, Kin Ian, et al.
Published: (2024)
by: Lo, Kin Ian, et al.
Published: (2024)
Retrieval-Augmented Linguistic Calibration
by: Yeh, Yi-Fan, et al.
Published: (2026)
by: Yeh, Yi-Fan, et al.
Published: (2026)
Linguistically-Controlled Paraphrase Generation
by: Elgaar, Mohamed, et al.
Published: (2024)
by: Elgaar, Mohamed, et al.
Published: (2024)
[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
by: Choshen, Leshem, et al.
Published: (2024)
by: Choshen, Leshem, et al.
Published: (2024)
Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
by: Warstadt, Alex, et al.
Published: (2025)
by: Warstadt, Alex, et al.
Published: (2025)
Limited-Resource Adapters Are Regularizers, Not Linguists
by: Fekete, Marcell, et al.
Published: (2025)
by: Fekete, Marcell, et al.
Published: (2025)
Detecting Linguistic Diversity on Social Media
by: Wong, Sidney, et al.
Published: (2025)
by: Wong, Sidney, et al.
Published: (2025)
Learning Phonotactics from Linguistic Informants
by: Breiss, Canaan, et al.
Published: (2024)
by: Breiss, Canaan, et al.
Published: (2024)
Similar Items
-
Predicting the Emergence of Induction Heads in Language Model Pretraining
by: Aoyama, Tatsuya, et al.
Published: (2025) -
Information-Theoretic Storage Cost in Sentence Comprehension
by: Kajikawa, Kohei, et al.
Published: (2026) -
Function Words as Statistical Cues for Language Learning
by: Yang, Xiulin, et al.
Published: (2026) -
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026) -
A Unified Assessment of the Poverty of the Stimulus Argument for Neural Language Models
by: Yang, Xiulin, et al.
Published: (2026)