Testing the Predictions of Surprisal Theory in 11 Languages
Fuente:
arXiv
Guardado en:
| Autores principales: | Wilcox, Ethan Gotlieb, Pimentel, Tiago, Meister, Clara, Cotterell, Ryan, Levy, Roger P. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Efficacy of Sampling Adapters
por: Meister, Clara, et al.
Publicado: (2023)
por: Meister, Clara, et al.
Publicado: (2023)
Towards a Similarity-adjusted Surprisal Theory
por: Meister, Clara, et al.
Publicado: (2024)
por: Meister, Clara, et al.
Publicado: (2024)
Analyzing Wrap-Up Effects through an Information-Theoretic Lens
por: Meister, Clara, et al.
Publicado: (2022)
por: Meister, Clara, et al.
Publicado: (2022)
On the Role of Context in Reading Time Prediction
por: Opedal, Andreas, et al.
Publicado: (2024)
por: Opedal, Andreas, et al.
Publicado: (2024)
Locally Typical Sampling
por: Meister, Clara, et al.
Publicado: (2022)
por: Meister, Clara, et al.
Publicado: (2022)
Surprise! Uniform Information Density Isn't the Whole Story: Predicting Surprisal Contours in Long-form Discourse
por: Tsipidi, Eleftheria, et al.
Publicado: (2024)
por: Tsipidi, Eleftheria, et al.
Publicado: (2024)
Predicting the Emergence of Induction Heads in Language Model Pretraining
por: Aoyama, Tatsuya, et al.
Publicado: (2025)
por: Aoyama, Tatsuya, et al.
Publicado: (2025)
Using Information Theory to Characterize Prosodic Typology: The Case of Tone, Pitch-Accent and Stress-Accent
por: Wilcox, Ethan Gotlieb, et al.
Publicado: (2025)
por: Wilcox, Ethan Gotlieb, et al.
Publicado: (2025)
How to Compute the Probability of a Word
por: Pimentel, Tiago, et al.
Publicado: (2024)
por: Pimentel, Tiago, et al.
Publicado: (2024)
Looking forward: Linguistic theory and methods
por: Mansfield, John, et al.
Publicado: (2025)
por: Mansfield, John, et al.
Publicado: (2025)
Reverse-Engineering the Reader
por: Kiegeland, Samuel, et al.
Publicado: (2024)
por: Kiegeland, Samuel, et al.
Publicado: (2024)
What Can String Probability Tell Us About Grammaticality?
por: Hu, Jennifer, et al.
Publicado: (2025)
por: Hu, Jennifer, et al.
Publicado: (2025)
On the Proper Treatment of Units in Surprisal Theory
por: Kiegeland, Samuel, et al.
Publicado: (2026)
por: Kiegeland, Samuel, et al.
Publicado: (2026)
Function Words as Statistical Cues for Language Learning
por: Yang, Xiulin, et al.
Publicado: (2026)
por: Yang, Xiulin, et al.
Publicado: (2026)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
por: Constantinescu, Ionut, et al.
Publicado: (2024)
por: Constantinescu, Ionut, et al.
Publicado: (2024)
What Language is This? Ask Your Tokenizer
por: Meister, Clara, et al.
Publicado: (2026)
por: Meister, Clara, et al.
Publicado: (2026)
Formal Aspects of Language Modeling
por: Cotterell, Ryan, et al.
Publicado: (2023)
por: Cotterell, Ryan, et al.
Publicado: (2023)
Information-Theoretic Storage Cost in Sentence Comprehension
por: Kajikawa, Kohei, et al.
Publicado: (2026)
por: Kajikawa, Kohei, et al.
Publicado: (2026)
Dual Alignment Between Language Model Layers and Human Sentence Processing
por: Kuribayashi, Tatsuki, et al.
Publicado: (2026)
por: Kuribayashi, Tatsuki, et al.
Publicado: (2026)
Modeling Bottom-up Information Quality during Language Processing
por: Ding, Cui, et al.
Publicado: (2025)
por: Ding, Cui, et al.
Publicado: (2025)
A Unified Assessment of the Poverty of the Stimulus Argument for Neural Language Models
por: Yang, Xiulin, et al.
Publicado: (2026)
por: Yang, Xiulin, et al.
Publicado: (2026)
Speakers Fill Lexical Semantic Gaps with Context
por: Pimentel, Tiago, et al.
Publicado: (2020)
por: Pimentel, Tiago, et al.
Publicado: (2020)
Probing for the Usage of Grammatical Number
por: Lasri, Karim, et al.
Publicado: (2022)
por: Lasri, Karim, et al.
Publicado: (2022)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
por: Hu, Michael Y., et al.
Publicado: (2024)
por: Hu, Michael Y., et al.
Publicado: (2024)
The Role of $n$-gram Smoothing in the Age of Neural Networks
por: Malagutti, Luca, et al.
Publicado: (2024)
por: Malagutti, Luca, et al.
Publicado: (2024)
What Do Prosody and Text Convey? Characterizing How Meaningful Information is Distributed Across Multiple Channels
por: Yadavalli, Aditya, et al.
Publicado: (2025)
por: Yadavalli, Aditya, et al.
Publicado: (2025)
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
por: Choshen, Leshem, et al.
Publicado: (2026)
por: Choshen, Leshem, et al.
Publicado: (2026)
Language Models Grow Less Humanlike beyond Phase Transition
por: Aoyama, Tatsuya, et al.
Publicado: (2025)
por: Aoyama, Tatsuya, et al.
Publicado: (2025)
Causal Estimation of Tokenisation Bias
por: Lesci, Pietro, et al.
Publicado: (2025)
por: Lesci, Pietro, et al.
Publicado: (2025)
Characterizing the Expressivity of Fixed-Precision Transformer Language Models
por: Li, Jiaoda, et al.
Publicado: (2025)
por: Li, Jiaoda, et al.
Publicado: (2025)
Towards Explainability in Legal Outcome Prediction Models
por: Valvoda, Josef, et al.
Publicado: (2024)
por: Valvoda, Josef, et al.
Publicado: (2024)
The Harmonic Structure of Information Contours
por: Tsipidi, Eleftheria, et al.
Publicado: (2025)
por: Tsipidi, Eleftheria, et al.
Publicado: (2025)
Transformers Can Represent $n$-gram Language Models
por: Svete, Anej, et al.
Publicado: (2024)
por: Svete, Anej, et al.
Publicado: (2024)
A Formal Perspective on Byte-Pair Encoding
por: Zouhar, Vilém, et al.
Publicado: (2023)
por: Zouhar, Vilém, et al.
Publicado: (2023)
N-gram-like Language Models Predict Reading Time Best
por: Michaelov, James A., et al.
Publicado: (2026)
por: Michaelov, James A., et al.
Publicado: (2026)
Characterizing the Expressivity of Local Attention in Transformers
por: Li, Jiaoda, et al.
Publicado: (2026)
por: Li, Jiaoda, et al.
Publicado: (2026)
Exact Hard Monotonic Attention for Character-Level Transduction
por: Wu, Shijie, et al.
Publicado: (2019)
por: Wu, Shijie, et al.
Publicado: (2019)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
por: Cotterell, Ryan, et al.
Publicado: (2024)
por: Cotterell, Ryan, et al.
Publicado: (2024)
Cross-lingual, Character-Level Neural Morphological Tagging
por: Cotterell, Ryan, et al.
Publicado: (2017)
por: Cotterell, Ryan, et al.
Publicado: (2017)
An Algebraic View of the Expressivity of Recurrent Language Models
por: Nowak, Franz, et al.
Publicado: (2026)
por: Nowak, Franz, et al.
Publicado: (2026)
Ejemplares similares
-
On the Efficacy of Sampling Adapters
por: Meister, Clara, et al.
Publicado: (2023) -
Towards a Similarity-adjusted Surprisal Theory
por: Meister, Clara, et al.
Publicado: (2024) -
Analyzing Wrap-Up Effects through an Information-Theoretic Lens
por: Meister, Clara, et al.
Publicado: (2022) -
On the Role of Context in Reading Time Prediction
por: Opedal, Andreas, et al.
Publicado: (2024) -
Locally Typical Sampling
por: Meister, Clara, et al.
Publicado: (2022)