Analyzing Wrap-Up Effects through an Information-Theoretic Lens
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meister, Clara, Pimentel, Tiago, Clark, Thomas Hikaru, Cotterell, Ryan, Levy, Roger |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Testing the Predictions of Surprisal Theory in 11 Languages
von: Wilcox, Ethan Gotlieb, et al.
Veröffentlicht: (2023)
von: Wilcox, Ethan Gotlieb, et al.
Veröffentlicht: (2023)
Locally Typical Sampling
von: Meister, Clara, et al.
Veröffentlicht: (2022)
von: Meister, Clara, et al.
Veröffentlicht: (2022)
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023)
von: Meister, Clara, et al.
Veröffentlicht: (2023)
How to Compute the Probability of a Word
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
von: Constantinescu, Ionut, et al.
Veröffentlicht: (2024)
von: Constantinescu, Ionut, et al.
Veröffentlicht: (2024)
Towards a Similarity-adjusted Surprisal Theory
von: Meister, Clara, et al.
Veröffentlicht: (2024)
von: Meister, Clara, et al.
Veröffentlicht: (2024)
Readers make targeted regressions to plausible errors in reanalysis of "noisy-channel garden-path" sentences
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
What Language is This? Ask Your Tokenizer
von: Meister, Clara, et al.
Veröffentlicht: (2026)
von: Meister, Clara, et al.
Veröffentlicht: (2026)
Formal Aspects of Language Modeling
von: Cotterell, Ryan, et al.
Veröffentlicht: (2023)
von: Cotterell, Ryan, et al.
Veröffentlicht: (2023)
Speakers Fill Lexical Semantic Gaps with Context
von: Pimentel, Tiago, et al.
Veröffentlicht: (2020)
von: Pimentel, Tiago, et al.
Veröffentlicht: (2020)
Probing for the Usage of Grammatical Number
von: Lasri, Karim, et al.
Veröffentlicht: (2022)
von: Lasri, Karim, et al.
Veröffentlicht: (2022)
Causal Estimation of Tokenisation Bias
von: Lesci, Pietro, et al.
Veröffentlicht: (2025)
von: Lesci, Pietro, et al.
Veröffentlicht: (2025)
The Role of $n$-gram Smoothing in the Age of Neural Networks
von: Malagutti, Luca, et al.
Veröffentlicht: (2024)
von: Malagutti, Luca, et al.
Veröffentlicht: (2024)
A Formal Perspective on Byte-Pair Encoding
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
Characterizing the Expressivity of Local Attention in Transformers
von: Li, Jiaoda, et al.
Veröffentlicht: (2026)
von: Li, Jiaoda, et al.
Veröffentlicht: (2026)
Exact Hard Monotonic Attention for Character-Level Transduction
von: Wu, Shijie, et al.
Veröffentlicht: (2019)
von: Wu, Shijie, et al.
Veröffentlicht: (2019)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
von: Cotterell, Ryan, et al.
Veröffentlicht: (2024)
von: Cotterell, Ryan, et al.
Veröffentlicht: (2024)
Cross-lingual, Character-Level Neural Morphological Tagging
von: Cotterell, Ryan, et al.
Veröffentlicht: (2017)
von: Cotterell, Ryan, et al.
Veröffentlicht: (2017)
Characterizing the Expressivity of Fixed-Precision Transformer Language Models
von: Li, Jiaoda, et al.
Veröffentlicht: (2025)
von: Li, Jiaoda, et al.
Veröffentlicht: (2025)
Transformers Can Represent $n$-gram Language Models
von: Svete, Anej, et al.
Veröffentlicht: (2024)
von: Svete, Anej, et al.
Veröffentlicht: (2024)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
von: Vargas, Francisco, et al.
Veröffentlicht: (2020)
von: Vargas, Francisco, et al.
Veröffentlicht: (2020)
Bearing Syntactic Fruit with Stack-Augmented Neural Networks
von: DuSell, Brian, et al.
Veröffentlicht: (2025)
von: DuSell, Brian, et al.
Veröffentlicht: (2025)
Towards Explainability in Legal Outcome Prediction Models
von: Valvoda, Josef, et al.
Veröffentlicht: (2024)
von: Valvoda, Josef, et al.
Veröffentlicht: (2024)
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
Joint Lemmatization and Morphological Tagging with LEMMING
von: Muller, Thomas, et al.
Veröffentlicht: (2024)
von: Muller, Thomas, et al.
Veröffentlicht: (2024)
Labeled Morphological Segmentation with Semi-Markov Models
von: Cotterell, Ryan, et al.
Veröffentlicht: (2024)
von: Cotterell, Ryan, et al.
Veröffentlicht: (2024)
Greedy or not, here I come: Language production under vocabulary constraints in humans and resource-rational models
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
A Simple Joint Model for Improved Contextual Neural Lemmatization
von: Malaviya, Chaitanya, et al.
Veröffentlicht: (2019)
von: Malaviya, Chaitanya, et al.
Veröffentlicht: (2019)
Hard Non-Monotonic Attention for Character-Level Transduction
von: Wu, Shijie, et al.
Veröffentlicht: (2018)
von: Wu, Shijie, et al.
Veröffentlicht: (2018)
An Algebraic View of the Expressivity of Recurrent Language Models
von: Nowak, Franz, et al.
Veröffentlicht: (2026)
von: Nowak, Franz, et al.
Veröffentlicht: (2026)
Log-linear Guardedness and its Implications
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2022)
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2022)
A Distributional Perspective on Word Learning in Neural Language Models
von: Ficarra, Filippo, et al.
Veröffentlicht: (2025)
von: Ficarra, Filippo, et al.
Veröffentlicht: (2025)
Structured Voronoi Sampling
von: Amini, Afra, et al.
Veröffentlicht: (2023)
von: Amini, Afra, et al.
Veröffentlicht: (2023)
On the Effect of (Near) Duplicate Subwords in Language Modelling
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
Revisiting Real-Time Digging-In Effects: No Evidence from NP/Z Garden-Paths
von: Maina-Kilaas, Amani, et al.
Veröffentlicht: (2026)
von: Maina-Kilaas, Amani, et al.
Veröffentlicht: (2026)
Local and Global Decoding in Text Generation
von: Gareev, Daniel, et al.
Veröffentlicht: (2024)
von: Gareev, Daniel, et al.
Veröffentlicht: (2024)
Efficiently Computing Susceptibility to Context in Language Models
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
On the Proper Treatment of Units in Surprisal Theory
von: Kiegeland, Samuel, et al.
Veröffentlicht: (2026)
von: Kiegeland, Samuel, et al.
Veröffentlicht: (2026)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
von: Nowak, Franz, et al.
Veröffentlicht: (2024)
von: Nowak, Franz, et al.
Veröffentlicht: (2024)
Unique Hard Attention: A Tale of Two Sides
von: Jerad, Selim, et al.
Veröffentlicht: (2025)
von: Jerad, Selim, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Testing the Predictions of Surprisal Theory in 11 Languages
von: Wilcox, Ethan Gotlieb, et al.
Veröffentlicht: (2023) -
Locally Typical Sampling
von: Meister, Clara, et al.
Veröffentlicht: (2022) -
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023) -
How to Compute the Probability of a Word
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024) -
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
von: Constantinescu, Ionut, et al.
Veröffentlicht: (2024)