Bearing Syntactic Fruit with Stack-Augmented Neural Networks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | DuSell, Brian, Cotterell, Ryan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns
par: DuSell, Brian, et autres
Publié: (2023)
par: DuSell, Brian, et autres
Publié: (2023)
Training Neural Networks as Recognizers of Formal Languages
par: Butoi, Alexandra, et autres
Publié: (2024)
par: Butoi, Alexandra, et autres
Publié: (2024)
Algorithms for Weighted Pushdown Automata
par: Butoi, Alexandra, et autres
Publié: (2022)
par: Butoi, Alexandra, et autres
Publié: (2022)
Information Locality as an Inductive Bias for Neural Language Models
par: Someya, Taiga, et autres
Publié: (2025)
par: Someya, Taiga, et autres
Publié: (2025)
On the Proper Treatment of Tokenization in Psycholinguistics
par: Giulianelli, Mario, et autres
Publié: (2024)
par: Giulianelli, Mario, et autres
Publié: (2024)
The Foundations of Tokenization: Statistical and Computational Concerns
par: Gastaldi, Juan Luis, et autres
Publié: (2024)
par: Gastaldi, Juan Luis, et autres
Publié: (2024)
PILA: A Historical-Linguistic Dataset of Proto-Italic and Latin
par: Bothwell, Stephen, et autres
Publié: (2024)
par: Bothwell, Stephen, et autres
Publié: (2024)
From Language Models over Tokens to Language Models over Characters
par: Vieira, Tim, et autres
Publié: (2024)
par: Vieira, Tim, et autres
Publié: (2024)
Syntactic Control of Language Models by Posterior Inference
par: Xefteri, Vicky, et autres
Publié: (2025)
par: Xefteri, Vicky, et autres
Publié: (2025)
Language Models over Canonical Byte-Pair Encodings
par: Vieira, Tim, et autres
Publié: (2025)
par: Vieira, Tim, et autres
Publié: (2025)
A Transformer with Stack Attention
par: Li, Jiaoda, et autres
Publié: (2024)
par: Li, Jiaoda, et autres
Publié: (2024)
Cross-lingual, Character-Level Neural Morphological Tagging
par: Cotterell, Ryan, et autres
Publié: (2017)
par: Cotterell, Ryan, et autres
Publié: (2017)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
par: Cotterell, Ryan, et autres
Publié: (2024)
par: Cotterell, Ryan, et autres
Publié: (2024)
On the Representational Capacity of Recurrent Neural Language Models
par: Nowak, Franz, et autres
Publié: (2023)
par: Nowak, Franz, et autres
Publié: (2023)
A Simple Joint Model for Improved Contextual Neural Lemmatization
par: Malaviya, Chaitanya, et autres
Publié: (2019)
par: Malaviya, Chaitanya, et autres
Publié: (2019)
Structured Voronoi Sampling
par: Amini, Afra, et autres
Publié: (2023)
par: Amini, Afra, et autres
Publié: (2023)
The Role of $n$-gram Smoothing in the Age of Neural Networks
par: Malagutti, Luca, et autres
Publié: (2024)
par: Malagutti, Luca, et autres
Publié: (2024)
A Distributional Perspective on Word Learning in Neural Language Models
par: Ficarra, Filippo, et autres
Publié: (2025)
par: Ficarra, Filippo, et autres
Publié: (2025)
Characterizing the Expressivity of Fixed-Precision Transformer Language Models
par: Li, Jiaoda, et autres
Publié: (2025)
par: Li, Jiaoda, et autres
Publié: (2025)
Characterizing the Expressivity of Local Attention in Transformers
par: Li, Jiaoda, et autres
Publié: (2026)
par: Li, Jiaoda, et autres
Publié: (2026)
Exact Hard Monotonic Attention for Character-Level Transduction
par: Wu, Shijie, et autres
Publié: (2019)
par: Wu, Shijie, et autres
Publié: (2019)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
par: Nowak, Franz, et autres
Publié: (2024)
par: Nowak, Franz, et autres
Publié: (2024)
Efficiently Computing Susceptibility to Context in Language Models
par: Liu, Tianyu, et autres
Publié: (2024)
par: Liu, Tianyu, et autres
Publié: (2024)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
par: Constantinescu, Ionut, et autres
Publié: (2024)
par: Constantinescu, Ionut, et autres
Publié: (2024)
Transformers Can Represent $n$-gram Language Models
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Lower Bounds on the Expressivity of Recurrent Neural Language Models
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
par: Vargas, Francisco, et autres
Publié: (2020)
par: Vargas, Francisco, et autres
Publié: (2020)
Towards Explainability in Legal Outcome Prediction Models
par: Valvoda, Josef, et autres
Publié: (2024)
par: Valvoda, Josef, et autres
Publié: (2024)
Emergence of Phonemic, Syntactic, and Semantic Representations in Artificial Neural Networks
par: Orhan, Pierre, et autres
Publié: (2026)
par: Orhan, Pierre, et autres
Publié: (2026)
Attending To Syntactic Information In Biomedical Event Extraction Via Graph Neural Networks
par: Noravesh, Farshad, et autres
Publié: (2025)
par: Noravesh, Farshad, et autres
Publié: (2025)
Hard Non-Monotonic Attention for Character-Level Transduction
par: Wu, Shijie, et autres
Publié: (2018)
par: Wu, Shijie, et autres
Publié: (2018)
The Causal Influence of Grammatical Gender on Distributional Semantics
par: Stańczak, Karolina, et autres
Publié: (2023)
par: Stańczak, Karolina, et autres
Publié: (2023)
Formal Aspects of Language Modeling
par: Cotterell, Ryan, et autres
Publié: (2023)
par: Cotterell, Ryan, et autres
Publié: (2023)
How Persuasive is Your Context?
par: Nguyen, Tu, et autres
Publié: (2025)
par: Nguyen, Tu, et autres
Publié: (2025)
Graph Neural Network Framework for Sentiment Analysis Using Syntactic Feature
par: Wu, Linxiao, et autres
Publié: (2024)
par: Wu, Linxiao, et autres
Publié: (2024)
Syntactic and Semantic Control of Large Language Models via Sequential Monte Carlo
par: Loula, João, et autres
Publié: (2025)
par: Loula, João, et autres
Publié: (2025)
Frege in the Flesh: Biolinguistics and the Neural Enforcement of Syntactic Structures
par: Murphy, Elliot
Publié: (2026)
par: Murphy, Elliot
Publié: (2026)
An Algebraic View of the Expressivity of Recurrent Language Models
par: Nowak, Franz, et autres
Publié: (2026)
par: Nowak, Franz, et autres
Publié: (2026)
Log-linear Guardedness and its Implications
par: Ravfogel, Shauli, et autres
Publié: (2022)
par: Ravfogel, Shauli, et autres
Publié: (2022)
Syntactic Learnability of Echo State Neural Language Models at Scale
par: Ueda, Ryo, et autres
Publié: (2025)
par: Ueda, Ryo, et autres
Publié: (2025)
Documents similaires
-
Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns
par: DuSell, Brian, et autres
Publié: (2023) -
Training Neural Networks as Recognizers of Formal Languages
par: Butoi, Alexandra, et autres
Publié: (2024) -
Algorithms for Weighted Pushdown Automata
par: Butoi, Alexandra, et autres
Publié: (2022) -
Information Locality as an Inductive Bias for Neural Language Models
par: Someya, Taiga, et autres
Publié: (2025) -
On the Proper Treatment of Tokenization in Psycholinguistics
par: Giulianelli, Mario, et autres
Publié: (2024)