Exact Hard Monotonic Attention for Character-Level Transduction
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Shijie, Cotterell, Ryan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2019
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hard Non-Monotonic Attention for Character-Level Transduction
di: Wu, Shijie, et al.
Pubblicazione: (2018)
di: Wu, Shijie, et al.
Pubblicazione: (2018)
Cross-lingual, Character-Level Neural Morphological Tagging
di: Cotterell, Ryan, et al.
Pubblicazione: (2017)
di: Cotterell, Ryan, et al.
Pubblicazione: (2017)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
di: Cotterell, Ryan, et al.
Pubblicazione: (2024)
di: Cotterell, Ryan, et al.
Pubblicazione: (2024)
Unique Hard Attention: A Tale of Two Sides
di: Jerad, Selim, et al.
Pubblicazione: (2025)
di: Jerad, Selim, et al.
Pubblicazione: (2025)
A Simple Joint Model for Improved Contextual Neural Lemmatization
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)
Characterizing the Expressivity of Local Attention in Transformers
di: Li, Jiaoda, et al.
Pubblicazione: (2026)
di: Li, Jiaoda, et al.
Pubblicazione: (2026)
A Transformer with Stack Attention
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
Characterizing the Expressivity of Fixed-Precision Transformer Language Models
di: Li, Jiaoda, et al.
Pubblicazione: (2025)
di: Li, Jiaoda, et al.
Pubblicazione: (2025)
Transformers Can Represent $n$-gram Language Models
di: Svete, Anej, et al.
Pubblicazione: (2024)
di: Svete, Anej, et al.
Pubblicazione: (2024)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
Bearing Syntactic Fruit with Stack-Augmented Neural Networks
di: DuSell, Brian, et al.
Pubblicazione: (2025)
di: DuSell, Brian, et al.
Pubblicazione: (2025)
Towards Explainability in Legal Outcome Prediction Models
di: Valvoda, Josef, et al.
Pubblicazione: (2024)
di: Valvoda, Josef, et al.
Pubblicazione: (2024)
An Algebraic View of the Expressivity of Recurrent Language Models
di: Nowak, Franz, et al.
Pubblicazione: (2026)
di: Nowak, Franz, et al.
Pubblicazione: (2026)
From Language Models over Tokens to Language Models over Characters
di: Vieira, Tim, et al.
Pubblicazione: (2024)
di: Vieira, Tim, et al.
Pubblicazione: (2024)
Log-linear Guardedness and its Implications
di: Ravfogel, Shauli, et al.
Pubblicazione: (2022)
di: Ravfogel, Shauli, et al.
Pubblicazione: (2022)
A Distributional Perspective on Word Learning in Neural Language Models
di: Ficarra, Filippo, et al.
Pubblicazione: (2025)
di: Ficarra, Filippo, et al.
Pubblicazione: (2025)
Structured Voronoi Sampling
di: Amini, Afra, et al.
Pubblicazione: (2023)
di: Amini, Afra, et al.
Pubblicazione: (2023)
Efficiently Computing Susceptibility to Context in Language Models
di: Liu, Tianyu, et al.
Pubblicazione: (2024)
di: Liu, Tianyu, et al.
Pubblicazione: (2024)
Joint Lemmatization and Morphological Tagging with LEMMING
di: Muller, Thomas, et al.
Pubblicazione: (2024)
di: Muller, Thomas, et al.
Pubblicazione: (2024)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
di: Constantinescu, Ionut, et al.
Pubblicazione: (2024)
di: Constantinescu, Ionut, et al.
Pubblicazione: (2024)
Labeled Morphological Segmentation with Semi-Markov Models
di: Cotterell, Ryan, et al.
Pubblicazione: (2024)
di: Cotterell, Ryan, et al.
Pubblicazione: (2024)
On the Proper Treatment of Units in Surprisal Theory
di: Kiegeland, Samuel, et al.
Pubblicazione: (2026)
di: Kiegeland, Samuel, et al.
Pubblicazione: (2026)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
di: Nowak, Franz, et al.
Pubblicazione: (2024)
di: Nowak, Franz, et al.
Pubblicazione: (2024)
Masked Hard-Attention Transformers Recognize Exactly the Star-Free Languages
di: Yang, Andy, et al.
Pubblicazione: (2023)
di: Yang, Andy, et al.
Pubblicazione: (2023)
Better Estimation of the Kullback--Leibler Divergence Between Language Models
di: Amini, Afra, et al.
Pubblicazione: (2025)
di: Amini, Afra, et al.
Pubblicazione: (2025)
Direct Preference Optimization with an Offset
di: Amini, Afra, et al.
Pubblicazione: (2024)
di: Amini, Afra, et al.
Pubblicazione: (2024)
Generalized Measures of Anticipation and Responsivity in Online Language Processing
di: Giulianelli, Mario, et al.
Pubblicazione: (2024)
di: Giulianelli, Mario, et al.
Pubblicazione: (2024)
Tokenization as Finite-State Transduction
di: Cognetta, Marco, et al.
Pubblicazione: (2024)
di: Cognetta, Marco, et al.
Pubblicazione: (2024)
Lower Bounds on the Expressivity of Recurrent Neural Language Models
di: Svete, Anej, et al.
Pubblicazione: (2024)
di: Svete, Anej, et al.
Pubblicazione: (2024)
Speakers Fill Lexical Semantic Gaps with Context
di: Pimentel, Tiago, et al.
Pubblicazione: (2020)
di: Pimentel, Tiago, et al.
Pubblicazione: (2020)
Simulating Hard Attention Using Soft Attention
di: Yang, Andy, et al.
Pubblicazione: (2024)
di: Yang, Andy, et al.
Pubblicazione: (2024)
Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't
di: Svete, Anej, et al.
Pubblicazione: (2026)
di: Svete, Anej, et al.
Pubblicazione: (2026)
On Efficiently Representing Regular Languages as RNNs
di: Svete, Anej, et al.
Pubblicazione: (2024)
di: Svete, Anej, et al.
Pubblicazione: (2024)
A Practical Method for Generating String Counterfactuals
di: Avitan, Matan, et al.
Pubblicazione: (2024)
di: Avitan, Matan, et al.
Pubblicazione: (2024)
Understanding the Ability of LLMs to Handle Character-Level Perturbation
di: Zhuo, Anyuan, et al.
Pubblicazione: (2025)
di: Zhuo, Anyuan, et al.
Pubblicazione: (2025)
SpeLLM: Character-Level Multi-Head Decoding
di: Ben-Artzy, Amit, et al.
Pubblicazione: (2025)
di: Ben-Artzy, Amit, et al.
Pubblicazione: (2025)
Locally Typical Sampling
di: Meister, Clara, et al.
Pubblicazione: (2022)
di: Meister, Clara, et al.
Pubblicazione: (2022)
On the Representational Capacity of Recurrent Neural Language Models
di: Nowak, Franz, et al.
Pubblicazione: (2023)
di: Nowak, Franz, et al.
Pubblicazione: (2023)
Kernelized Concept Erasure
di: Ravfogel, Shauli, et al.
Pubblicazione: (2022)
di: Ravfogel, Shauli, et al.
Pubblicazione: (2022)
What Do Language Models Learn in Context? The Structured Task Hypothesis
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hard Non-Monotonic Attention for Character-Level Transduction
di: Wu, Shijie, et al.
Pubblicazione: (2018) -
Cross-lingual, Character-Level Neural Morphological Tagging
di: Cotterell, Ryan, et al.
Pubblicazione: (2017) -
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
di: Cotterell, Ryan, et al.
Pubblicazione: (2024) -
Unique Hard Attention: A Tale of Two Sides
di: Jerad, Selim, et al.
Pubblicazione: (2025) -
A Simple Joint Model for Improved Contextual Neural Lemmatization
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)