GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling
Fuente:
arXiv
Saved in:
| Main Author: | Katsch, Tobias |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Linear Attention in Polynomial Time
by: Yau, Morris, et al.
Published: (2024)
by: Yau, Morris, et al.
Published: (2024)
On Characterizations for Language Generation: Interplay of Hallucinations, Breadth, and Stability
by: Kalavasis, Alkis, et al.
Published: (2024)
by: Kalavasis, Alkis, et al.
Published: (2024)
On the Limits of Language Generation: Trade-Offs Between Hallucination and Mode Collapse
by: Kalavasis, Alkis, et al.
Published: (2024)
by: Kalavasis, Alkis, et al.
Published: (2024)
On Language Generation in the Limit with Bounded Memory
by: Kleinberg, Jon, et al.
Published: (2026)
by: Kleinberg, Jon, et al.
Published: (2026)
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024)
by: Kleinberg, Jon, et al.
Published: (2024)
Differentially Private Language Generation and Identification in the Limit
by: Mehrotra, Anay, et al.
Published: (2026)
by: Mehrotra, Anay, et al.
Published: (2026)
Exploring Facets of Language Generation in the Limit
by: Charikar, Moses, et al.
Published: (2024)
by: Charikar, Moses, et al.
Published: (2024)
The CLRS-Text Algorithmic Reasoning Language Benchmark
by: Markeeva, Larisa, et al.
Published: (2024)
by: Markeeva, Larisa, et al.
Published: (2024)
Contrastive Identification and Generation in the Limit
by: Li, Xiaoyu, et al.
Published: (2026)
by: Li, Xiaoyu, et al.
Published: (2026)
Language Generation with Infinite Contamination
by: Mehrotra, Anay, et al.
Published: (2025)
by: Mehrotra, Anay, et al.
Published: (2025)
A Characterization of List Language Identification in the Limit
by: Charikar, Moses, et al.
Published: (2025)
by: Charikar, Moses, et al.
Published: (2025)
Pareto-optimal Non-uniform Language Generation
by: Charikar, Moses, et al.
Published: (2025)
by: Charikar, Moses, et al.
Published: (2025)
A Tighter Complexity Analysis of SparseGPT
by: Li, Xiaoyu, et al.
Published: (2024)
by: Li, Xiaoyu, et al.
Published: (2024)
The Library Theorem: How External Organization Governs Agentic Reasoning Capacity
by: Mainen, Zachary F.
Published: (2026)
by: Mainen, Zachary F.
Published: (2026)
Simulation of Graph Algorithms with Looped Transformers
by: de Luca, Artur Back, et al.
Published: (2024)
by: de Luca, Artur Back, et al.
Published: (2024)
HashEvict: A Pre-Attention KV Cache Eviction Strategy using Locality-Sensitive Hashing
by: Liu, Minghui, et al.
Published: (2024)
by: Liu, Minghui, et al.
Published: (2024)
An Algorithm for Learning Smaller Representations of Models With Scarce Data
by: de Wynter, Adrian
Published: (2020)
by: de Wynter, Adrian
Published: (2020)
Hallucination is a Consequence of Space-Optimality: A Rate-Distortion Theorem for Membership Testing
by: Guo, Anxin, et al.
Published: (2026)
by: Guo, Anxin, et al.
Published: (2026)
Constructing Decision Trees from Data Streams
by: Pham, Huy, et al.
Published: (2024)
by: Pham, Huy, et al.
Published: (2024)
Discovering Data Structures: Nearest Neighbor Search and Beyond
by: Salemohamed, Omar, et al.
Published: (2024)
by: Salemohamed, Omar, et al.
Published: (2024)
Uncovering Fairness through Data Complexity as an Early Indicator
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
Optimal Classification Trees for Continuous Feature Data Using Dynamic Programming with Branch-and-Bound
by: Brita, Catalin E., et al.
Published: (2025)
by: Brita, Catalin E., et al.
Published: (2025)
Model Stealing for Any Low-Rank Language Model
by: Liu, Allen, et al.
Published: (2024)
by: Liu, Allen, et al.
Published: (2024)
Linear-Time Algorithms for Front-Door Adjustment in Causal Graphs
by: Wienöbst, Marcel, et al.
Published: (2022)
by: Wienöbst, Marcel, et al.
Published: (2022)
Linear-Time Primitives for Algorithm Development in Graphical Causal Inference
by: Wienöbst, Marcel, et al.
Published: (2025)
by: Wienöbst, Marcel, et al.
Published: (2025)
Approximate Lifted Model Construction
by: Luttermann, Malte, et al.
Published: (2025)
by: Luttermann, Malte, et al.
Published: (2025)
Provably Learning from Modern Language Models via Low Logit Rank
by: Golowich, Noah, et al.
Published: (2025)
by: Golowich, Noah, et al.
Published: (2025)
Rethinking Model-based, Policy-based, and Value-based Reinforcement Learning via the Lens of Representation Complexity
by: Feng, Guhao, et al.
Published: (2023)
by: Feng, Guhao, et al.
Published: (2023)
Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm
by: Zanzotto, Fabio Massimo, et al.
Published: (2026)
by: Zanzotto, Fabio Massimo, et al.
Published: (2026)
A Partition Cover Approach to Tokenization
by: Lim, Jia Peng, et al.
Published: (2025)
by: Lim, Jia Peng, et al.
Published: (2025)
Algorithmically Establishing Trust in Evaluators
by: de Wynter, Adrian
Published: (2025)
by: de Wynter, Adrian
Published: (2025)
Welfare-Centric Clustering
by: Zhang, Claire Jie, et al.
Published: (2025)
by: Zhang, Claire Jie, et al.
Published: (2025)
Causal Equal Protection as Algorithmic Fairness
by: Di Bello, Marcello, et al.
Published: (2024)
by: Di Bello, Marcello, et al.
Published: (2024)
Theoretical limitations of multi-layer Transformer
by: Chen, Lijie, et al.
Published: (2024)
by: Chen, Lijie, et al.
Published: (2024)
Scalable Algorithms for Individual Preference Stable Clustering
by: Mosenzon, Ron, et al.
Published: (2024)
by: Mosenzon, Ron, et al.
Published: (2024)
Robust Fair Clustering with Group Membership Uncertainty Sets
by: Duppala, Sharmila, et al.
Published: (2024)
by: Duppala, Sharmila, et al.
Published: (2024)
Diversity-aware clustering: Computational Complexity and Approximation Algorithms
by: Thejaswi, Suhas, et al.
Published: (2024)
by: Thejaswi, Suhas, et al.
Published: (2024)
Fair Clustering: Critique, Caveats, and Future Directions
by: Dickerson, John, et al.
Published: (2024)
by: Dickerson, John, et al.
Published: (2024)
Compression Barriers for Autoregressive Transformers
by: Haris, Themistoklis, et al.
Published: (2025)
by: Haris, Themistoklis, et al.
Published: (2025)
Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods
by: Blum, Avrim, et al.
Published: (2025)
by: Blum, Avrim, et al.
Published: (2025)
Similar Items
-
Learning Linear Attention in Polynomial Time
by: Yau, Morris, et al.
Published: (2024) -
On Characterizations for Language Generation: Interplay of Hallucinations, Breadth, and Stability
by: Kalavasis, Alkis, et al.
Published: (2024) -
On the Limits of Language Generation: Trade-Offs Between Hallucination and Mode Collapse
by: Kalavasis, Alkis, et al.
Published: (2024) -
On Language Generation in the Limit with Bounded Memory
by: Kleinberg, Jon, et al.
Published: (2026) -
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024)