Learning Linear Attention in Polynomial Time
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yau, Morris, Akyürek, Ekin, Mao, Jiayuan, Tenenbaum, Joshua B., Jegelka, Stefanie, Andreas, Jacob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are Graph Neural Networks Optimal Approximation Algorithms?
von: Yau, Morris, et al.
Veröffentlicht: (2023)
von: Yau, Morris, et al.
Veröffentlicht: (2023)
Polynomial-Time Approximability of Constrained Reinforcement Learning
von: McMahan, Jeremy
Veröffentlicht: (2025)
von: McMahan, Jeremy
Veröffentlicht: (2025)
A Polynomial-Time Approximation for Pairwise Fair $k$-Median Clustering
von: Bandyapadhyay, Sayan, et al.
Veröffentlicht: (2024)
von: Bandyapadhyay, Sayan, et al.
Veröffentlicht: (2024)
GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling
von: Katsch, Tobias
Veröffentlicht: (2023)
von: Katsch, Tobias
Veröffentlicht: (2023)
Learning to Approximate Uniform Facility Location via Graph Neural Networks
von: Qian, Chendi, et al.
Veröffentlicht: (2026)
von: Qian, Chendi, et al.
Veröffentlicht: (2026)
Identification for Tree-shaped Structural Causal Models in Polynomial Time
von: Gupta, Aaryan, et al.
Veröffentlicht: (2023)
von: Gupta, Aaryan, et al.
Veröffentlicht: (2023)
HashEvict: A Pre-Attention KV Cache Eviction Strategy using Locality-Sensitive Hashing
von: Liu, Minghui, et al.
Veröffentlicht: (2024)
von: Liu, Minghui, et al.
Veröffentlicht: (2024)
Linear-Time Algorithms for Front-Door Adjustment in Causal Graphs
von: Wienöbst, Marcel, et al.
Veröffentlicht: (2022)
von: Wienöbst, Marcel, et al.
Veröffentlicht: (2022)
Linear-Time Primitives for Algorithm Development in Graphical Causal Inference
von: Wienöbst, Marcel, et al.
Veröffentlicht: (2025)
von: Wienöbst, Marcel, et al.
Veröffentlicht: (2025)
Streaming Attention Approximation via Discrepancy Theory
von: Kochetkova, Ekaterina, et al.
Veröffentlicht: (2025)
von: Kochetkova, Ekaterina, et al.
Veröffentlicht: (2025)
Positional Attention: Expressivity and Learnability of Algorithmic Computation
von: de Luca, Artur Back, et al.
Veröffentlicht: (2024)
von: de Luca, Artur Back, et al.
Veröffentlicht: (2024)
On Characterizations for Language Generation: Interplay of Hallucinations, Breadth, and Stability
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
On the Limits of Language Generation: Trade-Offs Between Hallucination and Mode Collapse
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
On Language Generation in the Limit with Bounded Memory
von: Kleinberg, Jon, et al.
Veröffentlicht: (2026)
von: Kleinberg, Jon, et al.
Veröffentlicht: (2026)
Language Generation in the Limit
von: Kleinberg, Jon, et al.
Veröffentlicht: (2024)
von: Kleinberg, Jon, et al.
Veröffentlicht: (2024)
Differentially Private Language Generation and Identification in the Limit
von: Mehrotra, Anay, et al.
Veröffentlicht: (2026)
von: Mehrotra, Anay, et al.
Veröffentlicht: (2026)
Exploring Facets of Language Generation in the Limit
von: Charikar, Moses, et al.
Veröffentlicht: (2024)
von: Charikar, Moses, et al.
Veröffentlicht: (2024)
The CLRS-Text Algorithmic Reasoning Language Benchmark
von: Markeeva, Larisa, et al.
Veröffentlicht: (2024)
von: Markeeva, Larisa, et al.
Veröffentlicht: (2024)
Contrastive Identification and Generation in the Limit
von: Li, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Li, Xiaoyu, et al.
Veröffentlicht: (2026)
Language Generation with Infinite Contamination
von: Mehrotra, Anay, et al.
Veröffentlicht: (2025)
von: Mehrotra, Anay, et al.
Veröffentlicht: (2025)
A Characterization of List Language Identification in the Limit
von: Charikar, Moses, et al.
Veröffentlicht: (2025)
von: Charikar, Moses, et al.
Veröffentlicht: (2025)
Pareto-optimal Non-uniform Language Generation
von: Charikar, Moses, et al.
Veröffentlicht: (2025)
von: Charikar, Moses, et al.
Veröffentlicht: (2025)
A Tighter Complexity Analysis of SparseGPT
von: Li, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Li, Xiaoyu, et al.
Veröffentlicht: (2024)
The Library Theorem: How External Organization Governs Agentic Reasoning Capacity
von: Mainen, Zachary F.
Veröffentlicht: (2026)
von: Mainen, Zachary F.
Veröffentlicht: (2026)
Extremely Simple Streaming Forest
von: Xu, Haoyin, et al.
Veröffentlicht: (2021)
von: Xu, Haoyin, et al.
Veröffentlicht: (2021)
Anytime-Constrained Equilibria in Polynomial Time
von: McMahan, Jeremy
Veröffentlicht: (2024)
von: McMahan, Jeremy
Veröffentlicht: (2024)
SubGen: Token Generation in Sublinear Time and Memory
von: Zandieh, Amir, et al.
Veröffentlicht: (2024)
von: Zandieh, Amir, et al.
Veröffentlicht: (2024)
Nearly Optimal Attention Coresets
von: Liberty, Edo, et al.
Veröffentlicht: (2026)
von: Liberty, Edo, et al.
Veröffentlicht: (2026)
Learning Algorithms in the Limit
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
Computing Optimal Regularizers for Online Linear Optimization
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2024)
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2024)
Anytime-Constrained Reinforcement Learning
von: McMahan, Jeremy, et al.
Veröffentlicht: (2023)
von: McMahan, Jeremy, et al.
Veröffentlicht: (2023)
Learning-Augmented Priority Queues
von: Benomar, Ziyad, et al.
Veröffentlicht: (2024)
von: Benomar, Ziyad, et al.
Veröffentlicht: (2024)
On Tradeoffs in Learning-Augmented Algorithms
von: Benomar, Ziyad, et al.
Veröffentlicht: (2025)
von: Benomar, Ziyad, et al.
Veröffentlicht: (2025)
Learning-Augmented Online Bipartite Fractional Matching
von: Choo, Davin, et al.
Veröffentlicht: (2025)
von: Choo, Davin, et al.
Veröffentlicht: (2025)
Learning-Based Algorithms for Graph Searching Problems
von: DePavia, Adela Frances, et al.
Veröffentlicht: (2024)
von: DePavia, Adela Frances, et al.
Veröffentlicht: (2024)
Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2026)
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2026)
A Partition Cover Approach to Tokenization
von: Lim, Jia Peng, et al.
Veröffentlicht: (2025)
von: Lim, Jia Peng, et al.
Veröffentlicht: (2025)
Algorithmically Establishing Trust in Evaluators
von: de Wynter, Adrian
Veröffentlicht: (2025)
von: de Wynter, Adrian
Veröffentlicht: (2025)
An Algorithm for Learning Smaller Representations of Models With Scarce Data
von: de Wynter, Adrian
Veröffentlicht: (2020)
von: de Wynter, Adrian
Veröffentlicht: (2020)
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
von: Li, Dongyue, et al.
Veröffentlicht: (2025)
von: Li, Dongyue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Are Graph Neural Networks Optimal Approximation Algorithms?
von: Yau, Morris, et al.
Veröffentlicht: (2023) -
Polynomial-Time Approximability of Constrained Reinforcement Learning
von: McMahan, Jeremy
Veröffentlicht: (2025) -
A Polynomial-Time Approximation for Pairwise Fair $k$-Median Clustering
von: Bandyapadhyay, Sayan, et al.
Veröffentlicht: (2024) -
GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling
von: Katsch, Tobias
Veröffentlicht: (2023) -
Learning to Approximate Uniform Facility Location via Graph Neural Networks
von: Qian, Chendi, et al.
Veröffentlicht: (2026)