Compression Barriers for Autoregressive Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Haris, Themistoklis, Onak, Krzysztof |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Theoretical limitations of multi-layer Transformer
di: Chen, Lijie, et al.
Pubblicazione: (2024)
di: Chen, Lijie, et al.
Pubblicazione: (2024)
$k$NN Attention Demystified: A Theoretical Exploration for Scalable Transformers
di: Haris, Themistoklis
Pubblicazione: (2024)
di: Haris, Themistoklis
Pubblicazione: (2024)
Efficient Algorithms for Adversarially Robust Approximate Nearest Neighbor Search
di: Andoni, Alexandr, et al.
Pubblicazione: (2026)
di: Andoni, Alexandr, et al.
Pubblicazione: (2026)
Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods
di: Blum, Avrim, et al.
Pubblicazione: (2025)
di: Blum, Avrim, et al.
Pubblicazione: (2025)
Diversity-aware clustering: Computational Complexity and Approximation Algorithms
di: Thejaswi, Suhas, et al.
Pubblicazione: (2024)
di: Thejaswi, Suhas, et al.
Pubblicazione: (2024)
Rethinking Model-based, Policy-based, and Value-based Reinforcement Learning via the Lens of Representation Complexity
di: Feng, Guhao, et al.
Pubblicazione: (2023)
di: Feng, Guhao, et al.
Pubblicazione: (2023)
$\mathrm{TIME}[t]\subseteq \mathrm{SPACE}[O(\sqrt{t})]$ via Tree Height Compression
di: Nye, Logan
Pubblicazione: (2025)
di: Nye, Logan
Pubblicazione: (2025)
Efficient Turing Machine Simulation with Transformers
di: Li, Qian, et al.
Pubblicazione: (2025)
di: Li, Qian, et al.
Pubblicazione: (2025)
Fast Approximation Algorithm for Non-Monotone DR-submodular Maximization under Size Constraint
di: Tran, Tan D., et al.
Pubblicazione: (2025)
di: Tran, Tan D., et al.
Pubblicazione: (2025)
Avoiding Obfuscation with Prover-Estimator Debate
di: Brown-Cohen, Jonah, et al.
Pubblicazione: (2025)
di: Brown-Cohen, Jonah, et al.
Pubblicazione: (2025)
Sorting by Strip Swaps is NP-Hard
di: Roy, Swapnoneel, et al.
Pubblicazione: (2025)
di: Roy, Swapnoneel, et al.
Pubblicazione: (2025)
Kidney Exchange: Faster Parameterized Algorithms and Tighter Lower Bounds
di: Banik, Aritra, et al.
Pubblicazione: (2025)
di: Banik, Aritra, et al.
Pubblicazione: (2025)
Fast-MWEM: Private Data Release in Sublinear Time
di: Haris, Themistoklis, et al.
Pubblicazione: (2026)
di: Haris, Themistoklis, et al.
Pubblicazione: (2026)
Efficient and Private Property Testing via Indistinguishability
di: Dwork, Cynthia, et al.
Pubblicazione: (2025)
di: Dwork, Cynthia, et al.
Pubblicazione: (2025)
A Distributional-Lifting Theorem for PAC Learning
di: Blanc, Guy, et al.
Pubblicazione: (2025)
di: Blanc, Guy, et al.
Pubblicazione: (2025)
Feature Selection and Junta Testing are Statistically Equivalent
di: Beretta, Lorenzo, et al.
Pubblicazione: (2025)
di: Beretta, Lorenzo, et al.
Pubblicazione: (2025)
Cascaded Learned Bloom Filter for Optimal Model-Filter Size Balance and Fast Rejection
di: Sato, Atsuki, et al.
Pubblicazione: (2025)
di: Sato, Atsuki, et al.
Pubblicazione: (2025)
Samplability makes learning easier
di: Blanc, Guy, et al.
Pubblicazione: (2025)
di: Blanc, Guy, et al.
Pubblicazione: (2025)
Is nasty noise actually harder than malicious noise?
di: Blanc, Guy, et al.
Pubblicazione: (2025)
di: Blanc, Guy, et al.
Pubblicazione: (2025)
Rate-optimal community detection near the KS threshold via node-robust algorithms
di: Ding, Jingqiu, et al.
Pubblicazione: (2025)
di: Ding, Jingqiu, et al.
Pubblicazione: (2025)
Computational-Statistical Tradeoffs from NP-hardness
di: Blanc, Guy, et al.
Pubblicazione: (2025)
di: Blanc, Guy, et al.
Pubblicazione: (2025)
Learning-Augmented Algorithms for Boolean Satisfiability
di: Attias, Idan, et al.
Pubblicazione: (2025)
di: Attias, Idan, et al.
Pubblicazione: (2025)
The Computational Complexity of Almost Stable Clustering with Penalties
di: Khodamoradi, Kamyar, et al.
Pubblicazione: (2025)
di: Khodamoradi, Kamyar, et al.
Pubblicazione: (2025)
Supersimulators
di: Dwork, Cynthia, et al.
Pubblicazione: (2025)
di: Dwork, Cynthia, et al.
Pubblicazione: (2025)
Superconstant Inapproximability of Decision Tree Learning
di: Koch, Caleb, et al.
Pubblicazione: (2024)
di: Koch, Caleb, et al.
Pubblicazione: (2024)
Exact and Approximate Algorithms for Polytree Learning
di: Harviainen, Juha, et al.
Pubblicazione: (2026)
di: Harviainen, Juha, et al.
Pubblicazione: (2026)
Differentially Private Verification of Distribution Properties
di: Du, Elbert, et al.
Pubblicazione: (2026)
di: Du, Elbert, et al.
Pubblicazione: (2026)
Fast decision tree learning solves hard coding-theoretic problems
di: Koch, Caleb, et al.
Pubblicazione: (2024)
di: Koch, Caleb, et al.
Pubblicazione: (2024)
Adaptive and oblivious statistical adversaries are equivalent
di: Blanc, Guy, et al.
Pubblicazione: (2024)
di: Blanc, Guy, et al.
Pubblicazione: (2024)
Private graphon estimation via sum-of-squares
di: Chen, Hongjie, et al.
Pubblicazione: (2024)
di: Chen, Hongjie, et al.
Pubblicazione: (2024)
Low-Degree Method Fails to Predict Robust Subspace Recovery
di: Jia, He, et al.
Pubblicazione: (2026)
di: Jia, He, et al.
Pubblicazione: (2026)
Omnipredictors for Regression and the Approximate Rank of Convex Functions
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
On the Power of Interactive Proofs for Learning
di: Gur, Tom, et al.
Pubblicazione: (2024)
di: Gur, Tom, et al.
Pubblicazione: (2024)
The Sample Complexity of Replicable Realizable PAC Learning
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
Active Learning for Decision Trees with Provable Guarantees
di: Moakhar, Arshia Soltani, et al.
Pubblicazione: (2026)
di: Moakhar, Arshia Soltani, et al.
Pubblicazione: (2026)
Low-degree phase transitions for detecting a planted clique in sublinear time
di: Mardia, Jay, et al.
Pubblicazione: (2024)
di: Mardia, Jay, et al.
Pubblicazione: (2024)
On the Hardness of Approximation of the Fair k-Center Problem
di: Thejaswi, Suhas
Pubblicazione: (2026)
di: Thejaswi, Suhas
Pubblicazione: (2026)
AdaBoost is not an Optimal Weak to Strong Learner
di: Høgsgaard, Mikael Møller, et al.
Pubblicazione: (2023)
di: Høgsgaard, Mikael Møller, et al.
Pubblicazione: (2023)
Hardness of Maximum Likelihood Learning of DPPs
di: Grigorescu, Elena, et al.
Pubblicazione: (2022)
di: Grigorescu, Elena, et al.
Pubblicazione: (2022)
Hardness of Learning Boolean Functions from Label Proportions
di: Guruswami, Venkatesan, et al.
Pubblicazione: (2024)
di: Guruswami, Venkatesan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Theoretical limitations of multi-layer Transformer
di: Chen, Lijie, et al.
Pubblicazione: (2024) -
$k$NN Attention Demystified: A Theoretical Exploration for Scalable Transformers
di: Haris, Themistoklis
Pubblicazione: (2024) -
Efficient Algorithms for Adversarially Robust Approximate Nearest Neighbor Search
di: Andoni, Alexandr, et al.
Pubblicazione: (2026) -
Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods
di: Blum, Avrim, et al.
Pubblicazione: (2025) -
Diversity-aware clustering: Computational Complexity and Approximation Algorithms
di: Thejaswi, Suhas, et al.
Pubblicazione: (2024)