Support Basis: Fast Attention Beyond Bounded Entries
Fuente:
arXiv
Saved in:
| Main Authors: | Aliakbarpour, Maryam, Braverman, Vladimir, Yin, Junze, Zhang, Haochen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
High-Dimensional Robust Mean Estimation with Untrusted Batches
by: Aliakbarpour, Maryam, et al.
Published: (2026)
by: Aliakbarpour, Maryam, et al.
Published: (2026)
CoVE: Compressed Vocabulary Expansion Makes Better LLM-based Recommender Systems
by: Zhang, Haochen, et al.
Published: (2025)
by: Zhang, Haochen, et al.
Published: (2025)
Breaking the Frozen Subspace: Importance Sampling for Low-Rank Optimization in LLM Pretraining
by: Zhang, Haochen, et al.
Published: (2025)
by: Zhang, Haochen, et al.
Published: (2025)
High-Probability Bounds For Heterogeneous Local Differential Privacy
by: Aliakbarpour, Maryam, et al.
Published: (2025)
by: Aliakbarpour, Maryam, et al.
Published: (2025)
Auditing Information Disclosure During LLM-Scale Gradient Descent Using Gradient Uniqueness
by: Abdelghafar, Sleem, et al.
Published: (2025)
by: Abdelghafar, Sleem, et al.
Published: (2025)
Conv-Basis: A New Paradigm for Efficient Attention Inference and Gradient Computation in Transformers
by: Liang, Yingyu, et al.
Published: (2024)
by: Liang, Yingyu, et al.
Published: (2024)
Solving Attention Kernel Regression Problem via Pre-conditioner
by: Song, Zhao, et al.
Published: (2023)
by: Song, Zhao, et al.
Published: (2023)
Optimal Prediction-Augmented Algorithms for Testing Independence of Distributions
by: Aliakbarpour, Maryam, et al.
Published: (2026)
by: Aliakbarpour, Maryam, et al.
Published: (2026)
Fast and Efficient Matching Algorithm with Deadline Instances
by: Song, Zhao, et al.
Published: (2023)
by: Song, Zhao, et al.
Published: (2023)
Optimal Algorithms for Augmented Testing of Discrete Distributions
by: Aliakbarpour, Maryam, et al.
Published: (2024)
by: Aliakbarpour, Maryam, et al.
Published: (2024)
Better Private Distribution Testing by Leveraging Unverified Auxiliary Data
by: Aliakbarpour, Maryam, et al.
Published: (2025)
by: Aliakbarpour, Maryam, et al.
Published: (2025)
How fast can you find a good hypothesis?
by: Aamand, Anders, et al.
Published: (2025)
by: Aamand, Anders, et al.
Published: (2025)
Privacy in Metalearning and Multitask Learning: Modeling and Separations
by: Aliakbarpour, Maryam, et al.
Published: (2024)
by: Aliakbarpour, Maryam, et al.
Published: (2024)
Fixed Design Analysis of Regularization-Based Continual Learning
by: Li, Haoran, et al.
Published: (2023)
by: Li, Haoran, et al.
Published: (2023)
Adversarially robust quantum state learning and testing
by: Aliakbarpour, Maryam, et al.
Published: (2025)
by: Aliakbarpour, Maryam, et al.
Published: (2025)
Nearly-Linear Time Private Hypothesis Selection with the Optimal Approximation Factor
by: Aliakbarpour, Maryam, et al.
Published: (2025)
by: Aliakbarpour, Maryam, et al.
Published: (2025)
Gap-Dependent Bounds for Federated $Q$-learning
by: Zhang, Haochen, et al.
Published: (2025)
by: Zhang, Haochen, et al.
Published: (2025)
Accelerating Attention with Basis Decomposition
by: Zhao, Jialin
Published: (2025)
by: Zhao, Jialin
Published: (2025)
Binary Hypothesis Testing for Softmax Models and Leverage Score Models
by: Gu, Yuzhou, et al.
Published: (2024)
by: Gu, Yuzhou, et al.
Published: (2024)
An Iterative Algorithm for Rescaled Hyperbolic Functions Regression
by: Gao, Yeqi, et al.
Published: (2023)
by: Gao, Yeqi, et al.
Published: (2023)
On the Structure of Replicable Hypothesis Testers
by: Aamand, Anders, et al.
Published: (2025)
by: Aamand, Anders, et al.
Published: (2025)
Enhancing Feature-Specific Data Protection via Bayesian Coordinate Differential Privacy
by: Aliakbarpour, Maryam, et al.
Published: (2024)
by: Aliakbarpour, Maryam, et al.
Published: (2024)
Memory-Statistics Tradeoff in Continual Learning with Structural Regularization
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Gap-Dependent Bounds for Q-Learning using Reference-Advantage Decomposition
by: Zheng, Zhong, et al.
Published: (2024)
by: Zheng, Zhong, et al.
Published: (2024)
Metalearning with Very Few Samples Per Task
by: Aliakbarpour, Maryam, et al.
Published: (2023)
by: Aliakbarpour, Maryam, et al.
Published: (2023)
BasisFormer: Attention-based Time Series Forecasting with Learnable and Interpretable Basis
by: Ni, Zelin, et al.
Published: (2023)
by: Ni, Zelin, et al.
Published: (2023)
Gap-Dependent Bounds for Nearly Minimax Optimal Reinforcement Learning with Linear Function Approximation
by: Zhang, Haochen, et al.
Published: (2026)
by: Zhang, Haochen, et al.
Published: (2026)
Gradient Shaping Beyond Clipping: A Functional Perspective on Update Magnitude Control
by: You, Haochen, et al.
Published: (2025)
by: You, Haochen, et al.
Published: (2025)
Efficient Alternating Minimization with Applications to Weighted Low Rank Approximation
by: Song, Zhao, et al.
Published: (2023)
by: Song, Zhao, et al.
Published: (2023)
Inverting the Leverage Score Gradient: An Efficient Approximate Newton Method
by: Li, Chenyang, et al.
Published: (2024)
by: Li, Chenyang, et al.
Published: (2024)
Chunk-wise Attention Transducers for Fast and Accurate Streaming Speech-to-Text
by: Xu, Hainan, et al.
Published: (2026)
by: Xu, Hainan, et al.
Published: (2026)
FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control
by: Xue, Jun, et al.
Published: (2026)
by: Xue, Jun, et al.
Published: (2026)
Entry-Specific Bounds for Low-Rank Matrix Completion under Highly Non-Uniform Sampling
by: Xi, Xumei, et al.
Published: (2024)
by: Xi, Xumei, et al.
Published: (2024)
Learning-augmented Maximum Independent Set
by: Braverman, Vladimir, et al.
Published: (2024)
by: Braverman, Vladimir, et al.
Published: (2024)
Dynamic Maintenance of Kernel Density Estimation Data Structure: From Practice to Theory
by: Liang, Jiehao, et al.
Published: (2022)
by: Liang, Jiehao, et al.
Published: (2022)
How to Inverting the Leverage Score Distribution?
by: Li, Zhihang, et al.
Published: (2024)
by: Li, Zhihang, et al.
Published: (2024)
Low Rank Matrix Completion via Robust Alternating Minimization in Nearly Linear Time
by: Gu, Yuzhou, et al.
Published: (2023)
by: Gu, Yuzhou, et al.
Published: (2023)
FastAttention: Extend FlashAttention2 to NPUs and Low-resource GPUs
by: Lin, Haoran, et al.
Published: (2024)
by: Lin, Haoran, et al.
Published: (2024)
Learning-Augmented Hierarchical Clustering
by: Braverman, Vladimir, et al.
Published: (2025)
by: Braverman, Vladimir, et al.
Published: (2025)
Computational Algebra with Attention: Transformer Oracles for Border Basis Algorithms
by: Kera, Hiroshi, et al.
Published: (2025)
by: Kera, Hiroshi, et al.
Published: (2025)
Similar Items
-
High-Dimensional Robust Mean Estimation with Untrusted Batches
by: Aliakbarpour, Maryam, et al.
Published: (2026) -
CoVE: Compressed Vocabulary Expansion Makes Better LLM-based Recommender Systems
by: Zhang, Haochen, et al.
Published: (2025) -
Breaking the Frozen Subspace: Importance Sampling for Low-Rank Optimization in LLM Pretraining
by: Zhang, Haochen, et al.
Published: (2025) -
High-Probability Bounds For Heterogeneous Local Differential Privacy
by: Aliakbarpour, Maryam, et al.
Published: (2025) -
Auditing Information Disclosure During LLM-Scale Gradient Descent Using Gradient Uniqueness
by: Abdelghafar, Sleem, et al.
Published: (2025)