Gong, P., Yi, J., Wang, S., Zhang, J., Jin, Z., Zhou, O., . . . Li, C. (2025). HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference.
Chicago Style (17th ed.) CitationGong, Ping, et al. HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference. 2025.
MLA (9th ed.) CitationGong, Ping, et al. HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference. 2025.
Warning: These citations may not always be 100% accurate.