Wang, J., Lin, M., An-Yeu, & Wu. (2024). LATTE: Low-Precision Approximate Attention with Head-wise Trainable Threshold for Efficient Transformer.
Chicago Style (17th ed.) CitationWang, Jiing-Ping, Ming-Guang Lin, An-Yeu, and Wu. LATTE: Low-Precision Approximate Attention with Head-wise Trainable Threshold for Efficient Transformer. 2024.
MLA (9th ed.) CitationWang, Jiing-Ping, et al. LATTE: Low-Precision Approximate Attention with Head-wise Trainable Threshold for Efficient Transformer. 2024.
Warning: These citations may not always be 100% accurate.