Kim, T., Lee, J., Ahn, D., Kim, S., Choi, J., Kim, M., & Kim, H. (2024). QUICK: Quantization-aware Interleaving and Conflict-free Kernel for efficient LLM inference.
Style de citation Chicago (17e éd.)Kim, Taesu, Jongho Lee, Daehyun Ahn, Sarang Kim, Jiwoong Choi, Minkyu Kim, et Hyungjun Kim. QUICK: Quantization-aware Interleaving and Conflict-free Kernel for Efficient LLM Inference. 2024.
Style de citation MLA (9e éd.)Kim, Taesu, et al. QUICK: Quantization-aware Interleaving and Conflict-free Kernel for Efficient LLM Inference. 2024.
Attention : ces citations peuvent ne pas être correctes à 100%.