Mu, Y., Wu, Y., Fan, Y., Wang, C., Li, H., Zeng, J., . . . Zhu, J. (2024). Cross-layer Attention Sharing for Pre-trained Large Language Models.
Style de citation Chicago (17e éd.)Mu, Yongyu, et al. Cross-layer Attention Sharing for Pre-trained Large Language Models. 2024.
Style de citation MLA (9e éd.)Mu, Yongyu, et al. Cross-layer Attention Sharing for Pre-trained Large Language Models. 2024.
Attention : ces citations peuvent ne pas être correctes à 100%.