Qiu, Z., Yu, S., Zhang, J., Zhang, S., Huang, X., Yang, J., & Lai, J. (2026). FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning.
Chicago Style (17th ed.) CitationQiu, Zhaopeng, Shuang Yu, Jingqi Zhang, Shuai Zhang, Xue Huang, Jingyi Yang, and Junjie Lai. FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning. 2026.
MLA (9th ed.) CitationQiu, Zhaopeng, et al. FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning. 2026.
Warning: These citations may not always be 100% accurate.