Liao, Z., Gao, Y., Yang, Y., Hu, Y., & Ding, J. (2026). MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models.
Chicago Style (17th ed.) CitationLiao, Zhaokang, Yingguo Gao, Yi Yang, Yongheng Hu, and Jingting Ding. MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models. 2026.
MLA (9th ed.) CitationLiao, Zhaokang, et al. MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models. 2026.
Warning: These citations may not always be 100% accurate.