APA (7th ed.) Citation

Liao, Z., Gao, Y., Yang, Y., Hu, Y., & Ding, J. (2026). MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models.

Chicago Style (17th ed.) Citation

Liao, Zhaokang, Yingguo Gao, Yi Yang, Yongheng Hu, and Jingting Ding. MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models. 2026.

MLA (9th ed.) Citation

Liao, Zhaokang, et al. MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models. 2026.

Warning: These citations may not always be 100% accurate.