zhang, R., li, z., Huang, J., Zhang, R., Xu, X., zhe, s., . . . Wang, C. (2026). From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning.
Chicago Style (17th ed.) Citationzhang, Ranxu, zeyang li, Jiacheng Huang, Rui Zhang, Xiaozhou Xu, sun zhe, Yanyong Zhang, and Chao Wang. From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning. 2026.
MLA (9th ed.) Citationzhang, Ranxu, et al. From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning. 2026.
Warning: These citations may not always be 100% accurate.