Kim, S., Kang, D., Kwon, T., Chae, H., Won, J., Lee, D., & Yeo, J. (2024). Evaluating Robustness of Reward Models for Mathematical Reasoning.
Chicago-Zitierstil (17. Ausg.)Kim, Sunghwan, Dongjin Kang, Taeyoon Kwon, Hyungjoo Chae, Jungsoo Won, Dongha Lee, und Jinyoung Yeo. Evaluating Robustness of Reward Models for Mathematical Reasoning. 2024.
MLA-Zitierstil (9. Ausg.)Kim, Sunghwan, et al. Evaluating Robustness of Reward Models for Mathematical Reasoning. 2024.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.