Qin, K., Liu, L., Liang, Y., Wang, L., Wang, Y., Zhang, Y., . . . Shi, D. (2026). ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework.
Style de citation Chicago (17e éd.)Qin, Kai, et al. ReflectRM: Boosting Generative Reward Models via Self-Reflection Within a Unified Judgment Framework. 2026.
Style de citation MLA (9e éd.)Qin, Kai, et al. ReflectRM: Boosting Generative Reward Models via Self-Reflection Within a Unified Judgment Framework. 2026.
Attention : ces citations peuvent ne pas être correctes à 100%.