Feng, Y., Kwiatkowski, A., Zheng, K., Kempe, J., & Duan, Y. (2025). PILAF: Optimal Human Preference Sampling for Reward Modeling.
Citazione stile Chigago Style (17a edizione)Feng, Yunzhen, Ariel Kwiatkowski, Kunhao Zheng, Julia Kempe, e Yaqi Duan. PILAF: Optimal Human Preference Sampling for Reward Modeling. 2025.
Citatione MLA (9a ed.)Feng, Yunzhen, et al. PILAF: Optimal Human Preference Sampling for Reward Modeling. 2025.
Attenzione: Queste citazioni potrebbero non essere precise al 100%.