Huang, H., Bu, X., Zhou, H., Qu, Y., Liu, J., Yang, M., . . . Zhao, T. (2024). An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4.
Citazione stile Chigago Style (17a edizione)Huang, Hui, Xingyuan Bu, Hongli Zhou, Yingqi Qu, Jing Liu, Muyun Yang, Bing Xu, e Tiejun Zhao. An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model Is Not a General Substitute for GPT-4. 2024.
Citatione MLA (9a ed.)Huang, Hui, et al. An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model Is Not a General Substitute for GPT-4. 2024.
Attenzione: Queste citazioni potrebbero non essere precise al 100%.