Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chern, Steffi, Chern, Ethan, Neubig, Graham, Liu, Pengfei
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!