Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Topsakal, Oguzhan, Edell, Colby Jacob, Harper, Jackson Bailey
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!