Can LLMs Review Scientific Papers?

Fuente: Zenodo
Guardado en:
Detalles Bibliográficos
Autor principal: Ban, Byunghyun
Formato: Recurso digital
Publicado: Zenodo 2026
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866901396688732160
author Ban, Byunghyun
author_facet Ban, Byunghyun
contents <p class="MsoNormal"><em><span>Maybe not.</span></em></p> <p class="MsoNormal"><span>Large language models (LLMs) can produce fluent and plausible peer-review reports, but fluency is not scholarly judgment. This paper presents a case study of three formal peer-review reports received during an actual submission to a Korea Citation Index (KCI) candidate journal. One review claimed that key methodological and experimental details were missing, although the manuscript explicitly contained information about validation data, hardware, software frameworks, model configuration, parameter count, algorithmic procedure, and limitations. After the author reported these grounding failures to the editor, the problematic review was not relied upon in the final decision, and the manuscript was subsequently accepted and published. The case is compared with two more grounded human reviews and an exploratory source-grounded LLM-generated review. The source-grounded LLM review was more aligned with the manuscript, but still produced minor factual errors and overstatements. The case suggests that grounding improves LLM-assisted review, but does not eliminate hallucination or accountability gaps.</span></p>
format Recurso digital
id zenodo_https___doi_org_10_5281_zenodo_20054277
institution Zenodo
language
publishDate 2026
publisher Zenodo
record_format zenodo
spellingShingle Can LLMs Review Scientific Papers?
Ban, Byunghyun
<p class="MsoNormal"><em><span>Maybe not.</span></em></p> <p class="MsoNormal"><span>Large language models (LLMs) can produce fluent and plausible peer-review reports, but fluency is not scholarly judgment. This paper presents a case study of three formal peer-review reports received during an actual submission to a Korea Citation Index (KCI) candidate journal. One review claimed that key methodological and experimental details were missing, although the manuscript explicitly contained information about validation data, hardware, software frameworks, model configuration, parameter count, algorithmic procedure, and limitations. After the author reported these grounding failures to the editor, the problematic review was not relied upon in the final decision, and the manuscript was subsequently accepted and published. The case is compared with two more grounded human reviews and an exploratory source-grounded LLM-generated review. The source-grounded LLM review was more aligned with the manuscript, but still produced minor factual errors and overstatements. The case suggests that grounding improves LLM-assisted review, but does not eliminate hallucination or accountability gaps.</span></p>
title Can LLMs Review Scientific Papers?
url https://doi.org/10.5281/zenodo.20054277