LLM for Comparative Narrative Analysis

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Kampen, Leo, Villarreal, Carlos Rabat, Yu, Louis, Karmaker, Santu, Feng, Dongji
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866909575256473600
author Kampen, Leo
Villarreal, Carlos Rabat
Yu, Louis
Karmaker, Santu
Feng, Dongji
author_facet Kampen, Leo
Villarreal, Carlos Rabat
Yu, Louis
Karmaker, Santu
Feng, Dongji
contents In this paper, we conducted a Multi-Perspective Comparative Narrative Analysis (CNA) on three prominent LLMs: GPT-3.5, PaLM2, and Llama2. We applied identical prompts and evaluated their outputs on specific tasks, ensuring an equitable and unbiased comparison between various LLMs. Our study revealed that the three LLMs generated divergent responses to the same prompt, indicating notable discrepancies in their ability to comprehend and analyze the given task. Human evaluation was used as the gold standard, evaluating four perspectives to analyze differences in LLM performance.
format Preprint
id arxiv_https___arxiv_org_abs_2504_08211
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle LLM for Comparative Narrative Analysis
Kampen, Leo
Villarreal, Carlos Rabat
Yu, Louis
Karmaker, Santu
Feng, Dongji
Computation and Language
Artificial Intelligence
In this paper, we conducted a Multi-Perspective Comparative Narrative Analysis (CNA) on three prominent LLMs: GPT-3.5, PaLM2, and Llama2. We applied identical prompts and evaluated their outputs on specific tasks, ensuring an equitable and unbiased comparison between various LLMs. Our study revealed that the three LLMs generated divergent responses to the same prompt, indicating notable discrepancies in their ability to comprehend and analyze the given task. Human evaluation was used as the gold standard, evaluating four perspectives to analyze differences in LLM performance.
title LLM for Comparative Narrative Analysis
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2504.08211