LLM for Comparative Narrative Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | , , , , |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
| _version_ | 1866909575256473600 |
|---|---|
| author | Kampen, Leo Villarreal, Carlos Rabat Yu, Louis Karmaker, Santu Feng, Dongji |
| author_facet | Kampen, Leo Villarreal, Carlos Rabat Yu, Louis Karmaker, Santu Feng, Dongji |
| contents | In this paper, we conducted a Multi-Perspective Comparative Narrative Analysis (CNA) on three prominent LLMs: GPT-3.5, PaLM2, and Llama2. We applied identical prompts and evaluated their outputs on specific tasks, ensuring an equitable and unbiased comparison between various LLMs. Our study revealed that the three LLMs generated divergent responses to the same prompt, indicating notable discrepancies in their ability to comprehend and analyze the given task. Human evaluation was used as the gold standard, evaluating four perspectives to analyze differences in LLM performance. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2504_08211 |
| institution | arXiv |
| publishDate | 2025 |
| record_format | arxiv |
| spellingShingle | LLM for Comparative Narrative Analysis Kampen, Leo Villarreal, Carlos Rabat Yu, Louis Karmaker, Santu Feng, Dongji Computation and Language Artificial Intelligence In this paper, we conducted a Multi-Perspective Comparative Narrative Analysis (CNA) on three prominent LLMs: GPT-3.5, PaLM2, and Llama2. We applied identical prompts and evaluated their outputs on specific tasks, ensuring an equitable and unbiased comparison between various LLMs. Our study revealed that the three LLMs generated divergent responses to the same prompt, indicating notable discrepancies in their ability to comprehend and analyze the given task. Human evaluation was used as the gold standard, evaluating four perspectives to analyze differences in LLM performance. |
| title | LLM for Comparative Narrative Analysis |
| topic | Computation and Language Artificial Intelligence |
| url | https://arxiv.org/abs/2504.08211 |