Beyond the Reported Cutoff: Where Large Language Models Fall Short on Financial Knowledge

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Shah, Agam, Ye, Liqin, Jaskowski, Sebastian, Xu, Wei, Chava, Sudheer
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866916867719823360
author Shah, Agam
Ye, Liqin
Jaskowski, Sebastian
Xu, Wei
Chava, Sudheer
author_facet Shah, Agam
Ye, Liqin
Jaskowski, Sebastian
Xu, Wei
Chava, Sudheer
contents Large Language Models (LLMs) are frequently utilized as sources of knowledge for question-answering. While it is known that LLMs may lack access to real-time data or newer data produced after the model's cutoff date, it is less clear how their knowledge spans across historical information. In this study, we assess the breadth of LLMs' knowledge using financial data of U.S. publicly traded companies by evaluating more than 197k questions and comparing model responses to factual data. We further explore the impact of company characteristics, such as size, retail investment, institutional attention, and readability of financial filings, on the accuracy of knowledge represented in LLMs. Our results reveal that LLMs are less informed about past financial performance, but they display a stronger awareness of larger companies and more recent information. Interestingly, at the same time, our analysis also reveals that LLMs are more likely to hallucinate for larger companies, especially for data from more recent years. The code, prompts, and model outputs are available on GitHub.
format Preprint
id arxiv_https___arxiv_org_abs_2504_00042
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Beyond the Reported Cutoff: Where Large Language Models Fall Short on Financial Knowledge
Shah, Agam
Ye, Liqin
Jaskowski, Sebastian
Xu, Wei
Chava, Sudheer
Computation and Language
Large Language Models (LLMs) are frequently utilized as sources of knowledge for question-answering. While it is known that LLMs may lack access to real-time data or newer data produced after the model's cutoff date, it is less clear how their knowledge spans across historical information. In this study, we assess the breadth of LLMs' knowledge using financial data of U.S. publicly traded companies by evaluating more than 197k questions and comparing model responses to factual data. We further explore the impact of company characteristics, such as size, retail investment, institutional attention, and readability of financial filings, on the accuracy of knowledge represented in LLMs. Our results reveal that LLMs are less informed about past financial performance, but they display a stronger awareness of larger companies and more recent information. Interestingly, at the same time, our analysis also reveals that LLMs are more likely to hallucinate for larger companies, especially for data from more recent years. The code, prompts, and model outputs are available on GitHub.
title Beyond the Reported Cutoff: Where Large Language Models Fall Short on Financial Knowledge
topic Computation and Language
url https://arxiv.org/abs/2504.00042