Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Kim, Minsu, Kim, Sangryul, Thorne, James
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:https://arxiv.org/abs/2504.19622
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866915262828118016
author Kim, Minsu
Kim, Sangryul
Thorne, James
author_facet Kim, Minsu
Kim, Sangryul
Thorne, James
contents This paper investigates the knowledge of language models from the perspective of Bayesian epistemology. We explore how language models adjust their confidence and responses when presented with evidence with varying levels of informativeness and reliability. To study these properties, we create a dataset with various types of evidence and analyze language models' responses and confidence using verbalized confidence, token probability, and sampling. We observed that language models do not consistently follow Bayesian epistemology: language models follow the Bayesian confirmation assumption well with true evidence but fail to adhere to other Bayesian assumptions when encountering different evidence types. Also, we demonstrated that language models can exhibit high confidence when given strong evidence, but this does not always guarantee high accuracy. Our analysis also reveals that language models are biased toward golden evidence and show varying performance depending on the degree of irrelevance, helping explain why they deviate from Bayesian assumptions.
format Preprint
id arxiv_https___arxiv_org_abs_2504_19622
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle From Evidence to Belief: A Bayesian Epistemology Approach to Language Models
Kim, Minsu
Kim, Sangryul
Thorne, James
Artificial Intelligence
This paper investigates the knowledge of language models from the perspective of Bayesian epistemology. We explore how language models adjust their confidence and responses when presented with evidence with varying levels of informativeness and reliability. To study these properties, we create a dataset with various types of evidence and analyze language models' responses and confidence using verbalized confidence, token probability, and sampling. We observed that language models do not consistently follow Bayesian epistemology: language models follow the Bayesian confirmation assumption well with true evidence but fail to adhere to other Bayesian assumptions when encountering different evidence types. Also, we demonstrated that language models can exhibit high confidence when given strong evidence, but this does not always guarantee high accuracy. Our analysis also reveals that language models are biased toward golden evidence and show varying performance depending on the degree of irrelevance, helping explain why they deviate from Bayesian assumptions.
title From Evidence to Belief: A Bayesian Epistemology Approach to Language Models
topic Artificial Intelligence
url https://arxiv.org/abs/2504.19622