DecoPrompt : Decoding Prompts Reduces Hallucinations when Large Language Models Meet False Premises

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Xu, Nan, Ma, Xuezhe
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866910792693055488
author Xu, Nan
Ma, Xuezhe
author_facet Xu, Nan
Ma, Xuezhe
contents While large language models (LLMs) have demonstrated increasing power, they have also called upon studies on their hallucinated outputs that deviate from factually correct statements. In this paper, we focus on one important scenario of false premises, where LLMs are distracted by misaligned claims although the model possesses the required factual knowledge to answer original questions accurately. Inspired by the observation that entropy of the false-premise prompt is closely related to its likelihood to elicit hallucination generation, we propose a new prompting algorithm, named DecoPrompt, to mitigate hallucination. DecoPrompt leverages LLMs to "decode" the false-premise prompts without really eliciting hallucination output from LLMs. We perform experiments on two datasets, demonstrating that DecoPrompt can reduce hallucinations effectively on outputs from different LLMs. Moreover, DecoPrompt exhibits cross-model transferability, which facilitates its applications to scenarios such as LLMs of large sizes or unavailable model logits.
format Preprint
id arxiv_https___arxiv_org_abs_2411_07457
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle DecoPrompt : Decoding Prompts Reduces Hallucinations when Large Language Models Meet False Premises
Xu, Nan
Ma, Xuezhe
Computation and Language
While large language models (LLMs) have demonstrated increasing power, they have also called upon studies on their hallucinated outputs that deviate from factually correct statements. In this paper, we focus on one important scenario of false premises, where LLMs are distracted by misaligned claims although the model possesses the required factual knowledge to answer original questions accurately. Inspired by the observation that entropy of the false-premise prompt is closely related to its likelihood to elicit hallucination generation, we propose a new prompting algorithm, named DecoPrompt, to mitigate hallucination. DecoPrompt leverages LLMs to "decode" the false-premise prompts without really eliciting hallucination output from LLMs. We perform experiments on two datasets, demonstrating that DecoPrompt can reduce hallucinations effectively on outputs from different LLMs. Moreover, DecoPrompt exhibits cross-model transferability, which facilitates its applications to scenarios such as LLMs of large sizes or unavailable model logits.
title DecoPrompt : Decoding Prompts Reduces Hallucinations when Large Language Models Meet False Premises
topic Computation and Language
url https://arxiv.org/abs/2411.07457