Efficient semantic uncertainty quantification in language models via diversity-steered sampling

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Park, Ji Won, Cho, Kyunghyun
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915750781911040
author Park, Ji Won
Cho, Kyunghyun
author_facet Park, Ji Won
Cho, Kyunghyun
contents Accurately estimating semantic aleatoric and epistemic uncertainties in large language models (LLMs) is particularly challenging in free-form question answering (QA), where obtaining stable estimates often requires many expensive generations. We introduce a diversity-steered sampler that discourages semantically redundant outputs during decoding, covers both autoregressive and masked diffusion paradigms, and yields substantial sample-efficiency gains. The key idea is to inject a continuous semantic-similarity penalty into the model's proposal distribution using a natural language inference (NLI) model lightly finetuned on partial prefixes or intermediate diffusion states. We debias downstream uncertainty estimates with importance reweighting and shrink their variance with control variates. Across four QA benchmarks, our method matches or surpasses baselines while covering more semantic clusters with the same number of samples. Being modular and requiring no gradient access to the base LLM, the framework promises to serve as a drop-in enhancement for uncertainty estimation in risk-sensitive model deployments.
format Preprint
id arxiv_https___arxiv_org_abs_2510_21310
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Efficient semantic uncertainty quantification in language models via diversity-steered sampling
Park, Ji Won
Cho, Kyunghyun
Computation and Language
Artificial Intelligence
Machine Learning
Accurately estimating semantic aleatoric and epistemic uncertainties in large language models (LLMs) is particularly challenging in free-form question answering (QA), where obtaining stable estimates often requires many expensive generations. We introduce a diversity-steered sampler that discourages semantically redundant outputs during decoding, covers both autoregressive and masked diffusion paradigms, and yields substantial sample-efficiency gains. The key idea is to inject a continuous semantic-similarity penalty into the model's proposal distribution using a natural language inference (NLI) model lightly finetuned on partial prefixes or intermediate diffusion states. We debias downstream uncertainty estimates with importance reweighting and shrink their variance with control variates. Across four QA benchmarks, our method matches or surpasses baselines while covering more semantic clusters with the same number of samples. Being modular and requiring no gradient access to the base LLM, the framework promises to serve as a drop-in enhancement for uncertainty estimation in risk-sensitive model deployments.
title Efficient semantic uncertainty quantification in language models via diversity-steered sampling
topic Computation and Language
Artificial Intelligence
Machine Learning
url https://arxiv.org/abs/2510.21310