Certain but not Probable? Differentiating Certainty from Probability in LLM Token Outputs for Probabilistic Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Toney-Wails, Autumn, Wails, Ryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
von: Ansell, Rebecca, et al.
Veröffentlicht: (2026)
von: Ansell, Rebecca, et al.
Veröffentlicht: (2026)
AI on AI: Exploring the Utility of GPT as an Expert Annotator of AI Publications
von: Toney-Wails, Autumn, et al.
Veröffentlicht: (2024)
von: Toney-Wails, Autumn, et al.
Veröffentlicht: (2024)
Understanding Token Probability Encoding in Output Embeddings
von: Cho, Hakaze, et al.
Veröffentlicht: (2024)
von: Cho, Hakaze, et al.
Veröffentlicht: (2024)
A Measurement of Genuine Tor Traces for Realistic Website Fingerprinting
von: Jansen, Rob, et al.
Veröffentlicht: (2024)
von: Jansen, Rob, et al.
Veröffentlicht: (2024)
Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference
von: Zhao, Stephen, et al.
Veröffentlicht: (2025)
von: Zhao, Stephen, et al.
Veröffentlicht: (2025)
Training-free LLM-generated Text Detection by Mining Token Probability Sequences
von: Xu, Yihuai, et al.
Veröffentlicht: (2024)
von: Xu, Yihuai, et al.
Veröffentlicht: (2024)
Probability Distributions Computed by Autoregressive Transformers
von: Yang, Andy, et al.
Veröffentlicht: (2025)
von: Yang, Andy, et al.
Veröffentlicht: (2025)
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
First Token Probability Guided RAG for Telecom Question Answering
von: Chen, Tingwei, et al.
Veröffentlicht: (2025)
von: Chen, Tingwei, et al.
Veröffentlicht: (2025)
TPA: Next Token Probability Attribution for Detecting Hallucinations in RAG
von: Lu, Pengqian, et al.
Veröffentlicht: (2025)
von: Lu, Pengqian, et al.
Veröffentlicht: (2025)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
von: Yang, Yunqiao, et al.
Veröffentlicht: (2025)
von: Yang, Yunqiao, et al.
Veröffentlicht: (2025)
On the Salience of Low-Probability Tokens for AI-Generated Text Detection: A Multiscale Uncertainty Perspective
von: Guo, Yikai, et al.
Veröffentlicht: (2026)
von: Guo, Yikai, et al.
Veröffentlicht: (2026)
Fact-Checking with Large Language Models via Probabilistic Certainty and Consistency
von: Wang, Haoran, et al.
Veröffentlicht: (2026)
von: Wang, Haoran, et al.
Veröffentlicht: (2026)
Do Not Let Low-Probability Tokens Over-Dominate in RL for LLMs
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
von: Du, Yanfan, et al.
Veröffentlicht: (2025)
von: Du, Yanfan, et al.
Veröffentlicht: (2025)
Alignment-Enhanced Decoding:Defending via Token-Level Adaptive Refining of Probability Distributions
von: Liu, Quan, et al.
Veröffentlicht: (2024)
von: Liu, Quan, et al.
Veröffentlicht: (2024)
Textual Entailment is not a Better Bias Metric than Token Probability
von: Felkner, Virginia K., et al.
Veröffentlicht: (2025)
von: Felkner, Virginia K., et al.
Veröffentlicht: (2025)
How to Compute the Probability of a Word
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection
von: Liu, Tao, et al.
Veröffentlicht: (2026)
von: Liu, Tao, et al.
Veröffentlicht: (2026)
Zonkey: A Hierarchical Diffusion Language Model with Differentiable Tokenization and Probabilistic Attention
von: Rozental, Alon
Veröffentlicht: (2026)
von: Rozental, Alon
Veröffentlicht: (2026)
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
Rehearsing Answers to Probable Questions with Perspective-Taking
von: Shih, Yung-Yu, et al.
Veröffentlicht: (2024)
von: Shih, Yung-Yu, et al.
Veröffentlicht: (2024)
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
A Probability--Quality Trade-off in Aligned Language Models and its Relation to Sampling Adaptors
von: Tan, Naaman, et al.
Veröffentlicht: (2024)
von: Tan, Naaman, et al.
Veröffentlicht: (2024)
Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
Detecting Hallucinations in Large Language Model Generation: A Token Probability Approach
von: Quevedo, Ernesto, et al.
Veröffentlicht: (2024)
von: Quevedo, Ernesto, et al.
Veröffentlicht: (2024)
Safeguarding RAG Pipelines with GMTP: A Gradient-based Masked Token Probability Method for Poisoned Document Detection
von: Kim, San, et al.
Veröffentlicht: (2025)
von: Kim, San, et al.
Veröffentlicht: (2025)
Calibrating Expressions of Certainty
von: Wang, Peiqi, et al.
Veröffentlicht: (2024)
von: Wang, Peiqi, et al.
Veröffentlicht: (2024)
How Should We Model the Probability of a Language?
von: Dent, Rasul, et al.
Veröffentlicht: (2026)
von: Dent, Rasul, et al.
Veröffentlicht: (2026)
Back to Basics: Revisiting Exploration in Reinforcement Learning for LLM Reasoning via Generative Probabilities
von: Li, Pengyi, et al.
Veröffentlicht: (2026)
von: Li, Pengyi, et al.
Veröffentlicht: (2026)
Improving LLM First-Token Predictions in Multiple-Choice Question Answering via Output Prefilling
von: Cappelletti, Silvia, et al.
Veröffentlicht: (2025)
von: Cappelletti, Silvia, et al.
Veröffentlicht: (2025)
Certainty robustness: Evaluating LLM stability under self-challenging prompts
von: Saadat, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Saadat, Mohammadreza, et al.
Veröffentlicht: (2026)
Calibrating Verbalized Probabilities for Large Language Models
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
Incoherent Probability Judgments in Large Language Models
von: Zhu, Jian-Qiao, et al.
Veröffentlicht: (2024)
von: Zhu, Jian-Qiao, et al.
Veröffentlicht: (2024)
PRISM: Probability Reallocation with In-Span Masking for Knowledge-Sensitive Alignment
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
Finding Replicable Human Evaluations via Stable Ranking Probability
von: Riley, Parker, et al.
Veröffentlicht: (2024)
von: Riley, Parker, et al.
Veröffentlicht: (2024)
Beyond Probabilities: Unveiling the Misalignment in Evaluating Large Language Models
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
von: Baan, Joris, et al.
Veröffentlicht: (2024)
von: Baan, Joris, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
von: Ansell, Rebecca, et al.
Veröffentlicht: (2026) -
AI on AI: Exploring the Utility of GPT as an Expert Annotator of AI Publications
von: Toney-Wails, Autumn, et al.
Veröffentlicht: (2024) -
Understanding Token Probability Encoding in Output Embeddings
von: Cho, Hakaze, et al.
Veröffentlicht: (2024) -
A Measurement of Genuine Tor Traces for Realistic Website Fingerprinting
von: Jansen, Rob, et al.
Veröffentlicht: (2024) -
Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference
von: Zhao, Stephen, et al.
Veröffentlicht: (2025)