A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions
Fuente:
arXiv
Saved in:
| Main Authors: | Shorinwa, Ola, Mei, Zhiting, Lidard, Justin, Ren, Allen Z., Majumdar, Anirudha |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
by: Yin, Tenny, et al.
Published: (2025)
by: Yin, Tenny, et al.
Published: (2025)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Video Generation Models in Robotics -- Applications, Research Challenges, Future Directions
by: Mei, Zhiting, et al.
Published: (2026)
by: Mei, Zhiting, et al.
Published: (2026)
Perceive With Confidence: Statistical Safety Assurances for Navigation with Learning-Based Perception
by: Mei, Zhiting, et al.
Published: (2024)
by: Mei, Zhiting, et al.
Published: (2024)
Consistency in Language Models: Current Landscape, Challenges, and Future Directions
by: Novikova, Jekaterina, et al.
Published: (2025)
by: Novikova, Jekaterina, et al.
Published: (2025)
A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future
by: Zhong, Jialun, et al.
Published: (2025)
by: Zhong, Jialun, et al.
Published: (2025)
PlayWorld: Learning Robot World Models from Autonomous Play
by: Yin, Tenny, et al.
Published: (2026)
by: Yin, Tenny, et al.
Published: (2026)
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
by: Shorinwa, Ola, et al.
Published: (2025)
by: Shorinwa, Ola, et al.
Published: (2025)
SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models
by: Gao, Xiang, et al.
Published: (2024)
by: Gao, Xiang, et al.
Published: (2024)
Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models
by: Zhou, Kevin, et al.
Published: (2025)
by: Zhou, Kevin, et al.
Published: (2025)
Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions
by: Zhang, Kun, et al.
Published: (2025)
by: Zhang, Kun, et al.
Published: (2025)
Compound AI Systems Optimization: A Survey of Methods, Challenges, and Future Directions
by: Lee, Yu-Ang, et al.
Published: (2025)
by: Lee, Yu-Ang, et al.
Published: (2025)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
by: Li, Huiqiong, et al.
Published: (2026)
by: Li, Huiqiong, et al.
Published: (2026)
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
by: Xiao, Quan, et al.
Published: (2025)
by: Xiao, Quan, et al.
Published: (2025)
Guiding Data Collection via Factored Scaling Curves
by: Zha, Lihan, et al.
Published: (2025)
by: Zha, Lihan, et al.
Published: (2025)
Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses
by: Chen, Tiejin, et al.
Published: (2025)
by: Chen, Tiejin, et al.
Published: (2025)
Uncertainty Quantification in Large Language Models Through Convex Hull Analysis
by: Catak, Ferhat Ozgur, et al.
Published: (2024)
by: Catak, Ferhat Ozgur, et al.
Published: (2024)
Exploring Large Language Models for Multimodal Sentiment Analysis: Challenges, Benchmarks, and Future Directions
by: Song, Shezheng
Published: (2024)
by: Song, Shezheng
Published: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
by: Hu, Yiran, et al.
Published: (2026)
by: Hu, Yiran, et al.
Published: (2026)
Survey of Natural Language Processing for Education: Taxonomy, Systematic Review, and Future Trends
by: Lan, Yunshi, et al.
Published: (2024)
by: Lan, Yunshi, et al.
Published: (2024)
SIMBA UQ: Similarity-Based Aggregation for Uncertainty Quantification in Large Language Models
by: Bhattacharjya, Debarun, et al.
Published: (2025)
by: Bhattacharjya, Debarun, et al.
Published: (2025)
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
by: Huang, Lei, et al.
Published: (2023)
by: Huang, Lei, et al.
Published: (2023)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
by: Cao, Qi, et al.
Published: (2026)
by: Cao, Qi, et al.
Published: (2026)
UQLM: A Python Package for Uncertainty Quantification in Large Language Models
by: Bouchard, Dylan, et al.
Published: (2025)
by: Bouchard, Dylan, et al.
Published: (2025)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
by: Grewal, Yashvir S., et al.
Published: (2024)
by: Grewal, Yashvir S., et al.
Published: (2024)
SciFi-Benchmark: Leveraging Science Fiction To Improve Robot Behavior
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
From Intentions to Techniques: A Comprehensive Taxonomy and Challenges in Text Watermarking for Large Language Models
by: Lalai, Harsh Nishant, et al.
Published: (2024)
by: Lalai, Harsh Nishant, et al.
Published: (2024)
Risk-Calibrated Human-Robot Interaction via Set-Valued Intent Prediction
by: Lidard, Justin, et al.
Published: (2024)
by: Lidard, Justin, et al.
Published: (2024)
Trustworthy Summarization via Uncertainty Quantification and Risk Awareness in Large Language Models
by: Pan, Shuaidong, et al.
Published: (2025)
by: Pan, Shuaidong, et al.
Published: (2025)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
by: Fan, Haozhi, et al.
Published: (2026)
by: Fan, Haozhi, et al.
Published: (2026)
URAG: A Benchmark for Uncertainty Quantification in Retrieval-Augmented Large Language Models
by: Nguyen, Vinh, et al.
Published: (2026)
by: Nguyen, Vinh, et al.
Published: (2026)
Semantic Density: Uncertainty Quantification for Large Language Models through Confidence Measurement in Semantic Space
by: Qiu, Xin, et al.
Published: (2024)
by: Qiu, Xin, et al.
Published: (2024)
Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Models in Clinical Text
by: Xiao, Bushi, et al.
Published: (2026)
by: Xiao, Bushi, et al.
Published: (2026)
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods
by: Cao, Yuji, et al.
Published: (2024)
by: Cao, Yuji, et al.
Published: (2024)
SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory
by: Zhao, Xingtao, et al.
Published: (2025)
by: Zhao, Xingtao, et al.
Published: (2025)
Similar Items
-
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
by: Mei, Zhiting, et al.
Published: (2025) -
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields
by: Mei, Zhiting, et al.
Published: (2025) -
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
by: Yin, Tenny, et al.
Published: (2025)