Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Bakman, Yavuz, Kang, Sungmin, Huang, Zhiqi, Yaldiz, Duygu Nur, Belém, Catarina G., Zhu, Chenyang, Kumar, Anoop, Samuel, Alfy, Avestimehr, Salman, Liu, Daben, Karimireddy, Sai Praneeth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026)
by: Bakman, Yavuz, et al.
Published: (2026)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)
by: Ziashahabi, Amir, et al.
Published: (2025)
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
by: Wang, Nien-Shao, et al.
Published: (2025)
by: Wang, Nien-Shao, et al.
Published: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
by: Oguz, Metehan, et al.
Published: (2025)
by: Oguz, Metehan, et al.
Published: (2025)
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
by: Mushtaq, Erum, et al.
Published: (2024)
by: Mushtaq, Erum, et al.
Published: (2024)
Confidence-Based Response Abstinence: Improving LLM Trustworthiness via Activation-Based Uncertainty Estimation
by: Huang, Zhiqi, et al.
Published: (2025)
by: Huang, Zhiqi, et al.
Published: (2025)
Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
by: Belem, Catarina G, et al.
Published: (2025)
by: Belem, Catarina G, et al.
Published: (2025)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
by: Mushtaq, Erum, et al.
Published: (2025)
by: Mushtaq, Erum, et al.
Published: (2025)
FB-RAG: Improving RAG with Forward and Backward Lookup
by: Chawla, Kushal, et al.
Published: (2025)
by: Chawla, Kushal, et al.
Published: (2025)
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
by: Lawton, Neal Gregory, et al.
Published: (2025)
by: Lawton, Neal Gregory, et al.
Published: (2025)
Uncertainty Quantification in Retrieval Augmented Question Answering
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
Improving Consistency in Retrieval-Augmented Systems with Group Similarity Rewards
by: Hamman, Faisal, et al.
Published: (2025)
by: Hamman, Faisal, et al.
Published: (2025)
Optimization with Access to Auxiliary Information
by: Chayti, El Mahdi, et al.
Published: (2022)
by: Chayti, El Mahdi, et al.
Published: (2022)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
by: Banayeeanzade, Amin, et al.
Published: (2026)
by: Banayeeanzade, Amin, et al.
Published: (2026)
Epistemic Wrapping for Uncertainty Quantification
by: Sultana, Maryam, et al.
Published: (2025)
by: Sultana, Maryam, et al.
Published: (2025)
Play by the Type Rules: Inferring Constraints for LLM Functions in Declarative Programs
by: Glenn, Parker, et al.
Published: (2025)
by: Glenn, Parker, et al.
Published: (2025)
LLM Optimization Unlocks Real-Time Pairwise Reranking
by: Wu, Jingyu, et al.
Published: (2025)
by: Wu, Jingyu, et al.
Published: (2025)
On the Limits of Momentum in Decentralized and Federated Optimization
by: Zaccone, Riccardo, et al.
Published: (2025)
by: Zaccone, Riccardo, et al.
Published: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
Improving Metacognition and Uncertainty Communication in Language Models
by: Steyvers, Mark, et al.
Published: (2025)
by: Steyvers, Mark, et al.
Published: (2025)
Label-wise Aleatoric and Epistemic Uncertainty Quantification
by: Sale, Yusuf, et al.
Published: (2024)
by: Sale, Yusuf, et al.
Published: (2024)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Collaborative Heterogeneous Causal Inference Beyond Meta-analysis
by: Guo, Tianyu, et al.
Published: (2024)
by: Guo, Tianyu, et al.
Published: (2024)
Defection-Free Collaboration between Competitors in a Learning System
by: Werner, Mariel, et al.
Published: (2024)
by: Werner, Mariel, et al.
Published: (2024)
Do Data Valuations Make Good Data Prices?
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
by: Jain, Neel, et al.
Published: (2024)
by: Jain, Neel, et al.
Published: (2024)
Mind the Ambiguity: Aleatoric Uncertainty Quantification in LLMs for Safe Medical Question Answering
by: Liu, Yaokun, et al.
Published: (2026)
by: Liu, Yaokun, et al.
Published: (2026)
Epistemic and Aleatoric Uncertainty Quantification in Weather and Climate Models
by: Mansfield, Laura A., et al.
Published: (2025)
by: Mansfield, Laura A., et al.
Published: (2025)
Laplacian Segmentation Networks Improve Epistemic Uncertainty Quantification
by: Zepf, Kilian, et al.
Published: (2023)
by: Zepf, Kilian, et al.
Published: (2023)
Epistemic Uncertainty Quantification For Pre-trained Neural Network
by: Wang, Hanjing, et al.
Published: (2024)
by: Wang, Hanjing, et al.
Published: (2024)
LIA: Privacy-Preserving Data Quality Evaluation in Federated Learning Using a Lazy Influence Approximation
by: Rokvic, Ljubomir, et al.
Published: (2022)
by: Rokvic, Ljubomir, et al.
Published: (2022)
Communication-Efficient Heterogeneous Federated Learning with Generalized Heavy-Ball Momentum
by: Zaccone, Riccardo, et al.
Published: (2023)
by: Zaccone, Riccardo, et al.
Published: (2023)
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
A Differentially Private Kaplan-Meier Estimator for Privacy-Preserving Survival Analysis
by: Veeraragavan, Narasimha Raghavan, et al.
Published: (2024)
by: Veeraragavan, Narasimha Raghavan, et al.
Published: (2024)
Alignment-Weighted DPO: A principled reasoning approach to improve safety alignment
by: Hu, Mengxuan, et al.
Published: (2026)
by: Hu, Mengxuan, et al.
Published: (2026)
Similar Items
-
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025) -
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
by: Yaldiz, Duygu Nur, et al.
Published: (2025) -
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)