IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Fan, Haozhi, Duan, Jinhao, Xu, Kaidi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
por: Duan, Jinhao, et al.
Publicado: (2023)
por: Duan, Jinhao, et al.
Publicado: (2023)
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
por: Wang, Zhiyuan, et al.
Publicado: (2024)
por: Wang, Zhiyuan, et al.
Publicado: (2024)
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
por: Wang, Zhiyuan, et al.
Publicado: (2024)
por: Wang, Zhiyuan, et al.
Publicado: (2024)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
por: Duan, Jinhao, et al.
Publicado: (2025)
por: Duan, Jinhao, et al.
Publicado: (2025)
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
por: Wang, Zhiyuan, et al.
Publicado: (2025)
por: Wang, Zhiyuan, et al.
Publicado: (2025)
Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
por: Bouchard, Dylan, et al.
Publicado: (2026)
por: Bouchard, Dylan, et al.
Publicado: (2026)
SConU: Selective Conformal Uncertainty in Large Language Models
por: Wang, Zhiyuan, et al.
Publicado: (2025)
por: Wang, Zhiyuan, et al.
Publicado: (2025)
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
por: Xiao, Quan, et al.
Publicado: (2025)
por: Xiao, Quan, et al.
Publicado: (2025)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
por: Cao, Qi, et al.
Publicado: (2026)
por: Cao, Qi, et al.
Publicado: (2026)
UQLM: A Python Package for Uncertainty Quantification in Large Language Models
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
por: Grewal, Yashvir S., et al.
Publicado: (2024)
por: Grewal, Yashvir S., et al.
Publicado: (2024)
Multi-group Uncertainty Quantification for Long-form Text Generation
por: Liu, Terrance, et al.
Publicado: (2024)
por: Liu, Terrance, et al.
Publicado: (2024)
Trustworthy Summarization via Uncertainty Quantification and Risk Awareness in Large Language Models
por: Pan, Shuaidong, et al.
Publicado: (2025)
por: Pan, Shuaidong, et al.
Publicado: (2025)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
por: Wang, Ziyu, et al.
Publicado: (2024)
por: Wang, Ziyu, et al.
Publicado: (2024)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
ESI: Epistemic Uncertainty Quantification via Semantic-preserving Intervention for Large Language Models
por: Li, Mingda, et al.
Publicado: (2025)
por: Li, Mingda, et al.
Publicado: (2025)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
por: Tao, Linwei, et al.
Publicado: (2025)
por: Tao, Linwei, et al.
Publicado: (2025)
Efficient Non-Parametric Uncertainty Quantification for Black-Box Large Language Models and Decision Planning
por: Tsai, Yao-Hung Hubert, et al.
Publicado: (2024)
por: Tsai, Yao-Hung Hubert, et al.
Publicado: (2024)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
por: Kirchhof, Michael, et al.
Publicado: (2025)
por: Kirchhof, Michael, et al.
Publicado: (2025)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
por: Pandit, Shrey, et al.
Publicado: (2025)
por: Pandit, Shrey, et al.
Publicado: (2025)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
por: Hu, Wenhao, et al.
Publicado: (2025)
por: Hu, Wenhao, et al.
Publicado: (2025)
Adaptive Uncertainty Quantification for Generative AI
por: Kim, Jungeum, et al.
Publicado: (2024)
por: Kim, Jungeum, et al.
Publicado: (2024)
Skewed Memorization in Large Language Models: Quantification and Decomposition
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
por: Duan, Jinhao, et al.
Publicado: (2024)
por: Duan, Jinhao, et al.
Publicado: (2024)
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
por: Santilli, Andrea, et al.
Publicado: (2025)
por: Santilli, Andrea, et al.
Publicado: (2025)
Uncertainty Quantification for Multimodal Large Language Models with Incoherence-adjusted Semantic Volume
por: Lau, Gregory Kang Ruey, et al.
Publicado: (2026)
por: Lau, Gregory Kang Ruey, et al.
Publicado: (2026)
Linguistic Calibration of Long-Form Generations
por: Band, Neil, et al.
Publicado: (2024)
por: Band, Neil, et al.
Publicado: (2024)
Calibrating Long-form Generations from Large Language Models
por: Huang, Yukun, et al.
Publicado: (2024)
por: Huang, Yukun, et al.
Publicado: (2024)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
por: Bouchard, Dylan, et al.
Publicado: (2026)
por: Bouchard, Dylan, et al.
Publicado: (2026)
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
por: Jiang, Mingjian, et al.
Publicado: (2024)
por: Jiang, Mingjian, et al.
Publicado: (2024)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
por: Nikitin, Alexander, et al.
Publicado: (2024)
por: Nikitin, Alexander, et al.
Publicado: (2024)
Uncertainty Quantification for Transformer Models for Dark-Pattern Detection
por: Muñoz, Javier, et al.
Publicado: (2024)
por: Muñoz, Javier, et al.
Publicado: (2024)
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
por: Hu, Zhiyuan, et al.
Publicado: (2024)
por: Hu, Zhiyuan, et al.
Publicado: (2024)
TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning
por: Zhang, Tunyu, et al.
Publicado: (2025)
por: Zhang, Tunyu, et al.
Publicado: (2025)
Representational Curvature Modulates Behavioral Uncertainty in Large Language Models
por: King, Jack, et al.
Publicado: (2026)
por: King, Jack, et al.
Publicado: (2026)
Benchmarking Benchmark Leakage in Large Language Models
por: Xu, Ruijie, et al.
Publicado: (2024)
por: Xu, Ruijie, et al.
Publicado: (2024)
An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms
por: Grünefeld, Nils, et al.
Publicado: (2026)
por: Grünefeld, Nils, et al.
Publicado: (2026)
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
por: Staliūnaitė, Ieva Raminta, et al.
Publicado: (2026)
por: Staliūnaitė, Ieva Raminta, et al.
Publicado: (2026)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
por: Obeso, Oscar, et al.
Publicado: (2025)
por: Obeso, Oscar, et al.
Publicado: (2025)
Ejemplares similares
-
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
por: Duan, Jinhao, et al.
Publicado: (2023) -
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
por: Wang, Zhiyuan, et al.
Publicado: (2024) -
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
por: Wang, Zhiyuan, et al.
Publicado: (2024) -
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
por: Duan, Jinhao, et al.
Publicado: (2025) -
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
por: Wang, Zhiyuan, et al.
Publicado: (2025)