Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
Fuente:
arXiv
Guardado en:
| Autores principales: | Bouchard, Dylan, Chauhan, Mohit Singh, Bajaj, Viren, Skarbrevik, David |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
UQLM: A Python Package for Uncertainty Quantification in Large Language Models
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
por: Bouchard, Dylan, et al.
Publicado: (2026)
por: Bouchard, Dylan, et al.
Publicado: (2026)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
por: Fan, Haozhi, et al.
Publicado: (2026)
por: Fan, Haozhi, et al.
Publicado: (2026)
Is Escalation Worth It? A Decision-Theoretic Characterization of LLM Cascades
por: Bouchard, Dylan
Publicado: (2026)
por: Bouchard, Dylan
Publicado: (2026)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
por: Jiang, Mingjian, et al.
Publicado: (2024)
por: Jiang, Mingjian, et al.
Publicado: (2024)
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
por: Duan, Jinhao, et al.
Publicado: (2023)
por: Duan, Jinhao, et al.
Publicado: (2023)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
por: Nikitin, Alexander, et al.
Publicado: (2024)
por: Nikitin, Alexander, et al.
Publicado: (2024)
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
por: Xiao, Quan, et al.
Publicado: (2025)
por: Xiao, Quan, et al.
Publicado: (2025)
Multi-group Uncertainty Quantification for Long-form Text Generation
por: Liu, Terrance, et al.
Publicado: (2024)
por: Liu, Terrance, et al.
Publicado: (2024)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
por: Cao, Qi, et al.
Publicado: (2026)
por: Cao, Qi, et al.
Publicado: (2026)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
por: Grewal, Yashvir S., et al.
Publicado: (2024)
por: Grewal, Yashvir S., et al.
Publicado: (2024)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
por: Wang, Ziyu, et al.
Publicado: (2024)
por: Wang, Ziyu, et al.
Publicado: (2024)
Trustworthy Summarization via Uncertainty Quantification and Risk Awareness in Large Language Models
por: Pan, Shuaidong, et al.
Publicado: (2025)
por: Pan, Shuaidong, et al.
Publicado: (2025)
ESI: Epistemic Uncertainty Quantification via Semantic-preserving Intervention for Large Language Models
por: Li, Mingda, et al.
Publicado: (2025)
por: Li, Mingda, et al.
Publicado: (2025)
Fine-Grained Interpretation of Political Opinions in Large Language Models
por: Hu, Jingyu, et al.
Publicado: (2025)
por: Hu, Jingyu, et al.
Publicado: (2025)
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
por: Santilli, Andrea, et al.
Publicado: (2025)
por: Santilli, Andrea, et al.
Publicado: (2025)
Efficient Non-Parametric Uncertainty Quantification for Black-Box Large Language Models and Decision Planning
por: Tsai, Yao-Hung Hubert, et al.
Publicado: (2024)
por: Tsai, Yao-Hung Hubert, et al.
Publicado: (2024)
Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs
por: Sahoo, Subramanyam
Publicado: (2026)
por: Sahoo, Subramanyam
Publicado: (2026)
FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models
por: Weng, Zixuan, et al.
Publicado: (2026)
por: Weng, Zixuan, et al.
Publicado: (2026)
Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning
por: Krishnan, Ranganath, et al.
Publicado: (2024)
por: Krishnan, Ranganath, et al.
Publicado: (2024)
LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models
por: Chen, Yukang, et al.
Publicado: (2023)
por: Chen, Yukang, et al.
Publicado: (2023)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
por: Kirchhof, Michael, et al.
Publicado: (2025)
por: Kirchhof, Michael, et al.
Publicado: (2025)
Perceptions of Linguistic Uncertainty by Language Models and Humans
por: Belem, Catarina G, et al.
Publicado: (2024)
por: Belem, Catarina G, et al.
Publicado: (2024)
Uncertainty Quantification for Transformer Models for Dark-Pattern Detection
por: Muñoz, Javier, et al.
Publicado: (2024)
por: Muñoz, Javier, et al.
Publicado: (2024)
SLOT: Structuring the Output of Large Language Models
por: Wang, Darren Yow-Bang, et al.
Publicado: (2025)
por: Wang, Darren Yow-Bang, et al.
Publicado: (2025)
FLAMES: Improving LLM Math Reasoning via a Fine-Grained Analysis of the Data Synthesis Pipeline
por: Seegmiller, Parker, et al.
Publicado: (2025)
por: Seegmiller, Parker, et al.
Publicado: (2025)
An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms
por: Grünefeld, Nils, et al.
Publicado: (2026)
por: Grünefeld, Nils, et al.
Publicado: (2026)
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
por: Staliūnaitė, Ieva Raminta, et al.
Publicado: (2026)
por: Staliūnaitė, Ieva Raminta, et al.
Publicado: (2026)
Adaptive Uncertainty Quantification for Generative AI
por: Kim, Jungeum, et al.
Publicado: (2024)
por: Kim, Jungeum, et al.
Publicado: (2024)
An Evaluation on Large Language Model Outputs: Discourse and Memorization
por: de Wynter, Adrian, et al.
Publicado: (2023)
por: de Wynter, Adrian, et al.
Publicado: (2023)
Leviathan: Decoupling Input and Output Representations in Language Models
por: Batley, Reza T., et al.
Publicado: (2026)
por: Batley, Reza T., et al.
Publicado: (2026)
Skewed Memorization in Large Language Models: Quantification and Decomposition
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models
por: Islam, Mohammed Saidul, et al.
Publicado: (2026)
por: Islam, Mohammed Saidul, et al.
Publicado: (2026)
Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context
por: Alizadeh, Keivan, et al.
Publicado: (2026)
por: Alizadeh, Keivan, et al.
Publicado: (2026)
A Study of Backdoors in Instruction Fine-tuned Language Models
por: Raghuram, Jayaram, et al.
Publicado: (2024)
por: Raghuram, Jayaram, et al.
Publicado: (2024)
Linguistic Calibration of Long-Form Generations
por: Band, Neil, et al.
Publicado: (2024)
por: Band, Neil, et al.
Publicado: (2024)
Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference
por: Zhao, Stephen, et al.
Publicado: (2025)
por: Zhao, Stephen, et al.
Publicado: (2025)
Ejemplares similares
-
UQLM: A Python Package for Uncertainty Quantification in Large Language Models
por: Bouchard, Dylan, et al.
Publicado: (2025) -
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
por: Bouchard, Dylan, et al.
Publicado: (2025) -
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
por: Bouchard, Dylan, et al.
Publicado: (2025) -
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
por: Bouchard, Dylan, et al.
Publicado: (2026) -
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
por: Fan, Haozhi, et al.
Publicado: (2026)