Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhiyuan, Duan, Jinhao, Yuan, Chenxi, Chen, Qingyu, Chen, Tianlong, Zhang, Yue, Wang, Ren, Shi, Xiaoshuang, Xu, Kaidi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
SConU: Selective Conformal Uncertainty in Large Language Models
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
di: Duan, Jinhao, et al.
Pubblicazione: (2023)
di: Duan, Jinhao, et al.
Pubblicazione: (2023)
Conformal Lesion Segmentation for 3D Medical Images
di: Tan, Binyu, et al.
Pubblicazione: (2025)
di: Tan, Binyu, et al.
Pubblicazione: (2025)
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
Free Form Medical Visual Question Answering in Radiology
di: Narayanan, Abhishek, et al.
Pubblicazione: (2024)
di: Narayanan, Abhishek, et al.
Pubblicazione: (2024)
Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs
di: Zhang, Zhuoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Zhuoxuan, et al.
Pubblicazione: (2025)
A Benchmark for Long-Form Medical Question Answering
di: Hosseini, Pedram, et al.
Pubblicazione: (2024)
di: Hosseini, Pedram, et al.
Pubblicazione: (2024)
Uncertainty Estimation of Large Language Models in Medical Question Answering
di: Wu, Jiaxin, et al.
Pubblicazione: (2024)
di: Wu, Jiaxin, et al.
Pubblicazione: (2024)
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
di: Kong, Fei, et al.
Pubblicazione: (2025)
di: Kong, Fei, et al.
Pubblicazione: (2025)
Beyond Redundancy: Diverse and Specialized Multi-Expert Sparse Autoencoder
di: Xu, Zhen, et al.
Pubblicazione: (2025)
di: Xu, Zhen, et al.
Pubblicazione: (2025)
Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversations
di: Liu, Zijie, et al.
Pubblicazione: (2025)
di: Liu, Zijie, et al.
Pubblicazione: (2025)
Trustworthy Medical Question Answering: An Evaluation-Centric Survey
di: Wang, Yinuo, et al.
Pubblicazione: (2025)
di: Wang, Yinuo, et al.
Pubblicazione: (2025)
Testing Question Answering Software with Context-Driven Question Generation
di: Liu, Shuang, et al.
Pubblicazione: (2025)
di: Liu, Shuang, et al.
Pubblicazione: (2025)
Atomic Consistency Preference Optimization for Long-Form Question Answering
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
Word Form Matters: LLMs' Semantic Reconstruction under Typoglycemia
di: Wang, Chenxi, et al.
Pubblicazione: (2025)
di: Wang, Chenxi, et al.
Pubblicazione: (2025)
Mind the Ambiguity: Aleatoric Uncertainty Quantification in LLMs for Safe Medical Question Answering
di: Liu, Yaokun, et al.
Pubblicazione: (2026)
di: Liu, Yaokun, et al.
Pubblicazione: (2026)
Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
di: Wang, Yubo, et al.
Pubblicazione: (2023)
di: Wang, Yubo, et al.
Pubblicazione: (2023)
GuideLLM: Exploring LLM-Guided Conversation with Applications in Autobiography Interviewing
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
Understanding Retrieval Augmentation for Long-Form Question Answering
di: Chen, Hung-Ting, et al.
Pubblicazione: (2023)
di: Chen, Hung-Ting, et al.
Pubblicazione: (2023)
Benchmarking Uncertainty Calibration in Large Language Model Long-Form Question Answering
di: Müller, Philip, et al.
Pubblicazione: (2026)
di: Müller, Philip, et al.
Pubblicazione: (2026)
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
di: Chen, Xupeng, et al.
Pubblicazione: (2026)
di: Chen, Xupeng, et al.
Pubblicazione: (2026)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
mForms : Multimodal Form-Filling with Question Answering
di: Heck, Larry, et al.
Pubblicazione: (2020)
di: Heck, Larry, et al.
Pubblicazione: (2020)
Caterpillar: A Pure-MLP Architecture with Shifted-Pillars-Concatenation
di: Sun, Jin, et al.
Pubblicazione: (2023)
di: Sun, Jin, et al.
Pubblicazione: (2023)
IndustryEQA: Pushing the Frontiers of Embodied Question Answering in Industrial Scenarios
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
A Dual-Attention Learning Network with Word and Sentence Embedding for Medical Visual Question Answering
di: Huang, Xiaofei, et al.
Pubblicazione: (2022)
di: Huang, Xiaofei, et al.
Pubblicazione: (2022)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
ESQA: Event Sequences Question Answering
di: Abdullaeva, Irina, et al.
Pubblicazione: (2024)
di: Abdullaeva, Irina, et al.
Pubblicazione: (2024)
TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination Via Latent Truthful-Guided Pre-Intervention
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models
di: Kong, Fei, et al.
Pubblicazione: (2023)
di: Kong, Fei, et al.
Pubblicazione: (2023)
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
di: Yao, Yifan, et al.
Pubblicazione: (2023)
di: Yao, Yifan, et al.
Pubblicazione: (2023)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
di: Hu, Wenhao, et al.
Pubblicazione: (2025)
di: Hu, Wenhao, et al.
Pubblicazione: (2025)
Toward Ethical AI Through Bayesian Uncertainty in Neural Question Answering
di: Di Sipio, Riccardo
Pubblicazione: (2025)
di: Di Sipio, Riccardo
Pubblicazione: (2025)
Prompt-based Personalized Federated Learning for Medical Visual Question Answering
di: Zhu, He, et al.
Pubblicazione: (2024)
di: Zhu, He, et al.
Pubblicazione: (2024)
Towards Probabilistic Question Answering Over Tabular Data
di: Shen, Chen, et al.
Pubblicazione: (2025)
di: Shen, Chen, et al.
Pubblicazione: (2025)
Uncertainty-Guided Self-Questioning and Answering for Video-Language Alignment
di: Chen, Jin, et al.
Pubblicazione: (2024)
di: Chen, Jin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025) -
SConU: Selective Conformal Uncertainty in Large Language Models
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025) -
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026) -
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
di: Duan, Jinhao, et al.
Pubblicazione: (2023) -
Conformal Lesion Segmentation for 3D Medical Images
di: Tan, Binyu, et al.
Pubblicazione: (2025)