Salvato in:
| Autori principali: | Zhang, Zhuoxuan, Duan, Jinhao, Kim, Edward, Xu, Kaidi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2509.13664 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
di: Hu, Wenhao, et al.
Pubblicazione: (2025)
di: Hu, Wenhao, et al.
Pubblicazione: (2025)
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
GuideLLM: Exploring LLM-Guided Conversation with Applications in Autobiography Interviewing
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
di: Zhang, Ziqian, et al.
Pubblicazione: (2026)
di: Zhang, Ziqian, et al.
Pubblicazione: (2026)
Mind the Ambiguity: Aleatoric Uncertainty Quantification in LLMs for Safe Medical Question Answering
di: Liu, Yaokun, et al.
Pubblicazione: (2026)
di: Liu, Yaokun, et al.
Pubblicazione: (2026)
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
di: Duan, Jinhao, et al.
Pubblicazione: (2023)
di: Duan, Jinhao, et al.
Pubblicazione: (2023)
Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
di: Hong, Junyuan, et al.
Pubblicazione: (2024)
di: Hong, Junyuan, et al.
Pubblicazione: (2024)
Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
di: Keluskar, Aryan, et al.
Pubblicazione: (2024)
di: Keluskar, Aryan, et al.
Pubblicazione: (2024)
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
LLMs can Find Mathematical Reasoning Mistakes by Pedagogical Chain-of-Thought
di: Jiang, Zhuoxuan, et al.
Pubblicazione: (2024)
di: Jiang, Zhuoxuan, et al.
Pubblicazione: (2024)
A$^2$Search: Ambiguity-Aware Question Answering with Reinforcement Learning
di: Zhang, Fengji, et al.
Pubblicazione: (2025)
di: Zhang, Fengji, et al.
Pubblicazione: (2025)
TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination Via Latent Truthful-Guided Pre-Intervention
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
di: Duan, Jinhao, et al.
Pubblicazione: (2025)
Beyond Surface Statistics: Robust Conformal Prediction for LLMs via Internal Representations
di: Wang, Yanli, et al.
Pubblicazione: (2026)
di: Wang, Yanli, et al.
Pubblicazione: (2026)
Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity
di: Wu, Xinwei, et al.
Pubblicazione: (2025)
di: Wu, Xinwei, et al.
Pubblicazione: (2025)
QPaug: Question and Passage Augmentation for Open-Domain Question Answering of LLMs
di: Kim, Minsang, et al.
Pubblicazione: (2024)
di: Kim, Minsang, et al.
Pubblicazione: (2024)
Resolving Intent Ambiguities by Retrieving Discriminative Clarifying Questions
di: Dhole, Kaustubh D.
Pubblicazione: (2020)
di: Dhole, Kaustubh D.
Pubblicazione: (2020)
Efficient Multi-Hop Question Answering over Knowledge Graphs via LLM Planning and Embedding-Guided Search
di: Shrestha, Manil, et al.
Pubblicazione: (2025)
di: Shrestha, Manil, et al.
Pubblicazione: (2025)
Dissecting Role Cognition in Medical LLMs via Neuronal Ablation
di: Liang, Xun, et al.
Pubblicazione: (2025)
di: Liang, Xun, et al.
Pubblicazione: (2025)
H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs
di: Gao, Cheng, et al.
Pubblicazione: (2025)
di: Gao, Cheng, et al.
Pubblicazione: (2025)
Interaction Dynamics as a Reward Signal for LLMs
di: Gooding, Sian, et al.
Pubblicazione: (2025)
di: Gooding, Sian, et al.
Pubblicazione: (2025)
Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs
di: Shore, Amber, et al.
Pubblicazione: (2025)
di: Shore, Amber, et al.
Pubblicazione: (2025)
From Answers to Questions: EQGBench for Evaluating LLMs' Educational Question Generation
di: Zhou, Chengliang, et al.
Pubblicazione: (2025)
di: Zhou, Chengliang, et al.
Pubblicazione: (2025)
NeuronTune: Fine-Grained Neuron Modulation for Balanced Safety-Utility Alignment in LLMs
di: Pan, Birong, et al.
Pubblicazione: (2025)
di: Pan, Birong, et al.
Pubblicazione: (2025)
Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversations
di: Liu, Zijie, et al.
Pubblicazione: (2025)
di: Liu, Zijie, et al.
Pubblicazione: (2025)
Can LLMs Ask Good Questions?
di: Zhang, Yueheng, et al.
Pubblicazione: (2025)
di: Zhang, Yueheng, et al.
Pubblicazione: (2025)
Towards Unbiased Evaluation of Detecting Unanswerable Questions in EHRSQL
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
ReasoningLM: Enabling Structural Subgraph Reasoning in Pre-trained Language Models for Question Answering over Knowledge Graph
di: Jiang, Jinhao, et al.
Pubblicazione: (2023)
di: Jiang, Jinhao, et al.
Pubblicazione: (2023)
HanjaBridge: Resolving Semantic Ambiguity in Korean LLMs via Hanja-Augmented Pre-Training
di: Choi, Seungho
Pubblicazione: (2025)
di: Choi, Seungho
Pubblicazione: (2025)
UD-English-CHILDES: A Collected Resource of Gold and Silver Universal Dependencies Trees for Child Language Interactions
di: Yang, Xiulin, et al.
Pubblicazione: (2025)
di: Yang, Xiulin, et al.
Pubblicazione: (2025)
Enhancing Multiple Dimensions of Trustworthiness in LLMs via Sparse Activation Control
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
Identifying Good and Bad Neurons for Task-Level Controllable LLMs
di: Li, Wenjie, et al.
Pubblicazione: (2026)
di: Li, Wenjie, et al.
Pubblicazione: (2026)
Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
Language Model Circuits Are Sparse in the Neuron Basis
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
Are We Asking the Right Questions? On Ambiguity in Natural Language Queries for Tabular Data Analysis
di: Gomm, Daniel, et al.
Pubblicazione: (2025)
di: Gomm, Daniel, et al.
Pubblicazione: (2025)
CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
Knowledge Tagging System on Math Questions via LLMs with Flexible Demonstration Retriever
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026) -
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
di: Duan, Jinhao, et al.
Pubblicazione: (2025) -
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025) -
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
di: Hu, Wenhao, et al.
Pubblicazione: (2025) -
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)