Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Xiaomin, Yu, Zhou, Zhang, Ziji, Zhuang, Yingying, Shah, Swair, Sadagopan, Narayanan, Beniwal, Anurag |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DARD: A Multi-Agent Approach for Task-Oriented Dialog Systems
di: Gupta, Aman, et al.
Pubblicazione: (2024)
di: Gupta, Aman, et al.
Pubblicazione: (2024)
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
TOD-ProcBench: Benchmarking Complex Instruction-Following in Task-Oriented Dialogues
di: Ghazarian, Sarik, et al.
Pubblicazione: (2025)
di: Ghazarian, Sarik, et al.
Pubblicazione: (2025)
How and Where to Translate? The Impact of Translation Strategies in Cross-lingual LLM Prompting
di: Gupta, Aman, et al.
Pubblicazione: (2025)
di: Gupta, Aman, et al.
Pubblicazione: (2025)
Multilingual Information Retrieval with a Monolingual Knowledge Base
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
Analysis of Indic Language Capabilities in LLMs
di: Vaidya, Aatman, et al.
Pubblicazione: (2025)
di: Vaidya, Aatman, et al.
Pubblicazione: (2025)
REIC: RAG-Enhanced Intent Classification at Scale
di: Zhang, Ziji, et al.
Pubblicazione: (2025)
di: Zhang, Ziji, et al.
Pubblicazione: (2025)
Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training
di: Ye, Chenlu, et al.
Pubblicazione: (2025)
di: Ye, Chenlu, et al.
Pubblicazione: (2025)
AXCEL: Automated eXplainable Consistency Evaluation using LLMs
di: Sreekar, P Aditya, et al.
Pubblicazione: (2024)
di: Sreekar, P Aditya, et al.
Pubblicazione: (2024)
Transparentize the Internal and External Knowledge Utilization in LLMs with Trustworthy Citation
di: Shen, Jiajun, et al.
Pubblicazione: (2025)
di: Shen, Jiajun, et al.
Pubblicazione: (2025)
Internal and External Impacts of Natural Language Processing Papers
di: Zhang, Yu
Pubblicazione: (2025)
di: Zhang, Yu
Pubblicazione: (2025)
Are LLMs Models of Distributional Semantics? A Case Study on Quantifiers
di: Enyan, Zhang, et al.
Pubblicazione: (2024)
di: Enyan, Zhang, et al.
Pubblicazione: (2024)
From Retrieval to Generation: Unifying External and Parametric Knowledge for Medical Question Answering
di: Li, Lei, et al.
Pubblicazione: (2025)
di: Li, Lei, et al.
Pubblicazione: (2025)
Enhancing Uncertainty Estimation in LLMs with Expectation of Aggregated Internal Belief
di: Xiao, Zeguan, et al.
Pubblicazione: (2025)
di: Xiao, Zeguan, et al.
Pubblicazione: (2025)
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
di: Yadav, Ankit, et al.
Pubblicazione: (2024)
di: Yadav, Ankit, et al.
Pubblicazione: (2024)
Is External Information Useful for Stance Detection with LLMs?
di: Nguyen, Quang Minh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang Minh, et al.
Pubblicazione: (2025)
Quantifying and Mitigating Premature Closure in Frontier LLMs
di: Handler, Rebecca, et al.
Pubblicazione: (2026)
di: Handler, Rebecca, et al.
Pubblicazione: (2026)
Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance
di: Yu, Jiachen, et al.
Pubblicazione: (2026)
di: Yu, Jiachen, et al.
Pubblicazione: (2026)
Bridging External and Parametric Knowledge: Mitigating Hallucination of LLMs with Shared-Private Semantic Synergy in Dual-Stream Knowledge
di: Sui, Yi, et al.
Pubblicazione: (2025)
di: Sui, Yi, et al.
Pubblicazione: (2025)
Can Language Model Understand Word Semantics as A Chatbot? An Empirical Study of Language Model Internal External Mismatch
di: Zhao, Jinman, et al.
Pubblicazione: (2024)
di: Zhao, Jinman, et al.
Pubblicazione: (2024)
Hallucination Detection with the Internal Layers of LLMs
di: Preiß, Martin
Pubblicazione: (2025)
di: Preiß, Martin
Pubblicazione: (2025)
Enhancing Uncertainty Modeling with Semantic Graph for Hallucination Detection
di: Chen, Kedi, et al.
Pubblicazione: (2025)
di: Chen, Kedi, et al.
Pubblicazione: (2025)
Do Internal Layers of LLMs Reveal Patterns for Jailbreak Detection?
di: Kadali, Sri Durga Sai Sowmya, et al.
Pubblicazione: (2025)
di: Kadali, Sri Durga Sai Sowmya, et al.
Pubblicazione: (2025)
INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection
di: Chen, Chao, et al.
Pubblicazione: (2024)
di: Chen, Chao, et al.
Pubblicazione: (2024)
Understanding Aha Moments: from External Observations to Internal Mechanisms
di: Yang, Shu, et al.
Pubblicazione: (2025)
di: Yang, Shu, et al.
Pubblicazione: (2025)
Using External knowledge to Enhanced PLM for Semantic Matching
di: Li, Min, et al.
Pubblicazione: (2025)
di: Li, Min, et al.
Pubblicazione: (2025)
Char-mander Use mBackdoor! A Study of Cross-lingual Backdoor Attacks in Multilingual LLMs
di: Beniwal, Himanshu, et al.
Pubblicazione: (2025)
di: Beniwal, Himanshu, et al.
Pubblicazione: (2025)
Detecting Hallucinations in Retrieval-Augmented Generation via Semantic-level Internal Reasoning Graph
di: Hu, Jianpeng, et al.
Pubblicazione: (2026)
di: Hu, Jianpeng, et al.
Pubblicazione: (2026)
Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis
di: Li, Li, et al.
Pubblicazione: (2025)
di: Li, Li, et al.
Pubblicazione: (2025)
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
di: Phillips, Edward, et al.
Pubblicazione: (2025)
di: Phillips, Edward, et al.
Pubblicazione: (2025)
TempPerturb-Eval: On the Joint Effects of Internal Temperature and External Perturbations in RAG Robustness
di: Zhou, Yongxin, et al.
Pubblicazione: (2025)
di: Zhou, Yongxin, et al.
Pubblicazione: (2025)
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs
di: Ao, Shuang, et al.
Pubblicazione: (2024)
di: Ao, Shuang, et al.
Pubblicazione: (2024)
SGR: A Stepwise Reasoning Framework for LLMs with External Subgraph Generation
di: Zhang, Xin, et al.
Pubblicazione: (2026)
di: Zhang, Xin, et al.
Pubblicazione: (2026)
Prompt-Guided Internal States for Hallucination Detection of Large Language Models
di: Zhang, Fujie, et al.
Pubblicazione: (2024)
di: Zhang, Fujie, et al.
Pubblicazione: (2024)
TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks
di: Kapoor, Vansh, et al.
Pubblicazione: (2026)
di: Kapoor, Vansh, et al.
Pubblicazione: (2026)
Quantifier Scope Interpretation in Language Learners and LLMs
di: Fang, Shaohua, et al.
Pubblicazione: (2025)
di: Fang, Shaohua, et al.
Pubblicazione: (2025)
CARES: Comprehensive Evaluation of Safety and Adversarial Robustness in Medical LLMs
di: Chen, Sijia, et al.
Pubblicazione: (2025)
di: Chen, Sijia, et al.
Pubblicazione: (2025)
Internal and External Knowledge Interactive Refinement Framework for Knowledge-Intensive Question Answering
di: Du, Haowei, et al.
Pubblicazione: (2024)
di: Du, Haowei, et al.
Pubblicazione: (2024)
Select to Know: An Internal-External Knowledge Self-Selection Framework for Domain-Specific Question Answering
di: He, Bolei, et al.
Pubblicazione: (2025)
di: He, Bolei, et al.
Pubblicazione: (2025)
DEPART: DEcomposing PARiTy across Multilingual LLMs
di: Uppadhyay, Manan, et al.
Pubblicazione: (2026)
di: Uppadhyay, Manan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
DARD: A Multi-Agent Approach for Task-Oriented Dialog Systems
di: Gupta, Aman, et al.
Pubblicazione: (2024) -
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
di: Li, Xiaomin, et al.
Pubblicazione: (2025) -
TOD-ProcBench: Benchmarking Complex Instruction-Following in Task-Oriented Dialogues
di: Ghazarian, Sarik, et al.
Pubblicazione: (2025) -
How and Where to Translate? The Impact of Translation Strategies in Cross-lingual LLM Prompting
di: Gupta, Aman, et al.
Pubblicazione: (2025) -
Multilingual Information Retrieval with a Monolingual Knowledge Base
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)