Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
Fuente:
arXiv
Guardado en:
| Autores principales: | Cheang, Chi Seng, Chan, Hou Pong, Zhang, Wenxuan, Deng, Yang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
por: Zhao, Raoyuan, et al.
Publicado: (2025)
por: Zhao, Raoyuan, et al.
Publicado: (2025)
Can AI Assistants Know What They Don't Know?
por: Cheng, Qinyuan, et al.
Publicado: (2024)
por: Cheng, Qinyuan, et al.
Publicado: (2024)
Do Retrieval Augmented Language Models Know When They Don't Know?
por: Zhou, Youchao, et al.
Publicado: (2025)
por: Zhou, Youchao, et al.
Publicado: (2025)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
por: Yeom, Jewon, et al.
Publicado: (2026)
por: Yeom, Jewon, et al.
Publicado: (2026)
Large Language Models Must Be Taught to Know What They Don't Know
por: Kapoor, Sanyam, et al.
Publicado: (2024)
por: Kapoor, Sanyam, et al.
Publicado: (2024)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
por: Mei, Zhiting, et al.
Publicado: (2025)
por: Mei, Zhiting, et al.
Publicado: (2025)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
por: Lee, Joosung, et al.
Publicado: (2026)
por: Lee, Joosung, et al.
Publicado: (2026)
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
por: Shi, Zeyu, et al.
Publicado: (2025)
por: Shi, Zeyu, et al.
Publicado: (2025)
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
por: Goloburda, Maiya, et al.
Publicado: (2026)
por: Goloburda, Maiya, et al.
Publicado: (2026)
Bayesian Mixture-of-Experts: Towards Making LLMs Know What They Don't Know
por: Li, Albus Yizhuo
Publicado: (2025)
por: Li, Albus Yizhuo
Publicado: (2025)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
por: Rezaeimanesh, Sara, et al.
Publicado: (2026)
por: Rezaeimanesh, Sara, et al.
Publicado: (2026)
I Don't Know: Explicit Modeling of Uncertainty with an [IDK] Token
por: Cohen, Roi, et al.
Publicado: (2024)
por: Cohen, Roi, et al.
Publicado: (2024)
Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say "I Don't Know"
por: Madhwal, Dhruv, et al.
Publicado: (2026)
por: Madhwal, Dhruv, et al.
Publicado: (2026)
Do LVLMs Know What They Know? A Systematic Study of Knowledge Boundary Perception in LVLMs
por: Ding, Zhikai, et al.
Publicado: (2025)
por: Ding, Zhikai, et al.
Publicado: (2025)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
por: Mayne, Harry, et al.
Publicado: (2025)
por: Mayne, Harry, et al.
Publicado: (2025)
Imagining What We Don't Know
por: Samuels, Lisa
Publicado: (2026)
por: Samuels, Lisa
Publicado: (2026)
Imagining What We Don't Know
por: Samuels, Lisa
Publicado: (2026)
por: Samuels, Lisa
Publicado: (2026)
Can LLMs Ground when they (Don't) Know: A Study on Direct and Loaded Political Questions
por: Lachenmaier, Clara, et al.
Publicado: (2025)
por: Lachenmaier, Clara, et al.
Publicado: (2025)
"Knowing When You Don't Know": A Multilingual Relevance Assessment Dataset for Robust Retrieval-Augmented Generation
por: Thakur, Nandan, et al.
Publicado: (2023)
por: Thakur, Nandan, et al.
Publicado: (2023)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
por: Zhang, Hanning, et al.
Publicado: (2023)
por: Zhang, Hanning, et al.
Publicado: (2023)
Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users
por: Balepur, Nishant, et al.
Publicado: (2026)
por: Balepur, Nishant, et al.
Publicado: (2026)
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing
por: Yang, Diji, et al.
Publicado: (2025)
por: Yang, Diji, et al.
Publicado: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
por: Cha, Sungguk, et al.
Publicado: (2024)
por: Cha, Sungguk, et al.
Publicado: (2024)
Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory
por: Cacioli, Jon-Paul
Publicado: (2026)
por: Cacioli, Jon-Paul
Publicado: (2026)
Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations
por: Wang, Haoran, et al.
Publicado: (2026)
por: Wang, Haoran, et al.
Publicado: (2026)
What We Know and What We Don't Know About the Function of γδ T Cells
por: Immo Prinz, et al.
Publicado: (2025)
por: Immo Prinz, et al.
Publicado: (2025)
Experts Don't Cheat: Learning What You Don't Know By Predicting Pairs
por: Johnson, Daniel D., et al.
Publicado: (2024)
por: Johnson, Daniel D., et al.
Publicado: (2024)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
por: Park, Young-Jin, et al.
Publicado: (2025)
por: Park, Young-Jin, et al.
Publicado: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
por: Bajpai, Divya Jyoti, et al.
Publicado: (2025)
por: Bajpai, Divya Jyoti, et al.
Publicado: (2025)
LLMs Learn Constructions That Humans Do Not Know
por: Dunn, Jonathan, et al.
Publicado: (2025)
por: Dunn, Jonathan, et al.
Publicado: (2025)
Do LLMs Know to Respect Copyright Notice?
por: Xu, Jialiang, et al.
Publicado: (2024)
por: Xu, Jialiang, et al.
Publicado: (2024)
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty
por: Ren, Jingyi, et al.
Publicado: (2026)
por: Ren, Jingyi, et al.
Publicado: (2026)
KnowRL: Teaching Language Models to Know What They Know
por: Kale, Sahil, et al.
Publicado: (2025)
por: Kale, Sahil, et al.
Publicado: (2025)
Truth Knows No Language: Evaluating Truthfulness Beyond English
por: Figueras, Blanca Calvo, et al.
Publicado: (2025)
por: Figueras, Blanca Calvo, et al.
Publicado: (2025)
Analyzing LLMs' Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations
por: Xiao, Chenghao, et al.
Publicado: (2025)
por: Xiao, Chenghao, et al.
Publicado: (2025)
Do LLMs Know about Hallucination? An Empirical Investigation of LLM's Hidden States
por: Duan, Hanyu, et al.
Publicado: (2024)
por: Duan, Hanyu, et al.
Publicado: (2024)
Knowing But Not Doing: Convergent Morality and Divergent Action in LLMs
por: Huang, Jen-tse, et al.
Publicado: (2026)
por: Huang, Jen-tse, et al.
Publicado: (2026)
Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method
por: Zhao, Yukun, et al.
Publicado: (2023)
por: Zhao, Yukun, et al.
Publicado: (2023)
Do Large Language Models Know What They Are Capable Of?
por: Barkan, Casey O., et al.
Publicado: (2025)
por: Barkan, Casey O., et al.
Publicado: (2025)
Ejemplares similares
-
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
por: Zhao, Raoyuan, et al.
Publicado: (2025) -
Can AI Assistants Know What They Don't Know?
por: Cheng, Qinyuan, et al.
Publicado: (2024) -
Do Retrieval Augmented Language Models Know When They Don't Know?
por: Zhou, Youchao, et al.
Publicado: (2025) -
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
por: Yeom, Jewon, et al.
Publicado: (2026) -
Large Language Models Must Be Taught to Know What They Don't Know
por: Kapoor, Sanyam, et al.
Publicado: (2024)