Do LLMs Know about Hallucination? An Empirical Investigation of LLM's Hidden States
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Duan, Hanyu, Yang, Yi, Tam, Kar Yan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Layer-wise Representation Dynamics: An Empirical Investigation Across Embedders and Base LLMs
von: Jiang, Jingzhou, et al.
Veröffentlicht: (2026)
von: Jiang, Jingzhou, et al.
Veröffentlicht: (2026)
LLM-Measure: Generating Valid, Consistent, and Reproducible Text-Based Measures for Social Science Research
von: Yang, Yi, et al.
Veröffentlicht: (2024)
von: Yang, Yi, et al.
Veröffentlicht: (2024)
Bias A-head? Analyzing Bias in Transformer-Based Language Model Attention Heads
von: Yang, Yi, et al.
Veröffentlicht: (2023)
von: Yang, Yi, et al.
Veröffentlicht: (2023)
Revealing the Numeracy Gap: An Empirical Investigation of Text Embedding Models
von: Deng, Ningyuan, et al.
Veröffentlicht: (2025)
von: Deng, Ningyuan, et al.
Veröffentlicht: (2025)
Evaluating and Aligning Human Economic Risk Preferences in LLMs
von: Liu, Jiaxin, et al.
Veröffentlicht: (2025)
von: Liu, Jiaxin, et al.
Veröffentlicht: (2025)
Beyond Surface Similarity: Detecting Subtle Semantic Shifts in Financial Narratives
von: Liu, Jiaxin, et al.
Veröffentlicht: (2024)
von: Liu, Jiaxin, et al.
Veröffentlicht: (2024)
Does Less Hallucination Mean Less Creativity? An Empirical Investigation in LLMs
von: Banerjee, Mohor, et al.
Veröffentlicht: (2025)
von: Banerjee, Mohor, et al.
Veröffentlicht: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
von: Cheang, Chi Seng, et al.
Veröffentlicht: (2025)
von: Cheang, Chi Seng, et al.
Veröffentlicht: (2025)
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
von: Madhusudhan, Nishanth, et al.
Veröffentlicht: (2024)
von: Madhusudhan, Nishanth, et al.
Veröffentlicht: (2024)
Do We Need Domain-Specific Embedding Models? An Empirical Investigation
von: Tang, Yixuan, et al.
Veröffentlicht: (2024)
von: Tang, Yixuan, et al.
Veröffentlicht: (2024)
ICR Probe: Tracking Hidden State Dynamics for Reliable Hallucination Detection in LLMs
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
Hallucination, Monofacts, and Miscalibration: An Empirical Investigation
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2025)
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2025)
Do LLMs Know to Respect Copyright Notice?
von: Xu, Jialiang, et al.
Veröffentlicht: (2024)
von: Xu, Jialiang, et al.
Veröffentlicht: (2024)
LLMs Learn Constructions That Humans Do Not Know
von: Dunn, Jonathan, et al.
Veröffentlicht: (2025)
von: Dunn, Jonathan, et al.
Veröffentlicht: (2025)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
FLARE: Task-agnostic embedding model evaluation through a normalization process
von: Jiang, Jingzhou, et al.
Veröffentlicht: (2026)
von: Jiang, Jingzhou, et al.
Veröffentlicht: (2026)
Do Language Models Know When They're Hallucinating References?
von: Agrawal, Ayush, et al.
Veröffentlicht: (2023)
von: Agrawal, Ayush, et al.
Veröffentlicht: (2023)
Knowing But Not Doing: Convergent Morality and Divergent Action in LLMs
von: Huang, Jen-tse, et al.
Veröffentlicht: (2026)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2026)
States Hidden in Hidden States: LLMs Emerge Discrete State Representations Implicitly
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2025)
Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting
von: Tan, Chenchen, et al.
Veröffentlicht: (2025)
von: Tan, Chenchen, et al.
Veröffentlicht: (2025)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
Investigating and Addressing Hallucinations of LLMs in Tasks Involving Negation
von: Varshney, Neeraj, et al.
Veröffentlicht: (2024)
von: Varshney, Neeraj, et al.
Veröffentlicht: (2024)
INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection
von: Chen, Chao, et al.
Veröffentlicht: (2024)
von: Chen, Chao, et al.
Veröffentlicht: (2024)
The LLM Already Knows: Estimating LLM-Perceived Question Difficulty via Hidden Representations
von: Zhu, Yubo, et al.
Veröffentlicht: (2025)
von: Zhu, Yubo, et al.
Veröffentlicht: (2025)
How Much Do LLMs Know About Chinese Zero Pronouns?
von: Li, Yifei, et al.
Veröffentlicht: (2026)
von: Li, Yifei, et al.
Veröffentlicht: (2026)
Know the Unknown: An Uncertainty-Sensitive Method for LLM Instruction Tuning
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
von: Tian, Yuchen, et al.
Veröffentlicht: (2024)
von: Tian, Yuchen, et al.
Veröffentlicht: (2024)
Do Large Language Models Know Conflict? Investigating Parametric vs. Non-Parametric Knowledge of LLMs for Conflict Forecasting
von: Nemkova, Apollinaire Poli, et al.
Veröffentlicht: (2025)
von: Nemkova, Apollinaire Poli, et al.
Veröffentlicht: (2025)
Do Language Models Know Theo Has a Wife? Investigating the Proviso Problem
von: Azin, Tara, et al.
Veröffentlicht: (2026)
von: Azin, Tara, et al.
Veröffentlicht: (2026)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
Forget What You Know about LLMs Evaluations -- LLMs are Like a Chameleon
von: Cohen-Inger, Nurit, et al.
Veröffentlicht: (2025)
von: Cohen-Inger, Nurit, et al.
Veröffentlicht: (2025)
Can LLMs Perceive Time? An Empirical Investigation
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
LLM Hallucination Detection: A Fast Fourier Transform Method Based on Hidden Layer Temporal Signals
von: Li, Jinxin, et al.
Veröffentlicht: (2025)
von: Li, Jinxin, et al.
Veröffentlicht: (2025)
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
von: Tanwar, Eshaan, et al.
Veröffentlicht: (2025)
von: Tanwar, Eshaan, et al.
Veröffentlicht: (2025)
Do Retrieval Augmented Language Models Know When They Don't Know?
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs
von: Bombieri, Marco, et al.
Veröffentlicht: (2026)
von: Bombieri, Marco, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Layer-wise Representation Dynamics: An Empirical Investigation Across Embedders and Base LLMs
von: Jiang, Jingzhou, et al.
Veröffentlicht: (2026) -
LLM-Measure: Generating Valid, Consistent, and Reproducible Text-Based Measures for Social Science Research
von: Yang, Yi, et al.
Veröffentlicht: (2024) -
Bias A-head? Analyzing Bias in Transformer-Based Language Model Attention Heads
von: Yang, Yi, et al.
Veröffentlicht: (2023) -
Revealing the Numeracy Gap: An Empirical Investigation of Text Embedding Models
von: Deng, Ningyuan, et al.
Veröffentlicht: (2025) -
Evaluating and Aligning Human Economic Risk Preferences in LLMs
von: Liu, Jiaxin, et al.
Veröffentlicht: (2025)