LLM Internal States Reveal Hallucination Risk Faced With a Query
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ji, Ziwei, Chen, Delong, Ishii, Etsuko, Cahyawijaya, Samuel, Bang, Yejin, Wilie, Bryan, Fung, Pascale |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
High-Dimension Human Value Representation in Large Language Models
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
Belief Revision: The Adaptability of Large Language Models Reasoning
par: Wilie, Bryan, et autres
Publié: (2024)
par: Wilie, Bryan, et autres
Publié: (2024)
High-Dimensional Interlingual Representations of Large Language Models
par: Wilie, Bryan, et autres
Publié: (2025)
par: Wilie, Bryan, et autres
Publié: (2025)
What Makes for Good Image Captions?
par: Chen, Delong, et autres
Publié: (2024)
par: Chen, Delong, et autres
Publié: (2024)
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
par: Lovenia, Holy, et autres
Publié: (2023)
par: Lovenia, Holy, et autres
Publié: (2023)
Survey of Hallucination in Natural Language Generation
par: Ji, Ziwei, et autres
Publié: (2022)
par: Ji, Ziwei, et autres
Publié: (2022)
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
par: Bang, Yejin, et autres
Publié: (2024)
par: Bang, Yejin, et autres
Publié: (2024)
HalluLens: LLM Hallucination Benchmark
par: Bang, Yejin, et autres
Publié: (2025)
par: Bang, Yejin, et autres
Publié: (2025)
Subobject-level Image Tokenization
par: Chen, Delong, et autres
Publié: (2024)
par: Chen, Delong, et autres
Publié: (2024)
LLMs Are Few-Shot In-Context Low-Resource Language Learners
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
WorldPrediction: A Benchmark for High-level World Modeling and Long-horizon Procedural Planning
par: Chen, Delong, et autres
Publié: (2025)
par: Chen, Delong, et autres
Publié: (2025)
Calibrating Verbal Uncertainty as a Linear Feature to Reduce Hallucinations
par: Ji, Ziwei, et autres
Publié: (2025)
par: Ji, Ziwei, et autres
Publié: (2025)
LLM for Everyone: Representing the Underrepresented in Large Language Models
par: Cahyawijaya, Samuel
Publié: (2024)
par: Cahyawijaya, Samuel
Publié: (2024)
Linguistic Minimal Pairs Elicit Linguistic Similarity in Large Language Models
par: Zhou, Xinyu, et autres
Publié: (2024)
par: Zhou, Xinyu, et autres
Publié: (2024)
Reward Prediction with Factorized World States
par: Shen, Yijun, et autres
Publié: (2026)
par: Shen, Yijun, et autres
Publié: (2026)
Planning with Reasoning using Vision Language World Model
par: Chen, Delong, et autres
Publié: (2025)
par: Chen, Delong, et autres
Publié: (2025)
Cendol: Open Instruction-tuned Generative Large Language Models for Indonesian Languages
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
par: Cahyawijaya, Samuel, et autres
Publié: (2024)
HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering
par: Tong, Chaodong, et autres
Publié: (2025)
par: Tong, Chaodong, et autres
Publié: (2025)
INSIDE: LLMs' Internal States Retain the Power of Hallucination Detection
par: Chen, Chao, et autres
Publié: (2024)
par: Chen, Chao, et autres
Publié: (2024)
HALoGEN: Fantastic LLM Hallucinations and Where to Find Them
par: Ravichander, Abhilasha, et autres
Publié: (2025)
par: Ravichander, Abhilasha, et autres
Publié: (2025)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
par: Mao, Nathan, et autres
Publié: (2026)
par: Mao, Nathan, et autres
Publié: (2026)
Prompt-Guided Internal States for Hallucination Detection of Large Language Models
par: Zhang, Fujie, et autres
Publié: (2024)
par: Zhang, Fujie, et autres
Publié: (2024)
The HalluRAG Dataset: Detecting Closed-Domain Hallucinations in RAG Applications Using an LLM's Internal States
par: Ridder, Fabian, et autres
Publié: (2024)
par: Ridder, Fabian, et autres
Publié: (2024)
The Law of Knowledge Overshadowing: Towards Understanding, Predicting, and Preventing LLM Hallucination
par: Zhang, Yuji, et autres
Publié: (2025)
par: Zhang, Yuji, et autres
Publié: (2025)
What Causes Knowledge Loss in Multilingual Language Models?
par: Khelli, Maria, et autres
Publié: (2025)
par: Khelli, Maria, et autres
Publié: (2025)
Probing LLM Hallucination from Within: Perturbation-Driven Approach via Internal Knowledge
par: Lee, Seongmin, et autres
Publié: (2024)
par: Lee, Seongmin, et autres
Publié: (2024)
Hallucination Detection via Internal States and Structured Reasoning Consistency in Large Language Models
par: Song, Yusheng, et autres
Publié: (2025)
par: Song, Yusheng, et autres
Publié: (2025)
WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries
par: Zhao, Wenting, et autres
Publié: (2024)
par: Zhao, Wenting, et autres
Publié: (2024)
ANAH: Analytical Annotation of Hallucinations in Large Language Models
par: Ji, Ziwei, et autres
Publié: (2024)
par: Ji, Ziwei, et autres
Publié: (2024)
Why and How LLMs Hallucinate: Connecting the Dots with Subsequence Associations
par: Sun, Yiyou, et autres
Publié: (2025)
par: Sun, Yiyou, et autres
Publié: (2025)
Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries
par: Curran, Damian, et autres
Publié: (2025)
par: Curran, Damian, et autres
Publié: (2025)
Action100M: A Large-scale Video Action Dataset
par: Chen, Delong, et autres
Publié: (2026)
par: Chen, Delong, et autres
Publié: (2026)
Hallucination Detection with the Internal Layers of LLMs
par: Preiß, Martin
Publié: (2025)
par: Preiß, Martin
Publié: (2025)
Enhancing LLM Reliability via Explicit Knowledge Boundary Modeling
par: Zheng, Hang, et autres
Publié: (2025)
par: Zheng, Hang, et autres
Publié: (2025)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
par: Liu, Jiarui, et autres
Publié: (2025)
par: Liu, Jiarui, et autres
Publié: (2025)
Disentangling Prompt Element Level Risk Factors for Hallucinations and Omissions in Mental Health LLM Responses
par: Ni, Congning, et autres
Publié: (2026)
par: Ni, Congning, et autres
Publié: (2026)
ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
par: Gu, Yuzhe, et autres
Publié: (2024)
par: Gu, Yuzhe, et autres
Publié: (2024)
Towards Efficient and Robust VQA-NLE Data Generation with Large Vision-Language Models
par: Irawan, Patrick Amadeus, et autres
Publié: (2024)
par: Irawan, Patrick Amadeus, et autres
Publié: (2024)
Delusions of Large Language Models
par: Xu, Hongshen, et autres
Publié: (2025)
par: Xu, Hongshen, et autres
Publié: (2025)
Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models
par: Zhang, Yuji, et autres
Publié: (2024)
par: Zhang, Yuji, et autres
Publié: (2024)
Documents similaires
-
High-Dimension Human Value Representation in Large Language Models
par: Cahyawijaya, Samuel, et autres
Publié: (2024) -
Belief Revision: The Adaptability of Large Language Models Reasoning
par: Wilie, Bryan, et autres
Publié: (2024) -
High-Dimensional Interlingual Representations of Large Language Models
par: Wilie, Bryan, et autres
Publié: (2025) -
What Makes for Good Image Captions?
par: Chen, Delong, et autres
Publié: (2024) -
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
par: Lovenia, Holy, et autres
Publié: (2023)