KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Jiawei, Xu, Chejian, Gai, Yu, Lecue, Freddy, Song, Dawn, Li, Bo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Halu-J: Critique-Based Hallucination Judge
por: Wang, Binjie, et al.
Publicado: (2024)
por: Wang, Binjie, et al.
Publicado: (2024)
Sanity Checks for Long-Form Hallucination Detection
por: Zollicoffer, Geigh, et al.
Publicado: (2026)
por: Zollicoffer, Geigh, et al.
Publicado: (2026)
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models
por: Agarwal, Vibhor, et al.
Publicado: (2024)
por: Agarwal, Vibhor, et al.
Publicado: (2024)
SymLoc: Symbolic Localization of Hallucination across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)
por: Lamba, Naveen, et al.
Publicado: (2025)
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)
por: Lamba, Naveen, et al.
Publicado: (2025)
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
por: Chen, Kedi, et al.
Publicado: (2024)
por: Chen, Kedi, et al.
Publicado: (2024)
ChatScene: Knowledge-Enabled Safety-Critical Scenario Generation for Autonomous Vehicles
por: Zhang, Jiawei, et al.
Publicado: (2024)
por: Zhang, Jiawei, et al.
Publicado: (2024)
On Early Detection of Hallucinations in Factual Question Answering
por: Snyder, Ben, et al.
Publicado: (2023)
por: Snyder, Ben, et al.
Publicado: (2023)
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
por: Pan, Wenbo, et al.
Publicado: (2025)
por: Pan, Wenbo, et al.
Publicado: (2025)
DyKnow: Dynamically Verifying Time-Sensitive Factual Knowledge in LLMs
por: Mousavi, Seyed Mahed, et al.
Publicado: (2024)
por: Mousavi, Seyed Mahed, et al.
Publicado: (2024)
Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
por: Rahman, Subhey Sadi, et al.
Publicado: (2025)
por: Rahman, Subhey Sadi, et al.
Publicado: (2025)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
por: Zhang, Xiaoying, et al.
Publicado: (2024)
por: Zhang, Xiaoying, et al.
Publicado: (2024)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
por: Yu, Lei, et al.
Publicado: (2024)
por: Yu, Lei, et al.
Publicado: (2024)
The First Token Knows: Single-Decode Confidence for Hallucination Detection
por: Gabriel, Mina
Publicado: (2026)
por: Gabriel, Mina
Publicado: (2026)
REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models
por: Lee, DongGeon, et al.
Publicado: (2025)
por: Lee, DongGeon, et al.
Publicado: (2025)
KG-FPQ: Evaluating Factuality Hallucination in LLMs with Knowledge Graph-based False Premise Questions
por: Zhu, Yanxu, et al.
Publicado: (2024)
por: Zhu, Yanxu, et al.
Publicado: (2024)
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
por: Li, Junliang, et al.
Publicado: (2025)
por: Li, Junliang, et al.
Publicado: (2025)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
por: Mao, Nathan, et al.
Publicado: (2026)
por: Mao, Nathan, et al.
Publicado: (2026)
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning
por: Wang, Shengyuan, et al.
Publicado: (2025)
por: Wang, Shengyuan, et al.
Publicado: (2025)
Graphusion: Leveraging Large Language Models for Scientific Knowledge Graph Fusion and Construction in NLP Education
por: Yang, Rui, et al.
Publicado: (2024)
por: Yang, Rui, et al.
Publicado: (2024)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
por: Ferrando, Javier, et al.
Publicado: (2024)
por: Ferrando, Javier, et al.
Publicado: (2024)
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality
por: Ren, Baochang, et al.
Publicado: (2025)
por: Ren, Baochang, et al.
Publicado: (2025)
AlignCheck: a Semantic Open-Domain Metric for Factual Consistency Assessment
por: Aghaebrahimian, Ahmad
Publicado: (2025)
por: Aghaebrahimian, Ahmad
Publicado: (2025)
WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries
por: Zhao, Wenting, et al.
Publicado: (2024)
por: Zhao, Wenting, et al.
Publicado: (2024)
MAD-Fact: A Multi-Agent Debate Framework for Long-Form Factuality Evaluation in LLMs
por: Ning, Yucheng, et al.
Publicado: (2025)
por: Ning, Yucheng, et al.
Publicado: (2025)
Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies
por: Ye, Yuxuan, et al.
Publicado: (2026)
por: Ye, Yuxuan, et al.
Publicado: (2026)
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models
por: Tran, Hieu, et al.
Publicado: (2024)
por: Tran, Hieu, et al.
Publicado: (2024)
Principled Detection of Hallucinations in Large Language Models via Multiple Testing
por: Li, Jiawei, et al.
Publicado: (2025)
por: Li, Jiawei, et al.
Publicado: (2025)
Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding
por: Zhu, Yifan, et al.
Publicado: (2026)
por: Zhu, Yifan, et al.
Publicado: (2026)
Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction
por: Wu, Shanglin, et al.
Publicado: (2025)
por: Wu, Shanglin, et al.
Publicado: (2025)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
por: Tian, Yuchen, et al.
Publicado: (2024)
por: Tian, Yuchen, et al.
Publicado: (2024)
Reasoning Factual Knowledge in Structured Data with Large Language Models
por: Huang, Sirui, et al.
Publicado: (2024)
por: Huang, Sirui, et al.
Publicado: (2024)
Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm
por: Wang, Haoyu, et al.
Publicado: (2026)
por: Wang, Haoyu, et al.
Publicado: (2026)
How Does Response Length Affect Long-Form Factuality
por: Zhao, James Xu, et al.
Publicado: (2025)
por: Zhao, James Xu, et al.
Publicado: (2025)
MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
por: Wu, Juncheng, et al.
Publicado: (2025)
por: Wu, Juncheng, et al.
Publicado: (2025)
TacoERE: Cluster-aware Compression for Event Relation Extraction
por: Guan, Yong, et al.
Publicado: (2024)
por: Guan, Yong, et al.
Publicado: (2024)
Graphusion: A RAG Framework for Knowledge Graph Construction with a Global Perspective
por: Yang, Rui, et al.
Publicado: (2024)
por: Yang, Rui, et al.
Publicado: (2024)
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
por: Sawczyn, Albert, et al.
Publicado: (2025)
por: Sawczyn, Albert, et al.
Publicado: (2025)
Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models
por: Li, Junyi, et al.
Publicado: (2025)
por: Li, Junyi, et al.
Publicado: (2025)
Editing Factual Knowledge and Explanatory Ability of Medical Large Language Models
por: Xu, Derong, et al.
Publicado: (2024)
por: Xu, Derong, et al.
Publicado: (2024)
Ejemplares similares
-
Halu-J: Critique-Based Hallucination Judge
por: Wang, Binjie, et al.
Publicado: (2024) -
Sanity Checks for Long-Form Hallucination Detection
por: Zollicoffer, Geigh, et al.
Publicado: (2026) -
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models
por: Agarwal, Vibhor, et al.
Publicado: (2024) -
SymLoc: Symbolic Localization of Hallucination across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025) -
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)