Span-Level Hallucination Detection for LLM-Generated Answers
Fuente:
arXiv
Saved in:
| Main Authors: | Elchafei, Passant, Abu-Elkheir, Mervet |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
by: Elchafei, Passant, et al.
Published: (2026)
by: Elchafei, Passant, et al.
Published: (2026)
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
by: Elchafei, Passant, et al.
Published: (2025)
by: Elchafei, Passant, et al.
Published: (2025)
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
by: Elchafei, Passant, et al.
Published: (2025)
by: Elchafei, Passant, et al.
Published: (2025)
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
by: Elchafei, Passant, et al.
Published: (2026)
by: Elchafei, Passant, et al.
Published: (2026)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
by: Yehuda, Yakir, et al.
Published: (2024)
by: Yehuda, Yakir, et al.
Published: (2024)
When Models Lie, We Learn: Multilingual Span-Level Hallucination Detection with PsiloQA
by: Rykov, Elisei, et al.
Published: (2025)
by: Rykov, Elisei, et al.
Published: (2025)
Learning to Reason for Hallucination Span Detection
by: Su, Hsuan, et al.
Published: (2025)
by: Su, Hsuan, et al.
Published: (2025)
Hallucinated Span Detection with Multi-View Attention Features
by: Ogasa, Yuya, et al.
Published: (2025)
by: Ogasa, Yuya, et al.
Published: (2025)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
by: Vemula, Saketh Reddy, et al.
Published: (2025)
by: Vemula, Saketh Reddy, et al.
Published: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
by: Wang, Changyue, et al.
Published: (2025)
by: Wang, Changyue, et al.
Published: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
LLM Hallucination Detection: HSAD
by: Li, JinXin, et al.
Published: (2025)
by: Li, JinXin, et al.
Published: (2025)
ATLANTIS at SemEval-2025 Task 3: Detecting Hallucinated Text Spans in Question Answering
by: Kobus, Catherine, et al.
Published: (2025)
by: Kobus, Catherine, et al.
Published: (2025)
Hallucination-Free Automatic Question & Answer Generation for Intuitive Learning
by: Wang, Nicholas X., et al.
Published: (2026)
by: Wang, Nicholas X., et al.
Published: (2026)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
by: Yao, Siyang, et al.
Published: (2026)
by: Yao, Siyang, et al.
Published: (2026)
Lowest Span Confidence: A Zero-Shot Metric for Efficient and Black-Box Hallucination Detection in LLMs
by: Qiao, Yitong, et al.
Published: (2026)
by: Qiao, Yitong, et al.
Published: (2026)
A Dataset of Open-Domain Question Answering with Multiple-Span Answers
by: Luo, Zhiyi, et al.
Published: (2024)
by: Luo, Zhiyi, et al.
Published: (2024)
Detecting Hallucinations in Authentic LLM-Human Interactions
by: Ren, Yujie, et al.
Published: (2025)
by: Ren, Yujie, et al.
Published: (2025)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
by: Yeom, Jewon, et al.
Published: (2026)
by: Yeom, Jewon, et al.
Published: (2026)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
by: Yeh, Min-Hsuan, et al.
Published: (2025)
by: Yeh, Min-Hsuan, et al.
Published: (2025)
Span Modeling for Idiomaticity and Figurative Language Detection with Span Contrastive Loss
by: Matheny, Blake, et al.
Published: (2026)
by: Matheny, Blake, et al.
Published: (2026)
Integrated Framework for LLM Evaluation with Answer Generation
by: Lee, Sujeong, et al.
Published: (2025)
by: Lee, Sujeong, et al.
Published: (2025)
Correct after Answer: Enhancing Multi-Span Question Answering with Post-Processing Method
by: Lin, Jiayi, et al.
Published: (2024)
by: Lin, Jiayi, et al.
Published: (2024)
GigaCheck: Detecting LLM-generated Content via Object-Centric Span Localization
by: Tolstykh, Irina, et al.
Published: (2024)
by: Tolstykh, Irina, et al.
Published: (2024)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
by: Powers, Maximus, et al.
Published: (2024)
by: Powers, Maximus, et al.
Published: (2024)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
by: Urlana, Ashok, et al.
Published: (2025)
by: Urlana, Ashok, et al.
Published: (2025)
Enhancing LLM-Based Short Answer Grading with Retrieval-Augmented Generation
by: Chu, Yucheng, et al.
Published: (2025)
by: Chu, Yucheng, et al.
Published: (2025)
How Does Beam Search improve Span-Level Confidence Estimation in Generative Sequence Labeling?
by: Hashimoto, Kazuma, et al.
Published: (2022)
by: Hashimoto, Kazuma, et al.
Published: (2022)
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
Span-Level Machine Translation Meta-Evaluation
by: Perrella, Stefano, et al.
Published: (2026)
by: Perrella, Stefano, et al.
Published: (2026)
Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads
by: Li, Guojing, et al.
Published: (2026)
by: Li, Guojing, et al.
Published: (2026)
Steer LLM Latents for Hallucination Detection
by: Park, Seongheon, et al.
Published: (2025)
by: Park, Seongheon, et al.
Published: (2025)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
by: Rahman, A B M Ashikur, et al.
Published: (2024)
by: Rahman, A B M Ashikur, et al.
Published: (2024)
Target Span Detection for Implicit Harmful Content
by: Jafari, Nazanin, et al.
Published: (2024)
by: Jafari, Nazanin, et al.
Published: (2024)
Semantic Search as Extractive Paraphrase Span Detection
by: Kanerva, Jenna, et al.
Published: (2021)
by: Kanerva, Jenna, et al.
Published: (2021)
Banishing LLM Hallucinations Requires Rethinking Generalization
by: Li, Johnny, et al.
Published: (2024)
by: Li, Johnny, et al.
Published: (2024)
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
by: Huang, Sicong, et al.
Published: (2025)
by: Huang, Sicong, et al.
Published: (2025)
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
by: Xu, Yangyifan, et al.
Published: (2024)
by: Xu, Yangyifan, et al.
Published: (2024)
Similar Items
-
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
by: Elchafei, Passant, et al.
Published: (2026) -
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
by: Elchafei, Passant, et al.
Published: (2025) -
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
by: Elchafei, Passant, et al.
Published: (2025) -
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
by: Elchafei, Passant, et al.
Published: (2026) -
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
by: Yehuda, Yakir, et al.
Published: (2024)