Span-Level Hallucination Detection for LLM-Generated Answers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Elchafei, Passant, Abu-Elkheir, Mervet |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
von: Elchafei, Passant, et al.
Veröffentlicht: (2026)
von: Elchafei, Passant, et al.
Veröffentlicht: (2026)
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
von: Elchafei, Passant, et al.
Veröffentlicht: (2026)
von: Elchafei, Passant, et al.
Veröffentlicht: (2026)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
von: Yehuda, Yakir, et al.
Veröffentlicht: (2024)
von: Yehuda, Yakir, et al.
Veröffentlicht: (2024)
When Models Lie, We Learn: Multilingual Span-Level Hallucination Detection with PsiloQA
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
Learning to Reason for Hallucination Span Detection
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
Hallucinated Span Detection with Multi-View Attention Features
von: Ogasa, Yuya, et al.
Veröffentlicht: (2025)
von: Ogasa, Yuya, et al.
Veröffentlicht: (2025)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
von: Vemula, Saketh Reddy, et al.
Veröffentlicht: (2025)
von: Vemula, Saketh Reddy, et al.
Veröffentlicht: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
von: Wei, Jiaheng, et al.
Veröffentlicht: (2024)
von: Wei, Jiaheng, et al.
Veröffentlicht: (2024)
LLM Hallucination Detection: HSAD
von: Li, JinXin, et al.
Veröffentlicht: (2025)
von: Li, JinXin, et al.
Veröffentlicht: (2025)
ATLANTIS at SemEval-2025 Task 3: Detecting Hallucinated Text Spans in Question Answering
von: Kobus, Catherine, et al.
Veröffentlicht: (2025)
von: Kobus, Catherine, et al.
Veröffentlicht: (2025)
Hallucination-Free Automatic Question & Answer Generation for Intuitive Learning
von: Wang, Nicholas X., et al.
Veröffentlicht: (2026)
von: Wang, Nicholas X., et al.
Veröffentlicht: (2026)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
von: Yao, Siyang, et al.
Veröffentlicht: (2026)
von: Yao, Siyang, et al.
Veröffentlicht: (2026)
Lowest Span Confidence: A Zero-Shot Metric for Efficient and Black-Box Hallucination Detection in LLMs
von: Qiao, Yitong, et al.
Veröffentlicht: (2026)
von: Qiao, Yitong, et al.
Veröffentlicht: (2026)
A Dataset of Open-Domain Question Answering with Multiple-Span Answers
von: Luo, Zhiyi, et al.
Veröffentlicht: (2024)
von: Luo, Zhiyi, et al.
Veröffentlicht: (2024)
Detecting Hallucinations in Authentic LLM-Human Interactions
von: Ren, Yujie, et al.
Veröffentlicht: (2025)
von: Ren, Yujie, et al.
Veröffentlicht: (2025)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
Span Modeling for Idiomaticity and Figurative Language Detection with Span Contrastive Loss
von: Matheny, Blake, et al.
Veröffentlicht: (2026)
von: Matheny, Blake, et al.
Veröffentlicht: (2026)
Integrated Framework for LLM Evaluation with Answer Generation
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
Correct after Answer: Enhancing Multi-Span Question Answering with Post-Processing Method
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
GigaCheck: Detecting LLM-generated Content via Object-Centric Span Localization
von: Tolstykh, Irina, et al.
Veröffentlicht: (2024)
von: Tolstykh, Irina, et al.
Veröffentlicht: (2024)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
von: Powers, Maximus, et al.
Veröffentlicht: (2024)
von: Powers, Maximus, et al.
Veröffentlicht: (2024)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
von: Urlana, Ashok, et al.
Veröffentlicht: (2025)
von: Urlana, Ashok, et al.
Veröffentlicht: (2025)
Enhancing LLM-Based Short Answer Grading with Retrieval-Augmented Generation
von: Chu, Yucheng, et al.
Veröffentlicht: (2025)
von: Chu, Yucheng, et al.
Veröffentlicht: (2025)
How Does Beam Search improve Span-Level Confidence Estimation in Generative Sequence Labeling?
von: Hashimoto, Kazuma, et al.
Veröffentlicht: (2022)
von: Hashimoto, Kazuma, et al.
Veröffentlicht: (2022)
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
Span-Level Machine Translation Meta-Evaluation
von: Perrella, Stefano, et al.
Veröffentlicht: (2026)
von: Perrella, Stefano, et al.
Veröffentlicht: (2026)
Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads
von: Li, Guojing, et al.
Veröffentlicht: (2026)
von: Li, Guojing, et al.
Veröffentlicht: (2026)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
von: Rahman, A B M Ashikur, et al.
Veröffentlicht: (2024)
von: Rahman, A B M Ashikur, et al.
Veröffentlicht: (2024)
Target Span Detection for Implicit Harmful Content
von: Jafari, Nazanin, et al.
Veröffentlicht: (2024)
von: Jafari, Nazanin, et al.
Veröffentlicht: (2024)
Semantic Search as Extractive Paraphrase Span Detection
von: Kanerva, Jenna, et al.
Veröffentlicht: (2021)
von: Kanerva, Jenna, et al.
Veröffentlicht: (2021)
Banishing LLM Hallucinations Requires Rethinking Generalization
von: Li, Johnny, et al.
Veröffentlicht: (2024)
von: Li, Johnny, et al.
Veröffentlicht: (2024)
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
von: Huang, Sicong, et al.
Veröffentlicht: (2025)
von: Huang, Sicong, et al.
Veröffentlicht: (2025)
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
von: Xu, Yangyifan, et al.
Veröffentlicht: (2024)
von: Xu, Yangyifan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
von: Elchafei, Passant, et al.
Veröffentlicht: (2026) -
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
von: Elchafei, Passant, et al.
Veröffentlicht: (2025) -
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
von: Elchafei, Passant, et al.
Veröffentlicht: (2025) -
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
von: Elchafei, Passant, et al.
Veröffentlicht: (2026) -
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
von: Yehuda, Yakir, et al.
Veröffentlicht: (2024)