Span-Level Hallucination Detection for LLM-Generated Answers
Fuente:
arXiv
Salvato in:
| Autori principali: | Elchafei, Passant, Abu-Elkheir, Mervet |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
di: Elchafei, Passant, et al.
Pubblicazione: (2026)
di: Elchafei, Passant, et al.
Pubblicazione: (2026)
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
di: Elchafei, Passant, et al.
Pubblicazione: (2026)
di: Elchafei, Passant, et al.
Pubblicazione: (2026)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
di: Yehuda, Yakir, et al.
Pubblicazione: (2024)
di: Yehuda, Yakir, et al.
Pubblicazione: (2024)
When Models Lie, We Learn: Multilingual Span-Level Hallucination Detection with PsiloQA
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
Learning to Reason for Hallucination Span Detection
di: Su, Hsuan, et al.
Pubblicazione: (2025)
di: Su, Hsuan, et al.
Pubblicazione: (2025)
Hallucinated Span Detection with Multi-View Attention Features
di: Ogasa, Yuya, et al.
Pubblicazione: (2025)
di: Ogasa, Yuya, et al.
Pubblicazione: (2025)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
di: Vemula, Saketh Reddy, et al.
Pubblicazione: (2025)
di: Vemula, Saketh Reddy, et al.
Pubblicazione: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
di: Wang, Changyue, et al.
Pubblicazione: (2025)
di: Wang, Changyue, et al.
Pubblicazione: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
di: Wei, Jiaheng, et al.
Pubblicazione: (2024)
di: Wei, Jiaheng, et al.
Pubblicazione: (2024)
LLM Hallucination Detection: HSAD
di: Li, JinXin, et al.
Pubblicazione: (2025)
di: Li, JinXin, et al.
Pubblicazione: (2025)
ATLANTIS at SemEval-2025 Task 3: Detecting Hallucinated Text Spans in Question Answering
di: Kobus, Catherine, et al.
Pubblicazione: (2025)
di: Kobus, Catherine, et al.
Pubblicazione: (2025)
Hallucination-Free Automatic Question & Answer Generation for Intuitive Learning
di: Wang, Nicholas X., et al.
Pubblicazione: (2026)
di: Wang, Nicholas X., et al.
Pubblicazione: (2026)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
di: Yao, Siyang, et al.
Pubblicazione: (2026)
di: Yao, Siyang, et al.
Pubblicazione: (2026)
Lowest Span Confidence: A Zero-Shot Metric for Efficient and Black-Box Hallucination Detection in LLMs
di: Qiao, Yitong, et al.
Pubblicazione: (2026)
di: Qiao, Yitong, et al.
Pubblicazione: (2026)
A Dataset of Open-Domain Question Answering with Multiple-Span Answers
di: Luo, Zhiyi, et al.
Pubblicazione: (2024)
di: Luo, Zhiyi, et al.
Pubblicazione: (2024)
Detecting Hallucinations in Authentic LLM-Human Interactions
di: Ren, Yujie, et al.
Pubblicazione: (2025)
di: Ren, Yujie, et al.
Pubblicazione: (2025)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
di: Du, Xuefeng, et al.
Pubblicazione: (2024)
di: Du, Xuefeng, et al.
Pubblicazione: (2024)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
Span Modeling for Idiomaticity and Figurative Language Detection with Span Contrastive Loss
di: Matheny, Blake, et al.
Pubblicazione: (2026)
di: Matheny, Blake, et al.
Pubblicazione: (2026)
Integrated Framework for LLM Evaluation with Answer Generation
di: Lee, Sujeong, et al.
Pubblicazione: (2025)
di: Lee, Sujeong, et al.
Pubblicazione: (2025)
Correct after Answer: Enhancing Multi-Span Question Answering with Post-Processing Method
di: Lin, Jiayi, et al.
Pubblicazione: (2024)
di: Lin, Jiayi, et al.
Pubblicazione: (2024)
GigaCheck: Detecting LLM-generated Content via Object-Centric Span Localization
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
di: Powers, Maximus, et al.
Pubblicazione: (2024)
di: Powers, Maximus, et al.
Pubblicazione: (2024)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
Enhancing LLM-Based Short Answer Grading with Retrieval-Augmented Generation
di: Chu, Yucheng, et al.
Pubblicazione: (2025)
di: Chu, Yucheng, et al.
Pubblicazione: (2025)
How Does Beam Search improve Span-Level Confidence Estimation in Generative Sequence Labeling?
di: Hashimoto, Kazuma, et al.
Pubblicazione: (2022)
di: Hashimoto, Kazuma, et al.
Pubblicazione: (2022)
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation
di: Li, Zixuan, et al.
Pubblicazione: (2024)
di: Li, Zixuan, et al.
Pubblicazione: (2024)
Span-Level Machine Translation Meta-Evaluation
di: Perrella, Stefano, et al.
Pubblicazione: (2026)
di: Perrella, Stefano, et al.
Pubblicazione: (2026)
Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads
di: Li, Guojing, et al.
Pubblicazione: (2026)
di: Li, Guojing, et al.
Pubblicazione: (2026)
Steer LLM Latents for Hallucination Detection
di: Park, Seongheon, et al.
Pubblicazione: (2025)
di: Park, Seongheon, et al.
Pubblicazione: (2025)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
di: Rahman, A B M Ashikur, et al.
Pubblicazione: (2024)
di: Rahman, A B M Ashikur, et al.
Pubblicazione: (2024)
Target Span Detection for Implicit Harmful Content
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
Semantic Search as Extractive Paraphrase Span Detection
di: Kanerva, Jenna, et al.
Pubblicazione: (2021)
di: Kanerva, Jenna, et al.
Pubblicazione: (2021)
Banishing LLM Hallucinations Requires Rethinking Generalization
di: Li, Johnny, et al.
Pubblicazione: (2024)
di: Li, Johnny, et al.
Pubblicazione: (2024)
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
di: Huang, Sicong, et al.
Pubblicazione: (2025)
di: Huang, Sicong, et al.
Pubblicazione: (2025)
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
di: Xu, Yangyifan, et al.
Pubblicazione: (2024)
di: Xu, Yangyifan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
di: Elchafei, Passant, et al.
Pubblicazione: (2026) -
Multimodal Arabic Captioning with Interpretable Visual Concept Integration
di: Elchafei, Passant, et al.
Pubblicazione: (2025) -
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
di: Elchafei, Passant, et al.
Pubblicazione: (2025) -
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
di: Elchafei, Passant, et al.
Pubblicazione: (2026) -
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
di: Yehuda, Yakir, et al.
Pubblicazione: (2024)