A review of faithfulness metrics for hallucination assessment in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Malin, Ben, Kalganova, Tatiana, Boulgouris, Nikoloas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
von: Malin, Ben, et al.
Veröffentlicht: (2025)
von: Malin, Ben, et al.
Veröffentlicht: (2025)
A comprehensive taxonomy of hallucinations in Large Language Models
von: Cossio, Manuel
Veröffentlicht: (2025)
von: Cossio, Manuel
Veröffentlicht: (2025)
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
A novel hallucination classification framework
von: Zavhorodnii, Maksym, et al.
Veröffentlicht: (2025)
von: Zavhorodnii, Maksym, et al.
Veröffentlicht: (2025)
`Generalization is hallucination' through the lens of tensor completions
von: Wong, Liang Ze
Veröffentlicht: (2025)
von: Wong, Liang Ze
Veröffentlicht: (2025)
A multilingual hallucination benchmark: MultiWikiQHalluA
von: Thoresen, Freja, et al.
Veröffentlicht: (2026)
von: Thoresen, Freja, et al.
Veröffentlicht: (2026)
Multilingual Large Language Models and Curse of Multilinguality
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
Probabilistic distances-based hallucination detection in LLMs with RAG
von: Oblovatny, Rodion, et al.
Veröffentlicht: (2025)
von: Oblovatny, Rodion, et al.
Veröffentlicht: (2025)
Large Language Models for Subjective Language Understanding: A Survey
von: Song, Changhao, et al.
Veröffentlicht: (2025)
von: Song, Changhao, et al.
Veröffentlicht: (2025)
Increasing faithfulness in human-human dialog summarization with Spoken Language Understanding tasks
von: Akani, Eunice, et al.
Veröffentlicht: (2024)
von: Akani, Eunice, et al.
Veröffentlicht: (2024)
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
Trustworthiness of Legal Considerations for the Use of LLMs in Education
von: Alaswad, Sara, et al.
Veröffentlicht: (2025)
von: Alaswad, Sara, et al.
Veröffentlicht: (2025)
BIRD: A Trustworthy Bayesian Inference Framework for Large Language Models
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
MHGraphBench: Knowledge Graph-Grounded Benchmarking of Mental Health Knowledge in Large Language Models
von: Liu, Weixin, et al.
Veröffentlicht: (2026)
von: Liu, Weixin, et al.
Veröffentlicht: (2026)
ACL-Verbatim: hallucination-free question answering for research
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
How do language models learn facts? Dynamics, curricula and hallucinations
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2025)
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2025)
Do LLM hallucination detectors suffer from low-resource effect?
von: Datta, Debtanu, et al.
Veröffentlicht: (2026)
von: Datta, Debtanu, et al.
Veröffentlicht: (2026)
Is Sarcasm Detection A Step-by-Step Reasoning Process in Large Language Models?
von: Yao, Ben, et al.
Veröffentlicht: (2024)
von: Yao, Ben, et al.
Veröffentlicht: (2024)
Diversity Helps Jailbreak Large Language Models
von: Zhao, Weiliang, et al.
Veröffentlicht: (2024)
von: Zhao, Weiliang, et al.
Veröffentlicht: (2024)
Large Language Models Are Human-Like Internally
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
von: Cui, Wendi, et al.
Veröffentlicht: (2024)
von: Cui, Wendi, et al.
Veröffentlicht: (2024)
Mixture-of-Agents Enhances Large Language Model Capabilities
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
Strong hallucinations from negation and how to fix them
von: Asher, Nicholas, et al.
Veröffentlicht: (2024)
von: Asher, Nicholas, et al.
Veröffentlicht: (2024)
Large Language Models for Multilingual Previously Fact-Checked Claim Detection
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
AILS-NTUA at SemEval-2024 Task 6: Efficient model tuning for hallucination detection and analysis
von: Grigoriadou, Natalia, et al.
Veröffentlicht: (2024)
von: Grigoriadou, Natalia, et al.
Veröffentlicht: (2024)
Towards Statistical Factuality Guarantee for Large Vision-Language Models
von: Li, Zhuohang, et al.
Veröffentlicht: (2025)
von: Li, Zhuohang, et al.
Veröffentlicht: (2025)
Ask-EDA: A Design Assistant Empowered by LLM, Hybrid RAG and Abbreviation De-hallucination
von: Shi, Luyao, et al.
Veröffentlicht: (2024)
von: Shi, Luyao, et al.
Veröffentlicht: (2024)
FABLES: Evaluating faithfulness and content selection in book-length summarization
von: Kim, Yekyung, et al.
Veröffentlicht: (2024)
von: Kim, Yekyung, et al.
Veröffentlicht: (2024)
The impact of fine tuning in LLaMA on hallucinations for named entity extraction in legal documentation
von: Vargas, Francisco, et al.
Veröffentlicht: (2025)
von: Vargas, Francisco, et al.
Veröffentlicht: (2025)
Bigger But Not Better: Small Neural Language Models Outperform Large Language Models in Detection of Thought Disorder
von: Li, Changye, et al.
Veröffentlicht: (2025)
von: Li, Changye, et al.
Veröffentlicht: (2025)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
Markov-Enhanced Clustering for Long Document Summarization: Tackling the 'Lost in the Middle' Challenge with Large Language Models
von: Amari, Aziz, et al.
Veröffentlicht: (2025)
von: Amari, Aziz, et al.
Veröffentlicht: (2025)
A Context-Aware Dual-Metric Framework for Confidence Estimation in Large Language Models
von: Yuan, Mingruo, et al.
Veröffentlicht: (2025)
von: Yuan, Mingruo, et al.
Veröffentlicht: (2025)
CPLLM: Clinical Prediction with Large Language Models
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2023)
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2023)
Beyond Textual Context: Structural Graph Encoding with Adaptive Space Alignment to alleviate the hallucination of LLMs
von: Zhang, Yifang, et al.
Veröffentlicht: (2025)
von: Zhang, Yifang, et al.
Veröffentlicht: (2025)
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2023)
Integrating Text and Time-Series into (Large) Language Models to Predict Medical Outcomes
von: Larbi, Iyadh Ben Cheikh, et al.
Veröffentlicht: (2025)
von: Larbi, Iyadh Ben Cheikh, et al.
Veröffentlicht: (2025)
Reducing hallucination in structured outputs via Retrieval-Augmented Generation
von: Béchard, Patrice, et al.
Veröffentlicht: (2024)
von: Béchard, Patrice, et al.
Veröffentlicht: (2024)
SLPL SHROOM at SemEval2024 Task 06: A comprehensive study on models ability to detect hallucination
von: Fallah, Pouya, et al.
Veröffentlicht: (2024)
von: Fallah, Pouya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
von: Malin, Ben, et al.
Veröffentlicht: (2025) -
A comprehensive taxonomy of hallucinations in Large Language Models
von: Cossio, Manuel
Veröffentlicht: (2025) -
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024) -
A novel hallucination classification framework
von: Zavhorodnii, Maksym, et al.
Veröffentlicht: (2025) -
`Generalization is hallucination' through the lens of tensor completions
von: Wong, Liang Ze
Veröffentlicht: (2025)