MultiHal: Multilingual Dataset for Knowledge-Graph Grounded Evaluation of LLM Hallucinations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lavrinovics, Ernests, Biswas, Russa, Hose, Katja, Bjerva, Johannes |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Knowledge Graphs, Large Language Models, and Hallucinations: An NLP Perspective
von: Lavrinovics, Ernests, et al.
Veröffentlicht: (2024)
von: Lavrinovics, Ernests, et al.
Veröffentlicht: (2024)
Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality
von: Zhang, Mike, et al.
Veröffentlicht: (2025)
von: Zhang, Mike, et al.
Veröffentlicht: (2025)
Against All Odds: Overcoming Typology, Script, and Language Confusion in Multilingual Embedding Inversion Attacks
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
Limited-Resource Adapters Are Regularizers, Not Linguists
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
MedHal: An Evaluation Dataset for Medical Hallucination Detection
von: Mehenni, Gaya, et al.
Veröffentlicht: (2025)
von: Mehenni, Gaya, et al.
Veröffentlicht: (2025)
Linguistically Grounded Analysis of Language Models using Shapley Head Values
von: Fekete, Marcell, et al.
Veröffentlicht: (2024)
von: Fekete, Marcell, et al.
Veröffentlicht: (2024)
Multilingual Gradient Word-Order Typology from Universal Dependencies
von: Baylor, Emi, et al.
Veröffentlicht: (2024)
von: Baylor, Emi, et al.
Veröffentlicht: (2024)
InFact: Informativeness Alignment for Improved LLM Factuality
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
Text Embedding Inversion Security for Multilingual Language Models
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
CreoleVal: Multilingual Multitask Benchmarks for Creoles
von: Lent, Heather, et al.
Veröffentlicht: (2023)
von: Lent, Heather, et al.
Veröffentlicht: (2023)
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2025)
Patterns of Persistence and Diffusibility across the World's Languages
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
KGHaluBench: A Knowledge Graph-Based Hallucination Benchmark for Evaluating the Breadth and Depth of LLM Knowledge
von: Robertson, Alex, et al.
Veröffentlicht: (2026)
von: Robertson, Alex, et al.
Veröffentlicht: (2026)
Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
von: Jiang, Chaoya, et al.
Veröffentlicht: (2024)
von: Jiang, Chaoya, et al.
Veröffentlicht: (2024)
Grounding LLM Reasoning with Knowledge Graphs
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2025)
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2025)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework
von: Sansford, Hannah, et al.
Veröffentlicht: (2024)
von: Sansford, Hannah, et al.
Veröffentlicht: (2024)
When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs
von: Fekete, Marcell, et al.
Veröffentlicht: (2026)
von: Fekete, Marcell, et al.
Veröffentlicht: (2026)
Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking
von: Hu, Songbo, et al.
Veröffentlicht: (2026)
von: Hu, Songbo, et al.
Veröffentlicht: (2026)
FactNet: A Billion-Scale Knowledge Graph for Multilingual Factual Grounding
von: Shen, Yingli, et al.
Veröffentlicht: (2026)
von: Shen, Yingli, et al.
Veröffentlicht: (2026)
ChartHal: A Fine-grained Framework Evaluating Hallucination of Large Vision Language Models in Chart Understanding
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors
von: Miyazato, Ryuhei, et al.
Veröffentlicht: (2026)
von: Miyazato, Ryuhei, et al.
Veröffentlicht: (2026)
Multilingual Knowledge Graph Completion via Efficient Multilingual Knowledge Sharing
von: Mao, Cunli, et al.
Veröffentlicht: (2025)
von: Mao, Cunli, et al.
Veröffentlicht: (2025)
Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
MASSIVE Multilingual Abstract Meaning Representation: A Dataset and Baselines for Hallucination Detection
von: Regan, Michael, et al.
Veröffentlicht: (2024)
von: Regan, Michael, et al.
Veröffentlicht: (2024)
How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP
von: Tatariya, Kushal, et al.
Veröffentlicht: (2024)
von: Tatariya, Kushal, et al.
Veröffentlicht: (2024)
ForMaT: Dataset for Visually-Grounded Multilingual PDF Translation
von: Ciesiółka, Michał, et al.
Veröffentlicht: (2026)
von: Ciesiółka, Michał, et al.
Veröffentlicht: (2026)
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
Enhancing Knowledge Graph Construction: Evaluating with Emphasis on Hallucination, Omission, and Graph Similarity Metrics
von: Ghanem, Hussam, et al.
Veröffentlicht: (2025)
von: Ghanem, Hussam, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Actionable Course Evaluation Student Feedback to Lecturers
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
What is "Typological Diversity" in NLP?
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
LAGO: Few-shot Crosslingual Embedding Inversion Attacks via Language Similarity-Aware Graph Optimization
von: Yu, Wenrui, et al.
Veröffentlicht: (2025)
von: Yu, Wenrui, et al.
Veröffentlicht: (2025)
Multilingual KokoroChat: A Multi-LLM Ensemble Translation Method for Creating a Multilingual Counseling Dialogue Dataset
von: Suzuki, Ryoma, et al.
Veröffentlicht: (2026)
von: Suzuki, Ryoma, et al.
Veröffentlicht: (2026)
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
A Principled Framework for Evaluating on Typologically Diverse Languages
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
Multilingual Power and Ideology Identification in the Parliament: a Reference Dataset and Simple Baselines
von: Çöltekin, Çağrı, et al.
Veröffentlicht: (2024)
von: Çöltekin, Çağrı, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Knowledge Graphs, Large Language Models, and Hallucinations: An NLP Perspective
von: Lavrinovics, Ernests, et al.
Veröffentlicht: (2024) -
Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality
von: Zhang, Mike, et al.
Veröffentlicht: (2025) -
Against All Odds: Overcoming Typology, Script, and Language Confusion in Multilingual Embedding Inversion Attacks
von: Chen, Yiyi, et al.
Veröffentlicht: (2024) -
Limited-Resource Adapters Are Regularizers, Not Linguists
von: Fekete, Marcell, et al.
Veröffentlicht: (2025) -
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)