A Unified Benchmark for Evaluating Knowledge Graph Construction Methods and Graph Neural Networks

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kabal, Othmane, Harzallah, Mounira, Guillet, Fabrice, Takeda, Hideaki, Ichise, Ryutaro
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910195530072064
author Kabal, Othmane
Harzallah, Mounira
Guillet, Fabrice
Takeda, Hideaki
Ichise, Ryutaro
author_facet Kabal, Othmane
Harzallah, Mounira
Guillet, Fabrice
Takeda, Hideaki
Ichise, Ryutaro
contents Knowledge graphs automatically constructed from text are increasingly used in real-world applications. However, their inherent noise, fragmentation, and semantic inconsistencies significantly affect the performance of Graph Neural Networks (GNNs) on downstream tasks. Assessing their performance and robustness remains difficult, as it is often unclear whether observed results stem from the learning model or from the quality of the constructed graph itself. In this work, we introduce a dual-purpose benchmark designed to jointly evaluate (i) the performance of GNNs on noisy, text-derived graphs and (ii) the effectiveness of graph construction methods on a downstream task. The benchmark is built in the biomedical domain from a single textual corpus and includes two automatically constructed graphs generated using different extraction methods, alongside a high-quality reference graph curated by experts that serves as an upper performance bound. This design enables controlled comparison of construction methods and systematic evaluation of GNN robustness through semi-supervised node classification. We further provide a standardized, reproducible, and extensible evaluation framework, facilitating the integration of new graph extraction methods and learning models.
format Preprint
id arxiv_https___arxiv_org_abs_2605_05476
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle A Unified Benchmark for Evaluating Knowledge Graph Construction Methods and Graph Neural Networks
Kabal, Othmane
Harzallah, Mounira
Guillet, Fabrice
Takeda, Hideaki
Ichise, Ryutaro
Machine Learning
Artificial Intelligence
Computation and Language
Knowledge graphs automatically constructed from text are increasingly used in real-world applications. However, their inherent noise, fragmentation, and semantic inconsistencies significantly affect the performance of Graph Neural Networks (GNNs) on downstream tasks. Assessing their performance and robustness remains difficult, as it is often unclear whether observed results stem from the learning model or from the quality of the constructed graph itself. In this work, we introduce a dual-purpose benchmark designed to jointly evaluate (i) the performance of GNNs on noisy, text-derived graphs and (ii) the effectiveness of graph construction methods on a downstream task. The benchmark is built in the biomedical domain from a single textual corpus and includes two automatically constructed graphs generated using different extraction methods, alongside a high-quality reference graph curated by experts that serves as an upper performance bound. This design enables controlled comparison of construction methods and systematic evaluation of GNN robustness through semi-supervised node classification. We further provide a standardized, reproducible, and extensible evaluation framework, facilitating the integration of new graph extraction methods and learning models.
title A Unified Benchmark for Evaluating Knowledge Graph Construction Methods and Graph Neural Networks
topic Machine Learning
Artificial Intelligence
Computation and Language
url https://arxiv.org/abs/2605.05476