Struct-X: Enhancing Large Language Models Reasoning with Structured Data

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Tan, Xiaoyu, Wang, Haoyu, Qiu, Xihe, Cheng, Yuan, Xu, Yinghui, Chu, Wei, Qi, Yuan
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929424504455168
author Tan, Xiaoyu
Wang, Haoyu
Qiu, Xihe
Cheng, Yuan
Xu, Yinghui
Chu, Wei
Qi, Yuan
author_facet Tan, Xiaoyu
Wang, Haoyu
Qiu, Xihe
Cheng, Yuan
Xu, Yinghui
Chu, Wei
Qi, Yuan
contents Structured data, rich in logical and relational information, has the potential to enhance the reasoning abilities of large language models (LLMs). Still, its integration poses a challenge due to the risk of overwhelming LLMs with excessive tokens and irrelevant context information. To address this, we propose Struct-X, a novel framework that operates through five key phases: ``read-model-fill-reflect-reason'' efficiently enabling LLMs to utilize structured data. It begins by encoding structured data into a topological space using graph embeddings, followed by filling in missing entity information with knowledge retrieval modules, and filtering out irrelevant tokens via a self-supervised module. The final phase involves constructing a topological network with selected tokens to further reduce the total token length for more effective LLM inference. Additionally, Struct-X includes an Auxiliary Module trained to generate prompts, aiding LLMs in analyzing structured data. Extensive experiments on benchmarks, including the knowledge graph question-answer task and the long document reading comprehension task, show that Struct-X notably improves LLM reasoning, demonstrating the effectiveness of structured data augmentation in improving LLM inference with complex input context.
format Preprint
id arxiv_https___arxiv_org_abs_2407_12522
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Struct-X: Enhancing Large Language Models Reasoning with Structured Data
Tan, Xiaoyu
Wang, Haoyu
Qiu, Xihe
Cheng, Yuan
Xu, Yinghui
Chu, Wei
Qi, Yuan
Computation and Language
Artificial Intelligence
Structured data, rich in logical and relational information, has the potential to enhance the reasoning abilities of large language models (LLMs). Still, its integration poses a challenge due to the risk of overwhelming LLMs with excessive tokens and irrelevant context information. To address this, we propose Struct-X, a novel framework that operates through five key phases: ``read-model-fill-reflect-reason'' efficiently enabling LLMs to utilize structured data. It begins by encoding structured data into a topological space using graph embeddings, followed by filling in missing entity information with knowledge retrieval modules, and filtering out irrelevant tokens via a self-supervised module. The final phase involves constructing a topological network with selected tokens to further reduce the total token length for more effective LLM inference. Additionally, Struct-X includes an Auxiliary Module trained to generate prompts, aiding LLMs in analyzing structured data. Extensive experiments on benchmarks, including the knowledge graph question-answer task and the long document reading comprehension task, show that Struct-X notably improves LLM reasoning, demonstrating the effectiveness of structured data augmentation in improving LLM inference with complex input context.
title Struct-X: Enhancing Large Language Models Reasoning with Structured Data
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2407.12522