TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kashmira, Savini, Dantanarayana, Jayanaka L., Brodsky, Joshua, Mahendra, Ashish, Kang, Yiping, Flautner, Krisztian, Tang, Lingjia, Mars, Jason
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914141662347264
author Kashmira, Savini
Dantanarayana, Jayanaka L.
Brodsky, Joshua
Mahendra, Ashish
Kang, Yiping
Flautner, Krisztian
Tang, Lingjia
Mars, Jason
author_facet Kashmira, Savini
Dantanarayana, Jayanaka L.
Brodsky, Joshua
Mahendra, Ashish
Kang, Yiping
Flautner, Krisztian
Tang, Lingjia
Mars, Jason
contents Retrieval-Augmented Generation (RAG) is one of the leading and most widely used techniques for enhancing LLM retrieval capabilities, but it still faces significant limitations in commercial use cases. RAG primarily relies on the query-chunk text-to-text similarity in the embedding space for retrieval and can fail to capture deeper semantic relationships across chunks, is highly sensitive to chunking strategies, and is prone to hallucinations. To address these challenges, we propose TOBUGraph, a graph-based retrieval framework that first constructs the knowledge graph from unstructured data dynamically and automatically. Using LLMs, TOBUGraph extracts structured knowledge and diverse relationships among data, going beyond RAG's text-to-text similarity. Retrieval is achieved through graph traversal, leveraging the extracted relationships and structures to enhance retrieval accuracy, eliminating the need for chunking configurations while reducing hallucination. We demonstrate TOBUGraph's effectiveness in TOBU, a real-world application in production for personal memory organization and retrieval. Our evaluation using real user data demonstrates that TOBUGraph outperforms multiple RAG implementations in both precision and recall, significantly improving user experience through improved retrieval accuracy.
format Preprint
id arxiv_https___arxiv_org_abs_2412_05447
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
Kashmira, Savini
Dantanarayana, Jayanaka L.
Brodsky, Joshua
Mahendra, Ashish
Kang, Yiping
Flautner, Krisztian
Tang, Lingjia
Mars, Jason
Machine Learning
Artificial Intelligence
Information Retrieval
Retrieval-Augmented Generation (RAG) is one of the leading and most widely used techniques for enhancing LLM retrieval capabilities, but it still faces significant limitations in commercial use cases. RAG primarily relies on the query-chunk text-to-text similarity in the embedding space for retrieval and can fail to capture deeper semantic relationships across chunks, is highly sensitive to chunking strategies, and is prone to hallucinations. To address these challenges, we propose TOBUGraph, a graph-based retrieval framework that first constructs the knowledge graph from unstructured data dynamically and automatically. Using LLMs, TOBUGraph extracts structured knowledge and diverse relationships among data, going beyond RAG's text-to-text similarity. Retrieval is achieved through graph traversal, leveraging the extracted relationships and structures to enhance retrieval accuracy, eliminating the need for chunking configurations while reducing hallucination. We demonstrate TOBUGraph's effectiveness in TOBU, a real-world application in production for personal memory organization and retrieval. Our evaluation using real user data demonstrates that TOBUGraph outperforms multiple RAG implementations in both precision and recall, significantly improving user experience through improved retrieval accuracy.
title TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
topic Machine Learning
Artificial Intelligence
Information Retrieval
url https://arxiv.org/abs/2412.05447