From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Gill, Gurbinder, Gupta, Ritvik, Lusson, Denis, Chandrashekar, Anand, Nguyen, Donald
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915514267205632
author Gill, Gurbinder
Gupta, Ritvik
Lusson, Denis
Chandrashekar, Anand
Nguyen, Donald
author_facet Gill, Gurbinder
Gupta, Ritvik
Lusson, Denis
Chandrashekar, Anand
Nguyen, Donald
contents Retrieval-Augmented Generation (RAG) has emerged as the standard paradigm for answering questions on enterprise data. Traditionally, RAG has centered on text-based semantic search and re-ranking. However, this approach falls short when dealing with questions beyond data summarization or non-text data. This has led to various attempts to supplement RAG to bridge the gap between RAG, the implementation paradigm, and the question answering problem that enterprise users expect it to solve. Given that contemporary RAG is a collection of techniques rather than a defined implementation, discussion of RAG and related question-answering systems benefits from a problem-oriented understanding. We propose a new classification framework (L1-L5) to categorize systems based on data modalities and task complexity of the underlying question answering problems: L1 (Surface Knowledge of Unstructured Data) through L4 (Reflective and Reasoned Knowledge) and the aspirational L5 (General Intelligence). We also introduce benchmarks aligned with these levels and evaluate four state-of-the-art platforms: LangChain, Azure AI Search, OpenAI, and Corvic AI. Our experiments highlight the value of multi-space retrieval and dynamic orchestration for enabling L1-L4 capabilities. We empirically validate our findings using diverse datasets indicative of enterprise use cases.
format Preprint
id arxiv_https___arxiv_org_abs_2509_21324
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
Gill, Gurbinder
Gupta, Ritvik
Lusson, Denis
Chandrashekar, Anand
Nguyen, Donald
Information Retrieval
Artificial Intelligence
I.2.1; I.2.7; I.2.8
Retrieval-Augmented Generation (RAG) has emerged as the standard paradigm for answering questions on enterprise data. Traditionally, RAG has centered on text-based semantic search and re-ranking. However, this approach falls short when dealing with questions beyond data summarization or non-text data. This has led to various attempts to supplement RAG to bridge the gap between RAG, the implementation paradigm, and the question answering problem that enterprise users expect it to solve. Given that contemporary RAG is a collection of techniques rather than a defined implementation, discussion of RAG and related question-answering systems benefits from a problem-oriented understanding. We propose a new classification framework (L1-L5) to categorize systems based on data modalities and task complexity of the underlying question answering problems: L1 (Surface Knowledge of Unstructured Data) through L4 (Reflective and Reasoned Knowledge) and the aspirational L5 (General Intelligence). We also introduce benchmarks aligned with these levels and evaluate four state-of-the-art platforms: LangChain, Azure AI Search, OpenAI, and Corvic AI. Our experiments highlight the value of multi-space retrieval and dynamic orchestration for enabling L1-L4 capabilities. We empirically validate our findings using diverse datasets indicative of enterprise use cases.
title From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
topic Information Retrieval
Artificial Intelligence
I.2.1; I.2.7; I.2.8
url https://arxiv.org/abs/2509.21324