Beyond Similarity Search: A Unified Data Layer for Production RAG Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Budigi, Venkata Krishna Prasanth, Sirigiri, Siri Chandana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
von: Bindschaedler, Laurent
Veröffentlicht: (2026)
von: Bindschaedler, Laurent
Veröffentlicht: (2026)
QVCache: A Query-Aware Vector Cache
von: Göçer, Anıl Eren, et al.
Veröffentlicht: (2026)
von: Göçer, Anıl Eren, et al.
Veröffentlicht: (2026)
HONEYBEE: Efficient Role-based Access Control for Vector Databases via Dynamic Partitioning[Technical Report]
von: Zhong, Hongbin, et al.
Veröffentlicht: (2025)
von: Zhong, Hongbin, et al.
Veröffentlicht: (2025)
LiveVectorLake: A Real-Time Versioned Knowledge Base Architecture for Streaming Vector Updates and Temporal Retrieval
von: Prajapati, Tarun
Veröffentlicht: (2025)
von: Prajapati, Tarun
Veröffentlicht: (2025)
Adaptive Prefiltering for High-Dimensional Similarity Search: A Frequency-Aware Approach
von: Calin, Teodor-Ioan
Veröffentlicht: (2025)
von: Calin, Teodor-Ioan
Veröffentlicht: (2025)
Grokers: Bottom-Up Inductive Comprehension and Write-Time Intelligence over Typed Knowledge Graphs
von: Magarshak, Gregory
Veröffentlicht: (2026)
von: Magarshak, Gregory
Veröffentlicht: (2026)
Passing the Baton: High Throughput Distributed Disk-Based Vector Search with BatANN
von: Dang, Nam Anh, et al.
Veröffentlicht: (2025)
von: Dang, Nam Anh, et al.
Veröffentlicht: (2025)
Siren Federate: Bridging document, relational, and graph models for exploratory graph analysis
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
Exact Trajectory Similarity Search With N-tree: An Efficient Metric Index for kNN and Range Queries
von: Güting, Ralf Hartmut, et al.
Veröffentlicht: (2024)
von: Güting, Ralf Hartmut, et al.
Veröffentlicht: (2024)
AlayaDB: The Data Foundation for Efficient and Effective Long-context LLM Inference
von: Deng, Yangshen, et al.
Veröffentlicht: (2025)
von: Deng, Yangshen, et al.
Veröffentlicht: (2025)
MisEdu-RAG: A Misconception-Aware Dual-Hypergraph RAG for Novice Math Teachers
von: Guo, Zhihan, et al.
Veröffentlicht: (2026)
von: Guo, Zhihan, et al.
Veröffentlicht: (2026)
Lambda: Learning Matchable Prior For Entity Alignment with Unlabeled Dangling Cases
von: Yin, Hang, et al.
Veröffentlicht: (2024)
von: Yin, Hang, et al.
Veröffentlicht: (2024)
Cost-Effective, Low Latency Vector Search with Azure Cosmos DB
von: Upreti, Nitish, et al.
Veröffentlicht: (2025)
von: Upreti, Nitish, et al.
Veröffentlicht: (2025)
Timehash: Hierarchical Time Indexing for Efficient Business Hours Search
von: Kim, Jinoh, et al.
Veröffentlicht: (2026)
von: Kim, Jinoh, et al.
Veröffentlicht: (2026)
flexvec: SQL Vector Retrieval with Programmatic Embedding Modulation
von: Delmas, Damian
Veröffentlicht: (2026)
von: Delmas, Damian
Veröffentlicht: (2026)
Space efficient implementation of hypergraph dualization in the D-basis algorithm
von: Homan, Skylar, et al.
Veröffentlicht: (2025)
von: Homan, Skylar, et al.
Veröffentlicht: (2025)
TRACE: A Time-Relational Approximate Cubing Engine for Fast Data Insights
von: Sivakumar, Suharsh, et al.
Veröffentlicht: (2024)
von: Sivakumar, Suharsh, et al.
Veröffentlicht: (2024)
UnWeaving the knots of GraphRAG -- turns out VectorRAG is almost enough
von: Tuora, Ryszard, et al.
Veröffentlicht: (2026)
von: Tuora, Ryszard, et al.
Veröffentlicht: (2026)
Transforming OPACs into Intelligent Discovery Systems: An AI-Powered, Knowledge Graph-Driven Smart OPAC for Digital Libraries
von: Rajeevan, M. S., et al.
Veröffentlicht: (2026)
von: Rajeevan, M. S., et al.
Veröffentlicht: (2026)
DatAasee -- A Metadata-Lake as Metadata Catalog for a Virtual Data-Lake
von: Himpe, Christian
Veröffentlicht: (2024)
von: Himpe, Christian
Veröffentlicht: (2024)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
von: Zhang, Liangliang, et al.
Veröffentlicht: (2025)
von: Zhang, Liangliang, et al.
Veröffentlicht: (2025)
Routing End User Queries to Enterprise Databases
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
From BM25 to Corrective RAG: Benchmarking Retrieval Strategies for Text-and-Table Documents
von: Akarsu, Meftun, et al.
Veröffentlicht: (2026)
von: Akarsu, Meftun, et al.
Veröffentlicht: (2026)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026)
von: Park, Sungho, et al.
Veröffentlicht: (2026)
Federated Semantic Knowledge Graphs for Laboratory Workflows: A Structured Expert Elicitation Methodology Demonstrated Through Bioanalytical Workflow Twins
von: Schachner, Luis F., et al.
Veröffentlicht: (2026)
von: Schachner, Luis F., et al.
Veröffentlicht: (2026)
Ontologies for Models and Algorithms in Applied Mathematics and Related Disciplines
von: Schembera, Björn, et al.
Veröffentlicht: (2023)
von: Schembera, Björn, et al.
Veröffentlicht: (2023)
Federated Continual Recommendation
von: Lim, Jaehyung, et al.
Veröffentlicht: (2025)
von: Lim, Jaehyung, et al.
Veröffentlicht: (2025)
Leveraging LLMs to Enable Natural Language Search on Go-to-market Platforms
von: Yao, Jesse, et al.
Veröffentlicht: (2024)
von: Yao, Jesse, et al.
Veröffentlicht: (2024)
LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help?
von: Takehi, Rikiya, et al.
Veröffentlicht: (2024)
von: Takehi, Rikiya, et al.
Veröffentlicht: (2024)
Large Language Models, Knowledge Graphs and Search Engines: A Crossroads for Answering Users' Questions
von: Hogan, Aidan, et al.
Veröffentlicht: (2025)
von: Hogan, Aidan, et al.
Veröffentlicht: (2025)
Diversification as Risk Minimization
von: Takehi, Rikiya, et al.
Veröffentlicht: (2025)
von: Takehi, Rikiya, et al.
Veröffentlicht: (2025)
H-MAPS: Hierarchical Memory-Augmented Proactive Search Assistant for Scientific Literature
von: Nishikawa, Koji, et al.
Veröffentlicht: (2026)
von: Nishikawa, Koji, et al.
Veröffentlicht: (2026)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
PRECEPT: Planning Resilience via Experience, Context Engineering & Probing Trajectories A Unified Framework for Test-Time Adaptation with Compositional Rule Learning and Pareto-Guided Prompt Evolution
von: Shahmansoori, Arash
Veröffentlicht: (2026)
von: Shahmansoori, Arash
Veröffentlicht: (2026)
Behavior-Aware Dual-Channel Preference Learning for Heterogeneous Sequential Recommendation
von: Xiao, Jing, et al.
Veröffentlicht: (2026)
von: Xiao, Jing, et al.
Veröffentlicht: (2026)
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
von: Bose, Joy
Veröffentlicht: (2026)
von: Bose, Joy
Veröffentlicht: (2026)
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
A Robust and Efficient Pipeline for Enterprise-Level Large-Scale Entity Resolution
von: Kannangara, Sandeepa, et al.
Veröffentlicht: (2025)
von: Kannangara, Sandeepa, et al.
Veröffentlicht: (2025)
EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
von: Sun, Yuhong, et al.
Veröffentlicht: (2026)
von: Sun, Yuhong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026) -
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
von: Bindschaedler, Laurent
Veröffentlicht: (2026) -
QVCache: A Query-Aware Vector Cache
von: Göçer, Anıl Eren, et al.
Veröffentlicht: (2026) -
HONEYBEE: Efficient Role-based Access Control for Vector Databases via Dynamic Partitioning[Technical Report]
von: Zhong, Hongbin, et al.
Veröffentlicht: (2025) -
LiveVectorLake: A Real-Time Versioned Knowledge Base Architecture for Streaming Vector Updates and Temporal Retrieval
von: Prajapati, Tarun
Veröffentlicht: (2025)