QVCache: A Query-Aware Vector Cache
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Göçer, Anıl Eren, Tsakalidou, Ioanna, Nicholson, Hamish, Kim, Kyoungmin, Ailamaki, Anastasia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
von: Bindschaedler, Laurent
Veröffentlicht: (2026)
von: Bindschaedler, Laurent
Veröffentlicht: (2026)
Beyond Similarity Search: A Unified Data Layer for Production RAG Systems
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
Routing End User Queries to Enterprise Databases
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
Low-Latency Stateful Stream Processing through Timely and Accurate Prefetching
von: Zapridou, Eleni, et al.
Veröffentlicht: (2026)
von: Zapridou, Eleni, et al.
Veröffentlicht: (2026)
LiveVectorLake: A Real-Time Versioned Knowledge Base Architecture for Streaming Vector Updates and Temporal Retrieval
von: Prajapati, Tarun
Veröffentlicht: (2025)
von: Prajapati, Tarun
Veröffentlicht: (2025)
HONEYBEE: Efficient Role-based Access Control for Vector Databases via Dynamic Partitioning[Technical Report]
von: Zhong, Hongbin, et al.
Veröffentlicht: (2025)
von: Zhong, Hongbin, et al.
Veröffentlicht: (2025)
BoomHQ: Learning to Boost Multiple Hybrid Queries on Vector DBMSs
von: Qiu, Ermu, et al.
Veröffentlicht: (2026)
von: Qiu, Ermu, et al.
Veröffentlicht: (2026)
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
Process Faster, Pay Less: Functional Isolation for Stream Processing
von: Zapridou, Eleni, et al.
Veröffentlicht: (2026)
von: Zapridou, Eleni, et al.
Veröffentlicht: (2026)
PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error Guarantees (Technical Report)
von: Zhu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhu, Yuxuan, et al.
Veröffentlicht: (2025)
Partial Adaptive Indexing for Approximate Query Answering
von: Maroulis, Stavros, et al.
Veröffentlicht: (2024)
von: Maroulis, Stavros, et al.
Veröffentlicht: (2024)
Grokers: Bottom-Up Inductive Comprehension and Write-Time Intelligence over Typed Knowledge Graphs
von: Magarshak, Gregory
Veröffentlicht: (2026)
von: Magarshak, Gregory
Veröffentlicht: (2026)
Deep Research is the New Analytics System: Towards Building the Runtime for AI-Driven Analytics
von: Russo, Matthew, et al.
Veröffentlicht: (2025)
von: Russo, Matthew, et al.
Veröffentlicht: (2025)
Bi-View Embedding Fusion: A Hybrid Learning Approach for Knowledge Graph's Nodes Classification Addressing Problems with Limited Data
von: Napoli, Rosario, et al.
Veröffentlicht: (2025)
von: Napoli, Rosario, et al.
Veröffentlicht: (2025)
OnPair: Short Strings Compression for Fast Random Access
von: Gargiulo, Francesco, et al.
Veröffentlicht: (2025)
von: Gargiulo, Francesco, et al.
Veröffentlicht: (2025)
Valori: A Deterministic Memory Substrate for AI Systems
von: Gudur, Varshith
Veröffentlicht: (2025)
von: Gudur, Varshith
Veröffentlicht: (2025)
Passing the Baton: High Throughput Distributed Disk-Based Vector Search with BatANN
von: Dang, Nam Anh, et al.
Veröffentlicht: (2025)
von: Dang, Nam Anh, et al.
Veröffentlicht: (2025)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
flexvec: SQL Vector Retrieval with Programmatic Embedding Modulation
von: Delmas, Damian
Veröffentlicht: (2026)
von: Delmas, Damian
Veröffentlicht: (2026)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026)
von: Park, Sungho, et al.
Veröffentlicht: (2026)
Siren Federate: Bridging document, relational, and graph models for exploratory graph analysis
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
SAM: A Stability-Aware Cache Manager for Multi-Tenant Embedded Databases
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
Clinical Knowledge Graph Construction and Evaluation with Multi-LLMs via Retrieval-Augmented Generation
von: Das, Udiptaman, et al.
Veröffentlicht: (2026)
von: Das, Udiptaman, et al.
Veröffentlicht: (2026)
Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing
von: Bayer, Thomas, et al.
Veröffentlicht: (2026)
von: Bayer, Thomas, et al.
Veröffentlicht: (2026)
Proving correctness for SQL implementations of OCL constraints
von: Nguyen, Hoang, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2024)
Risk-Aware GPU-Assisted Cardinality Estimation for Cost-Based Query Optimizers
von: Chang, Ilsun
Veröffentlicht: (2025)
von: Chang, Ilsun
Veröffentlicht: (2025)
The Case for Intent-Based Query Rewriting
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
AlayaDB: The Data Foundation for Efficient and Effective Long-context LLM Inference
von: Deng, Yangshen, et al.
Veröffentlicht: (2025)
von: Deng, Yangshen, et al.
Veröffentlicht: (2025)
Adaptive Prefiltering for High-Dimensional Similarity Search: A Frequency-Aware Approach
von: Calin, Teodor-Ioan
Veröffentlicht: (2025)
von: Calin, Teodor-Ioan
Veröffentlicht: (2025)
Bringing Private Reads to Hyperledger Fabric via Private Information Retrieval
von: Iasenovets, Artur, et al.
Veröffentlicht: (2025)
von: Iasenovets, Artur, et al.
Veröffentlicht: (2025)
Wave-Based Semantic Memory with Resonance-Based Retrieval: A Phase-Aware Alternative to Vector Embedding Stores
von: Listopad, Aleksandr
Veröffentlicht: (2025)
von: Listopad, Aleksandr
Veröffentlicht: (2025)
"My agent understands me better": Integrating Dynamic Human-like Memory Recall and Consolidation in LLM-Based Agents
von: Hou, Yuki, et al.
Veröffentlicht: (2024)
von: Hou, Yuki, et al.
Veröffentlicht: (2024)
Ontology-based Semantic Similarity Measures for Clustering Medical Concepts in Drug Safety
von: Painter, Jeffery L, et al.
Veröffentlicht: (2025)
von: Painter, Jeffery L, et al.
Veröffentlicht: (2025)
Semantic Similarity-Informed Bayesian Borrowing for Quantitative Signal Detection of Adverse Events
von: Haguinet, François, et al.
Veröffentlicht: (2025)
von: Haguinet, François, et al.
Veröffentlicht: (2025)
Optimizing Navigational Graph Queries
von: Mulder, Thomas, et al.
Veröffentlicht: (2024)
von: Mulder, Thomas, et al.
Veröffentlicht: (2024)
Avoiding Materialisation for Guarded Aggregate Queries
von: Lanzinger, Matthias, et al.
Veröffentlicht: (2024)
von: Lanzinger, Matthias, et al.
Veröffentlicht: (2024)
Tractable Conjunctive Queries over Static and Dynamic Relations
von: Kara, Ahmet, et al.
Veröffentlicht: (2024)
von: Kara, Ahmet, et al.
Veröffentlicht: (2024)
Yannakakis+: Practical Acyclic Query Evaluation with Theoretical Guarantees
von: Wang, Qichen, et al.
Veröffentlicht: (2025)
von: Wang, Qichen, et al.
Veröffentlicht: (2025)
Samyama: A Unified Graph-Vector Database with In-Database Optimization, Agentic Enrichment, and Hardware Acceleration
von: Mandarapu, Madhulatha, et al.
Veröffentlicht: (2026)
von: Mandarapu, Madhulatha, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
von: Bindschaedler, Laurent
Veröffentlicht: (2026) -
Beyond Similarity Search: A Unified Data Layer for Production RAG Systems
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026) -
Routing End User Queries to Enterprise Databases
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026) -
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026) -
Low-Latency Stateful Stream Processing through Timely and Accurate Prefetching
von: Zapridou, Eleni, et al.
Veröffentlicht: (2026)