BoomHQ: Learning to Boost Multiple Hybrid Queries on Vector DBMSs
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiu, Ermu, Chen, Tianyi, Gao, Jun, Wei, Xing, Tu, Yaofeng, Han, Yinjun, Lin, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QVCache: A Query-Aware Vector Cache
di: Göçer, Anıl Eren, et al.
Pubblicazione: (2026)
di: Göçer, Anıl Eren, et al.
Pubblicazione: (2026)
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
di: Bindschaedler, Laurent
Pubblicazione: (2026)
di: Bindschaedler, Laurent
Pubblicazione: (2026)
Cost-Effective, Low Latency Vector Search with Azure Cosmos DB
di: Upreti, Nitish, et al.
Pubblicazione: (2025)
di: Upreti, Nitish, et al.
Pubblicazione: (2025)
QuIVer: Rethinking ANN Graph Topology via Training-Free Binary Quantization
di: Xiao, Wenxuan, et al.
Pubblicazione: (2026)
di: Xiao, Wenxuan, et al.
Pubblicazione: (2026)
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
di: Wu, Zihao
Pubblicazione: (2025)
di: Wu, Zihao
Pubblicazione: (2025)
The Case for Intent-Based Query Rewriting
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
flexvec: SQL Vector Retrieval with Programmatic Embedding Modulation
di: Delmas, Damian
Pubblicazione: (2026)
di: Delmas, Damian
Pubblicazione: (2026)
Timehash: Hierarchical Time Indexing for Efficient Business Hours Search
di: Kim, Jinoh, et al.
Pubblicazione: (2026)
di: Kim, Jinoh, et al.
Pubblicazione: (2026)
Routing End User Queries to Enterprise Databases
di: Sudarshan, Saikrishna, et al.
Pubblicazione: (2026)
di: Sudarshan, Saikrishna, et al.
Pubblicazione: (2026)
Deep Research is the New Analytics System: Towards Building the Runtime for AI-Driven Analytics
di: Russo, Matthew, et al.
Pubblicazione: (2025)
di: Russo, Matthew, et al.
Pubblicazione: (2025)
TRACE: A Time-Relational Approximate Cubing Engine for Fast Data Insights
di: Sivakumar, Suharsh, et al.
Pubblicazione: (2024)
di: Sivakumar, Suharsh, et al.
Pubblicazione: (2024)
Space efficient implementation of hypergraph dualization in the D-basis algorithm
di: Homan, Skylar, et al.
Pubblicazione: (2025)
di: Homan, Skylar, et al.
Pubblicazione: (2025)
Beyond Similarity Search: A Unified Data Layer for Production RAG Systems
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
Exact Trajectory Similarity Search With N-tree: An Efficient Metric Index for kNN and Range Queries
di: Güting, Ralf Hartmut, et al.
Pubblicazione: (2024)
di: Güting, Ralf Hartmut, et al.
Pubblicazione: (2024)
AlayaDB: The Data Foundation for Efficient and Effective Long-context LLM Inference
di: Deng, Yangshen, et al.
Pubblicazione: (2025)
di: Deng, Yangshen, et al.
Pubblicazione: (2025)
Categorical Calculus and Algebra for Multi-Model Data
di: Lu, Jiaheng
Pubblicazione: (2026)
di: Lu, Jiaheng
Pubblicazione: (2026)
DatAasee -- A Metadata-Lake as Metadata Catalog for a Virtual Data-Lake
di: Himpe, Christian
Pubblicazione: (2024)
di: Himpe, Christian
Pubblicazione: (2024)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
di: Park, Sungho, et al.
Pubblicazione: (2026)
di: Park, Sungho, et al.
Pubblicazione: (2026)
A Robust and Efficient Pipeline for Enterprise-Level Large-Scale Entity Resolution
di: Kannangara, Sandeepa, et al.
Pubblicazione: (2025)
di: Kannangara, Sandeepa, et al.
Pubblicazione: (2025)
FLASC: A Flare-Sensitive Clustering Algorithm
di: Bot, D. M., et al.
Pubblicazione: (2023)
di: Bot, D. M., et al.
Pubblicazione: (2023)
Towards a FAIR Documentation of Workflows and Models in Applied Mathematics
di: Reidelbach, Marco, et al.
Pubblicazione: (2024)
di: Reidelbach, Marco, et al.
Pubblicazione: (2024)
LiveVectorLake: A Real-Time Versioned Knowledge Base Architecture for Streaming Vector Updates and Temporal Retrieval
di: Prajapati, Tarun
Pubblicazione: (2025)
di: Prajapati, Tarun
Pubblicazione: (2025)
Poisoning Learned Index Structures: Static and Dynamic Adversarial Attacks on ALEX
di: Jue, Allen
Pubblicazione: (2026)
di: Jue, Allen
Pubblicazione: (2026)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
LLM-based Query Expansion Fails for Unfamiliar and Ambiguous Queries
di: Abe, Kenya, et al.
Pubblicazione: (2025)
di: Abe, Kenya, et al.
Pubblicazione: (2025)
Automating Pharmacovigilance Evidence Generation: Using Large Language Models to Produce Context-Aware SQL
di: Painter, Jeffery L., et al.
Pubblicazione: (2024)
di: Painter, Jeffery L., et al.
Pubblicazione: (2024)
Covariance Structure and Coordinate Heterogeneity Govern Binary Quantization of Contrastive Embeddings
di: Xiao, Wenxuan
Pubblicazione: (2026)
di: Xiao, Wenxuan
Pubblicazione: (2026)
Implementing the draft Graph Query Language Standard
di: Crowe, Malcolm, et al.
Pubblicazione: (2024)
di: Crowe, Malcolm, et al.
Pubblicazione: (2024)
HONEYBEE: Efficient Role-based Access Control for Vector Databases via Dynamic Partitioning[Technical Report]
di: Zhong, Hongbin, et al.
Pubblicazione: (2025)
di: Zhong, Hongbin, et al.
Pubblicazione: (2025)
ColBERT's [MASK]-based Query Augmentation: Effects of Quadrupling the Query Input Length
di: Giacalone, Ben, et al.
Pubblicazione: (2024)
di: Giacalone, Ben, et al.
Pubblicazione: (2024)
From Item-Only to Query-Item: Query-Conditioned Generative Search with QGS in Quark
di: Song, Yanglong, et al.
Pubblicazione: (2026)
di: Song, Yanglong, et al.
Pubblicazione: (2026)
Generating Query Recommendations via LLMs
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
Doc2Query++: Topic-Coverage based Document Expansion and its Application to Dense Retrieval via Dual-Index Fusion
di: Kuo, Tzu-Lin, et al.
Pubblicazione: (2025)
di: Kuo, Tzu-Lin, et al.
Pubblicazione: (2025)
Selective Query Processing: a Risk-Sensitive Selection of System Configurations
di: Mothe, Josiane, et al.
Pubblicazione: (2023)
di: Mothe, Josiane, et al.
Pubblicazione: (2023)
Quantum Computing for Query Containment of Conjunctive Queries
di: Gerlach, Luisa, et al.
Pubblicazione: (2026)
di: Gerlach, Luisa, et al.
Pubblicazione: (2026)
Method for Aggregating Unstructured Data Using Large Language Models
di: Lazebnyi, Vsevolod, et al.
Pubblicazione: (2026)
di: Lazebnyi, Vsevolod, et al.
Pubblicazione: (2026)
Cost Trade-offs of Reasoning and Non-Reasoning Large Language Models in Text-to-SQL
di: Deochake, Saurabh, et al.
Pubblicazione: (2025)
di: Deochake, Saurabh, et al.
Pubblicazione: (2025)
A Brief Comparison of Training-Free Multi-Vector Sequence Compression Methods
di: Jha, Rohan, et al.
Pubblicazione: (2026)
di: Jha, Rohan, et al.
Pubblicazione: (2026)
From Internet of Things Data to Business Processes: Challenges and a Framework
di: Mangler, Juergen, et al.
Pubblicazione: (2024)
di: Mangler, Juergen, et al.
Pubblicazione: (2024)
Self-Aware Vector Embeddings for Retrieval-Augmented Generation: A Neuroscience-Inspired Framework for Temporal, Confidence-Weighted, and Relational Knowledge
di: Xu, Naizhong
Pubblicazione: (2026)
di: Xu, Naizhong
Pubblicazione: (2026)
Documenti analoghi
-
QVCache: A Query-Aware Vector Cache
di: Göçer, Anıl Eren, et al.
Pubblicazione: (2026) -
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
di: Bindschaedler, Laurent
Pubblicazione: (2026) -
Cost-Effective, Low Latency Vector Search with Azure Cosmos DB
di: Upreti, Nitish, et al.
Pubblicazione: (2025) -
QuIVer: Rethinking ANN Graph Topology via Training-Free Binary Quantization
di: Xiao, Wenxuan, et al.
Pubblicazione: (2026) -
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
di: Wu, Zihao
Pubblicazione: (2025)