Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wu, Zihao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diversification as Risk Minimization
von: Takehi, Rikiya, et al.
Veröffentlicht: (2025)
von: Takehi, Rikiya, et al.
Veröffentlicht: (2025)
Leveraging LLMs to Enable Natural Language Search on Go-to-market Platforms
von: Yao, Jesse, et al.
Veröffentlicht: (2024)
von: Yao, Jesse, et al.
Veröffentlicht: (2024)
Routing End User Queries to Enterprise Databases
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
von: Whittaker, Edward, et al.
Veröffentlicht: (2024)
von: Whittaker, Edward, et al.
Veröffentlicht: (2024)
The Case for Intent-Based Query Rewriting
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
Automating Pharmacovigilance Evidence Generation: Using Large Language Models to Produce Context-Aware SQL
von: Painter, Jeffery L., et al.
Veröffentlicht: (2024)
von: Painter, Jeffery L., et al.
Veröffentlicht: (2024)
A Grounded Memory System For Smart Personal Assistants
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
von: Sun, Yuhong, et al.
Veröffentlicht: (2026)
von: Sun, Yuhong, et al.
Veröffentlicht: (2026)
Large language models in finance : what is financial sentiment?
von: Kirtac, Kemal, et al.
Veröffentlicht: (2025)
von: Kirtac, Kemal, et al.
Veröffentlicht: (2025)
Method for Aggregating Unstructured Data Using Large Language Models
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026)
von: Park, Sungho, et al.
Veröffentlicht: (2026)
SEAR: Schema-Based Evaluation and Routing for LLM Gateways
von: Zhang, Zecheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zecheng, et al.
Veröffentlicht: (2026)
Scaling Multilingual Semantic Search in Uber Eats Delivery
von: Ling, Bo, et al.
Veröffentlicht: (2026)
von: Ling, Bo, et al.
Veröffentlicht: (2026)
Cost Trade-offs of Reasoning and Non-Reasoning Large Language Models in Text-to-SQL
von: Deochake, Saurabh, et al.
Veröffentlicht: (2025)
von: Deochake, Saurabh, et al.
Veröffentlicht: (2025)
flexvec: SQL Vector Retrieval with Programmatic Embedding Modulation
von: Delmas, Damian
Veröffentlicht: (2026)
von: Delmas, Damian
Veröffentlicht: (2026)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
REIS: A High-Performance and Energy-Efficient Retrieval System with In-Storage Processing
von: Chen, Kangqi, et al.
Veröffentlicht: (2025)
von: Chen, Kangqi, et al.
Veröffentlicht: (2025)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
von: Yao, Zhihui, et al.
Veröffentlicht: (2026)
von: Yao, Zhihui, et al.
Veröffentlicht: (2026)
Permanent Data Encoding (PDE): A Visual Language for Semantic Compression and Knowledge Preservation in 3-Character Units
von: Tsuyuki, Yoshiharu, et al.
Veröffentlicht: (2025)
von: Tsuyuki, Yoshiharu, et al.
Veröffentlicht: (2025)
Cultural Encoding in Large Language Models: The Existence Gap in AI-Mediated Brand Discovery
von: Junyao, Huang, et al.
Veröffentlicht: (2025)
von: Junyao, Huang, et al.
Veröffentlicht: (2025)
The Cultural Gene of Large Language Models: A Study on the Impact of Cross-Corpus Training on Model Values and Biases
von: Fenech-Borg, Emanuel Z., et al.
Veröffentlicht: (2025)
von: Fenech-Borg, Emanuel Z., et al.
Veröffentlicht: (2025)
Large Language Models for Relevance Judgment in Product Search
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
von: Iannelli, Michael, et al.
Veröffentlicht: (2024)
von: Iannelli, Michael, et al.
Veröffentlicht: (2024)
When Content is Goliath and Algorithm is David: The Style and Semantic Effects of Generative Search Engine
von: Ma, Lijia, et al.
Veröffentlicht: (2025)
von: Ma, Lijia, et al.
Veröffentlicht: (2025)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
von: Liu, Jianan, et al.
Veröffentlicht: (2026)
von: Liu, Jianan, et al.
Veröffentlicht: (2026)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
von: Kai, Zhang, et al.
Veröffentlicht: (2026)
von: Kai, Zhang, et al.
Veröffentlicht: (2026)
Session Context Embedding for Intent Understanding in Product Search
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
Reconnecting Fragmented Citation Networks with Semantic Augmentation
von: Huong, Vu Thi, et al.
Veröffentlicht: (2026)
von: Huong, Vu Thi, et al.
Veröffentlicht: (2026)
Real-World En Call Center Transcripts Dataset with PII Redaction
von: Dao, Ha, et al.
Veröffentlicht: (2025)
von: Dao, Ha, et al.
Veröffentlicht: (2025)
When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory
von: Shao, Jiaqi, et al.
Veröffentlicht: (2026)
von: Shao, Jiaqi, et al.
Veröffentlicht: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
von: Che, Jiarui, et al.
Veröffentlicht: (2026)
von: Che, Jiarui, et al.
Veröffentlicht: (2026)
A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering
von: Ros, Sereiwathna, et al.
Veröffentlicht: (2026)
von: Ros, Sereiwathna, et al.
Veröffentlicht: (2026)
ArcheType: A Novel Framework for Open-Source Column Type Annotation using Large Language Models
von: Feuer, Benjamin, et al.
Veröffentlicht: (2023)
von: Feuer, Benjamin, et al.
Veröffentlicht: (2023)
Towards Adaptive Context Management for Intelligent Conversational Question Answering
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2025)
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
von: Oruesagasti, Julen
Veröffentlicht: (2026)
von: Oruesagasti, Julen
Veröffentlicht: (2026)
From Native Memes to Global Moderation: Cross-Cultural Evaluation of Vision-Language Models for Hateful Meme Detection
von: Wang, Mo, et al.
Veröffentlicht: (2026)
von: Wang, Mo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Diversification as Risk Minimization
von: Takehi, Rikiya, et al.
Veröffentlicht: (2025) -
Leveraging LLMs to Enable Natural Language Search on Go-to-market Platforms
von: Yao, Jesse, et al.
Veröffentlicht: (2024) -
Routing End User Queries to Enterprise Databases
von: Sudarshan, Saikrishna, et al.
Veröffentlicht: (2026) -
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
von: Whittaker, Edward, et al.
Veröffentlicht: (2024) -
The Case for Intent-Based Query Rewriting
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)