ContextCache: Context-Aware Semantic Cache for Multi-Turn Queries in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yan, Jianxin, Ni, Wangze, Chen, Lei, Lin, Xuemin, Cheng, Peng, Qin, Zhan, Ren, Kui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAC: Relation-Aware Cache Replacement for Large Language Models
by: Wu, Yuchong, et al.
Published: (2026)
by: Wu, Yuchong, et al.
Published: (2026)
QCFuse: Query-Centric Cache Fusion for Efficient RAG Inference
by: Yan, Jianxin, et al.
Published: (2026)
by: Yan, Jianxin, et al.
Published: (2026)
StructRide: A Framework to Exploit the Structure Information of Shareability Graph in Ridesharing
by: Zhan, Jiexi, et al.
Published: (2024)
by: Zhan, Jiexi, et al.
Published: (2024)
ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs
by: Qi, Yanlin, et al.
Published: (2026)
by: Qi, Yanlin, et al.
Published: (2026)
KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction
by: Kim, Jang-Hyun, et al.
Published: (2025)
by: Kim, Jang-Hyun, et al.
Published: (2025)
Wait to be Faster: a Smart Pooling Framework for Dynamic Ridesharing
by: Zhong, Xiaoyao, et al.
Published: (2024)
by: Zhong, Xiaoyao, et al.
Published: (2024)
DCMF: A Dynamic Context Monitoring and Caching Framework for Context Management Platforms
by: Manchanda, Ashish, et al.
Published: (2025)
by: Manchanda, Ashish, et al.
Published: (2025)
Efficient Multiple Temporal Network Kernel Density Estimation
by: Shao, Yu, et al.
Published: (2025)
by: Shao, Yu, et al.
Published: (2025)
QVCache: A Query-Aware Vector Cache
by: Göçer, Anıl Eren, et al.
Published: (2026)
by: Göçer, Anıl Eren, et al.
Published: (2026)
CARPO: Leveraging Listwise Learning-to-Rank for Context-Aware Query Plan Optimization
by: Zhou, Wenrui, et al.
Published: (2025)
by: Zhou, Wenrui, et al.
Published: (2025)
Category-Aware Semantic Caching for Heterogeneous LLM Workloads
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
Structured Prompt Language: Declarative Context Management for LLMs
by: Gong, Wen G.
Published: (2026)
by: Gong, Wen G.
Published: (2026)
Hybrid Querying Over Relational Databases and Large Language Models
by: Zhao, Fuheng, et al.
Published: (2024)
by: Zhao, Fuheng, et al.
Published: (2024)
HyperJoin: LLM-augmented Hypergraph Link Prediction for Joinable Table Discovery
by: Liu, Shiyuan, et al.
Published: (2026)
by: Liu, Shiyuan, et al.
Published: (2026)
SQL-Encoder: Improving NL2SQL In-Context Learning Through a Context-Aware Encoder
by: Pourreza, Mohammadreza, et al.
Published: (2024)
by: Pourreza, Mohammadreza, et al.
Published: (2024)
Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)
by: Bindschaedler, Laurent
Published: (2026)
by: Bindschaedler, Laurent
Published: (2026)
Context-Driven Index Trimming: A Data Quality Perspective to Enhancing Precision of RALMs
by: Ma, Kexin, et al.
Published: (2024)
by: Ma, Kexin, et al.
Published: (2024)
LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models
by: Shi, Dachuan, et al.
Published: (2025)
by: Shi, Dachuan, et al.
Published: (2025)
One-Hop Sub-Query Result Caches for Graph Database Systems
by: Nguyen, Hieu, et al.
Published: (2024)
by: Nguyen, Hieu, et al.
Published: (2024)
Batch Hop-Constrained s-t Simple Path Query Processing in Large Graphs
by: Yuan, Long, et al.
Published: (2023)
by: Yuan, Long, et al.
Published: (2023)
Numerical Estimation of Spatial Distributions under Differential Privacy
by: Du, Leilei, et al.
Published: (2024)
by: Du, Leilei, et al.
Published: (2024)
ZipCache: A DRAM/SSD Cache with Built-in Transparent Compression
by: Xie, Rui, et al.
Published: (2024)
by: Xie, Rui, et al.
Published: (2024)
Context-Enriched Natural Language Descriptions of Vessel Trajectories
by: Patroumpas, Kostas, et al.
Published: (2026)
by: Patroumpas, Kostas, et al.
Published: (2026)
NAT-NL2GQL: A Novel Multi-Agent Framework for Translating Natural Language to Graph Query Language
by: Liang, Yuanyuan, et al.
Published: (2024)
by: Liang, Yuanyuan, et al.
Published: (2024)
Halo: Domain-Aware Query Optimization for Long-Context Question Answering
by: Chunduri, Pramod, et al.
Published: (2026)
by: Chunduri, Pramod, et al.
Published: (2026)
Graph Query Generation with Constraint-guided Large Language Agents
by: Wang, Mengying, et al.
Published: (2026)
by: Wang, Mengying, et al.
Published: (2026)
Can Language Models Enable In-Context Database?
by: Pan, Yu, et al.
Published: (2024)
by: Pan, Yu, et al.
Published: (2024)
TailorSQL: An NL2SQL System Tailored to Your Query Workload
by: Vaidya, Kapil, et al.
Published: (2025)
by: Vaidya, Kapil, et al.
Published: (2025)
Entity-Aware and Secure Query Optimization in Database Using Named Entity Recognition
by: Sultana, Azrin, et al.
Published: (2026)
by: Sultana, Azrin, et al.
Published: (2026)
Query Performance Explanation through Large Language Model for HTAP Systems
by: Xiu, Haibo, et al.
Published: (2024)
by: Xiu, Haibo, et al.
Published: (2024)
Natural Language Querying System Through Entity Enrichment
by: Amavi, Joshua, et al.
Published: (2024)
by: Amavi, Joshua, et al.
Published: (2024)
SAM: A Stability-Aware Cache Manager for Multi-Tenant Embedded Databases
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
LLM-R2: A Large Language Model Enhanced Rule-based Rewrite System for Boosting Query Efficiency
by: Li, Zhaodonghui, et al.
Published: (2024)
by: Li, Zhaodonghui, et al.
Published: (2024)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
by: Fu, Tianyu, et al.
Published: (2025)
by: Fu, Tianyu, et al.
Published: (2025)
Prompt Engineering Techniques for Context-dependent Text-to-SQL in Arabic
by: Almohaimeed, Saleh, et al.
Published: (2025)
by: Almohaimeed, Saleh, et al.
Published: (2025)
Relational Database Augmented Large Language Model
by: Qin, Zongyue, et al.
Published: (2024)
by: Qin, Zongyue, et al.
Published: (2024)
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering
by: Lin, Teng, et al.
Published: (2025)
by: Lin, Teng, et al.
Published: (2025)
Toward Multi-Database Query Reasoning for Text2Cypher
by: Ozsoy, Makbule Gulcin
Published: (2026)
by: Ozsoy, Makbule Gulcin
Published: (2026)
Constant-time Connectivity and 2-Edge Connectivity Querying in Dynamic Graphs
by: Xu, Lantian, et al.
Published: (2026)
by: Xu, Lantian, et al.
Published: (2026)
A Survey of Large Language Models on Generative Graph Analytics: Query, Learning, and Applications
by: Shang, Wenbo, et al.
Published: (2024)
by: Shang, Wenbo, et al.
Published: (2024)
Similar Items
-
RAC: Relation-Aware Cache Replacement for Large Language Models
by: Wu, Yuchong, et al.
Published: (2026) -
QCFuse: Query-Centric Cache Fusion for Efficient RAG Inference
by: Yan, Jianxin, et al.
Published: (2026) -
StructRide: A Framework to Exploit the Structure Information of Shareability Graph in Ridesharing
by: Zhan, Jiexi, et al.
Published: (2024) -
ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs
by: Qi, Yanlin, et al.
Published: (2026) -
KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction
by: Kim, Jang-Hyun, et al.
Published: (2025)