An Ensemble Embedding Approach for Improving Semantic Caching Performance in LLM-based Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Ghaffari, Shervin, Bahranifard, Zohre, Akbari, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature
by: Maiti, Agniva, et al.
Published: (2025)
by: Maiti, Agniva, et al.
Published: (2025)
An Integrated Framework of Prompt Engineering and Multidimensional Knowledge Graphs for Legal Dispute Analysis
by: Zhang, Mingda, et al.
Published: (2025)
by: Zhang, Mingda, et al.
Published: (2025)
Sentiment analysis of texts from social networks based on machine learning methods for monitoring public sentiment
by: Nurlanuly, Arsen Tolebay
Published: (2025)
by: Nurlanuly, Arsen Tolebay
Published: (2025)
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
by: Nguyen, Andy, et al.
Published: (2026)
by: Nguyen, Andy, et al.
Published: (2026)
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
by: Kumar, Aayush
Published: (2025)
by: Kumar, Aayush
Published: (2025)
Bridging Industrial Expertise and XR with LLM-Powered Conversational Agents
by: Tomkou, Despina, et al.
Published: (2025)
by: Tomkou, Despina, et al.
Published: (2025)
Multi-source Heterogeneous Public Opinion Analysis via Collaborative Reasoning and Adaptive Fusion: A Systematically Integrated Approach
by: Liu, Yi
Published: (2026)
by: Liu, Yi
Published: (2026)
A Computational Approach to Modeling Conversational Systems: Analyzing Large-Scale Quasi-Patterned Dialogue Flows
by: Ammar, Mohamed Achref Ben, et al.
Published: (2025)
by: Ammar, Mohamed Achref Ben, et al.
Published: (2025)
A Case Study of Balanced Query Recommendation on Wikipedia
by: Mishra, Harshit, et al.
Published: (2025)
by: Mishra, Harshit, et al.
Published: (2025)
MeVer at CheckThat! 2026: Cluster-Aware Hard-Negative Mining for Multilingual Scientific-Source Retrieval
by: Bakagianni, Juli, et al.
Published: (2026)
by: Bakagianni, Juli, et al.
Published: (2026)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
Optimizing Retrieval-Augmented Generation (RAG) for Colloquial Cantonese: A LoRA-Based Systematic Review
by: Calonge, David Santandreu, et al.
Published: (2025)
by: Calonge, David Santandreu, et al.
Published: (2025)
LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics
by: Peyronnet, Antoine, et al.
Published: (2026)
by: Peyronnet, Antoine, et al.
Published: (2026)
ARTAI: An Evaluation Platform to Assess Societal Risk of Recommender Algorithms
by: Ruan, Qin, et al.
Published: (2024)
by: Ruan, Qin, et al.
Published: (2024)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
by: Zhang, Liangliang, et al.
Published: (2025)
by: Zhang, Liangliang, et al.
Published: (2025)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
by: Da, Longchao, et al.
Published: (2025)
by: Da, Longchao, et al.
Published: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
by: Amamou, Hazem, et al.
Published: (2026)
by: Amamou, Hazem, et al.
Published: (2026)
When LLM meets Fuzzy-TOPSIS for Personnel Selection through Automated Profile Analysis
by: Hoque, Shahria, et al.
Published: (2026)
by: Hoque, Shahria, et al.
Published: (2026)
Exploring new Approaches for Information Retrieval through Natural Language Processing
by: Raj, Manak, et al.
Published: (2025)
by: Raj, Manak, et al.
Published: (2025)
Complementary Learning Approach for Text Classification using Large Language Models
by: Asgari, Navid, et al.
Published: (2025)
by: Asgari, Navid, et al.
Published: (2025)
Key Algorithms for Keyphrase Generation: Instruction-Based LLMs for Russian Scientific Keyphrases
by: Glazkova, Anna, et al.
Published: (2024)
by: Glazkova, Anna, et al.
Published: (2024)
Beyond Long Context: When Semantics Matter More than Tokens
by: Chawdhury, Tarun Kumar, et al.
Published: (2025)
by: Chawdhury, Tarun Kumar, et al.
Published: (2025)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
CANAL -- Cyber Activity News Alerting Language Model: Empirical Approach vs. Expensive LLM
by: Patel, Urjitkumar, et al.
Published: (2024)
by: Patel, Urjitkumar, et al.
Published: (2024)
Combating data scarcity in recommendation services: Integrating cognitive types of VARK and neural network technologies (LLM)
by: Zmanovskii, Nikita
Published: (2026)
by: Zmanovskii, Nikita
Published: (2026)
HumanMCP: A Human-Like Query Dataset for Evaluating MCP Tool Retrieval Performance
by: Laddha, Shubh, et al.
Published: (2025)
by: Laddha, Shubh, et al.
Published: (2025)
Fact Grounded Attention: Eliminating Hallucination in Large Language Models Through Attention Level Knowledge Integration
by: Gupta, Aayush
Published: (2025)
by: Gupta, Aayush
Published: (2025)
Why Agent Caching Fails and How to Fix It: Structured Intent Canonicalization with Few-Shot Learning
by: Basu, Abhinaba
Published: (2026)
by: Basu, Abhinaba
Published: (2026)
Controllable Evidence Selection in Retrieval-Augmented Question Answering via Deterministic Utility Gating
by: Unda, Victor P.
Published: (2026)
by: Unda, Victor P.
Published: (2026)
CLAP: Coreference-Linked Augmentation for Passage Retrieval
by: Xu, Huanwei, et al.
Published: (2025)
by: Xu, Huanwei, et al.
Published: (2025)
EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
by: Sun, Yuhong, et al.
Published: (2026)
by: Sun, Yuhong, et al.
Published: (2026)
Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding
by: Aljafari, Mohammed, et al.
Published: (2025)
by: Aljafari, Mohammed, et al.
Published: (2025)
AI Agents-as-Judge: Automated Assessment of Accuracy, Consistency, Completeness and Clarity for Enterprise Documents
by: Dasgupta, Sudip, et al.
Published: (2025)
by: Dasgupta, Sudip, et al.
Published: (2025)
Perception-Aware Bias Detection for Query Suggestions
by: Haak, Fabian, et al.
Published: (2026)
by: Haak, Fabian, et al.
Published: (2026)
Adaptive Multi-Stage Patent Claim Generation with Unified Quality Assessment
by: Liang, Chen-Wei, et al.
Published: (2026)
by: Liang, Chen-Wei, et al.
Published: (2026)
AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs
by: Perera, Manoj Madushanka, et al.
Published: (2026)
by: Perera, Manoj Madushanka, et al.
Published: (2026)
PolyTruth: Multilingual Disinformation Detection using Transformer-Based Language Models
by: Gouliev, Zaur, et al.
Published: (2025)
by: Gouliev, Zaur, et al.
Published: (2025)
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
by: Gondhalekar, Chinmay, et al.
Published: (2025)
by: Gondhalekar, Chinmay, et al.
Published: (2025)
Similar Items
-
Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature
by: Maiti, Agniva, et al.
Published: (2025) -
An Integrated Framework of Prompt Engineering and Multidimensional Knowledge Graphs for Legal Dispute Analysis
by: Zhang, Mingda, et al.
Published: (2025) -
Sentiment analysis of texts from social networks based on machine learning methods for monitoring public sentiment
by: Nurlanuly, Arsen Tolebay
Published: (2025) -
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
by: Nguyen, Andy, et al.
Published: (2026) -
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
by: Kumar, Aayush
Published: (2025)