Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability
Fuente:
arXiv
Saved in:
| Main Authors: | B, Gautam, Purwar, Anupam |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-augmented Retrieval: A Novel Framework for Fast Information Retrieval based Response Generation using Large Language Model
by: Ganesh, Sai, et al.
Published: (2024)
by: Ganesh, Sai, et al.
Published: (2024)
Evaluating Factual Density in Multi-Source RAG: A Study in Medical AI Accuracy
by: DeMarco, Michael R.
Published: (2026)
by: DeMarco, Michael R.
Published: (2026)
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
by: Zhu, Kunlun, et al.
Published: (2024)
by: Zhu, Kunlun, et al.
Published: (2024)
A Tale of Trust and Accuracy: Base vs. Instruct LLMs in RAG Systems
by: Cuconasu, Florin, et al.
Published: (2024)
by: Cuconasu, Florin, et al.
Published: (2024)
OpenRAG: Optimizing RAG End-to-End via In-Context Retrieval Learning
by: Zhou, Jiawei, et al.
Published: (2025)
by: Zhou, Jiawei, et al.
Published: (2025)
Can LLMs Outshine Conventional Recommenders? A Comparative Evaluation
by: Liu, Qijiong, et al.
Published: (2025)
by: Liu, Qijiong, et al.
Published: (2025)
Towards AI Evaluation in Domain-Specific RAG Systems: The AgriHubi Case Study
by: Hasan, Md. Toufique, et al.
Published: (2026)
by: Hasan, Md. Toufique, et al.
Published: (2026)
An Open-Source Web-Based Tool for Evaluating Open-Source Large Language Models Leveraging Information Retrieval from Custom Documents
by: I, Godfrey
Published: (2025)
by: I, Godfrey
Published: (2025)
Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework
by: Buskila, Avi-ad Avraam
Published: (2026)
by: Buskila, Avi-ad Avraam
Published: (2026)
RAG-IGBench: Innovative Evaluation for RAG-based Interleaved Generation in Open-domain Question Answering
by: Zhang, Rongyang, et al.
Published: (2025)
by: Zhang, Rongyang, et al.
Published: (2025)
Evaluating the Performance of LLMs on Technical Language Processing tasks
by: Kernycky, Andrew, et al.
Published: (2024)
by: Kernycky, Andrew, et al.
Published: (2024)
DistRAG: Towards Distance-Based Spatial Reasoning in LLMs
by: Schneider, Nicole R, et al.
Published: (2025)
by: Schneider, Nicole R, et al.
Published: (2025)
ARIA: Adaptive Retrieval Intelligence Assistant -- A Multimodal RAG Framework for Domain-Specific Engineering Education
by: Luo, Yue, et al.
Published: (2026)
by: Luo, Yue, et al.
Published: (2026)
Domain-Aware RAG: MoL-Enhanced RL for Efficient Training and Scalable Retrieval
by: Lin, Hao, et al.
Published: (2025)
by: Lin, Hao, et al.
Published: (2025)
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs
by: Ferraretto, Fernando, et al.
Published: (2024)
by: Ferraretto, Fernando, et al.
Published: (2024)
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Emotional RAG LLMs: Reading Comprehension for the Open Internet
by: Reichman, Benjamin, et al.
Published: (2024)
by: Reichman, Benjamin, et al.
Published: (2024)
Comparing Knowledge Sources for Open-Domain Scientific Claim Verification
by: Vladika, Juraj, et al.
Published: (2024)
by: Vladika, Juraj, et al.
Published: (2024)
UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG
by: Georgiev, Dobrik, et al.
Published: (2026)
by: Georgiev, Dobrik, et al.
Published: (2026)
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana
by: Filice, Simone, et al.
Published: (2025)
by: Filice, Simone, et al.
Published: (2025)
The Power of Noise: Redefining Retrieval for RAG Systems
by: Cuconasu, Florin, et al.
Published: (2024)
by: Cuconasu, Florin, et al.
Published: (2024)
A Hybrid RAG System with Comprehensive Enhancement on Complex Reasoning
by: Yuan, Ye, et al.
Published: (2024)
by: Yuan, Ye, et al.
Published: (2024)
Evaluating RAG-Fusion with RAGElo: an Automated Elo-based Framework
by: Rackauckas, Zackary, et al.
Published: (2024)
by: Rackauckas, Zackary, et al.
Published: (2024)
ClaimTrust: Propagation Trust Scoring for RAG Systems
by: Qian, Hangkai, et al.
Published: (2025)
by: Qian, Hangkai, et al.
Published: (2025)
FunnelRAG: A Coarse-to-Fine Progressive Retrieval Paradigm for RAG
by: Zhao, Xinping, et al.
Published: (2024)
by: Zhao, Xinping, et al.
Published: (2024)
Political Events using RAG with LLMs
by: Arslan, Muhammad, et al.
Published: (2025)
by: Arslan, Muhammad, et al.
Published: (2025)
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources
by: Xia, Yikuan, et al.
Published: (2025)
by: Xia, Yikuan, et al.
Published: (2025)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
by: Thakur, Nandan, et al.
Published: (2025)
by: Thakur, Nandan, et al.
Published: (2025)
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack
by: Gao, Yunfan, et al.
Published: (2025)
by: Gao, Yunfan, et al.
Published: (2025)
Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG
by: Bachyr, Omar El, et al.
Published: (2026)
by: Bachyr, Omar El, et al.
Published: (2026)
Rethinking the Evaluation of Dialogue Systems: Effects of User Feedback on Crowdworkers and LLMs
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
RAG based Question-Answering for Contextual Response Prediction System
by: Veturi, Sriram, et al.
Published: (2024)
by: Veturi, Sriram, et al.
Published: (2024)
Do RAG Systems Really Suffer From Positional Bias?
by: Cuconasu, Florin, et al.
Published: (2025)
by: Cuconasu, Florin, et al.
Published: (2025)
How Significant Are the Real Performance Gains? An Unbiased Evaluation Framework for GraphRAG
by: Zeng, Qiming, et al.
Published: (2025)
by: Zeng, Qiming, et al.
Published: (2025)
COS-Mix: Cosine Similarity and Distance Fusion for Improved Information Retrieval
by: Juvekar, Kush, et al.
Published: (2024)
by: Juvekar, Kush, et al.
Published: (2024)
Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG
by: Sun, Yiqun, et al.
Published: (2026)
by: Sun, Yiqun, et al.
Published: (2026)
LiveRAG: A diverse Q&A dataset with varying difficulty level for RAG evaluation
by: Carmel, David, et al.
Published: (2025)
by: Carmel, David, et al.
Published: (2025)
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
by: Khadilkar, Harshad, et al.
Published: (2025)
by: Khadilkar, Harshad, et al.
Published: (2025)
Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG
by: Li, Yubo, et al.
Published: (2026)
by: Li, Yubo, et al.
Published: (2026)
Similar Items
-
Context-augmented Retrieval: A Novel Framework for Fast Information Retrieval based Response Generation using Large Language Model
by: Ganesh, Sai, et al.
Published: (2024) -
Evaluating Factual Density in Multi-Source RAG: A Study in Medical AI Accuracy
by: DeMarco, Michael R.
Published: (2026) -
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
by: Zhu, Kunlun, et al.
Published: (2024) -
A Tale of Trust and Accuracy: Base vs. Instruct LLMs in RAG Systems
by: Cuconasu, Florin, et al.
Published: (2024) -
OpenRAG: Optimizing RAG End-to-End via In-Context Retrieval Learning
by: Zhou, Jiawei, et al.
Published: (2025)