RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Friel, Robert, Belyi, Masha, Sanyal, Atindriyo |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost
par: Belyi, Masha, et autres
Publié: (2024)
par: Belyi, Masha, et autres
Publié: (2024)
Benchmarking Retrieval-Augmented Generation for Medicine
par: Xiong, Guangzhi, et autres
Publié: (2024)
par: Xiong, Guangzhi, et autres
Publié: (2024)
LIT-RAGBench: Benchmarking Generator Capabilities of Large Language Models in Retrieval-Augmented Generation
par: Itai, Koki, et autres
Publié: (2026)
par: Itai, Koki, et autres
Publié: (2026)
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems
par: Thakur, Nandan, et autres
Publié: (2024)
par: Thakur, Nandan, et autres
Publié: (2024)
MRAG: Benchmarking Retrieval-Augmented Generation for Bio-medicine
par: Li, Liz, et autres
Publié: (2026)
par: Li, Liz, et autres
Publié: (2026)
MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems
par: Katsis, Yannis, et autres
Publié: (2025)
par: Katsis, Yannis, et autres
Publié: (2025)
Explainable Biomedical Hypothesis Generation via Retrieval Augmented Generation enabled Large Language Models
par: Pelletier, Alexander R., et autres
Publié: (2024)
par: Pelletier, Alexander R., et autres
Publié: (2024)
Benchmarking Retrieval-Augmented Generation for Chemistry
par: Zhong, Xianrui, et autres
Publié: (2025)
par: Zhong, Xianrui, et autres
Publié: (2025)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
par: Park, Chanhee, et autres
Publié: (2025)
par: Park, Chanhee, et autres
Publié: (2025)
Open-Source Reproduction and Explainability Analysis of Corrective Retrieval Augmented Generation
par: Yalavarthi, Surya Vardhan
Publié: (2026)
par: Yalavarthi, Surya Vardhan
Publié: (2026)
Towards Global Retrieval Augmented Generation: A Benchmark for Corpus-Level Reasoning
par: Luo, Qi, et autres
Publié: (2025)
par: Luo, Qi, et autres
Publié: (2025)
CanLegalRAGBench: Evaluating Retrieval-Augmented Generation on Canadian Case Law
par: Zhao, Ethan, et autres
Publié: (2026)
par: Zhao, Ethan, et autres
Publié: (2026)
XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation
par: Mao, Qianren, et autres
Publié: (2024)
par: Mao, Qianren, et autres
Publié: (2024)
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs
par: Omar, Reham, et autres
Publié: (2025)
par: Omar, Reham, et autres
Publié: (2025)
SciRerankBench: Benchmarking Rerankers Towards Scientific Retrieval-Augmented Generated LLMs
par: Chen, Haotian, et autres
Publié: (2025)
par: Chen, Haotian, et autres
Publié: (2025)
Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving
par: Zheng, Shunfeng, et autres
Publié: (2025)
par: Zheng, Shunfeng, et autres
Publié: (2025)
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation
par: Blandón, María Andrea Cruz, et autres
Publié: (2025)
par: Blandón, María Andrea Cruz, et autres
Publié: (2025)
On the Influence of Context Size and Model Choice in Retrieval-Augmented Generation Systems
par: Vladika, Juraj, et autres
Publié: (2025)
par: Vladika, Juraj, et autres
Publié: (2025)
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation
par: Ouyang, Jie, et autres
Publié: (2025)
par: Ouyang, Jie, et autres
Publié: (2025)
Semantic Tokens in Retrieval Augmented Generation
par: Suro, Joel
Publié: (2024)
par: Suro, Joel
Publié: (2024)
Clustered Retrieved Augmented Generation (CRAG)
par: Akesson, Simon, et autres
Publié: (2024)
par: Akesson, Simon, et autres
Publié: (2024)
Exploring Retrieval Augmented Generation in Arabic
par: El-Beltagy, Samhaa R., et autres
Publié: (2024)
par: El-Beltagy, Samhaa R., et autres
Publié: (2024)
Predictive Prefetching for Retrieval-Augmented Generation
par: Zhang, Wuyang, et autres
Publié: (2026)
par: Zhang, Wuyang, et autres
Publié: (2026)
Latent Abstraction for Retrieval-Augmented Generation
par: T, Ha Lan N., et autres
Publié: (2026)
par: T, Ha Lan N., et autres
Publié: (2026)
Retrieval-Augmented Generation with Hierarchical Knowledge
par: Huang, Haoyu, et autres
Publié: (2025)
par: Huang, Haoyu, et autres
Publié: (2025)
Retrieval-Augmented Generation with Conflicting Evidence
par: Wang, Han, et autres
Publié: (2025)
par: Wang, Han, et autres
Publié: (2025)
Aligning Extraction and Generation for Robust Retrieval-Augmented Generation
par: Song, Hwanjun, et autres
Publié: (2025)
par: Song, Hwanjun, et autres
Publié: (2025)
Improving Reliability and Explainability of Medical Question Answering through Atomic Fact Checking in Retrieval-Augmented LLMs
par: Vladika, Juraj, et autres
Publié: (2025)
par: Vladika, Juraj, et autres
Publié: (2025)
RAGViz: Diagnose and Visualize Retrieval-Augmented Generation
par: Wang, Tevin, et autres
Publié: (2024)
par: Wang, Tevin, et autres
Publié: (2024)
Retrieval-Augmented Generation-based Relation Extraction
par: Efeoglu, Sefika, et autres
Publié: (2024)
par: Efeoglu, Sefika, et autres
Publié: (2024)
DuetRAG: Collaborative Retrieval-Augmented Generation
par: Jiao, Dian, et autres
Publié: (2024)
par: Jiao, Dian, et autres
Publié: (2024)
Evaluation of Retrieval-Augmented Generation: A Survey
par: Yu, Hao, et autres
Publié: (2024)
par: Yu, Hao, et autres
Publié: (2024)
Knowledge Graph-Guided Retrieval Augmented Generation
par: Zhu, Xiangrong, et autres
Publié: (2025)
par: Zhu, Xiangrong, et autres
Publié: (2025)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
par: Zhou, Yujia, et autres
Publié: (2024)
par: Zhou, Yujia, et autres
Publié: (2024)
LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs -- No Silver Bullet for LC or RAG Routing
par: Li, Kuan, et autres
Publié: (2025)
par: Li, Kuan, et autres
Publié: (2025)
Towards Automated Smart Contract Generation: Evaluation, Benchmarking, and Retrieval-Augmented Repair
par: Chen, Zaoyu, et autres
Publié: (2025)
par: Chen, Zaoyu, et autres
Publié: (2025)
WeatherArchive-Bench: Benchmarking Retrieval-Augmented Reasoning for Historical Weather Archives
par: Yu, Yongan, et autres
Publié: (2025)
par: Yu, Yongan, et autres
Publié: (2025)
Retrieving, Rethinking and Revising: The Chain-of-Verification Can Improve Retrieval Augmented Generation
par: He, Bolei, et autres
Publié: (2024)
par: He, Bolei, et autres
Publié: (2024)
LLM-Confidence Reranker: A Training-Free Approach for Enhancing Retrieval-Augmented Generation Systems
par: Song, Zhipeng, et autres
Publié: (2026)
par: Song, Zhipeng, et autres
Publié: (2026)
Retrieval-Augmented Generation Systems for Intellectual Property via Synthetic Multi-Angle Fine-tuning
par: Ren, Runtao, et autres
Publié: (2025)
par: Ren, Runtao, et autres
Publié: (2025)
Documents similaires
-
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost
par: Belyi, Masha, et autres
Publié: (2024) -
Benchmarking Retrieval-Augmented Generation for Medicine
par: Xiong, Guangzhi, et autres
Publié: (2024) -
LIT-RAGBench: Benchmarking Generator Capabilities of Large Language Models in Retrieval-Augmented Generation
par: Itai, Koki, et autres
Publié: (2026) -
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems
par: Thakur, Nandan, et autres
Publié: (2024) -
MRAG: Benchmarking Retrieval-Augmented Generation for Bio-medicine
par: Li, Liz, et autres
Publié: (2026)