Evaluating RAG-Fusion with RAGElo: an Automated Elo-based Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Rackauckas, Zackary, Câmara, Arthur, Zavrel, Jakub |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAG-Fusion: a New Take on Retrieval-Augmented Generation
by: Rackauckas, Zackary
Published: (2024)
by: Rackauckas, Zackary
Published: (2024)
Self-Optimizing Multi-Agent Systems for Deep Research
by: Câmara, Arthur, et al.
Published: (2026)
by: Câmara, Arthur, et al.
Published: (2026)
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
by: Rackauckas, Zackary, et al.
Published: (2025)
by: Rackauckas, Zackary, et al.
Published: (2025)
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
by: Zhu, Kunlun, et al.
Published: (2024)
by: Zhu, Kunlun, et al.
Published: (2024)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2025)
by: Pradeep, Ronak, et al.
Published: (2025)
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework
by: Pradeep, Ronak, et al.
Published: (2024)
by: Pradeep, Ronak, et al.
Published: (2024)
In-depth Analysis of Graph-based RAG in a Unified Framework
by: Zhou, Yingli, et al.
Published: (2025)
by: Zhou, Yingli, et al.
Published: (2025)
Core-based Hierarchies for Efficient GraphRAG
by: Hossain, Jakir, et al.
Published: (2026)
by: Hossain, Jakir, et al.
Published: (2026)
RAG-IGBench: Innovative Evaluation for RAG-based Interleaved Generation in Open-domain Question Answering
by: Zhang, Rongyang, et al.
Published: (2025)
by: Zhang, Rongyang, et al.
Published: (2025)
FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing
by: Qian, Jingbin, et al.
Published: (2026)
by: Qian, Jingbin, et al.
Published: (2026)
FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation
by: Zhang, Zhuocheng, et al.
Published: (2025)
by: Zhang, Zhuocheng, et al.
Published: (2025)
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana
by: Filice, Simone, et al.
Published: (2025)
by: Filice, Simone, et al.
Published: (2025)
Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG
by: Bachyr, Omar El, et al.
Published: (2026)
by: Bachyr, Omar El, et al.
Published: (2026)
RAG-based Question Answering over Heterogeneous Data and Text
by: Christmann, Philipp, et al.
Published: (2024)
by: Christmann, Philipp, et al.
Published: (2024)
RAG based Question-Answering for Contextual Response Prediction System
by: Veturi, Sriram, et al.
Published: (2024)
by: Veturi, Sriram, et al.
Published: (2024)
TableRAG: A Retrieval Augmented Generation Framework for Heterogeneous Document Reasoning
by: Yu, Xiaohan, et al.
Published: (2025)
by: Yu, Xiaohan, et al.
Published: (2025)
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
by: Khadilkar, Harshad, et al.
Published: (2025)
by: Khadilkar, Harshad, et al.
Published: (2025)
How Significant Are the Real Performance Gains? An Unbiased Evaluation Framework for GraphRAG
by: Zeng, Qiming, et al.
Published: (2025)
by: Zeng, Qiming, et al.
Published: (2025)
FIT-RAG: Black-Box RAG with Factual Information and Token Reduction
by: Mao, Yuren, et al.
Published: (2024)
by: Mao, Yuren, et al.
Published: (2024)
FunnelRAG: A Coarse-to-Fine Progressive Retrieval Paradigm for RAG
by: Zhao, Xinping, et al.
Published: (2024)
by: Zhao, Xinping, et al.
Published: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
Evaluating Factual Density in Multi-Source RAG: A Study in Medical AI Accuracy
by: DeMarco, Michael R.
Published: (2026)
by: DeMarco, Michael R.
Published: (2026)
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack
by: Gao, Yunfan, et al.
Published: (2025)
by: Gao, Yunfan, et al.
Published: (2025)
BioRAG: A RAG-LLM Framework for Biological Question Reasoning
by: Wang, Chengrui, et al.
Published: (2024)
by: Wang, Chengrui, et al.
Published: (2024)
Modular RAG: Transforming RAG Systems into LEGO-like Reconfigurable Frameworks
by: Gao, Yunfan, et al.
Published: (2024)
by: Gao, Yunfan, et al.
Published: (2024)
MedCoT-RAG: Causal Chain-of-Thought RAG for Medical Question Answering
by: Wang, Ziyu, et al.
Published: (2025)
by: Wang, Ziyu, et al.
Published: (2025)
OpenRAG: Optimizing RAG End-to-End via In-Context Retrieval Learning
by: Zhou, Jiawei, et al.
Published: (2025)
by: Zhou, Jiawei, et al.
Published: (2025)
Evaluating Hybrid Retrieval Augmented Generation using Dynamic Test Sets: LiveRAG Challenge
by: Fensore, Chase, et al.
Published: (2025)
by: Fensore, Chase, et al.
Published: (2025)
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
MultiRAG: A Knowledge-guided Framework for Mitigating Hallucination in Multi-source Retrieval Augmented Generation
by: Wu, Wenlong, et al.
Published: (2025)
by: Wu, Wenlong, et al.
Published: (2025)
ARIA: Adaptive Retrieval Intelligence Assistant -- A Multimodal RAG Framework for Domain-Specific Engineering Education
by: Luo, Yue, et al.
Published: (2026)
by: Luo, Yue, et al.
Published: (2026)
LiveRAG: A diverse Q&A dataset with varying difficulty level for RAG evaluation
by: Carmel, David, et al.
Published: (2025)
by: Carmel, David, et al.
Published: (2025)
Scaling Retrieval Augmented Generation with RAG Fusion: Lessons from an Industry Deployment
by: Medrano, Luigi, et al.
Published: (2026)
by: Medrano, Luigi, et al.
Published: (2026)
MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG
by: Wang, Xihang, et al.
Published: (2026)
by: Wang, Xihang, et al.
Published: (2026)
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability
by: B, Gautam, et al.
Published: (2024)
by: B, Gautam, et al.
Published: (2024)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023)
by: Saad-Falcon, Jon, et al.
Published: (2023)
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems
by: Papadimitriou, Ioannis, et al.
Published: (2024)
by: Papadimitriou, Ioannis, et al.
Published: (2024)
H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations
by: Elchafei, Passant, et al.
Published: (2026)
by: Elchafei, Passant, et al.
Published: (2026)
Context Embeddings for Efficient Answer Generation in RAG
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Similar Items
-
RAG-Fusion: a New Take on Retrieval-Augmented Generation
by: Rackauckas, Zackary
Published: (2024) -
Self-Optimizing Multi-Agent Systems for Deep Research
by: Câmara, Arthur, et al.
Published: (2026) -
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
by: Rackauckas, Zackary, et al.
Published: (2025) -
RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
by: Zhu, Kunlun, et al.
Published: (2024) -
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2025)