Evaluating Multi-Hop Reasoning in RAG Systems: A Comparison of LLM-Based Retriever Evaluation Strategies
Fuente:
arXiv
Saved in:
| Main Authors: | Brehme, Lorenz, Ströhle, Thomas, Breu, Ruth |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Be Trusted for Evaluating RAG Systems? A Survey of Methods and Datasets
by: Brehme, Lorenz, et al.
Published: (2025)
by: Brehme, Lorenz, et al.
Published: (2025)
Retrieval-Augmented Generation in Industry: An Interview Study on Use Cases, Requirements, Challenges, and Evaluation
by: Brehme, Lorenz, et al.
Published: (2025)
by: Brehme, Lorenz, et al.
Published: (2025)
RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation
by: Brehme, Lorenz, et al.
Published: (2026)
by: Brehme, Lorenz, et al.
Published: (2026)
StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems
by: Patodiya, Aryan
Published: (2026)
by: Patodiya, Aryan
Published: (2026)
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
by: Islamaj, Rezarta, et al.
Published: (2026)
by: Islamaj, Rezarta, et al.
Published: (2026)
MHTS: Multi-Hop Tree Structure Framework for Generating Difficulty-Controllable QA Datasets for RAG Evaluation
by: Lee, Jeongsoo, et al.
Published: (2025)
by: Lee, Jeongsoo, et al.
Published: (2025)
Agentic RAG with Knowledge Graphs for Complex Multi-Hop Reasoning in Real-World Applications
by: Lelong, Jean, et al.
Published: (2025)
by: Lelong, Jean, et al.
Published: (2025)
OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
by: Liang, Jiaxuan, et al.
Published: (2025)
by: Liang, Jiaxuan, et al.
Published: (2025)
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems
by: Papadimitriou, Ioannis, et al.
Published: (2024)
by: Papadimitriou, Ioannis, et al.
Published: (2024)
ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
OPERA: A Reinforcement Learning--Enhanced Orchestrated Planner-Executor Architecture for Reasoning-Oriented Multi-Hop Retrieval
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation
by: Bolognesi, Giorgia, et al.
Published: (2026)
by: Bolognesi, Giorgia, et al.
Published: (2026)
A Multi-Source Retrieval Question Answering Framework Based on RAG
by: Wu, Ridong, et al.
Published: (2024)
by: Wu, Ridong, et al.
Published: (2024)
Reasoning RAG via System 1 or System 2: A Survey on Reasoning Agentic Retrieval-Augmented Generation for Industry Challenges
by: Liang, Jintao, et al.
Published: (2025)
by: Liang, Jintao, et al.
Published: (2025)
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective
by: Packowski, Sarah, et al.
Published: (2024)
by: Packowski, Sarah, et al.
Published: (2024)
LiteSemRAG: Lightweight LLM-Free Semantic-Aware Graph Retrieval for Robust RAG
by: Yue, Xiao, et al.
Published: (2026)
by: Yue, Xiao, et al.
Published: (2026)
InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning
by: Wang, Zheng, et al.
Published: (2025)
by: Wang, Zheng, et al.
Published: (2025)
S-Path-RAG: Semantic-Aware Shortest-Path Retrieval Augmented Generation for Multi-Hop Knowledge Graph Question Answering
by: Fu, Rong, et al.
Published: (2026)
by: Fu, Rong, et al.
Published: (2026)
Worse than Zero-shot? A Fact-Checking Dataset for Evaluating the Robustness of RAG Against Misleading Retrievals
by: Zeng, Linda, et al.
Published: (2025)
by: Zeng, Linda, et al.
Published: (2025)
AgenticRAG: Agentic Retrieval for Enterprise Knowledge Bases
by: Suresh, Susheel, et al.
Published: (2026)
by: Suresh, Susheel, et al.
Published: (2026)
OpenTCM: A GraphRAG-Empowered LLM-based System for Traditional Chinese Medicine Knowledge Retrieval and Diagnosis
by: He, Jinglin, et al.
Published: (2025)
by: He, Jinglin, et al.
Published: (2025)
DTKG: Dual-Track Knowledge Graph-Verified Reasoning Framework for Multi-Hop QA
by: Wang, Changhao, et al.
Published: (2025)
by: Wang, Changhao, et al.
Published: (2025)
BioRAG: A RAG-LLM Framework for Biological Question Reasoning
by: Wang, Chengrui, et al.
Published: (2024)
by: Wang, Chengrui, et al.
Published: (2024)
Progressive Searching for Retrieval in RAG
by: Jeong, Taehee, et al.
Published: (2026)
by: Jeong, Taehee, et al.
Published: (2026)
Hierarchical Lexical Graph for Enhanced Multi-Hop Retrieval
by: Ghassel, Abdellah, et al.
Published: (2025)
by: Ghassel, Abdellah, et al.
Published: (2025)
Resources for Automated Evaluation of Assistive RAG Systems that Help Readers with News Trustworthiness Assessment
by: Zhang, Dake, et al.
Published: (2026)
by: Zhang, Dake, et al.
Published: (2026)
Evaluating Chunking Strategies For Retrieval-Augmented Generation in Oil and Gas Enterprise Documents
by: Taiwo, Samuel, et al.
Published: (2026)
by: Taiwo, Samuel, et al.
Published: (2026)
Hyper-RAG: Combating LLM Hallucinations using Hypergraph-Driven Retrieval-Augmented Generation
by: Feng, Yifan, et al.
Published: (2025)
by: Feng, Yifan, et al.
Published: (2025)
Modernizing Facebook Scoped Search: Keyword and Embedding Hybrid Retrieval with LLM Evaluation
by: Su, Yongye, et al.
Published: (2025)
by: Su, Yongye, et al.
Published: (2025)
MG$^2$-RAG: Multi-Granularity Graph for Multimodal Retrieval-Augmented Generation
by: Dai, Sijun, et al.
Published: (2026)
by: Dai, Sijun, et al.
Published: (2026)
RAGtifier: Evaluating RAG Generation Approaches of State-of-the-Art RAG Systems for the SIGIR LiveRAG Competition
by: Cofala, Tim, et al.
Published: (2025)
by: Cofala, Tim, et al.
Published: (2025)
Enhancing Technical Documents Retrieval for RAG
by: Lai, Songjiang, et al.
Published: (2025)
by: Lai, Songjiang, et al.
Published: (2025)
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations
by: Lin, Jiaen, et al.
Published: (2025)
by: Lin, Jiaen, et al.
Published: (2025)
PRISM: Agentic Retrieval with LLMs for Multi-Hop Question Answering
by: Nahid, Md Mahadi Hasan, et al.
Published: (2025)
by: Nahid, Md Mahadi Hasan, et al.
Published: (2025)
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
by: Kashmira, Savini, et al.
Published: (2024)
by: Kashmira, Savini, et al.
Published: (2024)
Pseudo-Knowledge Graph: Meta-Path Guided Retrieval and In-Graph Text for RAG-Equipped LLM
by: Yang, Yuxin, et al.
Published: (2025)
by: Yang, Yuxin, et al.
Published: (2025)
HypRAG: Hyperbolic Dense Retrieval for Retrieval Augmented Generation
by: Madhu, Hiren, et al.
Published: (2026)
by: Madhu, Hiren, et al.
Published: (2026)
A2RAG: Adaptive Agentic Graph Retrieval for Cost-Aware and Reliable Reasoning
by: Liu, Jiate, et al.
Published: (2026)
by: Liu, Jiate, et al.
Published: (2026)
RAGXplain: From Explainable Evaluation to Actionable Guidance of RAG Pipelines
by: Cohen, Dvir, et al.
Published: (2025)
by: Cohen, Dvir, et al.
Published: (2025)
Clue-RAG: Towards Accurate and Cost-Efficient Graph-based RAG via Multi-Partite Graph and Query-Driven Iterative Retrieval
by: Su, Yaodong, et al.
Published: (2025)
by: Su, Yaodong, et al.
Published: (2025)
Similar Items
-
Can LLMs Be Trusted for Evaluating RAG Systems? A Survey of Methods and Datasets
by: Brehme, Lorenz, et al.
Published: (2025) -
Retrieval-Augmented Generation in Industry: An Interview Study on Use Cases, Requirements, Challenges, and Evaluation
by: Brehme, Lorenz, et al.
Published: (2025) -
RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation
by: Brehme, Lorenz, et al.
Published: (2026) -
StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems
by: Patodiya, Aryan
Published: (2026) -
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
by: Islamaj, Rezarta, et al.
Published: (2026)