Resources for Automated Evaluation of Assistive RAG Systems that Help Readers with News Trustworthiness Assessment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Dake, Smucker, Mark D., Clarke, Charles L. A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Extending MovieLens-32M to Provide New Evaluation Objectives
von: Smucker, Mark D., et al.
Veröffentlicht: (2025)
von: Smucker, Mark D., et al.
Veröffentlicht: (2025)
Can LLMs Be Trusted for Evaluating RAG Systems? A Survey of Methods and Datasets
von: Brehme, Lorenz, et al.
Veröffentlicht: (2025)
von: Brehme, Lorenz, et al.
Veröffentlicht: (2025)
Evaluating Multi-Hop Reasoning in RAG Systems: A Comparison of LLM-Based Retriever Evaluation Strategies
von: Brehme, Lorenz, et al.
Veröffentlicht: (2026)
von: Brehme, Lorenz, et al.
Veröffentlicht: (2026)
Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems
von: Fan, Dongzhe, et al.
Veröffentlicht: (2026)
von: Fan, Dongzhe, et al.
Veröffentlicht: (2026)
DMQR-RAG: Diverse Multi-Query Rewriting for RAG
von: Li, Zhicong, et al.
Veröffentlicht: (2024)
von: Li, Zhicong, et al.
Veröffentlicht: (2024)
RAGtifier: Evaluating RAG Generation Approaches of State-of-the-Art RAG Systems for the SIGIR LiveRAG Competition
von: Cofala, Tim, et al.
Veröffentlicht: (2025)
von: Cofala, Tim, et al.
Veröffentlicht: (2025)
StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems
von: Patodiya, Aryan
Veröffentlicht: (2026)
von: Patodiya, Aryan
Veröffentlicht: (2026)
R3A: Reinforced Reasoning for Relevance Assessment for RAG in User-Generated Content Platforms
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2025)
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2025)
Mobile-Agent-RAG: Driving Smart Multi-Agent Coordination with Contextual Knowledge Empowerment for Long-Horizon Mobile Automation
von: Zhou, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxiang, et al.
Veröffentlicht: (2025)
M-RAG: Making RAG Faster, Stronger, and More Efficient
von: Xu, Sun, et al.
Veröffentlicht: (2026)
von: Xu, Sun, et al.
Veröffentlicht: (2026)
RAGXplain: From Explainable Evaluation to Actionable Guidance of RAG Pipelines
von: Cohen, Dvir, et al.
Veröffentlicht: (2025)
von: Cohen, Dvir, et al.
Veröffentlicht: (2025)
Dynamic Evaluation Framework for Personalized and Trustworthy Agents: A Multi-Session Approach to Preference Adaptability
von: Shah, Chirag, et al.
Veröffentlicht: (2025)
von: Shah, Chirag, et al.
Veröffentlicht: (2025)
Evaluating Ensemble Methods for News Recommender Systems
von: Gray, Alexander, et al.
Veröffentlicht: (2024)
von: Gray, Alexander, et al.
Veröffentlicht: (2024)
Automated Interpretation of Non-Destructive Evaluation Contour Maps Using Large Language Models for Bridge Condition Assessment
von: Darji, Viraj Nishesh, et al.
Veröffentlicht: (2025)
von: Darji, Viraj Nishesh, et al.
Veröffentlicht: (2025)
AMAQA: A Metadata-based QA Dataset for RAG Systems
von: Bruni, Davide, et al.
Veröffentlicht: (2025)
von: Bruni, Davide, et al.
Veröffentlicht: (2025)
Leveraging Evidence-Guided LLMs to Enhance Trustworthy Depression Diagnosis
von: Yuan, Yining, et al.
Veröffentlicht: (2025)
von: Yuan, Yining, et al.
Veröffentlicht: (2025)
SearchRAG: Can Search Engines Be Helpful for LLM-based Medical Question Answering?
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
HugRAG: Hierarchical Causal Knowledge Graph Design for RAG
von: Wang, Nengbo, et al.
Veröffentlicht: (2026)
von: Wang, Nengbo, et al.
Veröffentlicht: (2026)
Worse than Zero-shot? A Fact-Checking Dataset for Evaluating the Robustness of RAG Against Misleading Retrievals
von: Zeng, Linda, et al.
Veröffentlicht: (2025)
von: Zeng, Linda, et al.
Veröffentlicht: (2025)
HF-RAG: Hierarchical Fusion-based RAG with Multiple Sources and Rankers
von: Santra, Payel, et al.
Veröffentlicht: (2025)
von: Santra, Payel, et al.
Veröffentlicht: (2025)
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective
von: Packowski, Sarah, et al.
Veröffentlicht: (2024)
von: Packowski, Sarah, et al.
Veröffentlicht: (2024)
LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation
von: Bolognesi, Giorgia, et al.
Veröffentlicht: (2026)
von: Bolognesi, Giorgia, et al.
Veröffentlicht: (2026)
Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation
von: Klearman, Andrew, et al.
Veröffentlicht: (2026)
von: Klearman, Andrew, et al.
Veröffentlicht: (2026)
OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
von: Liang, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Liang, Jiaxuan, et al.
Veröffentlicht: (2025)
Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth
von: Sahu, Gaurav, et al.
Veröffentlicht: (2026)
von: Sahu, Gaurav, et al.
Veröffentlicht: (2026)
AgenticRAG: Tool-Augmented Foundation Models for Zero-Shot Explainable Recommender Systems
von: Ma, Bo, et al.
Veröffentlicht: (2025)
von: Ma, Bo, et al.
Veröffentlicht: (2025)
EcphoryRAG: Re-Imagining Knowledge-Graph RAG via Human Associative Memory
von: Liao, Zirui
Veröffentlicht: (2025)
von: Liao, Zirui
Veröffentlicht: (2025)
Securing RAG: A Risk Assessment and Mitigation Framework
von: Ammann, Lukas, et al.
Veröffentlicht: (2025)
von: Ammann, Lukas, et al.
Veröffentlicht: (2025)
RAG-IGBench: Innovative Evaluation for RAG-based Interleaved Generation in Open-domain Question Answering
von: Zhang, Rongyang, et al.
Veröffentlicht: (2025)
von: Zhang, Rongyang, et al.
Veröffentlicht: (2025)
Graphs RAG at Scale: Beyond Retrieval-Augmented Generation With Labeled Property Graphs and Resource Description Framework for Complex and Unknown Search Spaces
von: Tadayon, Manie, et al.
Veröffentlicht: (2026)
von: Tadayon, Manie, et al.
Veröffentlicht: (2026)
Progressive Searching for Retrieval in RAG
von: Jeong, Taehee, et al.
Veröffentlicht: (2026)
von: Jeong, Taehee, et al.
Veröffentlicht: (2026)
MHTS: Multi-Hop Tree Structure Framework for Generating Difficulty-Controllable QA Datasets for RAG Evaluation
von: Lee, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Lee, Jeongsoo, et al.
Veröffentlicht: (2025)
MACA: A Framework for Distilling Trustworthy LLMs into Efficient Retrievers
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2026)
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2026)
Towards Trustworthy LLM-Based Recommendation via Rationale Integration
von: Park, Chung, et al.
Veröffentlicht: (2025)
von: Park, Chung, et al.
Veröffentlicht: (2025)
LiteSemRAG: Lightweight LLM-Free Semantic-Aware Graph Retrieval for Robust RAG
von: Yue, Xiao, et al.
Veröffentlicht: (2026)
von: Yue, Xiao, et al.
Veröffentlicht: (2026)
TrustRAG: An Information Assistant with Retrieval Augmented Generation
von: Fan, Yixing, et al.
Veröffentlicht: (2025)
von: Fan, Yixing, et al.
Veröffentlicht: (2025)
Modular RAG: Transforming RAG Systems into LEGO-like Reconfigurable Frameworks
von: Gao, Yunfan, et al.
Veröffentlicht: (2024)
von: Gao, Yunfan, et al.
Veröffentlicht: (2024)
Tree-based RAG-Agent Recommendation System: A Case Study in Medical Test Data
von: Yang, Yahe, et al.
Veröffentlicht: (2025)
von: Yang, Yahe, et al.
Veröffentlicht: (2025)
LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots
von: Sun, Haoran, et al.
Veröffentlicht: (2026)
von: Sun, Haoran, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Extending MovieLens-32M to Provide New Evaluation Objectives
von: Smucker, Mark D., et al.
Veröffentlicht: (2025) -
Can LLMs Be Trusted for Evaluating RAG Systems? A Survey of Methods and Datasets
von: Brehme, Lorenz, et al.
Veröffentlicht: (2025) -
Evaluating Multi-Hop Reasoning in RAG Systems: A Comparison of LLM-Based Retriever Evaluation Strategies
von: Brehme, Lorenz, et al.
Veröffentlicht: (2026) -
Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems
von: Fan, Dongzhe, et al.
Veröffentlicht: (2026) -
DMQR-RAG: Diverse Multi-Query Rewriting for RAG
von: Li, Zhicong, et al.
Veröffentlicht: (2024)