Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Guinet, Gauthier, Omidvar-Tehrani, Behrooz, Deoras, Anoop, Callot, Laurent |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAGVUE: A Diagnostic View for Explainable and Automated Evaluation of Retrieval-Augmented Generation
by: Murugaraj, Keerthana, et al.
Published: (2025)
by: Murugaraj, Keerthana, et al.
Published: (2025)
Evaluating Retrieval Quality in Retrieval-Augmented Generation
by: Salemi, Alireza, et al.
Published: (2024)
by: Salemi, Alireza, et al.
Published: (2024)
Generative Language Models with Retrieval Augmented Generation for Automated Short Answer Scoring
by: Wang, Zifan, et al.
Published: (2024)
by: Wang, Zifan, et al.
Published: (2024)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023)
by: Saad-Falcon, Jon, et al.
Published: (2023)
Entity Retrieval for Answering Entity-Centric Questions
by: Shavarani, Hassan S., et al.
Published: (2024)
by: Shavarani, Hassan S., et al.
Published: (2024)
Metacognitive Retrieval-Augmented Large Language Models
by: Zhou, Yujia, et al.
Published: (2024)
by: Zhou, Yujia, et al.
Published: (2024)
Contextual Compression in Retrieval-Augmented Generation for Large Language Models: A Survey
by: Verma, Sourav
Published: (2024)
by: Verma, Sourav
Published: (2024)
Large Language Model Critics for Execution-Free Evaluation of Code Changes
by: Yadavally, Aashish, et al.
Published: (2025)
by: Yadavally, Aashish, et al.
Published: (2025)
DRAGIN: Dynamic Retrieval Augmented Generation based on the Information Needs of Large Language Models
by: Su, Weihang, et al.
Published: (2024)
by: Su, Weihang, et al.
Published: (2024)
Evaluating the Robustness of Retrieval-Augmented Generation to Adversarial Evidence in the Health Domain
by: Amirshahi, Shakiba, et al.
Published: (2025)
by: Amirshahi, Shakiba, et al.
Published: (2025)
Chain-of-Retrieval Augmented Generation
by: Wang, Liang, et al.
Published: (2025)
by: Wang, Liang, et al.
Published: (2025)
Parametric Retrieval Augmented Generation
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
P-RAG: Progressive Retrieval Augmented Generation For Planning on Embodied Everyday Task
by: Xu, Weiye, et al.
Published: (2024)
by: Xu, Weiye, et al.
Published: (2024)
M-RAG: Reinforcing Large Language Model Performance through Retrieval-Augmented Generation with Multiple Partitions
by: Wang, Zheng, et al.
Published: (2024)
by: Wang, Zheng, et al.
Published: (2024)
A Large Language Model-based Framework for Semi-Structured Tender Document Retrieval-Augmented Generation
by: Zhao, Yilong, et al.
Published: (2024)
by: Zhao, Yilong, et al.
Published: (2024)
GeAR: Generation Augmented Retrieval
by: Liu, Haoyu, et al.
Published: (2025)
by: Liu, Haoyu, et al.
Published: (2025)
Dynamic and Parametric Retrieval-Augmented Generation
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
LLM-Specific Utility: A New Perspective for Retrieval-Augmented Generation
by: Zhang, Hengran, et al.
Published: (2025)
by: Zhang, Hengran, et al.
Published: (2025)
Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation
by: Salemi, Alireza, et al.
Published: (2024)
by: Salemi, Alireza, et al.
Published: (2024)
DS@GT at Touché: Large Language Models for Retrieval-Augmented Debate
by: Miyaguchi, Anthony, et al.
Published: (2025)
by: Miyaguchi, Anthony, et al.
Published: (2025)
Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
by: Ren, Ruiyang, et al.
Published: (2023)
by: Ren, Ruiyang, et al.
Published: (2023)
Retrieval Augmented Generation Systems: Automatic Dataset Creation, Evaluation and Boolean Agent Setup
by: Kenneweg, Tristan, et al.
Published: (2024)
by: Kenneweg, Tristan, et al.
Published: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
On Synthetic Data Strategies for Domain-Specific Generative Retrieval
by: Wen, Haoyang, et al.
Published: (2025)
by: Wen, Haoyang, et al.
Published: (2025)
Adaptive Retrieval-Augmented Generation for Conversational Systems
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
Loops On Retrieval Augmented Generation (LoRAG)
by: Thakur, Ayush, et al.
Published: (2024)
by: Thakur, Ayush, et al.
Published: (2024)
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models
by: Cong, Youan, et al.
Published: (2024)
by: Cong, Youan, et al.
Published: (2024)
Conversational Text Extraction with Large Language Models Using Retrieval-Augmented Systems
by: Roy, Soham, et al.
Published: (2025)
by: Roy, Soham, et al.
Published: (2025)
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning
by: Xu, Jian, et al.
Published: (2025)
by: Xu, Jian, et al.
Published: (2025)
DRAMA: Diverse Augmentation from Large Language Models to Smaller Dense Retrievers
by: Ma, Xueguang, et al.
Published: (2025)
by: Ma, Xueguang, et al.
Published: (2025)
G-Refer: Graph Retrieval-Augmented Large Language Model for Explainable Recommendation
by: Li, Yuhan, et al.
Published: (2025)
by: Li, Yuhan, et al.
Published: (2025)
Evaluating Hybrid Retrieval Augmented Generation using Dynamic Test Sets: LiveRAG Challenge
by: Fensore, Chase, et al.
Published: (2025)
by: Fensore, Chase, et al.
Published: (2025)
Evaluating Large Language Models for Cross-Lingual Retrieval
by: Zuo, Longfei, et al.
Published: (2025)
by: Zuo, Longfei, et al.
Published: (2025)
Resolving Conflicting Evidence in Automated Fact-Checking: A Study on Retrieval-Augmented LLMs
by: Ge, Ziyu, et al.
Published: (2025)
by: Ge, Ziyu, et al.
Published: (2025)
Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation
by: Shi, Teng, et al.
Published: (2025)
by: Shi, Teng, et al.
Published: (2025)
A Survey on Retrieval-Augmented Text Generation for Large Language Models
by: Huang, Yizheng, et al.
Published: (2024)
by: Huang, Yizheng, et al.
Published: (2024)
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
Evaluating the Effectiveness and Scalability of LLM-Based Data Augmentation for Retrieval
by: Chitale, Pranjal A., et al.
Published: (2025)
by: Chitale, Pranjal A., et al.
Published: (2025)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Similar Items
-
RAGVUE: A Diagnostic View for Explainable and Automated Evaluation of Retrieval-Augmented Generation
by: Murugaraj, Keerthana, et al.
Published: (2025) -
Evaluating Retrieval Quality in Retrieval-Augmented Generation
by: Salemi, Alireza, et al.
Published: (2024) -
Generative Language Models with Retrieval Augmented Generation for Automated Short Answer Scoring
by: Wang, Zifan, et al.
Published: (2024) -
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023) -
Entity Retrieval for Answering Entity-Centric Questions
by: Shavarani, Hassan S., et al.
Published: (2024)