Can we Evaluate RAGs with Synthetic Data?
Fuente:
arXiv
Saved in:
| Main Authors: | van Elburg, Jonas, van der Putten, Peter, Marx, Maarten |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agentic Large Language Models, a survey
by: Plaat, Aske, et al.
Published: (2025)
by: Plaat, Aske, et al.
Published: (2025)
RAGs to Riches: RAG-like Few-shot Learning for Large Language Model Role-playing
by: Rupprecht, Timothy, et al.
Published: (2025)
by: Rupprecht, Timothy, et al.
Published: (2025)
Generating High Quality Synthetic Data for Dutch Medical Conversations
by: Kuan, Cecilia, et al.
Published: (2026)
by: Kuan, Cecilia, et al.
Published: (2026)
From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries
by: Wadhwa, Hitesh, et al.
Published: (2024)
by: Wadhwa, Hitesh, et al.
Published: (2024)
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Reasoning Promotes Robustness in Theory of Mind Tasks
by: de Haan, Ian B., et al.
Published: (2026)
by: de Haan, Ian B., et al.
Published: (2026)
Facilitating Opinion Diversity through Hybrid NLP Approaches
by: van der Meer, Michiel
Published: (2024)
by: van der Meer, Michiel
Published: (2024)
The Impacts of AI Avatar Appearance and Disclosure on User Motivation
by: Visser, Boele, et al.
Published: (2024)
by: Visser, Boele, et al.
Published: (2024)
Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation
by: Cooray, Lakshan, et al.
Published: (2026)
by: Cooray, Lakshan, et al.
Published: (2026)
The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data
by: Pouplin, Thomas, et al.
Published: (2024)
by: Pouplin, Thomas, et al.
Published: (2024)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
by: Fanconi, Claudio, et al.
Published: (2025)
by: Fanconi, Claudio, et al.
Published: (2025)
Investigating the Role of Prompting and External Tools in Hallucination Rates of Large Language Models
by: Barkley, Liam, et al.
Published: (2024)
by: Barkley, Liam, et al.
Published: (2024)
Grounding Synthetic Data Evaluations of Language Models in Unsupervised Document Corpora
by: Majurski, Michael, et al.
Published: (2025)
by: Majurski, Michael, et al.
Published: (2025)
Can we trust the evaluation on ChatGPT?
by: Aiyappa, Rachith, et al.
Published: (2023)
by: Aiyappa, Rachith, et al.
Published: (2023)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
Can Large Language Models generalize analogy solving like children can?
by: Stevenson, Claire E., et al.
Published: (2024)
by: Stevenson, Claire E., et al.
Published: (2024)
Reasoning-Driven Synthetic Data Generation and Evaluation
by: Davidson, Tim R., et al.
Published: (2026)
by: Davidson, Tim R., et al.
Published: (2026)
Modular Techniques for Synthetic Long-Context Data Generation in Language Model Training and Evaluation
by: Subramanian, Seganrasan, et al.
Published: (2025)
by: Subramanian, Seganrasan, et al.
Published: (2025)
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method
by: Chen, Peter Baile, et al.
Published: (2025)
by: Chen, Peter Baile, et al.
Published: (2025)
Leveraging LLMs for Bangla Grammar Error Correction:Error Categorization, Synthetic Data, and Model Evaluation
by: Bhattacharyya, Pramit, et al.
Published: (2024)
by: Bhattacharyya, Pramit, et al.
Published: (2024)
Virus Infection Attack on LLMs: Your Poisoning Can Spread "VIA" Synthetic Data
by: Liang, Zi, et al.
Published: (2025)
by: Liang, Zi, et al.
Published: (2025)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
by: Ziki, Batsirayi Mupamhi, et al.
Published: (2026)
Scaling few-shot spoken word classification with generative meta-continual learning
by: Beyers, Louise, et al.
Published: (2026)
by: Beyers, Louise, et al.
Published: (2026)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
Undesirable Biases in NLP: Addressing Challenges of Measurement
by: van der Wal, Oskar, et al.
Published: (2022)
by: van der Wal, Oskar, et al.
Published: (2022)
What Has Been Lost with Synthetic Evaluation?
by: Gill, Alexander, et al.
Published: (2025)
by: Gill, Alexander, et al.
Published: (2025)
Evaluating Morphological Compositional Generalization in Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Scaling Laws of Synthetic Data for Language Models
by: Qin, Zeyu, et al.
Published: (2025)
by: Qin, Zeyu, et al.
Published: (2025)
CircuitSynth: Reliable Synthetic Data Generation
by: Cheng, Zehua, et al.
Published: (2026)
by: Cheng, Zehua, et al.
Published: (2026)
Language Bottleneck Models for Qualitative Knowledge State Modeling
by: Berthon, Antonin, et al.
Published: (2025)
by: Berthon, Antonin, et al.
Published: (2025)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
by: Sun, Hao, et al.
Published: (2025)
by: Sun, Hao, et al.
Published: (2025)
Evaluation of Multilingual Image Captioning: How far can we get with CLIP models?
by: Gomes, Gonçalo, et al.
Published: (2025)
by: Gomes, Gonçalo, et al.
Published: (2025)
Creativity in AI: Progresses and Challenges
by: Ismayilzada, Mete, et al.
Published: (2024)
by: Ismayilzada, Mete, et al.
Published: (2024)
Bridging Domain Knowledge and Process Discovery Using Large Language Models
by: Norouzifar, Ali, et al.
Published: (2024)
by: Norouzifar, Ali, et al.
Published: (2024)
A Universal Prompting Strategy for Extracting Process Model Information from Natural Language Text using Large Language Models
by: Neuberger, Julian, et al.
Published: (2024)
by: Neuberger, Julian, et al.
Published: (2024)
Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets
by: Xenofontos, Andreas, et al.
Published: (2026)
by: Xenofontos, Andreas, et al.
Published: (2026)
Is the Pope Catholic? Yes, the Pope is Catholic. Generative Evaluation of Non-Literal Intent Resolution in LLMs
by: Yerukola, Akhila, et al.
Published: (2024)
by: Yerukola, Akhila, et al.
Published: (2024)
Configurable Preference Tuning with Rubric-Guided Synthetic Data
by: Gallego, Víctor
Published: (2025)
by: Gallego, Víctor
Published: (2025)
Few-shot LLM Synthetic Data with Distribution Matching
by: Ren, Jiyuan, et al.
Published: (2025)
by: Ren, Jiyuan, et al.
Published: (2025)
Guided Persona-based AI Surveys: Can we replicate personal mobility preferences at scale using LLMs?
by: Tzachristas, Ioannis, et al.
Published: (2025)
by: Tzachristas, Ioannis, et al.
Published: (2025)
Similar Items
-
Agentic Large Language Models, a survey
by: Plaat, Aske, et al.
Published: (2025) -
RAGs to Riches: RAG-like Few-shot Learning for Large Language Model Role-playing
by: Rupprecht, Timothy, et al.
Published: (2025) -
Generating High Quality Synthetic Data for Dutch Medical Conversations
by: Kuan, Cecilia, et al.
Published: (2026) -
From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries
by: Wadhwa, Hitesh, et al.
Published: (2024) -
Evaluating Creative Short Story Generation in Humans and Large Language Models
by: Ismayilzada, Mete, et al.
Published: (2024)