The HalluRAG Dataset: Detecting Closed-Domain Hallucinations in RAG Applications Using an LLM's Internal States
Fuente:
arXiv
Saved in:
| Main Authors: | Ridder, Fabian, Schilling, Malte |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration
by: Ridder, Fabian, et al.
Published: (2026)
by: Ridder, Fabian, et al.
Published: (2026)
HalluGraph: Auditable Hallucination Detection for Legal RAG Systems via Knowledge Graph Alignment
by: Noël, Valentin, et al.
Published: (2025)
by: Noël, Valentin, et al.
Published: (2025)
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
by: Li, Jiarui, et al.
Published: (2024)
by: Li, Jiarui, et al.
Published: (2024)
The Semantic Illusion: Certified Limits of Embedding-Based Hallucination Detection in RAG Systems
by: Sinha, Debu
Published: (2025)
by: Sinha, Debu
Published: (2025)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
by: Pandit, Shrey, et al.
Published: (2025)
by: Pandit, Shrey, et al.
Published: (2025)
RAGProbe: An Automated Approach for Evaluating RAG Applications
by: Sivasothy, Shangeetha, et al.
Published: (2024)
by: Sivasothy, Shangeetha, et al.
Published: (2024)
Reward-RAG: Enhancing RAG with Reward Driven Supervision
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
LatentRAG: Latent Reasoning and Retrieval for Efficient Agentic RAG
by: Zheng, Yijia, et al.
Published: (2026)
by: Zheng, Yijia, et al.
Published: (2026)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
by: Liu, Emmy, et al.
Published: (2026)
by: Liu, Emmy, et al.
Published: (2026)
LLM-Independent Adaptive RAG: Let the Question Speak for Itself
by: Marina, Maria, et al.
Published: (2025)
by: Marina, Maria, et al.
Published: (2025)
Diversity Enhances an LLM's Performance in RAG and Long-context Task
by: Wang, Zhichao, et al.
Published: (2025)
by: Wang, Zhichao, et al.
Published: (2025)
ComoRAG: A Cognitive-Inspired Memory-Organized RAG for Stateful Long Narrative Reasoning
by: Wang, Juyuan, et al.
Published: (2025)
by: Wang, Juyuan, et al.
Published: (2025)
SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA
by: Xu, Haozhou, et al.
Published: (2025)
by: Xu, Haozhou, et al.
Published: (2025)
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
by: Binkowski, Jakub, et al.
Published: (2026)
by: Binkowski, Jakub, et al.
Published: (2026)
Structured RAG for Answering Aggregative Questions
by: Koshorek, Omri, et al.
Published: (2025)
by: Koshorek, Omri, et al.
Published: (2025)
Mitigating Bias in RAG: Controlling the Embedder
by: Kim, Taeyoun, et al.
Published: (2025)
by: Kim, Taeyoun, et al.
Published: (2025)
Highlight & Summarize: RAG without the jailbreaks
by: Cherubin, Giovanni, et al.
Published: (2025)
by: Cherubin, Giovanni, et al.
Published: (2025)
HalluSearch at SemEval-2025 Task 3: A Search-Enhanced RAG Pipeline for Hallucination Detection
by: Abdallah, Mohamed A., et al.
Published: (2025)
by: Abdallah, Mohamed A., et al.
Published: (2025)
P-RAG: Prompt-Enhanced Parametric RAG with LoRA and Selective CoT for Biomedical and Multi-Hop QA
by: Lyu, Xingda, et al.
Published: (2026)
by: Lyu, Xingda, et al.
Published: (2026)
PrismRAG: Boosting RAG Factuality with Distractor Resilience and Strategized Reasoning
by: Kachuee, Mohammad, et al.
Published: (2025)
by: Kachuee, Mohammad, et al.
Published: (2025)
Post-training an LLM for RAG? Train on Self-Generated Demonstrations
by: Finlayson, Matthew, et al.
Published: (2025)
by: Finlayson, Matthew, et al.
Published: (2025)
Biomedical Literature Q&A System Using Retrieval-Augmented Generation (RAG)
by: Garg, Mansi, et al.
Published: (2025)
by: Garg, Mansi, et al.
Published: (2025)
FinReflectKG -- HalluBench: GraphRAG Hallucination Benchmark for Financial Question Answering Systems
by: Kumar, Mahesh, et al.
Published: (2026)
by: Kumar, Mahesh, et al.
Published: (2026)
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
by: Jin, Bowen, et al.
Published: (2024)
by: Jin, Bowen, et al.
Published: (2024)
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG
by: Georgiev, Dobrik, et al.
Published: (2026)
by: Georgiev, Dobrik, et al.
Published: (2026)
AT-RAG: An Adaptive RAG Model Enhancing Query Efficiency with Topic Filtering and Iterative Reasoning
by: Rezaei, Mohammad Reza, et al.
Published: (2024)
by: Rezaei, Mohammad Reza, et al.
Published: (2024)
MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models
by: Xia, Peng, et al.
Published: (2024)
by: Xia, Peng, et al.
Published: (2024)
MetaRAG: Metamorphic Testing for Hallucination Detection in RAG Systems
by: Sok, Channdeth, et al.
Published: (2025)
by: Sok, Channdeth, et al.
Published: (2025)
FilterRAG: Zero-Shot Informed Retrieval-Augmented Generation to Mitigate Hallucinations in VQA
by: Sarwar, Nobin
Published: (2025)
by: Sarwar, Nobin
Published: (2025)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
by: Yehuda, Yakir, et al.
Published: (2024)
by: Yehuda, Yakir, et al.
Published: (2024)
Long Context RAG Performance of Large Language Models
by: Leng, Quinn, et al.
Published: (2024)
by: Leng, Quinn, et al.
Published: (2024)
Insight-RAG: Enhancing LLMs with Insight-Driven Augmentation
by: Pezeshkpour, Pouya, et al.
Published: (2025)
by: Pezeshkpour, Pouya, et al.
Published: (2025)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
by: Urlana, Ashok, et al.
Published: (2025)
by: Urlana, Ashok, et al.
Published: (2025)
DRAG: Distilling RAG for SLMs from LLMs to Transfer Knowledge and Mitigate Hallucination via Evidence and Graph-based Distillation
by: Chen, Jennifer, et al.
Published: (2025)
by: Chen, Jennifer, et al.
Published: (2025)
HalluDetect: Detecting, Mitigating, and Benchmarking Hallucinations in Conversational Systems in the Legal Domain
by: Anaokar, Spandan, et al.
Published: (2025)
by: Anaokar, Spandan, et al.
Published: (2025)
Towards Automated Safety Requirements Derivation Using Agent-based RAG
by: Balu, Balahari Vignesh, et al.
Published: (2025)
by: Balu, Balahari Vignesh, et al.
Published: (2025)
PersonaBOT: Bringing Customer Personas to Life with LLMs and RAG
by: Rizwan, Muhammed, et al.
Published: (2025)
by: Rizwan, Muhammed, et al.
Published: (2025)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
Integrating Domain Knowledge for Financial QA: A Multi-Retriever RAG Approach with LLMs
by: Zhang, Yukun, et al.
Published: (2025)
by: Zhang, Yukun, et al.
Published: (2025)
Similar Items
-
RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration
by: Ridder, Fabian, et al.
Published: (2026) -
HalluGraph: Auditable Hallucination Detection for Legal RAG Systems via Knowledge Graph Alignment
by: Noël, Valentin, et al.
Published: (2025) -
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
by: Li, Jiarui, et al.
Published: (2024) -
The Semantic Illusion: Certified Limits of Embedding-Based Hallucination Detection in RAG Systems
by: Sinha, Debu
Published: (2025) -
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
by: Pandit, Shrey, et al.
Published: (2025)