Bi'an: A Bilingual Benchmark and Model for Hallucination Detection in Retrieval-Augmented Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Jiang, Zhouyu, Sun, Mengshu, Zhang, Zhiqiang, Liang, Lei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Retrieve, Summarize, Plan: Advancing Multi-hop Question Answering with an Iterative Approach
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
Efficient Knowledge Infusion via KG-LLM Alignment
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024)
Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph Completion
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
Continual Few-shot Event Detection via Hierarchical Augmentation Networks
di: Zhang, Chenlong, et al.
Pubblicazione: (2024)
di: Zhang, Chenlong, et al.
Pubblicazione: (2024)
SKA-Bench: A Fine-Grained Benchmark for Evaluating Structured Knowledge Understanding of LLMs
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation
di: Zhang, Qianchi, et al.
Pubblicazione: (2026)
di: Zhang, Qianchi, et al.
Pubblicazione: (2026)
Detecting Hallucinations in Retrieval-Augmented Generation via Semantic-level Internal Reasoning Graph
di: Hu, Jianpeng, et al.
Pubblicazione: (2026)
di: Hu, Jianpeng, et al.
Pubblicazione: (2026)
InterpDetect: Interpretable Signals for Detecting Hallucinations in Retrieval-Augmented Generation
di: Tan, Likun, et al.
Pubblicazione: (2025)
di: Tan, Likun, et al.
Pubblicazione: (2025)
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability
di: Sun, Zhongxiang, et al.
Pubblicazione: (2024)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2024)
KAG: Boosting LLMs in Professional Domains via Knowledge Augmented Generation
di: Liang, Lei, et al.
Pubblicazione: (2024)
di: Liang, Lei, et al.
Pubblicazione: (2024)
SEReDeEP: Hallucination Detection in Retrieval-Augmented Models via Semantic Entropy and Context-Parameter Fusion
di: Wang, Lei
Pubblicazione: (2025)
di: Wang, Lei
Pubblicazione: (2025)
Rowen: Adaptive Retrieval-Augmented Generation for Hallucination Mitigation in LLMs
di: Ding, Hanxing, et al.
Pubblicazione: (2024)
di: Ding, Hanxing, et al.
Pubblicazione: (2024)
ChatUIE: Exploring Chat-based Unified Information Extraction using Large Language Models
di: Xu, Jun, et al.
Pubblicazione: (2024)
di: Xu, Jun, et al.
Pubblicazione: (2024)
OneGen: Efficient One-Pass Unified Generation and Retrieval for LLMs
di: Zhang, Jintian, et al.
Pubblicazione: (2024)
di: Zhang, Jintian, et al.
Pubblicazione: (2024)
Detecting Hallucination and Coverage Errors in Retrieval Augmented Generation for Controversial Topics
di: Chang, Tyler A., et al.
Pubblicazione: (2024)
di: Chang, Tyler A., et al.
Pubblicazione: (2024)
MAQInstruct: Instruction-based Unified Event Relation Extraction
di: Xu, Jun, et al.
Pubblicazione: (2025)
di: Xu, Jun, et al.
Pubblicazione: (2025)
MultiRAG: A Knowledge-guided Framework for Mitigating Hallucination in Multi-source Retrieval Augmented Generation
di: Wu, Wenlong, et al.
Pubblicazione: (2025)
di: Wu, Wenlong, et al.
Pubblicazione: (2025)
InstructIE: A Bilingual Instruction-based Information Extraction Dataset
di: Gui, Honghao, et al.
Pubblicazione: (2023)
di: Gui, Honghao, et al.
Pubblicazione: (2023)
PRGB Benchmark: A Robust Placeholder-Assisted Algorithm for Benchmarking Retrieval-Augmented Generation
di: Tan, Zhehao, et al.
Pubblicazione: (2025)
di: Tan, Zhehao, et al.
Pubblicazione: (2025)
Improving Natural Language Understanding for LLMs via Large-Scale Instruction Synthesis
di: Yuan, Lin, et al.
Pubblicazione: (2025)
di: Yuan, Lin, et al.
Pubblicazione: (2025)
RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models
di: Niu, Cheng, et al.
Pubblicazione: (2023)
di: Niu, Cheng, et al.
Pubblicazione: (2023)
Finetune-RAG: Fine-Tuning Language Models to Resist Hallucination in Retrieval-Augmented Generation
di: Lee, Zhan Peng, et al.
Pubblicazione: (2025)
di: Lee, Zhan Peng, et al.
Pubblicazione: (2025)
Benchmarking Retrieval-Augmented Generation for Medicine
di: Xiong, Guangzhi, et al.
Pubblicazione: (2024)
di: Xiong, Guangzhi, et al.
Pubblicazione: (2024)
Reducing Hallucinations of Medical Multimodal Large Language Models with Visual Retrieval-Augmented Generation
di: Chu, Yun-Wei, et al.
Pubblicazione: (2025)
di: Chu, Yun-Wei, et al.
Pubblicazione: (2025)
BianCang: A Traditional Chinese Medicine Large Language Model
di: Wei, Sibo, et al.
Pubblicazione: (2024)
di: Wei, Sibo, et al.
Pubblicazione: (2024)
CFVBench: A Comprehensive Video Benchmark for Fine-grained Multimodal Retrieval-Augmented Generation
di: Wei, Kaiwen, et al.
Pubblicazione: (2025)
di: Wei, Kaiwen, et al.
Pubblicazione: (2025)
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation
di: Hu, Haichuan, et al.
Pubblicazione: (2024)
di: Hu, Haichuan, et al.
Pubblicazione: (2024)
SciCUEval: A Comprehensive Dataset for Evaluating Scientific Context Understanding in Large Language Models
di: Yu, Jing, et al.
Pubblicazione: (2025)
di: Yu, Jing, et al.
Pubblicazione: (2025)
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination
di: Rakin, Salman, et al.
Pubblicazione: (2024)
di: Rakin, Salman, et al.
Pubblicazione: (2024)
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
di: Li, Yaocong, et al.
Pubblicazione: (2026)
di: Li, Yaocong, et al.
Pubblicazione: (2026)
HalluGuard: Evidence-Grounded Small Reasoning Models to Mitigate Hallucinations in Retrieval-Augmented Generation
di: Bergeron, Loris, et al.
Pubblicazione: (2025)
di: Bergeron, Loris, et al.
Pubblicazione: (2025)
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks
di: Tong, Chaodong, et al.
Pubblicazione: (2025)
di: Tong, Chaodong, et al.
Pubblicazione: (2025)
Removal of Hallucination on Hallucination: Debate-Augmented RAG
di: Hu, Wentao, et al.
Pubblicazione: (2025)
di: Hu, Wentao, et al.
Pubblicazione: (2025)
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
di: Li, Shanghao, et al.
Pubblicazione: (2025)
di: Li, Shanghao, et al.
Pubblicazione: (2025)
Predict the Retrieval! Test time adaptation for Retrieval Augmented Generation
di: Sun, Xin, et al.
Pubblicazione: (2026)
di: Sun, Xin, et al.
Pubblicazione: (2026)
OntoTune: Ontology-Driven Self-training for Aligning Large Language Models
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)
LIT-RAGBench: Benchmarking Generator Capabilities of Large Language Models in Retrieval-Augmented Generation
di: Itai, Koki, et al.
Pubblicazione: (2026)
di: Itai, Koki, et al.
Pubblicazione: (2026)
RoleEval: A Bilingual Role Evaluation Benchmark for Large Language Models
di: Shen, Tianhao, et al.
Pubblicazione: (2023)
di: Shen, Tianhao, et al.
Pubblicazione: (2023)
RAL2M: Retrieval Augmented Learning-To-Match Against Hallucination in Compliance-Guaranteed Service Systems
di: Hong, Mengze, et al.
Pubblicazione: (2026)
di: Hong, Mengze, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Retrieve, Summarize, Plan: Advancing Multi-hop Question Answering with an Iterative Approach
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024) -
Efficient Knowledge Infusion via KG-LLM Alignment
di: Jiang, Zhouyu, et al.
Pubblicazione: (2024) -
Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph Completion
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025) -
Continual Few-shot Event Detection via Hierarchical Augmentation Networks
di: Zhang, Chenlong, et al.
Pubblicazione: (2024) -
SKA-Bench: A Fine-Grained Benchmark for Evaluating Structured Knowledge Understanding of LLMs
di: Liu, Zhiqiang, et al.
Pubblicazione: (2025)