RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Niu, Cheng, Wu, Yuanhao, Zhu, Juno, Xu, Siliang, Shum, Kashun, Zhong, Randy, Song, Juntong, Zhang, Tong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VeraCT Scan: Retrieval-Augmented Fake News Detection with Justifiable Reasoning
por: Niu, Cheng, et al.
Publicado: (2024)
por: Niu, Cheng, et al.
Publicado: (2024)
OpenGenAlign: A Preference Dataset and Benchmark for Trustworthy Reward Modeling in Open-Ended, Long-Context Generation
por: Zhang, Hanning, et al.
Publicado: (2025)
por: Zhang, Hanning, et al.
Publicado: (2025)
DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning
por: Wu, Yuanhao, et al.
Publicado: (2025)
por: Wu, Yuanhao, et al.
Publicado: (2025)
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo
por: Wood, Michael C., et al.
Publicado: (2024)
por: Wood, Michael C., et al.
Publicado: (2024)
Unmasking Deceptive Visuals: Benchmarking Multimodal Large Language Models on Misleading Chart Question Answering
por: Chen, Zixin, et al.
Publicado: (2025)
por: Chen, Zixin, et al.
Publicado: (2025)
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation
por: Niu, Cheng, et al.
Publicado: (2024)
por: Niu, Cheng, et al.
Publicado: (2024)
Trustworthy Alignment of Retrieval-Augmented Large Language Models via Reinforcement Learning
por: Zhang, Zongmeng, et al.
Publicado: (2024)
por: Zhang, Zongmeng, et al.
Publicado: (2024)
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey
por: Ni, Bo, et al.
Publicado: (2025)
por: Ni, Bo, et al.
Publicado: (2025)
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
por: Kang, Haoqiang, et al.
Publicado: (2023)
por: Kang, Haoqiang, et al.
Publicado: (2023)
Synchronous Faithfulness Monitoring for Trustworthy Retrieval-Augmented Generation
por: Wu, Di, et al.
Publicado: (2024)
por: Wu, Di, et al.
Publicado: (2024)
ReARTeR: Retrieval-Augmented Reasoning with Trustworthy Process Rewarding
por: Sun, Zhongxiang, et al.
Publicado: (2025)
por: Sun, Zhongxiang, et al.
Publicado: (2025)
After Retrieval, Before Generation: Enhancing the Trustworthiness of Large Language Models in Retrieval-Augmented Generation
por: Dai, Xinbang, et al.
Publicado: (2025)
por: Dai, Xinbang, et al.
Publicado: (2025)
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models
por: Lyu, Yuanjie, et al.
Publicado: (2024)
por: Lyu, Yuanjie, et al.
Publicado: (2024)
EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation
por: Niu, Miaohe, et al.
Publicado: (2026)
por: Niu, Miaohe, et al.
Publicado: (2026)
Rowen: Adaptive Retrieval-Augmented Generation for Hallucination Mitigation in LLMs
por: Ding, Hanxing, et al.
Publicado: (2024)
por: Ding, Hanxing, et al.
Publicado: (2024)
Finetune-RAG: Fine-Tuning Language Models to Resist Hallucination in Retrieval-Augmented Generation
por: Lee, Zhan Peng, et al.
Publicado: (2025)
por: Lee, Zhan Peng, et al.
Publicado: (2025)
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability
por: Sun, Zhongxiang, et al.
Publicado: (2024)
por: Sun, Zhongxiang, et al.
Publicado: (2024)
FIRST: Teach A Reliable Large Language Model Through Efficient Trustworthy Distillation
por: Shum, KaShun, et al.
Publicado: (2024)
por: Shum, KaShun, et al.
Publicado: (2024)
RAGRouter: Learning to Route Queries to Multiple Retrieval-Augmented Language Models
por: Zhang, Jiarui, et al.
Publicado: (2025)
por: Zhang, Jiarui, et al.
Publicado: (2025)
Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
por: Shum, KaShun, et al.
Publicado: (2023)
por: Shum, KaShun, et al.
Publicado: (2023)
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
Predictive Data Selection: The Data That Predicts Is the Data That Teaches
por: Shum, Kashun, et al.
Publicado: (2025)
por: Shum, Kashun, et al.
Publicado: (2025)
QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation
por: Min, Dehai, et al.
Publicado: (2025)
por: Min, Dehai, et al.
Publicado: (2025)
Reducing Hallucinations of Medical Multimodal Large Language Models with Visual Retrieval-Augmented Generation
por: Chu, Yun-Wei, et al.
Publicado: (2025)
por: Chu, Yun-Wei, et al.
Publicado: (2025)
Train for Truth, Keep the Skills: Binary Retrieval-Augmented Reward Mitigates Hallucinations
por: Chen, Tong, et al.
Publicado: (2025)
por: Chen, Tong, et al.
Publicado: (2025)
Detecting Hallucinations in Retrieval-Augmented Generation via Semantic-level Internal Reasoning Graph
por: Hu, Jianpeng, et al.
Publicado: (2026)
por: Hu, Jianpeng, et al.
Publicado: (2026)
Harnessing Large Language Models for Knowledge Graph Question Answering via Adaptive Multi-Aspect Retrieval-Augmentation
por: Xu, Derong, et al.
Publicado: (2024)
por: Xu, Derong, et al.
Publicado: (2024)
TrustRAG: Enhancing Robustness and Trustworthiness in Retrieval-Augmented Generation
por: Zhou, Huichi, et al.
Publicado: (2025)
por: Zhou, Huichi, et al.
Publicado: (2025)
InterpDetect: Interpretable Signals for Detecting Hallucinations in Retrieval-Augmented Generation
por: Tan, Likun, et al.
Publicado: (2025)
por: Tan, Likun, et al.
Publicado: (2025)
Parallel Corpus Augmentation using Masked Language Models
por: Kumari, Vibhuti, et al.
Publicado: (2024)
por: Kumari, Vibhuti, et al.
Publicado: (2024)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
por: Yu, Xiaodong, et al.
Publicado: (2023)
por: Yu, Xiaodong, et al.
Publicado: (2023)
DuetRAG: Collaborative Retrieval-Augmented Generation
por: Jiao, Dian, et al.
Publicado: (2024)
por: Jiao, Dian, et al.
Publicado: (2024)
Fusion-Augmented Large Language Models: Boosting Diagnostic Trustworthiness via Model Consensus
por: Siam, Md Kamrul, et al.
Publicado: (2025)
por: Siam, Md Kamrul, et al.
Publicado: (2025)
HRDE: Retrieval-Augmented Large Language Models for Chinese Health Rumor Detection and Explainability
por: Chen, Yanfang, et al.
Publicado: (2024)
por: Chen, Yanfang, et al.
Publicado: (2024)
Graph Retrieval-Augmented Generation: A Survey
por: Peng, Boci, et al.
Publicado: (2024)
por: Peng, Boci, et al.
Publicado: (2024)
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation
por: Qi, Jirui, et al.
Publicado: (2024)
por: Qi, Jirui, et al.
Publicado: (2024)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
por: Zhou, Yujia, et al.
Publicado: (2024)
por: Zhou, Yujia, et al.
Publicado: (2024)
Bi'an: A Bilingual Benchmark and Model for Hallucination Detection in Retrieval-Augmented Generation
por: Jiang, Zhouyu, et al.
Publicado: (2025)
por: Jiang, Zhouyu, et al.
Publicado: (2025)
Mitigating Hallucinations of Large Language Models in Medical Information Extraction via Contrastive Decoding
por: Xu, Derong, et al.
Publicado: (2024)
por: Xu, Derong, et al.
Publicado: (2024)
Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation
por: Zhang, Qianchi, et al.
Publicado: (2026)
por: Zhang, Qianchi, et al.
Publicado: (2026)
Ejemplares similares
-
VeraCT Scan: Retrieval-Augmented Fake News Detection with Justifiable Reasoning
por: Niu, Cheng, et al.
Publicado: (2024) -
OpenGenAlign: A Preference Dataset and Benchmark for Trustworthy Reward Modeling in Open-Ended, Long-Context Generation
por: Zhang, Hanning, et al.
Publicado: (2025) -
DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning
por: Wu, Yuanhao, et al.
Publicado: (2025) -
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo
por: Wood, Michael C., et al.
Publicado: (2024) -
Unmasking Deceptive Visuals: Benchmarking Multimodal Large Language Models on Misleading Chart Question Answering
por: Chen, Zixin, et al.
Publicado: (2025)