Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Zhichao, Zhao, Yunxiao, Wang, Jiapu, Chen, Jiaoyan, Li, Xiaoli, Li, Ru, Pan, Jeff Z. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Atomic Fact Decomposition Helps Attributed Question Answering
von: Yan, Zhichao, et al.
Veröffentlicht: (2024)
von: Yan, Zhichao, et al.
Veröffentlicht: (2024)
Decomposing and Revising What Language Models Generate
von: Yan, Zhichao, et al.
Veröffentlicht: (2025)
von: Yan, Zhichao, et al.
Veröffentlicht: (2025)
Prompting Large Language Models with Partial Knowledge for Answering Questions with Unseen Entities
von: Yan, Zhichao, et al.
Veröffentlicht: (2025)
von: Yan, Zhichao, et al.
Veröffentlicht: (2025)
Beyond Factual Accuracy: Evaluating Coverage of Diverse Factual Information in Long-form Text Generation
von: Samarinas, Chris, et al.
Veröffentlicht: (2025)
von: Samarinas, Chris, et al.
Veröffentlicht: (2025)
Trustworthy Reasoning: Evaluating and Enhancing Factual Accuracy in LLM Intermediate Thought Processes
von: Jiao, Rui, et al.
Veröffentlicht: (2025)
von: Jiao, Rui, et al.
Veröffentlicht: (2025)
Explaining Black-box Language Models with Knowledge Probing Systems: A Post-hoc Explanation Perspective
von: Zhao, Yunxiao, et al.
Veröffentlicht: (2025)
von: Zhao, Yunxiao, et al.
Veröffentlicht: (2025)
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
von: He, Jie, et al.
Veröffentlicht: (2024)
von: He, Jie, et al.
Veröffentlicht: (2024)
Evaluating Mathematical Reasoning Beyond Accuracy
von: Xia, Shijie, et al.
Veröffentlicht: (2024)
von: Xia, Shijie, et al.
Veröffentlicht: (2024)
Evaluating Factual Density in Multi-Source RAG: A Study in Medical AI Accuracy
von: DeMarco, Michael R.
Veröffentlicht: (2026)
von: DeMarco, Michael R.
Veröffentlicht: (2026)
TrustScore: Reference-Free Evaluation of LLM Response Trustworthiness
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
Consistency-Aware Editing for Entity-level Unlearning in Language Models
von: Han, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Han, Xiaoqi, et al.
Veröffentlicht: (2025)
PrismRAG: Boosting RAG Factuality with Distractor Resilience and Strategized Reasoning
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
MoleculeQA: A Dataset to Evaluate Factual Accuracy in Molecular Comprehension
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
Long-Form Information Alignment Evaluation Beyond Atomic Facts
von: Zheng, Danna, et al.
Veröffentlicht: (2025)
von: Zheng, Danna, et al.
Veröffentlicht: (2025)
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
von: Li, Jiarui, et al.
Veröffentlicht: (2024)
von: Li, Jiarui, et al.
Veröffentlicht: (2024)
UniArk: Improving Generalisation and Consistency for Factual Knowledge Extraction through Debiasing
von: Yang, Yijun, et al.
Veröffentlicht: (2024)
von: Yang, Yijun, et al.
Veröffentlicht: (2024)
FS-RAG: A Frame Semantics Based Approach for Improved Factual Accuracy in Large Language Models
von: Madabushi, Harish Tayyar
Veröffentlicht: (2024)
von: Madabushi, Harish Tayyar
Veröffentlicht: (2024)
CLR-Fact: Evaluating the Complex Logical Reasoning Capability of Large Language Models over Factual Knowledge
von: Zheng, Tianshi, et al.
Veröffentlicht: (2024)
von: Zheng, Tianshi, et al.
Veröffentlicht: (2024)
ReaRAG: Knowledge-guided Reasoning Enhances Factuality of Large Reasoning Models with Iterative Retrieval Augmented Generation
von: Lee, Zhicheng, et al.
Veröffentlicht: (2025)
von: Lee, Zhicheng, et al.
Veröffentlicht: (2025)
Review of Inference-Time Scaling Strategies: Reasoning, Search and RAG
von: Wang, Zhichao, et al.
Veröffentlicht: (2025)
von: Wang, Zhichao, et al.
Veröffentlicht: (2025)
Beyond Scores: A Modular RAG-Based System for Automatic Short Answer Scoring with Feedback
von: Fateen, Menna, et al.
Veröffentlicht: (2024)
von: Fateen, Menna, et al.
Veröffentlicht: (2024)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
von: Clark, Peter, et al.
Veröffentlicht: (2023)
von: Clark, Peter, et al.
Veröffentlicht: (2023)
ClaimTrust: Propagation Trust Scoring for RAG Systems
von: Qian, Hangkai, et al.
Veröffentlicht: (2025)
von: Qian, Hangkai, et al.
Veröffentlicht: (2025)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
Can LLMs Evaluate Complex Attribution in QA? Automatic Benchmarking using Knowledge Graphs
von: Hu, Nan, et al.
Veröffentlicht: (2024)
von: Hu, Nan, et al.
Veröffentlicht: (2024)
FIT-RAG: Black-Box RAG with Factual Information and Token Reduction
von: Mao, Yuren, et al.
Veröffentlicht: (2024)
von: Mao, Yuren, et al.
Veröffentlicht: (2024)
Face4RAG: Factual Consistency Evaluation for Retrieval Augmented Generation in Chinese
von: Xu, Yunqi, et al.
Veröffentlicht: (2024)
von: Xu, Yunqi, et al.
Veröffentlicht: (2024)
SuperRAG: Beyond RAG with Layout-Aware Graph Modeling
von: Yang, Jeff, et al.
Veröffentlicht: (2025)
von: Yang, Jeff, et al.
Veröffentlicht: (2025)
CoTKR: Chain-of-Thought Enhanced Knowledge Rewriting for Complex Knowledge Graph Question Answering
von: Wu, Yike, et al.
Veröffentlicht: (2024)
von: Wu, Yike, et al.
Veröffentlicht: (2024)
Learning to Reason for Factuality
von: Chen, Xilun, et al.
Veröffentlicht: (2025)
von: Chen, Xilun, et al.
Veröffentlicht: (2025)
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation
von: Liu, Hao, et al.
Veröffentlicht: (2025)
von: Liu, Hao, et al.
Veröffentlicht: (2025)
Locomo-Plus: Beyond-Factual Cognitive Memory Evaluation Framework for LLM Agents
von: Li, Yifei, et al.
Veröffentlicht: (2026)
von: Li, Yifei, et al.
Veröffentlicht: (2026)
Are Large Language Models Really Good Logical Reasoners? A Comprehensive Evaluation and Beyond
von: Xu, Fangzhi, et al.
Veröffentlicht: (2023)
von: Xu, Fangzhi, et al.
Veröffentlicht: (2023)
GlobalRAG: Enhancing Global Reasoning in Multi-hop Question Answering via Reinforcement Learning
von: Luo, Jinchang, et al.
Veröffentlicht: (2025)
von: Luo, Jinchang, et al.
Veröffentlicht: (2025)
Toward Copyright Integrity and Verifiability via Multi-Bit Watermarking for Intelligent Transportation Systems
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
Knowledge Reasoning Language Model: Unifying Knowledge and Language for Inductive Knowledge Graph Reasoning
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Atomic Fact Decomposition Helps Attributed Question Answering
von: Yan, Zhichao, et al.
Veröffentlicht: (2024) -
Decomposing and Revising What Language Models Generate
von: Yan, Zhichao, et al.
Veröffentlicht: (2025) -
Prompting Large Language Models with Partial Knowledge for Answering Questions with Unseen Entities
von: Yan, Zhichao, et al.
Veröffentlicht: (2025) -
Beyond Factual Accuracy: Evaluating Coverage of Diverse Factual Information in Long-form Text Generation
von: Samarinas, Chris, et al.
Veröffentlicht: (2025) -
Trustworthy Reasoning: Evaluating and Enhancing Factual Accuracy in LLM Intermediate Thought Processes
von: Jiao, Rui, et al.
Veröffentlicht: (2025)