RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Kun, Li, Yunxiang, Zhang, Tianhua, Luo, Hongyin, Wu, Xixin, Glass, James, Meng, Helen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TreePS-RAG: Tree-based Process Supervision for Reinforcement Learning in Agentic RAG
von: Zhang, Tianhua, et al.
Veröffentlicht: (2026)
von: Zhang, Tianhua, et al.
Veröffentlicht: (2026)
Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed Chains
von: Li, Kun, et al.
Veröffentlicht: (2024)
von: Li, Kun, et al.
Veröffentlicht: (2024)
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers
von: Zhang, Tianhua, et al.
Veröffentlicht: (2024)
von: Zhang, Tianhua, et al.
Veröffentlicht: (2024)
Generate, Discriminate, Evolve: Enhancing Context Faithfulness via Fine-Grained Sentence-Level Self-Evolution
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
Natural Language Embedded Programs for Hybrid Language Symbolic Reasoning
von: Zhang, Tianhua, et al.
Veröffentlicht: (2023)
von: Zhang, Tianhua, et al.
Veröffentlicht: (2023)
End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning
von: Zheng, Qiaoyu, et al.
Veröffentlicht: (2025)
von: Zheng, Qiaoyu, et al.
Veröffentlicht: (2025)
OpenRAG: Optimizing RAG End-to-End via In-Context Retrieval Learning
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
Devising a Set of Compact and Explainable Spoken Language Feature for Screening Alzheimer's Disease
von: Li, Junan, et al.
Veröffentlicht: (2024)
von: Li, Junan, et al.
Veröffentlicht: (2024)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Stochastic RAG: End-to-End Retrieval-Augmented Generation through Expected Utility Maximization
von: Zamani, Hamed, et al.
Veröffentlicht: (2024)
von: Zamani, Hamed, et al.
Veröffentlicht: (2024)
EmoRAG: Evaluating RAG Robustness to Symbolic Perturbations
von: Zhou, Xinyun, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyun, et al.
Veröffentlicht: (2025)
Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
An End-to-End Ukrainian RAG for Local Deployment. Optimized Hybrid Search and Lightweight Generation
von: Trokhymovych, Mykola, et al.
Veröffentlicht: (2026)
von: Trokhymovych, Mykola, et al.
Veröffentlicht: (2026)
BioRAG: A RAG-LLM Framework for Biological Question Reasoning
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
THREAD: Thinking Deeper with Recursive Spawning
von: Schroeder, Philip, et al.
Veröffentlicht: (2024)
von: Schroeder, Philip, et al.
Veröffentlicht: (2024)
SIRAG: Towards Stable and Interpretable RAG with A Process-Supervised Multi-Agent Framework
von: Wang, Junlin, et al.
Veröffentlicht: (2025)
von: Wang, Junlin, et al.
Veröffentlicht: (2025)
Rethinking Machine Ethics -- Can LLMs Perform Moral Reasoning through the Lens of Moral Theories?
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
ROVER: Recursive Reasoning Over Videos with Vision-Language Models for Embodied Tasks
von: Schroeder, Philip, et al.
Veröffentlicht: (2025)
von: Schroeder, Philip, et al.
Veröffentlicht: (2025)
RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
CARE-RAG - Clinical Assessment and Reasoning in RAG
von: Potluri, Deepthi, et al.
Veröffentlicht: (2025)
von: Potluri, Deepthi, et al.
Veröffentlicht: (2025)
End-to-End Chatbot Evaluation with Adaptive Reasoning and Uncertainty Filtering
von: Dang, Nhi, et al.
Veröffentlicht: (2026)
von: Dang, Nhi, et al.
Veröffentlicht: (2026)
RuleRAG: Rule-Guided Retrieval-Augmented Generation with Language Models for Question Answering
von: Chen, Zhongwu, et al.
Veröffentlicht: (2024)
von: Chen, Zhongwu, et al.
Veröffentlicht: (2024)
UR$^2$: Unify RAG and Reasoning through Reinforcement Learning
von: Li, Weitao, et al.
Veröffentlicht: (2025)
von: Li, Weitao, et al.
Veröffentlicht: (2025)
Towards End-to-End Model-Agnostic Explanations for RAG Systems
von: Sudhi, Viju, et al.
Veröffentlicht: (2025)
von: Sudhi, Viju, et al.
Veröffentlicht: (2025)
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
von: Khadilkar, Harshad, et al.
Veröffentlicht: (2025)
von: Khadilkar, Harshad, et al.
Veröffentlicht: (2025)
LatentRAG: Latent Reasoning and Retrieval for Efficient Agentic RAG
von: Zheng, Yijia, et al.
Veröffentlicht: (2026)
von: Zheng, Yijia, et al.
Veröffentlicht: (2026)
Zero-RAG: Towards Retrieval-Augmented Generation with Zero Redundant Knowledge
von: Luo, Qi, et al.
Veröffentlicht: (2025)
von: Luo, Qi, et al.
Veröffentlicht: (2025)
Truth or Mirage? Towards End-to-End Factuality Evaluation with LLM-Oasis
von: Scirè, Alessandro, et al.
Veröffentlicht: (2024)
von: Scirè, Alessandro, et al.
Veröffentlicht: (2024)
URO-Bench: Towards Comprehensive Evaluation for End-to-End Spoken Dialogue Models
von: Yan, Ruiqi, et al.
Veröffentlicht: (2025)
von: Yan, Ruiqi, et al.
Veröffentlicht: (2025)
Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore
von: Yan, Zhichao, et al.
Veröffentlicht: (2026)
von: Yan, Zhichao, et al.
Veröffentlicht: (2026)
Injecting linguistic knowledge into BERT for Dialogue State Tracking
von: Feng, Xiaohan, et al.
Veröffentlicht: (2023)
von: Feng, Xiaohan, et al.
Veröffentlicht: (2023)
TextlessRAG: End-to-End Visual Document RAG by Speech Without Text
von: Xie, Peijin, et al.
Veröffentlicht: (2025)
von: Xie, Peijin, et al.
Veröffentlicht: (2025)
Purple-teaming LLMs with Adversarial Defender Training
von: Zhou, Jingyan, et al.
Veröffentlicht: (2024)
von: Zhou, Jingyan, et al.
Veröffentlicht: (2024)
Generating Leakage-Free Benchmarks for Robust RAG Evaluation
von: Liu, Jiayi, et al.
Veröffentlicht: (2026)
von: Liu, Jiayi, et al.
Veröffentlicht: (2026)
DistRAG: Towards Distance-Based Spatial Reasoning in LLMs
von: Schneider, Nicole R, et al.
Veröffentlicht: (2025)
von: Schneider, Nicole R, et al.
Veröffentlicht: (2025)
Retrieval is Not Enough: Enhancing RAG Reasoning through Test-Time Critique and Optimization
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
WavBench: Benchmarking Reasoning, Colloquialism, and Paralinguistics for End-to-End Spoken Dialogue Models
von: Li, Yangzhuo, et al.
Veröffentlicht: (2026)
von: Li, Yangzhuo, et al.
Veröffentlicht: (2026)
From Large to Super-Tiny: End-to-End Optimization for Cost-Efficient LLMs
von: Ni, Jiliang, et al.
Veröffentlicht: (2025)
von: Ni, Jiliang, et al.
Veröffentlicht: (2025)
The Multi-Round Diagnostic RAG Framework for Emulating Clinical Reasoning
von: Sun, Penglei, et al.
Veröffentlicht: (2025)
von: Sun, Penglei, et al.
Veröffentlicht: (2025)
PrismRAG: Boosting RAG Factuality with Distractor Resilience and Strategized Reasoning
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TreePS-RAG: Tree-based Process Supervision for Reinforcement Learning in Agentic RAG
von: Zhang, Tianhua, et al.
Veröffentlicht: (2026) -
Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed Chains
von: Li, Kun, et al.
Veröffentlicht: (2024) -
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers
von: Zhang, Tianhua, et al.
Veröffentlicht: (2024) -
Generate, Discriminate, Evolve: Enhancing Context Faithfulness via Fine-Grained Sentence-Level Self-Evolution
von: Li, Kun, et al.
Veröffentlicht: (2025) -
Natural Language Embedded Programs for Hybrid Language Symbolic Reasoning
von: Zhang, Tianhua, et al.
Veröffentlicht: (2023)