Can LLMs Evaluate Complex Attribution in QA? Automatic Benchmarking using Knowledge Graphs
Fuente:
arXiv
Salvato in:
| Autori principali: | Hu, Nan, Chen, Jiaoyan, Wu, Yike, Qi, Guilin, Wang, Hongru, Bi, Sheng, Chen, Yongrui, Wu, Tongtong, Pan, Jeff Z. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CoTKR: Chain-of-Thought Enhanced Knowledge Rewriting for Complex Knowledge Graph Question Answering
di: Wu, Yike, et al.
Pubblicazione: (2024)
di: Wu, Yike, et al.
Pubblicazione: (2024)
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
di: He, Jie, et al.
Pubblicazione: (2024)
di: He, Jie, et al.
Pubblicazione: (2024)
HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table Understanding
di: Jin, Rihui, et al.
Pubblicazione: (2024)
di: Jin, Rihui, et al.
Pubblicazione: (2024)
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
K-DeCore: Facilitating Knowledge Transfer in Continual Structured Knowledge Reasoning via Knowledge Decoupling
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
Harnessing Diverse Perspectives: A Multi-Agent Framework for Enhanced Error Detection in Knowledge Graphs
di: Li, Yu, et al.
Pubblicazione: (2025)
di: Li, Yu, et al.
Pubblicazione: (2025)
Start from Zero: Triple Set Prediction for Automatic Knowledge Graph Completion
di: Zhang, Wen, et al.
Pubblicazione: (2024)
di: Zhang, Wen, et al.
Pubblicazione: (2024)
Embedding Ontologies via Incorporating Extensional and Intensional Knowledge
di: Wang, Keyu, et al.
Pubblicazione: (2024)
di: Wang, Keyu, et al.
Pubblicazione: (2024)
Question Answering Over Spatio-Temporal Knowledge Graph
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
di: Dai, Xinbang, et al.
Pubblicazione: (2024)
Compound-QA: A Benchmark for Evaluating LLMs on Compound Questions
di: Hou, Yutao, et al.
Pubblicazione: (2024)
di: Hou, Yutao, et al.
Pubblicazione: (2024)
Magic Mushroom: A Customizable Benchmark for Fine-grained Analysis of Retrieval Noise Erosion in RAG Systems
di: Zhang, Yuxin, et al.
Pubblicazione: (2025)
di: Zhang, Yuxin, et al.
Pubblicazione: (2025)
StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models
di: Chen, Yongrui, et al.
Pubblicazione: (2026)
di: Chen, Yongrui, et al.
Pubblicazione: (2026)
MIKE: A New Benchmark for Fine-grained Multimodal Entity Knowledge Editing
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
Atomic Fact Decomposition Helps Attributed Question Answering
di: Yan, Zhichao, et al.
Pubblicazione: (2024)
di: Yan, Zhichao, et al.
Pubblicazione: (2024)
OneEval: Benchmarking LLM Knowledge-intensive Reasoning over Diverse Knowledge Bases
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
Noise-powered Multi-modal Knowledge Graph Representation Framework
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
Knowledge-Aware Neuron Interpretation for Scene Classification
di: Guan, Yong, et al.
Pubblicazione: (2024)
di: Guan, Yong, et al.
Pubblicazione: (2024)
Pandora: Leveraging Code-driven Knowledge Transfer for Unified Structured Knowledge Reasoning
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
From Superficial to Deep: Integrating External Knowledge for Follow-up Question Generation Using Knowledge Graph and LLM
di: Liu, Jianyu, et al.
Pubblicazione: (2025)
di: Liu, Jianyu, et al.
Pubblicazione: (2025)
Prompting Disentangled Embeddings for Knowledge Graph Completion with Pre-trained Language Model
di: Geng, Yuxia, et al.
Pubblicazione: (2023)
di: Geng, Yuxia, et al.
Pubblicazione: (2023)
KGroot: Enhancing Root Cause Analysis through Knowledge Graphs and Graph Convolutional Neural Networks
di: Wang, Tingting, et al.
Pubblicazione: (2024)
di: Wang, Tingting, et al.
Pubblicazione: (2024)
DEE: Dual-stage Explainable Evaluation Method for Text Generation
di: Zhang, Shenyu, et al.
Pubblicazione: (2024)
di: Zhang, Shenyu, et al.
Pubblicazione: (2024)
Multi-modal Knowledge Graph Generation with Semantics-enriched Prompts
di: Xu, Yajing, et al.
Pubblicazione: (2025)
di: Xu, Yajing, et al.
Pubblicazione: (2025)
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
di: Chen, Yongrui, et al.
Pubblicazione: (2025)
KCoEvo: A Knowledge Graph Augmented Framework for Evolutionary Code Generation
di: Kang, Jiazhen, et al.
Pubblicazione: (2026)
di: Kang, Jiazhen, et al.
Pubblicazione: (2026)
Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models
di: Jin, Rihui, et al.
Pubblicazione: (2025)
di: Jin, Rihui, et al.
Pubblicazione: (2025)
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data
di: Min, Dehai, et al.
Pubblicazione: (2024)
di: Min, Dehai, et al.
Pubblicazione: (2024)
Single Image Unlearning: Efficient Machine Unlearning in Multimodal Large Language Models
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
di: Li, Jiaqi, et al.
Pubblicazione: (2024)
Can LLMs be Good Graph Judge for Knowledge Graph Construction?
di: Huang, Haoyu, et al.
Pubblicazione: (2024)
di: Huang, Haoyu, et al.
Pubblicazione: (2024)
MATEval: A Multi-Agent Discussion Framework for Advancing Open-Ended Text Evaluation
di: Li, Yu, et al.
Pubblicazione: (2024)
di: Li, Yu, et al.
Pubblicazione: (2024)
PRIMO: Progressive Induction for Multi-hop Open Rule Generation
di: Liu, Jianyu, et al.
Pubblicazione: (2024)
di: Liu, Jianyu, et al.
Pubblicazione: (2024)
DoG-Instruct: Towards Premium Instruction-Tuning Data via Text-Grounded Instruction Wrapping
di: Chen, Yongrui, et al.
Pubblicazione: (2023)
di: Chen, Yongrui, et al.
Pubblicazione: (2023)
Large Language Models Meet Knowledge Graphs for Question Answering: Synthesis and Opportunities
di: Ma, Chuangtao, et al.
Pubblicazione: (2025)
di: Ma, Chuangtao, et al.
Pubblicazione: (2025)
Can LLMs Fool Graph Learning? Exploring Universal Adversarial Attacks on Text-Attributed Graphs
di: Chen, Zihui, et al.
Pubblicazione: (2026)
di: Chen, Zihui, et al.
Pubblicazione: (2026)
Can LLMs Solve ASP Problems? Insights from a Benchmarking Study (Extended Version)
di: Ren, Lin, et al.
Pubblicazione: (2025)
di: Ren, Lin, et al.
Pubblicazione: (2025)
MLDT: Multi-Level Decomposition for Complex Long-Horizon Robotic Task Planning with Open-Source Large Language Model
di: Wu, Yike, et al.
Pubblicazione: (2024)
di: Wu, Yike, et al.
Pubblicazione: (2024)
ADAG: Automatically Describing Attribution Graphs
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore
di: Yan, Zhichao, et al.
Pubblicazione: (2026)
di: Yan, Zhichao, et al.
Pubblicazione: (2026)
KGAMC: A Novel Knowledge Graph Driven Automatic Modulation Classification Scheme
di: Li, Yike, et al.
Pubblicazione: (2024)
di: Li, Yike, et al.
Pubblicazione: (2024)
Evaluating Knowledge Graph Based Retrieval Augmented Generation Methods under Knowledge Incompleteness
di: Zhou, Dongzhuoran, et al.
Pubblicazione: (2025)
di: Zhou, Dongzhuoran, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CoTKR: Chain-of-Thought Enhanced Knowledge Rewriting for Complex Knowledge Graph Question Answering
di: Wu, Yike, et al.
Pubblicazione: (2024) -
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
di: He, Jie, et al.
Pubblicazione: (2024) -
HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table Understanding
di: Jin, Rihui, et al.
Pubblicazione: (2024) -
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
di: Dai, Xinbang, et al.
Pubblicazione: (2024) -
K-DeCore: Facilitating Knowledge Transfer in Continual Structured Knowledge Reasoning via Knowledge Decoupling
di: Chen, Yongrui, et al.
Pubblicazione: (2025)