Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Shuhang, Zhou, Chuhao, Lin, Xiao, Dong, Zihan, Lu, Kuan, Peng, Zhencan, Yin, Jie, Metaxas, Dimitris N. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Theoretical Foundations of Latent Posterior Factors: Formal Guarantees for Multi-Evidence Reasoning
por: Alege, Aliyu Agboola
Publicado: (2026)
por: Alege, Aliyu Agboola
Publicado: (2026)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
por: Lei, Xiang, et al.
Publicado: (2025)
por: Lei, Xiang, et al.
Publicado: (2025)
ARCADIA: Scalable Causal Discovery for Corporate Bankruptcy Analysis Using Agentic AI
por: Maturo, Fabrizio, et al.
Publicado: (2025)
por: Maturo, Fabrizio, et al.
Publicado: (2025)
AssistedDS: Benchmarking How External Domain Knowledge Assists LLMs in Automated Data Science
por: Luo, An, et al.
Publicado: (2025)
por: Luo, An, et al.
Publicado: (2025)
Can Agentic AI Match the Performance of Human Data Scientists?
por: Luo, An, et al.
Publicado: (2025)
por: Luo, An, et al.
Publicado: (2025)
AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science
por: Luo, An, et al.
Publicado: (2026)
por: Luo, An, et al.
Publicado: (2026)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
por: Zhang, Li, et al.
Publicado: (2025)
por: Zhang, Li, et al.
Publicado: (2025)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
por: Alege, Aliyu Agboola
Publicado: (2026)
por: Alege, Aliyu Agboola
Publicado: (2026)
ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs
por: Basu, Abhinaba, et al.
Publicado: (2026)
por: Basu, Abhinaba, et al.
Publicado: (2026)
Conformal Prediction Sets for Next-Token Prediction in Large Language Models: Balancing Coverage Guarantees with Set Efficiency
por: Kotla, Yoshith Roy, et al.
Publicado: (2025)
por: Kotla, Yoshith Roy, et al.
Publicado: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
por: Haque, Md. Asraful, et al.
Publicado: (2026)
por: Haque, Md. Asraful, et al.
Publicado: (2026)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
por: Nwokocha, Caleb Princewill
Publicado: (2022)
por: Nwokocha, Caleb Princewill
Publicado: (2022)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
por: Zhang, Li, et al.
Publicado: (2026)
por: Zhang, Li, et al.
Publicado: (2026)
Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference
por: Du, Jin, et al.
Publicado: (2025)
por: Du, Jin, et al.
Publicado: (2025)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
por: Manokhin, Valery, et al.
Publicado: (2026)
por: Manokhin, Valery, et al.
Publicado: (2026)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
por: Zhang, Liangliang, et al.
Publicado: (2025)
por: Zhang, Liangliang, et al.
Publicado: (2025)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
por: Amamou, Hazem, et al.
Publicado: (2026)
por: Amamou, Hazem, et al.
Publicado: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
Bayesian Learning-driven Prototypical Contrastive Loss for Class-Incremental Learning
por: Raichur, Nisha L., et al.
Publicado: (2024)
por: Raichur, Nisha L., et al.
Publicado: (2024)
Federated Learning with MMD-based Early Stopping for Adaptive GNSS Interference Classification
por: Gaikwad, Nishant S., et al.
Publicado: (2024)
por: Gaikwad, Nishant S., et al.
Publicado: (2024)
KNIGHT: Knowledge Graph-Driven Multiple-Choice Question Generation with Adaptive Hardness Calibration
por: Amanlou, Mohammad, et al.
Publicado: (2026)
por: Amanlou, Mohammad, et al.
Publicado: (2026)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
por: Kaiser, Daniel, et al.
Publicado: (2025)
por: Kaiser, Daniel, et al.
Publicado: (2025)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
por: Cao, Deyu, et al.
Publicado: (2026)
por: Cao, Deyu, et al.
Publicado: (2026)
PLUGH: A Benchmark for Spatial Understanding and Reasoning in Large Language Models
por: Tikhonov, Alexey
Publicado: (2024)
por: Tikhonov, Alexey
Publicado: (2024)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
por: Cao, Deyu, et al.
Publicado: (2025)
por: Cao, Deyu, et al.
Publicado: (2025)
ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable
por: Chang, Sidi, et al.
Publicado: (2026)
por: Chang, Sidi, et al.
Publicado: (2026)
Customizing Graph Neural Networks using Path Reweighting
por: Chen, Jianpeng, et al.
Publicado: (2021)
por: Chen, Jianpeng, et al.
Publicado: (2021)
Chase Anonymisation: Privacy-Preserving Knowledge Graphs with Logical Reasoning
por: Bellomarini, Luigi, et al.
Publicado: (2024)
por: Bellomarini, Luigi, et al.
Publicado: (2024)
Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models
por: Wibbeke, Jelke, et al.
Publicado: (2025)
por: Wibbeke, Jelke, et al.
Publicado: (2025)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026)
por: Jia, Xiao
Publicado: (2026)
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
por: Badshah, Sher, et al.
Publicado: (2026)
por: Badshah, Sher, et al.
Publicado: (2026)
CAPE: Corrective Actions from Precondition Errors using Large Language Models
por: Raman, Shreyas Sundara, et al.
Publicado: (2022)
por: Raman, Shreyas Sundara, et al.
Publicado: (2022)
On measuring grounding and generalizing grounding problems
por: Quigley, Daniel, et al.
Publicado: (2025)
por: Quigley, Daniel, et al.
Publicado: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
por: Chanda, Prateek, et al.
Publicado: (2025)
por: Chanda, Prateek, et al.
Publicado: (2025)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
por: Yim, Wen-wai, et al.
Publicado: (2025)
por: Yim, Wen-wai, et al.
Publicado: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
por: Li, Yang, et al.
Publicado: (2026)
por: Li, Yang, et al.
Publicado: (2026)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
por: Patel, Urjitkumar, et al.
Publicado: (2025)
por: Patel, Urjitkumar, et al.
Publicado: (2025)
CLEV: LLM-Based Evaluation Through Lightweight Efficient Voting for Free-Form Question-Answering
por: Badshah, Sher, et al.
Publicado: (2025)
por: Badshah, Sher, et al.
Publicado: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
Ejemplares similares
-
Theoretical Foundations of Latent Posterior Factors: Formal Guarantees for Multi-Evidence Reasoning
por: Alege, Aliyu Agboola
Publicado: (2026) -
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
por: Lei, Xiang, et al.
Publicado: (2025) -
ARCADIA: Scalable Causal Discovery for Corporate Bankruptcy Analysis Using Agentic AI
por: Maturo, Fabrizio, et al.
Publicado: (2025) -
AssistedDS: Benchmarking How External Domain Knowledge Assists LLMs in Automated Data Science
por: Luo, An, et al.
Publicado: (2025) -
Can Agentic AI Match the Performance of Human Data Scientists?
por: Luo, An, et al.
Publicado: (2025)