FactCorrector: A Graph-Inspired Approach to Long-Form Factuality Correction of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Carnerero-Cano, Javier, Pronesti, Massimiliano, Marinescu, Radu, Tchrakian, Tigran, Barry, James, Gajcin, Jasmina, Hou, Yufang, Pascale, Alessandra, Daly, Elizabeth |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
von: Marinescu, Radu, et al.
Veröffentlicht: (2025)
von: Marinescu, Radu, et al.
Veröffentlicht: (2025)
WikiContradict: A Benchmark for Evaluating LLMs on Real-World Knowledge Conflicts from Wikipedia
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2025)
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2025)
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
von: Dejl, Adam, et al.
Veröffentlicht: (2025)
von: Dejl, Adam, et al.
Veröffentlicht: (2025)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2026)
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2026)
Who Sees the Risk? Stakeholder Conflicts and Explanatory Policies in LLM-based Risk Assessment
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
Query-driven Document-level Scientific Evidence Extraction from Biomedical Studies
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2025)
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2025)
Redefining Counterfactual Explanations for Reinforcement Learning: Overview, Challenges and Opportunities
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2022)
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2022)
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2024)
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2024)
Enhancing Study-Level Inference from Clinical Trial Papers via Reinforcement Learning-Based Numeric Reasoning
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2025)
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2025)
Semifactual Explanations for Reinforcement Learning
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2024)
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2024)
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
von: McCarthy, James, et al.
Veröffentlicht: (2025)
von: McCarthy, James, et al.
Veröffentlicht: (2025)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
von: Liu, Xin, et al.
Veröffentlicht: (2025)
von: Liu, Xin, et al.
Veröffentlicht: (2025)
Secure Change-Point Detection for Time Series under Homomorphic Encryption
von: Mazzone, Federico, et al.
Veröffentlicht: (2026)
von: Mazzone, Federico, et al.
Veröffentlicht: (2026)
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
von: Li, Junliang, et al.
Veröffentlicht: (2025)
von: Li, Junliang, et al.
Veröffentlicht: (2025)
MAD-Fact: A Multi-Agent Debate Framework for Long-Form Factuality Evaluation in LLMs
von: Ning, Yucheng, et al.
Veröffentlicht: (2025)
von: Ning, Yucheng, et al.
Veröffentlicht: (2025)
Generate, Evaluate, Iterate: Synthetic Data for Human-in-the-Loop Refinement of LLM Judges
von: Do, Hyo Jin, et al.
Veröffentlicht: (2025)
von: Do, Hyo Jin, et al.
Veröffentlicht: (2025)
Merging Facts, Crafting Fallacies: Evaluating the Contradictory Nature of Aggregated Factual Claims in Long-Form Generations
von: Chiang, Cheng-Han, et al.
Veröffentlicht: (2024)
von: Chiang, Cheng-Han, et al.
Veröffentlicht: (2024)
The effect of Skyrme--Chern-Simons dynamics on gauged Skyrmions in $2+1$ dimensions
von: Navarro-Lerida, Francisco, et al.
Veröffentlicht: (2023)
von: Navarro-Lerida, Francisco, et al.
Veröffentlicht: (2023)
Attractive and repulsive Yang-Mills--Higgs magnetic monopoles on $\mathbb{R}^3$
von: Navarro-Lérida, Francisco, et al.
Veröffentlicht: (2026)
von: Navarro-Lérida, Francisco, et al.
Veröffentlicht: (2026)
Gauged Skyrme analogue of Chern-Pontryagin
von: Tchrakian, D. H.
Veröffentlicht: (2024)
von: Tchrakian, D. H.
Veröffentlicht: (2024)
FactAlign: Long-form Factuality Alignment of Large Language Models
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
MAVEN-Fact: A Large-scale Event Factuality Detection Dataset
von: Li, Chunyang, et al.
Veröffentlicht: (2024)
von: Li, Chunyang, et al.
Veröffentlicht: (2024)
How Does Response Length Affect Long-Form Factuality
von: Zhao, James Xu, et al.
Veröffentlicht: (2025)
von: Zhao, James Xu, et al.
Veröffentlicht: (2025)
FaStfact: Faster, Stronger Long-Form Factuality Evaluations in LLMs
von: Wan, Yingjia, et al.
Veröffentlicht: (2025)
von: Wan, Yingjia, et al.
Veröffentlicht: (2025)
InFact: Informativeness Alignment for Improved LLM Factuality
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
Fact or Facsimile? Evaluating the Factual Robustness of Modern Retrievers
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
FactTest: Factuality Testing in Large Language Models with Finite-Sample and Distribution-Free Guarantees
von: Nie, Fan, et al.
Veröffentlicht: (2024)
von: Nie, Fan, et al.
Veröffentlicht: (2024)
Long-Form Information Alignment Evaluation Beyond Atomic Facts
von: Zheng, Danna, et al.
Veröffentlicht: (2025)
von: Zheng, Danna, et al.
Veröffentlicht: (2025)
FACTORY: A Challenging Human-Verified Prompt Set for Long-Form Factuality
von: Chen, Mingda, et al.
Veröffentlicht: (2025)
von: Chen, Mingda, et al.
Veröffentlicht: (2025)
DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
von: Wanner, Miriam, et al.
Veröffentlicht: (2024)
von: Wanner, Miriam, et al.
Veröffentlicht: (2024)
$SO(4)$ gauged $O(5)$ Skyrmion on $\mathbb{R}^4$
von: Navarro-Lerida, Francisco, et al.
Veröffentlicht: (2025)
von: Navarro-Lerida, Francisco, et al.
Veröffentlicht: (2025)
Excited solutions in a Skyrme--Chern-Simons model in $2+1$ dimensions
von: Navarro-Lérida, Francisco, et al.
Veröffentlicht: (2026)
von: Navarro-Lérida, Francisco, et al.
Veröffentlicht: (2026)
Factually: Exploring Wearable Fact-Checking for Augmented Truth Discernment
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2025)
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2025)
FactAppeal: Identifying Epistemic Factual Appeals in News Media
von: Mor-Lan, Guy, et al.
Veröffentlicht: (2025)
von: Mor-Lan, Guy, et al.
Veröffentlicht: (2025)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
The Curious Case of Factual (Mis)Alignment between LLMs' Short- and Long-Form Answers
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
Beyond Precision: Importance-Aware Recall for Factuality Evaluation in Long-Form LLM Generation
von: Jafari, Nazanin, et al.
Veröffentlicht: (2026)
von: Jafari, Nazanin, et al.
Veröffentlicht: (2026)
Think Through Uncertainty: Improving Long-Form Generation Factuality via Reasoning Calibration
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
Only Say What You Know: Calibration-Aware Generation for Long-Form Factuality
von: Luo, Wen, et al.
Veröffentlicht: (2026)
von: Luo, Wen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
von: Marinescu, Radu, et al.
Veröffentlicht: (2025) -
WikiContradict: A Benchmark for Evaluating LLMs on Real-World Knowledge Conflicts from Wikipedia
von: Hou, Yufang, et al.
Veröffentlicht: (2024) -
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
von: Gajcin, Jasmina, et al.
Veröffentlicht: (2025) -
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
von: Dejl, Adam, et al.
Veröffentlicht: (2025) -
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
von: Pronesti, Massimiliano, et al.
Veröffentlicht: (2026)