Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Yuxuan, Santos-Rodriguez, Raul, Simpson, Edwin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Using Similarity to Evaluate Factual Consistency in Summaries
von: Ye, Yuxuan, et al.
Veröffentlicht: (2024)
von: Ye, Yuxuan, et al.
Veröffentlicht: (2024)
Ta'keed: The First Generative Fact-Checking System for Arabic Claims
von: Althabiti, Saud, et al.
Veröffentlicht: (2024)
von: Althabiti, Saud, et al.
Veröffentlicht: (2024)
ClaimCheck: Real-Time Fact-Checking with Small Language Models
von: Putta, Akshith Reddy, et al.
Veröffentlicht: (2025)
von: Putta, Akshith Reddy, et al.
Veröffentlicht: (2025)
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
von: Rao, Delip, et al.
Veröffentlicht: (2026)
von: Rao, Delip, et al.
Veröffentlicht: (2026)
ClaimIQ at CheckThat! 2025: Comparing Prompted and Fine-Tuned Language Models for Verifying Numerical Claims
von: Anik, Anirban Saha, et al.
Veröffentlicht: (2025)
von: Anik, Anirban Saha, et al.
Veröffentlicht: (2025)
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
BiDeV: Bilateral Defusing Verification for Complex Claim Fact-Checking
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
von: Rahman, Subhey Sadi, et al.
Veröffentlicht: (2025)
von: Rahman, Subhey Sadi, et al.
Veröffentlicht: (2025)
DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs
von: Huang, Minghui
Veröffentlicht: (2025)
von: Huang, Minghui
Veröffentlicht: (2025)
AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
von: Ni, Jingwei, et al.
Veröffentlicht: (2024)
von: Ni, Jingwei, et al.
Veröffentlicht: (2024)
From Chaos to Clarity: Claim Normalization to Empower Fact-Checking
von: Sundriyal, Megha, et al.
Veröffentlicht: (2023)
von: Sundriyal, Megha, et al.
Veröffentlicht: (2023)
LLMTaxo: Leveraging Large Language Models for Constructing Taxonomy of Factual Claims from Social Media
von: Zhang, Haiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Haiqi, et al.
Veröffentlicht: (2025)
AlignCheck: a Semantic Open-Domain Metric for Factual Consistency Assessment
von: Aghaebrahimian, Ahmad
Veröffentlicht: (2025)
von: Aghaebrahimian, Ahmad
Veröffentlicht: (2025)
Language Models' Factuality Depends on the Language of Inquiry
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
Multimodal Claim Extraction for Fact-Checking
von: Teo, Joycelyn, et al.
Veröffentlicht: (2026)
von: Teo, Joycelyn, et al.
Veröffentlicht: (2026)
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
von: Tang, Liyan, et al.
Veröffentlicht: (2024)
von: Tang, Liyan, et al.
Veröffentlicht: (2024)
HealthFC: Verifying Health Claims with Evidence-Based Medical Fact-Checking
von: Vladika, Juraj, et al.
Veröffentlicht: (2023)
von: Vladika, Juraj, et al.
Veröffentlicht: (2023)
How well can LLMs Grade Essays in Arabic?
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2025)
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2025)
Automated essay scoring in Arabic: a dataset and analysis of a BERT-based system
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2024)
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2024)
Generating Benchmarks for Factuality Evaluation of Language Models
von: Muhlgay, Dor, et al.
Veröffentlicht: (2023)
von: Muhlgay, Dor, et al.
Veröffentlicht: (2023)
Factuality of Large Language Models: A Survey
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
Editing Factual Knowledge and Explanatory Ability of Medical Large Language Models
von: Xu, Derong, et al.
Veröffentlicht: (2024)
von: Xu, Derong, et al.
Veröffentlicht: (2024)
FactTest: Factuality Testing in Large Language Models with Finite-Sample and Distribution-Free Guarantees
von: Nie, Fan, et al.
Veröffentlicht: (2024)
von: Nie, Fan, et al.
Veröffentlicht: (2024)
Aligning Knowledge Graphs and Language Models for Factual Accuracy
von: Nishat, Nur A Zarin, et al.
Veröffentlicht: (2025)
von: Nishat, Nur A Zarin, et al.
Veröffentlicht: (2025)
FLAME: Factuality-Aware Alignment for Large Language Models
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
Towards Automated Fact-Checking of Real-World Claims: Exploring Task Formulation and Assessment with LLMs
von: Sahitaj, Premtim, et al.
Veröffentlicht: (2025)
von: Sahitaj, Premtim, et al.
Veröffentlicht: (2025)
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
von: Shin, Hagyeong, et al.
Veröffentlicht: (2025)
von: Shin, Hagyeong, et al.
Veröffentlicht: (2025)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
Reasoning Factual Knowledge in Structured Data with Large Language Models
von: Huang, Sirui, et al.
Veröffentlicht: (2024)
von: Huang, Sirui, et al.
Veröffentlicht: (2024)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
von: Kawano, Seiya, et al.
Veröffentlicht: (2024)
von: Kawano, Seiya, et al.
Veröffentlicht: (2024)
UNH at CheckThat! 2025: Fine-tuning Vs Prompting in Claim Extraction
von: Wilder, Joe, et al.
Veröffentlicht: (2025)
von: Wilder, Joe, et al.
Veröffentlicht: (2025)
Language Models with Conformal Factuality Guarantees
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
SciClaims: An End-to-End Generative System for Biomedical Claim Analysis
von: Ortega, Raúl, et al.
Veröffentlicht: (2025)
von: Ortega, Raúl, et al.
Veröffentlicht: (2025)
OpenFactCheck: A Unified Framework for Factuality Evaluation of LLMs
von: Iqbal, Hasan, et al.
Veröffentlicht: (2024)
von: Iqbal, Hasan, et al.
Veröffentlicht: (2024)
PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
von: Rooein, Donya, et al.
Veröffentlicht: (2026)
von: Rooein, Donya, et al.
Veröffentlicht: (2026)
Mixture of Reasonings: Teach Large Language Models to Reason with Adaptive Strategies
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality
von: Cheng, Aileen, et al.
Veröffentlicht: (2025)
von: Cheng, Aileen, et al.
Veröffentlicht: (2025)
FactAlign: Long-form Factuality Alignment of Large Language Models
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
Visuospatial Perspective Taking in Multimodal Language Models
von: Prunty, Jonathan, et al.
Veröffentlicht: (2026)
von: Prunty, Jonathan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Using Similarity to Evaluate Factual Consistency in Summaries
von: Ye, Yuxuan, et al.
Veröffentlicht: (2024) -
Ta'keed: The First Generative Fact-Checking System for Arabic Claims
von: Althabiti, Saud, et al.
Veröffentlicht: (2024) -
ClaimCheck: Real-Time Fact-Checking with Small Language Models
von: Putta, Akshith Reddy, et al.
Veröffentlicht: (2025) -
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
von: Rao, Delip, et al.
Veröffentlicht: (2026) -
ClaimIQ at CheckThat! 2025: Comparing Prompted and Fine-Tuned Language Models for Verifying Numerical Claims
von: Anik, Anirban Saha, et al.
Veröffentlicht: (2025)