MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts
Fuente:
arXiv
Saved in:
| Main Authors: | He, Jiayi, Huang, Yangmin, Du, Qianyun, Zhou, Xiangying, He, Zhiyang, Hu, Jiaxue, Tao, Xiaodong, Lai, Lixian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ProMedical: Hierarchical Fine-Grained Criteria Modeling for Medical LLM Alignment via Explicit Injection
by: Geng, He, et al.
Published: (2026)
by: Geng, He, et al.
Published: (2026)
MedFact: A Large-scale Chinese Dataset for Evidence-based Medical Fact-checking of LLM Responses
by: Chen, Tong, et al.
Published: (2025)
by: Chen, Tong, et al.
Published: (2025)
MedFact-R1: Towards Factual Medical Reasoning via Pseudo-Label Augmentation
by: Li, Gengliang, et al.
Published: (2025)
by: Li, Gengliang, et al.
Published: (2025)
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
CANDY: Benchmarking LLMs' Limitations and Assistive Potential in Chinese Misinformation Fact-Checking
by: Guo, Ruiling, et al.
Published: (2025)
by: Guo, Ruiling, et al.
Published: (2025)
Evaluating Large Language Model Capability in Vietnamese Fact-Checking Data Generation
by: To, Long Truong, et al.
Published: (2024)
by: To, Long Truong, et al.
Published: (2024)
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
by: Ahmad, Sarfraz, et al.
Published: (2025)
by: Ahmad, Sarfraz, et al.
Published: (2025)
CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval
by: Putta, Akshith Reddy, et al.
Published: (2026)
by: Putta, Akshith Reddy, et al.
Published: (2026)
TrendFact: A Benchmark for Explainable Hotspot Perception in Fact-Checking with Natural Language Explanation
by: Zhang, Xiaocheng, et al.
Published: (2024)
by: Zhang, Xiaocheng, et al.
Published: (2024)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
by: Wang, Yuxia, et al.
Published: (2024)
by: Wang, Yuxia, et al.
Published: (2024)
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
by: Wang, Shengkang, et al.
Published: (2024)
by: Wang, Shengkang, et al.
Published: (2024)
Towards Comprehensive Stage-wise Benchmarking of Large Language Models in Fact-Checking
by: Lin, Hongzhan, et al.
Published: (2026)
by: Lin, Hongzhan, et al.
Published: (2026)
FactSim: Fact-Checking for Opinion Summarization
by: Anghinoni, Leandro, et al.
Published: (2026)
by: Anghinoni, Leandro, et al.
Published: (2026)
Are Fact-Checking Tools Helpful? An Exploration of the Usability of Google Fact Check
by: Yang, Qiangeng, et al.
Published: (2024)
by: Yang, Qiangeng, et al.
Published: (2024)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024)
by: Hoa, Tran Thai, et al.
Published: (2024)
Fin-Fact: A Benchmark Dataset for Multimodal Financial Fact Checking and Explanation Generation
by: Rangapur, Aman, et al.
Published: (2023)
by: Rangapur, Aman, et al.
Published: (2023)
How LLMs Fail to Support Fact-Checking
by: Proma, Adiba Mahbub, et al.
Published: (2025)
by: Proma, Adiba Mahbub, et al.
Published: (2025)
Do We Need Language-Specific Fact-Checking Models? The Case of Chinese
by: Zhang, Caiqi, et al.
Published: (2024)
by: Zhang, Caiqi, et al.
Published: (2024)
Assessing Automated Fact-Checking for Medical LLM Responses with Knowledge Graphs
by: Zhou, Shasha, et al.
Published: (2025)
by: Zhou, Shasha, et al.
Published: (2025)
Application and Optimization of Large Models Based on Prompt Tuning for Fact-Check-Worthiness Estimation
by: Yu, Yinglong, et al.
Published: (2025)
by: Yu, Yinglong, et al.
Published: (2025)
FactIR: A Real-World Zero-shot Open-Domain Retrieval Benchmark for Fact-Checking
by: V, Venktesh, et al.
Published: (2025)
by: V, Venktesh, et al.
Published: (2025)
Check the Facts Step by Step
by: Alison Knopf
Published: (2024)
by: Alison Knopf
Published: (2024)
Multi-Agent Fact Checking
by: Verma, Ashwin, et al.
Published: (2025)
by: Verma, Ashwin, et al.
Published: (2025)
(Fact) Check Your Bias
by: Bakke, Eivind Morris, et al.
Published: (2025)
by: Bakke, Eivind Morris, et al.
Published: (2025)
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
by: Lai, Haoran, et al.
Published: (2026)
by: Lai, Haoran, et al.
Published: (2026)
Generative Large Language Models in Automated Fact-Checking: A Survey
by: Vykopal, Ivan, et al.
Published: (2024)
by: Vykopal, Ivan, et al.
Published: (2024)
Automated Fact-Checking of Climate Change Claims with Large Language Models
by: Leippold, Markus, et al.
Published: (2024)
by: Leippold, Markus, et al.
Published: (2024)
Multimodal Large Language Models to Support Real-World Fact-Checking
by: Geng, Jiahui, et al.
Published: (2024)
by: Geng, Jiahui, et al.
Published: (2024)
Fact-Checking with Large Language Models via Probabilistic Certainty and Consistency
by: Wang, Haoran, et al.
Published: (2026)
by: Wang, Haoran, et al.
Published: (2026)
Large Language Models for Multilingual Previously Fact-Checked Claim Detection
by: Vykopal, Ivan, et al.
Published: (2025)
by: Vykopal, Ivan, et al.
Published: (2025)
Heterogeneous Graph Reasoning for Fact Checking over Texts and Tables
by: Gong, Haisong, et al.
Published: (2024)
by: Gong, Haisong, et al.
Published: (2024)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
by: Russo, Daniel, et al.
Published: (2024)
by: Russo, Daniel, et al.
Published: (2024)
ClaimCheck: Real-Time Fact-Checking with Small Language Models
by: Putta, Akshith Reddy, et al.
Published: (2025)
by: Putta, Akshith Reddy, et al.
Published: (2025)
Community‐Driven Fact‐Checking on WhatsApp : Who Fact‐Checks Whom, Why, and With What Effect?
by: Kiran Garimella
Published: (2025)
by: Kiran Garimella
Published: (2025)
ECAMP: Entity-centered Context-aware Medical Vision Language Pre-training
by: Wang, Rongsheng, et al.
Published: (2023)
by: Wang, Rongsheng, et al.
Published: (2023)
Show Me the Work: Fact-Checkers' Requirements for Explainable Automated Fact-Checking
by: Warren, Greta, et al.
Published: (2025)
by: Warren, Greta, et al.
Published: (2025)
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
by: Sawczyn, Albert, et al.
Published: (2025)
by: Sawczyn, Albert, et al.
Published: (2025)
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models
by: Li, Miaoran, et al.
Published: (2023)
by: Li, Miaoran, et al.
Published: (2023)
Fact Checking Beyond Training Set
by: Karisani, Payam, et al.
Published: (2024)
by: Karisani, Payam, et al.
Published: (2024)
Multimodal Claim Extraction for Fact-Checking
by: Teo, Joycelyn, et al.
Published: (2026)
by: Teo, Joycelyn, et al.
Published: (2026)
Similar Items
-
ProMedical: Hierarchical Fine-Grained Criteria Modeling for Medical LLM Alignment via Explicit Injection
by: Geng, He, et al.
Published: (2026) -
MedFact: A Large-scale Chinese Dataset for Evidence-based Medical Fact-checking of LLM Responses
by: Chen, Tong, et al.
Published: (2025) -
MedFact-R1: Towards Factual Medical Reasoning via Pseudo-Label Augmentation
by: Li, Gengliang, et al.
Published: (2025) -
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking
by: Yang, Shuo, et al.
Published: (2025) -
CANDY: Benchmarking LLMs' Limitations and Assistive Potential in Chinese Misinformation Fact-Checking
by: Guo, Ruiling, et al.
Published: (2025)