A Benchmark for Open-Domain Numerical Fact-Checking Enhanced by Claim Decomposition
Fuente:
arXiv
Saved in:
| Main Authors: | Venktesh, V, Prabhu, Deepali, Anand, Avishek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DEXTER: A Benchmark for open-domain Complex Question Answering using LLMs
by: Prabhu, Venktesh V. Deepali, et al.
Published: (2024)
by: Prabhu, Venktesh V. Deepali, et al.
Published: (2024)
FactIR: A Real-World Zero-shot Open-Domain Retrieval Benchmark for Fact-Checking
by: V, Venktesh, et al.
Published: (2025)
by: V, Venktesh, et al.
Published: (2025)
FlashCheck: Exploration of Efficient Evidence Retrieval for Fast Fact-Checking
by: Nanekhan, Kevin, et al.
Published: (2025)
by: Nanekhan, Kevin, et al.
Published: (2025)
Decomposition Dilemmas: Does Claim Decomposition Boost or Burden Fact-Checking Performance?
by: Hu, Qisheng, et al.
Published: (2024)
by: Hu, Qisheng, et al.
Published: (2024)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024)
by: Hoa, Tran Thai, et al.
Published: (2024)
SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
by: Peng, Qiwei, et al.
Published: (2025)
by: Peng, Qiwei, et al.
Published: (2025)
SUNAR: Semantic Uncertainty based Neighborhood Aware Retrieval for Complex QA
by: Venktesh, V, et al.
Published: (2025)
by: Venktesh, V, et al.
Published: (2025)
When More Reformulations Hurt: Avoiding Drift using Ranker Feedback
by: Venktesh, V, et al.
Published: (2026)
by: Venktesh, V, et al.
Published: (2026)
+VeriRel: Verification Feedback to Enhance Document Retrieval for Scientific Fact Checking
by: Deng, Xingyu, et al.
Published: (2025)
by: Deng, Xingyu, et al.
Published: (2025)
Automated Justification Production for Claim Veracity in Fact Checking: A Survey on Architectures and Approaches
by: Eldifrawi, Islam, et al.
Published: (2024)
by: Eldifrawi, Islam, et al.
Published: (2024)
The Surprising Effectiveness of Rankers Trained on Expanded Queries
by: Anand, Abhijit, et al.
Published: (2024)
by: Anand, Abhijit, et al.
Published: (2024)
Understanding the User: An Intent-Based Ranking Dataset
by: Anand, Abhijit, et al.
Published: (2024)
by: Anand, Abhijit, et al.
Published: (2024)
Resolving Conflicting Evidence in Automated Fact-Checking: A Study on Retrieval-Augmented LLMs
by: Ge, Ziyu, et al.
Published: (2025)
by: Ge, Ziyu, et al.
Published: (2025)
Comparing Knowledge Sources for Open-Domain Scientific Claim Verification
by: Vladika, Juraj, et al.
Published: (2024)
by: Vladika, Juraj, et al.
Published: (2024)
Breaking the Lens of the Telescope: Online Relevance Estimation over Large Retrieval Sets
by: Rathee, Mandeep, et al.
Published: (2025)
by: Rathee, Mandeep, et al.
Published: (2025)
Reproducing Adaptive Reranking for Reasoning-Intensive IR
by: Rathee, Mandeep, et al.
Published: (2026)
by: Rathee, Mandeep, et al.
Published: (2026)
DS@GT at CheckThat! 2025: A Simple Retrieval-First, LLM-Backed Framework for Claim Normalization
by: Pramov, Aleksandar, et al.
Published: (2025)
by: Pramov, Aleksandar, et al.
Published: (2025)
Combining Evidence and Reasoning for Biomedical Fact-Checking
by: Barone, Mariano, et al.
Published: (2025)
by: Barone, Mariano, et al.
Published: (2025)
Fact Finder -- Enhancing Domain Expertise of Large Language Models by Incorporating Knowledge Graphs
by: Steinigen, Daniel, et al.
Published: (2024)
by: Steinigen, Daniel, et al.
Published: (2024)
Think Right, Not More: Test-Time Scaling for Numerical Claim Verification
by: Chungkham, Primakov, et al.
Published: (2025)
by: Chungkham, Primakov, et al.
Published: (2025)
It's High Time: A Survey of Temporal Question Answering
by: Piryani, Bhawna, et al.
Published: (2025)
by: Piryani, Bhawna, et al.
Published: (2025)
The Next Phase of Scientific Fact-Checking: Advanced Evidence Retrieval from Complex Structured Academic Papers
by: Deng, Xingyu, et al.
Published: (2025)
by: Deng, Xingyu, et al.
Published: (2025)
PARSE: An Open-Domain Reasoning Question Answering Benchmark for Persian
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions
by: Abdallah, Abdelrahman, et al.
Published: (2025)
by: Abdallah, Abdelrahman, et al.
Published: (2025)
DS@GT at CheckThat! 2025: Exploring Retrieval and Reranking Pipelines for Scientific Claim Source Retrieval on Social Media Discourse
by: Schofield, Jeanette, et al.
Published: (2025)
by: Schofield, Jeanette, et al.
Published: (2025)
UNH at CheckThat! 2025: Fine-tuning Vs Prompting in Claim Extraction
by: Wilder, Joe, et al.
Published: (2025)
by: Wilder, Joe, et al.
Published: (2025)
MultiMind at SemEval-2025 Task 7: Crosslingual Fact-Checked Claim Retrieval via Multi-Source Alignment
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking
by: Roitero, Kevin, et al.
Published: (2025)
by: Roitero, Kevin, et al.
Published: (2025)
Semi-automated Fact-checking in Portuguese: Corpora Enrichment using Retrieval with Claim extraction
by: Gomes, Juliana Resplande Sant'anna, et al.
Published: (2025)
by: Gomes, Juliana Resplande Sant'anna, et al.
Published: (2025)
Understanding Inequality of LLM Fact-Checking over Geographic Regions with Agent and Retrieval models
by: Coelho, Bruno, et al.
Published: (2025)
by: Coelho, Bruno, et al.
Published: (2025)
Large Language Models Require Curated Context for Reliable Political Fact-Checking -- Even with Reasoning and Web Search
by: DeVerna, Matthew R., et al.
Published: (2025)
by: DeVerna, Matthew R., et al.
Published: (2025)
FinVet: A Collaborative Framework of RAG and External Fact-Checking Agents for Financial Misinformation Detection
by: Araya, Daniel Berhane, et al.
Published: (2025)
by: Araya, Daniel Berhane, et al.
Published: (2025)
fact check AI at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-checked Claim Retrieval
by: Rastogi, Pranshu
Published: (2025)
by: Rastogi, Pranshu
Published: (2025)
Use of Retrieval-Augmented Large Language Model Agent for Long-Form COVID-19 Fact-Checking
by: Huang, Jingyi, et al.
Published: (2025)
by: Huang, Jingyi, et al.
Published: (2025)
Test-time Corpus Feedback: From Retrieval to RAG
by: Rathee, Mandeep, et al.
Published: (2025)
by: Rathee, Mandeep, et al.
Published: (2025)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
Peerispect: Claim Verification in Scientific Peer Reviews
by: Ghorbanpour, Ali, et al.
Published: (2026)
by: Ghorbanpour, Ali, et al.
Published: (2026)
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking
by: Akhtar, Mubashara, et al.
Published: (2024)
by: Akhtar, Mubashara, et al.
Published: (2024)
ClaimTrust: Propagation Trust Scoring for RAG Systems
by: Qian, Hangkai, et al.
Published: (2025)
by: Qian, Hangkai, et al.
Published: (2025)
AI Wizards at CheckThat! 2025: Enhancing Transformer-Based Embeddings with Sentiment for Subjectivity Detection in News Articles
by: Fasulo, Matteo, et al.
Published: (2025)
by: Fasulo, Matteo, et al.
Published: (2025)
Similar Items
-
DEXTER: A Benchmark for open-domain Complex Question Answering using LLMs
by: Prabhu, Venktesh V. Deepali, et al.
Published: (2024) -
FactIR: A Real-World Zero-shot Open-Domain Retrieval Benchmark for Fact-Checking
by: V, Venktesh, et al.
Published: (2025) -
FlashCheck: Exploration of Efficient Evidence Retrieval for Fast Fact-Checking
by: Nanekhan, Kevin, et al.
Published: (2025) -
Decomposition Dilemmas: Does Claim Decomposition Boost or Burden Fact-Checking Performance?
by: Hu, Qisheng, et al.
Published: (2024) -
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024)