Distill and Align Decomposition for Enhanced Claim Verification
Fuente:
arXiv
Salvato in:
| Autori principali: | Magomere, Jabez, Kochkina, Elena, Mensah, Samuel, Kaur, Simerjot, Acero, Fernando, Oncevay, Arturo, Smiley, Charese H., Liu, Xiaomo, Veloso, Manuela |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FinNLI: Novel Dataset for Multi-Genre Financial Natural Language Inference Benchmarking
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
A Variational Approach for Mitigating Entity Bias in Relation Extraction
di: Mensah, Samuel, et al.
Pubblicazione: (2025)
di: Mensah, Samuel, et al.
Pubblicazione: (2025)
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
di: Aguda, Toyin, et al.
Pubblicazione: (2024)
di: Aguda, Toyin, et al.
Pubblicazione: (2024)
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
Grounding LLM Reasoning with Knowledge Graphs
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025)
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025)
AI Analyst: Framework and Comprehensive Evaluation of Large Language Models for Financial Time Series Report Generation
di: Fons, Elizabeth, et al.
Pubblicazione: (2025)
di: Fons, Elizabeth, et al.
Pubblicazione: (2025)
Conservative Bias in Large Language Models: Measuring Relation Predictions
di: Aguda, Toyin, et al.
Pubblicazione: (2025)
di: Aguda, Toyin, et al.
Pubblicazione: (2025)
Scaling Crowdsourced Election Monitoring: Construction and Evaluation of Classification Models for Multilingual and Cross-Domain Classification Settings
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
When Claims Evolve: Evaluating and Enhancing the Robustness of Embedding Models Against Misinformation Edits
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research
di: Haque, Mirazul, et al.
Pubblicazione: (2026)
di: Haque, Mirazul, et al.
Pubblicazione: (2026)
Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
How Reliable are Confidence Estimators for Large Reasoning Models? A Systematic Benchmark on High-Stakes Domains
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
Calibrating LLM Confidence by Probing Perturbed Representation Stability
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2025)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2025)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
ExStrucTiny: A Benchmark for Schema-Variable Structured Information Extraction from Document Images
di: Sibue, Mathieu, et al.
Pubblicazione: (2026)
di: Sibue, Mathieu, et al.
Pubblicazione: (2026)
Optimizing Decomposition for Optimal Claim Verification
di: Lu, Yining, et al.
Pubblicazione: (2025)
di: Lu, Yining, et al.
Pubblicazione: (2025)
The Influence of Biomedical Research on Future Business Funding: Analyzing Scientific Impact and Content in Industrial Investments
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
The Alignment Bottleneck in Decomposition-Based Claim Verification
di: Akhter, Mahmud Elahi, et al.
Pubblicazione: (2026)
di: Akhter, Mahmud Elahi, et al.
Pubblicazione: (2026)
DocLLM: A layout-aware generative language model for multimodal document understanding
di: Wang, Dongsheng, et al.
Pubblicazione: (2023)
di: Wang, Dongsheng, et al.
Pubblicazione: (2023)
Perturb Your Data: Paraphrase-Guided Training Data Watermarking
di: Shetty, Pranav, et al.
Pubblicazione: (2025)
di: Shetty, Pranav, et al.
Pubblicazione: (2025)
A Claim Decomposition Benchmark for Long-form Answer Verification
di: Zhang, Zhihao, et al.
Pubblicazione: (2024)
di: Zhang, Zhihao, et al.
Pubblicazione: (2024)
"What is the value of {templates}?" Rethinking Document Information Extraction Datasets for LLMs
di: Zmigrod, Ran, et al.
Pubblicazione: (2024)
di: Zmigrod, Ran, et al.
Pubblicazione: (2024)
Pelican: Correcting Hallucination in Vision-LLMs via Claim Decomposition and Program of Thought Verification
di: Sahu, Pritish, et al.
Pubblicazione: (2024)
di: Sahu, Pritish, et al.
Pubblicazione: (2024)
MuSciClaims: Multimodal Scientific Claim Verification
di: Lal, Yash Kumar, et al.
Pubblicazione: (2025)
di: Lal, Yash Kumar, et al.
Pubblicazione: (2025)
GenPlanX. Generation of Plans and Execution
di: Borrajo, Daniel, et al.
Pubblicazione: (2025)
di: Borrajo, Daniel, et al.
Pubblicazione: (2025)
ClaimPKG: Enhancing Claim Verification via Pseudo-Subgraph Generation with Lightweight Specialized LLM
di: Pham, Hoang, et al.
Pubblicazione: (2025)
di: Pham, Hoang, et al.
Pubblicazione: (2025)
LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages
di: Bean, Andrew M., et al.
Pubblicazione: (2024)
di: Bean, Andrew M., et al.
Pubblicazione: (2024)
SciClaimEval: Cross-modal Claim Verification in Scientific Papers
di: Ho, Xanh, et al.
Pubblicazione: (2026)
di: Ho, Xanh, et al.
Pubblicazione: (2026)
MedScore: Generalizable Factuality Evaluation of Free-Form Medical Answers by Domain-adapted Claim Decomposition and Verification
di: Huang, Heyuan, et al.
Pubblicazione: (2025)
di: Huang, Heyuan, et al.
Pubblicazione: (2025)
A Closer Look at Claim Decomposition
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
Minimal Evidence Group Identification for Claim Verification
di: Li, Xiangci, et al.
Pubblicazione: (2024)
di: Li, Xiangci, et al.
Pubblicazione: (2024)
Complex Claim Verification with Evidence Retrieved in the Wild
di: Chen, Jifan, et al.
Pubblicazione: (2023)
di: Chen, Jifan, et al.
Pubblicazione: (2023)
Atomic Reasoning for Scientific Table Claim Verification
di: Zhang, Yuji, et al.
Pubblicazione: (2025)
di: Zhang, Yuji, et al.
Pubblicazione: (2025)
SciClaimHunt: A Large Dataset for Evidence-based Scientific Claim Verification
di: Kumar, Sujit, et al.
Pubblicazione: (2025)
di: Kumar, Sujit, et al.
Pubblicazione: (2025)
LETS-C: Leveraging Text Embedding for Time Series Classification
di: Kaur, Rachneet, et al.
Pubblicazione: (2024)
di: Kaur, Rachneet, et al.
Pubblicazione: (2024)
A Benchmark for Open-Domain Numerical Fact-Checking Enhanced by Claim Decomposition
di: Venktesh, V, et al.
Pubblicazione: (2025)
di: Venktesh, V, et al.
Pubblicazione: (2025)
Robust Claim Verification Through Fact Detection
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
Peerispect: Claim Verification in Scientific Peer Reviews
di: Ghorbanpour, Ali, et al.
Pubblicazione: (2026)
di: Ghorbanpour, Ali, et al.
Pubblicazione: (2026)
SynClaimEval: A Framework for Evaluating the Utility of Synthetic Data in Long-Context Claim Verification
di: Elaraby, Mohamed, et al.
Pubblicazione: (2025)
di: Elaraby, Mohamed, et al.
Pubblicazione: (2025)
Evergreen: Efficient Claim Verification for Semantic Aggregates
di: Lee, Alexander W., et al.
Pubblicazione: (2026)
di: Lee, Alexander W., et al.
Pubblicazione: (2026)
Documenti analoghi
-
FinNLI: Novel Dataset for Multi-Genre Financial Natural Language Inference Benchmarking
di: Magomere, Jabez, et al.
Pubblicazione: (2025) -
A Variational Approach for Mitigating Entity Bias in Relation Extraction
di: Mensah, Samuel, et al.
Pubblicazione: (2025) -
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
di: Aguda, Toyin, et al.
Pubblicazione: (2024) -
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline
di: Singh, Kuldeep, et al.
Pubblicazione: (2024) -
Grounding LLM Reasoning with Knowledge Graphs
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025)