Yesterday's News: Benchmarking Multi-Dimensional Out-of-Distribution Generalization of Misinformation Detection Models
Fuente:
arXiv
Saved in:
| Main Authors: | Verhoeven, Ivo, Mishra, Pushkar, Shutova, Ekaterina |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A (More) Realistic Evaluation Setup for Generalisation of Community Models on Malicious Content Detection
by: Verhoeven, Ivo, et al.
Published: (2024)
by: Verhoeven, Ivo, et al.
Published: (2024)
Multimodal Misinformation Detection using Large Vision-Language Models
by: Tahmasebi, Sahar, et al.
Published: (2024)
by: Tahmasebi, Sahar, et al.
Published: (2024)
MiDe22: An Annotated Multi-Event Tweet Dataset for Misinformation Detection
by: Toraman, Cagri, et al.
Published: (2022)
by: Toraman, Cagri, et al.
Published: (2022)
Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification
by: Barone, Mariano, et al.
Published: (2025)
by: Barone, Mariano, et al.
Published: (2025)
When Should Dense Retrievers Be Updated in Evolving Corpora? Detecting Out-of-Distribution Corpora Using GradNormIR
by: Ko, Dayoon, et al.
Published: (2025)
by: Ko, Dayoon, et al.
Published: (2025)
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation
by: Cheng, Yiruo, et al.
Published: (2024)
by: Cheng, Yiruo, et al.
Published: (2024)
Zero-Shot Warning Generation for Misinformative Multimodal Content
by: Delvecchio, Giovanni Pio, et al.
Published: (2025)
by: Delvecchio, Giovanni Pio, et al.
Published: (2025)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024)
by: Hoa, Tran Thai, et al.
Published: (2024)
EXCLAIM: An Explainable Cross-Modal Agentic System for Misinformation Detection with Hierarchical Retrieval
by: Wu, Yin, et al.
Published: (2025)
by: Wu, Yin, et al.
Published: (2025)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
by: Li, Haitao, et al.
Published: (2025)
by: Li, Haitao, et al.
Published: (2025)
StructText: A Synthetic Table-to-Text Approach for Benchmark Generation with Multi-Dimensional Evaluation
by: Kashyap, Satyananda, et al.
Published: (2025)
by: Kashyap, Satyananda, et al.
Published: (2025)
ClarifyMT-Bench: Benchmarking and Improving Multi-Turn Clarification for Conversational Large Language Models
by: Luo, Sichun, et al.
Published: (2025)
by: Luo, Sichun, et al.
Published: (2025)
Multi-Layer Ranking with Large Language Models for News Source Recommendation
by: Zhang, Wenjia, et al.
Published: (2024)
by: Zhang, Wenjia, et al.
Published: (2024)
A Survey on Proactive Defense Strategies Against Misinformation in Large Language Models
by: Liu, Shuliang, et al.
Published: (2025)
by: Liu, Shuliang, et al.
Published: (2025)
FinVet: A Collaborative Framework of RAG and External Fact-Checking Agents for Financial Misinformation Detection
by: Araya, Daniel Berhane, et al.
Published: (2025)
by: Araya, Daniel Berhane, et al.
Published: (2025)
Generative Multi-Modal Knowledge Retrieval with Large Language Models
by: Long, Xinwei, et al.
Published: (2024)
by: Long, Xinwei, et al.
Published: (2024)
SynDy: Synthetic Dynamic Dataset Generation Framework for Misinformation Tasks
by: Shliselberg, Michael, et al.
Published: (2024)
by: Shliselberg, Michael, et al.
Published: (2024)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Evidence-Grounded Multimodal Misinformation Detection with Attention-Based GNNs
by: Duwal, Sharad, et al.
Published: (2025)
by: Duwal, Sharad, et al.
Published: (2025)
MCiteBench: A Multimodal Benchmark for Generating Text with Citations
by: Hu, Caiyu, et al.
Published: (2025)
by: Hu, Caiyu, et al.
Published: (2025)
Seed-Guided Topic Discovery with Out-of-Vocabulary Seeds
by: Zhang, Yu, et al.
Published: (2022)
by: Zhang, Yu, et al.
Published: (2022)
GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing
by: Hu, Silan, et al.
Published: (2025)
by: Hu, Silan, et al.
Published: (2025)
Breaking It Down: Domain-Aware Semantic Segmentation for Retrieval Augmented Generation
by: Allamraju, Aparajitha, et al.
Published: (2025)
by: Allamraju, Aparajitha, et al.
Published: (2025)
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana
by: Filice, Simone, et al.
Published: (2025)
by: Filice, Simone, et al.
Published: (2025)
Building Russian Benchmark for Evaluation of Information Retrieval Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope Questions
by: Peng, Zhiyuan, et al.
Published: (2024)
by: Peng, Zhiyuan, et al.
Published: (2024)
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
by: Xi, Yunjia, et al.
Published: (2025)
by: Xi, Yunjia, et al.
Published: (2025)
Diagnosing LLM-based Rerankers in Cold-Start Recommender Systems: Coverage, Exposure and Practical Mitigations
by: Lemdiasova, Ekaterina, et al.
Published: (2026)
by: Lemdiasova, Ekaterina, et al.
Published: (2026)
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking
by: Qiao, Yixuan, et al.
Published: (2022)
by: Qiao, Yixuan, et al.
Published: (2022)
Detecting Generated Native Ads in Conversational Search
by: Schmidt, Sebastian, et al.
Published: (2024)
by: Schmidt, Sebastian, et al.
Published: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
CliniQ: A Multi-faceted Benchmark for Electronic Health Record Retrieval with Semantic Match Assessment
by: Zhao, Zhengyun, et al.
Published: (2025)
by: Zhao, Zhengyun, et al.
Published: (2025)
QP-OneModel: A Unified Generative LLM for Multi-Task Query Understanding in Xiaohongshu Search
by: Huang, Jianzhao, et al.
Published: (2026)
by: Huang, Jianzhao, et al.
Published: (2026)
Multi-Record Web Page Information Extraction From News Websites
by: Kustenkov, Alexander, et al.
Published: (2025)
by: Kustenkov, Alexander, et al.
Published: (2025)
Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering
by: Shi, Zhengliang, et al.
Published: (2024)
by: Shi, Zhengliang, et al.
Published: (2024)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Similar Items
-
A (More) Realistic Evaluation Setup for Generalisation of Community Models on Malicious Content Detection
by: Verhoeven, Ivo, et al.
Published: (2024) -
Multimodal Misinformation Detection using Large Vision-Language Models
by: Tahmasebi, Sahar, et al.
Published: (2024) -
MiDe22: An Annotated Multi-Event Tweet Dataset for Misinformation Detection
by: Toraman, Cagri, et al.
Published: (2022) -
Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification
by: Barone, Mariano, et al.
Published: (2025) -
When Should Dense Retrievers Be Updated in Evolving Corpora? Detecting Out-of-Distribution Corpora Using GradNormIR
by: Ko, Dayoon, et al.
Published: (2025)