A Guide to Misinformation Detection Data and Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Thibault, Camille, Tian, Jacob-Junqi, Peloquin-Skulski, Gabrielle, Curtis, Taylor Lynn, Zhou, James, Laflamme, Florence, Guan, Yuxiang, Rabbany, Reihaneh, Godbout, Jean-François, Pelrine, Kellin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uncertainty Resolution in Misinformation Detection
by: Orlovskiy, Yury, et al.
Published: (2024)
by: Orlovskiy, Yury, et al.
Published: (2024)
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation
by: Vergho, Tyler, et al.
Published: (2024)
by: Vergho, Tyler, et al.
Published: (2024)
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
by: Rivera, Mauricio, et al.
Published: (2024)
by: Rivera, Mauricio, et al.
Published: (2024)
Web Retrieval Agents for Evidence-Based Misinformation Detection
by: Tian, Jacob-Junqi, et al.
Published: (2024)
by: Tian, Jacob-Junqi, et al.
Published: (2024)
Epistemic Integrity in Large Language Models
by: Ghafouri, Bijean, et al.
Published: (2024)
by: Ghafouri, Bijean, et al.
Published: (2024)
Veracity: An Open-Source AI Fact-Checking System
by: Curtis, Taylor Lynn, et al.
Published: (2025)
by: Curtis, Taylor Lynn, et al.
Published: (2025)
Online Influence Campaigns: Strategies and Vulnerabilities
by: Musulan, Andreea, et al.
Published: (2024)
by: Musulan, Andreea, et al.
Published: (2024)
$\texttt{BluePrint}$: A Social Media User Dataset for LLM Persona Evaluation and Training
by: Bück-Kaeffer, Aurélien, et al.
Published: (2025)
by: Bück-Kaeffer, Aurélien, et al.
Published: (2025)
What do people want to fact-check?
by: Ghafouri, Bijean, et al.
Published: (2026)
by: Ghafouri, Bijean, et al.
Published: (2026)
From Intuition to Understanding: Using AI Peers to Overcome Physics Misconceptions
by: Weijers, Ruben, et al.
Published: (2025)
by: Weijers, Ruben, et al.
Published: (2025)
CrediBench: Building Web-Scale Network Datasets for Information Integrity
by: Kondrup, Emma, et al.
Published: (2025)
by: Kondrup, Emma, et al.
Published: (2025)
Towards Detecting Contextual Real-Time Toxicity for In-Game Chat
by: Yang, Zachary, et al.
Published: (2023)
by: Yang, Zachary, et al.
Published: (2023)
Emerging Vulnerabilities in Frontier Models: Multi-Turn Jailbreak Attacks
by: Gibbs, Tom, et al.
Published: (2024)
by: Gibbs, Tom, et al.
Published: (2024)
Regional and Temporal Patterns of Partisan Polarization during the COVID-19 Pandemic in the United States and Canada
by: Yang, Zachary, et al.
Published: (2024)
by: Yang, Zachary, et al.
Published: (2024)
A Simulation System Towards Solving Societal-Scale Manipulation
by: Touzel, Maximilian Puelma, et al.
Published: (2024)
by: Touzel, Maximilian Puelma, et al.
Published: (2024)
The Structural Safety Generalization Problem
by: Broomfield, Julius, et al.
Published: (2025)
by: Broomfield, Julius, et al.
Published: (2025)
Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks
by: Struppek, Lukas, et al.
Published: (2026)
by: Struppek, Lukas, et al.
Published: (2026)
Deepfakes in the 2025 Canadian Election: Prevalence, Partisanship, and Platform Dynamics
by: Livernoche, Victor, et al.
Published: (2025)
by: Livernoche, Victor, et al.
Published: (2025)
GPS-SSL: Guided Positive Sampling to Inject Prior Into Self-Supervised Learning
by: Feizi, Aarash, et al.
Published: (2024)
by: Feizi, Aarash, et al.
Published: (2024)
Accidental Vulnerability: Factors in Fine-Tuning that Shift Model Safeguards
by: Pandey, Punya Syon, et al.
Published: (2025)
by: Pandey, Punya Syon, et al.
Published: (2025)
OpenFake: An Open Dataset and Platform Toward Real-World Deepfake Detection
by: Livernoche, Victor, et al.
Published: (2025)
by: Livernoche, Victor, et al.
Published: (2025)
Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection
by: Yang, Zachary, et al.
Published: (2025)
by: Yang, Zachary, et al.
Published: (2025)
Jailbreak-Tuning: Models Efficiently Learn Jailbreak Susceptibility
by: Murphy, Brendan, et al.
Published: (2025)
by: Murphy, Brendan, et al.
Published: (2025)
Ask before you Build: Rethinking AI-for-Good in Human Trafficking Interventions
by: Nair, Pratheeksha, et al.
Published: (2025)
by: Nair, Pratheeksha, et al.
Published: (2025)
Exploiting Novel GPT-4 APIs
by: Pelrine, Kellin, et al.
Published: (2023)
by: Pelrine, Kellin, et al.
Published: (2023)
Enhancing Privacy in the Early Detection of Sexual Predators Through Federated Learning and Differential Privacy
by: Chehbouni, Khaoula, et al.
Published: (2025)
by: Chehbouni, Khaoula, et al.
Published: (2025)
The $\textit{Silicon Society}$ Cookbook: Design Space of LLM-based Social Simulations
by: Bück-Kaeffer, Aurélien, et al.
Published: (2026)
by: Bück-Kaeffer, Aurélien, et al.
Published: (2026)
EASE Configuration Facilitates A Reproducible Science of LLM Social Simulations
by: Sarangi, Sneheel, et al.
Published: (2026)
by: Sarangi, Sneheel, et al.
Published: (2026)
Sowing 'Seeds of Doubt': Cottage Industries of Election and Medical Misinformation in Brazil and the United States
by: Hassoun, Amelia, et al.
Published: (2023)
by: Hassoun, Amelia, et al.
Published: (2023)
Hallucination Detox: Sensitivity Dropout (SenD) for Large Language Model Training
by: Mohammadzadeh, Shahrad, et al.
Published: (2024)
by: Mohammadzadeh, Shahrad, et al.
Published: (2024)
Kurtosis-Guided Denoising Score Matching for Tabular Anomaly Detection
by: Livernoche, Victor, et al.
Published: (2026)
by: Livernoche, Victor, et al.
Published: (2026)
Weak Supervision for Real World Graphs
by: Nair, Pratheeksha, et al.
Published: (2025)
by: Nair, Pratheeksha, et al.
Published: (2025)
Large language models can effectively convince people to believe conspiracies
by: Costello, Thomas H., et al.
Published: (2026)
by: Costello, Thomas H., et al.
Published: (2026)
It's the Thought that Counts: Evaluating the Attempts of Frontier LLMs to Persuade on Harmful Topics
by: Kowal, Matthew, et al.
Published: (2025)
by: Kowal, Matthew, et al.
Published: (2025)
Are Large Language Models Good Temporal Graph Learners?
by: Huang, Shenyang, et al.
Published: (2025)
by: Huang, Shenyang, et al.
Published: (2025)
PairBench: Are Vision-Language Models Reliable at Comparing What They See?
by: Feizi, Aarash, et al.
Published: (2025)
by: Feizi, Aarash, et al.
Published: (2025)
Robust Misinformation Detection by Visiting Potential Commonsense Conflict
by: Wang, Bing, et al.
Published: (2025)
by: Wang, Bing, et al.
Published: (2025)
The Commodification of AI Sovereignty: Lessons from the Fight for Sovereign Oil
by: Yew, Rui-Jie, et al.
Published: (2026)
by: Yew, Rui-Jie, et al.
Published: (2026)
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
by: Wang, Bing, et al.
Published: (2024)
by: Wang, Bing, et al.
Published: (2024)
Stability of Extrinsic Cohesive-Zone Model with Penalty-Based Contact in Explicit Dynamic Fragmentation Simulations
by: Ghesquière-Diérickx, Thibault, et al.
Published: (2025)
by: Ghesquière-Diérickx, Thibault, et al.
Published: (2025)
Similar Items
-
Uncertainty Resolution in Misinformation Detection
by: Orlovskiy, Yury, et al.
Published: (2024) -
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation
by: Vergho, Tyler, et al.
Published: (2024) -
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
by: Rivera, Mauricio, et al.
Published: (2024) -
Web Retrieval Agents for Evidence-Based Misinformation Detection
by: Tian, Jacob-Junqi, et al.
Published: (2024) -
Epistemic Integrity in Large Language Models
by: Ghafouri, Bijean, et al.
Published: (2024)