MultiCW: A Large-Scale Balanced Benchmark Dataset for Training Robust Check-Worthiness Detection Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hyben, Martin, Kula, Sebastian, Cegin, Jan, Simko, Jakub, Srba, Ivan, Moro, Robert |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
von: Hyben, Martin, et al.
Veröffentlicht: (2023)
von: Hyben, Martin, et al.
Veröffentlicht: (2023)
RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets
von: Cegin, Jan, et al.
Veröffentlicht: (2025)
von: Cegin, Jan, et al.
Veröffentlicht: (2025)
A Generative-AI-Driven Claim Retrieval System Capable of Detecting and Retrieving Claims from Social Media Platforms in Multiple Languages
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
Fighting Randomness with Randomness: Mitigating Optimisation Instability of Fine-Tuning using Delayed Ensemble and Noisy Interpolation
von: Pecher, Branislav, et al.
Veröffentlicht: (2024)
von: Pecher, Branislav, et al.
Veröffentlicht: (2024)
Multilingual Models for Check-Worthy Social Media Posts Detection
von: Kula, Sebastian, et al.
Veröffentlicht: (2024)
von: Kula, Sebastian, et al.
Veröffentlicht: (2024)
Effects of diversity incentives on sample diversity and downstream model performance in LLM-based text augmentation
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
Use Random Selection for Now: Investigation of Few-Shot Selection Strategies in LLM-based Text Augmentation for Classification
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
LLMs vs Established Text Augmentation Techniques for Classification: When do the Benefits Outweight the Costs?
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
von: Cegin, Jan, et al.
Veröffentlicht: (2024)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
Autonomation, Not Automation: Activities and Needs of European Fact-checkers as a Basis for Designing Human-Centered AI Systems
von: Hrckova, Andrea, et al.
Veröffentlicht: (2022)
von: Hrckova, Andrea, et al.
Veröffentlicht: (2022)
Multilingual Previously Fact-Checked Claim Retrieval
von: Pikuliak, Matúš, et al.
Veröffentlicht: (2023)
von: Pikuliak, Matúš, et al.
Veröffentlicht: (2023)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
von: Macko, Dominik, et al.
Veröffentlicht: (2023)
von: Macko, Dominik, et al.
Veröffentlicht: (2023)
A Rigorous Evaluation of LLM Data Generation Strategies for Low-Resource Languages
von: Anikina, Tatiana, et al.
Veröffentlicht: (2025)
von: Anikina, Tatiana, et al.
Veröffentlicht: (2025)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
Claim Check-Worthiness Detection: How Well do LLMs Grasp Annotation Guidelines?
von: Majer, Laura, et al.
Veröffentlicht: (2024)
von: Majer, Laura, et al.
Veröffentlicht: (2024)
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
Large Language Models for Multilingual Previously Fact-Checked Claim Detection
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
Disinformation Capabilities of Large Language Models
von: Vykopal, Ivan, et al.
Veröffentlicht: (2023)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2023)
DelTriC: A Novel Clustering Method with Accurate Outlier
von: Javurek, Tomas, et al.
Veröffentlicht: (2025)
von: Javurek, Tomas, et al.
Veröffentlicht: (2025)
Generative Large Language Models in Automated Fact-Checking: A Survey
von: Vykopal, Ivan, et al.
Veröffentlicht: (2024)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2024)
Revisiting Prompt Sensitivity in Large Language Models for Text Classification: The Role of Prompt Underspecification
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
von: Zugecova, Aneta, et al.
Veröffentlicht: (2024)
von: Zugecova, Aneta, et al.
Veröffentlicht: (2024)
Application and Optimization of Large Models Based on Prompt Tuning for Fact-Check-Worthiness Estimation
von: Yu, Yinglong, et al.
Veröffentlicht: (2025)
von: Yu, Yinglong, et al.
Veröffentlicht: (2025)
PEFT-Factory: Unified Parameter-Efficient Fine-Tuning of Autoregressive Large Language Models
von: Belanec, Robert, et al.
Veröffentlicht: (2025)
von: Belanec, Robert, et al.
Veröffentlicht: (2025)
Interpretable Predictability-Based AI Text Detection: A Replication Study
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
Political Leaning and Politicalness Classification of Texts
von: Volf, Matous, et al.
Veröffentlicht: (2025)
von: Volf, Matous, et al.
Veröffentlicht: (2025)
HYBRINFOX at CheckThat! 2024 -- Task 1: Enhancing Language Models with Structured Information for Check-Worthiness Estimation
von: Faye, Géraud, et al.
Veröffentlicht: (2024)
von: Faye, Géraud, et al.
Veröffentlicht: (2024)
mdok-style at SemEval-2026 Task 9: Finetuning LLMs for Multilingual Polarization Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
PEFT-Bench: A Parameter-Efficient Fine-Tuning Methods Benchmark
von: Belanec, Robert, et al.
Veröffentlicht: (2025)
von: Belanec, Robert, et al.
Veröffentlicht: (2025)
BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages
von: Lucas, Jason, et al.
Veröffentlicht: (2026)
von: Lucas, Jason, et al.
Veröffentlicht: (2026)
Authorship Attribution in Multilingual Machine-Generated Texts
von: La Cava, Lucio, et al.
Veröffentlicht: (2025)
von: La Cava, Lucio, et al.
Veröffentlicht: (2025)
Investigating Language and Retrieval Bias in Multilingual Previously Fact-Checked Claim Detection
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025)
The Compressed Oracle is a Worthy (Multiplicative) Adversary
von: Jeffery, Stacey, et al.
Veröffentlicht: (2025)
von: Jeffery, Stacey, et al.
Veröffentlicht: (2025)
Bark beetles on logging residues of European larch: Effects of shading and diameter of logging residues on infestation density
von: Jakub Špoula, et al.
Veröffentlicht: (2024)
von: Jakub Špoula, et al.
Veröffentlicht: (2024)
Task Prompt Vectors: Effective Initialization through Multi-Task Soft-Prompt Transfer
von: Belanec, Robert, et al.
Veröffentlicht: (2024)
von: Belanec, Robert, et al.
Veröffentlicht: (2024)
TAPAAL SMC: Statistical Model Checking of Stochastic Timed-Arc Petri Nets
von: Dubois, Tanguy, et al.
Veröffentlicht: (2026)
von: Dubois, Tanguy, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
von: Hyben, Martin, et al.
Veröffentlicht: (2023) -
RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets
von: Cegin, Jan, et al.
Veröffentlicht: (2025) -
A Generative-AI-Driven Claim Retrieval System Capable of Detecting and Retrieving Claims from Social Media Platforms in Multiple Languages
von: Vykopal, Ivan, et al.
Veröffentlicht: (2025) -
Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification
von: Pecher, Branislav, et al.
Veröffentlicht: (2026) -
Fighting Randomness with Randomness: Mitigating Optimisation Instability of Fine-Tuning using Delayed Ensemble and Noisy Interpolation
von: Pecher, Branislav, et al.
Veröffentlicht: (2024)