BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages
Fuente:
arXiv
Saved in:
| Main Authors: | Lucas, Jason, Murtagh-White, Matt, Uchendu, Adaku, Al-Lawati, Ali, Yamashita, Michiharu, Macko, Dominik, Srba, Ivan, Moro, Robert, Lee, Dongwon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects
by: Lucas, Jason, et al.
Published: (2026)
by: Lucas, Jason, et al.
Published: (2026)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
by: Macko, Dominik, et al.
Published: (2023)
by: Macko, Dominik, et al.
Published: (2023)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
by: Macko, Dominik, et al.
Published: (2025)
by: Macko, Dominik, et al.
Published: (2025)
A Ship of Theseus: Curious Cases of Paraphrasing in LLM-Generated Texts
by: Tripto, Nafis Irtiza, et al.
Published: (2023)
by: Tripto, Nafis Irtiza, et al.
Published: (2023)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
by: Macko, Dominik, et al.
Published: (2025)
by: Macko, Dominik, et al.
Published: (2025)
GPT-who: An Information Density-based Machine-Generated Text Detector
by: Venkatraman, Saranya, et al.
Published: (2023)
by: Venkatraman, Saranya, et al.
Published: (2023)
TOPFORMER: Topology-Aware Authorship Attribution of Deepfake Texts with Diverse Writing Styles
by: Uchendu, Adaku, et al.
Published: (2023)
by: Uchendu, Adaku, et al.
Published: (2023)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
LLM Benchmark Datasets Should Be Contamination-Resistant
by: Al-Lawati, Ali, et al.
Published: (2026)
by: Al-Lawati, Ali, et al.
Published: (2026)
Topological Data Analysis Applications in Natural Language Processing: A Survey
by: Uchendu, Adaku, et al.
Published: (2024)
by: Uchendu, Adaku, et al.
Published: (2024)
Fake Resume Attacks: Data Poisoning on Online Job Platforms
by: Yamashita, Michiharu, et al.
Published: (2024)
by: Yamashita, Michiharu, et al.
Published: (2024)
Authorship Attribution in Multilingual Machine-Generated Texts
by: La Cava, Lucio, et al.
Published: (2025)
by: La Cava, Lucio, et al.
Published: (2025)
Disinformation Capabilities of Large Language Models
by: Vykopal, Ivan, et al.
Published: (2023)
by: Vykopal, Ivan, et al.
Published: (2023)
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection
by: Lee, Jooyoung, et al.
Published: (2024)
by: Lee, Jooyoung, et al.
Published: (2024)
Beemo: Benchmark of Expert-edited Machine-generated Outputs
by: Artemova, Ekaterina, et al.
Published: (2024)
by: Artemova, Ekaterina, et al.
Published: (2024)
Signature vs. Substance: Evaluating the Balance of Adversarial Resistance and Linguistic Quality in Watermarking Large Language Models
by: Guo, William, et al.
Published: (2025)
by: Guo, William, et al.
Published: (2025)
Unmasking Fake Careers: Detecting Machine-Generated Career Trajectories via Multi-layer Heterogeneous Graphs
by: Yamashita, Michiharu, et al.
Published: (2025)
by: Yamashita, Michiharu, et al.
Published: (2025)
Semantic Captioning: Benchmark Dataset and Graph-Aware Few-Shot In-Context Learning for SQL2Text
by: Al-Lawati, Ali, et al.
Published: (2025)
by: Al-Lawati, Ali, et al.
Published: (2025)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
by: Zugecova, Aneta, et al.
Published: (2024)
by: Zugecova, Aneta, et al.
Published: (2024)
Moltbook Moderation: Uncovering Hidden Intent Through Multi-Turn Dialogue
by: Al-Lawati, Ali, et al.
Published: (2026)
by: Al-Lawati, Ali, et al.
Published: (2026)
CAPER: Enhancing Career Trajectory Prediction using Temporal Knowledge Graph and Ternary Relationship
by: Lee, Yeon-Chang, et al.
Published: (2024)
by: Lee, Yeon-Chang, et al.
Published: (2024)
IMGTB: A Framework for Machine-Generated Text Detection Benchmarking
by: Spiegel, Michal, et al.
Published: (2023)
by: Spiegel, Michal, et al.
Published: (2023)
CEAID: Benchmark of Multilingual Machine-Generated Text Detection Methods for Central European Languages
by: Macko, Dominik, et al.
Published: (2025)
by: Macko, Dominik, et al.
Published: (2025)
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
by: Hyben, Martin, et al.
Published: (2023)
by: Hyben, Martin, et al.
Published: (2023)
Evaluation of Multilingual LLMs Personalized Text Generation Capabilities Targeting Groups and Social-Media Platforms
by: Macko, Dominik
Published: (2026)
by: Macko, Dominik
Published: (2026)
mdok of KInIT: Robustly Fine-tuned LLM for Binary and Multiclass AI-Generated Text Detection
by: Macko, Dominik
Published: (2025)
by: Macko, Dominik
Published: (2025)
mdok-style at SemEval-2026 Task 10: Finetuning LLMs for Conspiracy Detection
by: Macko, Dominik
Published: (2026)
by: Macko, Dominik
Published: (2026)
MultiCW: A Large-Scale Balanced Benchmark Dataset for Training Robust Check-Worthiness Detection Models
by: Hyben, Martin, et al.
Published: (2026)
by: Hyben, Martin, et al.
Published: (2026)
Graph-based Molecular In-context Learning Grounded on Morgan Fingerprints
by: Al-Lawati, Ali, et al.
Published: (2025)
by: Al-Lawati, Ali, et al.
Published: (2025)
Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification
by: Pecher, Branislav, et al.
Published: (2026)
by: Pecher, Branislav, et al.
Published: (2026)
Autonomation, Not Automation: Activities and Needs of European Fact-checkers as a Basis for Designing Human-Centered AI Systems
by: Hrckova, Andrea, et al.
Published: (2022)
by: Hrckova, Andrea, et al.
Published: (2022)
Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks
by: Al-Lawati, Ali, et al.
Published: (2026)
by: Al-Lawati, Ali, et al.
Published: (2026)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
by: Spiegel, Michal, et al.
Published: (2024)
by: Spiegel, Michal, et al.
Published: (2024)
PerQ: Efficient Evaluation of Multilingual Text Personalization Quality
by: Macko, Dominik, et al.
Published: (2025)
by: Macko, Dominik, et al.
Published: (2025)
RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets
by: Cegin, Jan, et al.
Published: (2025)
by: Cegin, Jan, et al.
Published: (2025)
PEFT-Bench: A Parameter-Efficient Fine-Tuning Methods Benchmark
by: Belanec, Robert, et al.
Published: (2025)
by: Belanec, Robert, et al.
Published: (2025)
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models
by: Ye, Yiran, et al.
Published: (2023)
by: Ye, Yiran, et al.
Published: (2023)
Masks and Mimicry: Strategic Obfuscation and Impersonation Attacks on Authorship Verification
by: Alperin, Kenneth, et al.
Published: (2025)
by: Alperin, Kenneth, et al.
Published: (2025)
Producción científica de la psicología vinculada a pequeños productores agropecuarios con énfasis en el ámbito del desarrollo rural
by: Sofía Murtagh
Published: (2011)
by: Sofía Murtagh
Published: (2011)
Similar Items
-
DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects
by: Lucas, Jason, et al.
Published: (2026) -
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
by: Macko, Dominik, et al.
Published: (2023) -
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024) -
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
by: Macko, Dominik, et al.
Published: (2025) -
A Ship of Theseus: Curious Cases of Paraphrasing in LLM-Generated Texts
by: Tripto, Nafis Irtiza, et al.
Published: (2023)