BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages
Fuente:
arXiv
Guardado en:
| Autores principales: | Lucas, Jason, Murtagh-White, Matt, Uchendu, Adaku, Al-Lawati, Ali, Yamashita, Michiharu, Macko, Dominik, Srba, Ivan, Moro, Robert, Lee, Dongwon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects
por: Lucas, Jason, et al.
Publicado: (2026)
por: Lucas, Jason, et al.
Publicado: (2026)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
por: Macko, Dominik, et al.
Publicado: (2023)
por: Macko, Dominik, et al.
Publicado: (2023)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
por: Macko, Dominik, et al.
Publicado: (2024)
por: Macko, Dominik, et al.
Publicado: (2024)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
por: Macko, Dominik, et al.
Publicado: (2025)
por: Macko, Dominik, et al.
Publicado: (2025)
A Ship of Theseus: Curious Cases of Paraphrasing in LLM-Generated Texts
por: Tripto, Nafis Irtiza, et al.
Publicado: (2023)
por: Tripto, Nafis Irtiza, et al.
Publicado: (2023)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
por: Macko, Dominik, et al.
Publicado: (2025)
por: Macko, Dominik, et al.
Publicado: (2025)
GPT-who: An Information Density-based Machine-Generated Text Detector
por: Venkatraman, Saranya, et al.
Publicado: (2023)
por: Venkatraman, Saranya, et al.
Publicado: (2023)
TOPFORMER: Topology-Aware Authorship Attribution of Deepfake Texts with Diverse Writing Styles
por: Uchendu, Adaku, et al.
Publicado: (2023)
por: Uchendu, Adaku, et al.
Publicado: (2023)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
por: Macko, Dominik, et al.
Publicado: (2024)
por: Macko, Dominik, et al.
Publicado: (2024)
LLM Benchmark Datasets Should Be Contamination-Resistant
por: Al-Lawati, Ali, et al.
Publicado: (2026)
por: Al-Lawati, Ali, et al.
Publicado: (2026)
Topological Data Analysis Applications in Natural Language Processing: A Survey
por: Uchendu, Adaku, et al.
Publicado: (2024)
por: Uchendu, Adaku, et al.
Publicado: (2024)
Fake Resume Attacks: Data Poisoning on Online Job Platforms
por: Yamashita, Michiharu, et al.
Publicado: (2024)
por: Yamashita, Michiharu, et al.
Publicado: (2024)
Authorship Attribution in Multilingual Machine-Generated Texts
por: La Cava, Lucio, et al.
Publicado: (2025)
por: La Cava, Lucio, et al.
Publicado: (2025)
Disinformation Capabilities of Large Language Models
por: Vykopal, Ivan, et al.
Publicado: (2023)
por: Vykopal, Ivan, et al.
Publicado: (2023)
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection
por: Lee, Jooyoung, et al.
Publicado: (2024)
por: Lee, Jooyoung, et al.
Publicado: (2024)
Beemo: Benchmark of Expert-edited Machine-generated Outputs
por: Artemova, Ekaterina, et al.
Publicado: (2024)
por: Artemova, Ekaterina, et al.
Publicado: (2024)
Signature vs. Substance: Evaluating the Balance of Adversarial Resistance and Linguistic Quality in Watermarking Large Language Models
por: Guo, William, et al.
Publicado: (2025)
por: Guo, William, et al.
Publicado: (2025)
Unmasking Fake Careers: Detecting Machine-Generated Career Trajectories via Multi-layer Heterogeneous Graphs
por: Yamashita, Michiharu, et al.
Publicado: (2025)
por: Yamashita, Michiharu, et al.
Publicado: (2025)
Semantic Captioning: Benchmark Dataset and Graph-Aware Few-Shot In-Context Learning for SQL2Text
por: Al-Lawati, Ali, et al.
Publicado: (2025)
por: Al-Lawati, Ali, et al.
Publicado: (2025)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
por: Zugecova, Aneta, et al.
Publicado: (2024)
por: Zugecova, Aneta, et al.
Publicado: (2024)
Moltbook Moderation: Uncovering Hidden Intent Through Multi-Turn Dialogue
por: Al-Lawati, Ali, et al.
Publicado: (2026)
por: Al-Lawati, Ali, et al.
Publicado: (2026)
CAPER: Enhancing Career Trajectory Prediction using Temporal Knowledge Graph and Ternary Relationship
por: Lee, Yeon-Chang, et al.
Publicado: (2024)
por: Lee, Yeon-Chang, et al.
Publicado: (2024)
IMGTB: A Framework for Machine-Generated Text Detection Benchmarking
por: Spiegel, Michal, et al.
Publicado: (2023)
por: Spiegel, Michal, et al.
Publicado: (2023)
CEAID: Benchmark of Multilingual Machine-Generated Text Detection Methods for Central European Languages
por: Macko, Dominik, et al.
Publicado: (2025)
por: Macko, Dominik, et al.
Publicado: (2025)
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
por: Hyben, Martin, et al.
Publicado: (2023)
por: Hyben, Martin, et al.
Publicado: (2023)
Evaluation of Multilingual LLMs Personalized Text Generation Capabilities Targeting Groups and Social-Media Platforms
por: Macko, Dominik
Publicado: (2026)
por: Macko, Dominik
Publicado: (2026)
mdok of KInIT: Robustly Fine-tuned LLM for Binary and Multiclass AI-Generated Text Detection
por: Macko, Dominik
Publicado: (2025)
por: Macko, Dominik
Publicado: (2025)
mdok-style at SemEval-2026 Task 10: Finetuning LLMs for Conspiracy Detection
por: Macko, Dominik
Publicado: (2026)
por: Macko, Dominik
Publicado: (2026)
MultiCW: A Large-Scale Balanced Benchmark Dataset for Training Robust Check-Worthiness Detection Models
por: Hyben, Martin, et al.
Publicado: (2026)
por: Hyben, Martin, et al.
Publicado: (2026)
Graph-based Molecular In-context Learning Grounded on Morgan Fingerprints
por: Al-Lawati, Ali, et al.
Publicado: (2025)
por: Al-Lawati, Ali, et al.
Publicado: (2025)
Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification
por: Pecher, Branislav, et al.
Publicado: (2026)
por: Pecher, Branislav, et al.
Publicado: (2026)
Autonomation, Not Automation: Activities and Needs of European Fact-checkers as a Basis for Designing Human-Centered AI Systems
por: Hrckova, Andrea, et al.
Publicado: (2022)
por: Hrckova, Andrea, et al.
Publicado: (2022)
Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks
por: Al-Lawati, Ali, et al.
Publicado: (2026)
por: Al-Lawati, Ali, et al.
Publicado: (2026)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
por: Spiegel, Michal, et al.
Publicado: (2024)
por: Spiegel, Michal, et al.
Publicado: (2024)
PerQ: Efficient Evaluation of Multilingual Text Personalization Quality
por: Macko, Dominik, et al.
Publicado: (2025)
por: Macko, Dominik, et al.
Publicado: (2025)
RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets
por: Cegin, Jan, et al.
Publicado: (2025)
por: Cegin, Jan, et al.
Publicado: (2025)
PEFT-Bench: A Parameter-Efficient Fine-Tuning Methods Benchmark
por: Belanec, Robert, et al.
Publicado: (2025)
por: Belanec, Robert, et al.
Publicado: (2025)
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models
por: Ye, Yiran, et al.
Publicado: (2023)
por: Ye, Yiran, et al.
Publicado: (2023)
Masks and Mimicry: Strategic Obfuscation and Impersonation Attacks on Authorship Verification
por: Alperin, Kenneth, et al.
Publicado: (2025)
por: Alperin, Kenneth, et al.
Publicado: (2025)
Producción científica de la psicología vinculada a pequeños productores agropecuarios con énfasis en el ámbito del desarrollo rural
por: Sofía Murtagh
Publicado: (2011)
por: Sofía Murtagh
Publicado: (2011)
Ejemplares similares
-
DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects
por: Lucas, Jason, et al.
Publicado: (2026) -
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
por: Macko, Dominik, et al.
Publicado: (2023) -
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
por: Macko, Dominik, et al.
Publicado: (2024) -
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
por: Macko, Dominik, et al.
Publicado: (2025) -
A Ship of Theseus: Curious Cases of Paraphrasing in LLM-Generated Texts
por: Tripto, Nafis Irtiza, et al.
Publicado: (2023)