Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch
Fuente:
arXiv
Saved in:
| Main Authors: | Strazda, Elza, Spanakis, Gerasimos |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Across Platforms and Languages: Dutch Influencers and Legal Disclosures on Instagram, YouTube and TikTok
by: Gui, Haoyang, et al.
Published: (2024)
by: Gui, Haoyang, et al.
Published: (2024)
From FusHa to Folk: Exploring Cross-Lingual Transfer in Arabic Language Models
by: Khalak, Abdulmuizz, et al.
Published: (2026)
by: Khalak, Abdulmuizz, et al.
Published: (2026)
Language Models as Artificial Learners: Investigating Crosslinguistic Influence
by: Issam, Abderrahmane, et al.
Published: (2026)
by: Issam, Abderrahmane, et al.
Published: (2026)
A Dutch Financial Large Language Model
by: Noels, Sander, et al.
Published: (2024)
by: Noels, Sander, et al.
Published: (2024)
A Representation Level Analysis of NMT Model Robustness to Grammatical Errors
by: Issam, Abderrahmane, et al.
Published: (2025)
by: Issam, Abderrahmane, et al.
Published: (2025)
Cross-Modal Robustness Transfer (CMRT): Training Robust Speech Translation Models Using Adversarial Text
by: Issam, Abderrahmane, et al.
Published: (2026)
by: Issam, Abderrahmane, et al.
Published: (2026)
Know When to Fuse: Investigating Non-English Hybrid Retrieval in the Legal Domain
by: Louis, Antoine, et al.
Published: (2024)
by: Louis, Antoine, et al.
Published: (2024)
Computational Studies in Influencer Marketing: A Systematic Literature Review
by: Gui, Haoyang, et al.
Published: (2025)
by: Gui, Haoyang, et al.
Published: (2025)
Traceable by Design: An LLM Pipeline and Dashboard for EU Regulatory Consultation Analysis
by: Bertaglia, Thales, et al.
Published: (2026)
by: Bertaglia, Thales, et al.
Published: (2026)
ColBERT-XM: A Modular Multi-Vector Representation Model for Zero-Shot Multilingual Information Retrieval
by: Louis, Antoine, et al.
Published: (2024)
by: Louis, Antoine, et al.
Published: (2024)
Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing
by: Gui, Haoyang, et al.
Published: (2025)
by: Gui, Haoyang, et al.
Published: (2025)
DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment
by: Issam, Abderrahmane, et al.
Published: (2025)
by: Issam, Abderrahmane, et al.
Published: (2025)
Fixed and Adaptive Simultaneous Machine Translation Strategies Using Adapters
by: Issam, Abderrahmane, et al.
Published: (2024)
by: Issam, Abderrahmane, et al.
Published: (2024)
Navigating WebAI: Training Agents to Complete Web Tasks with Large Language Models and Reinforcement Learning
by: Thil, Lucas-Andreï, et al.
Published: (2024)
by: Thil, Lucas-Andreï, et al.
Published: (2024)
Analyzing the Attention Heads for Pronoun Disambiguation in Context-aware Machine Translation Models
by: Mąka, Paweł, et al.
Published: (2024)
by: Mąka, Paweł, et al.
Published: (2024)
Automated HIV Screening on Dutch Electronic Health Records with Large Language Models
by: Zhou, Lang, et al.
Published: (2025)
by: Zhou, Lang, et al.
Published: (2025)
You Are What You Train: Effects of Data Composition on Training Context-aware Machine Translation Models
by: Mąka, Paweł, et al.
Published: (2025)
by: Mąka, Paweł, et al.
Published: (2025)
Fietje: An open, efficient LLM for Dutch
by: Vanroy, Bram
Published: (2024)
by: Vanroy, Bram
Published: (2024)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data
by: Saxena, Vageesh, et al.
Published: (2024)
by: Saxena, Vageesh, et al.
Published: (2024)
Language corpora for the Dutch medical domain
by: van Es, B.
Published: (2026)
by: van Es, B.
Published: (2026)
BEIR-NL: Zero-shot Information Retrieval Benchmark for the Dutch Language
by: Banar, Nikolay, et al.
Published: (2024)
by: Banar, Nikolay, et al.
Published: (2024)
Transforming Dutch: Debiasing Dutch Coreference Resolution Systems for Non-binary Pronouns
by: van Boven, Goya, et al.
Published: (2024)
by: van Boven, Goya, et al.
Published: (2024)
GEITje 7B Ultra: A Conversational Model for Dutch
by: Vanroy, Bram
Published: (2024)
by: Vanroy, Bram
Published: (2024)
Triple-Encoders: Representations That Fire Together, Wire Together
by: Erker, Justus-Jonas, et al.
Published: (2024)
by: Erker, Justus-Jonas, et al.
Published: (2024)
Sequence Shortening for Context-Aware Machine Translation
by: Mąka, Paweł, et al.
Published: (2024)
by: Mąka, Paweł, et al.
Published: (2024)
MTEB-NL and E5-NL: Embedding Benchmark and Models for Dutch
by: Banar, Nikolay, et al.
Published: (2025)
by: Banar, Nikolay, et al.
Published: (2025)
GPT-NL Public Corpus: A Permissively Licensed, Dutch-First Dataset for LLM Pre-training
by: van Oort, Jesse, et al.
Published: (2026)
by: van Oort, Jesse, et al.
Published: (2026)
Measuring Social Biases in Masked Language Models by Proxy of Prediction Quality
by: Zalkikar, Rahul, et al.
Published: (2024)
by: Zalkikar, Rahul, et al.
Published: (2024)
Robust Evaluation Measures for Evaluating Social Biases in Masked Language Models
by: Liu, Yang
Published: (2024)
by: Liu, Yang
Published: (2024)
ChocoLlama: Lessons Learned From Teaching Llamas Dutch
by: Meeus, Matthieu, et al.
Published: (2024)
by: Meeus, Matthieu, et al.
Published: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
by: Bauer, Julie, et al.
Published: (2025)
by: Bauer, Julie, et al.
Published: (2025)
Bilingual BSARD: Extending Statutory Article Retrieval to Dutch
by: Lotfi, Ehsan, et al.
Published: (2024)
by: Lotfi, Ehsan, et al.
Published: (2024)
FairPair: A Robust Evaluation of Biases in Language Models through Paired Perturbations
by: Dwivedi-Yu, Jane, et al.
Published: (2024)
by: Dwivedi-Yu, Jane, et al.
Published: (2024)
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models
by: Khandelwal, Khyati, et al.
Published: (2023)
by: Khandelwal, Khyati, et al.
Published: (2023)
Translating the Grievance Dictionary: a psychometric evaluation of Dutch, German, and Italian versions
by: van der Vegt, Isabelle, et al.
Published: (2025)
by: van der Vegt, Isabelle, et al.
Published: (2025)
BanStereoSet: A Dataset to Measure Stereotypical Social Biases in LLMs for Bangla
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
by: Huang, Jen-tse, et al.
Published: (2025)
by: Huang, Jen-tse, et al.
Published: (2025)
Generating High Quality Synthetic Data for Dutch Medical Conversations
by: Kuan, Cecilia, et al.
Published: (2026)
by: Kuan, Cecilia, et al.
Published: (2026)
Measuring Stereotype and Deviation Biases in Large Language Models
by: Wang, Daniel, et al.
Published: (2025)
by: Wang, Daniel, et al.
Published: (2025)
Similar Items
-
Across Platforms and Languages: Dutch Influencers and Legal Disclosures on Instagram, YouTube and TikTok
by: Gui, Haoyang, et al.
Published: (2024) -
From FusHa to Folk: Exploring Cross-Lingual Transfer in Arabic Language Models
by: Khalak, Abdulmuizz, et al.
Published: (2026) -
Language Models as Artificial Learners: Investigating Crosslinguistic Influence
by: Issam, Abderrahmane, et al.
Published: (2026) -
A Dutch Financial Large Language Model
by: Noels, Sander, et al.
Published: (2024) -
A Representation Level Analysis of NMT Model Robustness to Grammatical Errors
by: Issam, Abderrahmane, et al.
Published: (2025)