MiTTenS: A Dataset for Evaluating Gender Mistranslation
Fuente:
arXiv
Saved in:
| Main Authors: | Robinson, Kevin, Kudugunta, Sneha, Stella, Romina, Dev, Sunipa, Bastings, Jasmijn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
by: Bhutani, Mukul, et al.
Published: (2024)
by: Bhutani, Mukul, et al.
Published: (2024)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
by: Davani, Aida, et al.
Published: (2025)
by: Davani, Aida, et al.
Published: (2025)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
by: Jha, Akshita, et al.
Published: (2024)
by: Jha, Akshita, et al.
Published: (2024)
Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens
by: Nigatu, Hellina Hailu, et al.
Published: (2025)
by: Nigatu, Hellina Hailu, et al.
Published: (2025)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
MisgenderMender: A Community-Informed Approach to Interventions for Misgendering
by: Hossain, Tamanna, et al.
Published: (2024)
by: Hossain, Tamanna, et al.
Published: (2024)
A Unified Framework to Quantify Cultural Intelligence of AI
by: Dev, Sunipa, et al.
Published: (2026)
by: Dev, Sunipa, et al.
Published: (2026)
AI-generated podcasts: Synthetic Intimacy and Cultural Mistranslation in NotebookLM's Audio Overviews
by: Rettberg, Jill Walker
Published: (2025)
by: Rettberg, Jill Walker
Published: (2025)
GeniL: A Multilingual Dataset on Generalizing Language
by: Davani, Aida Mostafazadeh, et al.
Published: (2024)
by: Davani, Aida Mostafazadeh, et al.
Published: (2024)
(Mis)Fitting: A Survey of Scaling Laws
by: Li, Margaret, et al.
Published: (2025)
by: Li, Margaret, et al.
Published: (2025)
Low-Rank Adaptation for Multilingual Summarization: An Empirical Study
by: Whitehouse, Chenxi, et al.
Published: (2023)
by: Whitehouse, Chenxi, et al.
Published: (2023)
WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics
by: Maurya, Sneha, et al.
Published: (2026)
by: Maurya, Sneha, et al.
Published: (2026)
Slicing and Dicing: Configuring Optimal Mixtures of Experts
by: Li, Margaret, et al.
Published: (2026)
by: Li, Margaret, et al.
Published: (2026)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
SAFARI: A Community-Engaged Approach and Dataset of Stereotype Resources in the Sub-Saharan African Context
by: Verma, Aishwarya, et al.
Published: (2026)
by: Verma, Aishwarya, et al.
Published: (2026)
Leveraging Large Language Models to Measure Gender Representation Bias in Gendered Language Corpora
by: Derner, Erik, et al.
Published: (2024)
by: Derner, Erik, et al.
Published: (2024)
Local Contrastive Editing of Gender Stereotypes
by: Lutz, Marlene, et al.
Published: (2024)
by: Lutz, Marlene, et al.
Published: (2024)
For the Misgendered Chinese in Gender Bias Research: Multi-Task Learning with Knowledge Distillation for Pinyin Name-Gender Prediction
by: Du, Xiaocong, et al.
Published: (2024)
by: Du, Xiaocong, et al.
Published: (2024)
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
by: Ivetta, Guido, et al.
Published: (2025)
by: Ivetta, Guido, et al.
Published: (2025)
Gender Bias Detection in Court Decisions: A Brazilian Case Study
by: Benatti, Raysa, et al.
Published: (2024)
by: Benatti, Raysa, et al.
Published: (2024)
Quantifying Gender Biases Towards Politicians on Reddit
by: Marjanovic, Sara, et al.
Published: (2021)
by: Marjanovic, Sara, et al.
Published: (2021)
GECOBench: A Gender-Controlled Text Dataset and Benchmark for Quantifying Biases in Explanations
by: Wilming, Rick, et al.
Published: (2024)
by: Wilming, Rick, et al.
Published: (2024)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
by: Nejadgholi, Isar, et al.
Published: (2024)
by: Nejadgholi, Isar, et al.
Published: (2024)
Gendered Divides in Online Discussions about Reproductive Rights
by: Rao, Ashwin, et al.
Published: (2025)
by: Rao, Ashwin, et al.
Published: (2025)
Gender Bias in Emotion Recognition by Large Language Models
by: Herbert, Maureen, et al.
Published: (2025)
by: Herbert, Maureen, et al.
Published: (2025)
How Gender Interacts with Political Values: A Case Study on Czech BERT Models
by: Ali, Adnan Al, et al.
Published: (2024)
by: Ali, Adnan Al, et al.
Published: (2024)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
by: Chen, Liuliu, et al.
Published: (2026)
by: Chen, Liuliu, et al.
Published: (2026)
The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring
by: Armstrong, Lena, et al.
Published: (2024)
by: Armstrong, Lena, et al.
Published: (2024)
Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)
by: Girrbach, Leander, et al.
Published: (2024)
by: Girrbach, Leander, et al.
Published: (2024)
Gender Representation and Bias in Indian Civil Service Mock Interviews
by: Banerjee, Somonnoy, et al.
Published: (2024)
by: Banerjee, Somonnoy, et al.
Published: (2024)
ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review
by: Khowaja, Sunder Ali, et al.
Published: (2023)
by: Khowaja, Sunder Ali, et al.
Published: (2023)
Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images
by: Fraser, Kathleen C., et al.
Published: (2024)
by: Fraser, Kathleen C., et al.
Published: (2024)
A Multilingual Similarity Dataset for News Article Frame
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
by: Dorn, Rebecca, et al.
Published: (2024)
by: Dorn, Rebecca, et al.
Published: (2024)
Online Anti-sexist Speech: Identifying Resistance to Gender Bias in Political Discourse
by: Dutta, Aditi, et al.
Published: (2025)
by: Dutta, Aditi, et al.
Published: (2025)
Fair Play in the Newsroom: Actor-Based Filtering Gender Discrimination in Text Corpora
by: Urchs, Stefanie, et al.
Published: (2025)
by: Urchs, Stefanie, et al.
Published: (2025)
Gender and Positional Biases in LLM-Based Hiring Decisions: Evidence from Comparative CV/Résumé Evaluations
by: Rozado, David
Published: (2025)
by: Rozado, David
Published: (2025)
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation
by: Neumann, Terrence, et al.
Published: (2024)
by: Neumann, Terrence, et al.
Published: (2024)
Heard or Halted? Gender, Interruptions, and Emotional Tone in U.S. Supreme Court Oral Arguments
by: Tong, Yifei
Published: (2025)
by: Tong, Yifei
Published: (2025)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
Similar Items
-
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
by: Bhutani, Mukul, et al.
Published: (2024) -
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
by: Davani, Aida, et al.
Published: (2025) -
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
by: Jha, Akshita, et al.
Published: (2024) -
Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens
by: Nigatu, Hellina Hailu, et al.
Published: (2025) -
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
by: Cheng, Myra, et al.
Published: (2026)