Breaking Bias, Building Bridges: Evaluation and Mitigation of Social Biases in LLMs via Contact Hypothesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Raj, Chahat, Mukherjee, Anjishnu, Caliskan, Aylin, Anastasopoulos, Antonios, Zhu, Ziwei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BiasDora: Exploring Hidden Biased Associations in Vision-Language Models
por: Raj, Chahat, et al.
Publicado: (2024)
por: Raj, Chahat, et al.
Publicado: (2024)
VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
por: Raj, Chahat, et al.
Publicado: (2025)
por: Raj, Chahat, et al.
Publicado: (2025)
Talent or Luck? Evaluating Attribution Bias in Large Language Models
por: Raj, Chahat, et al.
Publicado: (2025)
por: Raj, Chahat, et al.
Publicado: (2025)
Purdah and Patriarchy: Evaluating and Mitigating South Asian Biases in Open-Ended Multilingual LLM Generations
por: Rinki, Mamnuya, et al.
Publicado: (2025)
por: Rinki, Mamnuya, et al.
Publicado: (2025)
Crossroads of Continents: Automated Artifact Extraction for Cultural Adaptation with Large Multimodal Models
por: Mukherjee, Anjishnu, et al.
Publicado: (2024)
por: Mukherjee, Anjishnu, et al.
Publicado: (2024)
Metadata Conditioned Large Language Models for Localization
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
What's Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
por: Pan, Jinhao, et al.
Publicado: (2025)
por: Pan, Jinhao, et al.
Publicado: (2025)
KnowBias: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement
por: Pan, Jinhao, et al.
Publicado: (2026)
por: Pan, Jinhao, et al.
Publicado: (2026)
Bias Association Discovery Framework for Open-Ended LLM Generations
por: Pan, Jinhao, et al.
Publicado: (2025)
por: Pan, Jinhao, et al.
Publicado: (2025)
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
por: Gueorguieva, Anna-Maria, et al.
Publicado: (2025)
por: Gueorguieva, Anna-Maria, et al.
Publicado: (2025)
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
por: Wilson, Kyra, et al.
Publicado: (2024)
por: Wilson, Kyra, et al.
Publicado: (2024)
Data-Augmentation-Based Dialectal Adaptation for LLMs
por: Faisal, Fahim, et al.
Publicado: (2024)
por: Faisal, Fahim, et al.
Publicado: (2024)
ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages
por: Ghosh, Sourojit, et al.
Publicado: (2023)
por: Ghosh, Sourojit, et al.
Publicado: (2023)
Urban Mobility Assessment Using LLMs
por: Bhandari, Prabin, et al.
Publicado: (2024)
por: Bhandari, Prabin, et al.
Publicado: (2024)
TigerCoder: A Novel Suite of LLMs for Code Generation in Bangla
por: Raihan, Nishat, et al.
Publicado: (2025)
por: Raihan, Nishat, et al.
Publicado: (2025)
Back to School: Translation Using Grammar Books
por: Hus, Jonathan, et al.
Publicado: (2024)
por: Hus, Jonathan, et al.
Publicado: (2024)
An Efficient Approach for Studying Cross-Lingual Transfer in Multilingual Language Models
por: Faisal, Fahim, et al.
Publicado: (2024)
por: Faisal, Fahim, et al.
Publicado: (2024)
GMU Systems for the IWSLT 2025 Low-Resource Speech Translation Shared Task
por: Meng, Chutong, et al.
Publicado: (2025)
por: Meng, Chutong, et al.
Publicado: (2025)
Biases Propagate in Encoder-based Vision-Language Models: A Systematic Analysis From Intrinsic Measures to Zero-shot Retrieval Outcomes
por: Ghate, Kshitish, et al.
Publicado: (2025)
por: Ghate, Kshitish, et al.
Publicado: (2025)
mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code Generation
por: Raihan, Nishat, et al.
Publicado: (2024)
por: Raihan, Nishat, et al.
Publicado: (2024)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
por: Long, Do Xuan, et al.
Publicado: (2024)
por: Long, Do Xuan, et al.
Publicado: (2024)
Speaking of Language: Reflections on Metalanguage Research in NLP
por: Schneider, Nathan, et al.
Publicado: (2026)
por: Schneider, Nathan, et al.
Publicado: (2026)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
por: Fayyazsanavi, Pooya, et al.
Publicado: (2024)
por: Fayyazsanavi, Pooya, et al.
Publicado: (2024)
Dialectal Toxicity Detection: Evaluating LLM-as-a-Judge Consistency Across Language Varieties
por: Faisal, Fahim, et al.
Publicado: (2024)
por: Faisal, Fahim, et al.
Publicado: (2024)
A Study on Scaling Up Multilingual News Framing Analysis
por: Akter, Syeda Sabrina, et al.
Publicado: (2024)
por: Akter, Syeda Sabrina, et al.
Publicado: (2024)
Dialect Normalization using Large Language Models and Morphological Rules
por: Dimakis, Antonios, et al.
Publicado: (2025)
por: Dimakis, Antonios, et al.
Publicado: (2025)
CODET: A Benchmark for Contrastive Dialectal Evaluation of Machine Translation
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2023)
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2023)
A Taxonomy of Stereotype Content in Large Language Models
por: Nicolas, Gandalf, et al.
Publicado: (2024)
por: Nicolas, Gandalf, et al.
Publicado: (2024)
Mitigating Biases in Language Models via Bias Unlearning
por: Liu, Dianqing, et al.
Publicado: (2025)
por: Liu, Dianqing, et al.
Publicado: (2025)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
por: Choi, Alexander S., et al.
Publicado: (2024)
por: Choi, Alexander S., et al.
Publicado: (2024)
A Case Study on Filtering for End-to-End Speech Translation
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2024)
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2024)
Script-Agnostic Language Identification
por: Agarwal, Milind, et al.
Publicado: (2024)
por: Agarwal, Milind, et al.
Publicado: (2024)
Developing a Mixed-Methods Pipeline for Community-Oriented Digitization of Kwak'wala Legacy Texts
por: Agarwal, Milind, et al.
Publicado: (2025)
por: Agarwal, Milind, et al.
Publicado: (2025)
Cross-Lingual Representation Alignment Through Contrastive Image-Caption Tuning
por: Krasner, Nathaniel, et al.
Publicado: (2025)
por: Krasner, Nathaniel, et al.
Publicado: (2025)
Automated Python Translation
por: Otten, Joshua, et al.
Publicado: (2025)
por: Otten, Joshua, et al.
Publicado: (2025)
No Thoughts Just AI: Biased LLM Hiring Recommendations Alter Human Decision Making and Limit Human Autonomy
por: Wilson, Kyra, et al.
Publicado: (2025)
por: Wilson, Kyra, et al.
Publicado: (2025)
LLMs are Biased Teachers: Evaluating LLM Bias in Personalized Education
por: Weissburg, Iain, et al.
Publicado: (2024)
por: Weissburg, Iain, et al.
Publicado: (2024)
Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers
por: Xie, Roy, et al.
Publicado: (2024)
por: Xie, Roy, et al.
Publicado: (2024)
A Morphologically-Aware Dictionary-based Data Augmentation Technique for Machine Translation of Under-Represented Languages
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2024)
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2024)
Ejemplares similares
-
BiasDora: Exploring Hidden Biased Associations in Vision-Language Models
por: Raj, Chahat, et al.
Publicado: (2024) -
VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
por: Raj, Chahat, et al.
Publicado: (2025) -
Talent or Luck? Evaluating Attribution Bias in Large Language Models
por: Raj, Chahat, et al.
Publicado: (2025) -
Purdah and Patriarchy: Evaluating and Mitigating South Asian Biases in Open-Ended Multilingual LLM Generations
por: Rinki, Mamnuya, et al.
Publicado: (2025) -
Crossroads of Continents: Automated Artifact Extraction for Cultural Adaptation with Large Multimodal Models
por: Mukherjee, Anjishnu, et al.
Publicado: (2024)