Man Made Language Models? Evaluating LLMs' Perpetuation of Masculine Generics Bias
Fuente:
arXiv
Guardado en:
| Autores principales: | Doyen, Enzo, Todirascu, Amalia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GeNRe: A French Gender-Neutral Rewriting System Using Collective Nouns
por: Doyen, Enzo, et al.
Publicado: (2025)
por: Doyen, Enzo, et al.
Publicado: (2025)
Chain-of-MetaWriting: Linguistic and Textual Analysis of How Small Language Models Write Young Students Texts
por: Buhnila, Ioana, et al.
Publicado: (2024)
por: Buhnila, Ioana, et al.
Publicado: (2024)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
por: Buscemi, Alessio, et al.
Publicado: (2025)
por: Buscemi, Alessio, et al.
Publicado: (2025)
ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages
por: Ghosh, Sourojit, et al.
Publicado: (2023)
por: Ghosh, Sourojit, et al.
Publicado: (2023)
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
por: Fayyaz, Hamed, et al.
Publicado: (2024)
por: Fayyaz, Hamed, et al.
Publicado: (2024)
Unpacking Human Preference for LLMs: Demographically Aware Evaluation with the HUMAINE Framework
por: Petrova, Nora, et al.
Publicado: (2026)
por: Petrova, Nora, et al.
Publicado: (2026)
No LLM is Free From Bias: A Comprehensive Study of Bias Evaluation in Large Language Models
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024)
por: Oi, Masanari, et al.
Publicado: (2024)
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models
por: Pombal, José, et al.
Publicado: (2026)
por: Pombal, José, et al.
Publicado: (2026)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
por: Nadeem, Afrozah, et al.
Publicado: (2025)
por: Nadeem, Afrozah, et al.
Publicado: (2025)
DataMan: Data Manager for Pre-training Large Language Models
por: Peng, Ru, et al.
Publicado: (2025)
por: Peng, Ru, et al.
Publicado: (2025)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
por: Wu, Yihao, et al.
Publicado: (2025)
por: Wu, Yihao, et al.
Publicado: (2025)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
por: Xu, Wenda, et al.
Publicado: (2025)
por: Xu, Wenda, et al.
Publicado: (2025)
Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs
por: Nadeem, Afrozah, et al.
Publicado: (2026)
por: Nadeem, Afrozah, et al.
Publicado: (2026)
Entropy-Based Block Pruning for Efficient Large Language Models
por: Yang, Liangwei, et al.
Publicado: (2025)
por: Yang, Liangwei, et al.
Publicado: (2025)
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models
por: Satriani, Dario, et al.
Publicado: (2025)
por: Satriani, Dario, et al.
Publicado: (2025)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
por: Saraf, Muskan, et al.
Publicado: (2025)
por: Saraf, Muskan, et al.
Publicado: (2025)
CEA-LIST at CheckThat! 2025: Evaluating LLMs as Detectors of Bias and Opinion in Text
por: Elbouanani, Akram, et al.
Publicado: (2025)
por: Elbouanani, Akram, et al.
Publicado: (2025)
Capturing Bias Diversity in LLMs
por: Gosavi, Purva Prasad, et al.
Publicado: (2024)
por: Gosavi, Purva Prasad, et al.
Publicado: (2024)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
por: Kumar, Divyanshu, et al.
Publicado: (2024)
por: Kumar, Divyanshu, et al.
Publicado: (2024)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
por: Neumann, Anna, et al.
Publicado: (2025)
por: Neumann, Anna, et al.
Publicado: (2025)
Where Fake Citations Are Made: Tracing Field-Level Hallucination to Specific Neurons in LLMs
por: Chen, Yuefei, et al.
Publicado: (2026)
por: Chen, Yuefei, et al.
Publicado: (2026)
Cultural Bias in Large Language Models: Evaluating AI Agents through Moral Questionnaires
por: Münker, Simon
Publicado: (2025)
por: Münker, Simon
Publicado: (2025)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
por: Wen, Yuchen, et al.
Publicado: (2024)
por: Wen, Yuchen, et al.
Publicado: (2024)
Investigating Bias: A Multilingual Pipeline for Generating, Solving, and Evaluating Math Problems with LLMs
por: Mahran, Mariam, et al.
Publicado: (2025)
por: Mahran, Mariam, et al.
Publicado: (2025)
Evaluate Bias without Manual Test Sets: A Concept Representation Perspective for LLMs
por: Gao, Lang, et al.
Publicado: (2025)
por: Gao, Lang, et al.
Publicado: (2025)
Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs
por: Bouchard, Dylan
Publicado: (2024)
por: Bouchard, Dylan
Publicado: (2024)
Regional Bias in Large Language Models
por: Gopinadh, M P V S, et al.
Publicado: (2026)
por: Gopinadh, M P V S, et al.
Publicado: (2026)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
por: Joshi, Ishika, et al.
Publicado: (2024)
por: Joshi, Ishika, et al.
Publicado: (2024)
CodeChain: Towards Modular Code Generation Through Chain of Self-revisions with Representative Sub-modules
por: Le, Hung, et al.
Publicado: (2023)
por: Le, Hung, et al.
Publicado: (2023)
Implicit Bias in LLMs: A Survey
por: Lin, Xinru, et al.
Publicado: (2025)
por: Lin, Xinru, et al.
Publicado: (2025)
Cognitive Bias in Decision-Making with LLMs
por: Echterhoff, Jessica, et al.
Publicado: (2024)
por: Echterhoff, Jessica, et al.
Publicado: (2024)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
por: Rao, Pooja S. B., et al.
Publicado: (2025)
por: Rao, Pooja S. B., et al.
Publicado: (2025)
MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
por: Luo, Ziyang, et al.
Publicado: (2025)
por: Luo, Ziyang, et al.
Publicado: (2025)
From Word to World: Evaluate and Mitigate Culture Bias in LLMs via Word Association Test
por: Dai, Xunlian, et al.
Publicado: (2025)
por: Dai, Xunlian, et al.
Publicado: (2025)
Social Bias Benchmark for Generation: A Comparison of Generation and QA-Based Evaluations
por: Jin, Jiho, et al.
Publicado: (2025)
por: Jin, Jiho, et al.
Publicado: (2025)
Assessing Political Bias in Large Language Models
por: Rettenberger, Luca, et al.
Publicado: (2024)
por: Rettenberger, Luca, et al.
Publicado: (2024)
Soft-prompt Tuning for Large Language Models to Evaluate Bias
por: Tian, Jacob-Junqi, et al.
Publicado: (2023)
por: Tian, Jacob-Junqi, et al.
Publicado: (2023)
Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
por: Shin, Jisu, et al.
Publicado: (2024)
por: Shin, Jisu, et al.
Publicado: (2024)
Ejemplares similares
-
GeNRe: A French Gender-Neutral Rewriting System Using Collective Nouns
por: Doyen, Enzo, et al.
Publicado: (2025) -
Chain-of-MetaWriting: Linguistic and Textual Analysis of How Small Language Models Write Young Students Texts
por: Buhnila, Ioana, et al.
Publicado: (2024) -
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026) -
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
por: Buscemi, Alessio, et al.
Publicado: (2025) -
ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages
por: Ghosh, Sourojit, et al.
Publicado: (2023)