Implicit Bias-Like Patterns in Reasoning Models
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Messi H. J., Lai, Calvin K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
More Distinctively Black and Feminine Faces Lead to Increased Stereotyping in Vision-Language Models
by: Lee, Messi H. J., et al.
Published: (2024)
by: Lee, Messi H. J., et al.
Published: (2024)
Implicit Bias in LLMs for Transgender Populations
by: Hirsch, Micaela, et al.
Published: (2026)
by: Hirsch, Micaela, et al.
Published: (2026)
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
by: Lee, Messi H. J., et al.
Published: (2024)
by: Lee, Messi H. J., et al.
Published: (2024)
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
by: Sun, Lihao, et al.
Published: (2025)
by: Sun, Lihao, et al.
Published: (2025)
From Bias Mitigation to Bias Negotiation: Governing Identity and Sociocultural Reasoning in Generative AI
by: Dunivin, Zackary Okun, et al.
Published: (2026)
by: Dunivin, Zackary Okun, et al.
Published: (2026)
Can Large Language Models Capture Public Opinion about Global Warming? An Empirical Assessment of Algorithmic Fidelity and Bias
by: Lee, S., et al.
Published: (2023)
by: Lee, S., et al.
Published: (2023)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
by: Jahara, Fatima, et al.
Published: (2025)
by: Jahara, Fatima, et al.
Published: (2025)
Giving AI Personalities Leads to More Human-Like Reasoning
by: Nighojkar, Animesh, et al.
Published: (2025)
by: Nighojkar, Animesh, et al.
Published: (2025)
Beyond Partisan Leaning: A Comparative Analysis of Political Bias in Large Language Models
by: Peng, Tai-Quan, et al.
Published: (2024)
by: Peng, Tai-Quan, et al.
Published: (2024)
Exploring Social Desirability Response Bias in Large Language Models: Evidence from GPT-4 Simulations
by: Lee, Sanguk, et al.
Published: (2024)
by: Lee, Sanguk, et al.
Published: (2024)
Human-like in-group bias in instruction-tuned language model agents
by: Lee, Messi H. J.
Published: (2026)
by: Lee, Messi H. J.
Published: (2026)
Measuring and Mitigating Bias in Code Generated by Large Language Models
by: Chen, Yuxi, et al.
Published: (2026)
by: Chen, Yuxi, et al.
Published: (2026)
Evaluation of Bias Towards Medical Professionals in Large Language Models
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
by: Shin, Jisu, et al.
Published: (2024)
by: Shin, Jisu, et al.
Published: (2024)
Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans
by: Lee, Messi H. J., et al.
Published: (2024)
by: Lee, Messi H. J., et al.
Published: (2024)
Fairness and Bias in Algorithmic Hiring: a Multidisciplinary Survey
by: Fabris, Alessandro, et al.
Published: (2023)
by: Fabris, Alessandro, et al.
Published: (2023)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
by: Kim, Kyuhee, et al.
Published: (2025)
by: Kim, Kyuhee, et al.
Published: (2025)
No Free Lunch in Language Model Bias Mitigation? Targeted Bias Reduction Can Exacerbate Unmitigated LLM Biases
by: Chand, Shireen, et al.
Published: (2025)
by: Chand, Shireen, et al.
Published: (2025)
Cross-Language Bias Examination in Large Language Models
by: Liang, Yuxuan, et al.
Published: (2025)
by: Liang, Yuxuan, et al.
Published: (2025)
The Technology of Outrage: Bias in Artificial Intelligence
by: Bridewell, Will, et al.
Published: (2024)
by: Bridewell, Will, et al.
Published: (2024)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
by: Itzhak, Itay, et al.
Published: (2023)
by: Itzhak, Itay, et al.
Published: (2023)
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
by: Wu, Shangxi, et al.
Published: (2023)
by: Wu, Shangxi, et al.
Published: (2023)
AccessEval: Benchmarking Disability Bias in Large Language Models
by: Panda, Srikant, et al.
Published: (2025)
by: Panda, Srikant, et al.
Published: (2025)
Gender Bias in Machine Translation and The Era of Large Language Models
by: Vanmassenhove, Eva
Published: (2024)
by: Vanmassenhove, Eva
Published: (2024)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
by: Ma, Mingyu Derek, et al.
Published: (2023)
by: Ma, Mingyu Derek, et al.
Published: (2023)
Are LLMs More Skeptical of Entertainment News?
by: Lai, Huiqian
Published: (2026)
by: Lai, Huiqian
Published: (2026)
Adaptive Generation of Bias-Eliciting Questions for LLMs
by: Staab, Robin, et al.
Published: (2025)
by: Staab, Robin, et al.
Published: (2025)
Global AI Bias Audit for Technical Governance
by: Hung, Jason
Published: (2026)
by: Hung, Jason
Published: (2026)
Equity Bias: An Ethical Framework for AI Design
by: Lockwood, Mary
Published: (2026)
by: Lockwood, Mary
Published: (2026)
Gender Bias of LLM in Economics: An Existentialism Perspective
by: Zhong, Hui, et al.
Published: (2024)
by: Zhong, Hui, et al.
Published: (2024)
Quantifying Gender Bias in Large Language Models: When ChatGPT Becomes a Hiring Manager
by: Gerszberg, Nina, et al.
Published: (2026)
by: Gerszberg, Nina, et al.
Published: (2026)
Bias-Aware AI Chatbot for Engineering Advising at the University of Maryland A. James Clark School of Engineering
by: Kartholy, Prarthana P., et al.
Published: (2025)
by: Kartholy, Prarthana P., et al.
Published: (2025)
Dissecting Bias of ChatGPT in College Major Recommendations
by: Zheng, Alex
Published: (2023)
by: Zheng, Alex
Published: (2023)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
by: Rao, Pooja S. B., et al.
Published: (2025)
by: Rao, Pooja S. B., et al.
Published: (2025)
Bias and Fairness in Large Language Models: A Survey
by: Gallegos, Isabel O., et al.
Published: (2023)
by: Gallegos, Isabel O., et al.
Published: (2023)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
by: Singh, Smriti, et al.
Published: (2024)
by: Singh, Smriti, et al.
Published: (2024)
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
by: Ko, Changgeon, et al.
Published: (2024)
by: Ko, Changgeon, et al.
Published: (2024)
Dr. Bias: Social Disparities in AI-Powered Medical Guidance
by: Kondrup, Emma, et al.
Published: (2025)
by: Kondrup, Emma, et al.
Published: (2025)
The Refutability Gap: Challenges in Validating Reasoning by Large Language Models
by: Mossel, Elchanan
Published: (2025)
by: Mossel, Elchanan
Published: (2025)
Identifying and Improving Disability Bias in GPT-Based Resume Screening
by: Glazko, Kate, et al.
Published: (2024)
by: Glazko, Kate, et al.
Published: (2024)
Similar Items
-
More Distinctively Black and Feminine Faces Lead to Increased Stereotyping in Vision-Language Models
by: Lee, Messi H. J., et al.
Published: (2024) -
Implicit Bias in LLMs for Transgender Populations
by: Hirsch, Micaela, et al.
Published: (2026) -
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
by: Lee, Messi H. J., et al.
Published: (2024) -
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
by: Sun, Lihao, et al.
Published: (2025) -
From Bias Mitigation to Bias Negotiation: Governing Identity and Sociocultural Reasoning in Generative AI
by: Dunivin, Zackary Okun, et al.
Published: (2026)