The Algorithmic Unconscious: Structural Mechanisms and Implicit Biases in Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Boisnard, Philippe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ethology of Latent Spaces
von: Boisnard, Philippe
Veröffentlicht: (2026)
von: Boisnard, Philippe
Veröffentlicht: (2026)
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
The Life Cycle of Large Language Models: A Review of Biases in Education
von: Lee, Jinsook, et al.
Veröffentlicht: (2024)
von: Lee, Jinsook, et al.
Veröffentlicht: (2024)
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
Understanding Intrinsic Socioeconomic Biases in Large Language Models
von: Arzaghi, Mina, et al.
Veröffentlicht: (2024)
von: Arzaghi, Mina, et al.
Veröffentlicht: (2024)
Large Language Models are Geographically Biased
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
PRISM: A Methodology for Auditing Biases in Large Language Models
von: Azzopardi, Leif, et al.
Veröffentlicht: (2024)
von: Azzopardi, Leif, et al.
Veröffentlicht: (2024)
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models
von: Khandelwal, Khyati, et al.
Veröffentlicht: (2023)
von: Khandelwal, Khyati, et al.
Veröffentlicht: (2023)
Generative Language Models Exhibit Social Identity Biases
von: Hu, Tiancheng, et al.
Veröffentlicht: (2023)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2023)
Inference-Time Reasoning Selectively Reduces Implicit Social Bias in Large Language Models
von: Apsel, Molly, et al.
Veröffentlicht: (2026)
von: Apsel, Molly, et al.
Veröffentlicht: (2026)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
von: Wu, Addison J., et al.
Veröffentlicht: (2025)
von: Wu, Addison J., et al.
Veröffentlicht: (2025)
A Toolbox for Surfacing Health Equity Harms and Biases in Large Language Models
von: Pfohl, Stephen R., et al.
Veröffentlicht: (2024)
von: Pfohl, Stephen R., et al.
Veröffentlicht: (2024)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
von: Jahara, Fatima, et al.
Veröffentlicht: (2025)
von: Jahara, Fatima, et al.
Veröffentlicht: (2025)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
von: Christian, Brian, et al.
Veröffentlicht: (2026)
von: Christian, Brian, et al.
Veröffentlicht: (2026)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)
von: Girrbach, Leander, et al.
Veröffentlicht: (2024)
von: Girrbach, Leander, et al.
Veröffentlicht: (2024)
Semantic and Structural Analysis of Implicit Biases in Large Language Models: An Interpretable Approach
von: Zhang, Renhan, et al.
Veröffentlicht: (2025)
von: Zhang, Renhan, et al.
Veröffentlicht: (2025)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
Covert Bias: The Severity of Social Views' Unalignment in Language Models Towards Implicit and Explicit Opinion
von: Aldayel, Abeer, et al.
Veröffentlicht: (2024)
von: Aldayel, Abeer, et al.
Veröffentlicht: (2024)
Prompt Perturbations Reveal Human-Like Biases in Large Language Model Survey Responses
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
Assessing Historical Structural Oppression Worldwide via Rule-Guided Prompting of Large Language Models
von: Chatterjee, Sreejato, et al.
Veröffentlicht: (2025)
von: Chatterjee, Sreejato, et al.
Veröffentlicht: (2025)
Empowering Many, Biasing a Few: Generalist Credit Scoring through Large Language Models
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
Exploring Bengali Religious Dialect Biases in Large Language Models with Evaluation Perspectives
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Motivation in Large Language Models
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
FairPair: A Robust Evaluation of Biases in Language Models through Paired Perturbations
von: Dwivedi-Yu, Jane, et al.
Veröffentlicht: (2024)
von: Dwivedi-Yu, Jane, et al.
Veröffentlicht: (2024)
A Comprehensive Study of Implicit and Explicit Biases in Large Language Models
von: Kazi, Fatima, et al.
Veröffentlicht: (2025)
von: Kazi, Fatima, et al.
Veröffentlicht: (2025)
Long-Tail Knowledge in Large Language Models: Taxonomy, Mechanisms, Interventions and Implications
von: Badhe, Sanket, et al.
Veröffentlicht: (2026)
von: Badhe, Sanket, et al.
Veröffentlicht: (2026)
The Biased Samaritan: LLM biases in Perceived Kindness
von: Fagan, Jack H, et al.
Veröffentlicht: (2025)
von: Fagan, Jack H, et al.
Veröffentlicht: (2025)
Quantifying Gender Biases Towards Politicians on Reddit
von: Marjanovic, Sara, et al.
Veröffentlicht: (2021)
von: Marjanovic, Sara, et al.
Veröffentlicht: (2021)
Language of Thought Shapes Output Diversity in Large Language Models
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
von: Xu, Shaoyang, et al.
Veröffentlicht: (2026)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
Large Language Models in the Abuse Detection Pipeline
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
Will Large Language Models Transform Clinical Prediction?
von: Yildiz, Yusuf, et al.
Veröffentlicht: (2025)
von: Yildiz, Yusuf, et al.
Veröffentlicht: (2025)
Climate Change from Large Language Models
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
Investigating Cultural Alignment of Large Language Models
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
von: AlKhamissi, Badr, et al.
Veröffentlicht: (2024)
Urban Computing in the Era of Large Language Models
von: Li, Zhonghang, et al.
Veröffentlicht: (2025)
von: Li, Zhonghang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Ethology of Latent Spaces
von: Boisnard, Philippe
Veröffentlicht: (2026) -
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024) -
The Life Cycle of Large Language Models: A Review of Biases in Education
von: Lee, Jinsook, et al.
Veröffentlicht: (2024) -
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024) -
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)