Large Language Models Develop Novel Social Biases Through Adaptive Exploration
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Addison J., Liu, Ryan, Bai, Xuechunzi, Griffiths, Thomas L. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
di: Wu, Addison J., et al.
Pubblicazione: (2026)
di: Wu, Addison J., et al.
Pubblicazione: (2026)
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
di: Bai, Xuechunzi, et al.
Pubblicazione: (2024)
di: Bai, Xuechunzi, et al.
Pubblicazione: (2024)
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
di: Sun, Lihao, et al.
Pubblicazione: (2025)
di: Sun, Lihao, et al.
Pubblicazione: (2025)
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
di: Liu, Ryan, et al.
Pubblicazione: (2024)
di: Liu, Ryan, et al.
Pubblicazione: (2024)
Large Language Models Assume People are More Rational than We Really are
di: Liu, Ryan, et al.
Pubblicazione: (2024)
di: Liu, Ryan, et al.
Pubblicazione: (2024)
Are Large Language Models Sensitive to the Motives Behind Communication?
di: Wu, Addison J., et al.
Pubblicazione: (2025)
di: Wu, Addison J., et al.
Pubblicazione: (2025)
A Systematic Analysis of Biases in Large Language Models
di: Zhang, Xulang, et al.
Pubblicazione: (2025)
di: Zhang, Xulang, et al.
Pubblicazione: (2025)
PRISM: A Methodology for Auditing Biases in Large Language Models
di: Azzopardi, Leif, et al.
Pubblicazione: (2024)
di: Azzopardi, Leif, et al.
Pubblicazione: (2024)
Large Language Models are Geographically Biased
di: Manvi, Rohin, et al.
Pubblicazione: (2024)
di: Manvi, Rohin, et al.
Pubblicazione: (2024)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
Levels of Analysis for Large Language Models
di: Ku, Alexander Y., et al.
Pubblicazione: (2025)
di: Ku, Alexander Y., et al.
Pubblicazione: (2025)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
di: Christian, Brian, et al.
Pubblicazione: (2026)
di: Christian, Brian, et al.
Pubblicazione: (2026)
An Exploration of Higher Education Course Evaluation by Large Language Models
di: Yuan, Bo, et al.
Pubblicazione: (2024)
di: Yuan, Bo, et al.
Pubblicazione: (2024)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
di: Ding, Junchen, et al.
Pubblicazione: (2025)
di: Ding, Junchen, et al.
Pubblicazione: (2025)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
di: Ren, Ruiping, et al.
Pubblicazione: (2024)
di: Ren, Ruiping, et al.
Pubblicazione: (2024)
White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
"Not in My Backyard": LLMs Uncover Online and Offline Social Biases Against Homelessness
di: Karr Jr., Jonathan A., et al.
Pubblicazione: (2025)
di: Karr Jr., Jonathan A., et al.
Pubblicazione: (2025)
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
di: Hirose, Manari, et al.
Pubblicazione: (2025)
di: Hirose, Manari, et al.
Pubblicazione: (2025)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
di: Wu, Sirui, et al.
Pubblicazione: (2025)
di: Wu, Sirui, et al.
Pubblicazione: (2025)
News and Load: A Quantitative Exploration of Natural Language Processing Applications for Forecasting Day-ahead Electricity System Demand
di: Bai, Yun, et al.
Pubblicazione: (2023)
di: Bai, Yun, et al.
Pubblicazione: (2023)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
di: Vida, Karina, et al.
Pubblicazione: (2024)
di: Vida, Karina, et al.
Pubblicazione: (2024)
Prompt Perturbations Reveal Human-Like Biases in Large Language Model Survey Responses
di: Rupprecht, Jens, et al.
Pubblicazione: (2025)
di: Rupprecht, Jens, et al.
Pubblicazione: (2025)
Empowering Many, Biasing a Few: Generalist Credit Scoring through Large Language Models
di: Feng, Duanyu, et al.
Pubblicazione: (2023)
di: Feng, Duanyu, et al.
Pubblicazione: (2023)
Measuring Machine Learning Harms from Stereotypes Requires Understanding Who Is Harmed by Which Errors in What Ways
di: Wang, Angelina, et al.
Pubblicazione: (2024)
di: Wang, Angelina, et al.
Pubblicazione: (2024)
Large Language Models for Education: A Survey
di: Xu, Hanyi, et al.
Pubblicazione: (2024)
di: Xu, Hanyi, et al.
Pubblicazione: (2024)
No Free Lunch in Language Model Bias Mitigation? Targeted Bias Reduction Can Exacerbate Unmitigated LLM Biases
di: Chand, Shireen, et al.
Pubblicazione: (2025)
di: Chand, Shireen, et al.
Pubblicazione: (2025)
Fine-Grained Behavior Simulation with Role-Playing Large Language Model on Social Media
di: Li, Kun, et al.
Pubblicazione: (2024)
di: Li, Kun, et al.
Pubblicazione: (2024)
Prompt Selection Matters: Enhancing Text Annotations for Social Sciences with Large Language Models
di: Abraham, Louis, et al.
Pubblicazione: (2024)
di: Abraham, Louis, et al.
Pubblicazione: (2024)
Self-Alignment of Large Language Models via Monopolylogue-based Social Scene Simulation
di: Pang, Xianghe, et al.
Pubblicazione: (2024)
di: Pang, Xianghe, et al.
Pubblicazione: (2024)
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
di: Hu, Yiran, et al.
Pubblicazione: (2026)
di: Hu, Yiran, et al.
Pubblicazione: (2026)
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
Sentient Agent as a Judge: Evaluating Higher-Order Social Cognition in Large Language Models
di: Zhang, Bang, et al.
Pubblicazione: (2025)
di: Zhang, Bang, et al.
Pubblicazione: (2025)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
IROTE: Human-like Traits Elicitation of Large Language Model via In-Context Self-Reflective Optimization
di: Bai, Yuzhuo, et al.
Pubblicazione: (2025)
di: Bai, Yuzhuo, et al.
Pubblicazione: (2025)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
di: Shieh, Evan, et al.
Pubblicazione: (2024)
di: Shieh, Evan, et al.
Pubblicazione: (2024)
Large Language Models Leverage External Knowledge to Extend Clinical Insight Beyond Language Boundaries
di: Wu, Jiageng, et al.
Pubblicazione: (2023)
di: Wu, Jiageng, et al.
Pubblicazione: (2023)
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
di: Bae, Suyoung, et al.
Pubblicazione: (2025)
di: Bae, Suyoung, et al.
Pubblicazione: (2025)
Large Language Model for Mental Health: A Systematic Review
di: Guo, Zhijun, et al.
Pubblicazione: (2024)
di: Guo, Zhijun, et al.
Pubblicazione: (2024)
Codebook Reduction and Saturation: Novel observations on Inductive Thematic Saturation for Large Language Models and initial coding in Thematic Analysis
di: De Paoli, Stefano, et al.
Pubblicazione: (2025)
di: De Paoli, Stefano, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
di: Wu, Addison J., et al.
Pubblicazione: (2026) -
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
di: Bai, Xuechunzi, et al.
Pubblicazione: (2024) -
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
di: Sun, Lihao, et al.
Pubblicazione: (2025) -
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
di: Liu, Ryan, et al.
Pubblicazione: (2024) -
Large Language Models Assume People are More Rational than We Really are
di: Liu, Ryan, et al.
Pubblicazione: (2024)