"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Junchen, Jiang, Penghao, Xu, Zihao, Ding, Ziqi, Zhu, Yichen, Jiang, Jiaojiao, Li, Yuekang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TombRaider: Entering the Vault of History to Jailbreak Large Language Models
by: Ding, Junchen, et al.
Published: (2025)
by: Ding, Junchen, et al.
Published: (2025)
Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-based Test Oracles
by: Xu, Zihao, et al.
Published: (2025)
by: Xu, Zihao, et al.
Published: (2025)
Investigating Political and Demographic Associations in Large Language Models Through Moral Foundations Theory
by: Smith-Vaniz, Nicole, et al.
Published: (2025)
by: Smith-Vaniz, Nicole, et al.
Published: (2025)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
by: Ren, Ruiping, et al.
Published: (2024)
by: Ren, Ruiping, et al.
Published: (2024)
NGCaptcha: A CAPTCHA Bridging the Past and the Future
by: Ding, Ziqi, et al.
Published: (2025)
by: Ding, Ziqi, et al.
Published: (2025)
The Algorithmic Unconscious: Structural Mechanisms and Implicit Biases in Large Language Models
by: Boisnard, Philippe
Published: (2026)
by: Boisnard, Philippe
Published: (2026)
The Life Cycle of Large Language Models: A Review of Biases in Education
by: Lee, Jinsook, et al.
Published: (2024)
by: Lee, Jinsook, et al.
Published: (2024)
A Systematic Analysis of Biases in Large Language Models
by: Zhang, Xulang, et al.
Published: (2025)
by: Zhang, Xulang, et al.
Published: (2025)
Understanding Intrinsic Socioeconomic Biases in Large Language Models
by: Arzaghi, Mina, et al.
Published: (2024)
by: Arzaghi, Mina, et al.
Published: (2024)
Whose Emotions and Moral Sentiments Do Language Models Reflect?
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Large Language Models are Geographically Biased
by: Manvi, Rohin, et al.
Published: (2024)
by: Manvi, Rohin, et al.
Published: (2024)
The Moral Machine Experiment on Large Language Models
by: Takemoto, Kazuhiro
Published: (2023)
by: Takemoto, Kazuhiro
Published: (2023)
Is Self-knowledge and Action Consistent or Not: Investigating Large Language Model's Personality
by: Ai, Yiming, et al.
Published: (2024)
by: Ai, Yiming, et al.
Published: (2024)
Investigating Cultural Alignment of Large Language Models
by: AlKhamissi, Badr, et al.
Published: (2024)
by: AlKhamissi, Badr, et al.
Published: (2024)
PRISM: A Methodology for Auditing Biases in Large Language Models
by: Azzopardi, Leif, et al.
Published: (2024)
by: Azzopardi, Leif, et al.
Published: (2024)
Ethics Whitepaper: Whitepaper on Ethical Research into Large Language Models
by: Ungless, Eddie L., et al.
Published: (2024)
by: Ungless, Eddie L., et al.
Published: (2024)
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models
by: Khandelwal, Khyati, et al.
Published: (2023)
by: Khandelwal, Khyati, et al.
Published: (2023)
Role-Play Paradox in Large Language Models: Reasoning Performance Gains and Ethical Dilemmas
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
by: Wu, Ya, et al.
Published: (2025)
by: Wu, Ya, et al.
Published: (2025)
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination
by: Li, Zihao
Published: (2023)
by: Li, Zihao
Published: (2023)
How Do Language Models Process Ethical Instructions? Deliberation, Consistency, and Other-Recognition Across Four Models
by: Fukui, Hiroki
Published: (2026)
by: Fukui, Hiroki
Published: (2026)
Generative Language Models Exhibit Social Identity Biases
by: Hu, Tiancheng, et al.
Published: (2023)
by: Hu, Tiancheng, et al.
Published: (2023)
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
by: Jiang, Han, et al.
Published: (2024)
by: Jiang, Han, et al.
Published: (2024)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
by: Vida, Karina, et al.
Published: (2024)
by: Vida, Karina, et al.
Published: (2024)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
by: Chen, Zhongren, et al.
Published: (2026)
by: Chen, Zhongren, et al.
Published: (2026)
The Moral Gap of Large Language Models
by: Skorski, Maciej, et al.
Published: (2025)
by: Skorski, Maciej, et al.
Published: (2025)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
by: Backmann, Steffen, et al.
Published: (2025)
by: Backmann, Steffen, et al.
Published: (2025)
PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models
by: Bao, Han, et al.
Published: (2026)
by: Bao, Han, et al.
Published: (2026)
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
by: Wu, Addison J., et al.
Published: (2025)
by: Wu, Addison J., et al.
Published: (2025)
A Toolbox for Surfacing Health Equity Harms and Biases in Large Language Models
by: Pfohl, Stephen R., et al.
Published: (2024)
by: Pfohl, Stephen R., et al.
Published: (2024)
White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs
by: Wan, Yixin, et al.
Published: (2024)
by: Wan, Yixin, et al.
Published: (2024)
Are Language Models Sensitive to Morally Irrelevant Distractors?
by: Shaw, Andrew, et al.
Published: (2026)
by: Shaw, Andrew, et al.
Published: (2026)
Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey
by: Deng, Chengyuan, et al.
Published: (2024)
by: Deng, Chengyuan, et al.
Published: (2024)
DiverseClaire: Simulating Students to Improve Introductory Programming Course Materials for All CS1 Learners
by: Wong, Wendy, et al.
Published: (2025)
by: Wong, Wendy, et al.
Published: (2025)
MoralBERT: A Fine-Tuned Language Model for Capturing Moral Values in Social Discussions
by: Preniqi, Vjosa, et al.
Published: (2024)
by: Preniqi, Vjosa, et al.
Published: (2024)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
by: Christian, Brian, et al.
Published: (2026)
by: Christian, Brian, et al.
Published: (2026)
Transformers and Cortical Waves: Encoders for Pulling In Context Across Time
by: Muller, Lyle, et al.
Published: (2024)
by: Muller, Lyle, et al.
Published: (2024)
Climate Change from Large Language Models
by: Zhu, Hongyin, et al.
Published: (2023)
by: Zhu, Hongyin, et al.
Published: (2023)
Denevil: Towards Deciphering and Navigating the Ethical Values of Large Language Models via Instruction Learning
by: Duan, Shitong, et al.
Published: (2023)
by: Duan, Shitong, et al.
Published: (2023)
Assessing the Impact of Conspiracy Theories Using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Similar Items
-
TombRaider: Entering the Vault of History to Jailbreak Large Language Models
by: Ding, Junchen, et al.
Published: (2025) -
Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-based Test Oracles
by: Xu, Zihao, et al.
Published: (2025) -
Investigating Political and Demographic Associations in Large Language Models Through Moral Foundations Theory
by: Smith-Vaniz, Nicole, et al.
Published: (2025) -
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
by: Ren, Ruiping, et al.
Published: (2024) -
NGCaptcha: A CAPTCHA Bridging the Past and the Future
by: Ding, Ziqi, et al.
Published: (2025)