LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Jiao, Junfeng, Afroogh, Saleh, Murali, Abhejay, Chen, Kevin, Atkinson, David, Dhurandhar, Amit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
Evaluating LLM Safety Across Child Development Stages: A Simulated Agent Approach
by: Murali, Abhejay, et al.
Published: (2025)
by: Murali, Abhejay, et al.
Published: (2025)
LLM Harms: A Taxonomy and Discussion
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
LLMs and Childhood Safety: Identifying Risks and Proposing a Protection Framework for Safe Child-LLM Interaction
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
The global landscape of academic guidelines for generative AI and Large Language Models
by: Jiao, Junfeng, et al.
Published: (2024)
by: Jiao, Junfeng, et al.
Published: (2024)
AI Empathy Erodes Cognitive Autonomy in Younger Users
by: Jiao, Junfeng, et al.
Published: (2026)
by: Jiao, Junfeng, et al.
Published: (2026)
Generative AI and LLMs in Industry: A text-mining Analysis and Critical Evaluation of Guidelines and Policy Statements Across Fourteen Industrial Sectors
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
IGGA: A Dataset of Industrial Guidelines and Policy Statements for Generative AIs
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
Evaluating the Effectiveness of OpenAI's Parental Control System
by: Ersoz, Kerem, et al.
Published: (2026)
by: Ersoz, Kerem, et al.
Published: (2026)
Navigating LLM Ethics: Advancements, Challenges, and Future Directions
by: Jiao, Junfeng, et al.
Published: (2024)
by: Jiao, Junfeng, et al.
Published: (2024)
Mapping out AI Functions in Intelligent Disaster (Mis)Management and AI-Caused Disasters
by: Pouresmaeil, Yasser, et al.
Published: (2025)
by: Pouresmaeil, Yasser, et al.
Published: (2025)
Intelligent Environmental Empathy (IEE): A new power and platform to fostering green obligation for climate peace and justice
by: Afroogh, Saleh, et al.
Published: (2024)
by: Afroogh, Saleh, et al.
Published: (2024)
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
by: Park, Jihyung, et al.
Published: (2025)
by: Park, Jihyung, et al.
Published: (2025)
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
by: Park, Jihyung, et al.
Published: (2026)
by: Park, Jihyung, et al.
Published: (2026)
Evaluating Retrieval-Augmented Generation Strategies for Large Language Models in Travel Mode Choice Prediction
by: Xu, Yiming, et al.
Published: (2025)
by: Xu, Yiming, et al.
Published: (2025)
A Task-Driven Human-AI Collaboration: When to Automate, When to Collaborate, When to Challenge
by: Afroogh, Saleh, et al.
Published: (2025)
by: Afroogh, Saleh, et al.
Published: (2025)
When Trust is Zero Sum: Automation Threat to Epistemic Agency
by: Malone, Emmie, et al.
Published: (2024)
by: Malone, Emmie, et al.
Published: (2024)
GreedLlama: Performance of Financial Value-Aligned Large Language Models in Moral Reasoning
by: Yu, Jeffy, et al.
Published: (2024)
by: Yu, Jeffy, et al.
Published: (2024)
Trust in AI: Progress, Challenges, and Future Directions
by: Afroogh, Saleh, et al.
Published: (2024)
by: Afroogh, Saleh, et al.
Published: (2024)
Ethical Risks of Large Language Models in Medical Consultation: An Assessment Based on Reproductive Ethics
by: Xu, Hanhui, et al.
Published: (2026)
by: Xu, Hanhui, et al.
Published: (2026)
Ethical Risks in Deploying Large Language Models: An Evaluation of Medical Ethics Jailbreaking
by: Huang, Chutian, et al.
Published: (2026)
by: Huang, Chutian, et al.
Published: (2026)
Student Perspectives on Using a Large Language Model (LLM) for an Assignment on Professional Ethics
by: Grande, Virginia, et al.
Published: (2024)
by: Grande, Virginia, et al.
Published: (2024)
The Convergent Ethics of AI? Analyzing Moral Foundation Priorities in Large Language Models with a Multi-Framework Approach
by: Coleman, Chad, et al.
Published: (2025)
by: Coleman, Chad, et al.
Published: (2025)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
by: Backmann, Steffen, et al.
Published: (2025)
by: Backmann, Steffen, et al.
Published: (2025)
Normative Evaluation of Large Language Models with Everyday Moral Dilemmas
by: Sachdeva, Pratik S., et al.
Published: (2025)
by: Sachdeva, Pratik S., et al.
Published: (2025)
ClarityEthic: Explainable Moral Judgment Utilizing Contrastive Ethical Insights from Large Language Models
by: Sun, Yuxi, et al.
Published: (2024)
by: Sun, Yuxi, et al.
Published: (2024)
AccessEval: Benchmarking Disability Bias in Large Language Models
by: Panda, Srikant, et al.
Published: (2025)
by: Panda, Srikant, et al.
Published: (2025)
Benchmarking Large Language Models on Homework Assessment in Circuit Analysis
by: Chen, Liangliang, et al.
Published: (2025)
by: Chen, Liangliang, et al.
Published: (2025)
The Moral Mind(s) of Large Language Models
by: Seror, Avner
Published: (2024)
by: Seror, Avner
Published: (2024)
Putting GenAI on Notice: GenAI Exceptionalism and Contract Law
by: Atkinson, David
Published: (2025)
by: Atkinson, David
Published: (2025)
Three Kinds of AI Ethics
by: Ratti, Emanuele
Published: (2025)
by: Ratti, Emanuele
Published: (2025)
Evaluating 21st-Century Competencies in Postsecondary Curricula with Large Language Models: Performance Benchmarking and Reasoning-Based Prompting Strategies
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
Differences in the Moral Foundations of Large Language Models
by: Kirgis, Peter
Published: (2025)
by: Kirgis, Peter
Published: (2025)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
by: Wu, Ya, et al.
Published: (2025)
by: Wu, Ya, et al.
Published: (2025)
CELL your Model: Contrastive Explanations for Large Language Models
by: Luss, Ronny, et al.
Published: (2024)
by: Luss, Ronny, et al.
Published: (2024)
Unfair Learning: GenAI Exceptionalism and Copyright Law
by: Atkinson, David
Published: (2025)
by: Atkinson, David
Published: (2025)
Ethics Whitepaper: Whitepaper on Ethical Research into Large Language Models
by: Ungless, Eddie L., et al.
Published: (2024)
by: Ungless, Eddie L., et al.
Published: (2024)
Inducing Human-like Biases in Moral Reasoning Language Models
by: Karpov, Artem, et al.
Published: (2024)
by: Karpov, Artem, et al.
Published: (2024)
Frontier AI Ethics: Anticipating and Evaluating the Societal Impacts of Language Model Agents
by: Lazar, Seth
Published: (2024)
by: Lazar, Seth
Published: (2024)
Similar Items
-
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
by: Jiao, Junfeng, et al.
Published: (2025) -
Evaluating LLM Safety Across Child Development Stages: A Simulated Agent Approach
by: Murali, Abhejay, et al.
Published: (2025) -
LLM Harms: A Taxonomy and Discussion
by: Chen, Kevin, et al.
Published: (2025) -
LLMs and Childhood Safety: Identifying Risks and Proposing a Protection Framework for Safe Child-LLM Interaction
by: Jiao, Junfeng, et al.
Published: (2025) -
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
by: Jiao, Junfeng, et al.
Published: (2025)