MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rosen, Simon, Singh, Siddarth, Gelo, Ebenezer, Robertson, Helen Sarah, Suder, Ibrahim, Williams, Victoria, Rosman, Benjamin, Tasse, Geraud Nangue, James, Steven |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RobocupGym: A challenging continuous control benchmark in Robocup
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
Unsupervised Hierarchical Skill Discovery
von: Harvey, Damion, et al.
Veröffentlicht: (2026)
von: Harvey, Damion, et al.
Veröffentlicht: (2026)
Counting Reward Automata: Sample Efficient Reinforcement Learning Through the Exploitation of Reward Function Structure
von: Bester, Tristan, et al.
Veröffentlicht: (2023)
von: Bester, Tristan, et al.
Veröffentlicht: (2023)
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2022)
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2022)
Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2025)
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2025)
Compositional Instruction Following with Language Models and Reinforcement Learning
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
Smart But Not Moral? Moral Alignment In Human-AI Decision-Making
von: Ernst, Christiane, et al.
Veröffentlicht: (2026)
von: Ernst, Christiane, et al.
Veröffentlicht: (2026)
ProgressGym: Alignment with a Millennium of Moral Progress
von: Qiu, Tianyi, et al.
Veröffentlicht: (2024)
von: Qiu, Tianyi, et al.
Veröffentlicht: (2024)
The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making
von: Garcia, Basile, et al.
Veröffentlicht: (2024)
von: Garcia, Basile, et al.
Veröffentlicht: (2024)
Building Interpretable Models for Moral Decision-Making
von: Goel, Mayank, et al.
Veröffentlicht: (2026)
von: Goel, Mayank, et al.
Veröffentlicht: (2026)
Heartificial Intelligence: Exploring Empathy in Language Models
von: Williams, Victoria, et al.
Veröffentlicht: (2025)
von: Williams, Victoria, et al.
Veröffentlicht: (2025)
Interoceptive Brain Processing Influences Moral Decision Making
von: Shengbin Cui, et al.
Veröffentlicht: (2024)
von: Shengbin Cui, et al.
Veröffentlicht: (2024)
MoralReason: Generalizable Moral Decision Alignment For LLM Agents Using Reasoning-Level Reinforcement Learning
von: An, Zhiyu, et al.
Veröffentlicht: (2025)
von: An, Zhiyu, et al.
Veröffentlicht: (2025)
Life and Death in Nazi Germany: Moral Decision Making in the Classroom.
von: Mork, Gordon R.
Veröffentlicht: (1983)
von: Mork, Gordon R.
Veröffentlicht: (1983)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
von: Hori, Toshiaki, et al.
Veröffentlicht: (2025)
von: Hori, Toshiaki, et al.
Veröffentlicht: (2025)
Making Moral Judgments
von: Forsyth, Donelson
Veröffentlicht: (2022)
von: Forsyth, Donelson
Veröffentlicht: (2022)
Systematic Deviations in Atomic Spectra: A New Analysis Method Based on the Harmonic Number Unit
von: Suder, Marek
Veröffentlicht: (2025)
von: Suder, Marek
Veröffentlicht: (2025)
An Equivalent Representation of Generalized Differentials
von: Suder, Valentin
Veröffentlicht: (2025)
von: Suder, Valentin
Veröffentlicht: (2025)
Universities, ‘Left Behind Places’ and the Making of a Moral Crisis
von: Sarah Chaytor, et al.
Veröffentlicht: (2025)
von: Sarah Chaytor, et al.
Veröffentlicht: (2025)
Integrating Reason-Based Moral Decision-Making in the Reinforcement Learning Architecture
von: Dargasz, Lisa
Veröffentlicht: (2025)
von: Dargasz, Lisa
Veröffentlicht: (2025)
Adolescents Are More Utilitarian Than Adults in Group Moral Decision‐Making
von: Yingying Jiang, et al.
Veröffentlicht: (2024)
von: Yingying Jiang, et al.
Veröffentlicht: (2024)
Histoires Morales: A French Dataset for Assessing Moral Alignment
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
Moral Alignment for LLM Agents
von: Tennant, Elizaveta, et al.
Veröffentlicht: (2024)
von: Tennant, Elizaveta, et al.
Veröffentlicht: (2024)
Addressing Moral Uncertainty using Large Language Models for Ethical Decision-Making
von: Dubey, Rohit K., et al.
Veröffentlicht: (2025)
von: Dubey, Rohit K., et al.
Veröffentlicht: (2025)
How Responses to Sexual Harassment and Moral Values Shape Investment Decision Making
von: Amal P. Abeysekera, et al.
Veröffentlicht: (2026)
von: Amal P. Abeysekera, et al.
Veröffentlicht: (2026)
The Greatest Good Benchmark: Measuring LLMs' Alignment with Utilitarian Moral Dilemmas
von: Marraffini, Giovanni Franco Gabriel, et al.
Veröffentlicht: (2025)
von: Marraffini, Giovanni Franco Gabriel, et al.
Veröffentlicht: (2025)
MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
von: Lin, Xiao, et al.
Veröffentlicht: (2025)
von: Lin, Xiao, et al.
Veröffentlicht: (2025)
Moral Economies of Corruption
von: Pierce, Steven
Veröffentlicht: (2016)
von: Pierce, Steven
Veröffentlicht: (2016)
GymPN: A Library for Decision-Making in Process Management Systems
von: Bianco, Riccardo Lo, et al.
Veröffentlicht: (2025)
von: Bianco, Riccardo Lo, et al.
Veröffentlicht: (2025)
MoralCLIP: Contrastive Alignment of Vision-and-Language Representations with Moral Foundations Theory
von: Condez, Ana Carolina, et al.
Veröffentlicht: (2025)
von: Condez, Ana Carolina, et al.
Veröffentlicht: (2025)
Moral Decision‐Making Under Ego Depletion in Virtual Reality: The Buffering Role of Emotion
von: Yanglei Cao, et al.
Veröffentlicht: (2026)
von: Yanglei Cao, et al.
Veröffentlicht: (2026)
Approaches to Ethical Decision‐Making: Contrasting Rationality‐Based Models Versus Moral Intuitionism
von: Seyda Deligonul, et al.
Veröffentlicht: (2025)
von: Seyda Deligonul, et al.
Veröffentlicht: (2025)
Can AI Model the Complexities of Human Moral Decision-Making? A Qualitative Study of Kidney Allocation Decisions
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
Making Moral Judgments on Adequate Grounds: From Transparent Evidence to Justified Moral Belief
von: Shlomit Wygoda Cohen
Veröffentlicht: (2025)
von: Shlomit Wygoda Cohen
Veröffentlicht: (2025)
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2026)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2026)
MTBBench: A Multimodal Sequential Clinical Decision-Making Benchmark in Oncology
von: Vasilev, Kiril, et al.
Veröffentlicht: (2025)
von: Vasilev, Kiril, et al.
Veröffentlicht: (2025)
KellyBench: A Benchmark for Long-Horizon Sequential Decision Making
von: Grady, Thomas, et al.
Veröffentlicht: (2026)
von: Grady, Thomas, et al.
Veröffentlicht: (2026)
A Note on Homelessness and the Moral Economy: How Thompson's Moral Economy Presents in Modern Day Homelessness
von: Sarah Werman
Veröffentlicht: (2024)
von: Sarah Werman
Veröffentlicht: (2024)
Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions
von: Kim, Jiseon, et al.
Veröffentlicht: (2026)
von: Kim, Jiseon, et al.
Veröffentlicht: (2026)
MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
Ähnliche Einträge
-
RobocupGym: A challenging continuous control benchmark in Robocup
von: Beukman, Michael, et al.
Veröffentlicht: (2024) -
Unsupervised Hierarchical Skill Discovery
von: Harvey, Damion, et al.
Veröffentlicht: (2026) -
Counting Reward Automata: Sample Efficient Reinforcement Learning Through the Exploitation of Reward Function Structure
von: Bester, Tristan, et al.
Veröffentlicht: (2023) -
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2022) -
Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
von: Tasse, Geraud Nangue, et al.
Veröffentlicht: (2025)