Discerning What Matters: A Multi-Dimensional Assessment of Moral Competence in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Kilov, Daniel, Hendy, Caroline, Guyot, Secil Yanik, Snoswell, Aaron J., Lazar, Seth |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMs
por: Yeh, Eric, et al.
Publicado: (2025)
por: Yeh, Eric, et al.
Publicado: (2025)
Beyond Mimicry: Preference Coherence in LLMs
por: Mikaelson, Luhan, et al.
Publicado: (2025)
por: Mikaelson, Luhan, et al.
Publicado: (2025)
The Continuity Layer: Why Intelligence Needs an Architecture for What It Carries Forward
por: Tanguturi, Samuel Sameer
Publicado: (2026)
por: Tanguturi, Samuel Sameer
Publicado: (2026)
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
por: Minai, Ali A.
Publicado: (2025)
por: Minai, Ali A.
Publicado: (2025)
Complete Implementation of WXF Chinese Chess Rules
por: Tan, Daniel, et al.
Publicado: (2024)
por: Tan, Daniel, et al.
Publicado: (2024)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
por: Bayarri-Planas, Jordi, et al.
Publicado: (2024)
por: Bayarri-Planas, Jordi, et al.
Publicado: (2024)
Moral Responsibility or Obedience: What Do We Want from AI?
por: Boland, Joseph
Publicado: (2025)
por: Boland, Joseph
Publicado: (2025)
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
por: Singh, Siddarth, et al.
Publicado: (2023)
por: Singh, Siddarth, et al.
Publicado: (2023)
How Data Quality Affects Machine Learning Models for Credit Risk Assessment
por: Maurino, Andrea
Publicado: (2025)
por: Maurino, Andrea
Publicado: (2025)
HCAST: Human-Calibrated Autonomy Software Tasks
por: Rein, David, et al.
Publicado: (2025)
por: Rein, David, et al.
Publicado: (2025)
What Does 'Human-Centred AI' Mean?
por: Guest, Olivia
Publicado: (2025)
por: Guest, Olivia
Publicado: (2025)
Beyond Direct Generation: A Decomposed Approach to Well-Crafted Screenwriting with LLMs
por: Lei, Hang, et al.
Publicado: (2025)
por: Lei, Hang, et al.
Publicado: (2025)
Mutagenesis screen to map the functions of parameters of Large Language Models
por: Hu, Yue, et al.
Publicado: (2024)
por: Hu, Yue, et al.
Publicado: (2024)
Appraisal-Guided Proximal Policy Optimization: Modeling Psychological Disorders in Dynamic Grid World
por: Prasad, Hari, et al.
Publicado: (2024)
por: Prasad, Hari, et al.
Publicado: (2024)
Machine Learning and Theory Ladenness -- A Phenomenological Account
por: Termine, Alberto, et al.
Publicado: (2024)
por: Termine, Alberto, et al.
Publicado: (2024)
A Case-Based Persistent Memory for a Large Language Model
por: Watson, Ian
Publicado: (2023)
por: Watson, Ian
Publicado: (2023)
Self-evolving expertise in complex non-verifiable subject domains: dialogue as implicit meta-RL
por: Bailey, Richard M.
Publicado: (2025)
por: Bailey, Richard M.
Publicado: (2025)
Reward is not enough: can we liberate AI from the reinforcement learning paradigm?
por: Glukhov, Vacslav
Publicado: (2022)
por: Glukhov, Vacslav
Publicado: (2022)
From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments
por: Luo, Lijing, et al.
Publicado: (2026)
por: Luo, Lijing, et al.
Publicado: (2026)
Do Chains-of-Thoughts of Large Language Models Suffer from Hallucinations, Cognitive Biases, or Phobias in Bayesian Reasoning?
por: Araya, Roberto
Publicado: (2025)
por: Araya, Roberto
Publicado: (2025)
Developing trustworthy AI applications with foundation models
por: Mock, Michael, et al.
Publicado: (2024)
por: Mock, Michael, et al.
Publicado: (2024)
Sensemaking in Novel Environments: How Human Cognition Can Inform Artificial Agents
por: Patterson, Robert E., et al.
Publicado: (2025)
por: Patterson, Robert E., et al.
Publicado: (2025)
How VADER is your AI? Towards a definition of artificial intelligence systems appropriate for regulation
por: Bezerra, Leonardo C. T., et al.
Publicado: (2024)
por: Bezerra, Leonardo C. T., et al.
Publicado: (2024)
What's my role? Modelling responsibility for AI-based safety-critical systems
por: Ryan, Philippa, et al.
Publicado: (2023)
por: Ryan, Philippa, et al.
Publicado: (2023)
Reasoning Beyond the Obvious: Evaluating Divergent and Convergent Thinking in LLMs for Financial Scenarios
por: Bok, Zhuang Qiang, et al.
Publicado: (2025)
por: Bok, Zhuang Qiang, et al.
Publicado: (2025)
MultiPruner: Balanced Structure Removal in Foundation Models
por: Muñoz, J. Pablo, et al.
Publicado: (2025)
por: Muñoz, J. Pablo, et al.
Publicado: (2025)
The Station: An Open-World Environment for AI-Driven Discovery
por: Chung, Stephen, et al.
Publicado: (2025)
por: Chung, Stephen, et al.
Publicado: (2025)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
por: Mahjoub, Omayma, et al.
Publicado: (2023)
por: Mahjoub, Omayma, et al.
Publicado: (2023)
Study of the Proper NNUE Dataset
por: Tan, Daniel, et al.
Publicado: (2024)
por: Tan, Daniel, et al.
Publicado: (2024)
Fanar: An Arabic-Centric Multimodal Generative AI Platform
por: Fanar Team, et al.
Publicado: (2025)
por: Fanar Team, et al.
Publicado: (2025)
EduQate: Generating Adaptive Curricula through RMABs in Education Settings
por: Tio, Sidney, et al.
Publicado: (2024)
por: Tio, Sidney, et al.
Publicado: (2024)
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
por: McCane, Brendan
Publicado: (2026)
por: McCane, Brendan
Publicado: (2026)
Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
por: Pant, Aakash, et al.
Publicado: (2026)
por: Pant, Aakash, et al.
Publicado: (2026)
Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems
por: Lavi, Gadi
Publicado: (2026)
por: Lavi, Gadi
Publicado: (2026)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
por: Kim, Sejin, et al.
Publicado: (2025)
por: Kim, Sejin, et al.
Publicado: (2025)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
por: Wu, Dekun, et al.
Publicado: (2023)
por: Wu, Dekun, et al.
Publicado: (2023)
Aristotle's Original Idea: For and Against Logic in the era of AI
por: Kakas, Antonis C.
Publicado: (2025)
por: Kakas, Antonis C.
Publicado: (2025)
Enhancing Multi-Agent Collaboration with Attention-Based Actor-Critic Policies
por: Belinchon, Hugo Garrido-Lestache, et al.
Publicado: (2025)
por: Belinchon, Hugo Garrido-Lestache, et al.
Publicado: (2025)
Personality-Driven Decision-Making in LLM-Based Autonomous Agents
por: Newsham, Lewis, et al.
Publicado: (2025)
por: Newsham, Lewis, et al.
Publicado: (2025)
Pose Matters: Evaluating Vision Transformers and CNNs for Human Action Recognition on Small COCO Subsets
por: Tang, MingZe, et al.
Publicado: (2025)
por: Tang, MingZe, et al.
Publicado: (2025)
Ejemplares similares
-
Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMs
por: Yeh, Eric, et al.
Publicado: (2025) -
Beyond Mimicry: Preference Coherence in LLMs
por: Mikaelson, Luhan, et al.
Publicado: (2025) -
The Continuity Layer: Why Intelligence Needs an Architecture for What It Carries Forward
por: Tanguturi, Samuel Sameer
Publicado: (2026) -
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
por: Minai, Ali A.
Publicado: (2025) -
Complete Implementation of WXF Chinese Chess Rules
por: Tan, Daniel, et al.
Publicado: (2024)