Kantian Deontology Meets AI Alignment: Towards Morally Grounded Fairness Metrics
Fuente:
arXiv
Saved in:
| Main Authors: | Mougan, Carlos, Brand, Joshua |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment
by: Huang, Allison, et al.
Published: (2024)
by: Huang, Allison, et al.
Published: (2024)
Kantian-Utilitarian XAI: Meta-Explained
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Combining Theory of Mind and Kindness for Self-Supervised Human-AI Alignment
by: Hewson, Joshua T. S.
Published: (2024)
by: Hewson, Joshua T. S.
Published: (2024)
Fairness Metric Design Exploration in Multi-Domain Moral Sentiment Classification using Transformer-Based Models
by: Naranbat, Battemuulen, et al.
Published: (2025)
by: Naranbat, Battemuulen, et al.
Published: (2025)
Moral Anchor System: A Predictive Framework for AI Value Alignment and Drift Prevention
by: Ravindran, Santhosh Kumar
Published: (2025)
by: Ravindran, Santhosh Kumar
Published: (2025)
Public Perceptions of Fairness Metrics Across Borders
by: Sasaki, Yuya, et al.
Published: (2024)
by: Sasaki, Yuya, et al.
Published: (2024)
Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment
by: Bajaj, Tanav Singh, et al.
Published: (2026)
by: Bajaj, Tanav Singh, et al.
Published: (2026)
Formalizing Kantian Ethics: Formula of the Universal Law Logic (FULL)
by: Olson, Taylor
Published: (2026)
by: Olson, Taylor
Published: (2026)
AI Alignment via Incentives and Correction
by: Agarwal, Rohit, et al.
Published: (2026)
by: Agarwal, Rohit, et al.
Published: (2026)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
by: Pasandi, Faezeh B., et al.
Published: (2026)
by: Pasandi, Faezeh B., et al.
Published: (2026)
Wide Reflective Equilibrium in LLM Alignment: Bridging Moral Epistemology and AI Safety
by: Brophy, Matthew
Published: (2025)
by: Brophy, Matthew
Published: (2025)
When Large Language Model Agents Meet 6G Networks: Perception, Grounding, and Alignment
by: Xu, Minrui, et al.
Published: (2024)
by: Xu, Minrui, et al.
Published: (2024)
Histoires Morales: A French Dataset for Assessing Moral Alignment
by: Leteno, Thibaud, et al.
Published: (2025)
by: Leteno, Thibaud, et al.
Published: (2025)
Towards Clinical AI Fairness: Filling Gaps in the Puzzle
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
Fair Clustering via Alignment
by: Kim, Kunwoong, et al.
Published: (2025)
by: Kim, Kunwoong, et al.
Published: (2025)
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
by: Bezou-Vrakatseli, Elfia, et al.
Published: (2024)
by: Bezou-Vrakatseli, Elfia, et al.
Published: (2024)
Probing the Probes: Methods and Metrics for Concept Alignment
by: Lysnæs-Larsen, Jacob, et al.
Published: (2025)
by: Lysnæs-Larsen, Jacob, et al.
Published: (2025)
A Review of Fairness and A Practical Guide to Selecting Context-Appropriate Fairness Metrics in Machine Learning
by: Barr, Caleb J. S., et al.
Published: (2024)
by: Barr, Caleb J. S., et al.
Published: (2024)
MoralReason: Generalizable Moral Decision Alignment For LLM Agents Using Reasoning-Level Reinforcement Learning
by: An, Zhiyu, et al.
Published: (2025)
by: An, Zhiyu, et al.
Published: (2025)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
by: Muslimani, Calarina, et al.
Published: (2025)
by: Muslimani, Calarina, et al.
Published: (2025)
Beyond Ethical Alignment: Evaluating LLMs as Artificial Moral Assistants
by: Galatolo, Alessio, et al.
Published: (2025)
by: Galatolo, Alessio, et al.
Published: (2025)
Moral Alignment for LLM Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Are Language Models Consequentialist or Deontological Moral Reasoners?
by: Samway, Keenan, et al.
Published: (2025)
by: Samway, Keenan, et al.
Published: (2025)
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
by: Okutomi, Akira
Published: (2025)
by: Okutomi, Akira
Published: (2025)
Towards Responsible AI: Advances in Safety, Fairness, and Accountability of Autonomous Systems
by: Cano, Filip
Published: (2025)
by: Cano, Filip
Published: (2025)
Metric-Fair Prompting: Treating Similar Samples Similarly
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
Resource Rational Contractualism Should Guide AI Alignment
by: Levine, Sydney, et al.
Published: (2025)
by: Levine, Sydney, et al.
Published: (2025)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
by: Tennant, Elizaveta, et al.
Published: (2023)
by: Tennant, Elizaveta, et al.
Published: (2023)
The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making
by: Garcia, Basile, et al.
Published: (2024)
by: Garcia, Basile, et al.
Published: (2024)
Towards Execution-Grounded Automated AI Research
by: Si, Chenglei, et al.
Published: (2026)
by: Si, Chenglei, et al.
Published: (2026)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
by: Chen, Benjamin Minhao, et al.
Published: (2026)
by: Chen, Benjamin Minhao, et al.
Published: (2026)
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
by: Rosen, Simon, et al.
Published: (2026)
by: Rosen, Simon, et al.
Published: (2026)
What's under the hood: Investigating Automatic Metrics on Meeting Summarization
by: Kirstein, Frederic, et al.
Published: (2024)
by: Kirstein, Frederic, et al.
Published: (2024)
Toward Equitable Recovery: A Fairness-Aware AI Framework for Prioritizing Post-Flood Aid in Bangladesh
by: Yesmin, Farjana, et al.
Published: (2025)
by: Yesmin, Farjana, et al.
Published: (2025)
Language-Grounded Multi-Agent Planning for Personalized and Fair Participatory Urban Sensing
by: Guo, Xusen, et al.
Published: (2026)
by: Guo, Xusen, et al.
Published: (2026)
Matrix Editing Meets Fair Clustering: Parameterized Algorithms and Complexity
by: Ganian, Robert, et al.
Published: (2025)
by: Ganian, Robert, et al.
Published: (2025)
Human-in-the-loop Fairness: Integrating Stakeholder Feedback to Incorporate Fairness Perspectives in Responsible AI
by: Taka, Evdoxia, et al.
Published: (2023)
by: Taka, Evdoxia, et al.
Published: (2023)
Contextual Moral Value Alignment Through Context-Based Aggregation
by: Dognin, Pierre, et al.
Published: (2024)
by: Dognin, Pierre, et al.
Published: (2024)
Position: Towards Bidirectional Human-AI Alignment
by: Shen, Hua, et al.
Published: (2024)
by: Shen, Hua, et al.
Published: (2024)
The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis
by: O'Doherty, Eoin, et al.
Published: (2025)
by: O'Doherty, Eoin, et al.
Published: (2025)
Similar Items
-
Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment
by: Huang, Allison, et al.
Published: (2024) -
Kantian-Utilitarian XAI: Meta-Explained
by: Atf, Zahra, et al.
Published: (2025) -
Combining Theory of Mind and Kindness for Self-Supervised Human-AI Alignment
by: Hewson, Joshua T. S.
Published: (2024) -
Fairness Metric Design Exploration in Multi-Domain Moral Sentiment Classification using Transformer-Based Models
by: Naranbat, Battemuulen, et al.
Published: (2025) -
Moral Anchor System: A Predictive Framework for AI Value Alignment and Drift Prevention
by: Ravindran, Santhosh Kumar
Published: (2025)