The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Benjamin Minhao, Xie, Xinyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AI-washing: The Asymmetric Effects of Its Two Types on Consumer Moral Judgments
von: Nyilasy, Greg, et al.
Veröffentlicht: (2025)
von: Nyilasy, Greg, et al.
Veröffentlicht: (2025)
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
Human Control Is the Anchor, Not the Answer: Early Divergence of Oversight in Agentic AI Communities
von: Shi, Hanjing, et al.
Veröffentlicht: (2026)
von: Shi, Hanjing, et al.
Veröffentlicht: (2026)
Between Rules and Reality: On the Context Sensitivity of LLM Moral Judgment
von: Sauter, Adrian, et al.
Veröffentlicht: (2026)
von: Sauter, Adrian, et al.
Veröffentlicht: (2026)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
von: Shen, Hua
Veröffentlicht: (2025)
von: Shen, Hua
Veröffentlicht: (2025)
Unilateral Relationship Revision Power in Human-AI Companion Interaction
von: Lange, Benjamin
Veröffentlicht: (2026)
von: Lange, Benjamin
Veröffentlicht: (2026)
Can AI Model the Complexities of Human Moral Decision-Making? A Qualitative Study of Kidney Allocation Decisions
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
von: Keswani, Vijay, et al.
Veröffentlicht: (2025)
Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment
von: Takahashi, Toru
Veröffentlicht: (2026)
von: Takahashi, Toru
Veröffentlicht: (2026)
On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods
von: Boerstler, Kyle, et al.
Veröffentlicht: (2024)
von: Boerstler, Kyle, et al.
Veröffentlicht: (2024)
Human/AI Collective Intelligence for Deliberative Democracy: A Human-Centred Design Approach
von: De Liddo, Anna, et al.
Veröffentlicht: (2026)
von: De Liddo, Anna, et al.
Veröffentlicht: (2026)
Autonomation, Not Automation: Activities and Needs of European Fact-checkers as a Basis for Designing Human-Centered AI Systems
von: Hrckova, Andrea, et al.
Veröffentlicht: (2022)
von: Hrckova, Andrea, et al.
Veröffentlicht: (2022)
Designing Human-AI Collaboration to Support Learning in Counterspeech Writing
von: Ding, Xiaohan, et al.
Veröffentlicht: (2024)
von: Ding, Xiaohan, et al.
Veröffentlicht: (2024)
Agentic AI as Undercover Teammates: Argumentative Knowledge Construction in Hybrid Human-AI Collaborative Learning
von: Yan, Lixiang, et al.
Veröffentlicht: (2025)
von: Yan, Lixiang, et al.
Veröffentlicht: (2025)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
von: Gaube, Susanne, et al.
Veröffentlicht: (2026)
von: Gaube, Susanne, et al.
Veröffentlicht: (2026)
Enhancing Large Language Model-Based Systems for End-to-End Circuit Analysis Problem Solving
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
Beyond Procedural Compliance: Human Oversight as a Dimension of Well-being Efficacy in AI Governance
von: Xie, Yao, et al.
Veröffentlicht: (2025)
von: Xie, Yao, et al.
Veröffentlicht: (2025)
Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment
von: Abe, Yoshia, et al.
Veröffentlicht: (2026)
von: Abe, Yoshia, et al.
Veröffentlicht: (2026)
Responding to Generative AI Technologies with Research-through-Design: The Ryelands AI Lab as an Exploratory Study
von: Benjamin, Jesse Josua, et al.
Veröffentlicht: (2024)
von: Benjamin, Jesse Josua, et al.
Veröffentlicht: (2024)
Enabling Multi-Agent Systems as Learning Designers: Applying Learning Sciences to AI Instructional Design
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
Culturally-Attuned Moral Machines: Implicit Learning of Human Value Systems by AI through Inverse Reinforcement Learning
von: Oliveira, Nigini, et al.
Veröffentlicht: (2023)
von: Oliveira, Nigini, et al.
Veröffentlicht: (2023)
Alignment Debt: The Hidden Work of Making AI Usable
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
Human-Centered Human-AI Collaboration (HCHAC)
von: Gao, Qi, et al.
Veröffentlicht: (2025)
von: Gao, Qi, et al.
Veröffentlicht: (2025)
Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models
von: Singh, Divyanshu Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Divyanshu Kumar, et al.
Veröffentlicht: (2026)
Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis
von: Chen, Benjamin M., et al.
Veröffentlicht: (2026)
von: Chen, Benjamin M., et al.
Veröffentlicht: (2026)
Exploring a Behavioral Model of "Positive Friction" in Human-AI Interaction
von: Chen, Zeya, et al.
Veröffentlicht: (2024)
von: Chen, Zeya, et al.
Veröffentlicht: (2024)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
von: Baum, Kevin
Veröffentlicht: (2025)
von: Baum, Kevin
Veröffentlicht: (2025)
Belief Offloading in Human-AI Interaction
von: Guingrich, Rose E., et al.
Veröffentlicht: (2026)
von: Guingrich, Rose E., et al.
Veröffentlicht: (2026)
Understanding Human-AI Trust in Education
von: Pitts, Griffin, et al.
Veröffentlicht: (2025)
von: Pitts, Griffin, et al.
Veröffentlicht: (2025)
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2024)
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2024)
Towards User-Centred Design of AI-Assisted Decision-Making in Law Enforcement
von: Nowack, Vesna, et al.
Veröffentlicht: (2025)
von: Nowack, Vesna, et al.
Veröffentlicht: (2025)
Lexical Anthropomorphization Influences on Moral Judgments of AI Bad Behavior
von: Banks, Jaime, et al.
Veröffentlicht: (2026)
von: Banks, Jaime, et al.
Veröffentlicht: (2026)
Generative AI User Experience: Developing Human--AI Epistemic Partnership
von: Zhai, Xiaoming
Veröffentlicht: (2026)
von: Zhai, Xiaoming
Veröffentlicht: (2026)
The Manipulation Problem: Conversational AI as a Threat to Epistemic Agency
von: Rosenberg, Louis
Veröffentlicht: (2023)
von: Rosenberg, Louis
Veröffentlicht: (2023)
Human-Centric eXplainable AI in Education
von: Maity, Subhankar, et al.
Veröffentlicht: (2024)
von: Maity, Subhankar, et al.
Veröffentlicht: (2024)
The Missing Knowledge Layer in AI: A Framework for Stable Human-AI Reasoning
von: Rosenbacke, Rikard, et al.
Veröffentlicht: (2026)
von: Rosenbacke, Rikard, et al.
Veröffentlicht: (2026)
Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships
von: De Freitas, Julian, et al.
Veröffentlicht: (2024)
von: De Freitas, Julian, et al.
Veröffentlicht: (2024)
Human-AI Interactions: Cognitive, Behavioral, and Emotional Impacts
von: Riley, Celeste, et al.
Veröffentlicht: (2025)
von: Riley, Celeste, et al.
Veröffentlicht: (2025)
Literary Narrative as Moral Probe : A Cross-System Framework for Evaluating AI Ethical Reasoning and Refusal Behavior
von: Flynn, David C.
Veröffentlicht: (2026)
von: Flynn, David C.
Veröffentlicht: (2026)
The Imbalanced User-AI Relationships as an Ethical Failure of Front-End Design in Healthcare AI
von: Mwadime, Maureen Mghambi
Veröffentlicht: (2026)
von: Mwadime, Maureen Mghambi
Veröffentlicht: (2026)
Ähnliche Einträge
-
AI-washing: The Asymmetric Effects of Its Two Types on Consumer Moral Judgments
von: Nyilasy, Greg, et al.
Veröffentlicht: (2025) -
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
von: Keswani, Vijay, et al.
Veröffentlicht: (2025) -
Human Control Is the Anchor, Not the Answer: Early Divergence of Oversight in Agentic AI Communities
von: Shi, Hanjing, et al.
Veröffentlicht: (2026) -
Between Rules and Reality: On the Context Sensitivity of LLM Moral Judgment
von: Sauter, Adrian, et al.
Veröffentlicht: (2026) -
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
von: Shen, Hua
Veröffentlicht: (2025)