Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Baum, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Constitutive vs. Corrective: A Causal Taxonomy of Human Runtime Involvement in AI Systems
von: Baum, Kevin, et al.
Veröffentlicht: (2026)
von: Baum, Kevin, et al.
Veröffentlicht: (2026)
AI Ethics and Governance in Practice: An Introduction
von: Leslie, David, et al.
Veröffentlicht: (2024)
von: Leslie, David, et al.
Veröffentlicht: (2024)
Dubito Ergo Sum: Exploring AI Ethics
von: Dorfler, Viktor, et al.
Veröffentlicht: (2025)
von: Dorfler, Viktor, et al.
Veröffentlicht: (2025)
EthicAlly: a Prototype for AI-Powered Research Ethics Support for the Social Sciences and Humanities
von: Grohmann, Steph
Veröffentlicht: (2025)
von: Grohmann, Steph
Veröffentlicht: (2025)
Aligning Trustworthy AI with Democracy: A Dual Taxonomy of Opportunities and Risks
von: Mentxaka, Oier, et al.
Veröffentlicht: (2025)
von: Mentxaka, Oier, et al.
Veröffentlicht: (2025)
Feeling Machines: Ethics, Culture, and the Rise of Emotional AI
von: Chavan, Vivek, et al.
Veröffentlicht: (2025)
von: Chavan, Vivek, et al.
Veröffentlicht: (2025)
Data Ethics Emergency Drill: A Toolbox for Discussing Responsible AI for Industry Teams
von: Hanschke, Vanessa Aisyahsari, et al.
Veröffentlicht: (2024)
von: Hanschke, Vanessa Aisyahsari, et al.
Veröffentlicht: (2024)
Mind the Gap! Pathways Towards Unifying AI Safety and Ethics Research
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
von: Gaube, Susanne, et al.
Veröffentlicht: (2026)
von: Gaube, Susanne, et al.
Veröffentlicht: (2026)
Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
Alignment Debt: The Hidden Work of Making AI Usable
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
von: Shen, Hua
Veröffentlicht: (2025)
von: Shen, Hua
Veröffentlicht: (2025)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
von: Vaccaro, Michelle, et al.
Veröffentlicht: (2026)
von: Vaccaro, Michelle, et al.
Veröffentlicht: (2026)
Oyster-I: Beyond Refusal -- Constructive Safety Alignment for Responsible Language Models
von: Duan, Ranjie, et al.
Veröffentlicht: (2025)
von: Duan, Ranjie, et al.
Veröffentlicht: (2025)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
von: Chen, Benjamin Minhao, et al.
Veröffentlicht: (2026)
von: Chen, Benjamin Minhao, et al.
Veröffentlicht: (2026)
Designing Beyond Language: Sociotechnical Barriers in AI Health Technologies for Limited English Proficiency
von: Huang, Michelle, et al.
Veröffentlicht: (2025)
von: Huang, Michelle, et al.
Veröffentlicht: (2025)
Beyond Detection: Governing GenAI in Academic Peer Review as a Sociotechnical Challenge
von: Chakravorti, Tatiana, et al.
Veröffentlicht: (2026)
von: Chakravorti, Tatiana, et al.
Veröffentlicht: (2026)
Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models
von: Singh, Divyanshu Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Divyanshu Kumar, et al.
Veröffentlicht: (2026)
Beyond Procedural Compliance: Human Oversight as a Dimension of Well-being Efficacy in AI Governance
von: Xie, Yao, et al.
Veröffentlicht: (2025)
von: Xie, Yao, et al.
Veröffentlicht: (2025)
Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment
von: Takahashi, Toru
Veröffentlicht: (2026)
von: Takahashi, Toru
Veröffentlicht: (2026)
Ethics Practices in AI Development: An Empirical Study Across Roles and Regions
von: Baldwin, Wilder, et al.
Veröffentlicht: (2025)
von: Baldwin, Wilder, et al.
Veröffentlicht: (2025)
Ethics and Technical Aspects of Generative AI Models in Digital Content Creation
von: Karagoz, Atahan
Veröffentlicht: (2024)
von: Karagoz, Atahan
Veröffentlicht: (2024)
The Doctor Will (Still) See You Now: On the Structural Limits of Agentic AI in Healthcare
von: Dias, Gabriela Aránguiz, et al.
Veröffentlicht: (2026)
von: Dias, Gabriela Aránguiz, et al.
Veröffentlicht: (2026)
AI Ethics and Social Norms: Exploring ChatGPT's Capabilities From What to How
von: Veisi, Omid, et al.
Veröffentlicht: (2025)
von: Veisi, Omid, et al.
Veröffentlicht: (2025)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
Beyond Models: A Framework for Contextual and Cultural Intelligence in African AI Deployment
von: Ndlovu, Qness
Veröffentlicht: (2025)
von: Ndlovu, Qness
Veröffentlicht: (2025)
Beyond single-channel agentic benchmarking
von: Radpour, Nelu D.
Veröffentlicht: (2026)
von: Radpour, Nelu D.
Veröffentlicht: (2026)
A Survey of AI Reliance
von: Eckhardt, Sven, et al.
Veröffentlicht: (2024)
von: Eckhardt, Sven, et al.
Veröffentlicht: (2024)
The Missing Knowledge Layer in AI: A Framework for Stable Human-AI Reasoning
von: Rosenbacke, Rikard, et al.
Veröffentlicht: (2026)
von: Rosenbacke, Rikard, et al.
Veröffentlicht: (2026)
Expanding AI Awareness Through Everyday Interactions with AI: A Reflective Journal Study
von: Hingle, Ashish, et al.
Veröffentlicht: (2024)
von: Hingle, Ashish, et al.
Veröffentlicht: (2024)
SafeSpace: An Integrated Web Application for Digital Safety and Emotional Well-being
von: Fatmi, Kayenat, et al.
Veröffentlicht: (2025)
von: Fatmi, Kayenat, et al.
Veröffentlicht: (2025)
Beyond Technocratic XAI: The Who, What & How in Explanation Design
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
DoYouTrustAI: A Tool to Teach Students About AI Misinformation and Prompt Engineering
von: Driscoll, Phillip, et al.
Veröffentlicht: (2025)
von: Driscoll, Phillip, et al.
Veröffentlicht: (2025)
A Rational Analysis of the Effects of Sycophantic AI
von: Batista, Rafael M., et al.
Veröffentlicht: (2026)
von: Batista, Rafael M., et al.
Veröffentlicht: (2026)
The Right to AI
von: Mushkani, Rashid, et al.
Veröffentlicht: (2025)
von: Mushkani, Rashid, et al.
Veröffentlicht: (2025)
AI and Identity
von: Tadimalla, Sri Yash, et al.
Veröffentlicht: (2024)
von: Tadimalla, Sri Yash, et al.
Veröffentlicht: (2024)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
von: Rastogi, Charvi, et al.
Veröffentlicht: (2026)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2026)
Digital Companionship: Overlapping Uses of AI Companions and AI Assistants
von: Manoli, Aikaterina, et al.
Veröffentlicht: (2025)
von: Manoli, Aikaterina, et al.
Veröffentlicht: (2025)
Artificial Intelligence: Beyound Ocularcentrism, the New Age of Humans Beyond the Spectacle
von: Moussaoui, Mustapha El
Veröffentlicht: (2026)
von: Moussaoui, Mustapha El
Veröffentlicht: (2026)
Generative AI Technologies, Techniques & Tensions: A Primer
von: Behrens, John T.
Veröffentlicht: (2026)
von: Behrens, John T.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Constitutive vs. Corrective: A Causal Taxonomy of Human Runtime Involvement in AI Systems
von: Baum, Kevin, et al.
Veröffentlicht: (2026) -
AI Ethics and Governance in Practice: An Introduction
von: Leslie, David, et al.
Veröffentlicht: (2024) -
Dubito Ergo Sum: Exploring AI Ethics
von: Dorfler, Viktor, et al.
Veröffentlicht: (2025) -
EthicAlly: a Prototype for AI-Powered Research Ethics Support for the Social Sciences and Humanities
von: Grohmann, Steph
Veröffentlicht: (2025) -
Aligning Trustworthy AI with Democracy: A Dual Taxonomy of Opportunities and Risks
von: Mentxaka, Oier, et al.
Veröffentlicht: (2025)