Understanding the Process of Human-AI Value Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | McKinlay, Jack, De Vos, Marina, Hoffmann, Janina A., Theodorou, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regulating AI: Applying insights from behavioural economics and psychology to the application of article 5 of the EU AI Act
by: Zhong, Huixin, et al.
Published: (2023)
by: Zhong, Huixin, et al.
Published: (2023)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
by: Motnikar, Lenart, et al.
Published: (2025)
by: Motnikar, Lenart, et al.
Published: (2025)
Context Matters: Contextual Value-Based Deliberation in Water Consumption Scenarios
by: Oliva-Felipe, Luis, et al.
Published: (2025)
by: Oliva-Felipe, Luis, et al.
Published: (2025)
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
by: Bakirtzis, Georgios, et al.
Published: (2024)
by: Bakirtzis, Georgios, et al.
Published: (2024)
AI and Human Oversight: A Risk-Based Framework for Alignment
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
An Evaluation of Cultural Value Alignment in LLM
by: Sukiennik, Nicholas, et al.
Published: (2025)
by: Sukiennik, Nicholas, et al.
Published: (2025)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
by: Pasandi, Faezeh B., et al.
Published: (2026)
by: Pasandi, Faezeh B., et al.
Published: (2026)
Understanding Human-AI Trust in Education
by: Pitts, Griffin, et al.
Published: (2025)
by: Pitts, Griffin, et al.
Published: (2025)
The AI Alignment Paradox
by: West, Robert, et al.
Published: (2024)
by: West, Robert, et al.
Published: (2024)
How Do Companies Manage the Environmental Sustainability of AI? An Interview Study About Green AI Efforts and Regulations
by: Sampatsing, Ashmita, et al.
Published: (2025)
by: Sampatsing, Ashmita, et al.
Published: (2025)
The Coming Crisis of Multi-Agent Misalignment: AI Alignment Must Be a Dynamic and Social Process
by: Carichon, Florian, et al.
Published: (2025)
by: Carichon, Florian, et al.
Published: (2025)
Dynamic Normativity: Necessary and Sufficient Conditions for Value Alignment
by: Corrêa, Nicholas Kluge
Published: (2024)
by: Corrêa, Nicholas Kluge
Published: (2024)
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI
by: Janowicz, Krzysztof, et al.
Published: (2025)
by: Janowicz, Krzysztof, et al.
Published: (2025)
Rethinking AI Cultural Alignment
by: Bravansky, Michal, et al.
Published: (2025)
by: Bravansky, Michal, et al.
Published: (2025)
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
by: Sornette, Didier, et al.
Published: (2026)
by: Sornette, Didier, et al.
Published: (2026)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
by: Shen, Hua
Published: (2025)
by: Shen, Hua
Published: (2025)
The Ethics of AI Value Chains
by: Attard-Frost, Blair, et al.
Published: (2023)
by: Attard-Frost, Blair, et al.
Published: (2023)
Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment
by: Takahashi, Toru
Published: (2026)
by: Takahashi, Toru
Published: (2026)
Justifications for Democratizing AI Alignment and Their Prospects
by: Steingrüber, André, et al.
Published: (2025)
by: Steingrüber, André, et al.
Published: (2025)
Antisocial Analagous Behavior, Alignment and Human Impact of Google AI Systems: Evaluating through the lens of modified Antisocial Behavior Criteria by Human Interaction, Independent LLM Analysis, and AI Self-Reflection
by: Ogilvie, Alan D.
Published: (2024)
by: Ogilvie, Alan D.
Published: (2024)
MAP: Multi-Human-Value Alignment Palette
by: Wang, Xinran, et al.
Published: (2024)
by: Wang, Xinran, et al.
Published: (2024)
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem
by: LaCroix, Travis
Published: (2026)
by: LaCroix, Travis
Published: (2026)
"I Am the One and Only, Your Cyber BFF": Understanding the Impact of GenAI Requires Understanding the Impact of Anthropomorphic AI
by: Cheng, Myra, et al.
Published: (2024)
by: Cheng, Myra, et al.
Published: (2024)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
by: Chen, Benjamin Minhao, et al.
Published: (2026)
by: Chen, Benjamin Minhao, et al.
Published: (2026)
TUX: Measuring Human--AI Tacit Understanding
by: Li, Yueshen, et al.
Published: (2026)
by: Li, Yueshen, et al.
Published: (2026)
Evaluating LLM Behavior in Hiring: Implicit Weights, Fairness Across Groups, and Alignment with Human Preferences
by: Hoffmann, Morgane, et al.
Published: (2026)
by: Hoffmann, Morgane, et al.
Published: (2026)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
by: Barthwal, Ankur, et al.
Published: (2025)
by: Barthwal, Ankur, et al.
Published: (2025)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
by: Alamdari, Parand A., et al.
Published: (2024)
by: Alamdari, Parand A., et al.
Published: (2024)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
by: Caputo, Nicholas A.
Published: (2024)
by: Caputo, Nicholas A.
Published: (2024)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
by: Tennant, Elizaveta, et al.
Published: (2023)
by: Tennant, Elizaveta, et al.
Published: (2023)
Struggle Premium : How Human Effort and Imperfection Drive Perceived Value in the Age of AI
by: Sultana, Nazneen, et al.
Published: (2026)
by: Sultana, Nazneen, et al.
Published: (2026)
Agentic AI and the Cyber Arms Race
by: Oesch, Sean, et al.
Published: (2025)
by: Oesch, Sean, et al.
Published: (2025)
Human-AI Collaborative Inductive Thematic Analysis: AI Guided Analysis and Human Interpretive Authority
by: Nyaaba, Matthew, et al.
Published: (2026)
by: Nyaaba, Matthew, et al.
Published: (2026)
PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization
by: Jiang, Han, et al.
Published: (2025)
by: Jiang, Han, et al.
Published: (2025)
Beyond Single-Sentence Prompts: Upgrading Value Alignment Benchmarks with Dialogues and Stories
by: Zhang, Yazhou, et al.
Published: (2025)
by: Zhang, Yazhou, et al.
Published: (2025)
From Descriptive to Prescriptive: Uncover the Social Value Alignment of LLM-based Agents
by: Qu, Jinxian, et al.
Published: (2026)
by: Qu, Jinxian, et al.
Published: (2026)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
by: Vishwarupe, Varad, et al.
Published: (2026)
by: Vishwarupe, Varad, et al.
Published: (2026)
Alignment, Agency and Autonomy in Frontier AI: A Systems Engineering Perspective
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
Using AI Alignment Theory to understand the potential pitfalls of regulatory frameworks
by: Tlaie, Alejandro
Published: (2024)
by: Tlaie, Alejandro
Published: (2024)
Similar Items
-
Regulating AI: Applying insights from behavioural economics and psychology to the application of article 5 of the EU AI Act
by: Zhong, Huixin, et al.
Published: (2023) -
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
by: Motnikar, Lenart, et al.
Published: (2025) -
Context Matters: Contextual Value-Based Deliberation in Water Consumption Scenarios
by: Oliva-Felipe, Luis, et al.
Published: (2025) -
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
by: Bakirtzis, Georgios, et al.
Published: (2024) -
AI and Human Oversight: A Risk-Based Framework for Alignment
by: Kandikatla, Laxmiraju, et al.
Published: (2025)