Justifications for Democratizing AI Alignment and Their Prospects
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Steingrüber, André, Baum, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
von: Baum, Kevin
Veröffentlicht: (2025)
von: Baum, Kevin
Veröffentlicht: (2025)
Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment
von: Jahn, Felix, et al.
Veröffentlicht: (2026)
von: Jahn, Felix, et al.
Veröffentlicht: (2026)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
von: Motnikar, Lenart, et al.
Veröffentlicht: (2025)
von: Motnikar, Lenart, et al.
Veröffentlicht: (2025)
Bringing AI Participation Down to Scale: A Comment on Open AIs Democratic Inputs to AI Project
von: Moats, David, et al.
Veröffentlicht: (2024)
von: Moats, David, et al.
Veröffentlicht: (2024)
Particip-AI: A Democratic Surveying Framework for Anticipating Future AI Use Cases, Harms and Benefits
von: Mun, Jimin, et al.
Veröffentlicht: (2024)
von: Mun, Jimin, et al.
Veröffentlicht: (2024)
Evaluating AI Evaluation: Perils and Prospects
von: Burden, John
Veröffentlicht: (2024)
von: Burden, John
Veröffentlicht: (2024)
The AI Alignment Paradox
von: West, Robert, et al.
Veröffentlicht: (2024)
von: West, Robert, et al.
Veröffentlicht: (2024)
Artificial Authority: From Machine Minds to Political Alignments. An Experimental Analysis of Democratic and Autocratic Biases in Large-Language Models
von: Ożegalska-Łukasik, Natalia, et al.
Veröffentlicht: (2025)
von: Ożegalska-Łukasik, Natalia, et al.
Veröffentlicht: (2025)
Rethinking AI Cultural Alignment
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
Fair Voting Methods as a Catalyst for Democratic Resilience: A Trilogy on Legitimacy, Impact and AI Safeguarding
von: Pournaras, Evangelos
Veröffentlicht: (2025)
von: Pournaras, Evangelos
Veröffentlicht: (2025)
Exploring AI Writers: Technology, Impact, and Future Prospects
von: Huang, Zhiqian
Veröffentlicht: (2025)
von: Huang, Zhiqian
Veröffentlicht: (2025)
Are There Exceptions to Goodhart's Law? On the Moral Justification of Fairness-Aware Machine Learning
von: Weerts, Hilde, et al.
Veröffentlicht: (2022)
von: Weerts, Hilde, et al.
Veröffentlicht: (2022)
Understanding the Process of Human-AI Value Alignment
von: McKinlay, Jack, et al.
Veröffentlicht: (2025)
von: McKinlay, Jack, et al.
Veröffentlicht: (2025)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
von: Barthwal, Ankur, et al.
Veröffentlicht: (2025)
von: Barthwal, Ankur, et al.
Veröffentlicht: (2025)
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI
von: Janowicz, Krzysztof, et al.
Veröffentlicht: (2025)
von: Janowicz, Krzysztof, et al.
Veröffentlicht: (2025)
AI-Based Reconstruction from Inherited Personal Data: Analysis, Feasibility, and Prospects
von: Zilberman, Mark
Veröffentlicht: (2025)
von: Zilberman, Mark
Veröffentlicht: (2025)
Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance
von: Gabelmann, Julius, et al.
Veröffentlicht: (2026)
von: Gabelmann, Julius, et al.
Veröffentlicht: (2026)
AI and Human Oversight: A Risk-Based Framework for Alignment
von: Kandikatla, Laxmiraju, et al.
Veröffentlicht: (2025)
von: Kandikatla, Laxmiraju, et al.
Veröffentlicht: (2025)
Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents
von: Baum, Kevin, et al.
Veröffentlicht: (2024)
von: Baum, Kevin, et al.
Veröffentlicht: (2024)
Alignment, Agency and Autonomy in Frontier AI: A Systems Engineering Perspective
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Using AI Alignment Theory to understand the potential pitfalls of regulatory frameworks
von: Tlaie, Alejandro
Veröffentlicht: (2024)
von: Tlaie, Alejandro
Veröffentlicht: (2024)
Generative AI Carries Non-Democratic Biases and Stereotypes: Representation of Women, Black Individuals, Age Groups, and People with Disability in AI-Generated Images across Occupations
von: Sadeghiani, Ayoob
Veröffentlicht: (2024)
von: Sadeghiani, Ayoob
Veröffentlicht: (2024)
Wide Reflective Equilibrium in LLM Alignment: Bridging Moral Epistemology and AI Safety
von: Brophy, Matthew
Veröffentlicht: (2025)
von: Brophy, Matthew
Veröffentlicht: (2025)
AI Alignment at Your Discretion
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
von: Pasandi, Faezeh B., et al.
Veröffentlicht: (2026)
von: Pasandi, Faezeh B., et al.
Veröffentlicht: (2026)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
von: Caputo, Nicholas A.
Veröffentlicht: (2024)
von: Caputo, Nicholas A.
Veröffentlicht: (2024)
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
von: Staufer, Leon, et al.
Veröffentlicht: (2026)
von: Staufer, Leon, et al.
Veröffentlicht: (2026)
Characterizing AI Agents for Alignment and Governance
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
The Coming Crisis of Multi-Agent Misalignment: AI Alignment Must Be a Dynamic and Social Process
von: Carichon, Florian, et al.
Veröffentlicht: (2025)
von: Carichon, Florian, et al.
Veröffentlicht: (2025)
An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models
von: Payne, Kenneth
Veröffentlicht: (2025)
von: Payne, Kenneth
Veröffentlicht: (2025)
Measuring AI Diffusion: A Population-Normalized Metric for Tracking Global AI Usage
von: Misra, Amit, et al.
Veröffentlicht: (2025)
von: Misra, Amit, et al.
Veröffentlicht: (2025)
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
von: Sornette, Didier, et al.
Veröffentlicht: (2026)
von: Sornette, Didier, et al.
Veröffentlicht: (2026)
Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
von: Konya, Andrew, et al.
Veröffentlicht: (2024)
The AI Pentad, the CHARME$^{2}$D Model, and an Assessment of Current-State AI Regulation
von: Gao, Di Kevin, et al.
Veröffentlicht: (2025)
von: Gao, Di Kevin, et al.
Veröffentlicht: (2025)
Do AI Companies Make Good on Voluntary Commitments to the White House?
von: Wang, Jennifer, et al.
Veröffentlicht: (2025)
von: Wang, Jennifer, et al.
Veröffentlicht: (2025)
Alignment Debt: The Hidden Work of Making AI Usable
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
von: Oyemike, Cumi, et al.
Veröffentlicht: (2025)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
von: Nghiem, Huy, et al.
Veröffentlicht: (2025)
von: Nghiem, Huy, et al.
Veröffentlicht: (2025)
Democratizing Differential Privacy: A Participatory AI Framework for Public Decision-Making
von: Yang, Wenjun, et al.
Veröffentlicht: (2025)
von: Yang, Wenjun, et al.
Veröffentlicht: (2025)
New Tools are Needed for Tracking Adherence to AI Model Behavioral Use Clauses
von: McDuff, Daniel, et al.
Veröffentlicht: (2025)
von: McDuff, Daniel, et al.
Veröffentlicht: (2025)
Towards Integrated Alignment
von: Reis, Ben Y., et al.
Veröffentlicht: (2025)
von: Reis, Ben Y., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
von: Baum, Kevin
Veröffentlicht: (2025) -
Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment
von: Jahn, Felix, et al.
Veröffentlicht: (2026) -
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
von: Motnikar, Lenart, et al.
Veröffentlicht: (2025) -
Bringing AI Participation Down to Scale: A Comment on Open AIs Democratic Inputs to AI Project
von: Moats, David, et al.
Veröffentlicht: (2024) -
Particip-AI: A Democratic Surveying Framework for Anticipating Future AI Use Cases, Harms and Benefits
von: Mun, Jimin, et al.
Veröffentlicht: (2024)