The Value of Disagreement in AI Design, Evaluation, and Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fazelpour, Sina, Fleisher, Will |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disciplining Deliberation: A Sociotechnical Perspective on Machine Learning Trade-offs
von: Fazelpour, Sina
Veröffentlicht: (2024)
von: Fazelpour, Sina
Veröffentlicht: (2024)
Aspirational Affordances of AI
von: Fazelpour, Sina, et al.
Veröffentlicht: (2025)
von: Fazelpour, Sina, et al.
Veröffentlicht: (2025)
Authenticity and exclusion: social media algorithms and the dynamics of belonging in epistemic communities
von: Akpinar, Nil-Jana, et al.
Veröffentlicht: (2024)
von: Akpinar, Nil-Jana, et al.
Veröffentlicht: (2024)
Should you use LLMs to simulate opinions? Quality checks for early-stage deliberation
von: Neumann, Terrence, et al.
Veröffentlicht: (2025)
von: Neumann, Terrence, et al.
Veröffentlicht: (2025)
Ambiguity Collapse by LLMs: A Taxonomy of Epistemic Risks
von: Gur-Arieh, Shira, et al.
Veröffentlicht: (2026)
von: Gur-Arieh, Shira, et al.
Veröffentlicht: (2026)
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation
von: Neumann, Terrence, et al.
Veröffentlicht: (2024)
von: Neumann, Terrence, et al.
Veröffentlicht: (2024)
Take Caution in Using LLMs as Human Surrogates: Scylla Ex Machina
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement
von: Kiet, Huynh Trung, et al.
Veröffentlicht: (2026)
von: Kiet, Huynh Trung, et al.
Veröffentlicht: (2026)
An Evaluation of Cultural Value Alignment in LLM
von: Sukiennik, Nicholas, et al.
Veröffentlicht: (2025)
von: Sukiennik, Nicholas, et al.
Veröffentlicht: (2025)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
von: Motnikar, Lenart, et al.
Veröffentlicht: (2025)
von: Motnikar, Lenart, et al.
Veröffentlicht: (2025)
Understanding the Process of Human-AI Value Alignment
von: McKinlay, Jack, et al.
Veröffentlicht: (2025)
von: McKinlay, Jack, et al.
Veröffentlicht: (2025)
The Values of Value in AI Adoption: Rethinking Efficiency in UX Designers' Workplaces
von: Cha, Inha, et al.
Veröffentlicht: (2026)
von: Cha, Inha, et al.
Veröffentlicht: (2026)
Negotiative Alignment: Embracing Disagreement to Achieve Fairer Outcomes -- Insights from Urban Studies
von: Mushkani, Rashid, et al.
Veröffentlicht: (2025)
von: Mushkani, Rashid, et al.
Veröffentlicht: (2025)
Dimensions of Generative AI Evaluation Design
von: Dow, P. Alex, et al.
Veröffentlicht: (2024)
von: Dow, P. Alex, et al.
Veröffentlicht: (2024)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
The Gray Area: Characterizing Moderator Disagreement on Reddit
von: Alipour, Shayan, et al.
Veröffentlicht: (2026)
von: Alipour, Shayan, et al.
Veröffentlicht: (2026)
EXAGREE: Mitigating Explanation Disagreement with Stakeholder-Aligned Models
von: Li, Sichao, et al.
Veröffentlicht: (2024)
von: Li, Sichao, et al.
Veröffentlicht: (2024)
AI Safety, Alignment, and Ethics (AI SAE)
von: Waldner, Dylan
Veröffentlicht: (2025)
von: Waldner, Dylan
Veröffentlicht: (2025)
Misconceptions, Pragmatism, and Value Tensions: Evaluating Students' Understanding and Perception of Generative AI for Education
von: Johri, Aditya, et al.
Veröffentlicht: (2024)
von: Johri, Aditya, et al.
Veröffentlicht: (2024)
Value Drifts: Tracing Value Alignment During LLM Post-Training
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
LocalValueBench: A Collaboratively Built and Extensible Benchmark for Evaluating Localized Value Alignment and Ethical Safety in Large Language Models
von: Meadows, Gwenyth Isobel, et al.
Veröffentlicht: (2024)
von: Meadows, Gwenyth Isobel, et al.
Veröffentlicht: (2024)
Legal Alignment for Safe and Ethical AI
von: Kolt, Noam, et al.
Veröffentlicht: (2026)
von: Kolt, Noam, et al.
Veröffentlicht: (2026)
AI Alignment vs. AI Ethical Treatment: 10 Challenges
von: Bradley, Adam, et al.
Veröffentlicht: (2025)
von: Bradley, Adam, et al.
Veröffentlicht: (2025)
Crafting Tomorrow's Evaluations: Assessment Design Strategies in the Era of Generative AI
von: Kadel, Rajan, et al.
Veröffentlicht: (2024)
von: Kadel, Rajan, et al.
Veröffentlicht: (2024)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
von: Tennant, Elizaveta, et al.
Veröffentlicht: (2023)
von: Tennant, Elizaveta, et al.
Veröffentlicht: (2023)
The AI Alignment Paradox
von: West, Robert, et al.
Veröffentlicht: (2024)
von: West, Robert, et al.
Veröffentlicht: (2024)
Behavior and Sublinear Algorithm for Opinion Disagreement on Noisy Social Networks
von: Xu, Wanyue, et al.
Veröffentlicht: (2026)
von: Xu, Wanyue, et al.
Veröffentlicht: (2026)
Dynamic Normativity: Necessary and Sufficient Conditions for Value Alignment
von: Corrêa, Nicholas Kluge
Veröffentlicht: (2024)
von: Corrêa, Nicholas Kluge
Veröffentlicht: (2024)
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem
von: LaCroix, Travis
Veröffentlicht: (2026)
von: LaCroix, Travis
Veröffentlicht: (2026)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
von: Chen, Benjamin Minhao, et al.
Veröffentlicht: (2026)
von: Chen, Benjamin Minhao, et al.
Veröffentlicht: (2026)
Seeking Human Security Consensus: A Unified Value Scale for Generative AI Value Safety
von: He, Ying, et al.
Veröffentlicht: (2026)
von: He, Ying, et al.
Veröffentlicht: (2026)
Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks
von: Greco, Candida M., et al.
Veröffentlicht: (2026)
von: Greco, Candida M., et al.
Veröffentlicht: (2026)
The Emotional Alignment Design Policy
von: Schwitzgebel, Eric, et al.
Veröffentlicht: (2025)
von: Schwitzgebel, Eric, et al.
Veröffentlicht: (2025)
Rethinking AI Cultural Alignment
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
Generative AI and Creative Work: Narratives, Values, and Impacts
von: Caramiaux, Baptiste, et al.
Veröffentlicht: (2025)
von: Caramiaux, Baptiste, et al.
Veröffentlicht: (2025)
ACE-Align: Attribute Causal Effect Alignment for Cultural Values under Varying Persona Granularities
von: Luo, Jiatang, et al.
Veröffentlicht: (2026)
von: Luo, Jiatang, et al.
Veröffentlicht: (2026)
The Ethics of AI Value Chains
von: Attard-Frost, Blair, et al.
Veröffentlicht: (2023)
von: Attard-Frost, Blair, et al.
Veröffentlicht: (2023)
Randomness, Not Representation: The Unreliability of Evaluating Cultural Alignment in LLMs
von: Khan, Ariba, et al.
Veröffentlicht: (2025)
von: Khan, Ariba, et al.
Veröffentlicht: (2025)
Predicting Disagreement with Human Raters in LLM-as-a-Judge Difficulty Assessment without Using Generation-Time Probability Signals
von: Ehara, Yo
Veröffentlicht: (2026)
von: Ehara, Yo
Veröffentlicht: (2026)
Ähnliche Einträge
-
Disciplining Deliberation: A Sociotechnical Perspective on Machine Learning Trade-offs
von: Fazelpour, Sina
Veröffentlicht: (2024) -
Aspirational Affordances of AI
von: Fazelpour, Sina, et al.
Veröffentlicht: (2025) -
Authenticity and exclusion: social media algorithms and the dynamics of belonging in epistemic communities
von: Akpinar, Nil-Jana, et al.
Veröffentlicht: (2024) -
Should you use LLMs to simulate opinions? Quality checks for early-stage deliberation
von: Neumann, Terrence, et al.
Veröffentlicht: (2025) -
Ambiguity Collapse by LLMs: A Taxonomy of Epistemic Risks
von: Gur-Arieh, Shira, et al.
Veröffentlicht: (2026)