Democratizing Reward Design for Personal and Representative Value-Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Blair, Carter, Larson, Kate, Law, Edith |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reflective Verbal Reward Design for Pluralistic Alignment
by: Blair, Carter, et al.
Published: (2025)
by: Blair, Carter, et al.
Published: (2025)
Unraveling the Dilemma of AI Errors: Exploring the Effectiveness of Human and Machine Explanations for Large Language Models
by: Pafla, Marvin, et al.
Published: (2024)
by: Pafla, Marvin, et al.
Published: (2024)
Value Alignment Tax: Measuring Value Trade-offs in LLM Alignment
by: Chen, Jiajun, et al.
Published: (2026)
by: Chen, Jiajun, et al.
Published: (2026)
Embedding Democratic Values into Social Media AIs via Societal Objective Functions
by: Jia, Chenyan, et al.
Published: (2023)
by: Jia, Chenyan, et al.
Published: (2023)
PreCare: Designing AI Assistants for Advance Care Planning (ACP) to Enhance Personal Value Exploration, Patient Knowledge, and Decisional Confidence
by: Hsu, Yu Lun, et al.
Published: (2025)
by: Hsu, Yu Lun, et al.
Published: (2025)
Towards Democratization of Subspeciality Medical Expertise
by: O'Sullivan, Jack W., et al.
Published: (2024)
by: O'Sullivan, Jack W., et al.
Published: (2024)
Learning Reward and Policy Jointly from Demonstration and Preference Improves Alignment
by: Li, Chenliang, et al.
Published: (2024)
by: Li, Chenliang, et al.
Published: (2024)
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
RoboPlayground: Democratizing Robotic Evaluation through Structured Physical Domains
by: Wang, Yi Ru, et al.
Published: (2026)
by: Wang, Yi Ru, et al.
Published: (2026)
Authors' Values and Attitudes Towards AI-bridged Scalable Personalization of Creative Language Arts
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
The Value-Sensitive Conversational Agent Co-Design Framework
by: Sadek, Malak, et al.
Published: (2023)
by: Sadek, Malak, et al.
Published: (2023)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
Designing AI Personalities: Enhancing Human-Agent Interaction Through Thoughtful Persona Design
by: Zargham, Nima, et al.
Published: (2024)
by: Zargham, Nima, et al.
Published: (2024)
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
by: Shen, Hua, et al.
Published: (2024)
by: Shen, Hua, et al.
Published: (2024)
Evaluating AI Alignment in LLMs: Output Analysis of Value Priorities Across 75 Models with Human Benchmarking
by: Lau, Gabriel Rongyang, et al.
Published: (2025)
by: Lau, Gabriel Rongyang, et al.
Published: (2025)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
by: Kim, Tae Soo, et al.
Published: (2025)
by: Kim, Tae Soo, et al.
Published: (2025)
WEIRD ICWSM: How Western, Educated, Industrialized, Rich, and Democratic is Social Computing Research?
by: Septiandri, Ali Akbar, et al.
Published: (2024)
by: Septiandri, Ali Akbar, et al.
Published: (2024)
ATWL: A Formal Language for Representing, Comparing, and Reusing Visual Analytics Workflows
by: Andrienko, Natalia, et al.
Published: (2026)
by: Andrienko, Natalia, et al.
Published: (2026)
TravelAgent: Generative Agents in the Built Environment
by: Noyman, Ariel, et al.
Published: (2024)
by: Noyman, Ariel, et al.
Published: (2024)
Interactive AI Alignment: Specification, Process, and Evaluation Alignment
by: Terry, Michael, et al.
Published: (2023)
by: Terry, Michael, et al.
Published: (2023)
From Stem to Stern: Contestability Along AI Value Chains
by: Balayn, Agathe, et al.
Published: (2024)
by: Balayn, Agathe, et al.
Published: (2024)
Context-Value-Action Architecture for Value-Driven Large Language Model Agents
by: Zhang, TianZe, et al.
Published: (2026)
by: Zhang, TianZe, et al.
Published: (2026)
PAL: Personal Adaptive Learner
by: Chakraborty, Megha, et al.
Published: (2026)
by: Chakraborty, Megha, et al.
Published: (2026)
Learning to Plan with Personalized Preferences
by: Xu, Manjie, et al.
Published: (2025)
by: Xu, Manjie, et al.
Published: (2025)
Building a "-Sensitive Design" Methodology from Political Philosophies or Ideologies
by: Maocheia-Ricci, Anthony, et al.
Published: (2026)
by: Maocheia-Ricci, Anthony, et al.
Published: (2026)
Alignment has a Fantasia Problem
by: Jo, Nathanael, et al.
Published: (2026)
by: Jo, Nathanael, et al.
Published: (2026)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
by: Chen, Benjamin Minhao, et al.
Published: (2026)
by: Chen, Benjamin Minhao, et al.
Published: (2026)
Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks
by: Rahman, Hasibur, et al.
Published: (2025)
by: Rahman, Hasibur, et al.
Published: (2025)
Heterogeneous Value Alignment Evaluation for Large Language Models
by: Zhang, Zhaowei, et al.
Published: (2023)
by: Zhang, Zhaowei, et al.
Published: (2023)
Unexploited Information Value in Human-AI Collaboration
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
Using Machine Mental Imagery for Representing Common Ground in Situated Dialogue
by: Mohapatra, Biswesh, et al.
Published: (2026)
by: Mohapatra, Biswesh, et al.
Published: (2026)
Customer Service Representative's Perception of the AI Assistant in an Organization's Call Center
by: Qin, Kai, et al.
Published: (2025)
by: Qin, Kai, et al.
Published: (2025)
Analysis, Modeling and Design of Personalized Digital Learning Environment
by: Khanal, Sanjaya, et al.
Published: (2024)
by: Khanal, Sanjaya, et al.
Published: (2024)
Emotion-Agent: Unsupervised Deep Reinforcement Learning with Distribution-Prototype Reward for Continuous Emotional EEG Analysis
by: Zhou, Zhihao, et al.
Published: (2024)
by: Zhou, Zhihao, et al.
Published: (2024)
Envisioning Possibilities and Challenges of AI for Personalized Cancer Care
by: Kong, Elaine, et al.
Published: (2024)
by: Kong, Elaine, et al.
Published: (2024)
The Impact of AI on Educational Assessment: A Framework for Constructive Alignment
by: Stokkink, Patrick
Published: (2025)
by: Stokkink, Patrick
Published: (2025)
Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment
by: Abe, Yoshia, et al.
Published: (2026)
by: Abe, Yoshia, et al.
Published: (2026)
Alignment-Process-Outcome: Rethinking How AIs and Humans Collaborate
by: Li, Haichang, et al.
Published: (2026)
by: Li, Haichang, et al.
Published: (2026)
Beyond Permissions: Investigating Mobile Personalization with Simulated Personas
by: Khalilov, Ibrahim, et al.
Published: (2025)
by: Khalilov, Ibrahim, et al.
Published: (2025)
Rude Humans and Vengeful Robots: Examining Human Perceptions of Robot Retaliatory Intentions in Professional Settings
by: Letheren, Kate, et al.
Published: (2025)
by: Letheren, Kate, et al.
Published: (2025)
Similar Items
-
Reflective Verbal Reward Design for Pluralistic Alignment
by: Blair, Carter, et al.
Published: (2025) -
Unraveling the Dilemma of AI Errors: Exploring the Effectiveness of Human and Machine Explanations for Large Language Models
by: Pafla, Marvin, et al.
Published: (2024) -
Value Alignment Tax: Measuring Value Trade-offs in LLM Alignment
by: Chen, Jiajun, et al.
Published: (2026) -
Embedding Democratic Values into Social Media AIs via Societal Objective Functions
by: Jia, Chenyan, et al.
Published: (2023) -
PreCare: Designing AI Assistants for Advance Care Planning (ACP) to Enhance Personal Value Exploration, Patient Knowledge, and Decisional Confidence
by: Hsu, Yu Lun, et al.
Published: (2025)