RLHF May Not Reflect Genuine Preferences
Fuente:
arXiv
Saved in:
| Main Authors: | Ghafouri, Bijean, Choi, Eun Cheol, Dey, Priyanka, Ferrara, Emilio |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lost Before Translation: Social Information Transmission and Survival in AI-AI Communication
by: Ghafouri, Bijean, et al.
Published: (2026)
by: Ghafouri, Bijean, et al.
Published: (2026)
The Variance Paradox: How AI Reduces Diversity but Increases Novelty
by: Ghafouri, Bijean
Published: (2025)
by: Ghafouri, Bijean
Published: (2025)
FACT-GPT: Fact-Checking Augmentation via Claim Matching with LLMs
by: Choi, Eun Cheol, et al.
Published: (2024)
by: Choi, Eun Cheol, et al.
Published: (2024)
What do people want to fact-check?
by: Ghafouri, Bijean, et al.
Published: (2026)
by: Ghafouri, Bijean, et al.
Published: (2026)
Limited Effectiveness of LLM-based Data Augmentation for COVID-19 Misinformation Stance Detection
by: Choi, Eun Cheol, et al.
Published: (2025)
by: Choi, Eun Cheol, et al.
Published: (2025)
Large Language Models for Wearable Sensor-Based Human Activity Recognition, Health Monitoring, and Behavioral Modeling: A Survey of Early Trends, Datasets, and Challenges
by: Ferrara, Emilio
Published: (2024)
by: Ferrara, Emilio
Published: (2024)
The Generative AI Paradox: GenAI and the Erosion of Trust, the Corrosion of Information Verification, and the Demise of Truth
by: Ferrara, Emilio
Published: (2026)
by: Ferrara, Emilio
Published: (2026)
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
Characterizing Human Actions in the Digital Platform by Temporal Context
by: Matsui, Akira, et al.
Published: (2022)
by: Matsui, Akira, et al.
Published: (2022)
GenAI Against Humanity: Nefarious Applications of Generative Artificial Intelligence and Large Language Models
by: Ferrara, Emilio
Published: (2023)
by: Ferrara, Emilio
Published: (2023)
"Can You Play Anything Else?" Understanding Play Style Flexibility in League of Legends
by: Chen, Emily, et al.
Published: (2024)
by: Chen, Emily, et al.
Published: (2024)
ReflectEd: Evaluating Reflection-Driven Learning in an AI-Assisted System
by: Sakib, Md Nazmus, et al.
Published: (2025)
by: Sakib, Md Nazmus, et al.
Published: (2025)
Epistemic Integrity in Large Language Models
by: Ghafouri, Bijean, et al.
Published: (2024)
by: Ghafouri, Bijean, et al.
Published: (2024)
From Daily Song to Daily Self: Supporting Reflective Songwriting of Deaf and Hard-of-Hearing Individuals through Generative Music AI
by: Choi, Youjin, et al.
Published: (2026)
by: Choi, Youjin, et al.
Published: (2026)
From Prompts to Reflection: Designing Reflective Play for GenAI Literacy
by: Ma, Qianou, et al.
Published: (2025)
by: Ma, Qianou, et al.
Published: (2025)
Feed-O-Meter: Investigating AI-Generated Mentee Personas as Interactive Agents for Scaffolding Design Feedback Practice
by: Lim, Hyunseung, et al.
Published: (2025)
by: Lim, Hyunseung, et al.
Published: (2025)
Proxona: Supporting Creators' Sensemaking and Ideation with LLM-Powered Audience Personas
by: Choi, Yoonseo, et al.
Published: (2024)
by: Choi, Yoonseo, et al.
Published: (2024)
Expandora: Broadening Design Exploration with Text-to-Image Model
by: Choi, DaEun, et al.
Published: (2025)
by: Choi, DaEun, et al.
Published: (2025)
Demystifying Tacit Knowledge in Graphic Design: Characteristics, Instances, Approaches, and Guidelines
by: Son, Kihoon, et al.
Published: (2024)
by: Son, Kihoon, et al.
Published: (2024)
Customizing ChatGPT for Second Language Speaking Practice: Genuine Support or Just a Marketing Gimmick?
by: Meng, Fanfei
Published: (2026)
by: Meng, Fanfei
Published: (2026)
IdeaBlocks: Expressing and Reusing Divergent Intents for Graphic Design Exploration using Generative AI
by: Choi, DaEun, et al.
Published: (2025)
by: Choi, DaEun, et al.
Published: (2025)
Adaptive LLM Agents: Toward Personalized Empathetic Care
by: Singh, Priyanka, et al.
Published: (2025)
by: Singh, Priyanka, et al.
Published: (2025)
State-Dependent Refusal and Learned Incapacity in RLHF-Aligned Language Models
by: Lee, TK
Published: (2025)
by: Lee, TK
Published: (2025)
CreativeConnect: Supporting Reference Recombination for Graphic Design Ideation with Generative AI
by: Choi, DaEun, et al.
Published: (2023)
by: Choi, DaEun, et al.
Published: (2023)
GenQuery: Supporting Expressive Visual Search with Generative Models
by: Son, Kihoon, et al.
Published: (2023)
by: Son, Kihoon, et al.
Published: (2023)
GET-Tok: A GenAI-Enriched Multimodal TikTok Dataset Documenting the 2022 Attempted Coup in Peru
by: Pinto, Gabriela, et al.
Published: (2024)
by: Pinto, Gabriela, et al.
Published: (2024)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
by: Lian, Zheng, et al.
Published: (2025)
by: Lian, Zheng, et al.
Published: (2025)
Reflecting Human Values in XAI: Emotional and Reflective Benefits in Creativity Support Tools
by: Cox, Samuel Rhys, et al.
Published: (2025)
by: Cox, Samuel Rhys, et al.
Published: (2025)
Reflections on Traceability for Visualization Research
by: Rogers, Jen, et al.
Published: (2026)
by: Rogers, Jen, et al.
Published: (2026)
Fast-Food Intimacy: How Chinese Women Navigate Soul's AI Boyfriend
by: Lai, Huiqian, et al.
Published: (2026)
by: Lai, Huiqian, et al.
Published: (2026)
Fulfillment of the Work Games: Warehouse Workers' Experiences with Algorithmic Management
by: Cheon, EunJeong, et al.
Published: (2025)
by: Cheon, EunJeong, et al.
Published: (2025)
Understanding User Perception and Intention to Use Smart Homes for Energy Efficiency: A Survey
by: Zharova, Alona, et al.
Published: (2022)
by: Zharova, Alona, et al.
Published: (2022)
Understanding Modality Preferences in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)
by: Tavakoli, Leila, et al.
Published: (2024)
Help Me Reflect: Leveraging Self-Reflection Interface Nudges to Enhance Deliberativeness on Online Deliberation Platforms
by: Yeo, Shun Yi, et al.
Published: (2024)
by: Yeo, Shun Yi, et al.
Published: (2024)
Two Modes of Reflection: How Temporal, Spatial, and Social Distances Affect Reflective Writing in Family Caregiving
by: Norihama, Shunpei, et al.
Published: (2025)
by: Norihama, Shunpei, et al.
Published: (2025)
Reflections on Visualization in Motion for Fitness Trackers
by: Islam, Alaul, et al.
Published: (2024)
by: Islam, Alaul, et al.
Published: (2024)
On Representing Humans' Soft-Ethics Preferences As Dispositions
by: Donati, Donatella, et al.
Published: (2024)
by: Donati, Donatella, et al.
Published: (2024)
Preference-Guided Multi-Objective UI Adaptation
by: Song, Yao, et al.
Published: (2025)
by: Song, Yao, et al.
Published: (2025)
Unveiling Disparities in Web Task Handling Between Human and Web Agent
by: Son, Kihoon, et al.
Published: (2024)
by: Son, Kihoon, et al.
Published: (2024)
Toward Pluralizing Reflection in HCI through Daoism
by: Zhu, Aaron Pengyu, et al.
Published: (2026)
by: Zhu, Aaron Pengyu, et al.
Published: (2026)
Similar Items
-
Lost Before Translation: Social Information Transmission and Survival in AI-AI Communication
by: Ghafouri, Bijean, et al.
Published: (2026) -
The Variance Paradox: How AI Reduces Diversity but Increases Novelty
by: Ghafouri, Bijean
Published: (2025) -
FACT-GPT: Fact-Checking Augmentation via Claim Matching with LLMs
by: Choi, Eun Cheol, et al.
Published: (2024) -
What do people want to fact-check?
by: Ghafouri, Bijean, et al.
Published: (2026) -
Limited Effectiveness of LLM-based Data Augmentation for COVID-19 Misinformation Stance Detection
by: Choi, Eun Cheol, et al.
Published: (2025)