Personalization Increases Affective Alignment but Has Role-Dependent Effects on Epistemic Independence in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kelley, Sean W., Riedl, Christoph |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Affective Computing Has Changed: The Foundation Model Disruption
von: Schuller, Björn, et al.
Veröffentlicht: (2024)
von: Schuller, Björn, et al.
Veröffentlicht: (2024)
PAD: Personalized Alignment of LLMs at Decoding-Time
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment
von: Pustejovsky, James, et al.
Veröffentlicht: (2026)
von: Pustejovsky, James, et al.
Veröffentlicht: (2026)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
When LLMs Team Up: The Emergence of Collaborative Affective Computing
von: Lai, Wenna, et al.
Veröffentlicht: (2025)
von: Lai, Wenna, et al.
Veröffentlicht: (2025)
LLMs with Personalities in Multi-issue Negotiation Games
von: Noh, Sean, et al.
Veröffentlicht: (2024)
von: Noh, Sean, et al.
Veröffentlicht: (2024)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
Can LLMs Truly Embody Human Personality? Analyzing AI and Human Behavior Alignment in Dispute Resolution
von: Kwon, Deuksin, et al.
Veröffentlicht: (2026)
von: Kwon, Deuksin, et al.
Veröffentlicht: (2026)
On Prompt Sensitivity of ChatGPT in Affective Computing
von: Amin, Mostafa M., et al.
Veröffentlicht: (2024)
von: Amin, Mostafa M., et al.
Veröffentlicht: (2024)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
von: Clark, Nicholas, et al.
Veröffentlicht: (2025)
von: Clark, Nicholas, et al.
Veröffentlicht: (2025)
Echoes of Automation: The Increasing Use of LLMs in Newsmaking
von: Ansari, Abolfazl, et al.
Veröffentlicht: (2025)
von: Ansari, Abolfazl, et al.
Veröffentlicht: (2025)
Orchestrating LLMs with Different Personalizations
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
Behavior Trees Enable Structured Programming of Language Model Agents
von: Kelley, Richard
Veröffentlicht: (2024)
von: Kelley, Richard
Veröffentlicht: (2024)
Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects
von: Peng, Ji-Lun, et al.
Veröffentlicht: (2026)
von: Peng, Ji-Lun, et al.
Veröffentlicht: (2026)
Persona-Aware Alignment Framework for Personalized Dialogue Generation
von: Li, Guanrong, et al.
Veröffentlicht: (2025)
von: Li, Guanrong, et al.
Veröffentlicht: (2025)
DPRF: A Generalizable Dynamic Persona Refinement Framework for Optimizing Behavior Alignment Between Personalized LLM Role-Playing Agents and Humans
von: Yao, Bingsheng, et al.
Veröffentlicht: (2025)
von: Yao, Bingsheng, et al.
Veröffentlicht: (2025)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
REAL: Response Embedding-based Alignment for LLMs
von: Zhang, Honggen, et al.
Veröffentlicht: (2024)
von: Zhang, Honggen, et al.
Veröffentlicht: (2024)
Fake Alignment: Are LLMs Really Aligned Well?
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
Flames: Benchmarking Value Alignment of LLMs in Chinese
von: Huang, Kexin, et al.
Veröffentlicht: (2023)
von: Huang, Kexin, et al.
Veröffentlicht: (2023)
From Personal to Collective: On the Role of Local and Global Memory in LLM Personalization
von: Wang, Zehong, et al.
Veröffentlicht: (2025)
von: Wang, Zehong, et al.
Veröffentlicht: (2025)
What Has Been Lost with Synthetic Evaluation?
von: Gill, Alexander, et al.
Veröffentlicht: (2025)
von: Gill, Alexander, et al.
Veröffentlicht: (2025)
Multimodality and Attention Increase Alignment in Natural Language Prediction Between Humans and Computational Models
von: Kewenig, Viktor, et al.
Veröffentlicht: (2023)
von: Kewenig, Viktor, et al.
Veröffentlicht: (2023)
Optimizing Alignment with Less: Leveraging Data Augmentation for Personalized Evaluation
von: Seraj, Javad, et al.
Veröffentlicht: (2024)
von: Seraj, Javad, et al.
Veröffentlicht: (2024)
Multi-Personality Generation of LLMs at Decoding-time
von: Chen, Rongxin, et al.
Veröffentlicht: (2025)
von: Chen, Rongxin, et al.
Veröffentlicht: (2025)
Guided Profile Generation Improves Personalization with LLMs
von: Zhang, Jiarui
Veröffentlicht: (2024)
von: Zhang, Jiarui
Veröffentlicht: (2024)
Advancing and Benchmarking Personalized Tool Invocation for LLMs
von: Huang, Xu, et al.
Veröffentlicht: (2025)
von: Huang, Xu, et al.
Veröffentlicht: (2025)
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
von: Sun, Lihao, et al.
Veröffentlicht: (2025)
von: Sun, Lihao, et al.
Veröffentlicht: (2025)
Affective-ROPTester: Capability and Bias Analysis of LLMs in Predicting Retinopathy of Prematurity
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
Poser: Unmasking Alignment Faking LLMs by Manipulating Their Internals
von: Clymer, Joshua, et al.
Veröffentlicht: (2024)
von: Clymer, Joshua, et al.
Veröffentlicht: (2024)
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
von: Chehade, Mohamad, et al.
Veröffentlicht: (2025)
von: Chehade, Mohamad, et al.
Veröffentlicht: (2025)
Unlocking Recursive Thinking of LLMs: Alignment via Refinement
von: Zhang, Haoke, et al.
Veröffentlicht: (2025)
von: Zhang, Haoke, et al.
Veröffentlicht: (2025)
Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges
von: Thakur, Aman Singh, et al.
Veröffentlicht: (2024)
von: Thakur, Aman Singh, et al.
Veröffentlicht: (2024)
Enhancing Reasoning Abilities of Small LLMs with Cognitive Alignment
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment
von: Cai, Hongru, et al.
Veröffentlicht: (2026)
von: Cai, Hongru, et al.
Veröffentlicht: (2026)
TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment
von: Nguyen, Thi-Nhung, et al.
Veröffentlicht: (2026)
von: Nguyen, Thi-Nhung, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Affective Computing Has Changed: The Foundation Model Disruption
von: Schuller, Björn, et al.
Veröffentlicht: (2024) -
PAD: Personalized Alignment of LLMs at Decoding-Time
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024) -
Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment
von: Pustejovsky, James, et al.
Veröffentlicht: (2026) -
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024) -
When LLMs Team Up: The Emergence of Collaborative Affective Computing
von: Lai, Wenna, et al.
Veröffentlicht: (2025)