Human Values Matter: Investigating How Misalignment Shapes Collective Behaviors in LLM Agent Communities
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Xiangxu, Wang, Jiamin, Zhao, Qinlin, Guo, Hanze, Li, Linzhuo, Yao, Jing, Zhou, Xiao, Yi, Xiaoyuan, Xie, Xing |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
On the Dynamics of Multi-Agent LLM Communities Driven by Value Diversity
par: Huang, Muhua, et autres
Publié: (2025)
par: Huang, Muhua, et autres
Publié: (2025)
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
par: Guo, Hanze, et autres
Publié: (2025)
par: Guo, Hanze, et autres
Publié: (2025)
Knowing Your Uncertainty -- On the application of LLM in social sciences
par: Zhang, Bolun, et autres
Publié: (2025)
par: Zhang, Bolun, et autres
Publié: (2025)
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses
par: Yao, Jing, et autres
Publié: (2024)
par: Yao, Jing, et autres
Publié: (2024)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
par: Lee, Jaehyeok, et autres
Publié: (2026)
par: Lee, Jaehyeok, et autres
Publié: (2026)
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
par: Jiang, Han, et autres
Publié: (2024)
par: Jiang, Han, et autres
Publié: (2024)
MoVa: Towards Generalizable Classification of Human Morals and Values
par: Chen, Ziyu, et autres
Publié: (2025)
par: Chen, Ziyu, et autres
Publié: (2025)
MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?
par: Yong, Xixian, et autres
Publié: (2025)
par: Yong, Xixian, et autres
Publié: (2025)
The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis
par: O'Doherty, Eoin, et autres
Publié: (2025)
par: O'Doherty, Eoin, et autres
Publié: (2025)
An Investigation into Value Misalignment in LLM-Generated Texts for Cultural Heritage
par: Bu, Fan, et autres
Publié: (2025)
par: Bu, Fan, et autres
Publié: (2025)
MoHoBench: Assessing Honesty of Multimodal Large Language Models via Unanswerable Visual Questions
par: Zhu, Yanxu, et autres
Publié: (2025)
par: Zhu, Yanxu, et autres
Publié: (2025)
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights
par: Choi, Sooyung, et autres
Publié: (2025)
par: Choi, Sooyung, et autres
Publié: (2025)
Beyond Human Norms: Unveiling Unique Values of Large Language Models through Interdisciplinary Approaches
par: Biedma, Pablo, et autres
Publié: (2024)
par: Biedma, Pablo, et autres
Publié: (2024)
Tricolore: Multi-Behavior User Profiling for Enhanced Candidate Generation in Recommender Systems
par: Zhou, Xiao, et autres
Publié: (2025)
par: Zhou, Xiao, et autres
Publié: (2025)
Contextualized Privacy Defense for LLM Agents
par: Wen, Yule, et autres
Publié: (2026)
par: Wen, Yule, et autres
Publié: (2026)
The Incomplete Bridge: How AI Research (Mis)Engages with Psychology
par: Jiang, Han, et autres
Publié: (2025)
par: Jiang, Han, et autres
Publié: (2025)
Dynamic Evaluation of Large Language Models by Meta Probing Agents
par: Zhu, Kaijie, et autres
Publié: (2024)
par: Zhu, Kaijie, et autres
Publié: (2024)
Architectural Vulnerability and Reliability Challenges in AI Text Annotation: A Survey-Inspired Framework with Independent Probability Assessment
par: li, Linzhuo
Publié: (2025)
par: li, Linzhuo
Publié: (2025)
Embedding an Ethical Mind: Aligning Text-to-Image Synthesis via Lightweight Value Optimization
par: Wang, Xingqi, et autres
Publié: (2024)
par: Wang, Xingqi, et autres
Publié: (2024)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
par: Wang, Xinyue, et autres
Publié: (2026)
par: Wang, Xinyue, et autres
Publié: (2026)
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
par: Naik, Akshat, et autres
Publié: (2025)
par: Naik, Akshat, et autres
Publié: (2025)
LLM Benchmark-User Need Misalignment for Climate Change
par: Liu, Oucheng, et autres
Publié: (2026)
par: Liu, Oucheng, et autres
Publié: (2026)
PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization
par: Jiang, Han, et autres
Publié: (2025)
par: Jiang, Han, et autres
Publié: (2025)
CAReDiO: Cultural Alignment via Representativeness and Distinctiveness Guided Data Optimization
par: Yao, Jing, et autres
Publié: (2025)
par: Yao, Jing, et autres
Publié: (2025)
AgentReview: Exploring Peer Review Dynamics with LLM Agents
par: Jin, Yiqiao, et autres
Publié: (2024)
par: Jin, Yiqiao, et autres
Publié: (2024)
AdAEM: An Adaptively and Automated Extensible Measurement of LLMs' Value Difference
par: Yao, Jing, et autres
Publié: (2025)
par: Yao, Jing, et autres
Publié: (2025)
Investigating Human Values in Online Communities
par: Borenstein, Nadav, et autres
Publié: (2024)
par: Borenstein, Nadav, et autres
Publié: (2024)
Value Compass Benchmarks: A Platform for Fundamental and Validated Evaluation of LLMs Values
par: Yao, Jing, et autres
Publié: (2025)
par: Yao, Jing, et autres
Publié: (2025)
Conformity Generates Collective Misalignment in AI Agents Societies
par: De Marzo, Giordano, et autres
Publié: (2026)
par: De Marzo, Giordano, et autres
Publié: (2026)
The Behavioral Fabric of LLM-Powered GUI Agents: Human Values and Interaction Outcomes
par: Gebreegziabher, Simret Araya, et autres
Publié: (2026)
par: Gebreegziabher, Simret Araya, et autres
Publié: (2026)
Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Real Online Customer Behavior Data
par: Lu, Yuxuan, et autres
Publié: (2025)
par: Lu, Yuxuan, et autres
Publié: (2025)
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models
par: Wang, Xinpeng, et autres
Publié: (2024)
par: Wang, Xinpeng, et autres
Publié: (2024)
Can Persona-Prompted LLMs Emulate Subgroup Values? An Empirical Analysis of Generalisability and Fairness in Cultural Alignment
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2026)
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2026)
The diverging effect of prestige and experience on the use of artificial intelligence knowledge
par: Nannan Zhao, et autres
Publié: (2026)
par: Nannan Zhao, et autres
Publié: (2026)
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents
par: Zhao, Qinlin, et autres
Publié: (2023)
par: Zhao, Qinlin, et autres
Publié: (2023)
SoREX: Towards Self-Explainable Social Recommendation with Relevant Ego-Path Extraction
par: Guo, Hanze, et autres
Publié: (2025)
par: Guo, Hanze, et autres
Publié: (2025)
Why not Collaborative Filtering in Dual View? Bridging Sparse and Dense Models
par: Guo, Hanze, et autres
Publié: (2026)
par: Guo, Hanze, et autres
Publié: (2026)
Denevil: Towards Deciphering and Navigating the Ethical Values of Large Language Models via Instruction Learning
par: Duan, Shitong, et autres
Publié: (2023)
par: Duan, Shitong, et autres
Publié: (2023)
Preemptive Detection and Correction of Misaligned Actions in LLM Agents
par: Fang, Haishuo, et autres
Publié: (2024)
par: Fang, Haishuo, et autres
Publié: (2024)
PromptBench: A Unified Library for Evaluation of Large Language Models
par: Zhu, Kaijie, et autres
Publié: (2023)
par: Zhu, Kaijie, et autres
Publié: (2023)
Documents similaires
-
On the Dynamics of Multi-Agent LLM Communities Driven by Value Diversity
par: Huang, Muhua, et autres
Publié: (2025) -
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
par: Guo, Hanze, et autres
Publié: (2025) -
Knowing Your Uncertainty -- On the application of LLM in social sciences
par: Zhang, Bolun, et autres
Publié: (2025) -
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses
par: Yao, Jing, et autres
Publié: (2024) -
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
par: Lee, Jaehyeok, et autres
Publié: (2026)