Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Cohen, Myke C., Su, Zhe, Kao, Hsien-Te, Nguyen, Daniel, Lynch, Spencer, Sap, Maarten, Volkova, Svitlana
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911113065529344
author Cohen, Myke C.
Su, Zhe
Kao, Hsien-Te
Nguyen, Daniel
Lynch, Spencer
Sap, Maarten
Volkova, Svitlana
author_facet Cohen, Myke C.
Su, Zhe
Kao, Hsien-Te
Nguyen, Daniel
Lynch, Spencer
Sap, Maarten
Volkova, Svitlana
contents This paper presents an evaluation framework for agentic AI systems in mission-critical negotiation contexts, addressing the need for AI agents that can adapt to diverse human operators and stakeholders. Using Sotopia as a simulation testbed, we present two experiments that systematically evaluated how personality traits and AI agent characteristics influence LLM-simulated social negotiation outcomes--a capability essential for a variety of applications involving cross-team coordination and civil-military interactions. Experiment 1 employs causal discovery methods to measure how personality traits impact price bargaining negotiations, through which we found that Agreeableness and Extraversion significantly affect believability, goal achievement, and knowledge acquisition outcomes. Sociocognitive lexical measures extracted from team communications detected fine-grained differences in agents' empathic communication, moral foundations, and opinion patterns, providing actionable insights for agentic AI systems that must operate reliably in high-stakes operational scenarios. Experiment 2 evaluates human-AI job negotiations by manipulating both simulated human personality and AI system characteristics, specifically transparency, competence, adaptability, demonstrating how AI agent trustworthiness impact mission effectiveness. These findings establish a repeatable evaluation methodology for experimenting with AI agent reliability across diverse operator personalities and human-agent team dynamics, directly supporting operational requirements for reliable AI systems. Our work advances the evaluation of agentic AI workflows by moving beyond standard performance metrics to incorporate social dynamics essential for mission success in complex operations.
format Preprint
id arxiv_https___arxiv_org_abs_2506_15928
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
Cohen, Myke C.
Su, Zhe
Kao, Hsien-Te
Nguyen, Daniel
Lynch, Spencer
Sap, Maarten
Volkova, Svitlana
Artificial Intelligence
Computation and Language
Human-Computer Interaction
This paper presents an evaluation framework for agentic AI systems in mission-critical negotiation contexts, addressing the need for AI agents that can adapt to diverse human operators and stakeholders. Using Sotopia as a simulation testbed, we present two experiments that systematically evaluated how personality traits and AI agent characteristics influence LLM-simulated social negotiation outcomes--a capability essential for a variety of applications involving cross-team coordination and civil-military interactions. Experiment 1 employs causal discovery methods to measure how personality traits impact price bargaining negotiations, through which we found that Agreeableness and Extraversion significantly affect believability, goal achievement, and knowledge acquisition outcomes. Sociocognitive lexical measures extracted from team communications detected fine-grained differences in agents' empathic communication, moral foundations, and opinion patterns, providing actionable insights for agentic AI systems that must operate reliably in high-stakes operational scenarios. Experiment 2 evaluates human-AI job negotiations by manipulating both simulated human personality and AI system characteristics, specifically transparency, competence, adaptability, demonstrating how AI agent trustworthiness impact mission effectiveness. These findings establish a repeatable evaluation methodology for experimenting with AI agent reliability across diverse operator personalities and human-agent team dynamics, directly supporting operational requirements for reliable AI systems. Our work advances the evaluation of agentic AI workflows by moving beyond standard performance metrics to incorporate social dynamics essential for mission success in complex operations.
title Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
topic Artificial Intelligence
Computation and Language
Human-Computer Interaction
url https://arxiv.org/abs/2506.15928