Effects of Theory of Mind and Prosocial Beliefs on Steering Human-Aligned Behaviors of LLMs in Ultimatum Games
Fuente:
arXiv
Saved in:
| Main Authors: | Yadav, Neemesh, Achananuparp, Palakorn, Jiang, Jing, Lim, Ee-Peng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models
by: Lee, Suhyun, et al.
Published: (2026)
by: Lee, Suhyun, et al.
Published: (2026)
A Multi-Stage Framework with Taxonomy-Guided Reasoning for Occupation Classification Using Large Language Models
by: Achananuparp, Palakorn, et al.
Published: (2025)
by: Achananuparp, Palakorn, et al.
Published: (2025)
Speaker Verification in Agent-Generated Conversations
by: Yang, Yizhe, et al.
Published: (2024)
by: Yang, Yizhe, et al.
Published: (2024)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
by: Arita, Takaya, et al.
Published: (2025)
by: Arita, Takaya, et al.
Published: (2025)
Theory of Mind and Self-Disclosure to CUIs
by: Cox, Samuel Rhys
Published: (2025)
by: Cox, Samuel Rhys
Published: (2025)
Toward Safe and Human-Aligned Game Conversational Recommendation via Multi-Agent Decomposition
by: Hui, Zheng, et al.
Published: (2025)
by: Hui, Zheng, et al.
Published: (2025)
Lexical Indicators of Mind Perception in Human-AI Companionship
by: Banks, Jaime, et al.
Published: (2026)
by: Banks, Jaime, et al.
Published: (2026)
OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
The Dynamics of Delusion: Modeling Bidirectional False Belief Amplification in Human-Chatbot Dialogue
by: Mehta, Ashish, et al.
Published: (2026)
by: Mehta, Ashish, et al.
Published: (2026)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
by: Choi, Alexander S., et al.
Published: (2024)
by: Choi, Alexander S., et al.
Published: (2024)
Exploring the Effects of Chatbot Anthropomorphism and Human Empathy on Human Prosocial Behavior Toward Chatbots
by: Li, Jingshu, et al.
Published: (2025)
by: Li, Jingshu, et al.
Published: (2025)
Evaluating Large Language Models in Theory of Mind Tasks
by: Kosinski, Michal
Published: (2023)
by: Kosinski, Michal
Published: (2023)
Using Contextually Aligned Online Reviews to Measure LLMs' Performance Disparities Across Language Varieties
by: Tang, Zixin, et al.
Published: (2025)
by: Tang, Zixin, et al.
Published: (2025)
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
by: Wu, Tongshuang, et al.
Published: (2023)
by: Wu, Tongshuang, et al.
Published: (2023)
TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews
by: Xu, Huimin, et al.
Published: (2025)
by: Xu, Huimin, et al.
Published: (2025)
Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders
by: Sun, Xin, et al.
Published: (2026)
by: Sun, Xin, et al.
Published: (2026)
Aligning LLMs with Individual Preferences via Interaction
by: Wu, Shujin, et al.
Published: (2024)
by: Wu, Shujin, et al.
Published: (2024)
Simulating Cooperative Prosocial Behavior with Multi-Agent LLMs: Evidence and Mechanisms for AI Agents to Inform Policy Decisions
by: Sreedhar, Karthik, et al.
Published: (2025)
by: Sreedhar, Karthik, et al.
Published: (2025)
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
by: Sarıtaş, Karahan, et al.
Published: (2025)
by: Sarıtaş, Karahan, et al.
Published: (2025)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
by: Huang, Dun-Ming, et al.
Published: (2024)
by: Huang, Dun-Ming, et al.
Published: (2024)
LLM-Augmented Semantic Steering of Text Embedding Projection Spaces
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Prototypical Human-AI Collaboration Behaviors from LLM-Assisted Writing in the Wild
by: Mysore, Sheshera, et al.
Published: (2025)
by: Mysore, Sheshera, et al.
Published: (2025)
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
Direct Advantage Regression: Aligning LLMs with Online AI Reward
by: He, Li, et al.
Published: (2025)
by: He, Li, et al.
Published: (2025)
The Effectiveness of Style Vectors for Steering Large Language Models: A Human Evaluation
by: Diallo, Diaoulé, et al.
Published: (2026)
by: Diallo, Diaoulé, et al.
Published: (2026)
Generating Educational Materials with Different Levels of Readability using LLMs
by: Huang, Chieh-Yang, et al.
Published: (2024)
by: Huang, Chieh-Yang, et al.
Published: (2024)
Aligning Language Models with Demonstrated Feedback
by: Shaikh, Omar, et al.
Published: (2024)
by: Shaikh, Omar, et al.
Published: (2024)
Leveraging Large Language Models for Career Mobility Analysis: A Study of Gender, Race, and Job Change Using U.S. Online Resume Profiles
by: Achananuparp, Palakorn, et al.
Published: (2025)
by: Achananuparp, Palakorn, et al.
Published: (2025)
Performance Gains of LLMs With Humans in a World of LLMs Versus Humans
by: McCullum, Lucas, et al.
Published: (2025)
by: McCullum, Lucas, et al.
Published: (2025)
Game Development as Human-LLM Interaction
by: Hong, Jiale, et al.
Published: (2024)
by: Hong, Jiale, et al.
Published: (2024)
The Benefits of Prosociality towards AI Agents: Examining the Effects of Helping AI Agents on Human Well-Being
by: Zhu, Zicheng, et al.
Published: (2025)
by: Zhu, Zicheng, et al.
Published: (2025)
The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs
by: Zhang, Xing, et al.
Published: (2026)
by: Zhang, Xing, et al.
Published: (2026)
Exploring Personalized Health Support through Data-Driven, Theory-Guided LLMs: A Case Study in Sleep Health
by: Wang, Xingbo, et al.
Published: (2025)
by: Wang, Xingbo, et al.
Published: (2025)
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
by: Liu, Naiming, et al.
Published: (2025)
by: Liu, Naiming, et al.
Published: (2025)
Thinking with Many Minds: Using Large Language Models for Multi-Perspective Problem-Solving
by: Park, Sanghyun, et al.
Published: (2025)
by: Park, Sanghyun, et al.
Published: (2025)
Exploring the Role of Theory of Mind in Human Decision Making: Cognitive, Spatial, and Emotional Influences in the Adversarial Rock-Paper-Scissors Game
by: Nguyen, Thuy Ngoc, et al.
Published: (2025)
by: Nguyen, Thuy Ngoc, et al.
Published: (2025)
Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages
by: Biswas, Shreyan, et al.
Published: (2025)
by: Biswas, Shreyan, et al.
Published: (2025)
Meta-Evaluating Local LLMs: Rethinking Performance Metrics for Serious Games
by: Isaza-Giraldo, Andrés, et al.
Published: (2025)
by: Isaza-Giraldo, Andrés, et al.
Published: (2025)
Reading Users' Minds from What They Say: An Investigation into LLM-based Empathic Mental Inference
by: Zhu, Qihao, et al.
Published: (2024)
by: Zhu, Qihao, et al.
Published: (2024)
What Do You Think I Think? Accounting for Human Beliefs Using Second-Order Theory of Mind
by: Callaghan, Patrick, et al.
Published: (2026)
by: Callaghan, Patrick, et al.
Published: (2026)
Similar Items
-
MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models
by: Lee, Suhyun, et al.
Published: (2026) -
A Multi-Stage Framework with Taxonomy-Guided Reasoning for Occupation Classification Using Large Language Models
by: Achananuparp, Palakorn, et al.
Published: (2025) -
Speaker Verification in Agent-Generated Conversations
by: Yang, Yizhe, et al.
Published: (2024) -
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
by: Arita, Takaya, et al.
Published: (2025) -
Theory of Mind and Self-Disclosure to CUIs
by: Cox, Samuel Rhys
Published: (2025)