SOTOPIA-S4: a user-friendly system for flexible, customizable, and large-scale social simulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Xuhui, Su, Zhe, Feng, Sophie, Zhou, Jiaxu, Huang, Jen-tse, Kao, Hsien-Te, Lynch, Spencer, Volkova, Svitlana, Wu, Tongshuang Sherry, Woolley, Anita, Zhu, Hao, Sap, Maarten |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
by: Cohen, Myke C., et al.
Published: (2025)
by: Cohen, Myke C., et al.
Published: (2025)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
by: Cohen, Myke C., et al.
Published: (2026)
by: Cohen, Myke C., et al.
Published: (2026)
SOTOPIA-TOM: Evaluating Information Management in Multi-Agent Interaction with Theory of Mind
by: YS, Yashwanth, et al.
Published: (2026)
by: YS, Yashwanth, et al.
Published: (2026)
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies
by: Zhou, Jiaxu, et al.
Published: (2025)
by: Zhou, Jiaxu, et al.
Published: (2025)
Exploratory Models of Human-AI Teams: Leveraging Human Digital Twins to Investigate Trust Development
by: Nguyen, Daniel, et al.
Published: (2024)
by: Nguyen, Daniel, et al.
Published: (2024)
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
by: Zhou, Xuhui, et al.
Published: (2023)
by: Zhou, Xuhui, et al.
Published: (2023)
On the Resilience of LLM-Based Multi-Agent Collaboration with Faulty Agents
by: Huang, Jen-tse, et al.
Published: (2024)
by: Huang, Jen-tse, et al.
Published: (2024)
SOTOPIA-$π$: Interactive Learning of Socially Intelligent Language Agents
by: Wang, Ruiyi, et al.
Published: (2024)
by: Wang, Ruiyi, et al.
Published: (2024)
Proactive Defense: Compound AI for Detecting Persuasion Attacks and Measuring Inoculation Effectiveness
by: Volkova, Svitlana, et al.
Published: (2025)
by: Volkova, Svitlana, et al.
Published: (2025)
Is this the real life? Is this just fantasy? The Misleading Success of Simulating Social Interactions With LLMs
by: Zhou, Xuhui, et al.
Published: (2024)
by: Zhou, Xuhui, et al.
Published: (2024)
Towards Safer Online Spaces: Simulating and Assessing Intervention Strategies for Eating Disorder Discussions
by: Penafiel, Louis, et al.
Published: (2024)
by: Penafiel, Louis, et al.
Published: (2024)
Mind the Sim2Real Gap in User Simulation for Agentic Tasks
by: Zhou, Xuhui, et al.
Published: (2026)
by: Zhou, Xuhui, et al.
Published: (2026)
Building Resilient Information Ecosystems: Large LLM-Generated Dataset of Persuasion Attacks
by: Kao, Hsien-Te, et al.
Published: (2025)
by: Kao, Hsien-Te, et al.
Published: (2025)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
1-2-3 Check: Enhancing Contextual Privacy in LLM via Multi-Agent Reasoning
by: Li, Wenkai, et al.
Published: (2025)
by: Li, Wenkai, et al.
Published: (2025)
Rethinking Theory of Mind Benchmarks for LLMs: Towards A User-Centered Perspective
by: Wang, Qiaosi, et al.
Published: (2025)
by: Wang, Qiaosi, et al.
Published: (2025)
GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses
by: Mun, Jimin, et al.
Published: (2026)
by: Mun, Jimin, et al.
Published: (2026)
Social World Models
by: Zhou, Xuhui, et al.
Published: (2025)
by: Zhou, Xuhui, et al.
Published: (2025)
Cross-Disciplinary Knowledge Retrieval and Synthesis: A Compound AI Architecture for Scientific Discovery
by: Volkova, Svitlana, et al.
Published: (2025)
by: Volkova, Svitlana, et al.
Published: (2025)
AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
by: Su, Zhe, et al.
Published: (2024)
by: Su, Zhe, et al.
Published: (2024)
Density-Guided Response Optimization: Community-Grounded Alignment via Implicit Acceptance Signals
by: Gerard, Patrick, et al.
Published: (2026)
by: Gerard, Patrick, et al.
Published: (2026)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
by: Fan, Xianzhe, et al.
Published: (2025)
by: Fan, Xianzhe, et al.
Published: (2025)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
by: Jain, Devansh, et al.
Published: (2024)
by: Jain, Devansh, et al.
Published: (2024)
BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data
by: Li, Wenkai, et al.
Published: (2024)
by: Li, Wenkai, et al.
Published: (2024)
TOM-SWE: User Mental Modeling For Software Engineering Agents
by: Zhou, Xuhui, et al.
Published: (2025)
by: Zhou, Xuhui, et al.
Published: (2025)
Words Like Knives: Backstory-Personalized Modeling and Detection of Violent Communication
by: Shen, Jocelyn, et al.
Published: (2025)
by: Shen, Jocelyn, et al.
Published: (2025)
EEMtoolbox : A user‐friendly R package for flexible ensemble ecosystem modelling
by: Luz Valerie Pascal, et al.
Published: (2025)
by: Luz Valerie Pascal, et al.
Published: (2025)
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
by: Fan, Xianzhe, et al.
Published: (2024)
by: Fan, Xianzhe, et al.
Published: (2024)
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
by: Fan, Xianzhe, et al.
Published: (2024)
by: Fan, Xianzhe, et al.
Published: (2024)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
by: Mireshghallah, Niloofar, et al.
Published: (2023)
by: Mireshghallah, Niloofar, et al.
Published: (2023)
EasyMetagenome: A user-friendly and flexible pipeline for shotgun metagenomic analysis in microbiome research.
by: Bai, Defeng, et al.
Published: (2025)
by: Bai, Defeng, et al.
Published: (2025)
Community-Aligned Behavior Under Uncertainty: Evidence of Epistemic Stance Transfer in LLMs
by: Gerard, Patrick, et al.
Published: (2025)
by: Gerard, Patrick, et al.
Published: (2025)
Training Proactive and Personalized LLM Agents
by: Sun, Weiwei, et al.
Published: (2025)
by: Sun, Weiwei, et al.
Published: (2025)
OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
by: Zhou, Kaitlyn, et al.
Published: (2024)
by: Zhou, Kaitlyn, et al.
Published: (2024)
Chapter 6 The user-friendly Galen
by: Vagelpohl, Uwe
Published: (2019)
by: Vagelpohl, Uwe
Published: (2019)
InterIntent: Investigating Social Intelligence of LLMs via Intention Understanding in an Interactive Game Context
by: Liu, Ziyi, et al.
Published: (2024)
by: Liu, Ziyi, et al.
Published: (2024)
Examining the Effect of Explanations of AI Privacy Redaction in AI-mediated Interactions
by: Kaushik, Roshni, et al.
Published: (2026)
by: Kaushik, Roshni, et al.
Published: (2026)
Science and technology. Thinking with flexible friends
Published: (1994)
Published: (1994)
Visualizing Complex Data Decisions: Design Study for Ethical Factors in AI Clinical Decision Support
by: Surodina, Svitlana, et al.
Published: (2024)
by: Surodina, Svitlana, et al.
Published: (2024)
Similar Items
-
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
by: Cohen, Myke C., et al.
Published: (2025) -
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
by: Cohen, Myke C., et al.
Published: (2026) -
SOTOPIA-TOM: Evaluating Information Management in Multi-Agent Interaction with Theory of Mind
by: YS, Yashwanth, et al.
Published: (2026) -
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies
by: Zhou, Jiaxu, et al.
Published: (2025) -
Exploratory Models of Human-AI Teams: Leveraging Human Digital Twins to Investigate Trust Development
by: Nguyen, Daniel, et al.
Published: (2024)