ToM-SSI: Evaluating Theory of Mind in Situated Social Interactions
Fuente:
arXiv
Saved in:
| Main Authors: | Bortoletto, Matteo, Ruhdorfer, Constantin, Bulling, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Brittle Minds, Fixable Activations: Understanding Belief Representations in Language Models
by: Bortoletto, Matteo, et al.
Published: (2024)
by: Bortoletto, Matteo, et al.
Published: (2024)
Explicit Modelling of Theory of Mind for Belief Prediction in Nonverbal Social Interactions
by: Bortoletto, Matteo, et al.
Published: (2024)
by: Bortoletto, Matteo, et al.
Published: (2024)
Limits of Theory of Mind Modelling in Dialogue-Based Collaborative Plan Acquisition
by: Bortoletto, Matteo, et al.
Published: (2024)
by: Bortoletto, Matteo, et al.
Published: (2024)
The Overcooked Generalisation Challenge: Evaluating Cooperation with Novel Partners in Unknown Environments Using Unsupervised Environment Design
by: Ruhdorfer, Constantin, et al.
Published: (2024)
by: Ruhdorfer, Constantin, et al.
Published: (2024)
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
by: Fan, Xianzhe, et al.
Published: (2025)
by: Fan, Xianzhe, et al.
Published: (2025)
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
DEL-ToM: Inference-Time Scaling for Theory-of-Mind Reasoning via Dynamic Epistemic Logic
by: Wu, Yuheng, et al.
Published: (2025)
by: Wu, Yuheng, et al.
Published: (2025)
ToM-LM: Delegating Theory of Mind Reasoning to External Symbolic Executors in Large Language Models
by: Tang, Weizhi, et al.
Published: (2024)
by: Tang, Weizhi, et al.
Published: (2024)
ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback
by: Bortoletto, Matteo, et al.
Published: (2025)
by: Bortoletto, Matteo, et al.
Published: (2025)
Decompose-ToM: Enhancing Theory of Mind Reasoning in Large Language Models through Simulation and Task Decomposition
by: Sarangi, Sneheel, et al.
Published: (2025)
by: Sarangi, Sneheel, et al.
Published: (2025)
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind
by: Shi, Haojun, et al.
Published: (2024)
by: Shi, Haojun, et al.
Published: (2024)
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
by: Gu, Yuling, et al.
Published: (2024)
by: Gu, Yuling, et al.
Published: (2024)
Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning
by: Ahmed, Nesreen K., et al.
Published: (2026)
by: Ahmed, Nesreen K., et al.
Published: (2026)
Exploring Next Token Prediction in Theory of Mind (ToM) Tasks: Comparative Experiments with GPT-2 and LLaMA-2 AI Models
by: Yadav, Pavan, et al.
Published: (2025)
by: Yadav, Pavan, et al.
Published: (2025)
Towards A Holistic Landscape of Situated Theory of Mind in Large Language Models
by: Ma, Ziqiao, et al.
Published: (2023)
by: Ma, Ziqiao, et al.
Published: (2023)
Rank-O-ToM: Unlocking Emotional Nuance Ranking to Enhance Affective Theory-of-Mind
by: Kim, JiHyun, et al.
Published: (2025)
by: Kim, JiHyun, et al.
Published: (2025)
Multi-ToM: Evaluating Multilingual Theory of Mind Capabilities in Large Language Models
by: Sadhu, Jayanta, et al.
Published: (2024)
by: Sadhu, Jayanta, et al.
Published: (2024)
Ontology-Guided Diffusion for Zero-Shot Visual Sim2Real Transfer
by: Youssef, Mohamed, et al.
Published: (2026)
by: Youssef, Mohamed, et al.
Published: (2026)
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
by: Xu, Hainiu, et al.
Published: (2024)
by: Xu, Hainiu, et al.
Published: (2024)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
by: Wagner, Eitan, et al.
Published: (2024)
by: Wagner, Eitan, et al.
Published: (2024)
Towards Safety Evaluations of Theory of Mind in Large Language Models
by: Aoshima, Tatsuhiro, et al.
Published: (2025)
by: Aoshima, Tatsuhiro, et al.
Published: (2025)
Evaluating Theory of (an uncertain) Mind: Predicting the Uncertain Beliefs of Others in Conversation Forecasting
by: Sicilia, Anthony, et al.
Published: (2024)
by: Sicilia, Anthony, et al.
Published: (2024)
UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind
by: Qian, Cheng, et al.
Published: (2026)
by: Qian, Cheng, et al.
Published: (2026)
Decomposing Theory of Mind: How Emotional Processing Mediates ToM Abilities in LLMs
by: Chulo, Ivan, et al.
Published: (2025)
by: Chulo, Ivan, et al.
Published: (2025)
Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting
by: Getachew, Nathaniel, et al.
Published: (2025)
by: Getachew, Nathaniel, et al.
Published: (2025)
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
by: Nguyen, Hieu Minh "Jord"
Published: (2025)
by: Nguyen, Hieu Minh "Jord"
Published: (2025)
TactfulToM: Do LLMs Have the Theory of Mind Ability to Understand White Lies?
by: Liu, Yiwei, et al.
Published: (2025)
by: Liu, Yiwei, et al.
Published: (2025)
NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
by: Chan, Chunkit, et al.
Published: (2024)
by: Chan, Chunkit, et al.
Published: (2024)
MindForge: Empowering Embodied Agents with Theory of Mind for Lifelong Cultural Learning
by: Lică, Mircea, et al.
Published: (2024)
by: Lică, Mircea, et al.
Published: (2024)
ToMChallenges: A Principle-Guided Dataset and Diverse Evaluation Tasks for Exploring Theory of Mind
by: Ma, Xiaomeng, et al.
Published: (2023)
by: Ma, Xiaomeng, et al.
Published: (2023)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
by: Amirizaniani, Maryam, et al.
Published: (2024)
by: Amirizaniani, Maryam, et al.
Published: (2024)
CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models
by: Li, Mengfan, et al.
Published: (2026)
by: Li, Mengfan, et al.
Published: (2026)
Theory of Mind in Large Language Models: Assessment and Enhancement
by: Chen, Ruirui, et al.
Published: (2025)
by: Chen, Ruirui, et al.
Published: (2025)
Visual Theory of Mind Enables the Invention of Proto-Writing
by: Spiegel, Benjamin A., et al.
Published: (2025)
by: Spiegel, Benjamin A., et al.
Published: (2025)
Theory of Mind and Self-Attributions of Mentality are Dissociable in LLMs
by: Kim, Junsol, et al.
Published: (2026)
by: Kim, Junsol, et al.
Published: (2026)
Probing the Robustness of Theory of Mind in Large Language Models
by: Nickel, Christian, et al.
Published: (2024)
by: Nickel, Christian, et al.
Published: (2024)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
by: Liang, Fangzhou, et al.
Published: (2025)
by: Liang, Fangzhou, et al.
Published: (2025)
Acquiring Grounded Representations of Words with Situated Interactive Instruction
by: Mohan, Shiwali, et al.
Published: (2025)
by: Mohan, Shiwali, et al.
Published: (2025)
Re-evaluating Theory of Mind evaluation in large language models
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Similar Items
-
Brittle Minds, Fixable Activations: Understanding Belief Representations in Language Models
by: Bortoletto, Matteo, et al.
Published: (2024) -
Explicit Modelling of Theory of Mind for Belief Prediction in Nonverbal Social Interactions
by: Bortoletto, Matteo, et al.
Published: (2024) -
Limits of Theory of Mind Modelling in Dialogue-Based Collaborative Plan Acquisition
by: Bortoletto, Matteo, et al.
Published: (2024) -
The Overcooked Generalisation Challenge: Evaluating Cooperation with Novel Partners in Unknown Environments Using Unsupervised Environment Design
by: Ruhdorfer, Constantin, et al.
Published: (2024) -
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork
by: Ruhdorfer, Constantin, et al.
Published: (2025)