Position: Theory of Mind Benchmarks are Broken for Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Riemer, Matthew, Ashktorab, Zahra, Bouneffouf, Djallel, Das, Payel, Liu, Miao, Weisz, Justin D., Campbell, Murray |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Scopes of Alignment
par: Varshney, Kush R., et autres
Publié: (2025)
par: Varshney, Kush R., et autres
Publié: (2025)
The Effectiveness of Approximate Regularized Replay for Efficient Supervised Fine-Tuning of Large Language Models
par: Riemer, Matthew, et autres
Publié: (2025)
par: Riemer, Matthew, et autres
Publié: (2025)
The Ultimate Test of Superintelligent AI Agents: Can an AI Balance Care and Control in Asymmetric Relationships?
par: Bouneffouf, Djallel, et autres
Publié: (2025)
par: Bouneffouf, Djallel, et autres
Publié: (2025)
Survey: Multi-Armed Bandits Meet Large Language Models
par: Bouneffouf, Djallel, et autres
Publié: (2025)
par: Bouneffouf, Djallel, et autres
Publié: (2025)
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems
par: Ide, Shun, et autres
Publié: (2024)
par: Ide, Shun, et autres
Publié: (2024)
Evaluating the Prompt Steerability of Large Language Models
par: Miehling, Erik, et autres
Publié: (2024)
par: Miehling, Erik, et autres
Publié: (2024)
Contextual Moral Value Alignment Through Context-Based Aggregation
par: Dognin, Pierre, et autres
Publié: (2024)
par: Dognin, Pierre, et autres
Publié: (2024)
Mitigating Misalignment Contagion by Steering with Implicit Traits
par: Chang, Maria, et autres
Publié: (2026)
par: Chang, Maria, et autres
Publié: (2026)
Agentic AI Needs a Systems Theory
par: Miehling, Erik, et autres
Publié: (2025)
par: Miehling, Erik, et autres
Publié: (2025)
COMPASS: Computational Mapping of Patient-Therapist Alliance Strategies with Language Modeling
par: Lin, Baihan, et autres
Publié: (2024)
par: Lin, Baihan, et autres
Publié: (2024)
ToMBench: Benchmarking Theory of Mind in Large Language Models
par: Chen, Zhuang, et autres
Publié: (2024)
par: Chen, Zhuang, et autres
Publié: (2024)
Combining Domain and Alignment Vectors to Achieve Better Knowledge-Safety Trade-offs in LLMs
par: Thakkar, Megh, et autres
Publié: (2024)
par: Thakkar, Megh, et autres
Publié: (2024)
Conversational Topic Recommendation in Counseling and Psychotherapy with Decision Transformer and Large Language Models
par: Gunal, Aylin, et autres
Publié: (2024)
par: Gunal, Aylin, et autres
Publié: (2024)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
par: Li, Huao, et autres
Publié: (2023)
par: Li, Huao, et autres
Publié: (2023)
From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models
par: Li, Xinyang, et autres
Publié: (2025)
par: Li, Xinyang, et autres
Publié: (2025)
CogToM: A Comprehensive Theory of Mind Benchmark inspired by Human Cognition for Large Language Models
par: Tong, Haibo, et autres
Publié: (2026)
par: Tong, Haibo, et autres
Publié: (2026)
Hypothetical Minds: Scaffolding Theory of Mind for Multi-Agent Tasks with Large Language Models
par: Cross, Logan, et autres
Publié: (2024)
par: Cross, Logan, et autres
Publié: (2024)
Needle in the Haystack for Memory Based Large Language Models
par: Nelson, Elliot, et autres
Publié: (2024)
par: Nelson, Elliot, et autres
Publié: (2024)
Fundamental Safety-Capability Trade-offs in Fine-tuning Large Language Models
par: Chen, Pin-Yu, et autres
Publié: (2025)
par: Chen, Pin-Yu, et autres
Publié: (2025)
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
par: Kim, Hyunwoo, et autres
Publié: (2025)
par: Kim, Hyunwoo, et autres
Publié: (2025)
Theory of Mind in Large Language Models: Assessment and Enhancement
par: Chen, Ruirui, et autres
Publié: (2025)
par: Chen, Ruirui, et autres
Publié: (2025)
Probing the Robustness of Theory of Mind in Large Language Models
par: Nickel, Christian, et autres
Publié: (2024)
par: Nickel, Christian, et autres
Publié: (2024)
SOCK: A Benchmark for Measuring Self-Replication in Large Language Models
par: Chavarria, Justin, et autres
Publié: (2025)
par: Chavarria, Justin, et autres
Publié: (2025)
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
par: Xu, Hainiu, et autres
Publié: (2024)
par: Xu, Hainiu, et autres
Publié: (2024)
EpMAN: Episodic Memory AttentioN for Generalizing to Longer Contexts
par: Chaudhury, Subhajit, et autres
Publié: (2025)
par: Chaudhury, Subhajit, et autres
Publié: (2025)
Mind Scramble: Unveiling Large Language Model Psychology Via Typoglycemia
par: Yu, Miao, et autres
Publié: (2024)
par: Yu, Miao, et autres
Publié: (2024)
Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models
par: Arif, Huzaifa, et autres
Publié: (2025)
par: Arif, Huzaifa, et autres
Publié: (2025)
Towards Safety Evaluations of Theory of Mind in Large Language Models
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
Codenames as a Benchmark for Large Language Models
par: Stephenson, Matthew, et autres
Publié: (2024)
par: Stephenson, Matthew, et autres
Publié: (2024)
Through the Theory of Mind's Eye: Reading Minds with Multimodal Video Large Language Models
par: Chen, Zhawnen, et autres
Publié: (2024)
par: Chen, Zhawnen, et autres
Publié: (2024)
Do Theory of Mind Benchmarks Need Explicit Human-like Reasoning in Language Models?
par: Lu, Yi-Long, et autres
Publié: (2025)
par: Lu, Yi-Long, et autres
Publié: (2025)
Proceedings of 1st Workshop on Advancing Artificial Intelligence through Theory of Mind
par: Abrini, Mouad, et autres
Publié: (2025)
par: Abrini, Mouad, et autres
Publié: (2025)
Dynamic Theory of Mind as a Temporal Memory Problem: Evidence from Large Language Models
par: Nguyen, Thuy Ngoc, et autres
Publié: (2026)
par: Nguyen, Thuy Ngoc, et autres
Publié: (2026)
Consolidation via Policy Information Regularization in Deep RL for Multi-Agent Games
par: Malloy, Tailia, et autres
Publié: (2020)
par: Malloy, Tailia, et autres
Publié: (2020)
Towards Machine Theory of Mind with Large Language Model-Augmented Inverse Planning
par: Gelpí, Rebekah A., et autres
Publié: (2025)
par: Gelpí, Rebekah A., et autres
Publié: (2025)
Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
par: Yang, Bo, et autres
Publié: (2025)
par: Yang, Bo, et autres
Publié: (2025)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
par: Nickel, Christian, et autres
Publié: (2026)
par: Nickel, Christian, et autres
Publié: (2026)
Towards A Holistic Landscape of Situated Theory of Mind in Large Language Models
par: Ma, Ziqiao, et autres
Publié: (2023)
par: Ma, Ziqiao, et autres
Publié: (2023)
SPAN: Benchmarking and Improving Cross-Calendar Temporal Reasoning of Large Language Models
par: Miao, Zhongjian, et autres
Publié: (2025)
par: Miao, Zhongjian, et autres
Publié: (2025)
OR-Bench: An Over-Refusal Benchmark for Large Language Models
par: Cui, Justin, et autres
Publié: (2024)
par: Cui, Justin, et autres
Publié: (2024)
Documents similaires
-
Scopes of Alignment
par: Varshney, Kush R., et autres
Publié: (2025) -
The Effectiveness of Approximate Regularized Replay for Efficient Supervised Fine-Tuning of Large Language Models
par: Riemer, Matthew, et autres
Publié: (2025) -
The Ultimate Test of Superintelligent AI Agents: Can an AI Balance Care and Control in Asymmetric Relationships?
par: Bouneffouf, Djallel, et autres
Publié: (2025) -
Survey: Multi-Armed Bandits Meet Large Language Models
par: Bouneffouf, Djallel, et autres
Publié: (2025) -
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems
par: Ide, Shun, et autres
Publié: (2024)