Probing the Robustness of Theory of Mind in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nickel, Christian, Schrewe, Laura, Flek, Lucie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
von: Nickel, Christian, et al.
Veröffentlicht: (2026)
von: Nickel, Christian, et al.
Veröffentlicht: (2026)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
von: Alavi, Khashayar, et al.
Veröffentlicht: (2025)
von: Alavi, Khashayar, et al.
Veröffentlicht: (2025)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
von: Welz, Simon, et al.
Veröffentlicht: (2025)
von: Welz, Simon, et al.
Veröffentlicht: (2025)
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
von: Fatimah, Shiza, et al.
Veröffentlicht: (2026)
von: Fatimah, Shiza, et al.
Veröffentlicht: (2026)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
Theory of Mind in Large Language Models: Assessment and Enhancement
von: Chen, Ruirui, et al.
Veröffentlicht: (2025)
von: Chen, Ruirui, et al.
Veröffentlicht: (2025)
Pitfalls of Conversational LLMs on News Debiasing
von: Schlicht, Ipek Baris, et al.
Veröffentlicht: (2024)
von: Schlicht, Ipek Baris, et al.
Veröffentlicht: (2024)
Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
von: Rawat, Shivam, et al.
Veröffentlicht: (2026)
von: Rawat, Shivam, et al.
Veröffentlicht: (2026)
ToMBench: Benchmarking Theory of Mind in Large Language Models
von: Chen, Zhuang, et al.
Veröffentlicht: (2024)
von: Chen, Zhuang, et al.
Veröffentlicht: (2024)
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2025)
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2025)
Towards Safety Evaluations of Theory of Mind in Large Language Models
von: Aoshima, Tatsuhiro, et al.
Veröffentlicht: (2025)
von: Aoshima, Tatsuhiro, et al.
Veröffentlicht: (2025)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
von: Li, Huao, et al.
Veröffentlicht: (2023)
von: Li, Huao, et al.
Veröffentlicht: (2023)
Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
von: Yang, Bo, et al.
Veröffentlicht: (2025)
von: Yang, Bo, et al.
Veröffentlicht: (2025)
Towards A Holistic Landscape of Situated Theory of Mind in Large Language Models
von: Ma, Ziqiao, et al.
Veröffentlicht: (2023)
von: Ma, Ziqiao, et al.
Veröffentlicht: (2023)
Zero, Finite, and Infinite Belief History of Theory of Mind Reasoning in Large Language Models
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
von: Nguyen, Hieu Minh "Jord"
Veröffentlicht: (2025)
von: Nguyen, Hieu Minh "Jord"
Veröffentlicht: (2025)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
von: Schlicht, Ipek Baris, et al.
Veröffentlicht: (2025)
von: Schlicht, Ipek Baris, et al.
Veröffentlicht: (2025)
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
von: Xu, Hainiu, et al.
Veröffentlicht: (2024)
von: Xu, Hainiu, et al.
Veröffentlicht: (2024)
ToM-LM: Delegating Theory of Mind Reasoning to External Symbolic Executors in Large Language Models
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models
von: Li, Mengfan, et al.
Veröffentlicht: (2026)
von: Li, Mengfan, et al.
Veröffentlicht: (2026)
Sensitivity Meets Sparsity: The Impact of Extremely Sparse Parameter Patterns on Theory-of-Mind of Large Language Models
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
To Think or Not To Think, That is The Question for Large Reasoning Models in Theory of Mind Tasks
von: Gong, Nanxu, et al.
Veröffentlicht: (2026)
von: Gong, Nanxu, et al.
Veröffentlicht: (2026)
Simulated Annealing Enhances Theory-of-Mind Reasoning in Autoregressive Language Models
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
Let's Put Ourselves in Sally's Shoes: Shoes-of-Others Prefilling Improves Theory of Mind in Large Language Models
von: Shinoda, Kazutoshi, et al.
Veröffentlicht: (2025)
von: Shinoda, Kazutoshi, et al.
Veröffentlicht: (2025)
Decompose-ToM: Enhancing Theory of Mind Reasoning in Large Language Models through Simulation and Task Decomposition
von: Sarangi, Sneheel, et al.
Veröffentlicht: (2025)
von: Sarangi, Sneheel, et al.
Veröffentlicht: (2025)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
von: Javadi, Vahid Sadiri, et al.
Veröffentlicht: (2024)
von: Javadi, Vahid Sadiri, et al.
Veröffentlicht: (2024)
Do Large Language Models Possess a Theory of Mind? A Comparative Evaluation Using the Strange Stories Paradigm
von: Babarczy, Anna, et al.
Veröffentlicht: (2026)
von: Babarczy, Anna, et al.
Veröffentlicht: (2026)
Tucano 2 Cool: Better Open Source LLMs for Portuguese
von: Corrêa, Nicholas Kluge, et al.
Veröffentlicht: (2026)
von: Corrêa, Nicholas Kluge, et al.
Veröffentlicht: (2026)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
von: Wagner, Eitan, et al.
Veröffentlicht: (2024)
Probing Causality Manipulation of Large Language Models
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
Probing Neural Topology of Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind
von: Ying, Lance, et al.
Veröffentlicht: (2024)
von: Ying, Lance, et al.
Veröffentlicht: (2024)
USDC: A Dataset of $\underline{U}$ser $\underline{S}$tance and $\underline{D}$ogmatism in Long $\underline{C}$onversations
von: Marreddy, Mounika, et al.
Veröffentlicht: (2024)
von: Marreddy, Mounika, et al.
Veröffentlicht: (2024)
LiveMind: Low-latency Large Language Models with Simultaneous Inference
von: Chen, Chuangtao, et al.
Veröffentlicht: (2024)
von: Chen, Chuangtao, et al.
Veröffentlicht: (2024)
Robust Knowledge Extraction from Large Language Models using Social Choice Theory
von: Potyka, Nico, et al.
Veröffentlicht: (2023)
von: Potyka, Nico, et al.
Veröffentlicht: (2023)
Do Theory of Mind Benchmarks Need Explicit Human-like Reasoning in Language Models?
von: Lu, Yi-Long, et al.
Veröffentlicht: (2025)
von: Lu, Yi-Long, et al.
Veröffentlicht: (2025)
Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting
von: Getachew, Nathaniel, et al.
Veröffentlicht: (2025)
von: Getachew, Nathaniel, et al.
Veröffentlicht: (2025)
Language-Informed Synthesis of Rational Agent Models for Grounded Theory-of-Mind Reasoning On-The-Fly
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
von: Nickel, Christian, et al.
Veröffentlicht: (2026) -
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
von: Alavi, Khashayar, et al.
Veröffentlicht: (2025) -
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025) -
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
von: Welz, Simon, et al.
Veröffentlicht: (2025) -
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
von: Fatimah, Shiza, et al.
Veröffentlicht: (2026)