Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Bo, Guo, Jiaxian, Iwasawa, Yusuke, Matsuo, Yutaka |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Suspicion-Agent: Playing Imperfect Information Games with Theory of Mind Aware GPT-4
par: Guo, Jiaxian, et autres
Publié: (2023)
par: Guo, Jiaxian, et autres
Publié: (2023)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
par: Gambardella, Andrew, et autres
Publié: (2024)
par: Gambardella, Andrew, et autres
Publié: (2024)
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
par: Wang, Ru, et autres
Publié: (2025)
par: Wang, Ru, et autres
Publié: (2025)
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
par: Gambardella, Andrew, et autres
Publié: (2025)
par: Gambardella, Andrew, et autres
Publié: (2025)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
par: Cao, Qi, et autres
Publié: (2026)
par: Cao, Qi, et autres
Publié: (2026)
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
par: Minegishi, Gouki, et autres
Publié: (2025)
par: Minegishi, Gouki, et autres
Publié: (2025)
Beyond Induction Heads: In-Context Meta Learning Induces Multi-Phase Circuit Emergence
par: Minegishi, Gouki, et autres
Publié: (2025)
par: Minegishi, Gouki, et autres
Publié: (2025)
Which Programming Language and What Features at Pre-training Stage Affect Downstream Logical Inference Performance?
par: Uchiyama, Fumiya, et autres
Publié: (2024)
par: Uchiyama, Fumiya, et autres
Publié: (2024)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
par: Choi, Younwoo, et autres
Publié: (2025)
par: Choi, Younwoo, et autres
Publié: (2025)
JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Japanese Large Language Models
par: Liu, Junyu, et autres
Publié: (2026)
par: Liu, Junyu, et autres
Publié: (2026)
Beyond In-Distribution Success: Scaling Curves of CoT Granularity for Language Model Generalization
par: Wang, Ru, et autres
Publié: (2025)
par: Wang, Ru, et autres
Publié: (2025)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
par: Li, Huao, et autres
Publié: (2023)
par: Li, Huao, et autres
Publié: (2023)
Automated Refinement of Essay Scoring Rubrics for Language Models via Reflect-and-Revise
par: Harada, Keno, et autres
Publié: (2025)
par: Harada, Keno, et autres
Publié: (2025)
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
par: Gu, Xiaojie, et autres
Publié: (2026)
par: Gu, Xiaojie, et autres
Publié: (2026)
From Chains to Graphs: Self-Structured Reasoning for General-Domain LLMs
par: Chen, Yingjian, et autres
Publié: (2026)
par: Chen, Yingjian, et autres
Publié: (2026)
Exposing Limitations of Language Model Agents in Sequential-Task Compositions on the Web
par: Furuta, Hiroki, et autres
Publié: (2023)
par: Furuta, Hiroki, et autres
Publié: (2023)
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
par: Kim, Hyunwoo, et autres
Publié: (2025)
par: Kim, Hyunwoo, et autres
Publié: (2025)
Theory of Mind in Large Language Models: Assessment and Enhancement
par: Chen, Ruirui, et autres
Publié: (2025)
par: Chen, Ruirui, et autres
Publié: (2025)
Probing the Robustness of Theory of Mind in Large Language Models
par: Nickel, Christian, et autres
Publié: (2024)
par: Nickel, Christian, et autres
Publié: (2024)
Towards Safety Evaluations of Theory of Mind in Large Language Models
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
ToMBench: Benchmarking Theory of Mind in Large Language Models
par: Chen, Zhuang, et autres
Publié: (2024)
par: Chen, Zhuang, et autres
Publié: (2024)
AGENTiGraph: A Multi-Agent Knowledge Graph Framework for Interactive, Domain-Specific LLM Chatbots
par: Zhao, Xinjie, et autres
Publié: (2025)
par: Zhao, Xinjie, et autres
Publié: (2025)
Cleansing the Artificial Mind: A Self-Reflective Detoxification Framework for Large Language Models
par: Zhang, Kaituo, et autres
Publié: (2026)
par: Zhang, Kaituo, et autres
Publié: (2026)
Sensitivity Meets Sparsity: The Impact of Extremely Sparse Parameter Patterns on Theory-of-Mind of Large Language Models
par: Wu, Yuheng, et autres
Publié: (2025)
par: Wu, Yuheng, et autres
Publié: (2025)
Counterfactual Token Generation in Large Language Models
par: Chatzi, Ivi, et autres
Publié: (2024)
par: Chatzi, Ivi, et autres
Publié: (2024)
CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models
par: Li, Mengfan, et autres
Publié: (2026)
par: Li, Mengfan, et autres
Publié: (2026)
CLOMO: Counterfactual Logical Modification with Large Language Models
par: Huang, Yinya, et autres
Publié: (2023)
par: Huang, Yinya, et autres
Publié: (2023)
Language-Informed Synthesis of Rational Agent Models for Grounded Theory-of-Mind Reasoning On-The-Fly
par: Ying, Lance, et autres
Publié: (2025)
par: Ying, Lance, et autres
Publié: (2025)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
par: Nickel, Christian, et autres
Publié: (2026)
par: Nickel, Christian, et autres
Publié: (2026)
Towards A Holistic Landscape of Situated Theory of Mind in Large Language Models
par: Ma, Ziqiao, et autres
Publié: (2023)
par: Ma, Ziqiao, et autres
Publié: (2023)
Topology of Reasoning: Understanding Large Reasoning Models through Reasoning Graph Properties
par: Minegishi, Gouki, et autres
Publié: (2025)
par: Minegishi, Gouki, et autres
Publié: (2025)
Aligning Large Language Models with Counterfactual DPO
par: Butcher, Bradley
Publié: (2024)
par: Butcher, Bradley
Publié: (2024)
Using Counterfactual Tasks to Evaluate the Generality of Analogical Reasoning in Large Language Models
par: Lewis, Martha, et autres
Publié: (2024)
par: Lewis, Martha, et autres
Publié: (2024)
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
par: Nguyen, Hieu Minh "Jord"
Publié: (2025)
par: Nguyen, Hieu Minh "Jord"
Publié: (2025)
Zero, Finite, and Infinite Belief History of Theory of Mind Reasoning in Large Language Models
par: Tang, Weizhi, et autres
Publié: (2024)
par: Tang, Weizhi, et autres
Publié: (2024)
On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons
par: Kojima, Takeshi, et autres
Publié: (2024)
par: Kojima, Takeshi, et autres
Publié: (2024)
UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind
par: Qian, Cheng, et autres
Publié: (2026)
par: Qian, Cheng, et autres
Publié: (2026)
Refactoring Programs Using Large Language Models with Few-Shot Examples
par: Shirafuji, Atsushi, et autres
Publié: (2023)
par: Shirafuji, Atsushi, et autres
Publié: (2023)
The Embodied World Model Based on LLM with Visual Information and Prediction-Oriented Prompts
par: Haijima, Wakana, et autres
Publié: (2024)
par: Haijima, Wakana, et autres
Publié: (2024)
MindForge: Empowering Embodied Agents with Theory of Mind for Lifelong Cultural Learning
par: Lică, Mircea, et autres
Publié: (2024)
par: Lică, Mircea, et autres
Publié: (2024)
Documents similaires
-
Suspicion-Agent: Playing Imperfect Information Games with Theory of Mind Aware GPT-4
par: Guo, Jiaxian, et autres
Publié: (2023) -
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
par: Gambardella, Andrew, et autres
Publié: (2024) -
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
par: Wang, Ru, et autres
Publié: (2025) -
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
par: Gambardella, Andrew, et autres
Publié: (2025) -
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
par: Cao, Qi, et autres
Publié: (2026)