Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
Fuente:
arXiv
Saved in:
| Main Authors: | Ruan, Minyuan, Wang, Ziyue, Liu, Kaiming, Lai, Yunghwei, Li, Peng, Liu, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Thinking with Visual Abstract: Enhancing Multimodal Reasoning via Visual Abstraction
by: Liu, Dairu, et al.
Published: (2025)
by: Liu, Dairu, et al.
Published: (2025)
The Dialogue That Heals: A Comprehensive Evaluation of Doctor Agents' Inquiry Capability
by: Gong, Linlu, et al.
Published: (2025)
by: Gong, Linlu, et al.
Published: (2025)
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty
by: Ren, Jingyi, et al.
Published: (2026)
by: Ren, Jingyi, et al.
Published: (2026)
UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind
by: Qian, Cheng, et al.
Published: (2026)
by: Qian, Cheng, et al.
Published: (2026)
TheraAgent: Self-Improving Therapeutic Agent for Precise and Comprehensive Treatment Planning
by: Li, Junkai, et al.
Published: (2026)
by: Li, Junkai, et al.
Published: (2026)
Doctor-R1: Mastering Clinical Inquiry with Experiential Agentic Reinforcement Learning
by: Lai, Yunghwei, et al.
Published: (2025)
by: Lai, Yunghwei, et al.
Published: (2025)
ToMBench: Benchmarking Theory of Mind in Large Language Models
by: Chen, Zhuang, et al.
Published: (2024)
by: Chen, Zhuang, et al.
Published: (2024)
Beyond One Path: Evaluating and Enhancing Divergent Thinking in Interactive LLM Agents
by: Park, Jihyeong, et al.
Published: (2026)
by: Park, Jihyeong, et al.
Published: (2026)
MUCAR: Benchmarking Multilingual Cross-Modal Ambiguity Resolution for Multimodal Large Language Models
by: Wang, Xiaolong, et al.
Published: (2025)
by: Wang, Xiaolong, et al.
Published: (2025)
Evaluating LLMs' Divergent Thinking Capabilities for Scientific Idea Generation with Minimal Context
by: Ruan, Kai, et al.
Published: (2024)
by: Ruan, Kai, et al.
Published: (2024)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
by: Zou, Henry Peng, et al.
Published: (2026)
by: Zou, Henry Peng, et al.
Published: (2026)
Patient-Zero: Scaling Synthetic Patient Agents to Real-World Distributions without Real Patient Data
by: Lai, Yunghwei, et al.
Published: (2025)
by: Lai, Yunghwei, et al.
Published: (2025)
MIST: Towards Multi-dimensional Implicit BiaS Evaluation of LLMs for Theory of Mind
by: Li, Yanlin, et al.
Published: (2025)
by: Li, Yanlin, et al.
Published: (2025)
Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind
by: Ying, Lance, et al.
Published: (2024)
by: Ying, Lance, et al.
Published: (2024)
Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
DEL-ToM: Inference-Time Scaling for Theory-of-Mind Reasoning via Dynamic Epistemic Logic
by: Wu, Yuheng, et al.
Published: (2025)
by: Wu, Yuheng, et al.
Published: (2025)
Scaling External Knowledge Input Beyond Context Windows of LLMs via Multi-Agent Collaboration
by: Liu, Zijun, et al.
Published: (2025)
by: Liu, Zijun, et al.
Published: (2025)
PersuasiveToM: A Benchmark for Evaluating Machine Theory of Mind in Persuasive Dialogues
by: Yu, Fangxu, et al.
Published: (2025)
by: Yu, Fangxu, et al.
Published: (2025)
Filling the Image Information Gap for VQA: Prompting Large Language Models to Proactively Ask Questions
by: Wang, Ziyue, et al.
Published: (2023)
by: Wang, Ziyue, et al.
Published: (2023)
Beyond the Surface: Enhancing LLM-as-a-Judge Alignment with Human via Internal Representations
by: Lai, Peng, et al.
Published: (2025)
by: Lai, Peng, et al.
Published: (2025)
Doing Things with Words: Rethinking Theory of Mind Simulation in Large Language Models
by: Lombardi, Agnese, et al.
Published: (2025)
by: Lombardi, Agnese, et al.
Published: (2025)
UserRL: Training Interactive User-Centric Agent via Reinforcement Learning
by: Qian, Cheng, et al.
Published: (2025)
by: Qian, Cheng, et al.
Published: (2025)
SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence
by: Bu, Yuyan, et al.
Published: (2026)
by: Bu, Yuyan, et al.
Published: (2026)
ToM-SSI: Evaluating Theory of Mind in Situated Social Interactions
by: Bortoletto, Matteo, et al.
Published: (2025)
by: Bortoletto, Matteo, et al.
Published: (2025)
Enhancing Conversational Agents with Theory of Mind: Aligning Beliefs, Desires, and Intentions for Human-Like Interaction
by: Jafari, Mehdi, et al.
Published: (2025)
by: Jafari, Mehdi, et al.
Published: (2025)
MobileWorld: Benchmarking Autonomous Mobile Agents in Agent-User Interactive and MCP-Augmented Environments
by: Kong, Quyu, et al.
Published: (2025)
by: Kong, Quyu, et al.
Published: (2025)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
by: Li, Huao, et al.
Published: (2023)
by: Li, Huao, et al.
Published: (2023)
MindFlow: Revolutionizing E-commerce Customer Support with Multimodal LLM Agents
by: Gong, Ming, et al.
Published: (2025)
by: Gong, Ming, et al.
Published: (2025)
Beyond Binary Gender: Evaluating Gender-Inclusive Machine Translation with Ambiguous Attitude Words
by: Chen, Yijie, et al.
Published: (2024)
by: Chen, Yijie, et al.
Published: (2024)
Infusing Theory of Mind into Socially Intelligent LLM Agents
by: Hwang, EunJeong, et al.
Published: (2025)
by: Hwang, EunJeong, et al.
Published: (2025)
Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation
by: Choi, Anna Seo Gyeong, et al.
Published: (2026)
by: Choi, Anna Seo Gyeong, et al.
Published: (2026)
EMemBench: Interactive Benchmarking of Episodic Memory for VLM Agents
by: Li, Xinze, et al.
Published: (2026)
by: Li, Xinze, et al.
Published: (2026)
UserBench: An Interactive Gym Environment for User-Centric Agents
by: Qian, Cheng, et al.
Published: (2025)
by: Qian, Cheng, et al.
Published: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
by: Choi, Younwoo, et al.
Published: (2025)
by: Choi, Younwoo, et al.
Published: (2025)
ReviewAgents: Bridging the Gap Between Human and AI-Generated Paper Reviews
by: Gao, Xian, et al.
Published: (2025)
by: Gao, Xian, et al.
Published: (2025)
MindDial: Belief Dynamics Tracking with Theory-of-Mind Modeling for Situated Neural Dialogue Generation
by: Qiu, Shuwen, et al.
Published: (2023)
by: Qiu, Shuwen, et al.
Published: (2023)
RebuttalAgent: Strategic Persuasion in Academic Rebuttal via Theory of Mind
by: He, Zhitao, et al.
Published: (2026)
by: He, Zhitao, et al.
Published: (2026)
Mind the Gap: Linguistic Divergence and Adaptation Strategies in Human-LLM Assistant vs. Human-Human Interactions
by: Zhang, Fulei, et al.
Published: (2025)
by: Zhang, Fulei, et al.
Published: (2025)
Me-Agent: A Personalized Mobile Agent with Two-Level User Habit Learning for Enhanced Interaction
by: Wang, Shuoxin, et al.
Published: (2026)
by: Wang, Shuoxin, et al.
Published: (2026)
MindForge: Empowering Embodied Agents with Theory of Mind for Lifelong Cultural Learning
by: Lică, Mircea, et al.
Published: (2024)
by: Lică, Mircea, et al.
Published: (2024)
Similar Items
-
Thinking with Visual Abstract: Enhancing Multimodal Reasoning via Visual Abstraction
by: Liu, Dairu, et al.
Published: (2025) -
The Dialogue That Heals: A Comprehensive Evaluation of Doctor Agents' Inquiry Capability
by: Gong, Linlu, et al.
Published: (2025) -
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty
by: Ren, Jingyi, et al.
Published: (2026) -
UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind
by: Qian, Cheng, et al.
Published: (2026) -
TheraAgent: Self-Improving Therapeutic Agent for Precise and Comprehensive Treatment Planning
by: Li, Junkai, et al.
Published: (2026)