MIST: Towards Multi-dimensional Implicit BiaS Evaluation of LLMs for Theory of Mind
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Yanlin, Liu, Hao, Liu, Huimin, Wang, Kun, Wei, Yinwei, Hu, Yupeng |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Implicit Bias in LLMs: A Survey
par: Lin, Xinru, et autres
Publié: (2025)
par: Lin, Xinru, et autres
Publié: (2025)
Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States
par: Xiao, Yang, et autres
Publié: (2025)
par: Xiao, Yang, et autres
Publié: (2025)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
par: Borah, Angana, et autres
Publié: (2024)
par: Borah, Angana, et autres
Publié: (2024)
Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
par: Ruan, Minyuan, et autres
Publié: (2026)
par: Ruan, Minyuan, et autres
Publié: (2026)
LLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue
par: Kowalyshyn, Katharine, et autres
Publié: (2025)
par: Kowalyshyn, Katharine, et autres
Publié: (2025)
DIF: A Framework for Benchmarking and Verifying Implicit Bias in LLMs
par: Yin, Lake, et autres
Publié: (2025)
par: Yin, Lake, et autres
Publié: (2025)
Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space
par: Chen, Chao, et autres
Publié: (2025)
par: Chen, Chao, et autres
Publié: (2025)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
par: Arita, Takaya, et autres
Publié: (2025)
par: Arita, Takaya, et autres
Publié: (2025)
Towards Safety Evaluations of Theory of Mind in Large Language Models
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
par: Aoshima, Tatsuhiro, et autres
Publié: (2025)
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
par: Yim, Yauwai, et autres
Publié: (2024)
par: Yim, Yauwai, et autres
Publié: (2024)
Bias Runs Deep: Implicit Reasoning Biases in Persona-Assigned LLMs
par: Gupta, Shashank, et autres
Publié: (2023)
par: Gupta, Shashank, et autres
Publié: (2023)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
par: Buscemi, Alessio, et autres
Publié: (2025)
par: Buscemi, Alessio, et autres
Publié: (2025)
QGEval: Benchmarking Multi-dimensional Evaluation for Question Generation
par: Fu, Weiping, et autres
Publié: (2024)
par: Fu, Weiping, et autres
Publié: (2024)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
par: Long, Do Xuan, et autres
Publié: (2024)
par: Long, Do Xuan, et autres
Publié: (2024)
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
par: Li, Kun, et autres
Publié: (2025)
par: Li, Kun, et autres
Publié: (2025)
Can LLMs Outshine Conventional Recommenders? A Comparative Evaluation
par: Liu, Qijiong, et autres
Publié: (2025)
par: Liu, Qijiong, et autres
Publié: (2025)
Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue
par: Hu, Junan, et autres
Publié: (2026)
par: Hu, Junan, et autres
Publié: (2026)
MIST: Jailbreaking Black-box Large Language Models via Iterative Semantic Tuning
par: Zheng, Muyang, et autres
Publié: (2025)
par: Zheng, Muyang, et autres
Publié: (2025)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
par: Amirizaniani, Maryam, et autres
Publié: (2024)
par: Amirizaniani, Maryam, et autres
Publié: (2024)
States Hidden in Hidden States: LLMs Emerge Discrete State Representations Implicitly
par: Chen, Junhao, et autres
Publié: (2024)
par: Chen, Junhao, et autres
Publié: (2024)
CreDes: Causal Reasoning Enhancement and Dual-End Searching for Solving Long-Range Reasoning Problems using LLMs
par: Wang, Kangsheng, et autres
Publié: (2024)
par: Wang, Kangsheng, et autres
Publié: (2024)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
par: Mirza, Imran, et autres
Publié: (2025)
par: Mirza, Imran, et autres
Publié: (2025)
MindDial: Belief Dynamics Tracking with Theory-of-Mind Modeling for Situated Neural Dialogue Generation
par: Qiu, Shuwen, et autres
Publié: (2023)
par: Qiu, Shuwen, et autres
Publié: (2023)
TactfulToM: Do LLMs Have the Theory of Mind Ability to Understand White Lies?
par: Liu, Yiwei, et autres
Publié: (2025)
par: Liu, Yiwei, et autres
Publié: (2025)
Multi-ToM: Evaluating Multilingual Theory of Mind Capabilities in Large Language Models
par: Sadhu, Jayanta, et autres
Publié: (2024)
par: Sadhu, Jayanta, et autres
Publié: (2024)
Evaluate Bias without Manual Test Sets: A Concept Representation Perspective for LLMs
par: Gao, Lang, et autres
Publié: (2025)
par: Gao, Lang, et autres
Publié: (2025)
Evaluating the Bias in LLMs for Surveying Opinion and Decision Making in Healthcare
par: Khaokaew, Yonchanok, et autres
Publié: (2025)
par: Khaokaew, Yonchanok, et autres
Publié: (2025)
AstroMind: A High-Fidelity Benchmark for Spacecraft Behavior Reasoning Based on Large Language Models
par: Liu, Hao, et autres
Publié: (2026)
par: Liu, Hao, et autres
Publié: (2026)
Towards Transfer Unlearning: Empirical Evidence of Cross-Domain Bias Mitigation
par: Lu, Huimin, et autres
Publié: (2024)
par: Lu, Huimin, et autres
Publié: (2024)
Theory of Mind and Self-Attributions of Mentality are Dissociable in LLMs
par: Kim, Junsol, et autres
Publié: (2026)
par: Kim, Junsol, et autres
Publié: (2026)
UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind
par: Qian, Cheng, et autres
Publié: (2026)
par: Qian, Cheng, et autres
Publié: (2026)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
par: Wen, Yuchen, et autres
Publié: (2024)
par: Wen, Yuchen, et autres
Publié: (2024)
CSSBench: Evaluating the Safety of Lightweight LLMs against Chinese-Specific Adversarial Patterns
par: Zhou, Zhenhong, et autres
Publié: (2026)
par: Zhou, Zhenhong, et autres
Publié: (2026)
Towards Multimodal Sentiment Analysis Debiasing via Bias Purification
par: Yang, Dingkang, et autres
Publié: (2024)
par: Yang, Dingkang, et autres
Publié: (2024)
McBE: A Multi-task Chinese Bias Evaluation Benchmark for Large Language Models
par: Lan, Tian, et autres
Publié: (2025)
par: Lan, Tian, et autres
Publié: (2025)
Promoting Equality in Large Language Models: Identifying and Mitigating the Implicit Bias based on Bayesian Theory
par: Deng, Yongxin, et autres
Publié: (2024)
par: Deng, Yongxin, et autres
Publié: (2024)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
par: Fan, Xianzhe, et autres
Publié: (2025)
par: Fan, Xianzhe, et autres
Publié: (2025)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
par: Fan, Zhiting, et autres
Publié: (2024)
par: Fan, Zhiting, et autres
Publié: (2024)
Reasoning with Graphs: Structuring Implicit Knowledge to Enhance LLMs Reasoning
par: Han, Haoyu, et autres
Publié: (2025)
par: Han, Haoyu, et autres
Publié: (2025)
CoMMET: To What Extent Can LLMs Perform Theory of Mind Tasks?
par: Chen, Ruirui, et autres
Publié: (2026)
par: Chen, Ruirui, et autres
Publié: (2026)
Documents similaires
-
Implicit Bias in LLMs: A Survey
par: Lin, Xinru, et autres
Publié: (2025) -
Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States
par: Xiao, Yang, et autres
Publié: (2025) -
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
par: Borah, Angana, et autres
Publié: (2024) -
Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
par: Ruan, Minyuan, et autres
Publié: (2026) -
LLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue
par: Kowalyshyn, Katharine, et autres
Publié: (2025)