Perceptions to Beliefs: Exploring Precursory Inferences for Theory of Mind in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jung, Chani, Kim, Dongkwan, Jin, Jiho, Kim, Jiseon, Seonwoo, Yeon, Choi, Yejin, Oh, Alice, Kim, Hyunwoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Cross-Cultural Differences in English Hate Speech Annotations: From Dataset Construction to Analysis
von: Lee, Nayeon, et al.
Veröffentlicht: (2023)
von: Lee, Nayeon, et al.
Veröffentlicht: (2023)
MUG-Eval: A Proxy Evaluation Framework for Multilingual Generation Capabilities in Any Language
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2025)
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2025)
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
Measuring Interest Group Positions on Legislation: An AI-Driven Analysis of Lobbying Reports
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses
von: Mun, Jimin, et al.
Veröffentlicht: (2026)
von: Mun, Jimin, et al.
Veröffentlicht: (2026)
FINEST: Improving LLM Responses to Sensitive Topics Through Fine-Grained Evaluation
von: Oh, Juhyun, et al.
Veröffentlicht: (2026)
von: Oh, Juhyun, et al.
Veröffentlicht: (2026)
Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language
von: Lee, Seungbeen, et al.
Veröffentlicht: (2025)
von: Lee, Seungbeen, et al.
Veröffentlicht: (2025)
PapersPlease: A Benchmark for Evaluating Motivational Values of Large Language Models Based on ERG Theory
von: Myung, Junho, et al.
Veröffentlicht: (2025)
von: Myung, Junho, et al.
Veröffentlicht: (2025)
LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
Generalizing Weisfeiler-Lehman Kernels to Subgraphs
von: Kim, Dongkwan, et al.
Veröffentlicht: (2024)
von: Kim, Dongkwan, et al.
Veröffentlicht: (2024)
Translating Subgraphs to Nodes Makes Simple GNNs Strong and Efficient for Subgraph Representation Learning
von: Kim, Dongkwan, et al.
Veröffentlicht: (2022)
von: Kim, Dongkwan, et al.
Veröffentlicht: (2022)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
Uncovering Factor Level Preferences to Improve Human-Model Alignment
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
Code-Switching In-Context Learning for Cross-Lingual Transfer of Large Language Models
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
Exploring Fine-Tuning of Large Audio Language Models for Spoken Language Understanding under Limited Speech Data
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
Causal Reasoning in Large Language Models: A Knowledge Graph Approach
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
von: Li, Huihan, et al.
Veröffentlicht: (2024)
von: Li, Huihan, et al.
Veröffentlicht: (2024)
Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
On the Effect of Uncertainty on Layer-wise Inference Dynamics
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
von: Gu, Yuling, et al.
Veröffentlicht: (2024)
von: Gu, Yuling, et al.
Veröffentlicht: (2024)
Metastable Hierarchy in Abstract Low-Temperature Lattice Models
von: Kim, Seonwoo
Veröffentlicht: (2025)
von: Kim, Seonwoo
Veröffentlicht: (2025)
Thanos: Enhancing Conversational Agents with Skill-of-Mind-Infused Large Language Model
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
OWQ: Outlier-Aware Weight Quantization for Efficient Fine-Tuning and Inference of Large Language Models
von: Lee, Changhun, et al.
Veröffentlicht: (2023)
von: Lee, Changhun, et al.
Veröffentlicht: (2023)
Zero, Finite, and Infinite Belief History of Theory of Mind Reasoning in Large Language Models
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
von: Tang, Weizhi, et al.
Veröffentlicht: (2024)
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
von: Acuna, David, et al.
Veröffentlicht: (2025)
von: Acuna, David, et al.
Veröffentlicht: (2025)
Anchoring LLM Gender Bias to Human Baselines: A Cross-Lingual Audit
von: Choi, Jiwoo, et al.
Veröffentlicht: (2026)
von: Choi, Jiwoo, et al.
Veröffentlicht: (2026)
XToM: Exploring the Multilingual Theory of Mind for Large Language Models
von: Chan, Chunkit, et al.
Veröffentlicht: (2025)
von: Chan, Chunkit, et al.
Veröffentlicht: (2025)
Personality Vector: Modulating Personality of Large Language Models by Model Merging
von: Sun, Seungjong, et al.
Veröffentlicht: (2025)
von: Sun, Seungjong, et al.
Veröffentlicht: (2025)
Natural Language Declarative Prompting (NLD-P): A Modular Governance Method for Prompt Design Under Model Drift
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2026)
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2026)
M3-SLU: Evaluating Speaker-Attributed Reasoning in Multimodal Large Language Models
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
Entangled in Representations: Mechanistic Investigation of Cultural Biases in Large Language Models
von: Yu, Haeun, et al.
Veröffentlicht: (2025)
von: Yu, Haeun, et al.
Veröffentlicht: (2025)
R2-KG: General-Purpose Dual-Agent Framework for Reliable Reasoning on Knowledge Graphs
von: Jo, Sumin, et al.
Veröffentlicht: (2025)
von: Jo, Sumin, et al.
Veröffentlicht: (2025)
Latent Preference Modeling for Cross-Session Personalized Tool Calling
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
von: Choi, Junhyuk, et al.
Veröffentlicht: (2026)
von: Choi, Junhyuk, et al.
Veröffentlicht: (2026)
Graph Elicitation for Guiding Multi-Step Reasoning in Large Language Models
von: Park, Jinyoung, et al.
Veröffentlicht: (2023)
von: Park, Jinyoung, et al.
Veröffentlicht: (2023)
Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploring Cross-Cultural Differences in English Hate Speech Annotations: From Dataset Construction to Analysis
von: Lee, Nayeon, et al.
Veröffentlicht: (2023) -
MUG-Eval: A Proxy Evaluation Framework for Multilingual Generation Capabilities in Any Language
von: Song, Seyoung, et al.
Veröffentlicht: (2025) -
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
von: Kim, Hyunwoo, et al.
Veröffentlicht: (2025) -
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023) -
Measuring Interest Group Positions on Legislation: An AI-Driven Analysis of Lobbying Reports
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)