XToM: Exploring the Multilingual Theory of Mind for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Chan, Chunkit, Yim, Yauwai, Zeng, Hongchuan, Zou, Zhiying, Cheng, Xinyuan, Sun, Zhifan, Deng, Zheye, Chung, Kawai, Ao, Yuzhuo, Fan, Yixiang, Jiayang, Cheng, Nie, Ercong, Wong, Ginny Y., Schmid, Helmut, Schütze, Hinrich, See, Simon, Song, Yangqiu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Persona Knowledge-Aligned Prompt Tuning Method for Online Debate
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
di: Yim, Yauwai, et al.
Pubblicazione: (2024)
di: Yim, Yauwai, et al.
Pubblicazione: (2024)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
Text-Tuple-Table: Towards Information Integration in Text-to-Table Generation via Global Tuple Extraction
di: Deng, Zheye, et al.
Pubblicazione: (2024)
di: Deng, Zheye, et al.
Pubblicazione: (2024)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
di: Nie, Ercong, et al.
Pubblicazione: (2025)
di: Nie, Ercong, et al.
Pubblicazione: (2025)
ActPlan-1K: Benchmarking the Procedural Planning Ability of Visual Language Models in Household Activities
di: Su, Ying, et al.
Pubblicazione: (2024)
di: Su, Ying, et al.
Pubblicazione: (2024)
Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models
di: Nie, Ercong, et al.
Pubblicazione: (2024)
di: Nie, Ercong, et al.
Pubblicazione: (2024)
GNNavi: Navigating the Information Flow in Large Language Models by Graph Neural Network
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
Structuring the Unstructured: A Systematic Review of Text-to-Structure Generation for Agentic AI with a Universal Evaluation Framework
di: Deng, Zheye, et al.
Pubblicazione: (2025)
di: Deng, Zheye, et al.
Pubblicazione: (2025)
DixitWorld: Evaluating Multimodal Abductive Reasoning in Vision-Language Models with Multi-Agent Dixit Gameplay
di: Mo, Yunxiang, et al.
Pubblicazione: (2025)
di: Mo, Yunxiang, et al.
Pubblicazione: (2025)
BMIKE-53: Investigating Cross-Lingual Knowledge Editing with In-Context Learning
di: Nie, Ercong, et al.
Pubblicazione: (2024)
di: Nie, Ercong, et al.
Pubblicazione: (2024)
Large Language Models as Neurolinguistic Subjects: Discrepancy between Performance and Competence
di: He, Linyang, et al.
Pubblicazione: (2024)
di: He, Linyang, et al.
Pubblicazione: (2024)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
EventGround: Narrative Reasoning by Grounding to Eventuality-centric Knowledge Graphs
di: Jiayang, Cheng, et al.
Pubblicazione: (2024)
di: Jiayang, Cheng, et al.
Pubblicazione: (2024)
Constrained Reasoning Chains for Enhancing Theory-of-Mind in Large Language Models
di: Lin, Zizheng, et al.
Pubblicazione: (2024)
di: Lin, Zizheng, et al.
Pubblicazione: (2024)
Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models
di: Wang, Mingyang, et al.
Pubblicazione: (2025)
di: Wang, Mingyang, et al.
Pubblicazione: (2025)
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks
di: Ma, Bolei, et al.
Pubblicazione: (2024)
di: Ma, Bolei, et al.
Pubblicazione: (2024)
ChatGPT Evaluation on Sentence Level Relations: A Focus on Temporal, Causal, and Discourse Relations
di: Chan, Chunkit, et al.
Pubblicazione: (2023)
di: Chan, Chunkit, et al.
Pubblicazione: (2023)
InteGround: On the Evaluation of Verification and Retrieval Planning in Integrative Grounding
di: Jiayang, Cheng, et al.
Pubblicazione: (2025)
di: Jiayang, Cheng, et al.
Pubblicazione: (2025)
Tracing Multilingual Factual Knowledge Acquisition in Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2025)
di: Liu, Yihong, et al.
Pubblicazione: (2025)
LogiDynamics: Unraveling the Dynamics of Inductive, Abductive and Deductive Logical Inferences in LLM Reasoning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
LLM in the Loop: Creating the ParaDeHate Dataset for Hate Speech Detoxification
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2025)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
CLR-Fact: Evaluating the Complex Logical Reasoning Capability of Large Language Models over Factual Knowledge
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
SAD: A Large-Scale Strategic Argumentative Dialogue Dataset
di: Liu, Yongkang, et al.
Pubblicazione: (2026)
di: Liu, Yongkang, et al.
Pubblicazione: (2026)
DiM\textsuperscript{3}: Bridging Multilingual and Multimodal Models via Direction- and Magnitude-Aware Merging
di: Wang, Zijing, et al.
Pubblicazione: (2026)
di: Wang, Zijing, et al.
Pubblicazione: (2026)
XCOMPS: A Multilingual Benchmark of Conceptual Minimal Pairs
di: He, Linyang, et al.
Pubblicazione: (2025)
di: He, Linyang, et al.
Pubblicazione: (2025)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
di: Xhelili, Orgest, et al.
Pubblicazione: (2024)
di: Xhelili, Orgest, et al.
Pubblicazione: (2024)
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models
di: Lin, Peiqin, et al.
Pubblicazione: (2024)
di: Lin, Peiqin, et al.
Pubblicazione: (2024)
Crosslingual On-Policy Self-Distillation for Multilingual Reasoning
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
OFA: A Framework of Initializing Unseen Subword Embeddings for Efficient Large-scale Multilingual Continued Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2023)
di: Liu, Yihong, et al.
Pubblicazione: (2023)
TransMI: A Framework to Create Strong Baselines from Multilingual Pretrained Language Models for Transliterated Data
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
$\mathbb{R}^{2k}$ is Theoretically Large Enough for Embedding-based Top-$k$ Retrieval
di: Wang, Zihao, et al.
Pubblicazione: (2026)
di: Wang, Zihao, et al.
Pubblicazione: (2026)
LangSAMP: Language-Script Aware Multilingual Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
ECon: On the Detection and Resolution of Evidence Conflicts
di: Jiayang, Cheng, et al.
Pubblicazione: (2024)
di: Jiayang, Cheng, et al.
Pubblicazione: (2024)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
di: Gerstner, Sebastian, et al.
Pubblicazione: (2026)
di: Gerstner, Sebastian, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Persona Knowledge-Aligned Prompt Tuning Method for Online Debate
di: Chan, Chunkit, et al.
Pubblicazione: (2024) -
NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
di: Chan, Chunkit, et al.
Pubblicazione: (2024) -
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
di: Yim, Yauwai, et al.
Pubblicazione: (2024) -
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
di: Liang, Fangzhou, et al.
Pubblicazione: (2025) -
Text-Tuple-Table: Towards Information Integration in Text-to-Table Generation via Global Tuple Extraction
di: Deng, Zheye, et al.
Pubblicazione: (2024)