Evaluating Large Language Models in Theory of Mind Tasks
Fuente:
arXiv
Salvato in:
| Autore principale: | Kosinski, Michal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
di: Sarıtaş, Karahan, et al.
Pubblicazione: (2025)
di: Sarıtaş, Karahan, et al.
Pubblicazione: (2025)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
di: Arita, Takaya, et al.
Pubblicazione: (2025)
di: Arita, Takaya, et al.
Pubblicazione: (2025)
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
di: Hirose, Manari, et al.
Pubblicazione: (2025)
di: Hirose, Manari, et al.
Pubblicazione: (2025)
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
di: Ibrahim, Lujain, et al.
Pubblicazione: (2025)
di: Ibrahim, Lujain, et al.
Pubblicazione: (2025)
Open Models, Closed Minds? On Agents Capabilities in Mimicking Human Personalities through Open Large Language Models
di: La Cava, Lucio, et al.
Pubblicazione: (2024)
di: La Cava, Lucio, et al.
Pubblicazione: (2024)
The Moral Machine Experiment on Large Language Models
di: Takemoto, Kazuhiro
Pubblicazione: (2023)
di: Takemoto, Kazuhiro
Pubblicazione: (2023)
The Perils & Promises of Fact-checking with Large Language Models
di: Quelle, Dorian, et al.
Pubblicazione: (2023)
di: Quelle, Dorian, et al.
Pubblicazione: (2023)
DaKultur: Evaluating the Cultural Awareness of Language Models for Danish with Native Speakers
di: Müller-Eberstein, Max, et al.
Pubblicazione: (2025)
di: Müller-Eberstein, Max, et al.
Pubblicazione: (2025)
Evaluating the Application of Large Language Models to Generate Feedback in Programming Education
di: Jacobs, Sven, et al.
Pubblicazione: (2024)
di: Jacobs, Sven, et al.
Pubblicazione: (2024)
ElectionSim: Massive Population Election Simulation Powered by Large Language Model Driven Agents
di: Zhang, Xinnong, et al.
Pubblicazione: (2024)
di: Zhang, Xinnong, et al.
Pubblicazione: (2024)
"Ownership, Not Just Happy Talk": Co-Designing a Participatory Large Language Model for Journalism
di: Tseng, Emily, et al.
Pubblicazione: (2025)
di: Tseng, Emily, et al.
Pubblicazione: (2025)
From Divergence to Consensus: Evaluating the Role of Large Language Models in Facilitating Agreement through Adaptive Strategies
di: Triantafyllopoulos, Loukas, et al.
Pubblicazione: (2025)
di: Triantafyllopoulos, Loukas, et al.
Pubblicazione: (2025)
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
di: Jones, Graham M., et al.
Pubblicazione: (2024)
di: Jones, Graham M., et al.
Pubblicazione: (2024)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
di: Badawi, Abeer, et al.
Pubblicazione: (2025)
The Moral Gap of Large Language Models
di: Skorski, Maciej, et al.
Pubblicazione: (2025)
di: Skorski, Maciej, et al.
Pubblicazione: (2025)
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data
di: Kursuncu, Ugur, et al.
Pubblicazione: (2025)
di: Kursuncu, Ugur, et al.
Pubblicazione: (2025)
Language Models as Critical Thinking Tools: A Case Study of Philosophers
di: Ye, Andre, et al.
Pubblicazione: (2024)
di: Ye, Andre, et al.
Pubblicazione: (2024)
Gender Trouble in Language Models: An Empirical Audit Guided by Gender Performativity Theory
di: Hafner, Franziska Sofia, et al.
Pubblicazione: (2025)
di: Hafner, Franziska Sofia, et al.
Pubblicazione: (2025)
Facial recognition technology and human raters can predict political orientation from images of expressionless faces even when controlling for demographics and self-presentation
di: Kosinski, Michal, et al.
Pubblicazione: (2023)
di: Kosinski, Michal, et al.
Pubblicazione: (2023)
More is More: Addition Bias in Large Language Models
di: Santagata, Luca, et al.
Pubblicazione: (2024)
di: Santagata, Luca, et al.
Pubblicazione: (2024)
Impacts of Anthropomorphizing Large Language Models in Learning Environments
di: Schaaff, Kristina, et al.
Pubblicazione: (2024)
di: Schaaff, Kristina, et al.
Pubblicazione: (2024)
Exploring the Human-LLM Synergy in Advancing Theory-driven Qualitative Analysis
di: Meng, Han, et al.
Pubblicazione: (2024)
di: Meng, Han, et al.
Pubblicazione: (2024)
Large Language Models as Psychological Simulators: A Methodological Guide
di: Lin, Zhicheng
Pubblicazione: (2025)
di: Lin, Zhicheng
Pubblicazione: (2025)
Evidence of conceptual mastery in the application of rules by Large Language Models
di: Nunes, José Luiz, et al.
Pubblicazione: (2025)
di: Nunes, José Luiz, et al.
Pubblicazione: (2025)
On the Reliability of Large Language Models to Misinformed and Demographically-Informed Prompts
di: Aremu, Toluwani, et al.
Pubblicazione: (2024)
di: Aremu, Toluwani, et al.
Pubblicazione: (2024)
Inadequacies of Large Language Model Benchmarks in the Era of Generative Artificial Intelligence
di: McIntosh, Timothy R., et al.
Pubblicazione: (2024)
di: McIntosh, Timothy R., et al.
Pubblicazione: (2024)
Empirical evidence of Large Language Model's influence on human spoken communication
di: Yakura, Hiromu, et al.
Pubblicazione: (2024)
di: Yakura, Hiromu, et al.
Pubblicazione: (2024)
NARRA-Gym for Evaluating Interactive Narrative Agents
di: Huang, Yue, et al.
Pubblicazione: (2026)
di: Huang, Yue, et al.
Pubblicazione: (2026)
What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma
di: Meng, Han, et al.
Pubblicazione: (2025)
di: Meng, Han, et al.
Pubblicazione: (2025)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
Getting in the Door: Streamlining Intake in Civil Legal Services with Large Language Models
di: Steenhuis, Quinten, et al.
Pubblicazione: (2024)
di: Steenhuis, Quinten, et al.
Pubblicazione: (2024)
Thinking with Many Minds: Using Large Language Models for Multi-Perspective Problem-Solving
di: Park, Sanghyun, et al.
Pubblicazione: (2025)
di: Park, Sanghyun, et al.
Pubblicazione: (2025)
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
di: Derner, Erik, et al.
Pubblicazione: (2026)
di: Derner, Erik, et al.
Pubblicazione: (2026)
Large-scale moral machine experiment on large language models
di: Ahmad, Muhammad Shahrul Zaim bin, et al.
Pubblicazione: (2024)
di: Ahmad, Muhammad Shahrul Zaim bin, et al.
Pubblicazione: (2024)
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
Conversational AI Powered by Large Language Models Amplifies False Memories in Witness Interviews
di: Chan, Samantha, et al.
Pubblicazione: (2024)
di: Chan, Samantha, et al.
Pubblicazione: (2024)
Theory of Mind and Self-Disclosure to CUIs
di: Cox, Samuel Rhys
Pubblicazione: (2025)
di: Cox, Samuel Rhys
Pubblicazione: (2025)
Understanding and Evaluating Trust in Generative AI and Large Language Models for Spreadsheets
di: Thorne, Simon
Pubblicazione: (2024)
di: Thorne, Simon
Pubblicazione: (2024)
Exploring Bengali Religious Dialect Biases in Large Language Models with Evaluation Perspectives
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2024)
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
di: Sarıtaş, Karahan, et al.
Pubblicazione: (2025) -
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
di: Arita, Takaya, et al.
Pubblicazione: (2025) -
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
di: Hirose, Manari, et al.
Pubblicazione: (2025) -
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
di: Ibrahim, Lujain, et al.
Pubblicazione: (2025) -
Open Models, Closed Minds? On Agents Capabilities in Mimicking Human Personalities through Open Large Language Models
di: La Cava, Lucio, et al.
Pubblicazione: (2024)