Is this the real life? Is this just fantasy? The Misleading Success of Simulating Social Interactions With LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Xuhui, Su, Zhe, Eisape, Tiwalayo, Kim, Hyunwoo, Sap, Maarten |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses
di: Mun, Jimin, et al.
Pubblicazione: (2026)
di: Mun, Jimin, et al.
Pubblicazione: (2026)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
di: Su, Zhe, et al.
Pubblicazione: (2024)
di: Su, Zhe, et al.
Pubblicazione: (2024)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
di: Fan, Xianzhe, et al.
Pubblicazione: (2025)
di: Fan, Xianzhe, et al.
Pubblicazione: (2025)
A Systematic Comparison of Syllogistic Reasoning in Humans and Language Models
di: Eisape, Tiwalayo, et al.
Pubblicazione: (2023)
di: Eisape, Tiwalayo, et al.
Pubblicazione: (2023)
Social World Models
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
Is the Pope Catholic? Yes, the Pope is Catholic. Generative Evaluation of Non-Literal Intent Resolution in LLMs
di: Yerukola, Akhila, et al.
Pubblicazione: (2024)
di: Yerukola, Akhila, et al.
Pubblicazione: (2024)
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
di: Zhou, Xuhui, et al.
Pubblicazione: (2023)
di: Zhou, Xuhui, et al.
Pubblicazione: (2023)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
di: Cohen, Myke C., et al.
Pubblicazione: (2026)
di: Cohen, Myke C., et al.
Pubblicazione: (2026)
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
di: Fan, Xianzhe, et al.
Pubblicazione: (2024)
di: Fan, Xianzhe, et al.
Pubblicazione: (2024)
Improving the Efficiency of Language Agent Teams with Adaptive Task Graphs
di: Mieczkowski, Elizabeth, et al.
Pubblicazione: (2026)
di: Mieczkowski, Elizabeth, et al.
Pubblicazione: (2026)
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
di: Cohen, Myke C., et al.
Pubblicazione: (2025)
di: Cohen, Myke C., et al.
Pubblicazione: (2025)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
Rethinking Theory of Mind Benchmarks for LLMs: Towards A User-Centered Perspective
di: Wang, Qiaosi, et al.
Pubblicazione: (2025)
di: Wang, Qiaosi, et al.
Pubblicazione: (2025)
Breaking mBad! Supervised Fine-tuning for Cross-Lingual Detoxification
di: Beniwal, Himanshu, et al.
Pubblicazione: (2025)
di: Beniwal, Himanshu, et al.
Pubblicazione: (2025)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
di: Kim, Jinhwa, et al.
Pubblicazione: (2025)
di: Kim, Jinhwa, et al.
Pubblicazione: (2025)
Training Proactive and Personalized LLM Agents
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2024)
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2024)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
di: Sun, Zhongxiang, et al.
Pubblicazione: (2026)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2026)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2024)
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2024)
Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations
di: Zheng, Mingqian, et al.
Pubblicazione: (2026)
di: Zheng, Mingqian, et al.
Pubblicazione: (2026)
Martingale Score: An Unsupervised Metric for Bayesian Rationality in LLM Reasoning
di: He, Zhonghao, et al.
Pubblicazione: (2025)
di: He, Zhonghao, et al.
Pubblicazione: (2025)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
Long Chain-of-Thought Reasoning Across Languages
di: Barua, Josh, et al.
Pubblicazione: (2025)
di: Barua, Josh, et al.
Pubblicazione: (2025)
Disparities in LLM Reasoning Accuracy and Explanations: A Case Study on African American English
di: Zhou, Runtao, et al.
Pubblicazione: (2025)
di: Zhou, Runtao, et al.
Pubblicazione: (2025)
CHARTOM: A Visual Theory-of-Mind Benchmark for LLMs on Misleading Charts
di: Bharti, Shubham, et al.
Pubblicazione: (2024)
di: Bharti, Shubham, et al.
Pubblicazione: (2024)
How Good (Or Bad) Are LLMs at Detecting Misleading Visualizations?
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
Mitigating Misleading Chain-of-Thought Reasoning with Selective Filtering
di: Wu, Yexin, et al.
Pubblicazione: (2024)
di: Wu, Yexin, et al.
Pubblicazione: (2024)
Mind the Gesture: Evaluating AI Sensitivity to Culturally Offensive Non-Verbal Gestures
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
Large Language Models as Misleading Assistants in Conversation
di: Hou, Betty Li, et al.
Pubblicazione: (2024)
di: Hou, Betty Li, et al.
Pubblicazione: (2024)
The LLM Fallacy: Misattribution in AI-Assisted Cognitive Workflows
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
Natural Language Declarative Prompting (NLD-P): A Modular Governance Method for Prompt Design Under Model Drift
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
di: Song, Maojia, et al.
Pubblicazione: (2025)
di: Song, Maojia, et al.
Pubblicazione: (2025)
Easier to Mislead Than to Correct: Harmful and Beneficial Revision in LLM Conformity
di: Qu, Jiaming, et al.
Pubblicazione: (2026)
di: Qu, Jiaming, et al.
Pubblicazione: (2026)
Fine-Grained and Thematic Evaluation of LLMs in Social Deduction Game
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
di: Gu, Yuling, et al.
Pubblicazione: (2024)
di: Gu, Yuling, et al.
Pubblicazione: (2024)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
di: Li, Jing-Jing, et al.
Pubblicazione: (2026)
di: Li, Jing-Jing, et al.
Pubblicazione: (2026)
Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog Generation
di: Park, Jinyoung, et al.
Pubblicazione: (2024)
di: Park, Jinyoung, et al.
Pubblicazione: (2024)
Learning Shortcuts: On the Misleading Promise of NLU in Language Models
di: Bihani, Geetanjali, et al.
Pubblicazione: (2024)
di: Bihani, Geetanjali, et al.
Pubblicazione: (2024)
Unmasking Deceptive Visuals: Benchmarking Multimodal Large Language Models on Misleading Chart Question Answering
di: Chen, Zixin, et al.
Pubblicazione: (2025)
di: Chen, Zixin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses
di: Mun, Jimin, et al.
Pubblicazione: (2026) -
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023) -
AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
di: Su, Zhe, et al.
Pubblicazione: (2024) -
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
di: Fan, Xianzhe, et al.
Pubblicazione: (2025) -
A Systematic Comparison of Syllogistic Reasoning in Humans and Language Models
di: Eisape, Tiwalayo, et al.
Pubblicazione: (2023)