An Implementation of Werewolf Agent That does not Truly Trust LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Sato, Takehiro, Ozaki, Shintaro, Yokoyama, Daisaku |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Strategy Adaptation in Large Language Model Werewolf Agents
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
Verbal Werewolf: Engage Users with Verbalized Agentic Werewolf Game Framework
di: Fan, Qihui, et al.
Pubblicazione: (2025)
di: Fan, Qihui, et al.
Pubblicazione: (2025)
Enhance Reasoning for Large Language Models in the Game Werewolf
di: Wu, Shuang, et al.
Pubblicazione: (2024)
di: Wu, Shuang, et al.
Pubblicazione: (2024)
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
di: Xu, Yuzhuang, et al.
Pubblicazione: (2023)
di: Xu, Yuzhuang, et al.
Pubblicazione: (2023)
Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
di: Qin, Chengwei, et al.
Pubblicazione: (2024)
di: Qin, Chengwei, et al.
Pubblicazione: (2024)
WereWolf-Plus: An Update of Werewolf Game setting Based on DSGBench
di: Xia, Xinyuan, et al.
Pubblicazione: (2025)
di: Xia, Xinyuan, et al.
Pubblicazione: (2025)
Enhancing Consistency of Werewolf AI through Dialogue Summarization and Persona Information
di: Tanaka, Yoshiki, et al.
Pubblicazione: (2026)
di: Tanaka, Yoshiki, et al.
Pubblicazione: (2026)
Identifying Influential N-grams in Confidence Calibration via Regression Analysis
di: Ozaki, Shintaro, et al.
Pubblicazione: (2026)
di: Ozaki, Shintaro, et al.
Pubblicazione: (2026)
Enhancing Dialogue Generation in Werewolf Game Through Situation Analysis and Persuasion Strategies
di: Qi, Zhiyang, et al.
Pubblicazione: (2024)
di: Qi, Zhiyang, et al.
Pubblicazione: (2024)
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction
di: Bailis, Suma, et al.
Pubblicazione: (2024)
di: Bailis, Suma, et al.
Pubblicazione: (2024)
Helmsman of the Masses? Evaluate the Opinion Leadership of Large Language Models in the Werewolf Game
di: Du, Silin, et al.
Pubblicazione: (2024)
di: Du, Silin, et al.
Pubblicazione: (2024)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
di: Choi, Alexander S., et al.
Pubblicazione: (2024)
di: Choi, Alexander S., et al.
Pubblicazione: (2024)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
di: Kim, Ahrii, et al.
Pubblicazione: (2026)
di: Kim, Ahrii, et al.
Pubblicazione: (2026)
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
di: Hayashi, Kazuki, et al.
Pubblicazione: (2025)
di: Hayashi, Kazuki, et al.
Pubblicazione: (2025)
BQA: Body Language Question Answering Dataset for Video Large Language Models
di: Ozaki, Shintaro, et al.
Pubblicazione: (2024)
di: Ozaki, Shintaro, et al.
Pubblicazione: (2024)
Towards Cross-Lingual Explanation of Artwork in Large-scale Vision Language Models
di: Ozaki, Shintaro, et al.
Pubblicazione: (2024)
di: Ozaki, Shintaro, et al.
Pubblicazione: (2024)
Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance
di: Ozaki, Shintaro, et al.
Pubblicazione: (2025)
di: Ozaki, Shintaro, et al.
Pubblicazione: (2025)
Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs
di: Galland, Lucie, et al.
Pubblicazione: (2026)
di: Galland, Lucie, et al.
Pubblicazione: (2026)
Decoding AI Authorship: Can LLMs Truly Mimic Human Style Across Literature and Politics?
di: Alsadhan, Nasser A
Pubblicazione: (2026)
di: Alsadhan, Nasser A
Pubblicazione: (2026)
Can AI Truly Represent Your Voice in Deliberations? A Comprehensive Study of Large-Scale Opinion Aggregation with LLMs
di: Zhu, Shenzhe, et al.
Pubblicazione: (2025)
di: Zhu, Shenzhe, et al.
Pubblicazione: (2025)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
di: Wang, Shouren, et al.
Pubblicazione: (2025)
di: Wang, Shouren, et al.
Pubblicazione: (2025)
In Agents We Trust, but Who Do Agents Trust? Latent Source Preferences Steer LLM Generations
di: Khan, Mohammad Aflah, et al.
Pubblicazione: (2026)
di: Khan, Mohammad Aflah, et al.
Pubblicazione: (2026)
AraTrust: An Evaluation of Trustworthiness for LLMs in Arabic
di: Alghamdi, Emad A., et al.
Pubblicazione: (2024)
di: Alghamdi, Emad A., et al.
Pubblicazione: (2024)
How does a Language-Specific Tokenizer affect LLMs?
di: Seo, Jean, et al.
Pubblicazione: (2025)
di: Seo, Jean, et al.
Pubblicazione: (2025)
ParsTranslit: Truly Versatile Tajik-Farsi Transliteration
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
Can LLMs Truly Embody Human Personality? Analyzing AI and Human Behavior Alignment in Dispute Resolution
di: Kwon, Deuksin, et al.
Pubblicazione: (2026)
di: Kwon, Deuksin, et al.
Pubblicazione: (2026)
Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs
di: Sakai, Shintaro, et al.
Pubblicazione: (2025)
di: Sakai, Shintaro, et al.
Pubblicazione: (2025)
When to Trust LLMs: Aligning Confidence with Response Quality
di: Tao, Shuchang, et al.
Pubblicazione: (2024)
di: Tao, Shuchang, et al.
Pubblicazione: (2024)
Are LLMs Truly Multilingual? Exploring Zero-Shot Multilingual Capability of LLMs for Information Retrieval: An Italian Healthcare Use Case
di: Kembu, Vignesh Kumar, et al.
Pubblicazione: (2025)
di: Kembu, Vignesh Kumar, et al.
Pubblicazione: (2025)
Do Large Language Models Truly Understand Geometric Structures?
di: Wang, Xiaofeng, et al.
Pubblicazione: (2025)
di: Wang, Xiaofeng, et al.
Pubblicazione: (2025)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
di: Yang, Haoyan, et al.
Pubblicazione: (2024)
di: Yang, Haoyan, et al.
Pubblicazione: (2024)
Building Trust in Clinical LLMs: Bias Analysis and Dataset Transparency
di: Maslenkova, Svetlana, et al.
Pubblicazione: (2025)
di: Maslenkova, Svetlana, et al.
Pubblicazione: (2025)
Trust, Safety, and Accuracy: Assessing LLMs for Routine Maternity Advice
di: Divya, V Sai, et al.
Pubblicazione: (2026)
di: Divya, V Sai, et al.
Pubblicazione: (2026)
Clinical knowledge in LLMs does not translate to human interactions
di: Bean, Andrew M., et al.
Pubblicazione: (2025)
di: Bean, Andrew M., et al.
Pubblicazione: (2025)
Do LLMs Truly Understand When a Precedent Is Overruled?
di: Zhang, Li, et al.
Pubblicazione: (2025)
di: Zhang, Li, et al.
Pubblicazione: (2025)
Do Large Language Models Truly Understand Cross-cultural Differences?
di: Guo, Shiwei, et al.
Pubblicazione: (2025)
di: Guo, Shiwei, et al.
Pubblicazione: (2025)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
di: Chern, Steffi, et al.
Pubblicazione: (2024)
di: Chern, Steffi, et al.
Pubblicazione: (2024)
Evaluation data contamination in LLMs: how do we measure it and (when) does it matter?
di: Singh, Aaditya K., et al.
Pubblicazione: (2024)
di: Singh, Aaditya K., et al.
Pubblicazione: (2024)
Are Large Language Models Truly Smarter Than Humans?
di: M, Eshwar Reddy, et al.
Pubblicazione: (2026)
di: M, Eshwar Reddy, et al.
Pubblicazione: (2026)
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
di: Sato, Takuma, et al.
Pubblicazione: (2025)
di: Sato, Takuma, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Strategy Adaptation in Large Language Model Werewolf Agents
di: Nakamori, Fuya, et al.
Pubblicazione: (2025) -
Verbal Werewolf: Engage Users with Verbalized Agentic Werewolf Game Framework
di: Fan, Qihui, et al.
Pubblicazione: (2025) -
Enhance Reasoning for Large Language Models in the Game Werewolf
di: Wu, Shuang, et al.
Pubblicazione: (2024) -
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
di: Xu, Yuzhuang, et al.
Pubblicazione: (2023) -
Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
di: Qin, Chengwei, et al.
Pubblicazione: (2024)