An Implementation of Werewolf Agent That does not Truly Trust LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sato, Takehiro, Ozaki, Shintaro, Yokoyama, Daisaku |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
Verbal Werewolf: Engage Users with Verbalized Agentic Werewolf Game Framework
von: Fan, Qihui, et al.
Veröffentlicht: (2025)
von: Fan, Qihui, et al.
Veröffentlicht: (2025)
Enhance Reasoning for Large Language Models in the Game Werewolf
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2023)
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2023)
Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
von: Qin, Chengwei, et al.
Veröffentlicht: (2024)
von: Qin, Chengwei, et al.
Veröffentlicht: (2024)
WereWolf-Plus: An Update of Werewolf Game setting Based on DSGBench
von: Xia, Xinyuan, et al.
Veröffentlicht: (2025)
von: Xia, Xinyuan, et al.
Veröffentlicht: (2025)
Enhancing Consistency of Werewolf AI through Dialogue Summarization and Persona Information
von: Tanaka, Yoshiki, et al.
Veröffentlicht: (2026)
von: Tanaka, Yoshiki, et al.
Veröffentlicht: (2026)
Identifying Influential N-grams in Confidence Calibration via Regression Analysis
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2026)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2026)
Enhancing Dialogue Generation in Werewolf Game Through Situation Analysis and Persuasion Strategies
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction
von: Bailis, Suma, et al.
Veröffentlicht: (2024)
von: Bailis, Suma, et al.
Veröffentlicht: (2024)
Helmsman of the Masses? Evaluate the Opinion Leadership of Large Language Models in the Werewolf Game
von: Du, Silin, et al.
Veröffentlicht: (2024)
von: Du, Silin, et al.
Veröffentlicht: (2024)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
von: Choi, Alexander S., et al.
Veröffentlicht: (2024)
von: Choi, Alexander S., et al.
Veröffentlicht: (2024)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
von: Kim, Ahrii, et al.
Veröffentlicht: (2026)
von: Kim, Ahrii, et al.
Veröffentlicht: (2026)
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
BQA: Body Language Question Answering Dataset for Video Large Language Models
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
Towards Cross-Lingual Explanation of Artwork in Large-scale Vision Language Models
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs
von: Galland, Lucie, et al.
Veröffentlicht: (2026)
von: Galland, Lucie, et al.
Veröffentlicht: (2026)
Decoding AI Authorship: Can LLMs Truly Mimic Human Style Across Literature and Politics?
von: Alsadhan, Nasser A
Veröffentlicht: (2026)
von: Alsadhan, Nasser A
Veröffentlicht: (2026)
Can AI Truly Represent Your Voice in Deliberations? A Comprehensive Study of Large-Scale Opinion Aggregation with LLMs
von: Zhu, Shenzhe, et al.
Veröffentlicht: (2025)
von: Zhu, Shenzhe, et al.
Veröffentlicht: (2025)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
von: Wang, Shouren, et al.
Veröffentlicht: (2025)
von: Wang, Shouren, et al.
Veröffentlicht: (2025)
In Agents We Trust, but Who Do Agents Trust? Latent Source Preferences Steer LLM Generations
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2026)
von: Khan, Mohammad Aflah, et al.
Veröffentlicht: (2026)
AraTrust: An Evaluation of Trustworthiness for LLMs in Arabic
von: Alghamdi, Emad A., et al.
Veröffentlicht: (2024)
von: Alghamdi, Emad A., et al.
Veröffentlicht: (2024)
How does a Language-Specific Tokenizer affect LLMs?
von: Seo, Jean, et al.
Veröffentlicht: (2025)
von: Seo, Jean, et al.
Veröffentlicht: (2025)
ParsTranslit: Truly Versatile Tajik-Farsi Transliteration
von: Merchant, Rayyan, et al.
Veröffentlicht: (2025)
von: Merchant, Rayyan, et al.
Veröffentlicht: (2025)
Can LLMs Truly Embody Human Personality? Analyzing AI and Human Behavior Alignment in Dispute Resolution
von: Kwon, Deuksin, et al.
Veröffentlicht: (2026)
von: Kwon, Deuksin, et al.
Veröffentlicht: (2026)
Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs
von: Sakai, Shintaro, et al.
Veröffentlicht: (2025)
von: Sakai, Shintaro, et al.
Veröffentlicht: (2025)
When to Trust LLMs: Aligning Confidence with Response Quality
von: Tao, Shuchang, et al.
Veröffentlicht: (2024)
von: Tao, Shuchang, et al.
Veröffentlicht: (2024)
Are LLMs Truly Multilingual? Exploring Zero-Shot Multilingual Capability of LLMs for Information Retrieval: An Italian Healthcare Use Case
von: Kembu, Vignesh Kumar, et al.
Veröffentlicht: (2025)
von: Kembu, Vignesh Kumar, et al.
Veröffentlicht: (2025)
Do Large Language Models Truly Understand Geometric Structures?
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2025)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
Building Trust in Clinical LLMs: Bias Analysis and Dataset Transparency
von: Maslenkova, Svetlana, et al.
Veröffentlicht: (2025)
von: Maslenkova, Svetlana, et al.
Veröffentlicht: (2025)
Trust, Safety, and Accuracy: Assessing LLMs for Routine Maternity Advice
von: Divya, V Sai, et al.
Veröffentlicht: (2026)
von: Divya, V Sai, et al.
Veröffentlicht: (2026)
Clinical knowledge in LLMs does not translate to human interactions
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Do Large Language Models Truly Understand Cross-cultural Differences?
von: Guo, Shiwei, et al.
Veröffentlicht: (2025)
von: Guo, Shiwei, et al.
Veröffentlicht: (2025)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
Evaluation data contamination in LLMs: how do we measure it and (when) does it matter?
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
Are Large Language Models Truly Smarter Than Humans?
von: M, Eshwar Reddy, et al.
Veröffentlicht: (2026)
von: M, Eshwar Reddy, et al.
Veröffentlicht: (2026)
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
von: Sato, Takuma, et al.
Veröffentlicht: (2025)
von: Sato, Takuma, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025) -
Verbal Werewolf: Engage Users with Verbalized Agentic Werewolf Game Framework
von: Fan, Qihui, et al.
Veröffentlicht: (2025) -
Enhance Reasoning for Large Language Models in the Game Werewolf
von: Wu, Shuang, et al.
Veröffentlicht: (2024) -
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2023) -
Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
von: Qin, Chengwei, et al.
Veröffentlicht: (2024)