WOLF: Werewolf-based Observations for LLM Deception and Falsehoods
Fuente:
arXiv
Saved in:
| Main Authors: | Agarwal, Mrinal, Rana, Saad, Sundoro, Theo, Berhe, Hermela, Kim, Spencer, Sharma, Vasu, O'Brien, Sean, Zhu, Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
by: Xu, Zelai, et al.
Published: (2023)
by: Xu, Zelai, et al.
Published: (2023)
Online Learning of Deceptive Policies under Intermittent Observation
by: Puthumanaillam, Gokul, et al.
Published: (2025)
by: Puthumanaillam, Gokul, et al.
Published: (2025)
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
by: Gho, Brendan, et al.
Published: (2025)
by: Gho, Brendan, et al.
Published: (2025)
Skill Description Deception Attack against Task Routing in Internet of Agents
by: He, Jiayi, et al.
Published: (2026)
by: He, Jiayi, et al.
Published: (2026)
Deception and Communication in Autonomous Multi-Agent Systems: An Experimental Study with Among Us
by: Milkowski, Maria, et al.
Published: (2026)
by: Milkowski, Maria, et al.
Published: (2026)
Reasoning Relay: Evaluating Stability and Interchangeability of Large Language Models in Mathematical Reasoning
by: Lu, Leo, et al.
Published: (2025)
by: Lu, Leo, et al.
Published: (2025)
CASPIAN: Online Detection and Attribution of Cascade Attacks in LLM Multi-Agent Systems via Cross-Channel Causal Monitoring
by: Venkatesh, Kavana, et al.
Published: (2026)
by: Venkatesh, Kavana, et al.
Published: (2026)
SpecBench: Evaluating Specification-Level Reasoning for Software Engineering LLM Agents
by: Hamblin, Grant, et al.
Published: (2026)
by: Hamblin, Grant, et al.
Published: (2026)
The Traitors: Deception and Trust in Multi-Agent Language Model Simulations
by: Curvo, Pedro M. P.
Published: (2025)
by: Curvo, Pedro M. P.
Published: (2025)
RAGAR, Your Falsehood Radar: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language Models
by: Khaliq, M. Abdul, et al.
Published: (2024)
by: Khaliq, M. Abdul, et al.
Published: (2024)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
by: Sinha, Aarush, et al.
Published: (2026)
by: Sinha, Aarush, et al.
Published: (2026)
Adaptive Accountability in Networked MAS: Tracing and Mitigating Emergent Norms at Scale
by: Alqithami, Saad
Published: (2025)
by: Alqithami, Saad
Published: (2025)
CH-MARL: Constrained Hierarchical Multiagent Reinforcement Learning for Sustainable Maritime Logistics
by: Alqithami, Saad
Published: (2025)
by: Alqithami, Saad
Published: (2025)
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
SMAGDi: Socratic Multi Agent Interaction Graph Distillation for Efficient High Accuracy Reasoning
by: Aluru, Aayush, et al.
Published: (2025)
by: Aluru, Aayush, et al.
Published: (2025)
Deception Analysis with Artificial Intelligence: An Interdisciplinary Perspective
by: Sarkadi, Stefan
Published: (2024)
by: Sarkadi, Stefan
Published: (2024)
Instruct, Not Assist: LLM-based Multi-Turn Planning and Hierarchical Questioning for Socratic Code Debugging
by: Kargupta, Priyanka, et al.
Published: (2024)
by: Kargupta, Priyanka, et al.
Published: (2024)
Human-guided Swarms: Impedance Control-inspired Influence in Virtual Reality Environments
by: Barclay, Spencer, et al.
Published: (2024)
by: Barclay, Spencer, et al.
Published: (2024)
Evaluating Creativity and Deception in Large Language Models: A Simulation Framework for Multi-Agent Balderdash
by: Hejabi, Parsa, et al.
Published: (2024)
by: Hejabi, Parsa, et al.
Published: (2024)
DC-Ada: Reward-Only Decentralized Sensor Adaptation for Heterogeneous Multi-Robot Teams
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Dynamic Strategy Adaptation in Multi-Agent Environments with Large Language Models
by: Mallampati, Shaurya, et al.
Published: (2025)
by: Mallampati, Shaurya, et al.
Published: (2025)
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
by: Gupta, Abhay, et al.
Published: (2025)
by: Gupta, Abhay, et al.
Published: (2025)
Self-Organizing Agent Network for LLM-based Workflow Automation
by: Xiong, Yiming, et al.
Published: (2025)
by: Xiong, Yiming, et al.
Published: (2025)
A speciation simulation that partly passes open-endedness tests
by: de Pinho, Théo, et al.
Published: (2026)
by: de Pinho, Théo, et al.
Published: (2026)
Bridging Perception and Action: A Lightweight Multimodal Meta-Planner Framework for Robust Earth Observation Agents
by: Xu, Jinghui, et al.
Published: (2026)
by: Xu, Jinghui, et al.
Published: (2026)
COREVQA: A Crowd Observation and Reasoning Entailment Visual Question Answering Benchmark
by: Chintapatla, Ishant, et al.
Published: (2025)
by: Chintapatla, Ishant, et al.
Published: (2025)
Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game
by: Mankarious, Saad
Published: (2026)
by: Mankarious, Saad
Published: (2026)
Soft Tournament Equilibrium
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Debate2Create: Robot Co-design via Multi-Agent LLM Debate
by: Qiu, Kevin, et al.
Published: (2025)
by: Qiu, Kevin, et al.
Published: (2025)
GOV-REK: Governed Reward Engineering Kernels for Designing Robust Multi-Agent Reinforcement Learning Systems
by: Rana, Ashish, et al.
Published: (2024)
by: Rana, Ashish, et al.
Published: (2024)
Deceptive Path Planning: A Bayesian Game Approach
by: Rostobaya, Violetta, et al.
Published: (2025)
by: Rostobaya, Violetta, et al.
Published: (2025)
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
by: Lin, Muhan, et al.
Published: (2025)
by: Lin, Muhan, et al.
Published: (2025)
LLM-ABM for Transportation: Assessing the Potential of LLM Agents in System Analysis
by: Liu, Tianming, et al.
Published: (2025)
by: Liu, Tianming, et al.
Published: (2025)
Inferring Occluded Agent Behavior in Dynamic Games from Noise Corrupted Observations
by: Qiu, Tianyu, et al.
Published: (2023)
by: Qiu, Tianyu, et al.
Published: (2023)
Is Your LLM-Based Multi-Agent a Reliable Real-World Planner? Exploring Fraud Detection in Travel Planning
by: Yao, Junchi, et al.
Published: (2025)
by: Yao, Junchi, et al.
Published: (2025)
$\aleph$-IPOMDP: Mitigating Deception in a Cognitive Hierarchy with Off-Policy Counterfactual Anomaly Detection
by: Alon, Nitay, et al.
Published: (2024)
by: Alon, Nitay, et al.
Published: (2024)
RMIO: A Model-Based MARL Framework for Scenarios with Observation Loss in Some Agents
by: Shi, Zifeng, et al.
Published: (2024)
by: Shi, Zifeng, et al.
Published: (2024)
LLM-ALSO: LLM-Driven Adaptive Learning-Signal Optimization for Multi-Agent Reinforcement Learning
by: Wu, Xiaoguang, et al.
Published: (2026)
by: Wu, Xiaoguang, et al.
Published: (2026)
MACRO-LLM: LLM-Empowered Multi-Agent Collaborative Reasoning under Spatiotemporal Partial Observability
by: Chen, Handi, et al.
Published: (2026)
by: Chen, Handi, et al.
Published: (2026)
The Observer-Situation Lattice: A Unified Formal Basis for Perspective-Aware Cognition
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Similar Items
-
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
by: Xu, Zelai, et al.
Published: (2023) -
Online Learning of Deceptive Policies under Intermittent Observation
by: Puthumanaillam, Gokul, et al.
Published: (2025) -
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
by: Gho, Brendan, et al.
Published: (2025) -
Skill Description Deception Attack against Task Routing in Internet of Agents
by: He, Jiayi, et al.
Published: (2026) -
Deception and Communication in Autonomous Multi-Agent Systems: An Experimental Study with Among Us
by: Milkowski, Maria, et al.
Published: (2026)