From Hallucination to Scheming: A Unified Taxonomy and Benchmark Analysis for LLM Deception
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Jerick, Zhang, Terry Jingcheng, Jin, Zhijing, Conitzer, Vincent |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cheap Talk, Empty Promise: Frontier LLMs easily break public promises for self-interest
by: Shi, Jerick, et al.
Published: (2026)
by: Shi, Jerick, et al.
Published: (2026)
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
by: Tewolde, Emanuel, et al.
Published: (2026)
by: Tewolde, Emanuel, et al.
Published: (2026)
Domain-Independent Deception: A New Taxonomy and Linguistic Analysis
by: Verma, Rakesh M., et al.
Published: (2024)
by: Verma, Rakesh M., et al.
Published: (2024)
Causality for Natural Language Processing
by: Jin, Zhijing
Published: (2025)
by: Jin, Zhijing
Published: (2025)
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
by: Horvitz, Eric, et al.
Published: (2024)
by: Horvitz, Eric, et al.
Published: (2024)
Banal Deception Human-AI Ecosystems: A Study of People's Perceptions of LLM-generated Deceptive Behaviour
by: Zhan, Xiao, et al.
Published: (2024)
by: Zhan, Xiao, et al.
Published: (2024)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
by: Cobben, Pepijn, et al.
Published: (2026)
by: Cobben, Pepijn, et al.
Published: (2026)
LLM Harms: A Taxonomy and Discussion
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
by: Li, Miles Q., et al.
Published: (2026)
by: Li, Miles Q., et al.
Published: (2026)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
by: Bhagat, Kirti, et al.
Published: (2025)
by: Bhagat, Kirti, et al.
Published: (2025)
Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach
by: Oketch, Kezia, et al.
Published: (2025)
by: Oketch, Kezia, et al.
Published: (2025)
LLM Agents in Law: Taxonomy, Applications, and Challenges
by: Liu, Shuang, et al.
Published: (2026)
by: Liu, Shuang, et al.
Published: (2026)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
by: Backmann, Steffen, et al.
Published: (2025)
by: Backmann, Steffen, et al.
Published: (2025)
Theorizing Deception: A Scoping Review of Theory in Research on Dark Patterns and Deceptive Design
by: Chang, Weichen Joe, et al.
Published: (2024)
by: Chang, Weichen Joe, et al.
Published: (2024)
Deception and Manipulation in Generative AI
by: Tarsney, Christian
Published: (2024)
by: Tarsney, Christian
Published: (2024)
On the Pros and Cons of Active Learning for Moral Preference Elicitation
by: Keswani, Vijay, et al.
Published: (2024)
by: Keswani, Vijay, et al.
Published: (2024)
Task-Dependent Evaluation of LLM Output Homogenization: A Taxonomy-Guided Framework
by: Jain, Shomik, et al.
Published: (2025)
by: Jain, Shomik, et al.
Published: (2025)
Market-Dependent Communication in Multi-Agent Alpha Generation
by: Shi, Jerick, et al.
Published: (2025)
by: Shi, Jerick, et al.
Published: (2025)
Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
by: Yang, Shujian, et al.
Published: (2025)
by: Yang, Shujian, et al.
Published: (2025)
Predictive Power of LLMs in Financial Markets
by: Shi, Jerick, et al.
Published: (2024)
by: Shi, Jerick, et al.
Published: (2024)
Ethical Implications of Training Deceptive AI
by: Starace, Jason, et al.
Published: (2026)
by: Starace, Jason, et al.
Published: (2026)
Access Over Deception: Fighting Deceptive Patterns through Accessibility
by: Pellkvist, Tobias, et al.
Published: (2026)
by: Pellkvist, Tobias, et al.
Published: (2026)
Deception Analysis with Artificial Intelligence: An Interdisciplinary Perspective
by: Sarkadi, Stefan
Published: (2024)
by: Sarkadi, Stefan
Published: (2024)
Can AI Model the Complexities of Human Moral Decision-Making? A Qualitative Study of Kidney Allocation Decisions
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods
by: Boerstler, Kyle, et al.
Published: (2024)
by: Boerstler, Kyle, et al.
Published: (2024)
Automatically Detecting Online Deceptive Patterns
by: Nayak, Asmit, et al.
Published: (2024)
by: Nayak, Asmit, et al.
Published: (2024)
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
by: Melis, Matteo, et al.
Published: (2025)
by: Melis, Matteo, et al.
Published: (2025)
The Patterns of Digital Deception
by: Dickinson, Gregory M.
Published: (2024)
by: Dickinson, Gregory M.
Published: (2024)
The AI-Fraud Diamond: A Novel Lens for Auditing Algorithmic Deception
by: Zweers, Benjamin, et al.
Published: (2025)
by: Zweers, Benjamin, et al.
Published: (2025)
The Complexity of Computing Robust Mediated Equilibria in Ordinal Games
by: Conitzer, Vincent
Published: (2024)
by: Conitzer, Vincent
Published: (2024)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
by: Jenny, David F., et al.
Published: (2023)
by: Jenny, David F., et al.
Published: (2023)
Mapping Compliance: A Taxonomy for Political Content Analysis under the EU's Digital Electoral Framework
by: Sekwenz, Marie-Therese, et al.
Published: (2025)
by: Sekwenz, Marie-Therese, et al.
Published: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
by: Choi, Younwoo, et al.
Published: (2025)
by: Choi, Younwoo, et al.
Published: (2025)
Stuck in the Turing Matrix: Inauthenticity, Deception and the Social Life of AI
by: Collins, Samuel Gerald
Published: (2026)
by: Collins, Samuel Gerald
Published: (2026)
The Traffickers' Pitch: Detecting Deceptive Recruitment in Online Job Boards
by: Zhou, Siyi, et al.
Published: (2026)
by: Zhou, Siyi, et al.
Published: (2026)
Towards A Framework for Levels of Anthropomorphic Deception in Robots and AI
by: Babel, Franziska, et al.
Published: (2026)
by: Babel, Franziska, et al.
Published: (2026)
Integrating Dark Pattern Taxonomies
by: Lewis, Frank, et al.
Published: (2024)
by: Lewis, Frank, et al.
Published: (2024)
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories
by: Ho, Jessee, et al.
Published: (2026)
by: Ho, Jessee, et al.
Published: (2026)
Similar Items
-
Cheap Talk, Empty Promise: Frontier LLMs easily break public promises for self-interest
by: Shi, Jerick, et al.
Published: (2026) -
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
by: Tewolde, Emanuel, et al.
Published: (2026) -
Domain-Independent Deception: A New Taxonomy and Linguistic Analysis
by: Verma, Rakesh M., et al.
Published: (2024) -
Causality for Natural Language Processing
by: Jin, Zhijing
Published: (2025) -
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
by: Horvitz, Eric, et al.
Published: (2024)