Saved in:
Bibliographic Details
Main Author: Rivasseau, Thomas
Format: Preprint
Published: 2026
Subjects:
Online Access:https://arxiv.org/abs/2604.02500
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917392152526848
author Rivasseau, Thomas
author_facet Rivasseau, Thomas
contents As ongoing research explores the ability of AI agents to be insider threats and act against company interests, we showcase the abilities of such agents to act against human well being in service of corporate authority. Building on Agentic Misalignment and AI scheming research, we present a scenario where the majority of evaluated state-of-the-art AI agents explicitly choose to suppress evidence of fraud and harm, in service of company profit. We test this scenario on 16 recent Large Language Models. Some models show remarkable resistance to our method and behave appropriately, but many do not, and instead aid and abet criminal activity. These experiments are simulations and were executed in a controlled virtual environment. No crime actually occurred.
format Preprint
id arxiv_https___arxiv_org_abs_2604_02500
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle I must delete the evidence: AI Agents Explicitly Cover up Fraud and Violent Crime
Rivasseau, Thomas
Artificial Intelligence
As ongoing research explores the ability of AI agents to be insider threats and act against company interests, we showcase the abilities of such agents to act against human well being in service of corporate authority. Building on Agentic Misalignment and AI scheming research, we present a scenario where the majority of evaluated state-of-the-art AI agents explicitly choose to suppress evidence of fraud and harm, in service of company profit. We test this scenario on 16 recent Large Language Models. Some models show remarkable resistance to our method and behave appropriately, but many do not, and instead aid and abet criminal activity. These experiments are simulations and were executed in a controlled virtual environment. No crime actually occurred.
title I must delete the evidence: AI Agents Explicitly Cover up Fraud and Violent Crime
topic Artificial Intelligence
url https://arxiv.org/abs/2604.02500