Staff View: :: Library Catalog

Saved in:

Bibliographic Details
Main Author:	Rivasseau, Thomas
Format:	Preprint
Published:	2026
Subjects:	Artificial Intelligence
Online Access:	https://arxiv.org/abs/2604.02500
Tags:	Add Tag No Tags, Be the first to tag this record!

_version_	1866917392152526848
author	Rivasseau, Thomas
author_facet	Rivasseau, Thomas
contents	As ongoing research explores the ability of AI agents to be insider threats and act against company interests, we showcase the abilities of such agents to act against human well being in service of corporate authority. Building on Agentic Misalignment and AI scheming research, we present a scenario where the majority of evaluated state-of-the-art AI agents explicitly choose to suppress evidence of fraud and harm, in service of company profit. We test this scenario on 16 recent Large Language Models. Some models show remarkable resistance to our method and behave appropriately, but many do not, and instead aid and abet criminal activity. These experiments are simulations and were executed in a controlled virtual environment. No crime actually occurred.
format	Preprint
id	arxiv_https___arxiv_org_abs_2604_02500
institution	arXiv
publishDate	2026
record_format	arxiv
spellingShingle	I must delete the evidence: AI Agents Explicitly Cover up Fraud and Violent Crime Rivasseau, Thomas Artificial Intelligence As ongoing research explores the ability of AI agents to be insider threats and act against company interests, we showcase the abilities of such agents to act against human well being in service of corporate authority. Building on Agentic Misalignment and AI scheming research, we present a scenario where the majority of evaluated state-of-the-art AI agents explicitly choose to suppress evidence of fraud and harm, in service of company profit. We test this scenario on 16 recent Large Language Models. Some models show remarkable resistance to our method and behave appropriately, but many do not, and instead aid and abet criminal activity. These experiments are simulations and were executed in a controlled virtual environment. No crime actually occurred.
title	I must delete the evidence: AI Agents Explicitly Cover up Fraud and Violent Crime
topic	Artificial Intelligence
url	https://arxiv.org/abs/2604.02500

Similar Items