Scaling Behavior Cloning Improves Causal Reasoning: An Open Model for Real-Time Video Game Playing

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yue, Yuguang, Salia, Irakli, Hunt, Samuel, Green, Chris, Shi, Wenzhe, Hunt, Jonathan J
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912856991072256
author Yue, Yuguang
Salia, Irakli
Hunt, Samuel
Green, Chris
Shi, Wenzhe
Hunt, Jonathan J
author_facet Yue, Yuguang
Salia, Irakli
Hunt, Samuel
Green, Chris
Shi, Wenzhe
Hunt, Jonathan J
contents Behavior cloning has seen a resurgence as scaling model and data sizes demonstrate strong performance. In this work, we introduce an open recipe for training a video game playing foundation model designed for inference in realtime on a consumer GPU. We release all data (8300+ hours of high quality human gameplay), training and inference code, and pretrained checkpoints under an open license. Empirically, we show that our best model achieves performance competitive with human players across a variety of 3D games. We use this recipe to investigate the scaling laws of behavior cloning, with a focus on causal reasoning. In a controlled toy setting, we first demonstrate that increasing training data and network depth leads to the model learning a more causal policy. We then validate these findings at scale, analyzing models up to 1.2 billion parameters. We observe that the causal improvements seen in the toy domain hold true as model size and training steps increase.
format Preprint
id arxiv_https___arxiv_org_abs_2601_04575
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Scaling Behavior Cloning Improves Causal Reasoning: An Open Model for Real-Time Video Game Playing
Yue, Yuguang
Salia, Irakli
Hunt, Samuel
Green, Chris
Shi, Wenzhe
Hunt, Jonathan J
Artificial Intelligence
Behavior cloning has seen a resurgence as scaling model and data sizes demonstrate strong performance. In this work, we introduce an open recipe for training a video game playing foundation model designed for inference in realtime on a consumer GPU. We release all data (8300+ hours of high quality human gameplay), training and inference code, and pretrained checkpoints under an open license. Empirically, we show that our best model achieves performance competitive with human players across a variety of 3D games. We use this recipe to investigate the scaling laws of behavior cloning, with a focus on causal reasoning. In a controlled toy setting, we first demonstrate that increasing training data and network depth leads to the model learning a more causal policy. We then validate these findings at scale, analyzing models up to 1.2 billion parameters. We observe that the causal improvements seen in the toy domain hold true as model size and training steps increase.
title Scaling Behavior Cloning Improves Causal Reasoning: An Open Model for Real-Time Video Game Playing
topic Artificial Intelligence
url https://arxiv.org/abs/2601.04575