Do We Need Transformers to Play FPS Video Games?

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Batth, Karmanbir, Sethi, Krish, Shariff, Aly, Shi, Leo, Patel, Hetul
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915258159857664
author Batth, Karmanbir
Sethi, Krish
Shariff, Aly
Shi, Leo
Patel, Hetul
author_facet Batth, Karmanbir
Sethi, Krish
Shariff, Aly
Shi, Leo
Patel, Hetul
contents In this paper, we explore the Transformer based architectures for reinforcement learning in both online and offline settings within the Doom game environment. Our investigation focuses on two primary approaches: Deep Transformer Q- learning Networks (DTQN) for online learning and Decision Transformers (DT) for offline reinforcement learning. DTQN leverages the sequential modelling capabilities of Transformers to enhance Q-learning in partially observable environments,while Decision Transformers repurpose sequence modelling techniques to enable offline agents to learn from past trajectories without direct interaction with the environment. We conclude that while Transformers might have performed well in Atari games, more traditional methods perform better than Transformer based method in both the settings in the VizDoom environment.
format Preprint
id arxiv_https___arxiv_org_abs_2504_17891
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Do We Need Transformers to Play FPS Video Games?
Batth, Karmanbir
Sethi, Krish
Shariff, Aly
Shi, Leo
Patel, Hetul
Machine Learning
In this paper, we explore the Transformer based architectures for reinforcement learning in both online and offline settings within the Doom game environment. Our investigation focuses on two primary approaches: Deep Transformer Q- learning Networks (DTQN) for online learning and Decision Transformers (DT) for offline reinforcement learning. DTQN leverages the sequential modelling capabilities of Transformers to enhance Q-learning in partially observable environments,while Decision Transformers repurpose sequence modelling techniques to enable offline agents to learn from past trajectories without direct interaction with the environment. We conclude that while Transformers might have performed well in Atari games, more traditional methods perform better than Transformer based method in both the settings in the VizDoom environment.
title Do We Need Transformers to Play FPS Video Games?
topic Machine Learning
url https://arxiv.org/abs/2504.17891