Performance Comparisons of Reinforcement Learning Algorithms for Sequential Experimental Design

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Barlas, Yasir Zubayr, Salako, Kizito
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918126899167232
author Barlas, Yasir Zubayr
Salako, Kizito
author_facet Barlas, Yasir Zubayr
Salako, Kizito
contents Recent developments in sequential experimental design look to construct a policy that can efficiently navigate the design space, in a way that maximises the expected information gain. Whilst there is work on achieving tractable policies for experimental design problems, there is significantly less work on obtaining policies that are able to generalise well - i.e. able to give good performance despite a change in the underlying statistical properties of the experiments. Conducting experiments sequentially has recently brought about the use of reinforcement learning, where an agent is trained to navigate the design space to select the most informative designs for experimentation. However, there is still a lack of understanding about the benefits and drawbacks of using certain reinforcement learning algorithms to train these agents. In our work, we investigate several reinforcement learning algorithms and their efficacy in producing agents that take maximally informative design decisions in sequential experimental design scenarios. We find that agent performance is impacted depending on the algorithm used for training, and that particular algorithms, using dropout or ensemble approaches, empirically showcase attractive generalisation properties.
format Preprint
id arxiv_https___arxiv_org_abs_2503_05905
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Performance Comparisons of Reinforcement Learning Algorithms for Sequential Experimental Design
Barlas, Yasir Zubayr
Salako, Kizito
Machine Learning
Recent developments in sequential experimental design look to construct a policy that can efficiently navigate the design space, in a way that maximises the expected information gain. Whilst there is work on achieving tractable policies for experimental design problems, there is significantly less work on obtaining policies that are able to generalise well - i.e. able to give good performance despite a change in the underlying statistical properties of the experiments. Conducting experiments sequentially has recently brought about the use of reinforcement learning, where an agent is trained to navigate the design space to select the most informative designs for experimentation. However, there is still a lack of understanding about the benefits and drawbacks of using certain reinforcement learning algorithms to train these agents. In our work, we investigate several reinforcement learning algorithms and their efficacy in producing agents that take maximally informative design decisions in sequential experimental design scenarios. We find that agent performance is impacted depending on the algorithm used for training, and that particular algorithms, using dropout or ensemble approaches, empirically showcase attractive generalisation properties.
title Performance Comparisons of Reinforcement Learning Algorithms for Sequential Experimental Design
topic Machine Learning
url https://arxiv.org/abs/2503.05905