An entropy penalized approach for stochastic control problems. Complete version

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bourdais, Thibaut, Oudjane, Nadia, Russo, Francesco
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916878479261696
author Bourdais, Thibaut
Oudjane, Nadia
Russo, Francesco
author_facet Bourdais, Thibaut
Oudjane, Nadia
Russo, Francesco
contents In this paper, we propose an original approach to stochastic control problems. We consider a weak formulation that is written as an optimization (minimization) problem on the space of probability measures. We then introduce a penalized version of this problem obtained by splitting the minimization variables and penalizing the discrepancy between the two variables via an entropy term. We show that the penalized problem provides a good approximation of the original problem when the weight of the entropy penalization term is large enough. Moreover, the penalized problem has the advantage of giving rise to two optimization subproblems that are easy to solve in each of the two optimization variables when the other is fixed. We take advantage of this property to propose an alternating optimization procedure that converges to the infimum of the penalized problem with a rate $O(1/k)$, where $k$ is the number of iterations. The relevance of this approach is illustrated by solving a high-dimensional stochastic control problem aimed at controlling consumption in electrical systems.
format Preprint
id arxiv_https___arxiv_org_abs_2309_01534
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle An entropy penalized approach for stochastic control problems. Complete version
Bourdais, Thibaut
Oudjane, Nadia
Russo, Francesco
Optimization and Control
Probability
In this paper, we propose an original approach to stochastic control problems. We consider a weak formulation that is written as an optimization (minimization) problem on the space of probability measures. We then introduce a penalized version of this problem obtained by splitting the minimization variables and penalizing the discrepancy between the two variables via an entropy term. We show that the penalized problem provides a good approximation of the original problem when the weight of the entropy penalization term is large enough. Moreover, the penalized problem has the advantage of giving rise to two optimization subproblems that are easy to solve in each of the two optimization variables when the other is fixed. We take advantage of this property to propose an alternating optimization procedure that converges to the infimum of the penalized problem with a rate $O(1/k)$, where $k$ is the number of iterations. The relevance of this approach is illustrated by solving a high-dimensional stochastic control problem aimed at controlling consumption in electrical systems.
title An entropy penalized approach for stochastic control problems. Complete version
topic Optimization and Control
Probability
url https://arxiv.org/abs/2309.01534