Constrained Reinforcement Learning for Safe Heat Pump Control

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Zhang, Baohe, Frison, Lilli, Brox, Thomas, Bödecker, Joschka
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866916414629085184
author Zhang, Baohe
Frison, Lilli
Brox, Thomas
Bödecker, Joschka
author_facet Zhang, Baohe
Frison, Lilli
Brox, Thomas
Bödecker, Joschka
contents Constrained Reinforcement Learning (RL) has emerged as a significant research area within RL, where integrating constraints with rewards is crucial for enhancing safety and performance across diverse control tasks. In the context of heating systems in the buildings, optimizing the energy efficiency while maintaining the residents' thermal comfort can be intuitively formulated as a constrained optimization problem. However, to solve it with RL may require large amount of data. Therefore, an accurate and versatile simulator is favored. In this paper, we propose a novel building simulator I4B which provides interfaces for different usages and apply a model-free constrained RL algorithm named constrained Soft Actor-Critic with Linear Smoothed Log Barrier function (CSAC-LB) to the heating optimization problem. Benchmarking against baseline algorithms demonstrates CSAC-LB's efficiency in data exploration, constraint satisfaction and performance.
format Preprint
id arxiv_https___arxiv_org_abs_2409_19716
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Constrained Reinforcement Learning for Safe Heat Pump Control
Zhang, Baohe
Frison, Lilli
Brox, Thomas
Bödecker, Joschka
Machine Learning
Artificial Intelligence
Systems and Control
Constrained Reinforcement Learning (RL) has emerged as a significant research area within RL, where integrating constraints with rewards is crucial for enhancing safety and performance across diverse control tasks. In the context of heating systems in the buildings, optimizing the energy efficiency while maintaining the residents' thermal comfort can be intuitively formulated as a constrained optimization problem. However, to solve it with RL may require large amount of data. Therefore, an accurate and versatile simulator is favored. In this paper, we propose a novel building simulator I4B which provides interfaces for different usages and apply a model-free constrained RL algorithm named constrained Soft Actor-Critic with Linear Smoothed Log Barrier function (CSAC-LB) to the heating optimization problem. Benchmarking against baseline algorithms demonstrates CSAC-LB's efficiency in data exploration, constraint satisfaction and performance.
title Constrained Reinforcement Learning for Safe Heat Pump Control
topic Machine Learning
Artificial Intelligence
Systems and Control
url https://arxiv.org/abs/2409.19716