Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions

Fuente: arXiv
Saved in:
Bibliographic Details
Main Author: Zhang, Harry
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916259775381504
author Zhang, Harry
author_facet Zhang, Harry
contents Model-based Reinforcement Learning (MBRL) has shown many desirable properties for intelligent control tasks. However, satisfying safety and stability constraints during training and rollout remains an open question. We propose a new Model-based RL framework to enable efficient policy learning with unknown dynamics based on learning model predictive control (LMPC) framework with mathematically provable guarantees of stability. We introduce and explore a novel method for adding safety constraints for model-based RL during training and policy learning. The new stability-augmented framework consists of a neural-network-based learner that learns to construct a Lyapunov function, and a model-based RL agent to consistently complete the tasks while satisfying user-specified constraints given only sub-optimal demonstrations and sparse-cost feedback. We demonstrate the capability of the proposed framework through simulated experiments.
format Preprint
id arxiv_https___arxiv_org_abs_2405_16184
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions
Zhang, Harry
Systems and Control
Artificial Intelligence
Machine Learning
Model-based Reinforcement Learning (MBRL) has shown many desirable properties for intelligent control tasks. However, satisfying safety and stability constraints during training and rollout remains an open question. We propose a new Model-based RL framework to enable efficient policy learning with unknown dynamics based on learning model predictive control (LMPC) framework with mathematically provable guarantees of stability. We introduce and explore a novel method for adding safety constraints for model-based RL during training and policy learning. The new stability-augmented framework consists of a neural-network-based learner that learns to construct a Lyapunov function, and a model-based RL agent to consistently complete the tasks while satisfying user-specified constraints given only sub-optimal demonstrations and sparse-cost feedback. We demonstrate the capability of the proposed framework through simulated experiments.
title Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions
topic Systems and Control
Artificial Intelligence
Machine Learning
url https://arxiv.org/abs/2405.16184