Learning Robust Policies for Uncertain Parametric Markov Decision Processes

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Rickard, Luke, Abate, Alessandro, Margellos, Kostas
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910448486449152
author Rickard, Luke
Abate, Alessandro
Margellos, Kostas
author_facet Rickard, Luke
Abate, Alessandro
Margellos, Kostas
contents Synthesising verifiably correct controllers for dynamical systems is crucial for safety-critical problems. To achieve this, it is important to account for uncertainty in a robust manner, while at the same time it is often of interest to avoid being overly conservative with the view of achieving a better cost. We propose a method for verifiably safe policy synthesis for a class of finite state models, under the presence of structural uncertainty. In particular, we consider uncertain parametric Markov decision processes (upMDPs), a special class of Markov decision processes, with parameterised transition functions, where such parameters are drawn from a (potentially) unknown distribution. Our framework leverages recent advancements in the so-called scenario approach theory, where we represent the uncertainty by means of scenarios, and provide guarantees on synthesised policies satisfying probabilistic computation tree logic (PCTL) formulae. We consider several common benchmarks/problems and compare our work to recent developments for verifying upMDPs.
format Preprint
id arxiv_https___arxiv_org_abs_2312_06344
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Learning Robust Policies for Uncertain Parametric Markov Decision Processes
Rickard, Luke
Abate, Alessandro
Margellos, Kostas
Systems and Control
Logic in Computer Science
Synthesising verifiably correct controllers for dynamical systems is crucial for safety-critical problems. To achieve this, it is important to account for uncertainty in a robust manner, while at the same time it is often of interest to avoid being overly conservative with the view of achieving a better cost. We propose a method for verifiably safe policy synthesis for a class of finite state models, under the presence of structural uncertainty. In particular, we consider uncertain parametric Markov decision processes (upMDPs), a special class of Markov decision processes, with parameterised transition functions, where such parameters are drawn from a (potentially) unknown distribution. Our framework leverages recent advancements in the so-called scenario approach theory, where we represent the uncertainty by means of scenarios, and provide guarantees on synthesised policies satisfying probabilistic computation tree logic (PCTL) formulae. We consider several common benchmarks/problems and compare our work to recent developments for verifying upMDPs.
title Learning Robust Policies for Uncertain Parametric Markov Decision Processes
topic Systems and Control
Logic in Computer Science
url https://arxiv.org/abs/2312.06344