DiffTune-MPC: Closed-Loop Learning for Model Predictive Control

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Tao, Ran, Cheng, Sheng, Wang, Xiaofeng, Wang, Shenlong, Hovakimyan, Naira
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911944957493248
author Tao, Ran
Cheng, Sheng
Wang, Xiaofeng
Wang, Shenlong
Hovakimyan, Naira
author_facet Tao, Ran
Cheng, Sheng
Wang, Xiaofeng
Wang, Shenlong
Hovakimyan, Naira
contents Model predictive control (MPC) has been applied to many platforms in robotics and autonomous systems for its capability to predict a system's future behavior while incorporating constraints that a system may have. To enhance the performance of a system with an MPC controller, one can manually tune the MPC's cost function. However, it can be challenging due to the possibly high dimension of the parameter space as well as the potential difference between the open-loop cost function in MPC and the overall closed-loop performance metric function. This paper presents DiffTune-MPC, a novel learning method, to learn the cost function of an MPC in a closed-loop manner. The proposed framework is compatible with the scenario where the time interval for performance evaluation and MPC's planning horizon have different lengths. We show the auxiliary problem whose solution admits the analytical gradients of MPC and discuss its variations in different MPC settings, including nonlinear MPCs that are solved using sequential quadratic programming. Simulation results demonstrate the learning capability of DiffTune-MPC and the generalization capability of the learned MPC parameters.
format Preprint
id arxiv_https___arxiv_org_abs_2312_11384
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle DiffTune-MPC: Closed-Loop Learning for Model Predictive Control
Tao, Ran
Cheng, Sheng
Wang, Xiaofeng
Wang, Shenlong
Hovakimyan, Naira
Robotics
Systems and Control
Model predictive control (MPC) has been applied to many platforms in robotics and autonomous systems for its capability to predict a system's future behavior while incorporating constraints that a system may have. To enhance the performance of a system with an MPC controller, one can manually tune the MPC's cost function. However, it can be challenging due to the possibly high dimension of the parameter space as well as the potential difference between the open-loop cost function in MPC and the overall closed-loop performance metric function. This paper presents DiffTune-MPC, a novel learning method, to learn the cost function of an MPC in a closed-loop manner. The proposed framework is compatible with the scenario where the time interval for performance evaluation and MPC's planning horizon have different lengths. We show the auxiliary problem whose solution admits the analytical gradients of MPC and discuss its variations in different MPC settings, including nonlinear MPCs that are solved using sequential quadratic programming. Simulation results demonstrate the learning capability of DiffTune-MPC and the generalization capability of the learned MPC parameters.
title DiffTune-MPC: Closed-Loop Learning for Model Predictive Control
topic Robotics
Systems and Control
url https://arxiv.org/abs/2312.11384