Generalized Model Predictive Path Integral Control as Expectation--Maximization

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Jiarui, Sharifi, Sina, Fazlyab, Mahyar
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911734953934848
author Wang, Jiarui
Sharifi, Sina
Fazlyab, Mahyar
author_facet Wang, Jiarui
Sharifi, Sina
Fazlyab, Mahyar
contents Model Predictive Path Integral (MPPI) control is a powerful sampling-based method for solving stochastic optimal control problems and has enabled real-time control in complex robotic systems. Despite its empirical success, its theoretical understanding remains limited. In this work, we show that MPPI can be interpreted as a special case of the Expectation-Maximization (EM) algorithm applied to a probabilistic inference formulation of optimal control. This perspective leads to a generalized EM-MPPI framework that extends MPPI beyond the commonly used Gaussian parameterization. We analyze the convergence behavior of this algorithm and characterize the local convergence rate in terms of the covariance of the posterior trajectory distribution and the exploration distribution. For exponential-family distributions, we establish a sufficient increase property of the log-likelihood when the log-partition function is strongly convex. Specializing the analysis to Gaussian MPPI yields explicit global and local convergence characterizations. The code for the experiments will be available upon acceptance.
format Preprint
id arxiv_https___arxiv_org_abs_2606_00317
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Generalized Model Predictive Path Integral Control as Expectation--Maximization
Wang, Jiarui
Sharifi, Sina
Fazlyab, Mahyar
Systems and Control
Optimization and Control
Model Predictive Path Integral (MPPI) control is a powerful sampling-based method for solving stochastic optimal control problems and has enabled real-time control in complex robotic systems. Despite its empirical success, its theoretical understanding remains limited. In this work, we show that MPPI can be interpreted as a special case of the Expectation-Maximization (EM) algorithm applied to a probabilistic inference formulation of optimal control. This perspective leads to a generalized EM-MPPI framework that extends MPPI beyond the commonly used Gaussian parameterization. We analyze the convergence behavior of this algorithm and characterize the local convergence rate in terms of the covariance of the posterior trajectory distribution and the exploration distribution. For exponential-family distributions, we establish a sufficient increase property of the log-likelihood when the log-partition function is strongly convex. Specializing the analysis to Gaussian MPPI yields explicit global and local convergence characterizations. The code for the experiments will be available upon acceptance.
title Generalized Model Predictive Path Integral Control as Expectation--Maximization
topic Systems and Control
Optimization and Control
url https://arxiv.org/abs/2606.00317