Motion Manipulation via Unsupervised Keypoint Positioning in Face Animation

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Li, Hong, Liu, Boyu, Liu, Xuhui, Zhang, Baochang
Format: Preprint
Publié: 2026
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866918369951744000
author Li, Hong
Liu, Boyu
Liu, Xuhui
Zhang, Baochang
author_facet Li, Hong
Liu, Boyu
Liu, Xuhui
Zhang, Baochang
contents Face animation deals with controlling and generating facial features with a wide range of applications. The methods based on unsupervised keypoint positioning can produce realistic and detailed virtual portraits. However, they cannot achieve controllable face generation since the existing keypoint decomposition pipelines fail to fully decouple identity semantics and intertwined motion information (e.g., rotation, translation, and expression). To address these issues, we present a new method, Motion Manipulation via unsupervised keypoint positioning in Face Animation (MMFA). We first introduce self-supervised representation learning to encode and decode expressions in the latent feature space and decouple them from other motion information. Secondly, we propose a new way to compute keypoints aiming to achieve arbitrary motion control. Moreover, we design a variational autoencoder to map expression features to a continuous Gaussian distribution, allowing us for the first time to interpolate facial expressions in an unsupervised framework. We have conducted extensive experiments on publicly available datasets to validate the effectiveness of MMFA, which show that MMFA offers pronounced advantages over prior arts in creating realistic animation and manipulating face motion.
format Preprint
id arxiv_https___arxiv_org_abs_2603_04302
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Motion Manipulation via Unsupervised Keypoint Positioning in Face Animation
Li, Hong
Liu, Boyu
Liu, Xuhui
Zhang, Baochang
Computer Vision and Pattern Recognition
Face animation deals with controlling and generating facial features with a wide range of applications. The methods based on unsupervised keypoint positioning can produce realistic and detailed virtual portraits. However, they cannot achieve controllable face generation since the existing keypoint decomposition pipelines fail to fully decouple identity semantics and intertwined motion information (e.g., rotation, translation, and expression). To address these issues, we present a new method, Motion Manipulation via unsupervised keypoint positioning in Face Animation (MMFA). We first introduce self-supervised representation learning to encode and decode expressions in the latent feature space and decouple them from other motion information. Secondly, we propose a new way to compute keypoints aiming to achieve arbitrary motion control. Moreover, we design a variational autoencoder to map expression features to a continuous Gaussian distribution, allowing us for the first time to interpolate facial expressions in an unsupervised framework. We have conducted extensive experiments on publicly available datasets to validate the effectiveness of MMFA, which show that MMFA offers pronounced advantages over prior arts in creating realistic animation and manipulating face motion.
title Motion Manipulation via Unsupervised Keypoint Positioning in Face Animation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2603.04302