DrFER: Learning Disentangled Representations for 3D Facial Expression Recognition

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Li, Hebeizi, Yang, Hongyu, Huang, Di
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866929274807648256
author Li, Hebeizi
Yang, Hongyu
Huang, Di
author_facet Li, Hebeizi
Yang, Hongyu
Huang, Di
contents Facial Expression Recognition (FER) has consistently been a focal point in the field of facial analysis. In the context of existing methodologies for 3D FER or 2D+3D FER, the extraction of expression features often gets entangled with identity information, compromising the distinctiveness of these features. To tackle this challenge, we introduce the innovative DrFER method, which brings the concept of disentangled representation learning to the field of 3D FER. DrFER employs a dual-branch framework to effectively disentangle expression information from identity information. Diverging from prior disentanglement endeavors in the 3D facial domain, we have carefully reconfigured both the loss functions and network structure to make the overall framework adaptable to point cloud data. This adaptation enhances the capability of the framework in recognizing facial expressions, even in cases involving varying head poses. Extensive evaluations conducted on the BU-3DFE and Bosphorus datasets substantiate that DrFER surpasses the performance of other 3D FER methods.
format Preprint
id arxiv_https___arxiv_org_abs_2403_08318
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle DrFER: Learning Disentangled Representations for 3D Facial Expression Recognition
Li, Hebeizi
Yang, Hongyu
Huang, Di
Computer Vision and Pattern Recognition
Facial Expression Recognition (FER) has consistently been a focal point in the field of facial analysis. In the context of existing methodologies for 3D FER or 2D+3D FER, the extraction of expression features often gets entangled with identity information, compromising the distinctiveness of these features. To tackle this challenge, we introduce the innovative DrFER method, which brings the concept of disentangled representation learning to the field of 3D FER. DrFER employs a dual-branch framework to effectively disentangle expression information from identity information. Diverging from prior disentanglement endeavors in the 3D facial domain, we have carefully reconfigured both the loss functions and network structure to make the overall framework adaptable to point cloud data. This adaptation enhances the capability of the framework in recognizing facial expressions, even in cases involving varying head poses. Extensive evaluations conducted on the BU-3DFE and Bosphorus datasets substantiate that DrFER surpasses the performance of other 3D FER methods.
title DrFER: Learning Disentangled Representations for 3D Facial Expression Recognition
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2403.08318