VoxelKeypointFusion: Generalizable Multi-View Multi-Person Pose Estimation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bermuth, Daniel, Poeppel, Alexander, Reif, Wolfgang
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910755172909056
author Bermuth, Daniel
Poeppel, Alexander
Reif, Wolfgang
author_facet Bermuth, Daniel
Poeppel, Alexander
Reif, Wolfgang
contents In the rapidly evolving field of computer vision, the task of accurately estimating the poses of multiple individuals from various viewpoints presents a formidable challenge, especially if the estimations should be reliable as well. This work presents an extensive evaluation of the generalization capabilities of multi-view multi-person pose estimators to unseen datasets and presents a new algorithm with strong performance in this task. It also studies the improvements by additionally using depth information. Since the new approach can not only generalize well to unseen datasets, but also to different keypoints, the first multi-view multi-person whole-body estimator is presented. To support further research on those topics, all of the work is publicly accessible.
format Preprint
id arxiv_https___arxiv_org_abs_2410_18723
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle VoxelKeypointFusion: Generalizable Multi-View Multi-Person Pose Estimation
Bermuth, Daniel
Poeppel, Alexander
Reif, Wolfgang
Computer Vision and Pattern Recognition
Human-Computer Interaction
In the rapidly evolving field of computer vision, the task of accurately estimating the poses of multiple individuals from various viewpoints presents a formidable challenge, especially if the estimations should be reliable as well. This work presents an extensive evaluation of the generalization capabilities of multi-view multi-person pose estimators to unseen datasets and presents a new algorithm with strong performance in this task. It also studies the improvements by additionally using depth information. Since the new approach can not only generalize well to unseen datasets, but also to different keypoints, the first multi-view multi-person whole-body estimator is presented. To support further research on those topics, all of the work is publicly accessible.
title VoxelKeypointFusion: Generalizable Multi-View Multi-Person Pose Estimation
topic Computer Vision and Pattern Recognition
Human-Computer Interaction
url https://arxiv.org/abs/2410.18723