BEDLAM2.0: Synthetic Humans and Cameras in Motion

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Tesch, Joachim, Becherini, Giorgio, Achar, Prerana, Yiannakidis, Anastasios, Kocabas, Muhammed, Patel, Priyanka, Black, Michael J.
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866918208191070208
author Tesch, Joachim
Becherini, Giorgio
Achar, Prerana
Yiannakidis, Anastasios
Kocabas, Muhammed
Patel, Priyanka
Black, Michael J.
author_facet Tesch, Joachim
Becherini, Giorgio
Achar, Prerana
Yiannakidis, Anastasios
Kocabas, Muhammed
Patel, Priyanka
Black, Michael J.
contents Inferring 3D human motion from video remains a challenging problem with many applications. While traditional methods estimate the human in image coordinates, many applications require human motion to be estimated in world coordinates. This is particularly challenging when there is both human and camera motion. Progress on this topic has been limited by the lack of rich video data with ground truth human and camera movement. We address this with BEDLAM2.0, a new dataset that goes beyond the popular BEDLAM dataset in important ways. In addition to introducing more diverse and realistic cameras and camera motions, BEDLAM2.0 increases diversity and realism of body shape, motions, clothing, hair, and 3D environments. Additionally, it adds shoes, which were missing in BEDLAM. BEDLAM has become a key resource for training 3D human pose and motion regressors today and we show that BEDLAM2.0 is significantly better, particularly for training methods that estimate humans in world coordinates. We compare state-of-the art methods trained on BEDLAM and BEDLAM2.0, and find that BEDLAM2.0 significantly improves accuracy over BEDLAM. For research purposes, we provide the rendered videos, ground truth body parameters, and camera motions. We also provide the 3D assets to which we have rights and links to those from third parties.
format Preprint
id arxiv_https___arxiv_org_abs_2511_14394
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle BEDLAM2.0: Synthetic Humans and Cameras in Motion
Tesch, Joachim
Becherini, Giorgio
Achar, Prerana
Yiannakidis, Anastasios
Kocabas, Muhammed
Patel, Priyanka
Black, Michael J.
Computer Vision and Pattern Recognition
Inferring 3D human motion from video remains a challenging problem with many applications. While traditional methods estimate the human in image coordinates, many applications require human motion to be estimated in world coordinates. This is particularly challenging when there is both human and camera motion. Progress on this topic has been limited by the lack of rich video data with ground truth human and camera movement. We address this with BEDLAM2.0, a new dataset that goes beyond the popular BEDLAM dataset in important ways. In addition to introducing more diverse and realistic cameras and camera motions, BEDLAM2.0 increases diversity and realism of body shape, motions, clothing, hair, and 3D environments. Additionally, it adds shoes, which were missing in BEDLAM. BEDLAM has become a key resource for training 3D human pose and motion regressors today and we show that BEDLAM2.0 is significantly better, particularly for training methods that estimate humans in world coordinates. We compare state-of-the art methods trained on BEDLAM and BEDLAM2.0, and find that BEDLAM2.0 significantly improves accuracy over BEDLAM. For research purposes, we provide the rendered videos, ground truth body parameters, and camera motions. We also provide the 3D assets to which we have rights and links to those from third parties.
title BEDLAM2.0: Synthetic Humans and Cameras in Motion
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2511.14394