BAG: Body-Aligned 3D Wearable Asset Generation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Luo, Zhongjin, Li, Yang, Zhang, Mingrui, Wang, Senbo, Yan, Han, Song, Xibin, Shang, Taizhang, Mao, Wei, Li, Hongdong, Han, Xiaoguang, Ji, Pan
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915123930595328
author Luo, Zhongjin
Li, Yang
Zhang, Mingrui
Wang, Senbo
Yan, Han
Song, Xibin
Shang, Taizhang
Mao, Wei
Li, Hongdong
Han, Xiaoguang
Ji, Pan
author_facet Luo, Zhongjin
Li, Yang
Zhang, Mingrui
Wang, Senbo
Yan, Han
Song, Xibin
Shang, Taizhang
Mao, Wei
Li, Hongdong
Han, Xiaoguang
Ji, Pan
contents While recent advancements have shown remarkable progress in general 3D shape generation models, the challenge of leveraging these approaches to automatically generate wearable 3D assets remains unexplored. To this end, we present BAG, a Body-aligned Asset Generation method to output 3D wearable asset that can be automatically dressed on given 3D human bodies. This is achived by controlling the 3D generation process using human body shape and pose information. Specifically, we first build a general single-image to consistent multiview image diffusion model, and train it on the large Objaverse dataset to achieve diversity and generalizability. Then we train a Controlnet to guide the multiview generator to produce body-aligned multiview images. The control signal utilizes the multiview 2D projections of the target human body, where pixel values represent the XYZ coordinates of the body surface in a canonical space. The body-conditioned multiview diffusion generates body-aligned multiview images, which are then fed into a native 3D diffusion model to produce the 3D shape of the asset. Finally, by recovering the similarity transformation using multiview silhouette supervision and addressing asset-body penetration with physics simulators, the 3D asset can be accurately fitted onto the target human body. Experimental results demonstrate significant advantages over existing methods in terms of image prompt-following capability, shape diversity, and shape quality. Our project page is available at https://bag-3d.github.io/.
format Preprint
id arxiv_https___arxiv_org_abs_2501_16177
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle BAG: Body-Aligned 3D Wearable Asset Generation
Luo, Zhongjin
Li, Yang
Zhang, Mingrui
Wang, Senbo
Yan, Han
Song, Xibin
Shang, Taizhang
Mao, Wei
Li, Hongdong
Han, Xiaoguang
Ji, Pan
Computer Vision and Pattern Recognition
Artificial Intelligence
Graphics
While recent advancements have shown remarkable progress in general 3D shape generation models, the challenge of leveraging these approaches to automatically generate wearable 3D assets remains unexplored. To this end, we present BAG, a Body-aligned Asset Generation method to output 3D wearable asset that can be automatically dressed on given 3D human bodies. This is achived by controlling the 3D generation process using human body shape and pose information. Specifically, we first build a general single-image to consistent multiview image diffusion model, and train it on the large Objaverse dataset to achieve diversity and generalizability. Then we train a Controlnet to guide the multiview generator to produce body-aligned multiview images. The control signal utilizes the multiview 2D projections of the target human body, where pixel values represent the XYZ coordinates of the body surface in a canonical space. The body-conditioned multiview diffusion generates body-aligned multiview images, which are then fed into a native 3D diffusion model to produce the 3D shape of the asset. Finally, by recovering the similarity transformation using multiview silhouette supervision and addressing asset-body penetration with physics simulators, the 3D asset can be accurately fitted onto the target human body. Experimental results demonstrate significant advantages over existing methods in terms of image prompt-following capability, shape diversity, and shape quality. Our project page is available at https://bag-3d.github.io/.
title BAG: Body-Aligned 3D Wearable Asset Generation
topic Computer Vision and Pattern Recognition
Artificial Intelligence
Graphics
url https://arxiv.org/abs/2501.16177