Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Wiedemann, Nina, Liu, Sainan, Leboutet, Quentin, Gao, Katelyn, Ummenhofer, Benjamin, Paulitsch, Michael, Yuan, Kai
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866912567050371072
author Wiedemann, Nina
Liu, Sainan
Leboutet, Quentin
Gao, Katelyn
Ummenhofer, Benjamin
Paulitsch, Michael
Yuan, Kai
author_facet Wiedemann, Nina
Liu, Sainan
Leboutet, Quentin
Gao, Katelyn
Ummenhofer, Benjamin
Paulitsch, Michael
Yuan, Kai
contents Following rapid advancements in text and image generation, research has increasingly shifted towards 3D generation. Unlike the well-established pixel-based representation in images, 3D representations remain diverse and fragmented, encompassing a wide variety of approaches such as voxel grids, neural radiance fields, signed distance functions, point clouds, or octrees, each offering distinct advantages and limitations. In this work, we present a unified evaluation framework designed to assess the performance of 3D representations in reconstruction and generation. We compare these representations based on multiple criteria: quality, computational efficiency, and generalization performance. Beyond standard model benchmarking, our experiments aim to derive best practices over all steps involved in the 3D generation pipeline, including preprocessing, mesh reconstruction, compression with autoencoders, and generation. Our findings highlight that reconstruction errors significantly impact overall performance, underscoring the need to evaluate generation and reconstruction jointly. We provide insights that can inform the selection of suitable 3D models for various applications, facilitating the development of more robust and application-specific solutions in 3D generation. The code for our framework is available at https://github.com/isl-org/unifi3d.
format Preprint
id arxiv_https___arxiv_org_abs_2509_02474
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework
Wiedemann, Nina
Liu, Sainan
Leboutet, Quentin
Gao, Katelyn
Ummenhofer, Benjamin
Paulitsch, Michael
Yuan, Kai
Graphics
Computer Vision and Pattern Recognition
Machine Learning
Following rapid advancements in text and image generation, research has increasingly shifted towards 3D generation. Unlike the well-established pixel-based representation in images, 3D representations remain diverse and fragmented, encompassing a wide variety of approaches such as voxel grids, neural radiance fields, signed distance functions, point clouds, or octrees, each offering distinct advantages and limitations. In this work, we present a unified evaluation framework designed to assess the performance of 3D representations in reconstruction and generation. We compare these representations based on multiple criteria: quality, computational efficiency, and generalization performance. Beyond standard model benchmarking, our experiments aim to derive best practices over all steps involved in the 3D generation pipeline, including preprocessing, mesh reconstruction, compression with autoencoders, and generation. Our findings highlight that reconstruction errors significantly impact overall performance, underscoring the need to evaluate generation and reconstruction jointly. We provide insights that can inform the selection of suitable 3D models for various applications, facilitating the development of more robust and application-specific solutions in 3D generation. The code for our framework is available at https://github.com/isl-org/unifi3d.
title Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework
topic Graphics
Computer Vision and Pattern Recognition
Machine Learning
url https://arxiv.org/abs/2509.02474