LEIA: Latent View-invariant Embeddings for Implicit 3D Articulation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Swaminathan, Archana, Gupta, Anubhav, Gupta, Kamal, Maiya, Shishira R., Agarwal, Vatsal, Shrivastava, Abhinav
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912021292777472
author Swaminathan, Archana
Gupta, Anubhav
Gupta, Kamal
Maiya, Shishira R.
Agarwal, Vatsal
Shrivastava, Abhinav
author_facet Swaminathan, Archana
Gupta, Anubhav
Gupta, Kamal
Maiya, Shishira R.
Agarwal, Vatsal
Shrivastava, Abhinav
contents Neural Radiance Fields (NeRFs) have revolutionized the reconstruction of static scenes and objects in 3D, offering unprecedented quality. However, extending NeRFs to model dynamic objects or object articulations remains a challenging problem. Previous works have tackled this issue by focusing on part-level reconstruction and motion estimation for objects, but they often rely on heuristics regarding the number of moving parts or object categories, which can limit their practical use. In this work, we introduce LEIA, a novel approach for representing dynamic 3D objects. Our method involves observing the object at distinct time steps or "states" and conditioning a hypernetwork on the current state, using this to parameterize our NeRF. This approach allows us to learn a view-invariant latent representation for each state. We further demonstrate that by interpolating between these states, we can generate novel articulation configurations in 3D space that were previously unseen. Our experimental results highlight the effectiveness of our method in articulating objects in a manner that is independent of the viewing angle and joint configuration. Notably, our approach outperforms previous methods that rely on motion information for articulation registration.
format Preprint
id arxiv_https___arxiv_org_abs_2409_06703
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle LEIA: Latent View-invariant Embeddings for Implicit 3D Articulation
Swaminathan, Archana
Gupta, Anubhav
Gupta, Kamal
Maiya, Shishira R.
Agarwal, Vatsal
Shrivastava, Abhinav
Computer Vision and Pattern Recognition
Neural Radiance Fields (NeRFs) have revolutionized the reconstruction of static scenes and objects in 3D, offering unprecedented quality. However, extending NeRFs to model dynamic objects or object articulations remains a challenging problem. Previous works have tackled this issue by focusing on part-level reconstruction and motion estimation for objects, but they often rely on heuristics regarding the number of moving parts or object categories, which can limit their practical use. In this work, we introduce LEIA, a novel approach for representing dynamic 3D objects. Our method involves observing the object at distinct time steps or "states" and conditioning a hypernetwork on the current state, using this to parameterize our NeRF. This approach allows us to learn a view-invariant latent representation for each state. We further demonstrate that by interpolating between these states, we can generate novel articulation configurations in 3D space that were previously unseen. Our experimental results highlight the effectiveness of our method in articulating objects in a manner that is independent of the viewing angle and joint configuration. Notably, our approach outperforms previous methods that rely on motion information for articulation registration.
title LEIA: Latent View-invariant Embeddings for Implicit 3D Articulation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2409.06703