Joint Optimization of Neural Radiance Fields and Continuous Camera Motion from a Monocular Video

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Nguyen, Hoang Chuong, Mao, Wei, Alvarez, Jose M., Liu, Miaomiao
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918002179440640
author Nguyen, Hoang Chuong
Mao, Wei
Alvarez, Jose M.
Liu, Miaomiao
author_facet Nguyen, Hoang Chuong
Mao, Wei
Alvarez, Jose M.
Liu, Miaomiao
contents Neural Radiance Fields (NeRF) has demonstrated its superior capability to represent 3D geometry but require accurately precomputed camera poses during training. To mitigate this requirement, existing methods jointly optimize camera poses and NeRF often relying on good pose initialisation or depth priors. However, these approaches struggle in challenging scenarios, such as large rotations, as they map each camera to a world coordinate system. We propose a novel method that eliminates prior dependencies by modeling continuous camera motions as time-dependent angular velocity and velocity. Relative motions between cameras are learned first via velocity integration, while camera poses can be obtained by aggregating such relative motions up to a world coordinate system defined at a single time step within the video. Specifically, accurate continuous camera movements are learned through a time-dependent NeRF, which captures local scene geometry and motion by training from neighboring frames for each time step. The learned motions enable fine-tuning the NeRF to represent the full scene geometry. Experiments on Co3D and Scannet show our approach achieves superior camera pose and depth estimation and comparable novel-view synthesis performance compared to state-of-the-art methods. Our code is available at https://github.com/HoangChuongNguyen/cope-nerf.
format Preprint
id arxiv_https___arxiv_org_abs_2504_19819
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Joint Optimization of Neural Radiance Fields and Continuous Camera Motion from a Monocular Video
Nguyen, Hoang Chuong
Mao, Wei
Alvarez, Jose M.
Liu, Miaomiao
Computer Vision and Pattern Recognition
Neural Radiance Fields (NeRF) has demonstrated its superior capability to represent 3D geometry but require accurately precomputed camera poses during training. To mitigate this requirement, existing methods jointly optimize camera poses and NeRF often relying on good pose initialisation or depth priors. However, these approaches struggle in challenging scenarios, such as large rotations, as they map each camera to a world coordinate system. We propose a novel method that eliminates prior dependencies by modeling continuous camera motions as time-dependent angular velocity and velocity. Relative motions between cameras are learned first via velocity integration, while camera poses can be obtained by aggregating such relative motions up to a world coordinate system defined at a single time step within the video. Specifically, accurate continuous camera movements are learned through a time-dependent NeRF, which captures local scene geometry and motion by training from neighboring frames for each time step. The learned motions enable fine-tuning the NeRF to represent the full scene geometry. Experiments on Co3D and Scannet show our approach achieves superior camera pose and depth estimation and comparable novel-view synthesis performance compared to state-of-the-art methods. Our code is available at https://github.com/HoangChuongNguyen/cope-nerf.
title Joint Optimization of Neural Radiance Fields and Continuous Camera Motion from a Monocular Video
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2504.19819