EvolvingGS: High-Fidelity Streamable Volumetric Video via Evolving 3D Gaussian Representation

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Zhang, Chao, Zhou, Yifeng, Wang, Shuheng, Li, Wenfa, Wang, Degang, Xu, Yi, Jiao, Shaohui
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866915185165336576
author Zhang, Chao
Zhou, Yifeng
Wang, Shuheng
Li, Wenfa
Wang, Degang
Xu, Yi
Jiao, Shaohui
author_facet Zhang, Chao
Zhou, Yifeng
Wang, Shuheng
Li, Wenfa
Wang, Degang
Xu, Yi
Jiao, Shaohui
contents We have recently seen great progress in 3D scene reconstruction through explicit point-based 3D Gaussian Splatting (3DGS), notable for its high quality and fast rendering speed. However, reconstructing dynamic scenes such as complex human performances with long durations remains challenging. Prior efforts fall short of modeling a long-term sequence with drastic motions, frequent topology changes or interactions with props, and resort to segmenting the whole sequence into groups of frames that are processed independently, which undermines temporal stability and thereby leads to an unpleasant viewing experience and inefficient storage footprint. In view of this, we introduce EvolvingGS, a two-stage strategy that first deforms the Gaussian model to coarsely align with the target frame, and then refines it with minimal point addition/subtraction, particularly in fast-changing areas. Owing to the flexibility of the incrementally evolving representation, our method outperforms existing approaches in terms of both per-frame and temporal quality metrics while maintaining fast rendering through its purely explicit representation. Moreover, by exploiting temporal coherence between successive frames, we propose a simple yet effective compression algorithm that achieves over 50x compression rate. Extensive experiments on both public benchmarks and challenging custom datasets demonstrate that our method significantly advances the state-of-the-art in dynamic scene reconstruction, particularly for extended sequences with complex human performances.
format Preprint
id arxiv_https___arxiv_org_abs_2503_05162
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle EvolvingGS: High-Fidelity Streamable Volumetric Video via Evolving 3D Gaussian Representation
Zhang, Chao
Zhou, Yifeng
Wang, Shuheng
Li, Wenfa
Wang, Degang
Xu, Yi
Jiao, Shaohui
Computer Vision and Pattern Recognition
We have recently seen great progress in 3D scene reconstruction through explicit point-based 3D Gaussian Splatting (3DGS), notable for its high quality and fast rendering speed. However, reconstructing dynamic scenes such as complex human performances with long durations remains challenging. Prior efforts fall short of modeling a long-term sequence with drastic motions, frequent topology changes or interactions with props, and resort to segmenting the whole sequence into groups of frames that are processed independently, which undermines temporal stability and thereby leads to an unpleasant viewing experience and inefficient storage footprint. In view of this, we introduce EvolvingGS, a two-stage strategy that first deforms the Gaussian model to coarsely align with the target frame, and then refines it with minimal point addition/subtraction, particularly in fast-changing areas. Owing to the flexibility of the incrementally evolving representation, our method outperforms existing approaches in terms of both per-frame and temporal quality metrics while maintaining fast rendering through its purely explicit representation. Moreover, by exploiting temporal coherence between successive frames, we propose a simple yet effective compression algorithm that achieves over 50x compression rate. Extensive experiments on both public benchmarks and challenging custom datasets demonstrate that our method significantly advances the state-of-the-art in dynamic scene reconstruction, particularly for extended sequences with complex human performances.
title EvolvingGS: High-Fidelity Streamable Volumetric Video via Evolving 3D Gaussian Representation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2503.05162