Guardado en:
Detalles Bibliográficos
Autores principales: Chien, Hao-Jen, Huang, Yi-Chuan, Wu, Chung-Ho, Chao, Wei-Lun, Liu, Yu-Lun
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:https://arxiv.org/abs/2512.05113
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866911306027630592
author Chien, Hao-Jen
Huang, Yi-Chuan
Wu, Chung-Ho
Chao, Wei-Lun
Liu, Yu-Lun
author_facet Chien, Hao-Jen
Huang, Yi-Chuan
Wu, Chung-Ho
Chao, Wei-Lun
Liu, Yu-Lun
contents Synthesizing high-fidelity frozen 3D scenes from monocular Mannequin-Challenge (MC) videos is a unique problem distinct from standard dynamic scene reconstruction. Instead of focusing on modeling motion, our goal is to create a frozen scene while strategically preserving subtle dynamics to enable user-controlled instant selection. To achieve this, we introduce a novel application of dynamic Gaussian splatting: the scene is modeled dynamically, which retains nearby temporal variation, and a static scene is rendered by fixing the model's time parameter. However, under this usage, monocular capture with sparse temporal supervision introduces artifacts like ghosting and blur for Gaussians that become unobserved or occluded at weakly supervised timestamps. We propose Splannequin, an architecture-agnostic regularization that detects two states of Gaussian primitives, hidden and defective, and applies temporal anchoring. Under predominantly forward camera motion, hidden states are anchored to their recent well-observed past states, while defective states are anchored to future states with stronger supervision. Our method integrates into existing dynamic Gaussian pipelines via simple loss terms, requires no architectural changes, and adds zero inference overhead. This results in markedly improved visual quality, enabling high-fidelity, user-selectable frozen-time renderings, validated by a 96% user preference. Project page: https://chien90190.github.io/splannequin/
format Preprint
id arxiv_https___arxiv_org_abs_2512_05113
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Splannequin: Freezing Monocular Mannequin-Challenge Footage with Dual-Detection Splatting
Chien, Hao-Jen
Huang, Yi-Chuan
Wu, Chung-Ho
Chao, Wei-Lun
Liu, Yu-Lun
Computer Vision and Pattern Recognition
Synthesizing high-fidelity frozen 3D scenes from monocular Mannequin-Challenge (MC) videos is a unique problem distinct from standard dynamic scene reconstruction. Instead of focusing on modeling motion, our goal is to create a frozen scene while strategically preserving subtle dynamics to enable user-controlled instant selection. To achieve this, we introduce a novel application of dynamic Gaussian splatting: the scene is modeled dynamically, which retains nearby temporal variation, and a static scene is rendered by fixing the model's time parameter. However, under this usage, monocular capture with sparse temporal supervision introduces artifacts like ghosting and blur for Gaussians that become unobserved or occluded at weakly supervised timestamps. We propose Splannequin, an architecture-agnostic regularization that detects two states of Gaussian primitives, hidden and defective, and applies temporal anchoring. Under predominantly forward camera motion, hidden states are anchored to their recent well-observed past states, while defective states are anchored to future states with stronger supervision. Our method integrates into existing dynamic Gaussian pipelines via simple loss terms, requires no architectural changes, and adds zero inference overhead. This results in markedly improved visual quality, enabling high-fidelity, user-selectable frozen-time renderings, validated by a 96% user preference. Project page: https://chien90190.github.io/splannequin/
title Splannequin: Freezing Monocular Mannequin-Challenge Footage with Dual-Detection Splatting
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2512.05113