Image-Plane Geometric Decoding for View-Invariant Indoor Scene Reconstruction

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Li, Mingyang, Fan, Yimeng, Liu, Changsong, Xu, Lixue, Wang, Xin, Liu, Yanyan, Zhang, Wei
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908613944016896
author Li, Mingyang
Fan, Yimeng
Liu, Changsong
Xu, Lixue
Wang, Xin
Liu, Yanyan
Zhang, Wei
author_facet Li, Mingyang
Fan, Yimeng
Liu, Changsong
Xu, Lixue
Wang, Xin
Liu, Yanyan
Zhang, Wei
contents Volume-based indoor scene reconstruction methods offer superior generalization capability and real-time deployment potential. However, existing methods rely on multi-view pixel back-projection ray intersections as weak geometric constraints to determine spatial positions. This dependence results in reconstruction quality being heavily influenced by input view density. Performance degrades in overlapping regions and unobserved areas.To address these limitations, we reduce dependency on inter-view geometric constraints by exploiting spatial information within individual views. We propose an image-plane decoding framework with three core components: Pixel-level Confidence Encoder, Affine Compensation Module, and Image-Plane Spatial Decoder. These modules decode three-dimensional structural information encoded in images through physical imaging processes. The framework effectively preserves spatial geometric features including edges, hollow structures, and complex textures. It significantly enhances view-invariant reconstruction.Experiments on indoor scene reconstruction datasets confirm superior reconstruction stability. Our method maintains nearly identical quality when view count reduces by 40%. It achieves a coefficient of variation of 0.24%, performance retention rate of 99.7%, and maximum performance drop of 0.42%. These results demonstrate that exploiting intra-view spatial information provides a robust solution for view-limited scenarios in practical applications.
format Preprint
id arxiv_https___arxiv_org_abs_2509_25744
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Image-Plane Geometric Decoding for View-Invariant Indoor Scene Reconstruction
Li, Mingyang
Fan, Yimeng
Liu, Changsong
Xu, Lixue
Wang, Xin
Liu, Yanyan
Zhang, Wei
Computer Vision and Pattern Recognition
Volume-based indoor scene reconstruction methods offer superior generalization capability and real-time deployment potential. However, existing methods rely on multi-view pixel back-projection ray intersections as weak geometric constraints to determine spatial positions. This dependence results in reconstruction quality being heavily influenced by input view density. Performance degrades in overlapping regions and unobserved areas.To address these limitations, we reduce dependency on inter-view geometric constraints by exploiting spatial information within individual views. We propose an image-plane decoding framework with three core components: Pixel-level Confidence Encoder, Affine Compensation Module, and Image-Plane Spatial Decoder. These modules decode three-dimensional structural information encoded in images through physical imaging processes. The framework effectively preserves spatial geometric features including edges, hollow structures, and complex textures. It significantly enhances view-invariant reconstruction.Experiments on indoor scene reconstruction datasets confirm superior reconstruction stability. Our method maintains nearly identical quality when view count reduces by 40%. It achieves a coefficient of variation of 0.24%, performance retention rate of 99.7%, and maximum performance drop of 0.42%. These results demonstrate that exploiting intra-view spatial information provides a robust solution for view-limited scenarios in practical applications.
title Image-Plane Geometric Decoding for View-Invariant Indoor Scene Reconstruction
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2509.25744