Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Bo, Lai, Qiuxia, Sun, Zeren, Shu, Xiangbo, Yao, Yazhou, Wang, Wenguan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Poly Kernel Inception Network for Remote Sensing Detection
by: Cai, Xinhao, et al.
Published: (2024)
by: Cai, Xinhao, et al.
Published: (2024)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Learning Clustering-based Prototypes for Compositional Zero-shot Learning
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Large Spatial Model: End-to-end Unposed Images to Semantic 3D
by: Fan, Zhiwen, et al.
Published: (2024)
by: Fan, Zhiwen, et al.
Published: (2024)
SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
by: Sheng, Yu, et al.
Published: (2025)
by: Sheng, Yu, et al.
Published: (2025)
Uni3R: Unified 3D Reconstruction and Semantic Understanding via Generalizable Gaussian Splatting from Unposed Multi-View Images
by: Sun, Xiangyu, et al.
Published: (2025)
by: Sun, Xiangyu, et al.
Published: (2025)
FTMoMamba: Motion Generation with Frequency and Text State Space Models
by: Li, Chengjian, et al.
Published: (2024)
by: Li, Chengjian, et al.
Published: (2024)
Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with Minimal 3D Knowledge
by: Wang, Haoru, et al.
Published: (2025)
by: Wang, Haoru, et al.
Published: (2025)
Learning 3D-Aware GANs from Unposed Images with Template Feature Field
by: Chen, Xinya, et al.
Published: (2024)
by: Chen, Xinya, et al.
Published: (2024)
A Conditional Probability Framework for Compositional Zero-shot Learning
by: Wu, Peng, et al.
Published: (2025)
by: Wu, Peng, et al.
Published: (2025)
Unposed-to-3D: Learning Simulation-Ready Vehicles from Real-World Images
by: Liu, Hongyuan, et al.
Published: (2026)
by: Liu, Hongyuan, et al.
Published: (2026)
LucidFusion: Reconstructing 3D Gaussians with Arbitrary Unposed Images
by: He, Hao, et al.
Published: (2024)
by: He, Hao, et al.
Published: (2024)
AdaFPP: Adapt-Focused Bi-Propagating Prototype Learning for Panoramic Activity Recognition
by: Cao, Meiqi, et al.
Published: (2024)
by: Cao, Meiqi, et al.
Published: (2024)
Spatio-temporal Decoupled Knowledge Compensator for Few-Shot Action Recognition
by: Qu, Hongyu, et al.
Published: (2026)
by: Qu, Hongyu, et al.
Published: (2026)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
by: Pei, Gensheng, et al.
Published: (2025)
by: Pei, Gensheng, et al.
Published: (2025)
RegGS: Unposed Sparse Views Gaussian Splatting with 3DGS Registration
by: Cheng, Chong, et al.
Published: (2025)
by: Cheng, Chong, et al.
Published: (2025)
Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
by: Cai, Xinhao, et al.
Published: (2025)
by: Cai, Xinhao, et al.
Published: (2025)
Information Bottleneck Approach to Spatial Attention Learning
by: Lai, Qiuxia, et al.
Published: (2021)
by: Lai, Qiuxia, et al.
Published: (2021)
Pseudo-View Enhancement via Confidence Fusion for Unposed Sparse-View Reconstruction
by: Zhao, Beizhen, et al.
Published: (2026)
by: Zhao, Beizhen, et al.
Published: (2026)
On-the-fly Reconstruction for Large-Scale Novel View Synthesis from Unposed Images
by: Meuleman, Andreas, et al.
Published: (2025)
by: Meuleman, Andreas, et al.
Published: (2025)
NVComposer: Boosting Generative Novel View Synthesis with Multiple Sparse and Unposed Images
by: Li, Lingen, et al.
Published: (2024)
by: Li, Lingen, et al.
Published: (2024)
Foster Adaptivity and Balance in Learning with Noisy Labels
by: Sheng, Mengmeng, et al.
Published: (2024)
by: Sheng, Mengmeng, et al.
Published: (2024)
Unposed 3DGS Reconstruction with Probabilistic Procrustes Mapping
by: Cheng, Chong, et al.
Published: (2025)
by: Cheng, Chong, et al.
Published: (2025)
No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos
by: Balice, Matteo, et al.
Published: (2026)
by: Balice, Matteo, et al.
Published: (2026)
UniSem: Generalizable Semantic 3D Reconstruction from Sparse Unposed Images
by: Liao, Guibiao, et al.
Published: (2026)
by: Liao, Guibiao, et al.
Published: (2026)
COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
by: Sun, Shaorong, et al.
Published: (2024)
by: Sun, Shaorong, et al.
Published: (2024)
UFO-4D: Unposed Feedforward 4D Reconstruction from Two Images
by: Hur, Junhwa, et al.
Published: (2026)
by: Hur, Junhwa, et al.
Published: (2026)
ZeroGS: Training 3D Gaussian Splatting from Unposed Images
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
UpFusion: Novel View Diffusion from Unposed Sparse View Observations
by: Kani, Bharath Raj Nagoor, et al.
Published: (2023)
by: Kani, Bharath Raj Nagoor, et al.
Published: (2023)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
DGGT: Feedforward 4D Reconstruction of Dynamic Driving Scenes using Unposed Images
by: Chen, Xiaoxue, et al.
Published: (2025)
by: Chen, Xiaoxue, et al.
Published: (2025)
Shape2Scene: 3D Scene Representation Learning Through Pre-training on Shape Data
by: Feng, Tuo, et al.
Published: (2024)
by: Feng, Tuo, et al.
Published: (2024)
Combating Noisy Labels through Fostering Self- and Neighbor-Consistency
by: Sun, Zeren, et al.
Published: (2026)
by: Sun, Zeren, et al.
Published: (2026)
VideoMAC: Video Masked Autoencoders Meet ConvNets
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Neural Clustering based Visual Representation Learning
by: Chen, Guikun, et al.
Published: (2024)
by: Chen, Guikun, et al.
Published: (2024)
Pragmatist: Multiview Conditional Diffusion Models for High-Fidelity 3D Reconstruction from Unposed Sparse Views
by: Zhang, Songchun, et al.
Published: (2024)
by: Zhang, Songchun, et al.
Published: (2024)
Similar Items
-
Poly Kernel Inception Network for Remote Sensing Detection
by: Cai, Xinhao, et al.
Published: (2024) -
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025) -
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026) -
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
by: Cai, Xinhao, et al.
Published: (2026) -
Learning Clustering-based Prototypes for Compositional Zero-shot Learning
by: Qu, Hongyu, et al.
Published: (2025)