VFM-Recon: Unlocking Cross-Domain Scene-Level Neural Reconstruction with Scale-Aligned Foundation Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Ming, Yuhang, Xi, Tingkang, Yang, Xingrui, Yang, Lixin, Peng, Yong, Lu, Cewu, Kong, Wanzeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning
by: Ming, Yuhang, et al.
Published: (2024)
by: Ming, Yuhang, et al.
Published: (2024)
Uncertainty Aware Human-machine Collaboration in Camouflaged Object Detection
by: Yang, Ziyue, et al.
Published: (2025)
by: Yang, Ziyue, et al.
Published: (2025)
PhyRecon: Physically Plausible Neural Scene Reconstruction
by: Ni, Junfeng, et al.
Published: (2024)
by: Ni, Junfeng, et al.
Published: (2024)
Time-Frequency Jointed Imperceptible Adversarial Attack to Brainprint Recognition with Deep Learning Models
by: Yi, Hangjie, et al.
Published: (2024)
by: Yi, Hangjie, et al.
Published: (2024)
GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction
by: Schmid, Katharina, et al.
Published: (2026)
by: Schmid, Katharina, et al.
Published: (2026)
Keyframe-Based Feed-Forward Visual Odometry
by: Dai, Weichen, et al.
Published: (2026)
by: Dai, Weichen, et al.
Published: (2026)
3D Scene-Camera Representation with Joint Camera Photometric Optimization
by: Dai, Weichen, et al.
Published: (2025)
by: Dai, Weichen, et al.
Published: (2025)
MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition
by: Xu, Zhengyi, et al.
Published: (2026)
by: Xu, Zhengyi, et al.
Published: (2026)
CUS-GS: A Compact Unified Structured Gaussian Splatting Framework for Multimodal Scene Representation
by: Ming, Yuhang, et al.
Published: (2025)
by: Ming, Yuhang, et al.
Published: (2025)
CSSSTN: A Class-sensitive Subject-to-subject Semantic Style Transfer Network for EEG Classification in RSVP Tasks
by: Yang, Ziyue, et al.
Published: (2025)
by: Yang, Ziyue, et al.
Published: (2025)
SHARE: Scene-Human Aligned Reconstruction
by: Li, Joshua, et al.
Published: (2025)
by: Li, Joshua, et al.
Published: (2025)
SemGrasp: Semantic Grasp Generation via Language Aligned Discretization
by: Li, Kailin, et al.
Published: (2024)
by: Li, Kailin, et al.
Published: (2024)
BioVFM-21M: Benchmarking and Scaling Self-Supervised Vision Foundation Models for Biomedical Image Analysis
by: Liu, Jiarun, et al.
Published: (2025)
by: Liu, Jiarun, et al.
Published: (2025)
Brain Fingerprint Identification
by: Kong, Wanzeng, et al.
Published: (2025)
by: Kong, Wanzeng, et al.
Published: (2025)
Multi-view Hand Reconstruction with a Point-Embedded Transformer
by: Yang, Lixin, et al.
Published: (2024)
by: Yang, Lixin, et al.
Published: (2024)
SLC$^2$-SLAM: Semantic-guided Loop Closure using Shared Latent Code for NeRF SLAM
by: Ming, Yuhang, et al.
Published: (2025)
by: Ming, Yuhang, et al.
Published: (2025)
ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
by: Zhao, Guosheng, et al.
Published: (2025)
by: Zhao, Guosheng, et al.
Published: (2025)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
by: Ni, Chaojun, et al.
Published: (2024)
by: Ni, Chaojun, et al.
Published: (2024)
SimRecon: SimReady Compositional Scene Reconstruction from Real Videos
by: Xia, Chong, et al.
Published: (2026)
by: Xia, Chong, et al.
Published: (2026)
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory
by: Yang, Changchuan, et al.
Published: (2025)
by: Yang, Changchuan, et al.
Published: (2025)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
by: Wang, Xinkai, et al.
Published: (2026)
by: Wang, Xinkai, et al.
Published: (2026)
ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction
by: Ni, Chaojun, et al.
Published: (2025)
by: Ni, Chaojun, et al.
Published: (2025)
ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
by: Liu, Fangfu, et al.
Published: (2024)
by: Liu, Fangfu, et al.
Published: (2024)
VFM-Loc: Zero-Shot Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies
by: Lu, Jun, et al.
Published: (2026)
by: Lu, Jun, et al.
Published: (2026)
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
by: Feng, Chen-Bin, et al.
Published: (2026)
by: Feng, Chen-Bin, et al.
Published: (2026)
DC-VLAQ: Query-Residual Aggregation for Robust Visual Place Recognition
by: Zhu, Hanyu, et al.
Published: (2026)
by: Zhu, Hanyu, et al.
Published: (2026)
Benchmarking Neural Radiance Fields for Autonomous Robots: An Overview
by: Ming, Yuhang, et al.
Published: (2024)
by: Ming, Yuhang, et al.
Published: (2024)
Decompositional Neural Scene Reconstruction with Generative Diffusion Prior
by: Ni, Junfeng, et al.
Published: (2025)
by: Ni, Junfeng, et al.
Published: (2025)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
by: Ma, Yanbiao, et al.
Published: (2025)
by: Ma, Yanbiao, et al.
Published: (2025)
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction
by: Xiong, Kai, et al.
Published: (2026)
by: Xiong, Kai, et al.
Published: (2026)
Revisit Human-Scene Interaction via Space Occupancy
by: Liu, Xinpeng, et al.
Published: (2023)
by: Liu, Xinpeng, et al.
Published: (2023)
Vox-Fusion++: Voxel-based Neural Implicit Dense Tracking and Mapping with Multi-maps
by: Zhai, Hongjia, et al.
Published: (2024)
by: Zhai, Hongjia, et al.
Published: (2024)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
by: Chen, Yutian, et al.
Published: (2026)
by: Chen, Yutian, et al.
Published: (2026)
ReconDrive: Fast Feed-Forward 4D Gaussian Splatting for Autonomous Driving Scene Reconstruction
by: Yu, Haibao, et al.
Published: (2026)
by: Yu, Haibao, et al.
Published: (2026)
DressRecon: Freeform 4D Human Reconstruction from Monocular Video
by: Tan, Jeff, et al.
Published: (2024)
by: Tan, Jeff, et al.
Published: (2024)
O$^2$-Recon: Completing 3D Reconstruction of Occluded Objects in the Scene with a Pre-trained 2D Diffusion Model
by: Hu, Yubin, et al.
Published: (2023)
by: Hu, Yubin, et al.
Published: (2023)
Scene Coordinate Reconstruction Priors
by: Bian, Wenjing, et al.
Published: (2025)
by: Bian, Wenjing, et al.
Published: (2025)
DISCO: Embodied Navigation and Interaction via Differentiable Scene Semantics and Dual-level Control
by: Xu, Xinyu, et al.
Published: (2024)
by: Xu, Xinyu, et al.
Published: (2024)
DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval
by: Yang, Ruohong, et al.
Published: (2025)
by: Yang, Ruohong, et al.
Published: (2025)
Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
Similar Items
-
VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning
by: Ming, Yuhang, et al.
Published: (2024) -
Uncertainty Aware Human-machine Collaboration in Camouflaged Object Detection
by: Yang, Ziyue, et al.
Published: (2025) -
PhyRecon: Physically Plausible Neural Scene Reconstruction
by: Ni, Junfeng, et al.
Published: (2024) -
Time-Frequency Jointed Imperceptible Adversarial Attack to Brainprint Recognition with Deep Learning Models
by: Yi, Hangjie, et al.
Published: (2024) -
GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction
by: Schmid, Katharina, et al.
Published: (2026)