MoonSeg3R: Monocular Online Zero-Shot Segment Anything in 3D with Reconstructive Foundation Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Du, Zhipeng, Danier, Duolikun, Lenssen, Jan Eric, Bilen, Hakan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024)
by: Danier, Duolikun, et al.
Published: (2024)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025)
by: Danier, Duolikun, et al.
Published: (2025)
Enhancing Deformable Convolution based Video Frame Interpolation with Coarse-to-fine 3D CNN
by: Danier, Duolikun, et al.
Published: (2022)
by: Danier, Duolikun, et al.
Published: (2022)
OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging
by: Tang, Yijie, et al.
Published: (2025)
by: Tang, Yijie, et al.
Published: (2025)
BVI-VFI: A Video Quality Database for Video Frame Interpolation
by: Danier, Duolikun, et al.
Published: (2022)
by: Danier, Duolikun, et al.
Published: (2022)
ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame Interpolation
by: Danier, Duolikun, et al.
Published: (2021)
by: Danier, Duolikun, et al.
Published: (2021)
LDMVFI: Video Frame Interpolation with Latent Diffusion Models
by: Danier, Duolikun, et al.
Published: (2023)
by: Danier, Duolikun, et al.
Published: (2023)
A Subjective Quality Study for Video Frame Interpolation
by: Danier, Duolikun, et al.
Published: (2022)
by: Danier, Duolikun, et al.
Published: (2022)
Looking 3D: Anomaly Detection with 2D-3D Alignment
by: Bhunia, Ankan, et al.
Published: (2024)
by: Bhunia, Ankan, et al.
Published: (2024)
Neural Parametric Gaussians for Monocular Non-Rigid Object Reconstruction
by: Das, Devikalyan, et al.
Published: (2023)
by: Das, Devikalyan, et al.
Published: (2023)
Sp2360: Sparse-view 360 Scene Reconstruction using Cascaded 2D Diffusion Priors
by: Paul, Soumava, et al.
Published: (2024)
by: Paul, Soumava, et al.
Published: (2024)
Zero-Shot Surgical Tool Segmentation in Monocular Video Using Segment Anything Model 2
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Rewis3d: Reconstruction Improves Weakly-Supervised Semantic Segmentation
by: Ernst, Jonas, et al.
Published: (2026)
by: Ernst, Jonas, et al.
Published: (2026)
RefAM: Attention Magnets for Zero-Shot Referral Segmentation
by: Kukleva, Anna, et al.
Published: (2025)
by: Kukleva, Anna, et al.
Published: (2025)
Spurfies: Sparse Surface Reconstruction using Local Geometry Priors
by: Raj, Kevin, et al.
Published: (2024)
by: Raj, Kevin, et al.
Published: (2024)
SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model
by: Zhang, Dingyuan, et al.
Published: (2023)
by: Zhang, Dingyuan, et al.
Published: (2023)
SAM3D: Zero-Shot Semi-Automatic Segmentation in 3D Medical Images with the Segment Anything Model
by: Chan, Trevor J., et al.
Published: (2024)
by: Chan, Trevor J., et al.
Published: (2024)
latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
by: Wewer, Christopher, et al.
Published: (2024)
by: Wewer, Christopher, et al.
Published: (2024)
HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos
by: Xiong, Weitao, et al.
Published: (2025)
by: Xiong, Weitao, et al.
Published: (2025)
Enhancing Monocular 3D Hand Reconstruction with Learned Texture Priors
by: Karvounas, Giorgos, et al.
Published: (2025)
by: Karvounas, Giorgos, et al.
Published: (2025)
SING3R-SLAM: Submap-based Indoor Monocular Gaussian SLAM with 3D Reconstruction Priors
by: Li, Kunyi, et al.
Published: (2025)
by: Li, Kunyi, et al.
Published: (2025)
SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
by: Asim, Mohammad, et al.
Published: (2026)
by: Asim, Mohammad, et al.
Published: (2026)
SAM-LAD: Segment Anything Model Meets Zero-Shot Logic Anomaly Detection
by: Peng, Yun, et al.
Published: (2024)
by: Peng, Yun, et al.
Published: (2024)
Jamais Vu: Exposing the Generalization Gap in Supervised Semantic Correspondence
by: Mariotti, Octave, et al.
Published: (2025)
by: Mariotti, Octave, et al.
Published: (2025)
MaterialSeg3D: Segmenting Dense Materials from 2D Priors for 3D Assets
by: Li, Zeyu, et al.
Published: (2024)
by: Li, Zeyu, et al.
Published: (2024)
SegGraph: Leveraging Graphs of SAM Segments for Few-Shot 3D Part Segmentation
by: Hu, Yueyang, et al.
Published: (2025)
by: Hu, Yueyang, et al.
Published: (2025)
NTO3D: Neural Target Object 3D Reconstruction with Segment Anything
by: Wei, Xiaobao, et al.
Published: (2023)
by: Wei, Xiaobao, et al.
Published: (2023)
FMGS-Avatar: Mesh-Guided 2D Gaussian Splatting with Foundation Model Priors for 3D Monocular Avatar Reconstruction
by: Fan, Jinlong, et al.
Published: (2025)
by: Fan, Jinlong, et al.
Published: (2025)
SimNP: Learning Self-Similarity Priors Between Neural Points
by: Wewer, Christopher, et al.
Published: (2023)
by: Wewer, Christopher, et al.
Published: (2023)
HumMorph: Generalized Dynamic Human Neural Fields from Few Views
by: Zadrożny, Jakub, et al.
Published: (2025)
by: Zadrożny, Jakub, et al.
Published: (2025)
RenderDiffusion: Image Diffusion for 3D Reconstruction, Inpainting and Generation
by: Anciukevičius, Titas, et al.
Published: (2022)
by: Anciukevičius, Titas, et al.
Published: (2022)
Spatially-Adaptive Hash Encodings For Neural Surface Reconstruction
by: Walker, Thomas, et al.
Published: (2024)
by: Walker, Thomas, et al.
Published: (2024)
BVI-Artefact: An Artefact Detection Benchmark Dataset for Streamed Videos
by: Feng, Chen, et al.
Published: (2023)
by: Feng, Chen, et al.
Published: (2023)
Universal representations:The missing link between faces, text, planktons, and cat breeds
by: Bilen, Hakan, et al.
Published: (2017)
by: Bilen, Hakan, et al.
Published: (2017)
MapAnything: Evaluating Monocular Metric Depth Models for 3D Urban Asset Localization
by: Carnot, Miriam Louise, et al.
Published: (2025)
by: Carnot, Miriam Louise, et al.
Published: (2025)
Improving 2D Feature Representations by 3D-Aware Fine-Tuning
by: Yue, Yuanwen, et al.
Published: (2024)
by: Yue, Yuanwen, et al.
Published: (2024)
GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction
by: Lin, Weiquan, et al.
Published: (2026)
by: Lin, Weiquan, et al.
Published: (2026)
MoGA: 3D Generative Avatar Prior for Monocular Gaussian Avatar Reconstruction
by: Dong, Zijian, et al.
Published: (2025)
by: Dong, Zijian, et al.
Published: (2025)
MonoMobility: Zero-Shot 3D Mobility Analysis from Monocular Videos
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
GazePrior: Zero-Shot AR/VR Eye Tracking via Learned 3D Gaze Reconstruction
by: Dumery, Corentin, et al.
Published: (2026)
by: Dumery, Corentin, et al.
Published: (2026)
Similar Items
-
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024) -
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025) -
Enhancing Deformable Convolution based Video Frame Interpolation with Coarse-to-fine 3D CNN
by: Danier, Duolikun, et al.
Published: (2022) -
OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging
by: Tang, Yijie, et al.
Published: (2025) -
BVI-VFI: A Video Quality Database for Video Frame Interpolation
by: Danier, Duolikun, et al.
Published: (2022)