CrossOver: 3D Scene Cross-Modal Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Sarkar, Sayan Deb, Miksik, Ondrej, Pollefeys, Marc, Barath, Daniel, Armeni, Iro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SGAligner++: Cross-Modal Language-Aided 3D Scene Graph Alignment
by: Singh, Binod, et al.
Published: (2025)
by: Singh, Binod, et al.
Published: (2025)
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
by: Sarkar, Sayan Deb, et al.
Published: (2026)
by: Sarkar, Sayan Deb, et al.
Published: (2026)
Volumetric Semantically Consistent 3D Panoptic Mapping
by: Miao, Yang, et al.
Published: (2023)
by: Miao, Yang, et al.
Published: (2023)
"Where am I?" Scene Retrieval with Language
by: Chen, Jiaqi, et al.
Published: (2024)
by: Chen, Jiaqi, et al.
Published: (2024)
Multiway Point Cloud Mosaicking with Diffusion and Global Optimization
by: Jin, Shengze, et al.
Published: (2024)
by: Jin, Shengze, et al.
Published: (2024)
HouseTour: A Virtual Real Estate A(I)gent
by: Çelen, Ata, et al.
Published: (2025)
by: Çelen, Ata, et al.
Published: (2025)
UnLoc: Leveraging Depth Uncertainties for Floorplan Localization
by: Wüest, Matthias, et al.
Published: (2025)
by: Wüest, Matthias, et al.
Published: (2025)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
Living Scenes: Multi-object Relocalization and Reconstruction in Changing 3D Environments
by: Zhu, Liyuan, et al.
Published: (2023)
by: Zhu, Liyuan, et al.
Published: (2023)
ReSpace: Text-Driven Autoregressive 3D Indoor Scene Synthesis and Editing
by: Bucher, Martin JJ., et al.
Published: (2025)
by: Bucher, Martin JJ., et al.
Published: (2025)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
by: Steiner, Emily, et al.
Published: (2026)
by: Steiner, Emily, et al.
Published: (2026)
Gravity-aligned Rotation Averaging with Circular Regression
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Learning to Make Keypoints Sub-Pixel Accurate
by: Kim, Shinjeong, et al.
Published: (2024)
by: Kim, Shinjeong, et al.
Published: (2024)
WildGS-SLAM: Monocular Gaussian Splatting SLAM in Dynamic Environments
by: Zheng, Jianhao, et al.
Published: (2025)
by: Zheng, Jianhao, et al.
Published: (2025)
Nothing Stands Still: A Spatiotemporal Benchmark on 3D Point Cloud Registration Under Large Geometric and Temporal Change
by: Sun, Tao, et al.
Published: (2023)
by: Sun, Tao, et al.
Published: (2023)
YoNoSplat: You Only Need One Model for Feedforward 3D Gaussian Splatting
by: Ye, Botao, et al.
Published: (2025)
by: Ye, Botao, et al.
Published: (2025)
ReStyle3D: Scene-Level Appearance Transfer with Semantic Correspondences
by: Zhu, Liyuan, et al.
Published: (2025)
by: Zhu, Liyuan, et al.
Published: (2025)
ReSplat: Learning Recurrent Gaussian Splatting
by: Xu, Haofei, et al.
Published: (2025)
by: Xu, Haofei, et al.
Published: (2025)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
by: Siegel, Peter, et al.
Published: (2025)
by: Siegel, Peter, et al.
Published: (2025)
DROID-SLAM in the Wild
by: Li, Moyang, et al.
Published: (2026)
by: Li, Moyang, et al.
Published: (2026)
Global Structure-from-Motion Revisited
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Cross-Modal Scene Semantic Alignment for Image Complexity Assessment
by: Luo, Yuqing, et al.
Published: (2025)
by: Luo, Yuqing, et al.
Published: (2025)
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025)
by: Lascheit, Kai, et al.
Published: (2025)
OVI-MAP:Open-Vocabulary Instance-Semantic Mapping
by: Deng, Zilong, et al.
Published: (2026)
by: Deng, Zilong, et al.
Published: (2026)
GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video Generator
by: Zhu, Liyuan, et al.
Published: (2026)
by: Zhu, Liyuan, et al.
Published: (2026)
FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
by: Delitzas, Alexandros, et al.
Published: (2026)
by: Delitzas, Alexandros, et al.
Published: (2026)
WildPose: A Unified Framework for Robust Pose Estimation in the Wild
by: Zheng, Jianhao, et al.
Published: (2026)
by: Zheng, Jianhao, et al.
Published: (2026)
LoopSplat: Loop Closure by Registering 3D Gaussian Splats
by: Zhu, Liyuan, et al.
Published: (2024)
by: Zhu, Liyuan, et al.
Published: (2024)
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
Sat2Scene: 3D Urban Scene Generation from Satellite Images with Diffusion
by: Li, Zuoyue, et al.
Published: (2024)
by: Li, Zuoyue, et al.
Published: (2024)
REACT3D: Recovering Articulations for Interactive Physical 3D Scenes
by: Huang, Zhao, et al.
Published: (2025)
by: Huang, Zhao, et al.
Published: (2025)
MAP-ADAPT: Real-Time Quality-Adaptive Semantic 3D Maps
by: Zheng, Jianhao, et al.
Published: (2024)
by: Zheng, Jianhao, et al.
Published: (2024)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
Facade Segmentation for Solar Photovoltaic Suitability
by: Duran, Ayca, et al.
Published: (2025)
by: Duran, Ayca, et al.
Published: (2025)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
by: Li, Jinlong, et al.
Published: (2025)
by: Li, Jinlong, et al.
Published: (2025)
Learning Modality Knowledge Alignment for Cross-Modality Transfer
by: Ma, Wenxuan, et al.
Published: (2024)
by: Ma, Wenxuan, et al.
Published: (2024)
VIGS-SLAM: Visual Inertial Gaussian Splatting SLAM
by: Zhu, Zihan, et al.
Published: (2025)
by: Zhu, Zihan, et al.
Published: (2025)
TORA: Topological Representation Alignment for 3D Shape Assembly
by: Lee, Nahyuk, et al.
Published: (2026)
by: Lee, Nahyuk, et al.
Published: (2026)
Similar Items
-
SGAligner++: Cross-Modal Language-Aided 3D Scene Graph Alignment
by: Singh, Binod, et al.
Published: (2025) -
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
by: Sarkar, Sayan Deb, et al.
Published: (2026) -
Volumetric Semantically Consistent 3D Panoptic Mapping
by: Miao, Yang, et al.
Published: (2023) -
"Where am I?" Scene Retrieval with Language
by: Chen, Jiaqi, et al.
Published: (2024) -
Multiway Point Cloud Mosaicking with Diffusion and Global Optimization
by: Jin, Shengze, et al.
Published: (2024)