"Where am I?" Scene Retrieval with Language
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Jiaqi, Barath, Daniel, Armeni, Iro, Pollefeys, Marc, Blum, Hermann |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HouseTour: A Virtual Real Estate A(I)gent
by: Çelen, Ata, et al.
Published: (2025)
by: Çelen, Ata, et al.
Published: (2025)
Multiway Point Cloud Mosaicking with Diffusion and Global Optimization
by: Jin, Shengze, et al.
Published: (2024)
by: Jin, Shengze, et al.
Published: (2024)
Volumetric Semantically Consistent 3D Panoptic Mapping
by: Miao, Yang, et al.
Published: (2023)
by: Miao, Yang, et al.
Published: (2023)
CrossOver: 3D Scene Cross-Modal Alignment
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
by: Di Giammarino, Luca, et al.
Published: (2024)
by: Di Giammarino, Luca, et al.
Published: (2024)
SGAligner++: Cross-Modal Language-Aided 3D Scene Graph Alignment
by: Singh, Binod, et al.
Published: (2025)
by: Singh, Binod, et al.
Published: (2025)
Gravity-aligned Rotation Averaging with Circular Regression
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Learning to Make Keypoints Sub-Pixel Accurate
by: Kim, Shinjeong, et al.
Published: (2024)
by: Kim, Shinjeong, et al.
Published: (2024)
WildGS-SLAM: Monocular Gaussian Splatting SLAM in Dynamic Environments
by: Zheng, Jianhao, et al.
Published: (2025)
by: Zheng, Jianhao, et al.
Published: (2025)
Living Scenes: Multi-object Relocalization and Reconstruction in Changing 3D Environments
by: Zhu, Liyuan, et al.
Published: (2023)
by: Zhu, Liyuan, et al.
Published: (2023)
DepthSplat: Connecting Gaussian Splatting and Depth
by: Xu, Haofei, et al.
Published: (2024)
by: Xu, Haofei, et al.
Published: (2024)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
ReSpace: Text-Driven Autoregressive 3D Indoor Scene Synthesis and Editing
by: Bucher, Martin JJ., et al.
Published: (2025)
by: Bucher, Martin JJ., et al.
Published: (2025)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
by: Steiner, Emily, et al.
Published: (2026)
by: Steiner, Emily, et al.
Published: (2026)
ReSplat: Learning Recurrent Gaussian Splatting
by: Xu, Haofei, et al.
Published: (2025)
by: Xu, Haofei, et al.
Published: (2025)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
by: Siegel, Peter, et al.
Published: (2025)
by: Siegel, Peter, et al.
Published: (2025)
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
by: Sarkar, Sayan Deb, et al.
Published: (2026)
by: Sarkar, Sayan Deb, et al.
Published: (2026)
Lost & Found: Tracking Changes from Egocentric Observations in 3D Dynamic Scene Graphs
by: Behrens, Tjark, et al.
Published: (2024)
by: Behrens, Tjark, et al.
Published: (2024)
DROID-SLAM in the Wild
by: Li, Moyang, et al.
Published: (2026)
by: Li, Moyang, et al.
Published: (2026)
YoNoSplat: You Only Need One Model for Feedforward 3D Gaussian Splatting
by: Ye, Botao, et al.
Published: (2025)
by: Ye, Botao, et al.
Published: (2025)
Global Structure-from-Motion Revisited
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
ARKit LabelMaker: A New Scale for Indoor 3D Scene Understanding
by: Ji, Guangda, et al.
Published: (2024)
by: Ji, Guangda, et al.
Published: (2024)
Nothing Stands Still: A Spatiotemporal Benchmark on 3D Point Cloud Registration Under Large Geometric and Temporal Change
by: Sun, Tao, et al.
Published: (2023)
by: Sun, Tao, et al.
Published: (2023)
OpenFrontier: General Navigation with Visual-Language Grounded Frontiers
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025)
by: Lascheit, Kai, et al.
Published: (2025)
UnLoc: Leveraging Depth Uncertainties for Floorplan Localization
by: Wüest, Matthias, et al.
Published: (2025)
by: Wüest, Matthias, et al.
Published: (2025)
OVI-MAP:Open-Vocabulary Instance-Semantic Mapping
by: Deng, Zilong, et al.
Published: (2026)
by: Deng, Zilong, et al.
Published: (2026)
ReStyle3D: Scene-Level Appearance Transfer with Semantic Correspondences
by: Zhu, Liyuan, et al.
Published: (2025)
by: Zhu, Liyuan, et al.
Published: (2025)
WildPose: A Unified Framework for Robust Pose Estimation in the Wild
by: Zheng, Jianhao, et al.
Published: (2026)
by: Zheng, Jianhao, et al.
Published: (2026)
FunFact: Building Probabilistic Functional 3D Scene Graphs via Factor-Graph Reasoning
by: Fu, Zhengyu, et al.
Published: (2026)
by: Fu, Zhengyu, et al.
Published: (2026)
Active Visual Localization for Multi-Agent Collaboration: A Data-Driven Approach
by: Hanlon, Matthew, et al.
Published: (2023)
by: Hanlon, Matthew, et al.
Published: (2023)
REACT3D: Recovering Articulations for Interactive Physical 3D Scenes
by: Huang, Zhao, et al.
Published: (2025)
by: Huang, Zhao, et al.
Published: (2025)
OpenDAS: Open-Vocabulary Domain Adaptation for 2D and 3D Segmentation
by: Yilmaz, Gonca, et al.
Published: (2024)
by: Yilmaz, Gonca, et al.
Published: (2024)
Where am I? Cross-View Geo-localization with Natural Language Descriptions
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
Facade Segmentation for Solar Photovoltaic Suitability
by: Duran, Ayca, et al.
Published: (2025)
by: Duran, Ayca, et al.
Published: (2025)
VIGS-SLAM: Visual Inertial Gaussian Splatting SLAM
by: Zhu, Zihan, et al.
Published: (2025)
by: Zhu, Zihan, et al.
Published: (2025)
Retrieval Robust to Object Motion Blur
by: Zou, Rong, et al.
Published: (2024)
by: Zou, Rong, et al.
Published: (2024)
GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video Generator
by: Zhu, Liyuan, et al.
Published: (2026)
by: Zhu, Liyuan, et al.
Published: (2026)
FrontierNet: Learning Visual Cues to Explore
by: Sun, Boyang, et al.
Published: (2025)
by: Sun, Boyang, et al.
Published: (2025)
Similar Items
-
HouseTour: A Virtual Real Estate A(I)gent
by: Çelen, Ata, et al.
Published: (2025) -
Multiway Point Cloud Mosaicking with Diffusion and Global Optimization
by: Jin, Shengze, et al.
Published: (2024) -
Volumetric Semantically Consistent 3D Panoptic Mapping
by: Miao, Yang, et al.
Published: (2023) -
CrossOver: 3D Scene Cross-Modal Alignment
by: Sarkar, Sayan Deb, et al.
Published: (2025) -
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
by: Di Giammarino, Luca, et al.
Published: (2024)