SpatiaLoc: Leveraging Multi-Level Spatial Enhanced Descriptors for Cross-Modal Localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shang, Tianyi, Xu, Pengjie, Deng, Zhaojun, Li, Zhenyu, Chen, Zhicong, Wu, Lijun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vehicle-Scene Interaction: A Text-Driven 3D Lidar Place Recognition Method for Autonomous Driving
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
Place Recognition Meet Multiple Modalitie: A Comprehensive Review, Current Challenges and Future Directions
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms
von: Shang, Tianyi, et al.
Veröffentlicht: (2024)
von: Shang, Tianyi, et al.
Veröffentlicht: (2024)
A2GC: Asymmetric Aggregation with Geometric Constraints for Locally Aggregated Descriptors
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)
Spatia: Video Generation with Updatable Spatial Memory
von: Zhao, Jinjing, et al.
Veröffentlicht: (2025)
von: Zhao, Jinjing, et al.
Veröffentlicht: (2025)
OptiCorNet: Optimizing Sequence-Based Context Correlation for Visual Place Recognition
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Robot Localization Systems via Deep Feature Perturbation
von: Li, Zhenyu, et al.
Veröffentlicht: (2026)
von: Li, Zhenyu, et al.
Veröffentlicht: (2026)
UnLoc: Leveraging Depth Uncertainties for Floorplan Localization
von: Wüest, Matthias, et al.
Veröffentlicht: (2025)
von: Wüest, Matthias, et al.
Veröffentlicht: (2025)
TransLocNet: Cross-Modal Attention for Aerial-Ground Vehicle Localization with Contrastive Learning
von: Pham, Phu, et al.
Veröffentlicht: (2025)
von: Pham, Phu, et al.
Veröffentlicht: (2025)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
von: Miao, Yang, et al.
Veröffentlicht: (2024)
von: Miao, Yang, et al.
Veröffentlicht: (2024)
Loc$^2$: Interpretable Cross-View Localization via Depth-Lifted Local Feature Matching
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
Riemannian and Symplectic Geometry for Hierarchical Text-Driven Place Recognition
von: Shang, Tianyi, et al.
Veröffentlicht: (2026)
von: Shang, Tianyi, et al.
Veröffentlicht: (2026)
Leveraging Modality Tags for Enhanced Cross-Modal Video Retrieval
von: Fragomeni, Adriano, et al.
Veröffentlicht: (2025)
von: Fragomeni, Adriano, et al.
Veröffentlicht: (2025)
SpatiaLQA: A Benchmark for Evaluating Spatial Logical Reasoning in Vision-Language Models
von: Xie, Yuechen, et al.
Veröffentlicht: (2026)
von: Xie, Yuechen, et al.
Veröffentlicht: (2026)
CurriculumLoc: Enhancing Cross-Domain Geolocalization through Multi-Stage Refinement
von: Hu, Boni, et al.
Veröffentlicht: (2023)
von: Hu, Boni, et al.
Veröffentlicht: (2023)
GLDesigner: Leveraging Multi-Modal LLMs as Designer for Enhanced Aesthetic Text Glyph Layouts
von: He, Junwen, et al.
Veröffentlicht: (2024)
von: He, Junwen, et al.
Veröffentlicht: (2024)
DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
MSG-Loc: Multi-Label Likelihood-based Semantic Graph Matching for Object-Level Global Localization
von: Lee, Gihyeon, et al.
Veröffentlicht: (2025)
von: Lee, Gihyeon, et al.
Veröffentlicht: (2025)
MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization
von: Xiao, Zhendong, et al.
Veröffentlicht: (2025)
von: Xiao, Zhendong, et al.
Veröffentlicht: (2025)
360Loc: A Dataset and Benchmark for Omnidirectional Visual Localization with Cross-device Queries
von: Huang, Huajian, et al.
Veröffentlicht: (2023)
von: Huang, Huajian, et al.
Veröffentlicht: (2023)
Leveraging Multi-Modal Information to Enhance Dataset Distillation
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
Decoupled Cross-Modal Alignment Network for Text-RGBT Person Retrieval and A High-Quality Benchmark
von: Deng, Yifei, et al.
Veröffentlicht: (2025)
von: Deng, Yifei, et al.
Veröffentlicht: (2025)
GSplatLoc: Grounding Keypoint Descriptors into 3D Gaussian Splatting for Improved Visual Localization
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
AstroLoc: Robust Space to Ground Image Localizer
von: Berton, Gabriele, et al.
Veröffentlicht: (2025)
von: Berton, Gabriele, et al.
Veröffentlicht: (2025)
DiaLoc: An Iterative Approach to Embodied Dialog Localization
von: Zhang, Chao, et al.
Veröffentlicht: (2024)
von: Zhang, Chao, et al.
Veröffentlicht: (2024)
LunarLoc: Segment-Based Global Localization on the Moon
von: Thomas, Annika, et al.
Veröffentlicht: (2025)
von: Thomas, Annika, et al.
Veröffentlicht: (2025)
VFM-Loc: Zero-Shot Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies
von: Lu, Jun, et al.
Veröffentlicht: (2026)
von: Lu, Jun, et al.
Veröffentlicht: (2026)
MultiLoc: Multi-view Guided Relative Pose Regression for Fast and Robust Visual Re-Localization
von: Dang, Nobel, et al.
Veröffentlicht: (2026)
von: Dang, Nobel, et al.
Veröffentlicht: (2026)
GSplatLoc: Ultra-Precise Camera Localization via 3D Gaussian Splatting
von: Zeller, Atticus J., et al.
Veröffentlicht: (2024)
von: Zeller, Atticus J., et al.
Veröffentlicht: (2024)
SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2026)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2026)
UAV-VisLoc: A Large-scale Dataset for UAV Visual Localization
von: Xu, Wenjia, et al.
Veröffentlicht: (2024)
von: Xu, Wenjia, et al.
Veröffentlicht: (2024)
UniLoc: Towards Universal Place Recognition Using Any Single Modality
von: Xia, Yan, et al.
Veröffentlicht: (2024)
von: Xia, Yan, et al.
Veröffentlicht: (2024)
Game4Loc: A UAV Geo-Localization Benchmark from Game Data
von: Ji, Yuxiang, et al.
Veröffentlicht: (2024)
von: Ji, Yuxiang, et al.
Veröffentlicht: (2024)
Spatial-Frequency Enhanced Mamba for Multi-Modal Image Fusion
von: Sun, Hui, et al.
Veröffentlicht: (2025)
von: Sun, Hui, et al.
Veröffentlicht: (2025)
Semantic-Enhanced Cross-Modal Place Recognition for Robust Robot Localization
von: Lin, Yujia, et al.
Veröffentlicht: (2025)
von: Lin, Yujia, et al.
Veröffentlicht: (2025)
ImLoc: Revisiting Visual Localization with Image-based Representation
von: Jiang, Xudong, et al.
Veröffentlicht: (2026)
von: Jiang, Xudong, et al.
Veröffentlicht: (2026)
COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
von: Wu, Hao, et al.
Veröffentlicht: (2024)
von: Wu, Hao, et al.
Veröffentlicht: (2024)
Beyond Semantics: Uncovering the Physics of Fakes via Universal Physical Descriptors for Cross-Modal Synthetic Detection
von: Qiu, Mei, et al.
Veröffentlicht: (2026)
von: Qiu, Mei, et al.
Veröffentlicht: (2026)
CLIP-Loc: Multi-modal Landmark Association for Global Localization in Object-based Maps
von: Matsuzaki, Shigemichi, et al.
Veröffentlicht: (2024)
von: Matsuzaki, Shigemichi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Vehicle-Scene Interaction: A Text-Driven 3D Lidar Place Recognition Method for Autonomous Driving
von: Shang, Tianyi, et al.
Veröffentlicht: (2025) -
Place Recognition Meet Multiple Modalitie: A Comprehensive Review, Current Challenges and Future Directions
von: Li, Zhenyu, et al.
Veröffentlicht: (2025) -
MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms
von: Shang, Tianyi, et al.
Veröffentlicht: (2024) -
A2GC: Asymmetric Aggregation with Geometric Constraints for Locally Aggregated Descriptors
von: Li, Zhenyu, et al.
Veröffentlicht: (2025) -
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition
von: Shang, Tianyi, et al.
Veröffentlicht: (2025)