MMS-VPR: Multimodal Street-Level Visual Place Recognition Dataset and Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Ou, Yiwei, Ren, Xiaobin, Sun, Ronggui, Gao, Guansong, Zhao, Kaiqi, Manfredini, Manfredo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Urban-ImageNet: A Large-Scale Multi-Modal Dataset and Evaluation Framework for Urban Space Perception
by: Ou, Yiwei, et al.
Published: (2026)
by: Ou, Yiwei, et al.
Published: (2026)
LaVPR: Benchmarking Language and Vision for Place Recognition
by: Idan, Ofer, et al.
Published: (2026)
by: Idan, Ofer, et al.
Published: (2026)
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
MeshVPR: Citywide Visual Place Recognition Using 3D Meshes
by: Berton, Gabriele, et al.
Published: (2024)
by: Berton, Gabriele, et al.
Published: (2024)
HypeVPR: Exploring Hyperbolic Space for Perspective to Equirectangular Visual Place Recognition
by: Woo, Suhan, et al.
Published: (2025)
by: Woo, Suhan, et al.
Published: (2025)
NYC-Indoor-VPR: A Long-Term Indoor Visual Place Recognition Dataset with Semi-Automatic Annotation
by: Sheng, Diwei, et al.
Published: (2024)
by: Sheng, Diwei, et al.
Published: (2024)
EffoVPR: Effective Foundation Model Utilization for Visual Place Recognition
by: Tzachor, Issar, et al.
Published: (2024)
by: Tzachor, Issar, et al.
Published: (2024)
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
by: Shen, Yanqing, et al.
Published: (2025)
by: Shen, Yanqing, et al.
Published: (2025)
SciceVPR: Stable Cross-Image Correlation Enhanced Model for Visual Place Recognition
by: Wan, Shanshan, et al.
Published: (2025)
by: Wan, Shanshan, et al.
Published: (2025)
CricaVPR: Cross-image Correlation-aware Representation Learning for Visual Place Recognition
by: Lu, Feng, et al.
Published: (2024)
by: Lu, Feng, et al.
Published: (2024)
NYC-Event-VPR: A Large-Scale High-Resolution Event-Based Visual Place Recognition Dataset in Dense Urban Environments
by: Pan, Taiyi, et al.
Published: (2024)
by: Pan, Taiyi, et al.
Published: (2024)
SelaVPR++: Towards Seamless Adaptation of Foundation Models for Efficient Place Recognition
by: Lu, Feng, et al.
Published: (2025)
by: Lu, Feng, et al.
Published: (2025)
Pair-VPR: Place-Aware Pre-training and Contrastive Pair Classification for Visual Place Recognition with Vision Transformers
by: Hausler, Stephen, et al.
Published: (2024)
by: Hausler, Stephen, et al.
Published: (2024)
UGNA-VPR: A Novel Training Paradigm for Visual Place Recognition Based on Uncertainty-Guided NeRF Augmentation
by: Shen, Yehui, et al.
Published: (2025)
by: Shen, Yehui, et al.
Published: (2025)
D$^{2}$-VPR: A Parameter-efficient Visual-foundation-model-based Visual Place Recognition Method via Knowledge Distillation and Deformable Aggregation
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
BEV$^2$PR: BEV-Enhanced Visual Place Recognition with Structural Cues
by: Ge, Fudong, et al.
Published: (2024)
by: Ge, Fudong, et al.
Published: (2024)
Text2Graph VPR: A Text-to-Graph Expert System for Explainable Place Recognition in Changing Environments
by: Yousefzadeh, Saeideh, et al.
Published: (2025)
by: Yousefzadeh, Saeideh, et al.
Published: (2025)
DiffPlace: Street View Generation via Place-Controllable Diffusion Model Enhancing Place Recognition
by: Li, Ji, et al.
Published: (2026)
by: Li, Ji, et al.
Published: (2026)
Feature Complementation Architecture for Visual Place Recognition
by: Wang, Weiwei, et al.
Published: (2025)
by: Wang, Weiwei, et al.
Published: (2025)
EPRBench: A High-Quality Benchmark Dataset for Event Stream Based Visual Place Recognition
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
Long-Term Visual Localization in Dynamic Benthic Environments: A Dataset, Footprint-Based Ground Truth, and Visual Place Recognition Benchmark
by: Larsen, Martin Kvisvik, et al.
Published: (2026)
by: Larsen, Martin Kvisvik, et al.
Published: (2026)
MMS-LLaMA: Efficient LLM-based Audio-Visual Speech Recognition with Minimal Multimodal Speech Tokens
by: Yeo, Jeong Hun, et al.
Published: (2025)
by: Yeo, Jeong Hun, et al.
Published: (2025)
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition
by: Wang, Teng, et al.
Published: (2024)
by: Wang, Teng, et al.
Published: (2024)
DisPlace: Discriminative Place Projections for Multi-Reference Visual Place Recognition
by: Rajani, Dhyey Manish, et al.
Published: (2026)
by: Rajani, Dhyey Manish, et al.
Published: (2026)
Optimal Transport Aggregation for Visual Place Recognition
by: Izquierdo, Sergio, et al.
Published: (2023)
by: Izquierdo, Sergio, et al.
Published: (2023)
Register assisted aggregation for Visual Place Recognition
by: Yu, Xuan, et al.
Published: (2024)
by: Yu, Xuan, et al.
Published: (2024)
Structured Pruning for Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2024)
by: Grainge, Oliver, et al.
Published: (2024)
Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets
by: Wu, Peng, et al.
Published: (2026)
by: Wu, Peng, et al.
Published: (2026)
VDNA-PR: Using General Dataset Representations for Robust Sequential Visual Place Recognition
by: Ramtoula, Benjamin, et al.
Published: (2024)
by: Ramtoula, Benjamin, et al.
Published: (2024)
EmbodiedPlace: Learning Mixture-of-Features with Embodied Constraints for Visual Place Recognition
by: Liu, Bingxi, et al.
Published: (2025)
by: Liu, Bingxi, et al.
Published: (2025)
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
by: Kassab, Hozaifa, et al.
Published: (2024)
by: Kassab, Hozaifa, et al.
Published: (2024)
Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition
by: Chen, Shunpeng, et al.
Published: (2026)
by: Chen, Shunpeng, et al.
Published: (2026)
VGGT-MPR: VGGT-Enhanced Multimodal Place Recognition in Autonomous Driving Environments
by: Xu, Jingyi, et al.
Published: (2026)
by: Xu, Jingyi, et al.
Published: (2026)
NocPlace: Nocturnal Visual Place Recognition via Generative and Inherited Knowledge Transfer
by: Liu, Bingxi, et al.
Published: (2024)
by: Liu, Bingxi, et al.
Published: (2024)
On the Estimation of Image-matching Uncertainty in Visual Place Recognition
by: Zaffar, Mubariz, et al.
Published: (2024)
by: Zaffar, Mubariz, et al.
Published: (2024)
Collaborative Visual Place Recognition through Federated Learning
by: Dutto, Mattia, et al.
Published: (2024)
by: Dutto, Mattia, et al.
Published: (2024)
Breaking the Frame: Visual Place Recognition by Overlap Prediction
by: Wei, Tong, et al.
Published: (2024)
by: Wei, Tong, et al.
Published: (2024)
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition
by: Jin, Tong, et al.
Published: (2024)
by: Jin, Tong, et al.
Published: (2024)
Focus on Local: Finding Reliable Discriminative Regions for Visual Place Recognition
by: Wang, Changwei, et al.
Published: (2025)
by: Wang, Changwei, et al.
Published: (2025)
Similar Items
-
Urban-ImageNet: A Large-Scale Multi-Modal Dataset and Evaluation Framework for Urban Space Perception
by: Ou, Yiwei, et al.
Published: (2026) -
LaVPR: Benchmarking Language and Vision for Place Recognition
by: Idan, Ofer, et al.
Published: (2026) -
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025) -
MeshVPR: Citywide Visual Place Recognition Using 3D Meshes
by: Berton, Gabriele, et al.
Published: (2024) -
HypeVPR: Exploring Hyperbolic Space for Perspective to Equirectangular Visual Place Recognition
by: Woo, Suhan, et al.
Published: (2025)