Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?
Fuente:
arXiv
Saved in:
| Main Authors: | Waheed, Sania, Ferrarini, Bruno, Milford, Michael, Ramchurn, Sarvapali D., Ehsan, Shoaib |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
Through the Lens of Doubt: Robust and Efficient Uncertainty Estimation for Visual Place Recognition
by: Miller, Emily, et al.
Published: (2025)
by: Miller, Emily, et al.
Published: (2025)
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
Structured Pruning for Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2024)
by: Grainge, Oliver, et al.
Published: (2024)
Joint Multi-Condition Representation Modelling via Matrix Factorisation for Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2025)
by: Ismagilov, Timur, et al.
Published: (2025)
On Motion Blur and Deblurring in Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2024)
by: Ismagilov, Timur, et al.
Published: (2024)
One Channel to Rule Them All: Rethinking Input Representation for Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2026)
by: Ismagilov, Timur, et al.
Published: (2026)
Multi-Technique Sequential Information Consistency For Dynamic Visual Place Recognition In Changing Environments
by: Arcanjo, Bruno, et al.
Published: (2024)
by: Arcanjo, Bruno, et al.
Published: (2024)
A Survey of Language-Based Communication in Robotics
by: Hunt, William, et al.
Published: (2024)
by: Hunt, William, et al.
Published: (2024)
Continuous Object State Recognition for Cooking Robots Using Pre-Trained Vision-Language Models and Black-box Optimization
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
Adversarial Attacks and Detection in Visual Place Recognition for Safer Robot Navigation
by: Malone, Connor, et al.
Published: (2025)
by: Malone, Connor, et al.
Published: (2025)
Image-Based Relocalization and Alignment for Long-Term Monitoring of Dynamic Underwater Environments
by: Gorry, Beverley, et al.
Published: (2025)
by: Gorry, Beverley, et al.
Published: (2025)
DTP: A Simple yet Effective Distracting Token Pruning Framework for Vision-Language Action Models
by: Li, Chenyang, et al.
Published: (2026)
by: Li, Chenyang, et al.
Published: (2026)
Distributed Nash Equilibrium Seeking Algorithm in Aggregative Games for Heterogeneous Multi-Robot Systems
by: Dong, Yi, et al.
Published: (2025)
by: Dong, Yi, et al.
Published: (2025)
Ensemble-Based Event Camera Place Recognition Under Varying Illumination
by: Joseph, Therese, et al.
Published: (2025)
by: Joseph, Therese, et al.
Published: (2025)
Applications of Spiking Neural Networks in Visual Place Recognition
by: Hussaini, Somayeh, et al.
Published: (2023)
by: Hussaini, Somayeh, et al.
Published: (2023)
Less is More: Lean yet Powerful Vision-Language Model for Autonomous Driving
by: Yang, Sheng, et al.
Published: (2025)
by: Yang, Sheng, et al.
Published: (2025)
Going Places: Place Recognition in Artificial and Natural Systems
by: Milford, Michael, et al.
Published: (2025)
by: Milford, Michael, et al.
Published: (2025)
Employing Universal Voting Schemes for Improved Visual Place Recognition Performance
by: Waheed, Maria, et al.
Published: (2024)
by: Waheed, Maria, et al.
Published: (2024)
Enhancing Visual Place Recognition via Fast and Slow Adaptive Biasing in Event Cameras
by: Nair, Gokul B., et al.
Published: (2024)
by: Nair, Gokul B., et al.
Published: (2024)
Quantile Transfer for Reliable Operating Point Selection in Visual Place Recognition
by: Rajani, Dhyey Manish, et al.
Published: (2026)
by: Rajani, Dhyey Manish, et al.
Published: (2026)
DisPlace: Discriminative Place Projections for Multi-Reference Visual Place Recognition
by: Rajani, Dhyey Manish, et al.
Published: (2026)
by: Rajani, Dhyey Manish, et al.
Published: (2026)
Improving Visual Place Recognition Based Robot Navigation By Verifying Localization Estimates
by: Claxton, Owen, et al.
Published: (2024)
by: Claxton, Owen, et al.
Published: (2024)
Robotic State Recognition with Image-to-Text Retrieval Task of Pre-Trained Vision-Language Model and Black-Box Optimization
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
Bridging Human Oversight and Black-box Driver Assistance: Vision-Language Models for Predictive Alerting in Lane Keeping Assist Systems
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
Image Embedding Sampling Method for Diverse Captioning
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
by: Han, Xiaofeng, et al.
Published: (2025)
by: Han, Xiaofeng, et al.
Published: (2025)
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
by: Shao, Rui, et al.
Published: (2025)
by: Shao, Rui, et al.
Published: (2025)
Saliency-Guided Domain Adaptation for Left-Hand Driving in Autonomous Steering
by: Mehraban, Zahra, et al.
Published: (2025)
by: Mehraban, Zahra, et al.
Published: (2025)
Long-Term Multi-Session 3D Reconstruction Under Substantial Appearance Change
by: Gorry, Beverley, et al.
Published: (2026)
by: Gorry, Beverley, et al.
Published: (2026)
Look Ma, No Ground Truth! Ground-Truth-Free Tuning of Structure from Motion and Visual SLAM
by: Fontan, Alejandro, et al.
Published: (2024)
by: Fontan, Alejandro, et al.
Published: (2024)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
by: Ding, Pengxiang, et al.
Published: (2023)
by: Ding, Pengxiang, et al.
Published: (2023)
Robotic Environmental State Recognition with Pre-Trained Vision-Language Models and Black-Box Optimization
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
FUSELOC: Fusing Global and Local Descriptors to Disambiguate 2D-3D Matching in Visual Localization
by: Nguyen, Son Tung, et al.
Published: (2024)
by: Nguyen, Son Tung, et al.
Published: (2024)
Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation
by: Nair, Gokul B., et al.
Published: (2025)
by: Nair, Gokul B., et al.
Published: (2025)
Generalized Robot 3D Vision-Language Model with Fast Rendering and Pre-Training Vision-Language Alignment
by: Liu, Kangcheng, et al.
Published: (2023)
by: Liu, Kangcheng, et al.
Published: (2023)
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
by: Cheng, An-Chieh, et al.
Published: (2024)
by: Cheng, An-Chieh, et al.
Published: (2024)
What Matters in Building Vision-Language-Action Models for Generalist Robots
by: Li, Xinghang, et al.
Published: (2024)
by: Li, Xinghang, et al.
Published: (2024)
Similar Items
-
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
by: Waheed, Sania, et al.
Published: (2025) -
Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
by: Grainge, Oliver, et al.
Published: (2025) -
Through the Lens of Doubt: Robust and Efficient Uncertainty Estimation for Visual Place Recognition
by: Miller, Emily, et al.
Published: (2025) -
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025) -
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)