Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Grainge, Oliver, Waheed, Sania, Stilgoe, Jack, Milford, Michael, Ehsan, Shoaib |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
Structured Pruning for Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2024)
by: Grainge, Oliver, et al.
Published: (2024)
Employing Universal Voting Schemes for Improved Visual Place Recognition Performance
by: Waheed, Maria, et al.
Published: (2024)
by: Waheed, Maria, et al.
Published: (2024)
Image Embedding Sampling Method for Diverse Captioning
by: Waheed, Sania, et al.
Published: (2025)
by: Waheed, Sania, et al.
Published: (2025)
Through the Lens of Doubt: Robust and Efficient Uncertainty Estimation for Visual Place Recognition
by: Miller, Emily, et al.
Published: (2025)
by: Miller, Emily, et al.
Published: (2025)
Multi-Technique Sequential Information Consistency For Dynamic Visual Place Recognition In Changing Environments
by: Arcanjo, Bruno, et al.
Published: (2024)
by: Arcanjo, Bruno, et al.
Published: (2024)
Evaluating Precise Geolocation Inference Capabilities of Vision Language Models
by: Jay, Neel, et al.
Published: (2025)
by: Jay, Neel, et al.
Published: (2025)
Joint Multi-Condition Representation Modelling via Matrix Factorisation for Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2025)
by: Ismagilov, Timur, et al.
Published: (2025)
One Channel to Rule Them All: Rethinking Input Representation for Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2026)
by: Ismagilov, Timur, et al.
Published: (2026)
Evaluation of Geolocation Capabilities of Multimodal Large Language Models and Analysis of Associated Privacy Risks
by: Zhang, Xian, et al.
Published: (2025)
by: Zhang, Xian, et al.
Published: (2025)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
by: Sathe, Ashutosh, et al.
Published: (2024)
by: Sathe, Ashutosh, et al.
Published: (2024)
Granular Privacy Control for Geolocation with Vision Language Models
by: Mendes, Ethan, et al.
Published: (2024)
by: Mendes, Ethan, et al.
Published: (2024)
Skill-Conditioned Visual Geolocation for Vision-Language Models
by: Yang, Chenjie, et al.
Published: (2026)
by: Yang, Chenjie, et al.
Published: (2026)
On Motion Blur and Deblurring in Visual Place Recognition
by: Ismagilov, Timur, et al.
Published: (2024)
by: Ismagilov, Timur, et al.
Published: (2024)
Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization
by: Bharadwaj, Siddhant, et al.
Published: (2026)
by: Bharadwaj, Siddhant, et al.
Published: (2026)
Where on Earth? A Vision-Language Benchmark for Probing Model Geolocation Skills Across Scales
by: Qian, Zhaofang, et al.
Published: (2025)
by: Qian, Zhaofang, et al.
Published: (2025)
Image-Based Geolocation Using Large Vision-Language Models
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
by: Zheng, Xiangxi, et al.
Published: (2025)
by: Zheng, Xiangxi, et al.
Published: (2025)
Zero-shot Vision-Language Reranking for Cross-View Geolocalization
by: Erzurumlu, Yunus Talha, et al.
Published: (2026)
by: Erzurumlu, Yunus Talha, et al.
Published: (2026)
Capability $\neq$ Interpretability: Human Interpretability of Vision Foundation Models
by: Colin, Julien, et al.
Published: (2026)
by: Colin, Julien, et al.
Published: (2026)
LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild
by: Wang, Zhiqiang, et al.
Published: (2024)
by: Wang, Zhiqiang, et al.
Published: (2024)
Statewide Visual Geolocalization in the Wild
by: Fervers, Florian, et al.
Published: (2024)
by: Fervers, Florian, et al.
Published: (2024)
GeoShield: Safeguarding Geolocation Privacy from Vision-Language Models via Adversarial Perturbations
by: Liu, Xinwei, et al.
Published: (2025)
by: Liu, Xinwei, et al.
Published: (2025)
Improving Visual Place Recognition with Sequence-Matching Receptiveness Prediction
by: Hussaini, Somayeh, et al.
Published: (2025)
by: Hussaini, Somayeh, et al.
Published: (2025)
Automatic Map Density Selection for Locally-Performant Visual Place Recognition
by: Hussaini, Somayeh, et al.
Published: (2026)
by: Hussaini, Somayeh, et al.
Published: (2026)
Assessing and Learning Alignment of Unimodal Vision and Language Models
by: Zhang, Le, et al.
Published: (2024)
by: Zhang, Le, et al.
Published: (2024)
CHOICE: Benchmarking the Remote Sensing Capabilities of Large Vision-Language Models
by: An, Xiao, et al.
Published: (2024)
by: An, Xiao, et al.
Published: (2024)
Street-Level Geolocalization Using Multimodal Large Language Models and Retrieval-Augmented Generation
by: Bicakci, Yunus Serhat, et al.
Published: (2025)
by: Bicakci, Yunus Serhat, et al.
Published: (2025)
From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models
by: Gong, Weile, et al.
Published: (2026)
by: Gong, Weile, et al.
Published: (2026)
Sparkle: Mastering Basic Spatial Capabilities in Vision Language Models Elicits Generalization to Spatial Reasoning
by: Tang, Yihong, et al.
Published: (2024)
by: Tang, Yihong, et al.
Published: (2024)
Application Of Vision-Language Models For Assessing Osteoarthritis Disease Severity
by: Felfeliyan, Banafshe, et al.
Published: (2024)
by: Felfeliyan, Banafshe, et al.
Published: (2024)
Ensemble-Based Event Camera Place Recognition Under Varying Illumination
by: Joseph, Therese, et al.
Published: (2025)
by: Joseph, Therese, et al.
Published: (2025)
Applications of Spiking Neural Networks in Visual Place Recognition
by: Hussaini, Somayeh, et al.
Published: (2023)
by: Hussaini, Somayeh, et al.
Published: (2023)
Going Places: Place Recognition in Artificial and Natural Systems
by: Milford, Michael, et al.
Published: (2025)
by: Milford, Michael, et al.
Published: (2025)
Unlocking the Capabilities of Large Vision-Language Models for Generalizable and Explainable Deepfake Detection
by: Yu, Peipeng, et al.
Published: (2025)
by: Yu, Peipeng, et al.
Published: (2025)
Exploring the Zero-Shot Capabilities of Vision-Language Models for Improving Gaze Following
by: Gupta, Anshul, et al.
Published: (2024)
by: Gupta, Anshul, et al.
Published: (2024)
Unleashing the Capabilities of Large Vision-Language Models for Intelligent Perception of Roadside Infrastructure
by: Fu, Luxuan, et al.
Published: (2026)
by: Fu, Luxuan, et al.
Published: (2026)
Similar Items
-
Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?
by: Waheed, Sania, et al.
Published: (2025) -
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
by: Waheed, Sania, et al.
Published: (2025) -
TAT-VPR: Ternary Adaptive Transformer for Dynamic and Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025) -
TeTRA-VPR: A Ternary Transformer Approach for Compact Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2025) -
Structured Pruning for Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2024)