Saved in:
| Main Authors: | Durrieu, Emilie, Hurter, Christophe, Muller, Philippe, Boutin, Victor |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.00912 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CityGuessr: City-Level Video Geo-Localization on a Global Scale
by: Kulkarni, Parth Parag, et al.
Published: (2024)
by: Kulkarni, Parth Parag, et al.
Published: (2024)
SpikeCLR: Contrastive Self-Supervised Learning for Few-Shot Event-Based Vision using Spiking Neural Networks
by: Vaillant, Maxime, et al.
Published: (2026)
by: Vaillant, Maxime, et al.
Published: (2026)
GeoRouter: Dynamic Paradigm Routing for Worldwide Image Geolocalization
by: Jia, Pengyue, et al.
Published: (2026)
by: Jia, Pengyue, et al.
Published: (2026)
GeoRanker: Distance-Aware Ranking for Worldwide Image Geolocalization
by: Jia, Pengyue, et al.
Published: (2025)
by: Jia, Pengyue, et al.
Published: (2025)
LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild
by: Wang, Zhiqiang, et al.
Published: (2024)
by: Wang, Zhiqiang, et al.
Published: (2024)
PIGEON: Predicting Image Geolocations
by: Haas, Lukas, et al.
Published: (2023)
by: Haas, Lukas, et al.
Published: (2023)
GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization
by: Wang, Yikun, et al.
Published: (2025)
by: Wang, Yikun, et al.
Published: (2025)
Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization
by: Bharadwaj, Siddhant, et al.
Published: (2026)
by: Bharadwaj, Siddhant, et al.
Published: (2026)
Street-Level Geolocalization Using Multimodal Large Language Models and Retrieval-Augmented Generation
by: Bicakci, Yunus Serhat, et al.
Published: (2025)
by: Bicakci, Yunus Serhat, et al.
Published: (2025)
Image-Based Geolocation Using Large Vision-Language Models
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
by: Ghasemi, Narges, et al.
Published: (2025)
by: Ghasemi, Narges, et al.
Published: (2025)
Measuring the Impact of Scene Level Objects on Object Detection: Towards Quantitative Explanations of Detection Decisions
by: Haar, Lynn Vonder, et al.
Published: (2024)
by: Haar, Lynn Vonder, et al.
Published: (2024)
Statewide Visual Geolocalization in the Wild
by: Fervers, Florian, et al.
Published: (2024)
by: Fervers, Florian, et al.
Published: (2024)
Enhancing Worldwide Image Geolocation by Ensembling Satellite-Based Ground-Level Attribute Predictors
by: Bianco, Michael J., et al.
Published: (2024)
by: Bianco, Michael J., et al.
Published: (2024)
A Variational Approach for Joint Image Recovery and Feature Extraction Based on Spatially-Varying Generalised Gaussian Models
by: Chouzenoux, Emilie, et al.
Published: (2022)
by: Chouzenoux, Emilie, et al.
Published: (2022)
From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models
by: Li, Lingyao, et al.
Published: (2025)
by: Li, Lingyao, et al.
Published: (2025)
Combi-CAM: A Novel Multi-Layer Approach for Explainable Image Geolocalization
by: Faget, David, et al.
Published: (2026)
by: Faget, David, et al.
Published: (2026)
Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
by: Zhou, Zhongliang, et al.
Published: (2024)
by: Zhou, Zhongliang, et al.
Published: (2024)
Examining the Impact of Optical Aberrations to Image Classification and Object Detection Models
by: Müller, Patrick, et al.
Published: (2025)
by: Müller, Patrick, et al.
Published: (2025)
CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping
by: Lebailly, Tim, et al.
Published: (2023)
by: Lebailly, Tim, et al.
Published: (2023)
Interpretable Perception and Reasoning for Audiovisual Geolocation
by: Su, Yiyang, et al.
Published: (2026)
by: Su, Yiyang, et al.
Published: (2026)
Granular Privacy Control for Geolocation with Vision Language Models
by: Mendes, Ethan, et al.
Published: (2024)
by: Mendes, Ethan, et al.
Published: (2024)
Cross-Modal Redundancy and the Geometry of Vision-Language Embeddings
by: Dhimoïla, Grégoire, et al.
Published: (2026)
by: Dhimoïla, Grégoire, et al.
Published: (2026)
Learning to Wander: Improving the Global Image Geolocation Ability of LMMs via Actionable Reasoning
by: Zheng, Yushuo, et al.
Published: (2026)
by: Zheng, Yushuo, et al.
Published: (2026)
Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion
by: Bethell, Adam, et al.
Published: (2024)
by: Bethell, Adam, et al.
Published: (2024)
GeoShield: Safeguarding Geolocation Privacy from Vision-Language Models via Adversarial Perturbations
by: Liu, Xinwei, et al.
Published: (2025)
by: Liu, Xinwei, et al.
Published: (2025)
GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow Prediction
by: Lehyeh, Ayesh Abu, et al.
Published: (2026)
by: Lehyeh, Ayesh Abu, et al.
Published: (2026)
GeoSearch: Augmenting Worldwide Geolocalization with Web-Scale Reverse Image Search and Image Matching
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
Choosing the right basis for interpretability: Psychophysical comparison between neuron-based and dictionary-based representations
by: Colin, Julien, et al.
Published: (2024)
by: Colin, Julien, et al.
Published: (2024)
Instance-Level Moving Object Segmentation from a Single Image with Events
by: Wan, Zhexiong, et al.
Published: (2025)
by: Wan, Zhexiong, et al.
Published: (2025)
Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
by: Grainge, Oliver, et al.
Published: (2025)
by: Grainge, Oliver, et al.
Published: (2025)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
by: Chen, Junzhe, et al.
Published: (2024)
by: Chen, Junzhe, et al.
Published: (2024)
GeoRC: A Benchmark for Geolocation Reasoning Chains
by: Talreja, Mohit, et al.
Published: (2026)
by: Talreja, Mohit, et al.
Published: (2026)
3D Object Manipulation in a Single Image using Generative Models
by: Zhao, Ruisi, et al.
Published: (2025)
by: Zhao, Ruisi, et al.
Published: (2025)
Accurate Explanation Model for Image Classifiers using Class Association Embedding
by: Xie, Ruitao, et al.
Published: (2024)
by: Xie, Ruitao, et al.
Published: (2024)
Skill-Conditioned Visual Geolocation for Vision-Language Models
by: Yang, Chenjie, et al.
Published: (2026)
by: Yang, Chenjie, et al.
Published: (2026)
GaGA: Towards Interactive Global Geolocation Assistant
by: Dou, Zhiyang, et al.
Published: (2024)
by: Dou, Zhiyang, et al.
Published: (2024)
Unified Category-Level Object Detection and Pose Estimation from RGB Images using 3D Prototypes
by: Fischer, Tom, et al.
Published: (2025)
by: Fischer, Tom, et al.
Published: (2025)
Detecting Text Manipulation in Images using Vision Language Models
by: Vidit, Vidit, et al.
Published: (2025)
by: Vidit, Vidit, et al.
Published: (2025)
SMGeo: Cross-View Object Geo-Localization with Grid-Level Mixture-of-Experts
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
Similar Items
-
CityGuessr: City-Level Video Geo-Localization on a Global Scale
by: Kulkarni, Parth Parag, et al.
Published: (2024) -
SpikeCLR: Contrastive Self-Supervised Learning for Few-Shot Event-Based Vision using Spiking Neural Networks
by: Vaillant, Maxime, et al.
Published: (2026) -
GeoRouter: Dynamic Paradigm Routing for Worldwide Image Geolocalization
by: Jia, Pengyue, et al.
Published: (2026) -
GeoRanker: Distance-Aware Ranking for Worldwide Image Geolocalization
by: Jia, Pengyue, et al.
Published: (2025) -
LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild
by: Wang, Zhiqiang, et al.
Published: (2024)