PIGEON: Predicting Image Geolocations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Haas, Lukas, Skreta, Michal, Alberti, Silas, Finn, Chelsea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
Tripod: Three Complementary Inductive Biases for Disentangled Representation Learning
von: Hsu, Kyle, et al.
Veröffentlicht: (2024)
von: Hsu, Kyle, et al.
Veröffentlicht: (2024)
AutoFT: Learning an Objective for Robust Fine-Tuning
von: Choi, Caroline, et al.
Veröffentlicht: (2024)
von: Choi, Caroline, et al.
Veröffentlicht: (2024)
Aligning Modalities in Vision Large Language Models via Preference Fine-tuning
von: Zhou, Yiyang, et al.
Veröffentlicht: (2024)
von: Zhou, Yiyang, et al.
Veröffentlicht: (2024)
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
Enhancing Worldwide Image Geolocation by Ensembling Satellite-Based Ground-Level Attribute Predictors
von: Bianco, Michael J., et al.
Veröffentlicht: (2024)
von: Bianco, Michael J., et al.
Veröffentlicht: (2024)
Helpful DoggyBot: Open-World Object Fetching using Legged Robots and Vision-Language Models
von: Wu, Qi, et al.
Veröffentlicht: (2024)
von: Wu, Qi, et al.
Veröffentlicht: (2024)
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
von: Fu, Zipeng, et al.
Veröffentlicht: (2024)
von: Fu, Zipeng, et al.
Veröffentlicht: (2024)
Evaluating Precise Geolocation Inference Capabilities of Vision Language Models
von: Jay, Neel, et al.
Veröffentlicht: (2025)
von: Jay, Neel, et al.
Veröffentlicht: (2025)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
von: Zhou, Yiyang, et al.
Veröffentlicht: (2023)
von: Zhou, Yiyang, et al.
Veröffentlicht: (2023)
Deepfake Detection of Face Images based on a Convolutional Neural Network
von: Kroiß, Lukas, et al.
Veröffentlicht: (2025)
von: Kroiß, Lukas, et al.
Veröffentlicht: (2025)
GeoRC: A Benchmark for Geolocation Reasoning Chains
von: Talreja, Mohit, et al.
Veröffentlicht: (2026)
von: Talreja, Mohit, et al.
Veröffentlicht: (2026)
Conformal Prediction for Image Segmentation Using Morphological Prediction Sets
von: Mossina, Luca, et al.
Veröffentlicht: (2025)
von: Mossina, Luca, et al.
Veröffentlicht: (2025)
Pitfalls of Conformal Predictions for Medical Image Classification
von: Mehrtens, Hendrik, et al.
Veröffentlicht: (2025)
von: Mehrtens, Hendrik, et al.
Veröffentlicht: (2025)
No Safe Dose: How Training Data Drives Unsafe Image Generation
von: Friedrich, Felix, et al.
Veröffentlicht: (2026)
von: Friedrich, Felix, et al.
Veröffentlicht: (2026)
Text-Guided Image Clustering
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
Controllable Image Generation with Composed Parallel Token Prediction
von: Stirling, Jamie, et al.
Veröffentlicht: (2024)
von: Stirling, Jamie, et al.
Veröffentlicht: (2024)
PatchMorph: A Stochastic Deep Learning Approach for Unsupervised 3D Brain Image Registration with Small Patches
von: Skibbe, Henrik, et al.
Veröffentlicht: (2023)
von: Skibbe, Henrik, et al.
Veröffentlicht: (2023)
Origins of Creativity in Attention-Based Diffusion Models
von: Finn, Emma, et al.
Veröffentlicht: (2025)
von: Finn, Emma, et al.
Veröffentlicht: (2025)
Fast Fishing: Approximating BAIT for Efficient and Scalable Deep Active Image Classification
von: Huseljic, Denis, et al.
Veröffentlicht: (2024)
von: Huseljic, Denis, et al.
Veröffentlicht: (2024)
HumanPlus: Humanoid Shadowing and Imitation from Humans
von: Fu, Zipeng, et al.
Veröffentlicht: (2024)
von: Fu, Zipeng, et al.
Veröffentlicht: (2024)
MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
Confidence-Based Task Prediction in Continual Disease Classification Using Probability Distribution
von: Verma, Tanvi, et al.
Veröffentlicht: (2024)
von: Verma, Tanvi, et al.
Veröffentlicht: (2024)
Adaptive Conformal Prediction for Reliable and Explainable Medical Image Classification
von: Octadion, One, et al.
Veröffentlicht: (2026)
von: Octadion, One, et al.
Veröffentlicht: (2026)
CASPFormer: Trajectory Prediction from BEV Images with Deformable Attention
von: Yadav, Harsh, et al.
Veröffentlicht: (2024)
von: Yadav, Harsh, et al.
Veröffentlicht: (2024)
Controlling False Positives in Image Segmentation via Conformal Prediction
von: Mossina, Luca, et al.
Veröffentlicht: (2025)
von: Mossina, Luca, et al.
Veröffentlicht: (2025)
Predictive Digital Twin for Condition Monitoring Using Thermal Imaging
von: Menges, Daniel, et al.
Veröffentlicht: (2024)
von: Menges, Daniel, et al.
Veröffentlicht: (2024)
Text-Guided Alternative Image Clustering
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
PIGEON: VLM-Driven Object Navigation via Points of Interest Selection
von: Peng, Cheng, et al.
Veröffentlicht: (2025)
von: Peng, Cheng, et al.
Veröffentlicht: (2025)
Visual Analysis of Prediction Uncertainty in Neural Networks for Deep Image Synthesis
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Improving Predictive Confidence in Medical Imaging via Online Label Smoothing
von: Choudhury, Kushan, et al.
Veröffentlicht: (2025)
von: Choudhury, Kushan, et al.
Veröffentlicht: (2025)
Conformal Semantic Image Segmentation: Post-hoc Quantification of Predictive Uncertainty
von: Mossina, Luca, et al.
Veröffentlicht: (2024)
von: Mossina, Luca, et al.
Veröffentlicht: (2024)
Entropy Minimization without Model Collapse: Mitigating Prediction Bias in Medical Imaging
von: Nielen, Tim, et al.
Veröffentlicht: (2026)
von: Nielen, Tim, et al.
Veröffentlicht: (2026)
Integrating Artificial Intelligence Models and Synthetic Image Data for Enhanced Asset Inspection and Defect Identification
von: Mandati, Reddy, et al.
Veröffentlicht: (2024)
von: Mandati, Reddy, et al.
Veröffentlicht: (2024)
PolypNextLSTM: A lightweight and fast polyp video segmentation network using ConvNext and ConvLSTM
von: Bhattacharya, Debayan, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Debayan, et al.
Veröffentlicht: (2024)
Localized adversarial artifacts for compressed sensing MRI
von: Alaifari, Rima, et al.
Veröffentlicht: (2022)
von: Alaifari, Rima, et al.
Veröffentlicht: (2022)
Device Image-IV Mapping using Variational Autoencoder for Inverse Design and Forward Prediction
von: Lu, Thomas, et al.
Veröffentlicht: (2023)
von: Lu, Thomas, et al.
Veröffentlicht: (2023)
ThermoPore: Predicting Part Porosity Based on Thermal Images Using Deep Learning
von: Pak, Peter Myung-Won, et al.
Veröffentlicht: (2024)
von: Pak, Peter Myung-Won, et al.
Veröffentlicht: (2024)
xCG: Explainable Cell Graphs for Survival Prediction in Non-Small Cell Lung Cancer
von: Sextro, Marvin, et al.
Veröffentlicht: (2024)
von: Sextro, Marvin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025) -
Tripod: Three Complementary Inductive Biases for Disentangled Representation Learning
von: Hsu, Kyle, et al.
Veröffentlicht: (2024) -
AutoFT: Learning an Objective for Robust Fine-Tuning
von: Choi, Caroline, et al.
Veröffentlicht: (2024) -
Aligning Modalities in Vision Large Language Models via Preference Fine-tuning
von: Zhou, Yiyang, et al.
Veröffentlicht: (2024) -
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)