Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
Fuente:
arXiv
Saved in:
| Main Authors: | Dufour, Nicolas, Picard, David, Kalogeiton, Vicky, Landrieu, Loic |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
by: Dufour, Nicolas, et al.
Published: (2024)
by: Dufour, Nicolas, et al.
Published: (2024)
MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency
by: Dufour, Nicolas, et al.
Published: (2025)
by: Dufour, Nicolas, et al.
Published: (2025)
Training-Free Synthetic Data Generation with Dual IP-Adapter Guidance
by: Boudier, Luc, et al.
Published: (2025)
by: Boudier, Luc, et al.
Published: (2025)
Analysis of Classifier-Free Guidance Weight Schedulers
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
PoM: Efficient Image and Video Generation with the Polynomial Mixer
by: Picard, David, et al.
Published: (2024)
by: Picard, David, et al.
Published: (2024)
T-REGS: Minimum Spanning Tree Regularization for Self-Supervised Learning
by: Mordacq, Julie, et al.
Published: (2025)
by: Mordacq, Julie, et al.
Published: (2025)
Di$\mathtt{[M]}$O: Distilling Masked Diffusion Models into One-step Generator
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
Soft-Di[M]O: Improving One-Step Discrete Image Generation with Soft Embeddings
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
OpenStreetView-5M: The Many Roads to Global Visual Geolocation
by: Astruc, Guillaume, et al.
Published: (2024)
by: Astruc, Guillaume, et al.
Published: (2024)
ADAPT: Multimodal Learning for Detecting Physiological Changes under Missing Modalities
by: Mordacq, Julie, et al.
Published: (2024)
by: Mordacq, Julie, et al.
Published: (2024)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2023)
by: Benigmim, Yasser, et al.
Published: (2023)
Diffusion Reinforcement Learning via Centered Reward Distillation
by: Zhu, Yuanzhi, et al.
Published: (2026)
by: Zhu, Yuanzhi, et al.
Published: (2026)
One-step Diffusion Models with Bregman Density Ratio Matching
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
E.T. the Exceptional Trajectories: Text-to-camera-trajectory generation with character awareness
by: Courant, Robin, et al.
Published: (2024)
by: Courant, Robin, et al.
Published: (2024)
Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach
by: Liu, Yaofang, et al.
Published: (2024)
by: Liu, Yaofang, et al.
Published: (2024)
How far can we go with ImageNet for Text-to-Image generation?
by: Degeorge, L., et al.
Published: (2025)
by: Degeorge, L., et al.
Published: (2025)
Rethinking Timesteps Samplers and Prediction Types
by: Xie, Bin, et al.
Published: (2025)
by: Xie, Bin, et al.
Published: (2025)
PIGEON: Predicting Image Geolocations
by: Haas, Lukas, et al.
Published: (2023)
by: Haas, Lukas, et al.
Published: (2023)
Activation-wise Propagation: A One-Timestep Strategy for Spiking Neural Networks
by: Song, Jian, et al.
Published: (2025)
by: Song, Jian, et al.
Published: (2025)
Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training
by: Kim, Myunsoo, et al.
Published: (2024)
by: Kim, Myunsoo, et al.
Published: (2024)
Formalizing the Sampling Design Space of Diffusion-Based Generative Models via Adaptive Solvers and Wasserstein-Bounded Timesteps
by: Jo, Sangwoo, et al.
Published: (2026)
by: Jo, Sangwoo, et al.
Published: (2026)
AnySat: One Earth Observation Model for Many Resolutions, Scales, and Modalities
by: Astruc, Guillaume, et al.
Published: (2024)
by: Astruc, Guillaume, et al.
Published: (2024)
OmniSat: Self-Supervised Modality Fusion for Earth Observation
by: Astruc, Guillaume, et al.
Published: (2024)
by: Astruc, Guillaume, et al.
Published: (2024)
Quantised Global Autoencoder: A Holistic Approach to Representing Visual Data
by: Elsner, Tim, et al.
Published: (2024)
by: Elsner, Tim, et al.
Published: (2024)
TMPQ-DM: Joint Timestep Reduction and Quantization Precision Selection for Efficient Diffusion Models
by: Sun, Haojun, et al.
Published: (2024)
by: Sun, Haojun, et al.
Published: (2024)
Scalable 3D Panoptic Segmentation As Superpoint Graph Clustering
by: Robert, Damien, et al.
Published: (2024)
by: Robert, Damien, et al.
Published: (2024)
EZ-SP: Fast and Lightweight Superpoint-Based 3D Segmentation
by: Geist, Louis, et al.
Published: (2025)
by: Geist, Louis, et al.
Published: (2025)
One View Is Enough! Monocular Training for In-the-Wild Novel View Generation
by: Rahary, Adrien Ramanana, et al.
Published: (2026)
by: Rahary, Adrien Ramanana, et al.
Published: (2026)
CoDEx: Combining Domain Expertise for Spatial Generalization in Satellite Image Analysis
by: Kuriyal, Abhishek, et al.
Published: (2025)
by: Kuriyal, Abhishek, et al.
Published: (2025)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
by: Delatolas, Thanos, et al.
Published: (2025)
by: Delatolas, Thanos, et al.
Published: (2025)
The Disappearance of Timestep Embedding in Modern Time-Dependent Neural Networks
by: Kim, Bum Jun, et al.
Published: (2024)
by: Kim, Bum Jun, et al.
Published: (2024)
SF20K Competition 2025: Summary and findings
by: Ghermi, Ridouane, et al.
Published: (2026)
by: Ghermi, Ridouane, et al.
Published: (2026)
Enhancing Worldwide Image Geolocation by Ensembling Satellite-Based Ground-Level Attribute Predictors
by: Bianco, Michael J., et al.
Published: (2024)
by: Bianco, Michael J., et al.
Published: (2024)
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
by: Neuhaus, Yannic, et al.
Published: (2026)
by: Neuhaus, Yannic, et al.
Published: (2026)
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
by: You, Haoran, et al.
Published: (2024)
by: You, Haoran, et al.
Published: (2024)
Order Matters: 3D Shape Generation from Sequential VR Sketches
by: Chen, Yizi, et al.
Published: (2025)
by: Chen, Yizi, et al.
Published: (2025)
Unified Text-Image-to-Video Generation: A Training-Free Approach to Flexible Visual Conditioning
by: Lai, Bolin, et al.
Published: (2025)
by: Lai, Bolin, et al.
Published: (2025)
Adapting Vision Transformers to Ultra-High Resolution Semantic Segmentation with Relay Tokens
by: Perron, Yohann, et al.
Published: (2026)
by: Perron, Yohann, et al.
Published: (2026)
Learnable Earth Parser: Discovering 3D Prototypes in Aerial Scans
by: Loiseau, Romain, et al.
Published: (2023)
by: Loiseau, Romain, et al.
Published: (2023)
GeoRC: A Benchmark for Geolocation Reasoning Chains
by: Talreja, Mohit, et al.
Published: (2026)
by: Talreja, Mohit, et al.
Published: (2026)
Similar Items
-
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
by: Dufour, Nicolas, et al.
Published: (2024) -
MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency
by: Dufour, Nicolas, et al.
Published: (2025) -
Training-Free Synthetic Data Generation with Dual IP-Adapter Guidance
by: Boudier, Luc, et al.
Published: (2025) -
Analysis of Classifier-Free Guidance Weight Schedulers
by: Wang, Xi, et al.
Published: (2024) -
PoM: Efficient Image and Video Generation with the Polynomial Mixer
by: Picard, David, et al.
Published: (2024)