Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency
Fuente:
arXiv
Saved in:
| Main Authors: | Kalantidis, Yannis, Sarıyıldız, Mert Bülent, Rezende, Rafael S., Weinzaepfel, Philippe, Larlus, Diane, Csurka, Gabriela |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UNIC: Universal Classification Models via Multi-teacher Distillation
by: Sariyildiz, Mert Bulent, et al.
Published: (2024)
by: Sariyildiz, Mert Bulent, et al.
Published: (2024)
DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
Task Alignment: A simple and effective proxy for model merging in computer vision
by: de Jorge, Pau, et al.
Published: (2026)
by: de Jorge, Pau, et al.
Published: (2026)
What could go wrong? Discovering and describing failure modes in computer vision
by: Csurka, Gabriela, et al.
Published: (2024)
by: Csurka, Gabriela, et al.
Published: (2024)
Kinaema: a recurrent sequence model for memory and pose in motion
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
by: Suma, Pavel, et al.
Published: (2026)
by: Suma, Pavel, et al.
Published: (2026)
Gaussian Splatting Feature Fields for Privacy-Preserving Visual Localization
by: Pietrantoni, Maxime, et al.
Published: (2025)
by: Pietrantoni, Maxime, et al.
Published: (2025)
Label Propagation for Zero-shot Classification with Vision-Language Models
by: Stojnić, Vladan, et al.
Published: (2024)
by: Stojnić, Vladan, et al.
Published: (2024)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
by: Jiang, Zeren, et al.
Published: (2025)
by: Jiang, Zeren, et al.
Published: (2025)
RANa: Retrieval-Augmented Navigation
by: Monaci, Gianluca, et al.
Published: (2025)
by: Monaci, Gianluca, et al.
Published: (2025)
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
by: Stojnić, Vladan, et al.
Published: (2025)
by: Stojnić, Vladan, et al.
Published: (2025)
On Good Practices for Task-Specific Distillation of Large Pretrained Visual Models
by: Marrie, Juliette, et al.
Published: (2024)
by: Marrie, Juliette, et al.
Published: (2024)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
by: Tschernezki, Vadim, et al.
Published: (2025)
by: Tschernezki, Vadim, et al.
Published: (2025)
PanSt3R: Multi-view Consistent Panoptic Segmentation
by: Zust, Lojze, et al.
Published: (2025)
by: Zust, Lojze, et al.
Published: (2025)
Self-supervised Learning of Neural Implicit Feature Fields for Camera Pose Refinement
by: Pietrantoni, Maxime, et al.
Published: (2024)
by: Pietrantoni, Maxime, et al.
Published: (2024)
Can we make NeRF-based visual localization privacy-preserving?
by: Pietrantoni, Maxime, et al.
Published: (2025)
by: Pietrantoni, Maxime, et al.
Published: (2025)
What does really matter in image goal navigation?
by: Monaci, Gianluca, et al.
Published: (2025)
by: Monaci, Gianluca, et al.
Published: (2025)
LUDVIG: Learning-Free Uplifting of 2D Visual Features to Gaussian Splatting Scenes
by: Marrie, Juliette, et al.
Published: (2024)
by: Marrie, Juliette, et al.
Published: (2024)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
Win-Win: Training High-Resolution Vision Transformers from Two Windows
by: Leroy, Vincent, et al.
Published: (2023)
by: Leroy, Vincent, et al.
Published: (2023)
PoseEmbroider: Towards a 3D, Visual, Semantic-aware Human Pose Representation
by: Delmas, Ginger, et al.
Published: (2024)
by: Delmas, Ginger, et al.
Published: (2024)
PoseFix: Correcting 3D Human Poses with Natural Language
by: Delmas, Ginger, et al.
Published: (2023)
by: Delmas, Ginger, et al.
Published: (2023)
Purposer: Putting Human Motion Generation in Context
by: Ugrinovic, Nicolas, et al.
Published: (2024)
by: Ugrinovic, Nicolas, et al.
Published: (2024)
SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
Pow3R: Empowering Unconstrained 3D Reconstruction with Camera and Scene Priors
by: Jang, Wonbong, et al.
Published: (2025)
by: Jang, Wonbong, et al.
Published: (2025)
Enhancing Cross-View Geo-Localization Generalization via Global-Local Consistency and Geometric Equivariance
by: Wang, Xiaowei, et al.
Published: (2025)
by: Wang, Xiaowei, et al.
Published: (2025)
PoseScript: Linking 3D Human Poses and Natural Language
by: Delmas, Ginger, et al.
Published: (2022)
by: Delmas, Ginger, et al.
Published: (2022)
Test-time Vocabulary Adaptation for Language-driven Object Detection
by: Liu, Mingxuan, et al.
Published: (2025)
by: Liu, Mingxuan, et al.
Published: (2025)
ControlVP: Interactive Geometric Refinement of AI-Generated Images with Consistent Vanishing Points
by: Okumura, Ryota, et al.
Published: (2025)
by: Okumura, Ryota, et al.
Published: (2025)
MASt3R-SfM: a Fully-Integrated Solution for Unconstrained Structure-from-Motion
by: Duisterhof, Bardienus, et al.
Published: (2024)
by: Duisterhof, Bardienus, et al.
Published: (2024)
HAMSt3R: Human-Aware Multi-view Stereo 3D Reconstruction
by: Rojas, Sara, et al.
Published: (2025)
by: Rojas, Sara, et al.
Published: (2025)
Sora Generates Videos with Stunning Geometrical Consistency
by: Li, Xuanyi, et al.
Published: (2024)
by: Li, Xuanyi, et al.
Published: (2024)
EPIC Fields: Marrying 3D Geometry and Video Understanding
by: Tschernezki, Vadim, et al.
Published: (2023)
by: Tschernezki, Vadim, et al.
Published: (2023)
MUSt3R: Multi-view Network for Stereo 3D Reconstruction
by: Cabon, Yohann, et al.
Published: (2025)
by: Cabon, Yohann, et al.
Published: (2025)
Text-Video Retrieval with Global-Local Semantic Consistent Learning
by: Zhang, Haonan, et al.
Published: (2024)
by: Zhang, Haonan, et al.
Published: (2024)
Grab-3D: Detecting AI-Generated Videos from 3D Geometric Temporal Consistency
by: Chen, Wenhan, et al.
Published: (2025)
by: Chen, Wenhan, et al.
Published: (2025)
CondiMen: Conditional Multi-Person Mesh Recovery
by: Romain, Brégier, et al.
Published: (2024)
by: Romain, Brégier, et al.
Published: (2024)
Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot
by: Baradel, Fabien, et al.
Published: (2024)
by: Baradel, Fabien, et al.
Published: (2024)
How Could Generative AI Support Compliance with the EU AI Act? A Review for Safe Automated Driving Perception
by: Keser, Mert, et al.
Published: (2024)
by: Keser, Mert, et al.
Published: (2024)
GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation
by: Ackermann, Jan, et al.
Published: (2026)
by: Ackermann, Jan, et al.
Published: (2026)
Similar Items
-
UNIC: Universal Classification Models via Multi-teacher Distillation
by: Sariyildiz, Mert Bulent, et al.
Published: (2024) -
DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers
by: Sariyildiz, Mert Bulent, et al.
Published: (2025) -
Task Alignment: A simple and effective proxy for model merging in computer vision
by: de Jorge, Pau, et al.
Published: (2026) -
What could go wrong? Discovering and describing failure modes in computer vision
by: Csurka, Gabriela, et al.
Published: (2024) -
Kinaema: a recurrent sequence model for memory and pose in motion
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)