A Unified Hierarchical Framework for Fine-grained Cross-view Geo-localization over Large-scale Scenarios
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Zhuo, Zhang, Ye, Li, Kunhong, Wang, Longguang, Guo, Yulan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PanMatch: Unleashing the Potential of Large Vision Models for Unified Matching Models
by: Zhang, Yongjian, et al.
Published: (2025)
by: Zhang, Yongjian, et al.
Published: (2025)
DropoutGS: Dropping Out Gaussians for Better Sparse-view Rendering
by: Xu, Yexing, et al.
Published: (2025)
by: Xu, Yexing, et al.
Published: (2025)
AIQViT: Architecture-Informed Post-Training Quantization for Vision Transformers
by: Jiang, Runqing, et al.
Published: (2025)
by: Jiang, Runqing, et al.
Published: (2025)
Pluggable Style Representation Learning for Multi-Style Transfer
by: Liu, Hongda, et al.
Published: (2025)
by: Liu, Hongda, et al.
Published: (2025)
SaMam: Style-aware State Space Model for Arbitrary Image Style Transfer
by: Liu, Hongda, et al.
Published: (2025)
by: Liu, Hongda, et al.
Published: (2025)
Preserving Full Degradation Details for Blind Image Super-Resolution
by: Liu, Hongda, et al.
Published: (2024)
by: Liu, Hongda, et al.
Published: (2024)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
by: Chen, Minglin, et al.
Published: (2025)
by: Chen, Minglin, et al.
Published: (2025)
Learning Cross-view Visual Geo-localization without Ground Truth
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Large-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding
by: Shui, Zhongyi, et al.
Published: (2025)
by: Shui, Zhongyi, et al.
Published: (2025)
UPGS: Unified Pose-aware Gaussian Splatting for Dynamic Scene Deblurring
by: Wu, Zhijing, et al.
Published: (2025)
by: Wu, Zhijing, et al.
Published: (2025)
CrossView-GS: Cross-view Gaussian Splatting For Large-scale Scene Reconstruction
by: Zhang, Chenhao, et al.
Published: (2025)
by: Zhang, Chenhao, et al.
Published: (2025)
3rd Place Solution to Large-scale Fine-grained Food Recognition
by: Zhong, Yang, et al.
Published: (2025)
by: Zhong, Yang, et al.
Published: (2025)
VideoDirector: Precise Video Editing via Text-to-Video Models
by: Wang, Yukun, et al.
Published: (2024)
by: Wang, Yukun, et al.
Published: (2024)
Anchor-free Cross-view Object Geo-localization with Gaussian Position Encoding and Cross-view Association
by: Ling, Xingtao, et al.
Published: (2025)
by: Ling, Xingtao, et al.
Published: (2025)
Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-Tuning
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
by: Zhang, Ye, et al.
Published: (2026)
by: Zhang, Ye, et al.
Published: (2026)
ConGeo: Robust Cross-view Geo-localization across Ground View Variations
by: Mi, Li, et al.
Published: (2024)
by: Mi, Li, et al.
Published: (2024)
MobileGeo: Exploring Hierarchical Knowledge Distillation for Resource-Efficient Cross-view Drone Geo-Localization
by: Sun, Jian, et al.
Published: (2025)
by: Sun, Jian, et al.
Published: (2025)
Pack-PTQ: Advancing Post-training Quantization of Neural Networks by Pack-wise Reconstruction
by: Li, Changjun, et al.
Published: (2025)
by: Li, Changjun, et al.
Published: (2025)
Language-guided Hierarchical Fine-grained Image Forgery Detection and Localization
by: Guo, Xiao, et al.
Published: (2024)
by: Guo, Xiao, et al.
Published: (2024)
PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials
by: Hu, Wenmiao, et al.
Published: (2024)
by: Hu, Wenmiao, et al.
Published: (2024)
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
by: Zhang, Shi-Xue, et al.
Published: (2025)
by: Zhang, Shi-Xue, et al.
Published: (2025)
JRN-Geo: A Joint Perception Network based on RGB and Normal images for Cross-view Geo-localization
by: Zhou, Hongyu, et al.
Published: (2025)
by: Zhou, Hongyu, et al.
Published: (2025)
Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
by: Ye, Hualin, et al.
Published: (2025)
by: Ye, Hualin, et al.
Published: (2025)
Deep Lookup Network
by: Guo, Yulan, et al.
Published: (2025)
by: Guo, Yulan, et al.
Published: (2025)
Achieving Fine-grained Cross-modal Understanding through Brain-inspired Hierarchical Representation Learning
by: You, Weihang, et al.
Published: (2026)
by: You, Weihang, et al.
Published: (2026)
Personalized Federated Learning for Cross-view Geo-localization
by: Anagnostopoulos, Christos, et al.
Published: (2024)
by: Anagnostopoulos, Christos, et al.
Published: (2024)
FORGE: Fine-grained Multimodal Evaluation for Manufacturing Scenarios
by: Jian, Xiangru, et al.
Published: (2026)
by: Jian, Xiangru, et al.
Published: (2026)
Probing Deep into Temporal Profile Makes the Infrared Small Target Detector Much Better
by: Li, Ruojing, et al.
Published: (2025)
by: Li, Ruojing, et al.
Published: (2025)
Improving Cross-view Object Geo-localization: A Dual Attention Approach with Cross-view Interaction and Multi-Scale Spatial Features
by: Zhu, Xingtao Ling Yingying
Published: (2025)
by: Zhu, Xingtao Ling Yingying
Published: (2025)
Cross-view image geo-localization with Panorama-BEV Co-Retrieval Network
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
Multi-weather Cross-view Geo-localization Using Denoising Diffusion Models
by: Feng, Tongtong, et al.
Published: (2024)
by: Feng, Tongtong, et al.
Published: (2024)
Cross-Hierarchical Bidirectional Consistency Learning for Fine-Grained Visual Classification
by: Gao, Pengxiang, et al.
Published: (2025)
by: Gao, Pengxiang, et al.
Published: (2025)
OpenSatMap: A Fine-grained High-resolution Satellite Dataset for Large-scale Map Construction
by: Zhao, Hongbo, et al.
Published: (2024)
by: Zhao, Hongbo, et al.
Published: (2024)
NTIRE 2024 Challenge on Stereo Image Super-Resolution: Methods and Results
by: Wang, Longguang, et al.
Published: (2024)
by: Wang, Longguang, et al.
Published: (2024)
Dr.V: A Hierarchical Perception-Temporal-Cognition Framework to Diagnose Video Hallucination by Fine-grained Spatial-Temporal Grounding
by: Luo, Meng, et al.
Published: (2025)
by: Luo, Meng, et al.
Published: (2025)
CoVis: A Collaborative Framework for Fine-grained Graphic Visual Understanding
by: Deng, Xiaoyu, et al.
Published: (2024)
by: Deng, Xiaoyu, et al.
Published: (2024)
RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception
by: Zhu, Xiaosu, et al.
Published: (2024)
by: Zhu, Xiaosu, et al.
Published: (2024)
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
by: Zhang, Yancheng, et al.
Published: (2026)
by: Zhang, Yancheng, et al.
Published: (2026)
Unifying UAV Cross-View Geo-Localization via 3D Geometric Perception
by: Li, Haoyuan, et al.
Published: (2026)
by: Li, Haoyuan, et al.
Published: (2026)
Similar Items
-
PanMatch: Unleashing the Potential of Large Vision Models for Unified Matching Models
by: Zhang, Yongjian, et al.
Published: (2025) -
DropoutGS: Dropping Out Gaussians for Better Sparse-view Rendering
by: Xu, Yexing, et al.
Published: (2025) -
AIQViT: Architecture-Informed Post-Training Quantization for Vision Transformers
by: Jiang, Runqing, et al.
Published: (2025) -
Pluggable Style Representation Learning for Multi-Style Transfer
by: Liu, Hongda, et al.
Published: (2025) -
SaMam: Style-aware State Space Model for Arbitrary Image Style Transfer
by: Liu, Hongda, et al.
Published: (2025)