Cross-view image geo-localization with Panorama-BEV Co-Retrieval Network
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ye, Junyan, Lv, Zhutao, Li, Weijia, Yu, Jinhua, Yang, Haote, Zhong, Huaping, He, Conghui |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Leveraging BEV Paradigm for Ground-to-Aerial Image Synthesis
par: Ye, Junyan, et autres
Publié: (2024)
par: Ye, Junyan, et autres
Publié: (2024)
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
par: Ye, Junyan, et autres
Publié: (2024)
par: Ye, Junyan, et autres
Publié: (2024)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
par: Li, Weijia, et autres
Publié: (2024)
par: Li, Weijia, et autres
Publié: (2024)
Earth-Agent: Unlocking the Full Landscape of Earth Observation with Agents
par: Feng, Peilin, et autres
Publié: (2025)
par: Feng, Peilin, et autres
Publié: (2025)
UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban Scenarios
par: Zhou, Baichuan, et autres
Publié: (2024)
par: Zhou, Baichuan, et autres
Publié: (2024)
Where am I? Cross-View Geo-localization with Natural Language Descriptions
par: Ye, Junyan, et autres
Publié: (2024)
par: Ye, Junyan, et autres
Publié: (2024)
FakeVLM-R1: Internalizing Physical Laws via CoT for Synthetic Image Detection
par: Zhu, Leqi, et autres
Publié: (2026)
par: Zhu, Leqi, et autres
Publié: (2026)
Cross-view geo-localization: a survey
par: Durgam, Abhilash, et autres
Publié: (2024)
par: Durgam, Abhilash, et autres
Publié: (2024)
3D Building Reconstruction from Monocular Remote Sensing Images with Multi-level Supervisions
par: Li, Weijia, et autres
Publié: (2024)
par: Li, Weijia, et autres
Publié: (2024)
OmniAID: Decoupling Semantic and Artifacts for Universal AI-Generated Image Detection in the Wild
par: Guo, Yuncheng, et autres
Publié: (2025)
par: Guo, Yuncheng, et autres
Publié: (2025)
Cross-view geo-localization, Image retrieval, Multiscale geometric modeling, Frequency domain enhancement
par: Zhang, Hongying, et autres
Publié: (2026)
par: Zhang, Hongying, et autres
Publié: (2026)
AgenticOCR: Parsing Only What You Need for Efficient Retrieval-Augmented Generation
par: Wang, Zhengren, et autres
Publié: (2026)
par: Wang, Zhengren, et autres
Publié: (2026)
BLINK-Twice: You see, but do you observe? A Reasoning Benchmark on Visual Perception
par: Ye, Junyan, et autres
Publié: (2025)
par: Ye, Junyan, et autres
Publié: (2025)
Fine-Grained Building Function Recognition from Street-View Images via Geometry-Aware Semi-Supervised Learning
par: Li, Weijia, et autres
Publié: (2024)
par: Li, Weijia, et autres
Publié: (2024)
GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image Generation
par: Yan, Zhiyuan, et autres
Publié: (2025)
par: Yan, Zhiyuan, et autres
Publié: (2025)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
par: Zhang, Jinqing, et autres
Publié: (2024)
par: Zhang, Jinqing, et autres
Publié: (2024)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
par: Tang, Tao, et autres
Publié: (2024)
par: Tang, Tao, et autres
Publié: (2024)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
par: Xu, Shaoqing, et autres
Publié: (2024)
par: Xu, Shaoqing, et autres
Publié: (2024)
Spot the Fake: Large Multimodal Model-Based Synthetic Image Detection with Artifact Explanation
par: Wen, Siwei, et autres
Publié: (2025)
par: Wen, Siwei, et autres
Publié: (2025)
GenClaw: Code-Driven Agentic Image Generation
par: Ye, Junyan, et autres
Publié: (2026)
par: Ye, Junyan, et autres
Publié: (2026)
Retrieval-guided Cross-view Image Synthesis
par: Yang, Hongji, et autres
Publié: (2024)
par: Yang, Hongji, et autres
Publié: (2024)
LEGION: Learning to Ground and Explain for Synthetic Image Detection
par: Kang, Hengrui, et autres
Publié: (2025)
par: Kang, Hengrui, et autres
Publié: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
par: Ye, Junyan, et autres
Publié: (2025)
par: Ye, Junyan, et autres
Publié: (2025)
VIGC: Visual Instruction Generation and Correction
par: Wang, Bin, et autres
Publié: (2023)
par: Wang, Bin, et autres
Publié: (2023)
JRN-Geo: A Joint Perception Network based on RGB and Normal images for Cross-view Geo-localization
par: Zhou, Hongyu, et autres
Publié: (2025)
par: Zhou, Hongyu, et autres
Publié: (2025)
Learning Cross-view Visual Geo-localization without Ground Truth
par: Li, Haoyuan, et autres
Publié: (2024)
par: Li, Haoyuan, et autres
Publié: (2024)
CLAIR: CLIP-Aided Weakly Supervised Zero-Shot Cross-Domain Image Retrieval
par: Tan, Chor Boon, et autres
Publié: (2025)
par: Tan, Chor Boon, et autres
Publié: (2025)
Window-to-Window BEV Representation Learning for Limited FoV Cross-View Geo-localization
par: Cheng, Lei, et autres
Publié: (2024)
par: Cheng, Lei, et autres
Publié: (2024)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
par: Chen, Zeming, et autres
Publié: (2025)
par: Chen, Zeming, et autres
Publié: (2025)
DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception
par: Man, Yunze, et autres
Publié: (2023)
par: Man, Yunze, et autres
Publié: (2023)
MVPbev: Multi-view Perspective Image Generation from BEV with Test-time Controllability and Generalizability
par: Liu, Buyu, et autres
Publié: (2024)
par: Liu, Buyu, et autres
Publié: (2024)
Scene4U: Hierarchical Layered 3D Scene Reconstruction from Single Panoramic Image for Your Immerse Exploration
par: Huang, Zilong, et autres
Publié: (2025)
par: Huang, Zilong, et autres
Publié: (2025)
FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection
par: Jiang, Zheng, et autres
Publié: (2024)
par: Jiang, Zheng, et autres
Publié: (2024)
Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem?
par: Wen, Zichen, et autres
Publié: (2025)
par: Wen, Zichen, et autres
Publié: (2025)
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
par: Ye, Junyan, et autres
Publié: (2025)
par: Ye, Junyan, et autres
Publié: (2025)
Parallel Cross Strip Attention Network for Single Image Dehazing
par: Tong, Lihan, et autres
Publié: (2024)
par: Tong, Lihan, et autres
Publié: (2024)
SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors
par: Fan, Ruijie, et autres
Publié: (2025)
par: Fan, Ruijie, et autres
Publié: (2025)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
par: Zhao, Xiao, et autres
Publié: (2024)
par: Zhao, Xiao, et autres
Publié: (2024)
PM4Bench: Benchmarking Large Vision-Language Models with Parallel Multilingual Multi-Modal Multi-task Corpus
par: Gao, Junyuan, et autres
Publié: (2025)
par: Gao, Junyuan, et autres
Publié: (2025)
Anchor-free Cross-view Object Geo-localization with Gaussian Position Encoding and Cross-view Association
par: Ling, Xingtao, et autres
Publié: (2025)
par: Ling, Xingtao, et autres
Publié: (2025)
Documents similaires
-
Leveraging BEV Paradigm for Ground-to-Aerial Image Synthesis
par: Ye, Junyan, et autres
Publié: (2024) -
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
par: Ye, Junyan, et autres
Publié: (2024) -
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
par: Li, Weijia, et autres
Publié: (2024) -
Earth-Agent: Unlocking the Full Landscape of Earth Observation with Agents
par: Feng, Peilin, et autres
Publié: (2025) -
UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban Scenarios
par: Zhou, Baichuan, et autres
Publié: (2024)