Multiple-environment Self-adaptive Network for Aerial-view Geo-localization
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Tingyu, Zheng, Zhedong, Sun, Yaoqi, Yan, Chenggang, Yang, Yi, Chua, Tat-Seng |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching
by: Chu, Meng, et al.
Published: (2023)
by: Chu, Meng, et al.
Published: (2023)
3D Magic Mirror: Clothing Reconstruction from a Single Image via a Causal Perspective
by: Zheng, Zhedong, et al.
Published: (2022)
by: Zheng, Zhedong, et al.
Published: (2022)
Scale-adaptive UAV Geo-localization via Height-aware Partition Learning
by: Chen, Quan, et al.
Published: (2024)
by: Chen, Quan, et al.
Published: (2024)
SDPL: Shifting-Dense Partition Learning for UAV-View Geo-Localization
by: Chen, Quan, et al.
Published: (2024)
by: Chen, Quan, et al.
Published: (2024)
Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization
by: Chen, Yiyang, et al.
Published: (2022)
by: Chen, Yiyang, et al.
Published: (2022)
Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse
by: Fang, Yunsong, et al.
Published: (2026)
by: Fang, Yunsong, et al.
Published: (2026)
Salient Object Detection in Complex Weather Conditions via Noise Indicators
by: Chen, Quan, et al.
Published: (2025)
by: Chen, Quan, et al.
Published: (2025)
Instilling Multi-round Thinking to Text-guided Image Generation
by: Zeng, Lidong, et al.
Published: (2024)
by: Zeng, Lidong, et al.
Published: (2024)
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
by: Jin, Zhe, et al.
Published: (2025)
by: Jin, Zhe, et al.
Published: (2025)
StepNet: Spatial-temporal Part-aware Network for Isolated Sign Language Recognition
by: Shen, Xiaolong, et al.
Published: (2022)
by: Shen, Xiaolong, et al.
Published: (2022)
Progressive Depth Decoupling and Modulating for Flexible Depth Completion
by: Yang, Zhiwen, et al.
Published: (2024)
by: Yang, Zhiwen, et al.
Published: (2024)
Universal Scene Graph Generation
by: Wu, Shengqiong, et al.
Published: (2025)
by: Wu, Shengqiong, et al.
Published: (2025)
VGNC: Reducing the Overfitting of Sparse-view 3DGS via Validation-guided Gaussian Number Control
by: Lin, Lifeng, et al.
Published: (2025)
by: Lin, Lifeng, et al.
Published: (2025)
PiPa++: Towards Unification of Domain Adaptive Semantic Segmentation via Self-supervised Learning
by: Chen, Mu, et al.
Published: (2024)
by: Chen, Mu, et al.
Published: (2024)
EQ-TAA: Equivariant Traffic Accident Anticipation via Diffusion-Based Accident Video Synthesis
by: Fang, Jianwu, et al.
Published: (2025)
by: Fang, Jianwu, et al.
Published: (2025)
Video2BEV: Transforming Drone Videos to BEVs for Video-based Geo-localization
by: Ju, Hao, et al.
Published: (2024)
by: Ju, Hao, et al.
Published: (2024)
Quality-aware Selective Fusion Network for V-D-T Salient Object Detection
by: Bao, Liuxin, et al.
Published: (2024)
by: Bao, Liuxin, et al.
Published: (2024)
Understanding Long Videos via LLM-Powered Entity Relation Graphs
by: Chu, Meng, et al.
Published: (2025)
by: Chu, Meng, et al.
Published: (2025)
Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation
by: Chen, Mu, et al.
Published: (2023)
by: Chen, Mu, et al.
Published: (2023)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
ITS3D: Inference-Time Scaling for Text-Guided 3D Diffusion Models
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
Multi-Granularity Class Prototype Topology Distillation for Class-Incremental Source-Free Unsupervised Domain Adaptation
by: Deng, Peihua, et al.
Published: (2024)
by: Deng, Peihua, et al.
Published: (2024)
Few-Shot Generative Model Adaption via Identity Injection and Preservation
by: He, Yeqi, et al.
Published: (2026)
by: He, Yeqi, et al.
Published: (2026)
Disentangling Masked Autoencoders for Unsupervised Domain Generalization
by: Zhang, An, et al.
Published: (2024)
by: Zhang, An, et al.
Published: (2024)
Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
Extending Visual Dynamics for Video-to-Music Generation
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
Combating Multimodal LLM Hallucination via Bottom-Up Holistic Reasoning
by: Wu, Shengqiong, et al.
Published: (2024)
by: Wu, Shengqiong, et al.
Published: (2024)
Global Commander and Local Operative: A Dual-Agent Framework for Scene Navigation
by: Jin, Kaiming, et al.
Published: (2026)
by: Jin, Kaiming, et al.
Published: (2026)
Exploring the Impact of Synthetic Data for Aerial-view Human Detection
by: Lee, Hyungtae, et al.
Published: (2024)
by: Lee, Hyungtae, et al.
Published: (2024)
Thinking with Blueprints: Assisting Vision-Language Models in Spatial Reasoning via Structured Object Representation
by: Ma, Weijian, et al.
Published: (2026)
by: Ma, Weijian, et al.
Published: (2026)
AnchorFlow: Training-Free 3D Editing via Latent Anchor-Aligned Flows
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
Collaborative Group: Composed Image Retrieval via Consensus Learning from Noisy Annotations
by: Zhang, Xu, et al.
Published: (2023)
by: Zhang, Xu, et al.
Published: (2023)
Harnessing Weak Pair Uncertainty for Text-based Person Search
by: Sun, Jintao, et al.
Published: (2026)
by: Sun, Jintao, et al.
Published: (2026)
WeatherPrompt: Multi-modality Representation Learning for All-Weather Drone Visual Geo-Localization
by: Wen, Jiahao, et al.
Published: (2025)
by: Wen, Jiahao, et al.
Published: (2025)
Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
Can I Trust Your Answer? Visually Grounded Video Question Answering
by: Xiao, Junbin, et al.
Published: (2023)
by: Xiao, Junbin, et al.
Published: (2023)
Towards Semantic Equivalence of Tokenization in Multimodal LLM
by: Wu, Shengqiong, et al.
Published: (2024)
by: Wu, Shengqiong, et al.
Published: (2024)
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
SeCap: Self-Calibrating and Adaptive Prompts for Cross-view Person Re-Identification in Aerial-Ground Networks
by: Wang, Shining, et al.
Published: (2025)
by: Wang, Shining, et al.
Published: (2025)
Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models
by: Liang, Xiao, et al.
Published: (2025)
by: Liang, Xiao, et al.
Published: (2025)
Similar Items
-
Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching
by: Chu, Meng, et al.
Published: (2023) -
3D Magic Mirror: Clothing Reconstruction from a Single Image via a Causal Perspective
by: Zheng, Zhedong, et al.
Published: (2022) -
Scale-adaptive UAV Geo-localization via Height-aware Partition Learning
by: Chen, Quan, et al.
Published: (2024) -
SDPL: Shifting-Dense Partition Learning for UAV-View Geo-Localization
by: Chen, Quan, et al.
Published: (2024) -
Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization
by: Chen, Yiyang, et al.
Published: (2022)