UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhong, Siru, Hao, Xixuan, Yan, Yibo, Zhang, Ying, Song, Yangqiu, Liang, Yuxuan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction
by: Hao, Xixuan, et al.
Published: (2024)
by: Hao, Xixuan, et al.
Published: (2024)
Cross Space and Time: A Spatio-Temporal Unitized Model for Traffic Flow Forecasting
by: Ruan, Weilin, et al.
Published: (2024)
by: Ruan, Weilin, et al.
Published: (2024)
Vision-Enhanced Time Series Forecasting via Latent Diffusion Models
by: Ruan, Weilin, et al.
Published: (2025)
by: Ruan, Weilin, et al.
Published: (2025)
DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception
by: Man, Yunze, et al.
Published: (2023)
by: Man, Yunze, et al.
Published: (2023)
Graph-Based Cross-Domain Knowledge Distillation for Cross-Dataset Text-to-Image Person Retrieval
by: Luo, Bingjun, et al.
Published: (2025)
by: Luo, Bingjun, et al.
Published: (2025)
VimTS: A Unified Video and Image Text Spotter for Enhancing the Cross-domain Generalization
by: Liu, Yuliang, et al.
Published: (2024)
by: Liu, Yuliang, et al.
Published: (2024)
Mask-aware Text-to-Image Retrieval: Referring Expression Segmentation Meets Cross-modal Retrieval
by: Shen, Li-Cheng, et al.
Published: (2025)
by: Shen, Li-Cheng, et al.
Published: (2025)
Adaptive Domain Learning for Cross-domain Image Denoising
by: Qian, Zian, et al.
Published: (2024)
by: Qian, Zian, et al.
Published: (2024)
GEA: Generation-Enhanced Alignment for Text-to-Image Person Retrieval
by: Zou, Hao, et al.
Published: (2025)
by: Zou, Hao, et al.
Published: (2025)
ASR-enhanced Multimodal Representation Learning for Cross-Domain Product Retrieval
by: Zhao, Ruixiang, et al.
Published: (2024)
by: Zhao, Ruixiang, et al.
Published: (2024)
Analyzing the Impact of Low-Rank Adaptation for Cross-Domain Few-Shot Object Detection in Aerial Images
by: Talaoubrid, Hicham, et al.
Published: (2025)
by: Talaoubrid, Hicham, et al.
Published: (2025)
FunOTTA: On-the-Fly Adaptation on Cross-Domain Fundus Image via Stable Test-time Training
by: Zeng, Qian, et al.
Published: (2024)
by: Zeng, Qian, et al.
Published: (2024)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
by: Zhu, Mengdan, et al.
Published: (2025)
by: Zhu, Mengdan, et al.
Published: (2025)
Pseudo Label Refinery for Unsupervised Domain Adaptation on Cross-dataset 3D Object Detection
by: Zhang, Zhanwei, et al.
Published: (2024)
by: Zhang, Zhanwei, et al.
Published: (2024)
SynCDR : Training Cross Domain Retrieval Models with Synthetic Data
by: Mishra, Samarth, et al.
Published: (2023)
by: Mishra, Samarth, et al.
Published: (2023)
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
by: Li, Yuanzhe, et al.
Published: (2025)
by: Li, Yuanzhe, et al.
Published: (2025)
Unsupervised Domain Adaptation via Similarity-based Prototypes for Cross-Modality Segmentation
by: Ye, Ziyu, et al.
Published: (2025)
by: Ye, Ziyu, et al.
Published: (2025)
Text-Phase Synergy Network with Dual Priors for Unsupervised Cross-Domain Image Retrieval
by: Yang, Jing, et al.
Published: (2026)
by: Yang, Jing, et al.
Published: (2026)
Transcending Domains through Text-to-Image Diffusion: A Source-Free Approach to Domain Adaptation
by: Chopra, Shivang, et al.
Published: (2023)
by: Chopra, Shivang, et al.
Published: (2023)
FairDomain: Achieving Fairness in Cross-Domain Medical Image Segmentation and Classification
by: Tian, Yu, et al.
Published: (2024)
by: Tian, Yu, et al.
Published: (2024)
AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval
by: Wang, Yihan, et al.
Published: (2026)
by: Wang, Yihan, et al.
Published: (2026)
FBCIR: Balancing Cross-Modal Focuses in Composed Image Retrieval
by: Zhao, Chenchen, et al.
Published: (2026)
by: Zhao, Chenchen, et al.
Published: (2026)
Boosting Generalizability towards Zero-Shot Cross-Dataset Single-Image Indoor Depth by Meta-Initialization
by: Wu, Cho-Ying, et al.
Published: (2024)
by: Wu, Cho-Ying, et al.
Published: (2024)
TrueCity: Real and Simulated Urban Data for Cross-Domain 3D Scene Understanding
by: Nguyen, Duc, et al.
Published: (2025)
by: Nguyen, Duc, et al.
Published: (2025)
TCSA-UDA: Text-Driven Cross-Semantic Alignment for Unsupervised Domain Adaptation in Medical Image Segmentation
by: Maurya, Lalit, et al.
Published: (2025)
by: Maurya, Lalit, et al.
Published: (2025)
Masked Contrastive Reconstruction for Cross-modal Medical Image-Report Retrieval
by: Wei, Zeqiang, et al.
Published: (2023)
by: Wei, Zeqiang, et al.
Published: (2023)
Geospatial Representation Learning: A Survey from Deep Learning to The LLM Era
by: Hao, Xixuan, et al.
Published: (2025)
by: Hao, Xixuan, et al.
Published: (2025)
Multi-task Cross-modal Learning for Chest X-ray Image Retrieval
by: Liang, Zhaohui, et al.
Published: (2026)
by: Liang, Zhaohui, et al.
Published: (2026)
A Generalized Label Shift Perspective for Cross-Domain Gaze Estimation
by: Yang, Hao-Ran, et al.
Published: (2025)
by: Yang, Hao-Ran, et al.
Published: (2025)
Enhancing Whole Slide Image Classification through Supervised Contrastive Domain Adaptation
by: Carretero, Ilán, et al.
Published: (2024)
by: Carretero, Ilán, et al.
Published: (2024)
Cross-modal Full-mode Fine-grained Alignment for Text-to-Image Person Retrieval
by: Yin, Hao, et al.
Published: (2025)
by: Yin, Hao, et al.
Published: (2025)
Text or Image? What is More Important in Cross-Domain Generalization Capabilities of Hate Meme Detection Models?
by: Aggarwal, Piush, et al.
Published: (2024)
by: Aggarwal, Piush, et al.
Published: (2024)
Robust Building Damage Detection in Cross-Disaster Settings Using Domain Adaptation
by: Mouradi, Asmae, et al.
Published: (2026)
by: Mouradi, Asmae, et al.
Published: (2026)
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?
by: Chen, Yuxin, et al.
Published: (2024)
by: Chen, Yuxin, et al.
Published: (2024)
Rethink Predicting the Optical Flow with the Kinetics Perspective
by: Cheng, Yuhao, et al.
Published: (2024)
by: Cheng, Yuhao, et al.
Published: (2024)
RAAP: Retrieval-Augmented Affordance Prediction with Cross-Image Action Alignment
by: Zhuang, Qiyuan, et al.
Published: (2026)
by: Zhuang, Qiyuan, et al.
Published: (2026)
Unsupervised Cross-Domain Image Retrieval via Prototypical Optimal Transport
by: Li, Bin, et al.
Published: (2024)
by: Li, Bin, et al.
Published: (2024)
Caption-Matching: A Multimodal Approach for Cross-Domain Image Retrieval
by: Iijima, Lucas, et al.
Published: (2024)
by: Iijima, Lucas, et al.
Published: (2024)
Image Translation-Based Unsupervised Cross-Modality Domain Adaptation for Medical Image Segmentation
by: Yang, Tao, et al.
Published: (2025)
by: Yang, Tao, et al.
Published: (2025)
Cross-Modal Adapter for Vision-Language Retrieval
by: Jiang, Haojun, et al.
Published: (2022)
by: Jiang, Haojun, et al.
Published: (2022)
Similar Items
-
UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction
by: Hao, Xixuan, et al.
Published: (2024) -
Cross Space and Time: A Spatio-Temporal Unitized Model for Traffic Flow Forecasting
by: Ruan, Weilin, et al.
Published: (2024) -
Vision-Enhanced Time Series Forecasting via Latent Diffusion Models
by: Ruan, Weilin, et al.
Published: (2025) -
DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception
by: Man, Yunze, et al.
Published: (2023) -
Graph-Based Cross-Domain Knowledge Distillation for Cross-Dataset Text-to-Image Person Retrieval
by: Luo, Bingjun, et al.
Published: (2025)