Zero-shot sketch-based remote sensing image retrieval based on multi-level and attention-guided tokenization
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Bo, Wang, Chen, Ma, Xiaoshuang, Song, Beiping, Liu, Zhuang, Sun, Fangde |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval
by: Su, Hanwen, et al.
Published: (2024)
by: Su, Hanwen, et al.
Published: (2024)
SAAN: Similarity-aware attention flow network for change detection with VHR remote sensing images
by: Guo, Haonan, et al.
Published: (2023)
by: Guo, Haonan, et al.
Published: (2023)
Deep learning-based interactive segmentation in remote sensing
by: Wang, Zhe, et al.
Published: (2023)
by: Wang, Zhe, et al.
Published: (2023)
Towards a multimodal framework for remote sensing image change retrieval and captioning
by: Ferrod, Roger, et al.
Published: (2024)
by: Ferrod, Roger, et al.
Published: (2024)
Fus-MAE: A cross-attention-based data fusion approach for Masked Autoencoders in remote sensing
by: Chan-To-Hing, Hugo, et al.
Published: (2024)
by: Chan-To-Hing, Hugo, et al.
Published: (2024)
Mask-guided cross-image attention for zero-shot in-silico histopathologic image generation with a diffusion model
by: Winter, Dominik, et al.
Published: (2024)
by: Winter, Dominik, et al.
Published: (2024)
OBSUM: An object-based spatial unmixing model for spatiotemporal fusion of remote sensing images
by: Guo, Houcai, et al.
Published: (2023)
by: Guo, Houcai, et al.
Published: (2023)
Dynamic Multi-level Weighted Alignment Network for Zero-shot Sketch-based Image Retrieval
by: Su, Hanwen, et al.
Published: (2025)
by: Su, Hanwen, et al.
Published: (2025)
MergeSAM: Unsupervised change detection of remote sensing images based on the Segment Anything Model
by: Hu, Meiqi, et al.
Published: (2025)
by: Hu, Meiqi, et al.
Published: (2025)
ZeroShape: Regression-based Zero-shot Shape Reconstruction
by: Huang, Zixuan, et al.
Published: (2023)
by: Huang, Zixuan, et al.
Published: (2023)
Few-shot multi-token DreamBooth with LoRa for style-consistent character generation
by: Pascual, Ruben, et al.
Published: (2025)
by: Pascual, Ruben, et al.
Published: (2025)
Toward Motion Robustness: A masked attention regularization framework in remote photoplethysmography
by: Zhao, Pengfei, et al.
Published: (2024)
by: Zhao, Pengfei, et al.
Published: (2024)
Leveraging feature communication in federated learning for remote sensing image classification
by: Duong, Anh-Kiet, et al.
Published: (2024)
by: Duong, Anh-Kiet, et al.
Published: (2024)
Zero-shot domain adaptation based on dual-level mix and contrast
by: Zhe, Yu, et al.
Published: (2024)
by: Zhe, Yu, et al.
Published: (2024)
Enhancing Zero-shot Counting via Language-guided Exemplar Learning
by: Wang, Mingjie, et al.
Published: (2024)
by: Wang, Mingjie, et al.
Published: (2024)
Learning Clustering-based Prototypes for Compositional Zero-shot Learning
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
by: Song, Jiawei, et al.
Published: (2024)
by: Song, Jiawei, et al.
Published: (2024)
Contrastive ground-level image and remote sensing pre-training improves representation learning for natural world imagery
by: Huynh, Andy V., et al.
Published: (2024)
by: Huynh, Andy V., et al.
Published: (2024)
LightFormer: A lightweight and efficient decoder for remote sensing image segmentation
by: Chen, Sihang, et al.
Published: (2025)
by: Chen, Sihang, et al.
Published: (2025)
GraVITON: Graph based garment warping with attention guided inversion for Virtual-tryon
by: Pathak, Sanhita, et al.
Published: (2024)
by: Pathak, Sanhita, et al.
Published: (2024)
Graph-guided Cross-composition Feature Disentanglement for Compositional Zero-shot Learning
by: Geng, Yuxia, et al.
Published: (2024)
by: Geng, Yuxia, et al.
Published: (2024)
Text-guided Synthetic Geometric Augmentation for Zero-shot 3D Understanding
by: Torimi, Kohei, et al.
Published: (2025)
by: Torimi, Kohei, et al.
Published: (2025)
Contextual Interaction via Primitive-based Adversarial Training For Compositional Zero-shot Learning
by: Li, Suyi, et al.
Published: (2024)
by: Li, Suyi, et al.
Published: (2024)
Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition
by: Rao, Mingxing, et al.
Published: (2024)
by: Rao, Mingxing, et al.
Published: (2024)
Compositional Zero-shot Learning via Progressive Language-based Observations
by: Li, Lin, et al.
Published: (2023)
by: Li, Lin, et al.
Published: (2023)
LOGCAN++: Adaptive Local-global class-aware network for semantic segmentation of remote sensing imagery
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
PUGS: Zero-shot Physical Understanding with Gaussian Splatting
by: Shuai, Yinghao, et al.
Published: (2025)
by: Shuai, Yinghao, et al.
Published: (2025)
A class-driven hierarchical ResNet for classification of multispectral remote sensing images
by: Weikmann, Giulio, et al.
Published: (2025)
by: Weikmann, Giulio, et al.
Published: (2025)
Leveraging knowledge distillation for partial multi-task learning from multiple remote sensing datasets
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
sketch2symm: Symmetry-aware sketch-to-shape generation via semantic bridging
by: Zhou, Yan, et al.
Published: (2025)
by: Zhou, Yan, et al.
Published: (2025)
Semi-supervised reference-based sketch extraction using a contrastive learning framework
by: Seo, Chang Wook, et al.
Published: (2024)
by: Seo, Chang Wook, et al.
Published: (2024)
VGDiffZero: Text-to-image Diffusion Models Can Be Zero-shot Visual Grounders
by: Liu, Xuyang, et al.
Published: (2023)
by: Liu, Xuyang, et al.
Published: (2023)
Sissi: Zero-shot Style-guided Image Synthesis via Semantic-style Integration
by: Deng, Yingying, et al.
Published: (2026)
by: Deng, Yingying, et al.
Published: (2026)
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
by: Wu, Pengying, et al.
Published: (2024)
by: Wu, Pengying, et al.
Published: (2024)
Joint multi-dimensional dynamic attention and transformer for general image restoration
by: Zhang, Huan, et al.
Published: (2024)
by: Zhang, Huan, et al.
Published: (2024)
Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval
by: Lyou, Eunyi, et al.
Published: (2024)
by: Lyou, Eunyi, et al.
Published: (2024)
Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
by: Du, Yongchao, et al.
Published: (2024)
by: Du, Yongchao, et al.
Published: (2024)
Zero-shot HOI Detection with MLLM-based Detector-agnostic Interaction Recognition
by: Xuan, Shiyu, et al.
Published: (2026)
by: Xuan, Shiyu, et al.
Published: (2026)
Multi-method Integration with Confidence-based Weighting for Zero-shot Image Classification
by: Yin, Siqi, et al.
Published: (2024)
by: Yin, Siqi, et al.
Published: (2024)
Gradient-based multi-focus image fusion with focus-aware saliency enhancement
by: Li, Haoyu, et al.
Published: (2025)
by: Li, Haoyu, et al.
Published: (2025)
Similar Items
-
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval
by: Su, Hanwen, et al.
Published: (2024) -
SAAN: Similarity-aware attention flow network for change detection with VHR remote sensing images
by: Guo, Haonan, et al.
Published: (2023) -
Deep learning-based interactive segmentation in remote sensing
by: Wang, Zhe, et al.
Published: (2023) -
Towards a multimodal framework for remote sensing image change retrieval and captioning
by: Ferrod, Roger, et al.
Published: (2024) -
Fus-MAE: A cross-attention-based data fusion approach for Masked Autoencoders in remote sensing
by: Chan-To-Hing, Hugo, et al.
Published: (2024)