SenCLIP: Enhancing zero-shot land-use mapping for Sentinel-2 with ground-level prompting
Fuente:
arXiv
Saved in:
| Main Authors: | Jain, Pallavi, Ienco, Dino, Interdonato, Roberto, Berchoux, Tristan, Marcos, Diego |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TimeSenCLIP: A Time Series Vision-Language Model for Remote Sensing
by: Jain, Pallavi, et al.
Published: (2025)
by: Jain, Pallavi, et al.
Published: (2025)
Multi-modal Co-learning for Earth Observation: Enhancing single-modality models via modality collaboration
by: Mena, Francisco, et al.
Published: (2025)
by: Mena, Francisco, et al.
Published: (2025)
VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation
by: Qu, Zhen, et al.
Published: (2024)
by: Qu, Zhen, et al.
Published: (2024)
Two-stage Vision Transformers and Hard Masking offer Robust Object Representations
by: Aniraj, Ananthu, et al.
Published: (2025)
by: Aniraj, Ananthu, et al.
Published: (2025)
PDiscoFormer: Relaxing Part Discovery Constraints with Vision Transformers
by: Aniraj, Ananthu, et al.
Published: (2024)
by: Aniraj, Ananthu, et al.
Published: (2024)
Metonymy in vision models undermines attention-based interpretability
by: Aniraj, Ananthu, et al.
Published: (2026)
by: Aniraj, Ananthu, et al.
Published: (2026)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
by: Ienco, Dino, et al.
Published: (2024)
by: Ienco, Dino, et al.
Published: (2024)
Towards a multimodal framework for remote sensing image change retrieval and captioning
by: Ferrod, Roger, et al.
Published: (2024)
by: Ferrod, Roger, et al.
Published: (2024)
Atomizer: Generalizing to new modalities by breaking satellite images down to a set of scalars
by: de Turckheim, Hugo Riffaud, et al.
Published: (2025)
by: de Turckheim, Hugo Riffaud, et al.
Published: (2025)
InfraDiffusion: zero-shot depth map restoration with diffusion models and prompted segmentation from sparse infrastructure point clouds
by: Jing, Yixiong, et al.
Published: (2025)
by: Jing, Yixiong, et al.
Published: (2025)
RAMEN: Resolution-Adjustable Multimodal Encoder for Earth Observation
by: Houdré, Nicolas, et al.
Published: (2025)
by: Houdré, Nicolas, et al.
Published: (2025)
Revisiting Cross-Modal Knowledge Distillation: A Disentanglement Approach for RGBD Semantic Segmentation
by: Ferrod, Roger, et al.
Published: (2025)
by: Ferrod, Roger, et al.
Published: (2025)
Sen2Fire: A Challenging Benchmark Dataset for Wildfire Detection using Sentinel Data
by: Xu, Yonghao, et al.
Published: (2024)
by: Xu, Yonghao, et al.
Published: (2024)
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
EcoWikiRS: Learning Ecological Representation of Satellite Images from Weak Supervision with Species Observations and Wikipedia
by: Zermatten, Valerie, et al.
Published: (2025)
by: Zermatten, Valerie, et al.
Published: (2025)
COOkeD: Ensemble-based OOD detection in the era of zero-shot CLIP
by: Humblot-Renaux, Galadrielle, et al.
Published: (2025)
by: Humblot-Renaux, Galadrielle, et al.
Published: (2025)
CLIP4Sketch: Enhancing Sketch to Mugshot Matching through Dataset Augmentation using Diffusion Models
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
by: Shiri, Mahshid, et al.
Published: (2025)
by: Shiri, Mahshid, et al.
Published: (2025)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
by: Ma, Wenxin, et al.
Published: (2025)
by: Ma, Wenxin, et al.
Published: (2025)
Geographical Context Matters: Bridging Fine and Coarse Spatial Information to Enhance Continental Land Cover Mapping
by: Ghassemi, Babak, et al.
Published: (2025)
by: Ghassemi, Babak, et al.
Published: (2025)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
by: Kim, Donghyeong, et al.
Published: (2025)
by: Kim, Donghyeong, et al.
Published: (2025)
MediCLIP: Adapting CLIP for Few-shot Medical Image Anomaly Detection
by: Zhang, Ximiao, et al.
Published: (2024)
by: Zhang, Ximiao, et al.
Published: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
by: Xie, Jingyou, et al.
Published: (2024)
by: Xie, Jingyou, et al.
Published: (2024)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
FADE: Few-shot/zero-shot Anomaly Detection Engine using Large Vision-Language Model
by: Li, Yuanwei, et al.
Published: (2024)
by: Li, Yuanwei, et al.
Published: (2024)
SAM-MPA: Applying SAM to Few-shot Medical Image Segmentation using Mask Propagation and Auto-prompting
by: Xu, Jie, et al.
Published: (2024)
by: Xu, Jie, et al.
Published: (2024)
CLIP-driven Zero-shot Learning with Ambiguous Labels
by: Fan, Jinfu, et al.
Published: (2026)
by: Fan, Jinfu, et al.
Published: (2026)
Explaining CLIP Zero-shot Predictions Through Concepts
by: Ozdemir, Onat, et al.
Published: (2026)
by: Ozdemir, Onat, et al.
Published: (2026)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
by: Pang, Li, et al.
Published: (2025)
by: Pang, Li, et al.
Published: (2025)
CardiacCLIP: Video-based CLIP Adaptation for LVEF Prediction in a Few-shot Manner
by: Du, Yao, et al.
Published: (2025)
by: Du, Yao, et al.
Published: (2025)
CLIP's Visual Embedding Projector is a Few-shot Cornucopia
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
DiffCLIP: Few-shot Language-driven Multimodal Classifier
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023)
by: Wang, Xiang, et al.
Published: (2023)
CLIP is Strong Enough to Fight Back: Test-time Counterattacks towards Zero-shot Adversarial Robustness of CLIP
by: Xing, Songlong, et al.
Published: (2025)
by: Xing, Songlong, et al.
Published: (2025)
Noise-aware few-shot learning through bi-directional multi-view prompt alignment
by: Niu, Lu, et al.
Published: (2026)
by: Niu, Lu, et al.
Published: (2026)
EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
by: Duan, Zicheng, et al.
Published: (2024)
by: Duan, Zicheng, et al.
Published: (2024)
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
by: Zenotto, Andrea, et al.
Published: (2026)
by: Zenotto, Andrea, et al.
Published: (2026)
MulCLIP: A Multi-level Alignment Framework for Enhancing Fine-grained Long-context CLIP
by: Truong, Chau, et al.
Published: (2025)
by: Truong, Chau, et al.
Published: (2025)
Partial CLIP is Enough: Chimera-Seg for Zero-shot Semantic Segmentation
by: Chen, Jialei, et al.
Published: (2025)
by: Chen, Jialei, et al.
Published: (2025)
Adaptive Multi-prompt Contrastive Network for Few-shot Out-of-distribution Detection
by: Fang, Xiang, et al.
Published: (2025)
by: Fang, Xiang, et al.
Published: (2025)
Similar Items
-
TimeSenCLIP: A Time Series Vision-Language Model for Remote Sensing
by: Jain, Pallavi, et al.
Published: (2025) -
Multi-modal Co-learning for Earth Observation: Enhancing single-modality models via modality collaboration
by: Mena, Francisco, et al.
Published: (2025) -
VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation
by: Qu, Zhen, et al.
Published: (2024) -
Two-stage Vision Transformers and Hard Masking offer Robust Object Representations
by: Aniraj, Ananthu, et al.
Published: (2025) -
PDiscoFormer: Relaxing Part Discovery Constraints with Vision Transformers
by: Aniraj, Ananthu, et al.
Published: (2024)