SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xue, Cao, Shengting, Li, Shenglin, Gong, Jiaqi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SatSwinMAE: Efficient Autoencoding for Multiscale Time-series Satellite Imagery
by: Nakayama, Yohei, et al.
Published: (2024)
by: Nakayama, Yohei, et al.
Published: (2024)
Sat2Flow: A Structure-Aware Diffusion Framework for Human Flow Generation from Satellite Imagery
by: Wang, Xiangxu, et al.
Published: (2025)
by: Wang, Xiangxu, et al.
Published: (2025)
DiffusionSat: A Generative Foundation Model for Satellite Imagery
by: Khanna, Samar, et al.
Published: (2023)
by: Khanna, Samar, et al.
Published: (2023)
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes
by: Liu, Jiaqi, et al.
Published: (2025)
by: Liu, Jiaqi, et al.
Published: (2025)
SatVision-TOA: A Geospatial Foundation Model for Coarse-Resolution All-Sky Remote Sensing Imagery
by: Spradlin, Caleb S., et al.
Published: (2024)
by: Spradlin, Caleb S., et al.
Published: (2024)
SatCLIP: Global, General-Purpose Location Embeddings with Satellite Imagery
by: Klemmer, Konstantin, et al.
Published: (2023)
by: Klemmer, Konstantin, et al.
Published: (2023)
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
by: Ma, Qiwei, et al.
Published: (2025)
by: Ma, Qiwei, et al.
Published: (2025)
In-Context Learning Improves Compositional Understanding of Vision-Language Models
by: Nulli, Matteo, et al.
Published: (2024)
by: Nulli, Matteo, et al.
Published: (2024)
Satellite-to-Street: Synthesizing Post-Disaster Views from Satellite Imagery via Generative Vision Models
by: Yang, Yifan, et al.
Published: (2026)
by: Yang, Yifan, et al.
Published: (2026)
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
by: Limbu, Manshi, et al.
Published: (2025)
by: Limbu, Manshi, et al.
Published: (2025)
BLIP3-KALE: Knowledge Augmented Large-Scale Dense Captions
by: Awadalla, Anas, et al.
Published: (2024)
by: Awadalla, Anas, et al.
Published: (2024)
Lightweight Multimodal Adaptation of Vision Language Models for Species Recognition and Habitat Context Interpretation in Drone Thermal Imagery
by: Chen, Hao, et al.
Published: (2026)
by: Chen, Hao, et al.
Published: (2026)
Deep Learning for Pavement Condition Evaluation Using Satellite Imagery
by: Lebaku, Prathyush Kumar Reddy, et al.
Published: (2025)
by: Lebaku, Prathyush Kumar Reddy, et al.
Published: (2025)
Vision-Aware Text Features in Referring Image Segmentation: From Object Understanding to Context Understanding
by: Nguyen-Truong, Hai, et al.
Published: (2024)
by: Nguyen-Truong, Hai, et al.
Published: (2024)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
by: Luo, Junwei, et al.
Published: (2025)
by: Luo, Junwei, et al.
Published: (2025)
Position Prediction Self-Supervised Learning for Multimodal Satellite Imagery Semantic Segmentation
by: Waithaka, John, et al.
Published: (2025)
by: Waithaka, John, et al.
Published: (2025)
Reframing Image Difference Captioning with BLIP2IDC and Synthetic Augmentation
by: Evennou, Gautier, et al.
Published: (2024)
by: Evennou, Gautier, et al.
Published: (2024)
Landsat30-AU: A Vision-Language Dataset for Australian Landsat Imagery
by: Ma, Sai, et al.
Published: (2025)
by: Ma, Sai, et al.
Published: (2025)
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
by: Chen, Jiuhai, et al.
Published: (2025)
by: Chen, Jiuhai, et al.
Published: (2025)
RAU: Reference-based Anatomical Understanding with Vision Language Models
by: Li, Yiwei, et al.
Published: (2025)
by: Li, Yiwei, et al.
Published: (2025)
Bootstrapping Rare Object Detection in High-Resolution Satellite Imagery
by: Zaytar, Akram, et al.
Published: (2024)
by: Zaytar, Akram, et al.
Published: (2024)
M4-BLIP: Advancing Multi-Modal Media Manipulation Detection through Face-Enhanced Local Analysis
by: Wu, Hang, et al.
Published: (2025)
by: Wu, Hang, et al.
Published: (2025)
CLIP-based Camera-Agnostic Feature Learning for Intra-camera Person Re-Identification
by: Tan, Xuan, et al.
Published: (2024)
by: Tan, Xuan, et al.
Published: (2024)
BLIP-FusePPO: A Vision-Language Deep Reinforcement Learning Framework for Lane Keeping in Autonomous Vehicles
by: Miangoleh, Seyed Ahmad Hosseini, et al.
Published: (2025)
by: Miangoleh, Seyed Ahmad Hosseini, et al.
Published: (2025)
Abn-BLIP: Abnormality-aligned Bootstrapping Language-Image Pre-training for Pulmonary Embolism Diagnosis and Report Generation from CTPA
by: Zhong, Zhusi, et al.
Published: (2025)
by: Zhong, Zhusi, et al.
Published: (2025)
An Empirical Study of Methods for Small Object Detection from Satellite Imagery
by: Yuan, Xiaohui, et al.
Published: (2025)
by: Yuan, Xiaohui, et al.
Published: (2025)
SAR-AE-SFP: SAR Imagery Adversarial Example in Real Physics domain with Target Scattering Feature Parameters
by: Cui, Jiahao, et al.
Published: (2024)
by: Cui, Jiahao, et al.
Published: (2024)
STAR: A First-Ever Dataset and A Large-Scale Benchmark for Scene Graph Generation in Large-Size Satellite Imagery
by: Li, Yansheng, et al.
Published: (2024)
by: Li, Yansheng, et al.
Published: (2024)
SD-VLM: Spatial Measuring and Understanding with Depth-Encoded Vision-Language Models
by: Chen, Pingyi, et al.
Published: (2025)
by: Chen, Pingyi, et al.
Published: (2025)
AllClear: A Comprehensive Dataset and Benchmark for Cloud Removal in Satellite Imagery
by: Zhou, Hangyu, et al.
Published: (2024)
by: Zhou, Hangyu, et al.
Published: (2024)
TS-SatFire: A Multi-Task Satellite Image Time-Series Dataset for Wildfire Detection and Prediction
by: Zhao, Yu, et al.
Published: (2024)
by: Zhao, Yu, et al.
Published: (2024)
Sat3DGen: Comprehensive Street-Level 3D Scene Generation from Single Satellite Image
by: Qian, Ming, et al.
Published: (2026)
by: Qian, Ming, et al.
Published: (2026)
Understanding Cross Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
by: Gong, Changqing, et al.
Published: (2025)
by: Gong, Changqing, et al.
Published: (2025)
IGraSS: Learning to Identify Infrastructure Networks from Satellite Imagery by Iterative Graph-constrained Semantic Segmentation
by: Hoque, Oishee Bintey, et al.
Published: (2025)
by: Hoque, Oishee Bintey, et al.
Published: (2025)
BehaviorVLM: Unified Finetuning-Free Behavioral Understanding with Vision-Language Reasoning
by: Ke, Jingyang, et al.
Published: (2026)
by: Ke, Jingyang, et al.
Published: (2026)
From Pixels to Progress: Generating Road Network from Satellite Imagery for Socioeconomic Insights in Impoverished Areas
by: Xi, Yanxin, et al.
Published: (2024)
by: Xi, Yanxin, et al.
Published: (2024)
Think or Not? Selective Reasoning via Reinforcement Learning for Vision-Language Models
by: Wang, Jiaqi, et al.
Published: (2025)
by: Wang, Jiaqi, et al.
Published: (2025)
Feature Identification for Hierarchical Contrastive Learning
by: Ott, Julius, et al.
Published: (2025)
by: Ott, Julius, et al.
Published: (2025)
BiasICL: In-Context Learning and Demographic Biases of Vision Language Models
by: Xu, Sonnet, et al.
Published: (2025)
by: Xu, Sonnet, et al.
Published: (2025)
Generating Satellite Imagery Data for Wildfire Detection through Mask-Conditioned Generative AI
by: Martin, Valeria, et al.
Published: (2026)
by: Martin, Valeria, et al.
Published: (2026)
Similar Items
-
SatSwinMAE: Efficient Autoencoding for Multiscale Time-series Satellite Imagery
by: Nakayama, Yohei, et al.
Published: (2024) -
Sat2Flow: A Structure-Aware Diffusion Framework for Human Flow Generation from Satellite Imagery
by: Wang, Xiangxu, et al.
Published: (2025) -
DiffusionSat: A Generative Foundation Model for Satellite Imagery
by: Khanna, Samar, et al.
Published: (2023) -
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes
by: Liu, Jiaqi, et al.
Published: (2025) -
SatVision-TOA: A Geospatial Foundation Model for Coarse-Resolution All-Sky Remote Sensing Imagery
by: Spradlin, Caleb S., et al.
Published: (2024)