FUSAR-KLIP: Towards Multimodal Foundation Models for Remote Sensing
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yi, Zhang, Xiaokun, Fang, Qingchen, Liu, Jing, Ye, Ziqi, Li, Rui, Liu, Li, Wang, Haipeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FUSAR-GPT : A Spatiotemporal Feature-Embedded and Two-Stage Decoupled Visual Language Model for SAR Imagery
by: Zhang, Xiaokun, et al.
Published: (2026)
by: Zhang, Xiaokun, et al.
Published: (2026)
Object Fidelity Diffusion for Remote Sensing Image Generation
by: Ye, Ziqi, et al.
Published: (2025)
by: Ye, Ziqi, et al.
Published: (2025)
A Survey on Remote Sensing Foundation Models: From Vision to Multimodality
by: Huang, Ziyue, et al.
Published: (2025)
by: Huang, Ziyue, et al.
Published: (2025)
VFM-ISRefiner: Towards Better Adapting Vision Foundation Models for Interactive Segmentation of Remote Sensing Images
by: Wang, Deliang, et al.
Published: (2025)
by: Wang, Deliang, et al.
Published: (2025)
OpenRSD: Towards Open-prompts for Object Detection in Remote Sensing Images
by: Huang, Ziyue, et al.
Published: (2025)
by: Huang, Ziyue, et al.
Published: (2025)
SpectralGPT: Spectral Remote Sensing Foundation Model
by: Hong, Danfeng, et al.
Published: (2023)
by: Hong, Danfeng, et al.
Published: (2023)
Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
by: Hong, Danfeng, et al.
Published: (2026)
by: Hong, Danfeng, et al.
Published: (2026)
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
by: Liu, Fan, et al.
Published: (2023)
by: Liu, Fan, et al.
Published: (2023)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
by: Liu, Yi, et al.
Published: (2026)
by: Liu, Yi, et al.
Published: (2026)
RS-vHeat: Heat Conduction Guided Efficient Remote Sensing Foundation Model
by: Hu, Huiyang, et al.
Published: (2024)
by: Hu, Huiyang, et al.
Published: (2024)
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining
by: Wang, Di, et al.
Published: (2024)
by: Wang, Di, et al.
Published: (2024)
LSKNet: A Foundation Lightweight Backbone for Remote Sensing
by: Li, Yuxuan, et al.
Published: (2024)
by: Li, Yuxuan, et al.
Published: (2024)
SeaMo: A Season-Aware Multimodal Foundation Model for Remote Sensing
by: Li, Xuyang, et al.
Published: (2024)
by: Li, Xuyang, et al.
Published: (2024)
Falcon: A Remote Sensing Vision-Language Foundation Model (Technical Report)
by: Yao, Kelu, et al.
Published: (2025)
by: Yao, Kelu, et al.
Published: (2025)
Towards Privacy-preserved Pre-training of Remote Sensing Foundation Models with Federated Mutual-guidance Learning
by: Tan, Jieyi, et al.
Published: (2025)
by: Tan, Jieyi, et al.
Published: (2025)
SkySense: A Multi-Modal Remote Sensing Foundation Model Towards Universal Interpretation for Earth Observation Imagery
by: Guo, Xin, et al.
Published: (2023)
by: Guo, Xin, et al.
Published: (2023)
FlexiMo: A Flexible Remote Sensing Foundation Model
by: Li, Xuyang, et al.
Published: (2025)
by: Li, Xuyang, et al.
Published: (2025)
MapGlue: Multimodal Remote Sensing Image Matching
by: Wu, Peihao, et al.
Published: (2025)
by: Wu, Peihao, et al.
Published: (2025)
Decoding the Delta: Unifying Remote Sensing Change Detection and Understanding with Multimodal Large Language Models
by: Li, Xiaohe, et al.
Published: (2026)
by: Li, Xiaohe, et al.
Published: (2026)
Towards Knowledge Guided Pretraining Approaches for Multimodal Foundation Models: Applications in Remote Sensing
by: Ravirathinam, Praveen, et al.
Published: (2024)
by: Ravirathinam, Praveen, et al.
Published: (2024)
RSRefSeg: Referring Remote Sensing Image Segmentation with Foundation Models
by: Chen, Keyan, et al.
Published: (2025)
by: Chen, Keyan, et al.
Published: (2025)
Any-Optical-Model: A Universal Foundation Model for Optical Remote Sensing
by: Li, Xuyang, et al.
Published: (2025)
by: Li, Xuyang, et al.
Published: (2025)
DynamicVis: Dynamic Visual Perception for Efficient Remote Sensing Foundation Models
by: Chen, Keyan, et al.
Published: (2025)
by: Chen, Keyan, et al.
Published: (2025)
Cross-modal Context-aware Learning for Visual Prompt Guided Multimodal Image Understanding in Remote Sensing
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
SIGMAE: A Spectral-Index-Guided Foundation Model for Multispectral Remote Sensing
by: Zhang, Xiaokang, et al.
Published: (2026)
by: Zhang, Xiaokang, et al.
Published: (2026)
RSBuilding: Towards General Remote Sensing Image Building Extraction and Change Detection with Foundation Model
by: Wang, Mingze, et al.
Published: (2024)
by: Wang, Mingze, et al.
Published: (2024)
MGIMM: Multi-Granularity Instruction Multimodal Model for Attribute-Guided Remote Sensing Image Detailed Description
by: Yang, Cong, et al.
Published: (2024)
by: Yang, Cong, et al.
Published: (2024)
Adapting Vision Foundation Models for Robust Cloud Segmentation in Remote Sensing Images
by: Zou, Xuechao, et al.
Published: (2024)
by: Zou, Xuechao, et al.
Published: (2024)
ATRNet-STAR: A Large Dataset and Benchmark Towards Remote Sensing Object Recognition in the Wild
by: Liu, Yongxiang, et al.
Published: (2025)
by: Liu, Yongxiang, et al.
Published: (2025)
HuiYanEarth-SAR: A Foundation Model for High-Fidelity and Low-Cost Global Remote Sensing Imagery Generation
by: Liu, Yongxiang, et al.
Published: (2026)
by: Liu, Yongxiang, et al.
Published: (2026)
Foundation Model-Driven Semantic Change Detection in Remote Sensing Imagery
by: Shen, Hengtong, et al.
Published: (2026)
by: Shen, Hengtong, et al.
Published: (2026)
SkySense V2: A Unified Foundation Model for Multi-modal Remote Sensing
by: Zhang, Yingying, et al.
Published: (2025)
by: Zhang, Yingying, et al.
Published: (2025)
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
by: Wang, Fengxiang, et al.
Published: (2025)
by: Wang, Fengxiang, et al.
Published: (2025)
A Spatial-Spectral-Frequency Interactive Network for Multimodal Remote Sensing Classification
by: Liu, Hao, et al.
Published: (2025)
by: Liu, Hao, et al.
Published: (2025)
SpectralX: Parameter-efficient Domain Generalization for Spectral Remote Sensing Foundation Models
by: Zhang, Yuxiang, et al.
Published: (2025)
by: Zhang, Yuxiang, et al.
Published: (2025)
RSRefSeg 2: Decoupling Referring Remote Sensing Image Segmentation with Foundation Models
by: Chen, Keyan, et al.
Published: (2025)
by: Chen, Keyan, et al.
Published: (2025)
Integrating Reinforcement Learning with Visual Generative Models: Foundations and Advances
by: Liang, Yuanzhi, et al.
Published: (2025)
by: Liang, Yuanzhi, et al.
Published: (2025)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
by: Yang, Shuai, et al.
Published: (2026)
by: Yang, Shuai, et al.
Published: (2026)
MetaEarth: A Generative Foundation Model for Global-Scale Remote Sensing Image Generation
by: Yu, Zhiping, et al.
Published: (2024)
by: Yu, Zhiping, et al.
Published: (2024)
RS-DFM: A Remote Sensing Distributed Foundation Model for Diverse Downstream Tasks
by: Wang, Zhechao, et al.
Published: (2024)
by: Wang, Zhechao, et al.
Published: (2024)
Similar Items
-
FUSAR-GPT : A Spatiotemporal Feature-Embedded and Two-Stage Decoupled Visual Language Model for SAR Imagery
by: Zhang, Xiaokun, et al.
Published: (2026) -
Object Fidelity Diffusion for Remote Sensing Image Generation
by: Ye, Ziqi, et al.
Published: (2025) -
A Survey on Remote Sensing Foundation Models: From Vision to Multimodality
by: Huang, Ziyue, et al.
Published: (2025) -
VFM-ISRefiner: Towards Better Adapting Vision Foundation Models for Interactive Segmentation of Remote Sensing Images
by: Wang, Deliang, et al.
Published: (2025) -
OpenRSD: Towards Open-prompts for Object Detection in Remote Sensing Images
by: Huang, Ziyue, et al.
Published: (2025)