MSSDF: Modality-Shared Self-supervised Distillation for High-Resolution Multi-modal Remote Sensing Image Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Tong, Chen, Guanzhou, Zhang, Xiaodong, Liu, Chenxi, Wang, Jiaqi, Tan, Xiaoliang, Guo, Wenchao, Yang, Qingyuan, Zhang, Kaiqi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STARS: Shared-specific Translation and Alignment for missing-modality Remote Sensing Semantic Segmentation
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
LMFNet: An Efficient Multimodal Fusion Approach for Semantic Segmentation in High-Resolution Remote Sensing
by: Wang, Tong, et al.
Published: (2024)
by: Wang, Tong, et al.
Published: (2024)
Segment Change Model (SCM) for Unsupervised Change detection in VHR Remote Sensing Images: a Case Study of Buildings
by: Tan, Xiaoliang, et al.
Published: (2023)
by: Tan, Xiaoliang, et al.
Published: (2023)
S3Net: Innovating Stereo Matching and Semantic Segmentation with a Single-Branch Semantic Stereo Network in Satellite Epipolar Imagery
by: Yang, Qingyuan, et al.
Published: (2024)
by: Yang, Qingyuan, et al.
Published: (2024)
BFA-YOLO: A balanced multiscale object detection network for building façade attachments detection
by: Chen, Yangguang, et al.
Published: (2024)
by: Chen, Yangguang, et al.
Published: (2024)
Enhancing Terrestrial Net Primary Productivity Estimation with EXP-CASA: A Novel Light Use Efficiency Model Approach
by: Chen, Guanzhou, et al.
Published: (2024)
by: Chen, Guanzhou, et al.
Published: (2024)
Multi-modal Uncertainty Robust Tree Cover Segmentation For High-Resolution Remote Sensing Images
by: Gui, Yuanyuan, et al.
Published: (2025)
by: Gui, Yuanyuan, et al.
Published: (2025)
Balanced Multi-modal Federated Learning via Cross-Modal Infiltration
by: Fan, Yunfeng, et al.
Published: (2023)
by: Fan, Yunfeng, et al.
Published: (2023)
Expediting Building Footprint Extraction from High-resolution Remote Sensing Images via progressive lenient supervision
by: Guo, Haonan, et al.
Published: (2023)
by: Guo, Haonan, et al.
Published: (2023)
DeepLight: Reconstructing High-Resolution Observations of Nighttime Light With Multi-Modal Remote Sensing Data
by: Zhang, Lixian, et al.
Published: (2024)
by: Zhang, Lixian, et al.
Published: (2024)
ImageRAG: Enhancing Ultra High Resolution Remote Sensing Imagery Analysis with ImageRAG
by: Zhang, Zilun, et al.
Published: (2024)
by: Zhang, Zilun, et al.
Published: (2024)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
Remote Sensing Image Classification with Decoupled Knowledge Distillation
by: He, Yaping, et al.
Published: (2025)
by: He, Yaping, et al.
Published: (2025)
Overcome Modal Bias in Multi-modal Federated Learning via Balanced Modality Selection
by: Fan, Yunfeng, et al.
Published: (2023)
by: Fan, Yunfeng, et al.
Published: (2023)
RSCC: A Large-Scale Remote Sensing Change Caption Dataset for Disaster Events
by: Chen, Zhenyuan, et al.
Published: (2025)
by: Chen, Zhenyuan, et al.
Published: (2025)
A Benchmark for Ultra-High-Resolution Remote Sensing MLLMs
by: Dang, Yunkai, et al.
Published: (2025)
by: Dang, Yunkai, et al.
Published: (2025)
SlimDiffSR: Toward Lightweight and Efficient Remote Sensing Image Super-Resolution via Diffusion Model Distillation
by: Wang, Ce, et al.
Published: (2026)
by: Wang, Ce, et al.
Published: (2026)
EarthGPT: A Universal Multi-modal Large Language Model for Multi-sensor Image Comprehension in Remote Sensing Domain
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding
by: Xu, Linrui, et al.
Published: (2024)
by: Xu, Linrui, et al.
Published: (2024)
Any2Any: Unified Arbitrary Modality Translation for Remote Sensing
by: Chen, Haoyang, et al.
Published: (2026)
by: Chen, Haoyang, et al.
Published: (2026)
Self-Supervised Quantization-Aware Knowledge Distillation
by: Zhao, Kaiqi, et al.
Published: (2024)
by: Zhao, Kaiqi, et al.
Published: (2024)
Remote Sensing with High Spatial Resolution
by: Sandmann, André, et al.
Published: (2024)
by: Sandmann, André, et al.
Published: (2024)
RingMoE: Mixture-of-Modality-Experts Multi-Modal Foundation Models for Universal Remote Sensing Image Interpretation
by: Bi, Hanbo, et al.
Published: (2025)
by: Bi, Hanbo, et al.
Published: (2025)
Multimodal Remote Sensing Image Matching Method Based on Improved Self‐Similarity Index Map and Absolute Phase Direction
by: Xiaodong Niu, et al.
Published: (2025)
by: Xiaodong Niu, et al.
Published: (2025)
Self-supervised Audiovisual Representation Learning for Remote Sensing Data
by: Heidler, Konrad, et al.
Published: (2021)
by: Heidler, Konrad, et al.
Published: (2021)
Frequency-Assisted Mamba for Remote Sensing Image Super-Resolution
by: Xiao, Yi, et al.
Published: (2024)
by: Xiao, Yi, et al.
Published: (2024)
UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
by: Zhu, Enze, et al.
Published: (2024)
by: Zhu, Enze, et al.
Published: (2024)
Learning a Cross-modality Anomaly Detector for Remote Sensing Imagery
by: Li, Jingtao, et al.
Published: (2023)
by: Li, Jingtao, et al.
Published: (2023)
EarthMarker: A Visual Prompting Multi-modal Large Language Model for Remote Sensing
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
Rethinking Multi-view Representation Learning via Distilled Disentangling
by: Ke, Guanzhou, et al.
Published: (2024)
by: Ke, Guanzhou, et al.
Published: (2024)
Self-Supervised Cross-Modal Text-Image Time Series Retrieval in Remote Sensing
by: Hoxha, Genc, et al.
Published: (2025)
by: Hoxha, Genc, et al.
Published: (2025)
Distilling Cross-Modal Knowledge via Feature Disentanglement
by: Liu, Junhong, et al.
Published: (2025)
by: Liu, Junhong, et al.
Published: (2025)
XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?
by: Wang, Fengxiang, et al.
Published: (2025)
by: Wang, Fengxiang, et al.
Published: (2025)
RemoteDet-Mamba: A Hybrid Mamba-CNN Network for Multi-modal Object Detection in Remote Sensing Images
by: Ren, Kejun, et al.
Published: (2024)
by: Ren, Kejun, et al.
Published: (2024)
Ultrathin Graphene Strain Sensor Arrays for High‐Sensitivity Multifunctional Sensing with Millimeter‐Scale Resolution
by: Wenchao Luo, et al.
Published: (2025)
by: Wenchao Luo, et al.
Published: (2025)
CM-MaskSD: Cross-Modality Masked Self-Distillation for Referring Image Segmentation
by: Wang, Wenxuan, et al.
Published: (2023)
by: Wang, Wenxuan, et al.
Published: (2023)
HANet: A Hierarchical Attention Network for Change Detection With Bitemporal Very-High-Resolution Remote Sensing Images
by: Han, Chengxi, et al.
Published: (2024)
by: Han, Chengxi, et al.
Published: (2024)
SDRNET: Stacked Deep Residual Network for Accurate Semantic Segmentation of Fine-Resolution Remotely Sensed Images
by: Wambugu, Naftaly, et al.
Published: (2025)
by: Wambugu, Naftaly, et al.
Published: (2025)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
by: Lai, Songning, et al.
Published: (2023)
by: Lai, Songning, et al.
Published: (2023)
GeoEyes: On-Demand Visual Focusing for Evidence-Grounded Understanding of Ultra-High-Resolution Remote Sensing Imagery
by: Wang, Fengxiang, et al.
Published: (2026)
by: Wang, Fengxiang, et al.
Published: (2026)
Similar Items
-
STARS: Shared-specific Translation and Alignment for missing-modality Remote Sensing Semantic Segmentation
by: Wang, Tong, et al.
Published: (2026) -
LMFNet: An Efficient Multimodal Fusion Approach for Semantic Segmentation in High-Resolution Remote Sensing
by: Wang, Tong, et al.
Published: (2024) -
Segment Change Model (SCM) for Unsupervised Change detection in VHR Remote Sensing Images: a Case Study of Buildings
by: Tan, Xiaoliang, et al.
Published: (2023) -
S3Net: Innovating Stereo Matching and Semantic Segmentation with a Single-Branch Semantic Stereo Network in Satellite Epipolar Imagery
by: Yang, Qingyuan, et al.
Published: (2024) -
BFA-YOLO: A balanced multiscale object detection network for building façade attachments detection
by: Chen, Yangguang, et al.
Published: (2024)