SARLANG-1M: A Benchmark for Vision-Language Modeling in SAR Image Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yimin, Xiao, Aoran, Ren, Yexian, Zhu, Yuting, Chen, Hongruixuan, Xia, Junshi, Yokoya, Naoto |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
by: Wei, Yimin, et al.
Published: (2026)
by: Wei, Yimin, et al.
Published: (2026)
GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing
by: Xiao, Aoran, et al.
Published: (2026)
by: Xiao, Aoran, et al.
Published: (2026)
OpenEarthMap-SAR: A Benchmark Synthetic Aperture Radar Dataset for Global High-Resolution Land Cover Mapping
by: Xia, Junshi, et al.
Published: (2025)
by: Xia, Junshi, et al.
Published: (2025)
SynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
by: Song, Jian, et al.
Published: (2024)
by: Song, Jian, et al.
Published: (2024)
Generalized Few-Shot Semantic Segmentation in Remote Sensing: Challenge and Benchmark
by: Broni-Bediako, Clifford, et al.
Published: (2024)
by: Broni-Bediako, Clifford, et al.
Published: (2024)
ChangeMamba: Remote Sensing Change Detection With Spatiotemporal State Space Model
by: Chen, Hongruixuan, et al.
Published: (2024)
by: Chen, Hongruixuan, et al.
Published: (2024)
DynamicVL: Benchmarking Multimodal Large Language Models for Dynamic City Understanding
by: Xuan, Weihao, et al.
Published: (2025)
by: Xuan, Weihao, et al.
Published: (2025)
Unsupervised Domain Adaptation Architecture Search with Self-Training for Land Cover Mapping
by: Broni-Bediako, Clifford, et al.
Published: (2024)
by: Broni-Bediako, Clifford, et al.
Published: (2024)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
by: Chen, Hongruixuan, et al.
Published: (2023)
by: Chen, Hongruixuan, et al.
Published: (2023)
Enhancing Monocular Height Estimation via Sparse LiDAR-Guided Correction
by: Song, Jian, et al.
Published: (2025)
by: Song, Jian, et al.
Published: (2025)
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
by: Wang, Junjue, et al.
Published: (2025)
by: Wang, Junjue, et al.
Published: (2025)
Change Detection Between Optical Remote Sensing Imagery and Map Data via Segment Anything Model (SAM)
by: Chen, Hongruixuan, et al.
Published: (2024)
by: Chen, Hongruixuan, et al.
Published: (2024)
A Vision Centric Remote Sensing Benchmark
by: Adejumo, Abduljaleel, et al.
Published: (2025)
by: Adejumo, Abduljaleel, et al.
Published: (2025)
Segment Anything with Multiple Modalities
by: Xiao, Aoran, et al.
Published: (2024)
by: Xiao, Aoran, et al.
Published: (2024)
Foundation Models for Remote Sensing and Earth Observation: A Survey
by: Xiao, Aoran, et al.
Published: (2024)
by: Xiao, Aoran, et al.
Published: (2024)
Geo3DVQA: Evaluating Vision-Language Models for 3D Geospatial Reasoning from Aerial Imagery
by: Tsujimoto, Mai, et al.
Published: (2025)
by: Tsujimoto, Mai, et al.
Published: (2025)
Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models
by: Xuan, Weihao, et al.
Published: (2025)
by: Xuan, Weihao, et al.
Published: (2025)
BRIGHT: A globally distributed multimodal building damage assessment dataset with very-high-resolution for all-weather disaster response
by: Chen, Hongruixuan, et al.
Published: (2025)
by: Chen, Hongruixuan, et al.
Published: (2025)
MP-HSIR: A Multi-Prompt Framework for Universal Hyperspectral Image Restoration
by: Wu, Zhehui, et al.
Published: (2025)
by: Wu, Zhehui, et al.
Published: (2025)
CrossEarth: Geospatial Vision Foundation Model for Domain Generalizable Remote Sensing Semantic Segmentation
by: Gong, Ziyang, et al.
Published: (2024)
by: Gong, Ziyang, et al.
Published: (2024)
A Survey of Sample-Efficient Deep Learning for Change Detection in Remote Sensing: Tasks, Strategies, and Challenges
by: Ding, Lei, et al.
Published: (2025)
by: Ding, Lei, et al.
Published: (2025)
Experience-Driven Multi-Agent Systems Are Training-free Context-aware Earth Observers
by: Dai, Pengyu, et al.
Published: (2026)
by: Dai, Pengyu, et al.
Published: (2026)
SARU: A Shadow-Aware and Removal Unified Framework for Remote Sensing Images with New Benchmarks
by: Bo, Zi-Yang, et al.
Published: (2026)
by: Bo, Zi-Yang, et al.
Published: (2026)
Robust Self-Supervised Cross-Modal Super-Resolution against Real-World Misaligned Observations
by: Dong, Xiaoyu, et al.
Published: (2026)
by: Dong, Xiaoyu, et al.
Published: (2026)
Enhancing 3D LiDAR Segmentation by Shaping Dense and Accurate 2D Semantic Predictions
by: Dong, Xiaoyu, et al.
Published: (2026)
by: Dong, Xiaoyu, et al.
Published: (2026)
MambaX: Image Super-Resolution with State Predictive Control
by: Li, Chenyu, et al.
Published: (2025)
by: Li, Chenyu, et al.
Published: (2025)
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
by: Ma, Qiwei, et al.
Published: (2025)
by: Ma, Qiwei, et al.
Published: (2025)
Language-Informed Hyperspectral Image Synthesis for Imbalanced-Small Sample Classification via Semi-Supervised Conditional Diffusion Model
by: Zhu, Yimin, et al.
Published: (2025)
by: Zhu, Yimin, et al.
Published: (2025)
CrossEarth-SAR: A SAR-Centric and Billion-Scale Geospatial Foundation Model for Domain Generalizable Semantic Segmentation
by: Ye, Ziqi, et al.
Published: (2026)
by: Ye, Ziqi, et al.
Published: (2026)
UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing
by: Ma, Lichen, et al.
Published: (2026)
by: Ma, Lichen, et al.
Published: (2026)
CHOICE: Benchmarking the Remote Sensing Capabilities of Large Vision-Language Models
by: An, Xiao, et al.
Published: (2024)
by: An, Xiao, et al.
Published: (2024)
Iterative Self-Improvement of Vision Language Models for Image Scoring and Self-Explanation
by: Tanji, Naoto, et al.
Published: (2025)
by: Tanji, Naoto, et al.
Published: (2025)
Building Extraction from Remote Sensing Imagery under Hazy and Low-light Conditions: Benchmark and Baseline
by: Sang, Feifei, et al.
Published: (2026)
by: Sang, Feifei, et al.
Published: (2026)
Direction-aware 3D Large Multimodal Models
by: Liu, Quan, et al.
Published: (2026)
by: Liu, Quan, et al.
Published: (2026)
EVLM: An Efficient Vision-Language Model for Visual Understanding
by: Chen, Kaibing, et al.
Published: (2024)
by: Chen, Kaibing, et al.
Published: (2024)
Local-to-Global Cross-Modal Attention-Aware Fusion for HSI-X Semantic Segmentation
by: Zhang, Xuming, et al.
Published: (2024)
by: Zhang, Xuming, et al.
Published: (2024)
V4d: voxel for 4d novel view synthesis
by: Gan, Wanshui, et al.
Published: (2022)
by: Gan, Wanshui, et al.
Published: (2022)
GaussianOcc: Fully Self-supervised and Efficient 3D Occupancy Estimation with Gaussian Splatting
by: Gan, Wanshui, et al.
Published: (2024)
by: Gan, Wanshui, et al.
Published: (2024)
Towards SAR Automatic Target Recognition MultiCategory SAR Image Classification Based on Light Weight Vision Transformer
by: Zhao, Guibin, et al.
Published: (2024)
by: Zhao, Guibin, et al.
Published: (2024)
Flooding Regularization for Stable Training of Generative Adversarial Networks
by: Yahiro, Iu, et al.
Published: (2023)
by: Yahiro, Iu, et al.
Published: (2023)
Similar Items
-
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
by: Wei, Yimin, et al.
Published: (2026) -
GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing
by: Xiao, Aoran, et al.
Published: (2026) -
OpenEarthMap-SAR: A Benchmark Synthetic Aperture Radar Dataset for Global High-Resolution Land Cover Mapping
by: Xia, Junshi, et al.
Published: (2025) -
SynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
by: Song, Jian, et al.
Published: (2024) -
Generalized Few-Shot Semantic Segmentation in Remote Sensing: Challenge and Benchmark
by: Broni-Bediako, Clifford, et al.
Published: (2024)