GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Xiao, Aoran, Cheng, Shihao, Xu, Yonghao, Ren, Yexian, Chen, Hongruixuan, Yokoya, Naoto |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
par: Wei, Yimin, et autres
Publié: (2026)
par: Wei, Yimin, et autres
Publié: (2026)
SARLANG-1M: A Benchmark for Vision-Language Modeling in SAR Image Understanding
par: Wei, Yimin, et autres
Publié: (2025)
par: Wei, Yimin, et autres
Publié: (2025)
Change Detection Between Optical Remote Sensing Imagery and Map Data via Segment Anything Model (SAM)
par: Chen, Hongruixuan, et autres
Publié: (2024)
par: Chen, Hongruixuan, et autres
Publié: (2024)
ChangeMamba: Remote Sensing Change Detection With Spatiotemporal State Space Model
par: Chen, Hongruixuan, et autres
Publié: (2024)
par: Chen, Hongruixuan, et autres
Publié: (2024)
SynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
par: Song, Jian, et autres
Publié: (2024)
par: Song, Jian, et autres
Publié: (2024)
A Vision Centric Remote Sensing Benchmark
par: Adejumo, Abduljaleel, et autres
Publié: (2025)
par: Adejumo, Abduljaleel, et autres
Publié: (2025)
Generalized Few-Shot Semantic Segmentation in Remote Sensing: Challenge and Benchmark
par: Broni-Bediako, Clifford, et autres
Publié: (2024)
par: Broni-Bediako, Clifford, et autres
Publié: (2024)
Foundation Models for Remote Sensing and Earth Observation: A Survey
par: Xiao, Aoran, et autres
Publié: (2024)
par: Xiao, Aoran, et autres
Publié: (2024)
Enhancing Monocular Height Estimation via Sparse LiDAR-Guided Correction
par: Song, Jian, et autres
Publié: (2025)
par: Song, Jian, et autres
Publié: (2025)
VLM2GeoVec: Toward Universal Multimodal Embeddings for Remote Sensing
par: Aimar, Emanuel Sánchez, et autres
Publié: (2025)
par: Aimar, Emanuel Sánchez, et autres
Publié: (2025)
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing
par: Zhang, Zilun, et autres
Publié: (2025)
par: Zhang, Zilun, et autres
Publié: (2025)
A Survey of Sample-Efficient Deep Learning for Change Detection in Remote Sensing: Tasks, Strategies, and Challenges
par: Ding, Lei, et autres
Publié: (2025)
par: Ding, Lei, et autres
Publié: (2025)
Segment Anything with Multiple Modalities
par: Xiao, Aoran, et autres
Publié: (2024)
par: Xiao, Aoran, et autres
Publié: (2024)
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
par: Wang, Junjue, et autres
Publié: (2025)
par: Wang, Junjue, et autres
Publié: (2025)
Towards Realistic Remote Sensing Dataset Distillation with Discriminative Prototype-guided Diffusion
par: Xu, Yonghao, et autres
Publié: (2026)
par: Xu, Yonghao, et autres
Publié: (2026)
Geo3DVQA: Evaluating Vision-Language Models for 3D Geospatial Reasoning from Aerial Imagery
par: Tsujimoto, Mai, et autres
Publié: (2025)
par: Tsujimoto, Mai, et autres
Publié: (2025)
ChangeBridge: Spatiotemporal Image Generation with Multimodal Controls for Remote Sensing
par: Zhao, Zhenghui, et autres
Publié: (2025)
par: Zhao, Zhenghui, et autres
Publié: (2025)
GeoHeight-Bench: Towards Height-Aware Multimodal Reasoning in Remote Sensing
par: Hu, Xuran, et autres
Publié: (2026)
par: Hu, Xuran, et autres
Publié: (2026)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
par: Chen, Hongruixuan, et autres
Publié: (2023)
par: Chen, Hongruixuan, et autres
Publié: (2023)
CrossEarth: Geospatial Vision Foundation Model for Domain Generalizable Remote Sensing Semantic Segmentation
par: Gong, Ziyang, et autres
Publié: (2024)
par: Gong, Ziyang, et autres
Publié: (2024)
OpenEarthMap-SAR: A Benchmark Synthetic Aperture Radar Dataset for Global High-Resolution Land Cover Mapping
par: Xia, Junshi, et autres
Publié: (2025)
par: Xia, Junshi, et autres
Publié: (2025)
AlignMMBench: Evaluating Chinese Multimodal Alignment in Large Vision-Language Models
par: Wu, Yuhang, et autres
Publié: (2024)
par: Wu, Yuhang, et autres
Publié: (2024)
Creation-MMBench: Assessing Context-Aware Creative Intelligence in MLLM
par: Fang, Xinyu, et autres
Publié: (2025)
par: Fang, Xinyu, et autres
Publié: (2025)
Experience-Driven Multi-Agent Systems Are Training-free Context-aware Earth Observers
par: Dai, Pengyu, et autres
Publié: (2026)
par: Dai, Pengyu, et autres
Publié: (2026)
GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing
par: Shabbir, Akashah, et autres
Publié: (2025)
par: Shabbir, Akashah, et autres
Publié: (2025)
GeoR-Bench: Evaluating Geoscience Visual Reasoning
par: Zheng, Yushuo, et autres
Publié: (2026)
par: Zheng, Yushuo, et autres
Publié: (2026)
GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI
par: Chen, Pengcheng, et autres
Publié: (2024)
par: Chen, Pengcheng, et autres
Publié: (2024)
On the Adversarial Vulnerabilities of Transfer Learning in Remote Sensing
par: Bai, Tao, et autres
Publié: (2025)
par: Bai, Tao, et autres
Publié: (2025)
Universal Adversarial Defense in Remote Sensing Based on Pre-trained Denoising Diffusion Models
par: Yu, Weikang, et autres
Publié: (2023)
par: Yu, Weikang, et autres
Publié: (2023)
Bridging Supervision Gaps: A Unified Framework for Remote Sensing Change Detection
par: Jiang, Kaixuan, et autres
Publié: (2026)
par: Jiang, Kaixuan, et autres
Publié: (2026)
Building Extraction from Remote Sensing Imagery under Hazy and Low-light Conditions: Benchmark and Baseline
par: Sang, Feifei, et autres
Publié: (2026)
par: Sang, Feifei, et autres
Publié: (2026)
GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing
par: Hasan, Maram, et autres
Publié: (2026)
par: Hasan, Maram, et autres
Publié: (2026)
HANet: A Hierarchical Attention Network for Change Detection With Bitemporal Very-High-Resolution Remote Sensing Images
par: Han, Chengxi, et autres
Publié: (2024)
par: Han, Chengxi, et autres
Publié: (2024)
AdaptMMBench: Benchmarking Adaptive Multimodal Reasoning for Mode Selection and Reasoning Process
par: Zhang, Xintong, et autres
Publié: (2026)
par: Zhang, Xintong, et autres
Publié: (2026)
SpectralGPT: Spectral Remote Sensing Foundation Model
par: Hong, Danfeng, et autres
Publié: (2023)
par: Hong, Danfeng, et autres
Publié: (2023)
SARU: A Shadow-Aware and Removal Unified Framework for Remote Sensing Images with New Benchmarks
par: Bo, Zi-Yang, et autres
Publié: (2026)
par: Bo, Zi-Yang, et autres
Publié: (2026)
SMGeo: Cross-View Object Geo-Localization with Grid-Level Mixture-of-Experts
par: Zhang, Fan, et autres
Publié: (2025)
par: Zhang, Fan, et autres
Publié: (2025)
Direction-aware 3D Large Multimodal Models
par: Liu, Quan, et autres
Publié: (2026)
par: Liu, Quan, et autres
Publié: (2026)
Change Guiding Network: Incorporating Change Prior to Guide Change Detection in Remote Sensing Imagery
par: Han, Chengxi, et autres
Publié: (2024)
par: Han, Chengxi, et autres
Publié: (2024)
Flooding Regularization for Stable Training of Generative Adversarial Networks
par: Yahiro, Iu, et autres
Publié: (2023)
par: Yahiro, Iu, et autres
Publié: (2023)
Documents similaires
-
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
par: Wei, Yimin, et autres
Publié: (2026) -
SARLANG-1M: A Benchmark for Vision-Language Modeling in SAR Image Understanding
par: Wei, Yimin, et autres
Publié: (2025) -
Change Detection Between Optical Remote Sensing Imagery and Map Data via Segment Anything Model (SAM)
par: Chen, Hongruixuan, et autres
Publié: (2024) -
ChangeMamba: Remote Sensing Change Detection With Spatiotemporal State Space Model
par: Chen, Hongruixuan, et autres
Publié: (2024) -
SynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
par: Song, Jian, et autres
Publié: (2024)