Towards Efficient Benchmarking of Foundation Models in Remote Sensing: A Capabilities Encoding Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Adorni, Pierre, Pham, Minh-Tan, May, Stéphane, Lefèvre, Sébastien |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EoS-FM: Can an Ensemble of Specialist Models act as a Generalist Feature Extractor?
by: Adorni, Pierre, et al.
Published: (2025)
by: Adorni, Pierre, et al.
Published: (2025)
GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing
by: Hasan, Maram, et al.
Published: (2026)
by: Hasan, Maram, et al.
Published: (2026)
Contributions to Label-Efficient Learning in Computer Vision and Remote Sensing
by: Pham, Minh-Tan
Published: (2025)
by: Pham, Minh-Tan
Published: (2025)
Leveraging Membership Inference Attacks for Privacy Measurement in Federated Learning for Remote Sensing Images
by: Duong, Anh-Kiet, et al.
Published: (2026)
by: Duong, Anh-Kiet, et al.
Published: (2026)
Mapping earth mounds from space
by: Uzun, Baki, et al.
Published: (2024)
by: Uzun, Baki, et al.
Published: (2024)
OceanMAE: A Foundation Model for Ocean Remote Sensing
by: Stamer, Viola-Joanna, et al.
Published: (2026)
by: Stamer, Viola-Joanna, et al.
Published: (2026)
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
by: Wang, Fengxiang, et al.
Published: (2025)
by: Wang, Fengxiang, et al.
Published: (2025)
REMSA: Foundation Model Selection for Remote Sensing via a Constraint-Aware Agent
by: Chen, Binger, et al.
Published: (2025)
by: Chen, Binger, et al.
Published: (2025)
LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation
by: Wang, Jun, et al.
Published: (2026)
by: Wang, Jun, et al.
Published: (2026)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales
by: Rodriguez, Jorge L., et al.
Published: (2026)
by: Rodriguez, Jorge L., et al.
Published: (2026)
RingMo-Aerial: An Aerial Remote Sensing Foundation Model With Affine Transformation Contrastive Learning
by: Diao, Wenhui, et al.
Published: (2024)
by: Diao, Wenhui, et al.
Published: (2024)
A Billion-scale Foundation Model for Remote Sensing Images
by: Cha, Keumgang, et al.
Published: (2023)
by: Cha, Keumgang, et al.
Published: (2023)
SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models
by: Zhong, Chen, et al.
Published: (2026)
by: Zhong, Chen, et al.
Published: (2026)
Exploring Efficient Open-Vocabulary Segmentation in the Remote Sensing
by: Li, Bingyu, et al.
Published: (2025)
by: Li, Bingyu, et al.
Published: (2025)
Detecting Brick Kiln Infrastructure at Scale: Graph, Foundation, and Remote Sensing Models for Satellite Imagery Data
by: Nazir, Usman, et al.
Published: (2026)
by: Nazir, Usman, et al.
Published: (2026)
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
by: Luo, Zhiming, et al.
Published: (2026)
by: Luo, Zhiming, et al.
Published: (2026)
HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing
by: Zhang, Xinyu, et al.
Published: (2026)
by: Zhang, Xinyu, et al.
Published: (2026)
Rethinking Electro-Optical Vision Foundation Models for Remote Sensing Retrieval: A Controlled Comparison with Generalist VFM
by: Park, Hyobin, et al.
Published: (2026)
by: Park, Hyobin, et al.
Published: (2026)
Text2Seg: Remote Sensing Image Semantic Segmentation via Text-Guided Visual Foundation Models
by: Zhang, Jielu, et al.
Published: (2023)
by: Zhang, Jielu, et al.
Published: (2023)
Towards Scalable Foundation Model for Multi-modal and Hyperspectral Geospatial Data
by: Si, Haozhe, et al.
Published: (2025)
by: Si, Haozhe, et al.
Published: (2025)
Efficient Adaptation For Remote Sensing Visual Grounding
by: Moughnieh, Hasan, et al.
Published: (2025)
by: Moughnieh, Hasan, et al.
Published: (2025)
CGEarthEye:A High-Resolution Remote Sensing Vision Foundation Model Based on the Jilin-1 Satellite Constellation
by: Yi, Zhiwei, et al.
Published: (2025)
by: Yi, Zhiwei, et al.
Published: (2025)
A Benchmark for Ultra-High-Resolution Remote Sensing MLLMs
by: Dang, Yunkai, et al.
Published: (2025)
by: Dang, Yunkai, et al.
Published: (2025)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
by: Truong, Khang, et al.
Published: (2025)
by: Truong, Khang, et al.
Published: (2025)
DC3DCD: unsupervised learning for multiclass 3D point cloud change detection
by: de Gélis, Iris, et al.
Published: (2023)
by: de Gélis, Iris, et al.
Published: (2023)
SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding
by: Luo, Junwei, et al.
Published: (2024)
by: Luo, Junwei, et al.
Published: (2024)
DGTRSD & DGTRS-CLIP: A Dual-Granularity Remote Sensing Image-Text Dataset and Vision Language Foundation Model for Alignment
by: Chen, Weizhi, et al.
Published: (2025)
by: Chen, Weizhi, et al.
Published: (2025)
Towards Neural Foundation Models for Vision: Aligning EEG, MEG, and fMRI Representations for Decoding, Encoding, and Modality Conversion
by: Ferrante, Matteo, et al.
Published: (2024)
by: Ferrante, Matteo, et al.
Published: (2024)
How Much 3D Do Video Foundation Models Encode?
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Neurosymbolic Inference On Foundation Models For Remote Sensing Text-to-image Retrieval With Complex Queries
by: Mezzi, Emanuele, et al.
Published: (2025)
by: Mezzi, Emanuele, et al.
Published: (2025)
A Medical Data-Effective Learning Benchmark for Highly Efficient Pre-training of Foundation Models
by: Yang, Wenxuan, et al.
Published: (2024)
by: Yang, Wenxuan, et al.
Published: (2024)
Benchmarking Foundation Models and Parameter-Efficient Fine-Tuning for Prognosis Prediction in Medical Imaging
by: Ruffini, Filippo, et al.
Published: (2025)
by: Ruffini, Filippo, et al.
Published: (2025)
HIMOSA: Efficient Remote Sensing Image Super-Resolution with Hierarchical Mixture of Sparse Attention
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models
by: Pan, Chenbin, et al.
Published: (2025)
by: Pan, Chenbin, et al.
Published: (2025)
BabyVLM-V2: Toward Developmentally Grounded Pretraining and Benchmarking of Vision Foundation Models
by: Wang, Shengao, et al.
Published: (2025)
by: Wang, Shengao, et al.
Published: (2025)
EPEE: Towards Efficient and Effective Foundation Models in Biomedicine
by: Zhan, Zaifu, et al.
Published: (2025)
by: Zhan, Zaifu, et al.
Published: (2025)
Remote Sensing Retrieval-Augmented Generation: Bridging Remote Sensing Imagery and Comprehensive Knowledge with a Multi-Modal Dataset and Retrieval-Augmented Generation Model
by: Wen, Congcong, et al.
Published: (2025)
by: Wen, Congcong, et al.
Published: (2025)
FUSE-RSVLM: Feature Fusion Vision-Language Model for Remote Sensing
by: Dang, Yunkai, et al.
Published: (2025)
by: Dang, Yunkai, et al.
Published: (2025)
Vision-Language Models in Remote Sensing: Current Progress and Future Trends
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Similar Items
-
EoS-FM: Can an Ensemble of Specialist Models act as a Generalist Feature Extractor?
by: Adorni, Pierre, et al.
Published: (2025) -
GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing
by: Hasan, Maram, et al.
Published: (2026) -
Contributions to Label-Efficient Learning in Computer Vision and Remote Sensing
by: Pham, Minh-Tan
Published: (2025) -
Leveraging Membership Inference Attacks for Privacy Measurement in Federated Learning for Remote Sensing Images
by: Duong, Anh-Kiet, et al.
Published: (2026) -
Mapping earth mounds from space
by: Uzun, Baki, et al.
Published: (2024)