Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Danfeng, Li, Chenyu, Li, Xuyang, Camps-Valls, Gustau, Chanussot, Jocelyn |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UrbanSAM: Learning Invariance-Inspired Adapters for Segment Anything Models in Urban Construction
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
Any-Optical-Model: A Universal Foundation Model for Optical Remote Sensing
von: Li, Xuyang, et al.
Veröffentlicht: (2025)
von: Li, Xuyang, et al.
Veröffentlicht: (2025)
SeaMo: A Season-Aware Multimodal Foundation Model for Remote Sensing
von: Li, Xuyang, et al.
Veröffentlicht: (2024)
von: Li, Xuyang, et al.
Veröffentlicht: (2024)
FlexiMo: A Flexible Remote Sensing Foundation Model
von: Li, Xuyang, et al.
Veröffentlicht: (2025)
von: Li, Xuyang, et al.
Veröffentlicht: (2025)
SpectralGPT: Spectral Remote Sensing Foundation Model
von: Hong, Danfeng, et al.
Veröffentlicht: (2023)
von: Hong, Danfeng, et al.
Veröffentlicht: (2023)
SpectralMamba: Efficient Mamba for Hyperspectral Image Classification
von: Yao, Jing, et al.
Veröffentlicht: (2024)
von: Yao, Jing, et al.
Veröffentlicht: (2024)
KANO: Kolmogorov-Arnold Neural Operator for Image Super-Resolution
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
A Survey of Sample-Efficient Deep Learning for Change Detection in Remote Sensing: Tasks, Strategies, and Challenges
von: Ding, Lei, et al.
Veröffentlicht: (2025)
von: Ding, Lei, et al.
Veröffentlicht: (2025)
MambaX: Image Super-Resolution with State Predictive Control
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
Low-Rank Representations Meets Deep Unfolding: A Generalized and Interpretable Network for Hyperspectral Anomaly Detection
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
Investigating Traffic Accident Detection Using Multimodal Large Language Models
von: Skender, Ilhan, et al.
Veröffentlicht: (2025)
von: Skender, Ilhan, et al.
Veröffentlicht: (2025)
MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems
von: Li, Kaixin, et al.
Veröffentlicht: (2024)
von: Li, Kaixin, et al.
Veröffentlicht: (2024)
CUARewardBench: A Benchmark for Evaluating Reward Models on Computer-using Agent
von: Lin, Haojia, et al.
Veröffentlicht: (2025)
von: Lin, Haojia, et al.
Veröffentlicht: (2025)
Feature Extraction in the Remote Sensing Data Value Chain: A Systematic Review of Methods and Applications
von: Mankovich, Nathan, et al.
Veröffentlicht: (2025)
von: Mankovich, Nathan, et al.
Veröffentlicht: (2025)
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
von: Meng, Fanqing, et al.
Veröffentlicht: (2026)
von: Meng, Fanqing, et al.
Veröffentlicht: (2026)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
Can Vision-Language Models Handle Long-Context Code? An Empirical Study on Visual Compression
von: Zhong, Jianping, et al.
Veröffentlicht: (2026)
von: Zhong, Jianping, et al.
Veröffentlicht: (2026)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
von: Khadka, Krishna, et al.
Veröffentlicht: (2026)
von: Khadka, Krishna, et al.
Veröffentlicht: (2026)
GUing: A Mobile GUI Search Engine using a Vision-Language Model
von: Wei, Jialiang, et al.
Veröffentlicht: (2024)
von: Wei, Jialiang, et al.
Veröffentlicht: (2024)
Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
von: Lyu, Yunbo, et al.
Veröffentlicht: (2025)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2025)
ROMAN: Reward-Orchestrated Multi-Head Attention Network for Autonomous Driving System Testing
von: Chi, Jianlei, et al.
Veröffentlicht: (2026)
von: Chi, Jianlei, et al.
Veröffentlicht: (2026)
Ear-Keeper: A Cross-Platform AI System for Rapid and Accurate Ear Disease Diagnosis
von: Lu, Feiyan, et al.
Veröffentlicht: (2023)
von: Lu, Feiyan, et al.
Veröffentlicht: (2023)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
von: Khalid, Waleed, et al.
Veröffentlicht: (2025)
von: Khalid, Waleed, et al.
Veröffentlicht: (2025)
How Far Can VLMs Go for Visual Bug Detection? Studying 19,738 Keyframes from 41 Hours of Gameplay Videos
von: Lu, Wentao, et al.
Veröffentlicht: (2026)
von: Lu, Wentao, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Deep Learning for Wildfire Danger Forecasting
von: Kondylatos, Spyros, et al.
Veröffentlicht: (2025)
von: Kondylatos, Spyros, et al.
Veröffentlicht: (2025)
JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
Earth Embeddings as Products: Taxonomy, Ecosystem, and Standardized Access
von: Fang, Heng, et al.
Veröffentlicht: (2026)
von: Fang, Heng, et al.
Veröffentlicht: (2026)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
von: Abbasishahkoo, Amin, et al.
Veröffentlicht: (2026)
von: Abbasishahkoo, Amin, et al.
Veröffentlicht: (2026)
DOne: Decoupling Structure and Rendering for High-Fidelity Design-to-Code Generation
von: Huang, Xinhao, et al.
Veröffentlicht: (2026)
von: Huang, Xinhao, et al.
Veröffentlicht: (2026)
MVOS_HSI: A Python Library for Preprocessing Agricultural Crop Hyperspectral Data
von: Aggarwal, Rishik, et al.
Veröffentlicht: (2026)
von: Aggarwal, Rishik, et al.
Veröffentlicht: (2026)
ITKIT: Feasible CT Image Analysis based on SimpleITK and MMEngine
von: Zhang, Yiqin, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqin, et al.
Veröffentlicht: (2026)
SWAN -- Enabling Fast and Mobile Histopathology Image Annotation through Swipeable Interfaces
von: Banerjee, Sweta, et al.
Veröffentlicht: (2025)
von: Banerjee, Sweta, et al.
Veröffentlicht: (2025)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
Cross-Breed Pig Identification Using Auricular Vein Pattern Recognition: A Machine Learning Approach for Small-Scale Farming Applications
von: Nsengiyumvaa, Emmanuel, et al.
Veröffentlicht: (2025)
von: Nsengiyumvaa, Emmanuel, et al.
Veröffentlicht: (2025)
Natural Adversaries: Fuzzing Autonomous Vehicles with Realistic Roadside Object Placements
von: Sun, Yang, et al.
Veröffentlicht: (2024)
von: Sun, Yang, et al.
Veröffentlicht: (2024)
Interpretable Gallbladder Ultrasound Diagnosis: A Lightweight Web-Mobile Software Platform with Real-Time XAI
von: Bhoyan, Fuyad Hasan, et al.
Veröffentlicht: (2025)
von: Bhoyan, Fuyad Hasan, et al.
Veröffentlicht: (2025)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
von: Lambertenghi, Stefano Carlo, et al.
Veröffentlicht: (2025)
von: Lambertenghi, Stefano Carlo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UrbanSAM: Learning Invariance-Inspired Adapters for Segment Anything Models in Urban Construction
von: Li, Chenyu, et al.
Veröffentlicht: (2025) -
Any-Optical-Model: A Universal Foundation Model for Optical Remote Sensing
von: Li, Xuyang, et al.
Veröffentlicht: (2025) -
SeaMo: A Season-Aware Multimodal Foundation Model for Remote Sensing
von: Li, Xuyang, et al.
Veröffentlicht: (2024) -
FlexiMo: A Flexible Remote Sensing Foundation Model
von: Li, Xuyang, et al.
Veröffentlicht: (2025) -
SpectralGPT: Spectral Remote Sensing Foundation Model
von: Hong, Danfeng, et al.
Veröffentlicht: (2023)