SkyMoE: A Vision-Language Foundation Model for Enhancing Geospatial Interpretation with Mixture of Experts
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Jiaqi, Fu, Ronghao, Sun, Lang, Liu, Haoran, Yang, Xiao, Zhang, Weipeng, Na, Xu, Duan, Zhuoran, Yang, Bo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GeoDiT: A Diffusion-based Vision-Language Model for Geospatial Understanding
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
SkyNative: A Native Multimodal Framework for Remote Sensing Visual Evidence Reasoning
di: Yang, Xiao, et al.
Pubblicazione: (2026)
di: Yang, Xiao, et al.
Pubblicazione: (2026)
Towards Faithful Reasoning in Remote Sensing: A Perceptually-Grounded GeoSpatial Chain-of-Thought for Vision-Language Models
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
GeoSolver: Scaling Test-Time Reasoning in Remote Sensing with Fine-Grained Process Supervision
di: Sun, Lang, et al.
Pubblicazione: (2026)
di: Sun, Lang, et al.
Pubblicazione: (2026)
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
GeoAlignCLIP: Enhancing Fine-Grained Vision-Language Alignment in Remote Sensing via Multi-Granular Consistency Learning
di: Yang, Xiao, et al.
Pubblicazione: (2026)
di: Yang, Xiao, et al.
Pubblicazione: (2026)
ECG-MoE: Mixture-of-Expert Electrocardiogram Foundation Model
di: Xu, Yuhao, et al.
Pubblicazione: (2026)
di: Xu, Yuhao, et al.
Pubblicazione: (2026)
GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding
di: Zhu, Jiashun, et al.
Pubblicazione: (2026)
di: Zhu, Jiashun, et al.
Pubblicazione: (2026)
Fair-MoE: Fairness-Oriented Mixture of Experts in Vision-Language Models
di: Wang, Peiran, et al.
Pubblicazione: (2025)
di: Wang, Peiran, et al.
Pubblicazione: (2025)
MoPD: Mixture-of-Prompts Distillation for Vision-Language Models
di: Chen, Yang, et al.
Pubblicazione: (2024)
di: Chen, Yang, et al.
Pubblicazione: (2024)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
di: Yang, Cheng, et al.
Pubblicazione: (2024)
di: Yang, Cheng, et al.
Pubblicazione: (2024)
PWC-MoE: Privacy-Aware Wireless Collaborative Mixture of Experts
di: Su, Yang, et al.
Pubblicazione: (2025)
di: Su, Yang, et al.
Pubblicazione: (2025)
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
di: Du, Zhiying, et al.
Pubblicazione: (2025)
di: Du, Zhiying, et al.
Pubblicazione: (2025)
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
di: Lin, Bin, et al.
Pubblicazione: (2024)
di: Lin, Bin, et al.
Pubblicazione: (2024)
RingMoE: Mixture-of-Modality-Experts Multi-Modal Foundation Models for Universal Remote Sensing Image Interpretation
di: Bi, Hanbo, et al.
Pubblicazione: (2025)
di: Bi, Hanbo, et al.
Pubblicazione: (2025)
BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma
di: Yang, Junlin, et al.
Pubblicazione: (2026)
di: Yang, Junlin, et al.
Pubblicazione: (2026)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
di: Liu, Huaize, et al.
Pubblicazione: (2025)
di: Liu, Huaize, et al.
Pubblicazione: (2025)
MoLAE: Mixture of Latent Experts for Parameter-Efficient Language Models
di: Liu, Zehua, et al.
Pubblicazione: (2025)
di: Liu, Zehua, et al.
Pubblicazione: (2025)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
di: Zong, Zhuofan, et al.
Pubblicazione: (2024)
di: Zong, Zhuofan, et al.
Pubblicazione: (2024)
OpenMoE: An Early Effort on Open Mixture-of-Experts Language Models
di: Xue, Fuzhao, et al.
Pubblicazione: (2024)
di: Xue, Fuzhao, et al.
Pubblicazione: (2024)
LAER-MoE: Load-Adaptive Expert Re-layout for Efficient Mixture-of-Experts Training
di: Liu, Xinyi, et al.
Pubblicazione: (2026)
di: Liu, Xinyi, et al.
Pubblicazione: (2026)
Moirai-MoE: Empowering Time Series Foundation Models with Sparse Mixture of Experts
di: Liu, Xu, et al.
Pubblicazione: (2024)
di: Liu, Xu, et al.
Pubblicazione: (2024)
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
di: Jiang, Songtao, et al.
Pubblicazione: (2024)
di: Jiang, Songtao, et al.
Pubblicazione: (2024)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
MoEScore: Mixture-of-Experts-Based Text-Audio Relevance Score Prediction for Text-to-Audio System Evaluation
di: Sun, Bochao, et al.
Pubblicazione: (2026)
di: Sun, Bochao, et al.
Pubblicazione: (2026)
WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting
di: Wu, Shunyu, et al.
Pubblicazione: (2026)
di: Wu, Shunyu, et al.
Pubblicazione: (2026)
Equipping Vision Foundation Model with Mixture of Experts for Out-of-Distribution Detection
di: Zhao, Shizhen, et al.
Pubblicazione: (2025)
di: Zhao, Shizhen, et al.
Pubblicazione: (2025)
MoBiLE: Efficient Mixture-of-Experts Inference on Consumer GPU with Mixture of Big Little Experts
di: Zhao, Yushu, et al.
Pubblicazione: (2025)
di: Zhao, Yushu, et al.
Pubblicazione: (2025)
Earth-Adapter: Bridge the Geospatial Domain Gaps with Mixture of Frequency Adaptation
di: Hu, Xiaoxing, et al.
Pubblicazione: (2025)
di: Hu, Xiaoxing, et al.
Pubblicazione: (2025)
Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation
di: He, Xiao, et al.
Pubblicazione: (2025)
di: He, Xiao, et al.
Pubblicazione: (2025)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
di: Yang, Zhenjie, et al.
Pubblicazione: (2025)
di: Yang, Zhenjie, et al.
Pubblicazione: (2025)
MoEQuant: Enhancing Quantization for Mixture-of-Experts Large Language Models via Expert-Balanced Sampling and Affinity Guidance
di: Hu, Xing, et al.
Pubblicazione: (2025)
di: Hu, Xing, et al.
Pubblicazione: (2025)
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
di: Han, Xumeng, et al.
Pubblicazione: (2024)
di: Han, Xumeng, et al.
Pubblicazione: (2024)
InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation
di: Xiao, Jinqi, et al.
Pubblicazione: (2025)
di: Xiao, Jinqi, et al.
Pubblicazione: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
di: Jin, Peng, et al.
Pubblicazione: (2024)
di: Jin, Peng, et al.
Pubblicazione: (2024)
Mixture-of-Linear-Experts for Long-term Time Series Forecasting
di: Ni, Ronghao, et al.
Pubblicazione: (2023)
di: Ni, Ronghao, et al.
Pubblicazione: (2023)
FreqMoE: Enhancing Time Series Forecasting through Frequency Decomposition Mixture of Experts
di: Liu, Ziqi
Pubblicazione: (2025)
di: Liu, Ziqi
Pubblicazione: (2025)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
di: Feng, Yuchen, et al.
Pubblicazione: (2025)
di: Feng, Yuchen, et al.
Pubblicazione: (2025)
MoAPT: Mixture of Adversarial Prompt Tuning for Vision-Language Models
di: Zhao, Shiji, et al.
Pubblicazione: (2025)
di: Zhao, Shiji, et al.
Pubblicazione: (2025)
MoST: Mixing Speech and Text with Modality-Aware Mixture of Experts
di: Lou, Yuxuan, et al.
Pubblicazione: (2026)
di: Lou, Yuxuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
GeoDiT: A Diffusion-based Vision-Language Model for Geospatial Understanding
di: Liu, Jiaqi, et al.
Pubblicazione: (2025) -
SkyNative: A Native Multimodal Framework for Remote Sensing Visual Evidence Reasoning
di: Yang, Xiao, et al.
Pubblicazione: (2026) -
Towards Faithful Reasoning in Remote Sensing: A Perceptually-Grounded GeoSpatial Chain-of-Thought for Vision-Language Models
di: Liu, Jiaqi, et al.
Pubblicazione: (2025) -
GeoSolver: Scaling Test-Time Reasoning in Remote Sensing with Fine-Grained Process Supervision
di: Sun, Lang, et al.
Pubblicazione: (2026) -
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks
di: Fu, Ronghao, et al.
Pubblicazione: (2026)