Equipping Vision Foundation Model with Mixture of Experts for Out-of-Distribution Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Shizhen, Liu, Jiahui, Wen, Xin, Tan, Haoru, Qi, Xiaojuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can OOD Object Detectors Learn from Foundation Models?
by: Liu, Jiahui, et al.
Published: (2024)
by: Liu, Jiahui, et al.
Published: (2024)
Data Pruning by Information Maximization
by: Tan, Haoru, et al.
Published: (2025)
by: Tan, Haoru, et al.
Published: (2025)
Learning from Neighbors: Category Extrapolation for Long-Tail Learning
by: Zhao, Shizhen, et al.
Published: (2024)
by: Zhao, Shizhen, et al.
Published: (2024)
ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation
by: Zhou, Shengchao, et al.
Published: (2025)
by: Zhou, Shengchao, et al.
Published: (2025)
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)
by: Zheng, Anlin, et al.
Published: (2026)
A Mixture of Exemplars Approach for Efficient Out-of-Distribution Detection with Foundation Models
by: Mannix, Evelyn, et al.
Published: (2023)
by: Mannix, Evelyn, et al.
Published: (2023)
Debiasing Text-to-Image Diffusion Models
by: He, Ruifei, et al.
Published: (2024)
by: He, Ruifei, et al.
Published: (2024)
A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
Aligning Effective Tokens with Video Anomaly in Large Language Models
by: Chen, Yingxian, et al.
Published: (2025)
by: Chen, Yingxian, et al.
Published: (2025)
SkyMoE: A Vision-Language Foundation Model for Enhancing Geospatial Interpretation with Mixture of Experts
by: Liu, Jiaqi, et al.
Published: (2025)
by: Liu, Jiaqi, et al.
Published: (2025)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
by: Zheng, Anlin, et al.
Published: (2025)
by: Zheng, Anlin, et al.
Published: (2025)
Knowledge Regularized Negative Feature Tuning of Vision-Language Models for Out-of-Distribution Detection
by: Zhu, Wenjie, et al.
Published: (2025)
by: Zhu, Wenjie, et al.
Published: (2025)
Long-Tailed Distribution-Aware Router For Mixture-of-Experts in Large Vision-Language Model
by: Cai, Chaoxiang, et al.
Published: (2025)
by: Cai, Chaoxiang, et al.
Published: (2025)
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
by: Han, Xumeng, et al.
Published: (2024)
by: Han, Xumeng, et al.
Published: (2024)
Can 3D Vision-Language Models Truly Understand Natural Language?
by: Deng, Weipeng, et al.
Published: (2024)
by: Deng, Weipeng, et al.
Published: (2024)
Delving into Out-of-Distribution Detection with Medical Vision-Language Models
by: Ju, Lie, et al.
Published: (2025)
by: Ju, Lie, et al.
Published: (2025)
Learning with Mixture of Prototypes for Out-of-Distribution Detection
by: Lu, Haodong, et al.
Published: (2024)
by: Lu, Haodong, et al.
Published: (2024)
Logit Mixture Outlier Exposure for Fine-grained Out-of-Distribution Detection
by: Shinohara, Akito, et al.
Published: (2025)
by: Shinohara, Akito, et al.
Published: (2025)
Fair-MoE: Fairness-Oriented Mixture of Experts in Vision-Language Models
by: Wang, Peiran, et al.
Published: (2025)
by: Wang, Peiran, et al.
Published: (2025)
Mixpert: Mitigating Multimodal Learning Conflicts with Efficient Mixture-of-Vision-Experts
by: He, Xin, et al.
Published: (2025)
by: He, Xin, et al.
Published: (2025)
Mamba as a Bridge: Where Vision Foundation Models Meet Vision Language Models for Domain-Generalized Semantic Segmentation
by: Zhang, Xin, et al.
Published: (2025)
by: Zhang, Xin, et al.
Published: (2025)
Agentic AIs Are the Missing Paradigm for Out-of-Distribution Generalization in Foundation Models
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
Rethinking Efficient Mixture-of-Experts for Remote Sensing Modality-Missing Classification
by: Gao, Qinghao, et al.
Published: (2025)
by: Gao, Qinghao, et al.
Published: (2025)
Training-Free Out-Of-Distribution Segmentation With Foundation Models
by: Nayal, Laith, et al.
Published: (2025)
by: Nayal, Laith, et al.
Published: (2025)
Harnessing Large Language and Vision-Language Models for Robust Out-of-Distribution Detection
by: Lee, Pei-Kang, et al.
Published: (2025)
by: Lee, Pei-Kang, et al.
Published: (2025)
FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
by: Lu, Xinhua, et al.
Published: (2025)
by: Lu, Xinhua, et al.
Published: (2025)
CSMoE: An Efficient Remote Sensing Foundation Model with Soft Mixture-of-Experts
by: Hackel, Leonard, et al.
Published: (2025)
by: Hackel, Leonard, et al.
Published: (2025)
Vision Also You Need: Navigating Out-of-Distribution Detection with Multimodal Large Language Model
by: Xu, Haoran, et al.
Published: (2026)
by: Xu, Haoran, et al.
Published: (2026)
TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
by: Liu, Huaize, et al.
Published: (2025)
by: Liu, Huaize, et al.
Published: (2025)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
by: Zong, Zhuofan, et al.
Published: (2024)
by: Zong, Zhuofan, et al.
Published: (2024)
Towards Vision Mixture of Experts for Wildlife Monitoring on the Edge
by: Mensah, Emmanuel Azuh, et al.
Published: (2024)
by: Mensah, Emmanuel Azuh, et al.
Published: (2024)
Teacher-Guided Routing for Sparse Vision Mixture-of-Experts
by: Kada, Masahiro, et al.
Published: (2026)
by: Kada, Masahiro, et al.
Published: (2026)
Backdooring Vision-Language Models with Out-Of-Distribution Data
by: Lyu, Weimin, et al.
Published: (2024)
by: Lyu, Weimin, et al.
Published: (2024)
Specializing Foundation Models via Mixture of Low-Rank Experts for Comprehensive Head CT Analysis
by: Yoo, Youngjin, et al.
Published: (2026)
by: Yoo, Youngjin, et al.
Published: (2026)
"Principal Components" Enable A New Language of Images
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
Solving Token Gradient Conflict in Mixture-of-Experts for Large Vision-Language Model
by: Yang, Longrong, et al.
Published: (2024)
by: Yang, Longrong, et al.
Published: (2024)
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
by: Lin, Bin, et al.
Published: (2024)
by: Lin, Bin, et al.
Published: (2024)
Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts Adapters
by: Yu, Jiazuo, et al.
Published: (2024)
by: Yu, Jiazuo, et al.
Published: (2024)
Understanding Data Influence with Differential Approximation
by: Tan, Haoru, et al.
Published: (2025)
by: Tan, Haoru, et al.
Published: (2025)
Similar Items
-
Can OOD Object Detectors Learn from Foundation Models?
by: Liu, Jiahui, et al.
Published: (2024) -
Data Pruning by Information Maximization
by: Tan, Haoru, et al.
Published: (2025) -
Learning from Neighbors: Category Extrapolation for Long-Tail Learning
by: Zhao, Shizhen, et al.
Published: (2024) -
ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation
by: Zhou, Shengchao, et al.
Published: (2025) -
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)