Massively Multimodal Foundation Models: A Framework for Capturing Interactions with Specialized Mixture-of-Experts
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Xing, Chung, Hsing-Huan, Ghosh, Joydeep, Liang, Paul Pu, Saria, Suchi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling
by: Chung, Hsing-Huan, et al.
Published: (2026)
by: Chung, Hsing-Huan, et al.
Published: (2026)
Between Linear and Sinusoidal: Rethinking the Time Encoder in Dynamic Graph Learning
by: Chung, Hsing-Huan, et al.
Published: (2025)
by: Chung, Hsing-Huan, et al.
Published: (2025)
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
FuseMoE: Mixture-of-Experts Transformers for Fleximodal Fusion
by: Han, Xing, et al.
Published: (2024)
by: Han, Xing, et al.
Published: (2024)
On the Invariance and Generality of Neural Scaling Laws
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
Novel Node Category Detection Under Subpopulation Shift
by: Chung, Hsing-Huan, et al.
Published: (2024)
by: Chung, Hsing-Huan, et al.
Published: (2024)
WATCH: Adaptive Monitoring for AI Deployments via Weighted-Conformal Martingales
by: Prinster, Drew, et al.
Published: (2025)
by: Prinster, Drew, et al.
Published: (2025)
Open-Set Domain Adaptation Under Background Distribution Shift: Challenges and A Provably Efficient Solution
by: Chaudhari, Shravan, et al.
Published: (2025)
by: Chaudhari, Shravan, et al.
Published: (2025)
Conformal Validity Guarantees Exist for Any Data Distribution (and How to Find Them)
by: Prinster, Drew, et al.
Published: (2024)
by: Prinster, Drew, et al.
Published: (2024)
Data Augmentations for Improved (Large) Language Model Generalization
by: Feder, Amir, et al.
Published: (2023)
by: Feder, Amir, et al.
Published: (2023)
MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping
by: Shan, Xiaojun, et al.
Published: (2025)
by: Shan, Xiaojun, et al.
Published: (2025)
Improving Coverage in Combined Prediction Sets with Weighted p-values
by: Wong, Gina, et al.
Published: (2025)
by: Wong, Gina, et al.
Published: (2025)
Achieving Fairness Across Local and Global Models in Federated Learning
by: Makhija, Disha, et al.
Published: (2024)
by: Makhija, Disha, et al.
Published: (2024)
TriForecaster: A Mixture of Experts Framework for Multi-Region Electric Load Forecasting with Tri-dimensional Specialization
by: Zhu, Zhaoyang, et al.
Published: (2025)
by: Zhu, Zhaoyang, et al.
Published: (2025)
MultiMed: Massively Multimodal and Multitask Medical Understanding
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
Toward Structural Multimodal Representations: Specialization, Selection, and Sparsification via Mixture-of-Experts
by: Choi, Hahyeon, et al.
Published: (2026)
by: Choi, Hahyeon, et al.
Published: (2026)
Mixture-of-Clustered-Experts: Advancing Expert Specialization and Generalization in Instruction Tuning
by: Eo, Sugyeong, et al.
Published: (2025)
by: Eo, Sugyeong, et al.
Published: (2025)
MMCTOP: A Multimodal Textualization and Mixture-of-Experts Framework for Clinical Trial Outcome Prediction
by: Aparício, Carolina, et al.
Published: (2025)
by: Aparício, Carolina, et al.
Published: (2025)
AnyExperts: On-Demand Expert Allocation for Multimodal Language Models with Mixture of Expert
by: Gao, Yuting, et al.
Published: (2025)
by: Gao, Yuting, et al.
Published: (2025)
GMoPE:A Prompt-Expert Mixture Framework for Graph Foundation Models
by: Wang, Zhibin, et al.
Published: (2025)
by: Wang, Zhibin, et al.
Published: (2025)
Moirai-MoE: Empowering Time Series Foundation Models with Sparse Mixture of Experts
by: Liu, Xu, et al.
Published: (2024)
by: Liu, Xu, et al.
Published: (2024)
Time Tracker: Mixture-of-Experts-Enhanced Foundation Time Series Forecasting Model with Decoupled Training Pipelines
by: Liang, Aobo, et al.
Published: (2025)
by: Liang, Aobo, et al.
Published: (2025)
Quadratic Gating Mixture of Experts: Statistical Insights into Self-Attention
by: Akbarian, Pedram, et al.
Published: (2024)
by: Akbarian, Pedram, et al.
Published: (2024)
A Mixture of Experts Foundation Model for Scanning Electron Microscopy Image Analysis
by: Ahmed, Sk Miraj, et al.
Published: (2026)
by: Ahmed, Sk Miraj, et al.
Published: (2026)
Exploring Expert Specialization through Unsupervised Training in Sparse Mixture of Experts
by: Nikolic, Strahinja, et al.
Published: (2025)
by: Nikolic, Strahinja, et al.
Published: (2025)
Multilinear Mixture of Experts: Scalable Expert Specialization through Factorization
by: Oldfield, James, et al.
Published: (2024)
by: Oldfield, James, et al.
Published: (2024)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
Efficiently Editing Mixture-of-Experts Models with Compressed Experts
by: He, Yifei, et al.
Published: (2025)
by: He, Yifei, et al.
Published: (2025)
How Many Experts Are Enough? Towards Optimal Semantic Specialization for Mixture-of-Experts
by: Park, Sumin, et al.
Published: (2025)
by: Park, Sumin, et al.
Published: (2025)
The Illusion of Specialization: Unveiling the Domain-Invariant "Standing Committee" in Mixture-of-Experts Models
by: Wang, Yan, et al.
Published: (2026)
by: Wang, Yan, et al.
Published: (2026)
OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration
by: Li, Shijun, et al.
Published: (2025)
by: Li, Shijun, et al.
Published: (2025)
A Stronger Mixture of Low-Rank Experts for Fine-Tuning Foundation Models
by: Sun, Mengyang, et al.
Published: (2025)
by: Sun, Mengyang, et al.
Published: (2025)
Toward Inference-optimal Mixture-of-Expert Large Language Models
by: Yun, Longfei, et al.
Published: (2024)
by: Yun, Longfei, et al.
Published: (2024)
Domain-Specialized Object Detection via Model-Level Mixtures of Experts
by: Pavlitska, Svetlana, et al.
Published: (2026)
by: Pavlitska, Svetlana, et al.
Published: (2026)
Sparse Mixture-of-Experts for Compositional Generalization: Empirical Evidence and Theoretical Foundations of Optimal Sparsity
by: Zhao, Jinze, et al.
Published: (2024)
by: Zhao, Jinze, et al.
Published: (2024)
MoE-Health: A Mixture of Experts Framework for Robust Multimodal Healthcare Prediction
by: Wang, Xiaoyang, et al.
Published: (2025)
by: Wang, Xiaoyang, et al.
Published: (2025)
SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning
by: Xie, Zhen-Hao, et al.
Published: (2026)
by: Xie, Zhen-Hao, et al.
Published: (2026)
Conformal Policy Control
by: Prinster, Drew, et al.
Published: (2026)
by: Prinster, Drew, et al.
Published: (2026)
MixtureKit: A General Framework for Composing, Training, and Visualizing Mixture-of-Experts Models
by: Chamma, Ahmad, et al.
Published: (2025)
by: Chamma, Ahmad, et al.
Published: (2025)
Similar Items
-
MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling
by: Chung, Hsing-Huan, et al.
Published: (2026) -
Between Linear and Sinusoidal: Rethinking the Time Encoder in Dynamic Graph Learning
by: Chung, Hsing-Huan, et al.
Published: (2025) -
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
by: Han, Xing, et al.
Published: (2026) -
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
by: Nguyen, Huy, et al.
Published: (2024) -
FuseMoE: Mixture-of-Experts Transformers for Fleximodal Fusion
by: Han, Xing, et al.
Published: (2024)