MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Hongyu, Xu, Jiayu, Wang, Ruiping, Feng, Yan, Zhai, Yitao, Pei, Peng, Cai, Xunliang, Chen, Xilin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer
von: Zhu, Minghao, et al.
Veröffentlicht: (2024)
von: Zhu, Minghao, et al.
Veröffentlicht: (2024)
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
von: Romero, Miguel, et al.
Veröffentlicht: (2025)
von: Romero, Miguel, et al.
Veröffentlicht: (2025)
MoTE: Mixture of Task-specific Experts for Pre-Trained ModelBased Class-incremental Learning
von: Li, Linjie, et al.
Veröffentlicht: (2025)
von: Li, Linjie, et al.
Veröffentlicht: (2025)
M4U: Evaluating Multilingual Understanding and Reasoning for Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
Glance and Focus: Memory Prompting for Multi-Event Video Question Answering
von: Bai, Ziyi, et al.
Veröffentlicht: (2024)
von: Bai, Ziyi, et al.
Veröffentlicht: (2024)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation
von: Lu, Yujie, et al.
Veröffentlicht: (2026)
von: Lu, Yujie, et al.
Veröffentlicht: (2026)
Blocks as Probes: Dissecting Categorization Ability of Large Multimodal Models
von: Fu, Bin, et al.
Veröffentlicht: (2024)
von: Fu, Bin, et al.
Veröffentlicht: (2024)
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
RoboPCA: Pose-centered Affordance Learning from Human Demonstrations for Robot Manipulation
von: Xiao, Zhanqi, et al.
Veröffentlicht: (2026)
von: Xiao, Zhanqi, et al.
Veröffentlicht: (2026)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
von: Zong, Zhuofan, et al.
Veröffentlicht: (2024)
von: Zong, Zhuofan, et al.
Veröffentlicht: (2024)
MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models
von: Wang, Dianyi, et al.
Veröffentlicht: (2025)
von: Wang, Dianyi, et al.
Veröffentlicht: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts
von: Xin, Jiayi, et al.
Veröffentlicht: (2025)
von: Xin, Jiayi, et al.
Veröffentlicht: (2025)
MoME: Mixture of Multimodal Experts for Cancer Survival Prediction
von: Xiong, Conghao, et al.
Veröffentlicht: (2024)
von: Xiong, Conghao, et al.
Veröffentlicht: (2024)
A Survey on Interpretability in Visual Recognition
von: Wan, Qiyang, et al.
Veröffentlicht: (2025)
von: Wan, Qiyang, et al.
Veröffentlicht: (2025)
VisKnow: Constructing Visual Knowledge Base for Object Understanding
von: Yao, Ziwei, et al.
Veröffentlicht: (2025)
von: Yao, Ziwei, et al.
Veröffentlicht: (2025)
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
von: Lin, Bin, et al.
Veröffentlicht: (2024)
von: Lin, Bin, et al.
Veröffentlicht: (2024)
Aria: An Open Multimodal Native Mixture-of-Experts Model
von: Li, Dongxu, et al.
Veröffentlicht: (2024)
von: Li, Dongxu, et al.
Veröffentlicht: (2024)
R^2MoE: Redundancy-Removal Mixture of Experts for Lifelong Concept Learning
von: Guo, Xiaohan, et al.
Veröffentlicht: (2025)
von: Guo, Xiaohan, et al.
Veröffentlicht: (2025)
GEAR: GEometry-motion Alternating Refinement for Articulated Object Modeling with Gaussian Splatting
von: Li, Jialin, et al.
Veröffentlicht: (2026)
von: Li, Jialin, et al.
Veröffentlicht: (2026)
GM-MoE: Low-Light Enhancement with Gated-Mechanism Mixture-of-Experts
von: Liao, Minwen, et al.
Veröffentlicht: (2025)
von: Liao, Minwen, et al.
Veröffentlicht: (2025)
MoRE: 3D Visual Geometry Reconstruction Meets Mixture-of-Experts
von: Gao, Jingnan, et al.
Veröffentlicht: (2025)
von: Gao, Jingnan, et al.
Veröffentlicht: (2025)
Semi-MoE: Mixture-of-Experts meets Semi-Supervised Histopathology Segmentation
von: Vu, Nguyen Lan Vi, et al.
Veröffentlicht: (2025)
von: Vu, Nguyen Lan Vi, et al.
Veröffentlicht: (2025)
TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts
von: Xu, Yu, et al.
Veröffentlicht: (2026)
von: Xu, Yu, et al.
Veröffentlicht: (2026)
OpenSubject: Leveraging Video-Derived Identity and Diversity Priors for Subject-driven Image Generation and Manipulation
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
MoE3D: A Mixture-of-Experts Module for 3D Reconstruction
von: Wang, Zichen, et al.
Veröffentlicht: (2026)
von: Wang, Zichen, et al.
Veröffentlicht: (2026)
Fair-MoE: Fairness-Oriented Mixture of Experts in Vision-Language Models
von: Wang, Peiran, et al.
Veröffentlicht: (2025)
von: Wang, Peiran, et al.
Veröffentlicht: (2025)
MoPE: Mixture of Prompt Experts for Parameter-Efficient and Scalable Multimodal Fusion
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
Mixture-of-Modality-Experts with Holistic Token Learning for Fine-Grained Multimodal Visual Analytics in Driver Action Recognition
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
MambaMoE: Mixture-of-Spectral-Spatial-Experts State Space Model for Hyperspectral Image Classification
von: Xu, Yichu, et al.
Veröffentlicht: (2025)
von: Xu, Yichu, et al.
Veröffentlicht: (2025)
GS-LTS: 3D Gaussian Splatting-Based Adaptive Modeling for Long-Term Service Robots
von: Fu, Bin, et al.
Veröffentlicht: (2025)
von: Fu, Bin, et al.
Veröffentlicht: (2025)
MoME: Mixture of Visual Language Medical Experts for Medical Imaging Segmentation
von: Rezvani, Arghavan, et al.
Veröffentlicht: (2025)
von: Rezvani, Arghavan, et al.
Veröffentlicht: (2025)
CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering
von: Huai, Tianyu, et al.
Veröffentlicht: (2025)
von: Huai, Tianyu, et al.
Veröffentlicht: (2025)
RingMoE: Mixture-of-Modality-Experts Multi-Modal Foundation Models for Universal Remote Sensing Image Interpretation
von: Bi, Hanbo, et al.
Veröffentlicht: (2025)
von: Bi, Hanbo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer
von: Zhu, Minghao, et al.
Veröffentlicht: (2024) -
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
von: Liu, Zhili, et al.
Veröffentlicht: (2024) -
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
von: Romero, Miguel, et al.
Veröffentlicht: (2025) -
MoTE: Mixture of Task-specific Experts for Pre-Trained ModelBased Class-incremental Learning
von: Li, Linjie, et al.
Veröffentlicht: (2025) -
M4U: Evaluating Multilingual Understanding and Reasoning for Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)