MoEController: Instruction-based Arbitrary Image Manipulation with Mixture-of-Expert Controllers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Sijia, Chen, Chen, Lu, Haonan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image Generation
von: He, Xuehai, et al.
Veröffentlicht: (2024)
von: He, Xuehai, et al.
Veröffentlicht: (2024)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
von: Jiang, Songtao, et al.
Veröffentlicht: (2024)
von: Jiang, Songtao, et al.
Veröffentlicht: (2024)
PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation
von: Ma, Jian, et al.
Veröffentlicht: (2023)
von: Ma, Jian, et al.
Veröffentlicht: (2023)
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
MoPE: Mixture of Prompt Experts for Parameter-Efficient and Scalable Multimodal Fusion
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
ChartMoE: Mixture of Diversely Aligned Expert Connector for Chart Understanding
von: Xu, Zhengzhuo, et al.
Veröffentlicht: (2024)
von: Xu, Zhengzhuo, et al.
Veröffentlicht: (2024)
X2Edit: Revisiting Arbitrary-Instruction Image Editing through Self-Constructed Data and Task-Aware Representation Learning
von: Ma, Jian, et al.
Veröffentlicht: (2025)
von: Ma, Jian, et al.
Veröffentlicht: (2025)
Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language Models
von: Luo, Jun, et al.
Veröffentlicht: (2024)
von: Luo, Jun, et al.
Veröffentlicht: (2024)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation
von: Lu, Yujie, et al.
Veröffentlicht: (2026)
von: Lu, Yujie, et al.
Veröffentlicht: (2026)
MoPD: Mixture-of-Prompts Distillation for Vision-Language Models
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
LLaVA-MoLE: Sparse Mixture of LoRA Experts for Mitigating Data Conflicts in Instruction Finetuning MLLMs
von: Chen, Shaoxiang, et al.
Veröffentlicht: (2024)
von: Chen, Shaoxiang, et al.
Veröffentlicht: (2024)
Multimodal LLM With Hierarchical Mixture-of-Experts for VQA on 3D Brain MRI
von: Vepa, Arvind Murari, et al.
Veröffentlicht: (2025)
von: Vepa, Arvind Murari, et al.
Veröffentlicht: (2025)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
A High-Quality Text-Rich Image Instruction Tuning Dataset via Hybrid Instruction Generation
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
MoExtend: Tuning New Experts for Modality and Task Extension
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
Mixture of Group Experts for Learning Invariant Representations
von: Kang, Lei, et al.
Veröffentlicht: (2025)
von: Kang, Lei, et al.
Veröffentlicht: (2025)
LiME: Lightweight Mixture of Experts for Efficient Multimodal Multi-task Learning
von: Kowsher, Md, et al.
Veröffentlicht: (2026)
von: Kowsher, Md, et al.
Veröffentlicht: (2026)
InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation
von: Xiao, Jinqi, et al.
Veröffentlicht: (2025)
von: Xiao, Jinqi, et al.
Veröffentlicht: (2025)
3D-MoE: A Mixture-of-Experts Multi-modal LLM for 3D Vision and Pose Diffusion via Rectified Flow
von: Ma, Yueen, et al.
Veröffentlicht: (2025)
von: Ma, Yueen, et al.
Veröffentlicht: (2025)
A Novel Trustworthy Video Summarization Algorithm Through a Mixture of LoRA Experts
von: Du, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Du, Wenzhuo, et al.
Veröffentlicht: (2025)
Learning from Fine-Grained Visual Discrepancies: Mitigating Multimodal Hallucinations via In-Context Visual Contrastive Optimization
von: Deng, Haolin, et al.
Veröffentlicht: (2026)
von: Deng, Haolin, et al.
Veröffentlicht: (2026)
AesBench: An Expert Benchmark for Multimodal Large Language Models on Image Aesthetics Perception
von: Huang, Yipo, et al.
Veröffentlicht: (2024)
von: Huang, Yipo, et al.
Veröffentlicht: (2024)
SP-MoMamba: Superpixel-driven Mixture of State Space Experts for Efficient Image Super-Resolution
von: Zou, Wenbin, et al.
Veröffentlicht: (2026)
von: Zou, Wenbin, et al.
Veröffentlicht: (2026)
p-MoD: Building Mixture-of-Depths MLLMs via Progressive Ratio Decay
von: Zhang, Jun, et al.
Veröffentlicht: (2024)
von: Zhang, Jun, et al.
Veröffentlicht: (2024)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
Omni-SMoLA: Boosting Generalist Multimodal Models with Soft Mixture of Low-rank Experts
von: Wu, Jialin, et al.
Veröffentlicht: (2023)
von: Wu, Jialin, et al.
Veröffentlicht: (2023)
TrueMoE: Dual-Routing Mixture of Discriminative Experts for Synthetic Image Detection
von: Zhang, Laixin, et al.
Veröffentlicht: (2025)
von: Zhang, Laixin, et al.
Veröffentlicht: (2025)
PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding
von: Wang, Nan, et al.
Veröffentlicht: (2026)
von: Wang, Nan, et al.
Veröffentlicht: (2026)
GazeMoE: Perception of Gaze Target with Mixture-of-Experts
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2026)
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2026)
Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
MoME: Mixture of Visual Language Medical Experts for Medical Imaging Segmentation
von: Rezvani, Arghavan, et al.
Veröffentlicht: (2025)
von: Rezvani, Arghavan, et al.
Veröffentlicht: (2025)
Low-Rank Mixture-of-Experts for Continual Medical Image Segmentation
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
ICM-Assistant: Instruction-tuning Multimodal Large Language Models for Rule-based Explainable Image Content Moderation
von: Wu, Mengyang, et al.
Veröffentlicht: (2024)
von: Wu, Mengyang, et al.
Veröffentlicht: (2024)
MMMG: A Massive, Multidisciplinary, Multi-Tier Generation Benchmark for Text-to-Image Reasoning
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image Generation
von: He, Xuehai, et al.
Veröffentlicht: (2024) -
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
von: Huang, Yushi, et al.
Veröffentlicht: (2025) -
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
von: Jiang, Songtao, et al.
Veröffentlicht: (2024) -
PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation
von: Ma, Jian, et al.
Veröffentlicht: (2023) -
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)