MoE-GRPO: Optimizing Mixture-of-Experts via Reinforcement Learning in Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Ko, Dohwan, Park, Jinyoung, Choi, Seoung, Lee, Sanghyeok, Lee, Seohyun, Kim, Hyunwoo J. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
por: Choi, Joonmyung, et al.
Publicado: (2026)
por: Choi, Joonmyung, et al.
Publicado: (2026)
DeepVideo-R1: Video Reinforcement Fine-Tuning via Difficulty-aware Regressive GRPO
por: Park, Jinyoung, et al.
Publicado: (2025)
por: Park, Jinyoung, et al.
Publicado: (2025)
ODPG: Outfitting Diffusion with Pose Guided Condition
por: Lee, Seohyun, et al.
Publicado: (2025)
por: Lee, Seohyun, et al.
Publicado: (2025)
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
por: Lee, Sanghyeok, et al.
Publicado: (2024)
por: Lee, Sanghyeok, et al.
Publicado: (2024)
EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality
por: Lee, Sanghyeok, et al.
Publicado: (2024)
por: Lee, Sanghyeok, et al.
Publicado: (2024)
Prompt Learning via Meta-Regularization
por: Park, Jinyoung, et al.
Publicado: (2024)
por: Park, Jinyoung, et al.
Publicado: (2024)
Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval
por: Ko, Dohwan, et al.
Publicado: (2025)
por: Ko, Dohwan, et al.
Publicado: (2025)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
por: Park, Jihwan, et al.
Publicado: (2025)
por: Park, Jihwan, et al.
Publicado: (2025)
Fair-MoE: Fairness-Oriented Mixture of Experts in Vision-Language Models
por: Wang, Peiran, et al.
Publicado: (2025)
por: Wang, Peiran, et al.
Publicado: (2025)
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
por: Lin, Bin, et al.
Publicado: (2024)
por: Lin, Bin, et al.
Publicado: (2024)
LLaMo: Large Language Model-based Molecular Graph Assistant
por: Park, Jinyoung, et al.
Publicado: (2024)
por: Park, Jinyoung, et al.
Publicado: (2024)
MoE-GS: Mixture of Experts for Dynamic Gaussian Splatting
por: Jin, In-Hwan, et al.
Publicado: (2025)
por: Jin, In-Hwan, et al.
Publicado: (2025)
Representation Shift: Unifying Token Compression with FlashAttention
por: Choi, Joonmyung, et al.
Publicado: (2025)
por: Choi, Joonmyung, et al.
Publicado: (2025)
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
por: Jiang, Songtao, et al.
Publicado: (2024)
por: Jiang, Songtao, et al.
Publicado: (2024)
SocialNav-MoE: A Mixture-of-Experts Vision Language Model for Socially Compliant Navigation with Reinforcement Fine-Tuning
por: Kawabata, Tomohito, et al.
Publicado: (2025)
por: Kawabata, Tomohito, et al.
Publicado: (2025)
Enhancing Mixture-of-Experts Specialization via Cluster-Aware Upcycling
por: Chu, Sanghyeok, et al.
Publicado: (2026)
por: Chu, Sanghyeok, et al.
Publicado: (2026)
MoE3D: A Mixture-of-Experts Module for 3D Reconstruction
por: Wang, Zichen, et al.
Publicado: (2026)
por: Wang, Zichen, et al.
Publicado: (2026)
vid-TLDR: Training Free Token merging for Light-weight Video Transformer
por: Choi, Joonmyung, et al.
Publicado: (2024)
por: Choi, Joonmyung, et al.
Publicado: (2024)
ST-VLM: Kinematic Instruction Tuning for Spatio-Temporal Reasoning in Vision-Language Models
por: Ko, Dohwan, et al.
Publicado: (2025)
por: Ko, Dohwan, et al.
Publicado: (2025)
MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks
por: Zhu, Xingkui, et al.
Publicado: (2024)
por: Zhu, Xingkui, et al.
Publicado: (2024)
TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts
por: Xu, Yu, et al.
Publicado: (2026)
por: Xu, Yu, et al.
Publicado: (2026)
R^2MoE: Redundancy-Removal Mixture of Experts for Lifelong Concept Learning
por: Guo, Xiaohan, et al.
Publicado: (2025)
por: Guo, Xiaohan, et al.
Publicado: (2025)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
por: Cha, Juhan, et al.
Publicado: (2024)
por: Cha, Juhan, et al.
Publicado: (2024)
I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts
por: Xin, Jiayi, et al.
Publicado: (2025)
por: Xin, Jiayi, et al.
Publicado: (2025)
MoAI: Mixture of All Intelligence for Large Language and Vision Models
por: Lee, Byung-Kwan, et al.
Publicado: (2024)
por: Lee, Byung-Kwan, et al.
Publicado: (2024)
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
por: Lee, Ji Soo, et al.
Publicado: (2025)
por: Lee, Ji Soo, et al.
Publicado: (2025)
Semi-MoE: Mixture-of-Experts meets Semi-Supervised Histopathology Segmentation
por: Vu, Nguyen Lan Vi, et al.
Publicado: (2025)
por: Vu, Nguyen Lan Vi, et al.
Publicado: (2025)
GM-MoE: Low-Light Enhancement with Gated-Mechanism Mixture-of-Experts
por: Liao, Minwen, et al.
Publicado: (2025)
por: Liao, Minwen, et al.
Publicado: (2025)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
por: Liu, Yang, et al.
Publicado: (2026)
por: Liu, Yang, et al.
Publicado: (2026)
BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma
por: Yang, Junlin, et al.
Publicado: (2026)
por: Yang, Junlin, et al.
Publicado: (2026)
MoE-FFD: Mixture of Experts for Generalized and Parameter-Efficient Face Forgery Detection
por: Kong, Chenqi, et al.
Publicado: (2024)
por: Kong, Chenqi, et al.
Publicado: (2024)
CBDES MoE: Hierarchically Decoupled Mixture-of-Experts for Functional Modules in Autonomous Driving
por: Xiang, Qi, et al.
Publicado: (2025)
por: Xiang, Qi, et al.
Publicado: (2025)
RS-MoE: A Vision-Language Model with Mixture of Experts for Remote Sensing Image Captioning and Visual Question Answering
por: Lin, Hui, et al.
Publicado: (2024)
por: Lin, Hui, et al.
Publicado: (2024)
Adapted-MoE: Mixture of Experts with Test-Time Adaption for Anomaly Detection
por: Lei, Tianwu, et al.
Publicado: (2024)
por: Lei, Tianwu, et al.
Publicado: (2024)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
por: Kim, Jongha, et al.
Publicado: (2024)
por: Kim, Jongha, et al.
Publicado: (2024)
Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts
por: Zhang, Yue, et al.
Publicado: (2025)
por: Zhang, Yue, et al.
Publicado: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
por: Li, Yunxin, et al.
Publicado: (2024)
por: Li, Yunxin, et al.
Publicado: (2024)
MoE3D: Mixture of Experts meets Multi-Modal 3D Understanding
por: Li, Yu, et al.
Publicado: (2025)
por: Li, Yu, et al.
Publicado: (2025)
TabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token Focusing
por: Kim, Jongha, et al.
Publicado: (2025)
por: Kim, Jongha, et al.
Publicado: (2025)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
por: Zhang, Jihai, et al.
Publicado: (2024)
por: Zhang, Jihai, et al.
Publicado: (2024)
Ejemplares similares
-
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
por: Choi, Joonmyung, et al.
Publicado: (2026) -
DeepVideo-R1: Video Reinforcement Fine-Tuning via Difficulty-aware Regressive GRPO
por: Park, Jinyoung, et al.
Publicado: (2025) -
ODPG: Outfitting Diffusion with Pose Guided Condition
por: Lee, Seohyun, et al.
Publicado: (2025) -
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
por: Lee, Sanghyeok, et al.
Publicado: (2024) -
EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality
por: Lee, Sanghyeok, et al.
Publicado: (2024)