Efficiently Democratizing Medical LLMs for 50 Languages via a Mixture of Language Family Experts
Fuente:
arXiv
Guardado en:
| Autores principales: | Zheng, Guorui, Wang, Xidong, Liang, Juhao, Chen, Nuo, Zheng, Yuping, Wang, Benyou |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Smurfs: Multi-Agent System using Context-Efficient DFSDT for Tool Planning
por: Chen, Junzhi, et al.
Publicado: (2024)
por: Chen, Junzhi, et al.
Publicado: (2024)
Apollo: A Lightweight Multilingual Medical LLM towards Democratizing Medical AI to 6B People
por: Wang, Xidong, et al.
Publicado: (2024)
por: Wang, Xidong, et al.
Publicado: (2024)
LLMs for Doctors: Leveraging Medical LLMs to Assist Doctors, Not Replace Them
por: Xie, Wenya, et al.
Publicado: (2024)
por: Xie, Wenya, et al.
Publicado: (2024)
Online Training of Large Language Models: Learn while chatting
por: Liang, Juhao, et al.
Publicado: (2024)
por: Liang, Juhao, et al.
Publicado: (2024)
Enabling Doctor-Centric Medical AI with LLMs through Workflow-Aligned Tasks and Benchmarks
por: Xie, Wenya, et al.
Publicado: (2025)
por: Xie, Wenya, et al.
Publicado: (2025)
Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs
por: Song, Dingjie, et al.
Publicado: (2024)
por: Song, Dingjie, et al.
Publicado: (2024)
LongLLaVA: Scaling Multi-modal LLMs to 1000 Images Efficiently via a Hybrid Architecture
por: Wang, Xidong, et al.
Publicado: (2024)
por: Wang, Xidong, et al.
Publicado: (2024)
HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs
por: Chen, Junying, et al.
Publicado: (2024)
por: Chen, Junying, et al.
Publicado: (2024)
HiMed: Incentivizing Hindi Reasoning in Medical LLMs
por: Jiang, Dingfeng, et al.
Publicado: (2026)
por: Jiang, Dingfeng, et al.
Publicado: (2026)
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion
por: Zhu, Jianqing, et al.
Publicado: (2024)
por: Zhu, Jianqing, et al.
Publicado: (2024)
Less, but Better: Efficient Multilingual Expansion for LLMs via Layer-wise Mixture-of-Experts
por: Zhang, Xue, et al.
Publicado: (2025)
por: Zhang, Xue, et al.
Publicado: (2025)
Understanding and Leveraging the Expert Specialization of Context Faithfulness in Mixture-of-Experts LLMs
por: Bai, Jun, et al.
Publicado: (2025)
por: Bai, Jun, et al.
Publicado: (2025)
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
por: Jiang, Songtao, et al.
Publicado: (2024)
por: Jiang, Songtao, et al.
Publicado: (2024)
GigaChat Family: Efficient Russian Language Modeling Through Mixture of Experts Architecture
por: GigaChat team, et al.
Publicado: (2025)
por: GigaChat team, et al.
Publicado: (2025)
Alignment at Pre-training! Towards Native Alignment for Arabic LLMs
por: Liang, Juhao, et al.
Publicado: (2024)
por: Liang, Juhao, et al.
Publicado: (2024)
Roadmap towards Superhuman Speech Understanding using Large Language Models
por: Bu, Fan, et al.
Publicado: (2024)
por: Bu, Fan, et al.
Publicado: (2024)
CoD, Towards an Interpretable Medical Agent using Chain of Diagnosis
por: Chen, Junying, et al.
Publicado: (2024)
por: Chen, Junying, et al.
Publicado: (2024)
When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models
por: Yoon, Youngsik, et al.
Publicado: (2026)
por: Yoon, Youngsik, et al.
Publicado: (2026)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
por: Jing, Linglin, et al.
Publicado: (2025)
por: Jing, Linglin, et al.
Publicado: (2025)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
por: Guo, Hongcheng, et al.
Publicado: (2025)
por: Guo, Hongcheng, et al.
Publicado: (2025)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
por: Feng, Yuchen, et al.
Publicado: (2025)
por: Feng, Yuchen, et al.
Publicado: (2025)
MING-MOE: Enhancing Medical Multi-Task Learning in Large Language Models with Sparse Mixture of Low-Rank Adapter Experts
por: Liao, Yusheng, et al.
Publicado: (2024)
por: Liao, Yusheng, et al.
Publicado: (2024)
From Beginner to Expert: Modeling Medical Knowledge into General LLMs
por: Li, Qiang, et al.
Publicado: (2023)
por: Li, Qiang, et al.
Publicado: (2023)
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
por: DeepSeek-AI, et al.
Publicado: (2024)
por: DeepSeek-AI, et al.
Publicado: (2024)
Understanding Multilingualism in Mixture-of-Experts LLMs: Routing Mechanism, Expert Specialization, and Layerwise Steering
por: Chen, Yuxin, et al.
Publicado: (2026)
por: Chen, Yuxin, et al.
Publicado: (2026)
Unveiling Super Experts in Mixture-of-Experts Large Language Models
por: Su, Zunhai, et al.
Publicado: (2025)
por: Su, Zunhai, et al.
Publicado: (2025)
FactorLLM: Factorizing Knowledge via Mixture of Experts for Large Language Models
por: Zhao, Zhongyu, et al.
Publicado: (2024)
por: Zhao, Zhongyu, et al.
Publicado: (2024)
MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria
por: Ge, Wentao, et al.
Publicado: (2023)
por: Ge, Wentao, et al.
Publicado: (2023)
LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages
por: Huang, Xuhan, et al.
Publicado: (2024)
por: Huang, Xuhan, et al.
Publicado: (2024)
Mixture of Heterogeneous Grouped Experts for Language Modeling
por: Ma, Zhicheng, et al.
Publicado: (2026)
por: Ma, Zhicheng, et al.
Publicado: (2026)
OpenMoE: An Early Effort on Open Mixture-of-Experts Language Models
por: Xue, Fuzhao, et al.
Publicado: (2024)
por: Xue, Fuzhao, et al.
Publicado: (2024)
HuatuoGPT-II, One-stage Training for Medical Adaption of LLMs
por: Chen, Junying, et al.
Publicado: (2023)
por: Chen, Junying, et al.
Publicado: (2023)
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
por: Chen, Qizhou, et al.
Publicado: (2024)
por: Chen, Qizhou, et al.
Publicado: (2024)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
por: Wang, An, et al.
Publicado: (2024)
por: Wang, An, et al.
Publicado: (2024)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
por: Huang, Yushi, et al.
Publicado: (2025)
por: Huang, Yushi, et al.
Publicado: (2025)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
por: Tang, Yehui, et al.
Publicado: (2025)
por: Tang, Yehui, et al.
Publicado: (2025)
A Survey on Mixture of Experts in Large Language Models
por: Cai, Weilin, et al.
Publicado: (2024)
por: Cai, Weilin, et al.
Publicado: (2024)
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
por: Jiang, Fan, et al.
Publicado: (2026)
por: Jiang, Fan, et al.
Publicado: (2026)
HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
por: Chen, Junying, et al.
Publicado: (2024)
por: Chen, Junying, et al.
Publicado: (2024)
Mixture of Experts for Low-Resource LLMs
por: Joseph, Ori Bar, et al.
Publicado: (2026)
por: Joseph, Ori Bar, et al.
Publicado: (2026)
Ejemplares similares
-
Smurfs: Multi-Agent System using Context-Efficient DFSDT for Tool Planning
por: Chen, Junzhi, et al.
Publicado: (2024) -
Apollo: A Lightweight Multilingual Medical LLM towards Democratizing Medical AI to 6B People
por: Wang, Xidong, et al.
Publicado: (2024) -
LLMs for Doctors: Leveraging Medical LLMs to Assist Doctors, Not Replace Them
por: Xie, Wenya, et al.
Publicado: (2024) -
Online Training of Large Language Models: Learn while chatting
por: Liang, Juhao, et al.
Publicado: (2024) -
Enabling Doctor-Centric Medical AI with LLMs through Workflow-Aligned Tasks and Benchmarks
por: Xie, Wenya, et al.
Publicado: (2025)