MING-MOE: Enhancing Medical Multi-Task Learning in Large Language Models with Sparse Mixture of Low-Rank Adapter Experts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liao, Yusheng, Jiang, Shuyang, Wang, Yu, Wang, Yanfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TAIA: Large Language Models are Out-of-Distribution Data Learners
von: Jiang, Shuyang, et al.
Veröffentlicht: (2024)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2024)
Towards Omni-RAG: Comprehensive Retrieval-Augmented Generation for Large Language Models in Medical Applications
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
MedCare: Advancing Medical LLMs through Decoupling Clinical Alignment and Knowledge Aggregation
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
MedS$^3$: Towards Medical Slow Thinking with Self-Evolved Soft Dual-sided Process Supervision
von: Jiang, Shuyang, et al.
Veröffentlicht: (2025)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2025)
Leveraging Diverse Modeling Contexts with Collaborating Learning for Neural Machine Translation
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
Overthinking Reduction with Decoupled Rewards and Curriculum Data Scheduling
von: Jiang, Shuyang, et al.
Veröffentlicht: (2025)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2025)
HeteroRAG: A Heterogeneous Retrieval-Augmented Generation Framework for Medical Vision Language Tasks
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
Automatic Interactive Evaluation for Large Language Models with State Aware Patient Simulator
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
MALoRA: Mixture of Asymmetric Low-Rank Adaptation for Enhanced Multi-Task Learning
von: Wang, Xujia, et al.
Veröffentlicht: (2024)
von: Wang, Xujia, et al.
Veröffentlicht: (2024)
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
MM-SAP: A Comprehensive Benchmark for Assessing Self-Awareness of Multimodal Large Language Models in Perception
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
Training-Free Bayesianization for Low-Rank Adapters of Large Language Models
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
Ensembles of Low-Rank Expert Adapters
von: Li, Yinghao, et al.
Veröffentlicht: (2025)
von: Li, Yinghao, et al.
Veröffentlicht: (2025)
Med-PMC: Medical Personalized Multi-modal Consultation with a Proactive Ask-First-Observe-Next Paradigm
von: Liu, Hongcheng, et al.
Veröffentlicht: (2024)
von: Liu, Hongcheng, et al.
Veröffentlicht: (2024)
Miner:Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Models
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
DSVD: Dynamic Self-Verify Decoding for Faithful Generation in Large Language Models
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
DiffLoRA: Differential Low-Rank Adapters for Large Language Models
von: Misrahi, Alexandre, et al.
Veröffentlicht: (2025)
von: Misrahi, Alexandre, et al.
Veröffentlicht: (2025)
Block Circulant Adapter for Large Language Models
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
EHR-R1: A Reasoning-Enhanced Foundational Language Model for Electronic Health Record Analysis
von: Liao, Yusheng, et al.
Veröffentlicht: (2025)
von: Liao, Yusheng, et al.
Veröffentlicht: (2025)
DictLLM: Harnessing Key-Value Data Structures with Large Language Models for Enhanced Medical Diagnostics
von: Guo, YiQiu, et al.
Veröffentlicht: (2024)
von: Guo, YiQiu, et al.
Veröffentlicht: (2024)
AdaMoLE: Fine-Tuning Large Language Models with Adaptive Mixture of Low-Rank Adaptation Experts
von: Liu, Zefang, et al.
Veröffentlicht: (2024)
von: Liu, Zefang, et al.
Veröffentlicht: (2024)
A Survey on Mixture of Experts in Large Language Models
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
Inducing Generalization across Languages and Tasks using Featurized Low-Rank Mixtures
von: Lin, Chu-Cheng, et al.
Veröffentlicht: (2024)
von: Lin, Chu-Cheng, et al.
Veröffentlicht: (2024)
SparseDoctor: Towards Efficient Chat Doctor with Mixture of Experts Enhanced Large Language Models
von: Zhang, Jianbin, et al.
Veröffentlicht: (2025)
von: Zhang, Jianbin, et al.
Veröffentlicht: (2025)
Multiple Choice Learning of Low-Rank Adapters for Language Modeling
von: Letzelter, Victor, et al.
Veröffentlicht: (2025)
von: Letzelter, Victor, et al.
Veröffentlicht: (2025)
Unveiling Super Experts in Mixture-of-Experts Large Language Models
von: Su, Zunhai, et al.
Veröffentlicht: (2025)
von: Su, Zunhai, et al.
Veröffentlicht: (2025)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
von: Jing, Linglin, et al.
Veröffentlicht: (2025)
von: Jing, Linglin, et al.
Veröffentlicht: (2025)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
Diversifying the Expert Knowledge for Task-Agnostic Pruning in Sparse Mixture-of-Experts
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism
von: Sun, Mengyang, et al.
Veröffentlicht: (2026)
von: Sun, Mengyang, et al.
Veröffentlicht: (2026)
Enhancing AI Safety Through the Fusion of Low Rank Adapters
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2024)
von: Gudipudi, Satya Swaroop, et al.
Veröffentlicht: (2024)
Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach
von: Li, Haolin, et al.
Veröffentlicht: (2026)
von: Li, Haolin, et al.
Veröffentlicht: (2026)
Selecting Auxiliary Data via Neural Tangent Kernels for Low-Resource Domains
von: Wang, Pingjie, et al.
Veröffentlicht: (2025)
von: Wang, Pingjie, et al.
Veröffentlicht: (2025)
Drawing the Line: Enhancing Trustworthiness of MLLMs Through the Power of Refusal
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
KG-Rank: Enhancing Large Language Models for Medical QA with Knowledge Graphs and Ranking Techniques
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
DICE: Structured Reasoning in LLMs through SLM-Guided Chain-of-Thought Correction
von: Li, Yiqi, et al.
Veröffentlicht: (2025)
von: Li, Yiqi, et al.
Veröffentlicht: (2025)
Text-Routed Sparse Mixture-of-Experts Model with Explanation and Temporal Alignment for Multi-Modal Sentiment Analysis
von: Rao, Dongning, et al.
Veröffentlicht: (2025)
von: Rao, Dongning, et al.
Veröffentlicht: (2025)
A Sequential Optimal Learning Approach to Automated Prompt Engineering in Large Language Models
von: Wang, Shuyang, et al.
Veröffentlicht: (2025)
von: Wang, Shuyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TAIA: Large Language Models are Out-of-Distribution Data Learners
von: Jiang, Shuyang, et al.
Veröffentlicht: (2024) -
Towards Omni-RAG: Comprehensive Retrieval-Augmented Generation for Large Language Models in Medical Applications
von: Chen, Zhe, et al.
Veröffentlicht: (2025) -
MedCare: Advancing Medical LLMs through Decoupling Clinical Alignment and Knowledge Aggregation
von: Liao, Yusheng, et al.
Veröffentlicht: (2024) -
ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents
von: Liao, Yusheng, et al.
Veröffentlicht: (2024) -
MedS$^3$: Towards Medical Slow Thinking with Self-Evolved Soft Dual-sided Process Supervision
von: Jiang, Shuyang, et al.
Veröffentlicht: (2025)