MEMoE: Enhancing Model Editing with Mixture of Experts Adaptors
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Renzhi, Li, Piji |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
par: Wang, Renzhi, et autres
Publié: (2024)
par: Wang, Renzhi, et autres
Publié: (2024)
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning
par: Wang, Renzhi, et autres
Publié: (2024)
par: Wang, Renzhi, et autres
Publié: (2024)
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
par: Xia, Runze, et autres
Publié: (2025)
par: Xia, Runze, et autres
Publié: (2025)
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
par: Chen, Qizhou, et autres
Publié: (2024)
par: Chen, Qizhou, et autres
Publié: (2024)
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
par: Ni, Xuanfan, et autres
Publié: (2024)
par: Ni, Xuanfan, et autres
Publié: (2024)
Temporal Guidance for Large Language Models
par: Zheng, Hong-Kai, et autres
Publié: (2026)
par: Zheng, Hong-Kai, et autres
Publié: (2026)
MMoE: Enhancing Multimodal Models with Mixtures of Multimodal Interaction Experts
par: Yu, Haofei, et autres
Publié: (2023)
par: Yu, Haofei, et autres
Publié: (2023)
Generating Diverse Training Samples for Relation Extraction with Large Language Models
par: Li, Zexuan, et autres
Publié: (2025)
par: Li, Zexuan, et autres
Publié: (2025)
M-BRe: Discovering Training Samples for Relation Extraction from Unlabeled Texts with Large Language Models
par: Li, Zexuan, et autres
Publié: (2025)
par: Li, Zexuan, et autres
Publié: (2025)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
par: Jing, Linglin, et autres
Publié: (2025)
par: Jing, Linglin, et autres
Publié: (2025)
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA
par: Li, Jiaang, et autres
Publié: (2024)
par: Li, Jiaang, et autres
Publié: (2024)
ReXMoE: Reusing Experts with Minimal Overhead in Mixture-of-Experts
par: Tan, Zheyue, et autres
Publié: (2025)
par: Tan, Zheyue, et autres
Publié: (2025)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
par: Wang, An, et autres
Publié: (2024)
par: Wang, An, et autres
Publié: (2024)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
par: Guo, Hongcheng, et autres
Publié: (2025)
par: Guo, Hongcheng, et autres
Publié: (2025)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
par: Feng, Yuchen, et autres
Publié: (2025)
par: Feng, Yuchen, et autres
Publié: (2025)
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
par: Dai, Damai, et autres
Publié: (2024)
par: Dai, Damai, et autres
Publié: (2024)
Unveiling Super Experts in Mixture-of-Experts Large Language Models
par: Su, Zunhai, et autres
Publié: (2025)
par: Su, Zunhai, et autres
Publié: (2025)
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
par: Wang, Zihan, et autres
Publié: (2025)
par: Wang, Zihan, et autres
Publié: (2025)
Characteristic AI Agents via Large Language Models
par: Wang, Xi, et autres
Publié: (2024)
par: Wang, Xi, et autres
Publié: (2024)
5W1H Extraction With Large Language Models
par: Cao, Yang, et autres
Publié: (2024)
par: Cao, Yang, et autres
Publié: (2024)
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
par: Christoforos, Alexandros, et autres
Publié: (2025)
par: Christoforos, Alexandros, et autres
Publié: (2025)
Improve Language Model and Brain Alignment via Associative Memory
par: Yin, Congchi, et autres
Publié: (2025)
par: Yin, Congchi, et autres
Publié: (2025)
Scalable Model Editing via Customized Expert Networks
par: Yao, Zihan, et autres
Publié: (2024)
par: Yao, Zihan, et autres
Publié: (2024)
Decoding the Echoes of Vision from fMRI: Memory Disentangling for Past Semantic Information
par: Xia, Runze, et autres
Publié: (2024)
par: Xia, Runze, et autres
Publié: (2024)
Hallucination Mitigating for Medical Report Generation
par: Zhao, Ruoqing, et autres
Publié: (2026)
par: Zhao, Ruoqing, et autres
Publié: (2026)
CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning
par: Lan, Yangsong, et autres
Publié: (2026)
par: Lan, Yangsong, et autres
Publié: (2026)
An Empirical Investigation of Domain Adaptation Ability for Chinese Spelling Check Models
par: Wang, Xi, et autres
Publié: (2024)
par: Wang, Xi, et autres
Publié: (2024)
Mixture of Lookup Experts
par: Jie, Shibo, et autres
Publié: (2025)
par: Jie, Shibo, et autres
Publié: (2025)
LLaVA-CMoE: Towards Continual Mixture of Experts for Large Vision-Language Models
par: Zhao, Hengyuan, et autres
Publié: (2025)
par: Zhao, Hengyuan, et autres
Publié: (2025)
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
par: Lin, Hongzhan, et autres
Publié: (2024)
par: Lin, Hongzhan, et autres
Publié: (2024)
Bridging Sequence-Structure Alignment in RNA Foundation Models
par: Yang, Heng, et autres
Publié: (2024)
par: Yang, Heng, et autres
Publié: (2024)
Mixture of Neuron Experts
par: Cheng, Runxi, et autres
Publié: (2025)
par: Cheng, Runxi, et autres
Publié: (2025)
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
par: Huang, Hukai, et autres
Publié: (2024)
par: Huang, Hukai, et autres
Publié: (2024)
MoDEM: Mixture of Domain Expert Models
par: Simonds, Toby, et autres
Publié: (2024)
par: Simonds, Toby, et autres
Publié: (2024)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
par: Tang, Yehui, et autres
Publié: (2025)
par: Tang, Yehui, et autres
Publié: (2025)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
par: Takashiro, Shota, et autres
Publié: (2026)
par: Takashiro, Shota, et autres
Publié: (2026)
MH-MoE: Multi-Head Mixture-of-Experts
par: Huang, Shaohan, et autres
Publié: (2024)
par: Huang, Shaohan, et autres
Publié: (2024)
MoBiLE: Efficient Mixture-of-Experts Inference on Consumer GPU with Mixture of Big Little Experts
par: Zhao, Yushu, et autres
Publié: (2025)
par: Zhao, Yushu, et autres
Publié: (2025)
Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel
par: Zheng, Chuanyang, et autres
Publié: (2025)
par: Zheng, Chuanyang, et autres
Publié: (2025)
Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling
par: Ran, Junfeng, et autres
Publié: (2025)
par: Ran, Junfeng, et autres
Publié: (2025)
Documents similaires
-
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
par: Wang, Renzhi, et autres
Publié: (2024) -
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning
par: Wang, Renzhi, et autres
Publié: (2024) -
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
par: Xia, Runze, et autres
Publié: (2025) -
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
par: Chen, Qizhou, et autres
Publié: (2024) -
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
par: Ni, Xuanfan, et autres
Publié: (2024)