MEMoE: Enhancing Model Editing with Mixture of Experts Adaptors
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Renzhi, Li, Piji |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
por: Wang, Renzhi, et al.
Publicado: (2024)
por: Wang, Renzhi, et al.
Publicado: (2024)
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning
por: Wang, Renzhi, et al.
Publicado: (2024)
por: Wang, Renzhi, et al.
Publicado: (2024)
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
por: Xia, Runze, et al.
Publicado: (2025)
por: Xia, Runze, et al.
Publicado: (2025)
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
por: Chen, Qizhou, et al.
Publicado: (2024)
por: Chen, Qizhou, et al.
Publicado: (2024)
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
por: Ni, Xuanfan, et al.
Publicado: (2024)
por: Ni, Xuanfan, et al.
Publicado: (2024)
Temporal Guidance for Large Language Models
por: Zheng, Hong-Kai, et al.
Publicado: (2026)
por: Zheng, Hong-Kai, et al.
Publicado: (2026)
MMoE: Enhancing Multimodal Models with Mixtures of Multimodal Interaction Experts
por: Yu, Haofei, et al.
Publicado: (2023)
por: Yu, Haofei, et al.
Publicado: (2023)
Generating Diverse Training Samples for Relation Extraction with Large Language Models
por: Li, Zexuan, et al.
Publicado: (2025)
por: Li, Zexuan, et al.
Publicado: (2025)
M-BRe: Discovering Training Samples for Relation Extraction from Unlabeled Texts with Large Language Models
por: Li, Zexuan, et al.
Publicado: (2025)
por: Li, Zexuan, et al.
Publicado: (2025)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
por: Jing, Linglin, et al.
Publicado: (2025)
por: Jing, Linglin, et al.
Publicado: (2025)
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA
por: Li, Jiaang, et al.
Publicado: (2024)
por: Li, Jiaang, et al.
Publicado: (2024)
ReXMoE: Reusing Experts with Minimal Overhead in Mixture-of-Experts
por: Tan, Zheyue, et al.
Publicado: (2025)
por: Tan, Zheyue, et al.
Publicado: (2025)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
por: Wang, An, et al.
Publicado: (2024)
por: Wang, An, et al.
Publicado: (2024)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
por: Guo, Hongcheng, et al.
Publicado: (2025)
por: Guo, Hongcheng, et al.
Publicado: (2025)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
por: Feng, Yuchen, et al.
Publicado: (2025)
por: Feng, Yuchen, et al.
Publicado: (2025)
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
por: Dai, Damai, et al.
Publicado: (2024)
por: Dai, Damai, et al.
Publicado: (2024)
Unveiling Super Experts in Mixture-of-Experts Large Language Models
por: Su, Zunhai, et al.
Publicado: (2025)
por: Su, Zunhai, et al.
Publicado: (2025)
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
por: Wang, Zihan, et al.
Publicado: (2025)
por: Wang, Zihan, et al.
Publicado: (2025)
Characteristic AI Agents via Large Language Models
por: Wang, Xi, et al.
Publicado: (2024)
por: Wang, Xi, et al.
Publicado: (2024)
5W1H Extraction With Large Language Models
por: Cao, Yang, et al.
Publicado: (2024)
por: Cao, Yang, et al.
Publicado: (2024)
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
por: Christoforos, Alexandros, et al.
Publicado: (2025)
por: Christoforos, Alexandros, et al.
Publicado: (2025)
Improve Language Model and Brain Alignment via Associative Memory
por: Yin, Congchi, et al.
Publicado: (2025)
por: Yin, Congchi, et al.
Publicado: (2025)
Scalable Model Editing via Customized Expert Networks
por: Yao, Zihan, et al.
Publicado: (2024)
por: Yao, Zihan, et al.
Publicado: (2024)
Decoding the Echoes of Vision from fMRI: Memory Disentangling for Past Semantic Information
por: Xia, Runze, et al.
Publicado: (2024)
por: Xia, Runze, et al.
Publicado: (2024)
Hallucination Mitigating for Medical Report Generation
por: Zhao, Ruoqing, et al.
Publicado: (2026)
por: Zhao, Ruoqing, et al.
Publicado: (2026)
CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning
por: Lan, Yangsong, et al.
Publicado: (2026)
por: Lan, Yangsong, et al.
Publicado: (2026)
An Empirical Investigation of Domain Adaptation Ability for Chinese Spelling Check Models
por: Wang, Xi, et al.
Publicado: (2024)
por: Wang, Xi, et al.
Publicado: (2024)
Mixture of Lookup Experts
por: Jie, Shibo, et al.
Publicado: (2025)
por: Jie, Shibo, et al.
Publicado: (2025)
LLaVA-CMoE: Towards Continual Mixture of Experts for Large Vision-Language Models
por: Zhao, Hengyuan, et al.
Publicado: (2025)
por: Zhao, Hengyuan, et al.
Publicado: (2025)
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
por: Lin, Hongzhan, et al.
Publicado: (2024)
por: Lin, Hongzhan, et al.
Publicado: (2024)
Bridging Sequence-Structure Alignment in RNA Foundation Models
por: Yang, Heng, et al.
Publicado: (2024)
por: Yang, Heng, et al.
Publicado: (2024)
Mixture of Neuron Experts
por: Cheng, Runxi, et al.
Publicado: (2025)
por: Cheng, Runxi, et al.
Publicado: (2025)
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
por: Huang, Hukai, et al.
Publicado: (2024)
por: Huang, Hukai, et al.
Publicado: (2024)
MoDEM: Mixture of Domain Expert Models
por: Simonds, Toby, et al.
Publicado: (2024)
por: Simonds, Toby, et al.
Publicado: (2024)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
por: Tang, Yehui, et al.
Publicado: (2025)
por: Tang, Yehui, et al.
Publicado: (2025)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
por: Takashiro, Shota, et al.
Publicado: (2026)
por: Takashiro, Shota, et al.
Publicado: (2026)
MH-MoE: Multi-Head Mixture-of-Experts
por: Huang, Shaohan, et al.
Publicado: (2024)
por: Huang, Shaohan, et al.
Publicado: (2024)
MoBiLE: Efficient Mixture-of-Experts Inference on Consumer GPU with Mixture of Big Little Experts
por: Zhao, Yushu, et al.
Publicado: (2025)
por: Zhao, Yushu, et al.
Publicado: (2025)
Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel
por: Zheng, Chuanyang, et al.
Publicado: (2025)
por: Zheng, Chuanyang, et al.
Publicado: (2025)
Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling
por: Ran, Junfeng, et al.
Publicado: (2025)
por: Ran, Junfeng, et al.
Publicado: (2025)
Ejemplares similares
-
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
por: Wang, Renzhi, et al.
Publicado: (2024) -
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning
por: Wang, Renzhi, et al.
Publicado: (2024) -
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
por: Xia, Runze, et al.
Publicado: (2025) -
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
por: Chen, Qizhou, et al.
Publicado: (2024) -
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
por: Ni, Xuanfan, et al.
Publicado: (2024)