Salvato in:
| Autori principali: | Rokah, Adam, Veress, Daniel, Caulk, Caleb, Sharan, Sourav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2601.15021 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EMoE: Eigenbasis-Guided Routing for Mixture-of-Experts
di: Cheng, Anzhe, et al.
Pubblicazione: (2026)
di: Cheng, Anzhe, et al.
Pubblicazione: (2026)
Stable Routing for Mixture-of-Experts in Class-Incremental Learning
di: Guo, Zirui, et al.
Pubblicazione: (2026)
di: Guo, Zirui, et al.
Pubblicazione: (2026)
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
di: Lou, Meng, et al.
Pubblicazione: (2026)
di: Lou, Meng, et al.
Pubblicazione: (2026)
Expert Race: A Flexible Routing Strategy for Scaling Diffusion Transformer with Mixture of Experts
di: Yuan, Yike, et al.
Pubblicazione: (2025)
di: Yuan, Yike, et al.
Pubblicazione: (2025)
Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language Models
di: Luo, Jun, et al.
Pubblicazione: (2024)
di: Luo, Jun, et al.
Pubblicazione: (2024)
Routers in Vision Mixture of Experts: An Empirical Study
di: Liu, Tianlin, et al.
Pubblicazione: (2024)
di: Liu, Tianlin, et al.
Pubblicazione: (2024)
Multilinear Mixture of Experts: Scalable Expert Specialization through Factorization
di: Oldfield, James, et al.
Pubblicazione: (2024)
di: Oldfield, James, et al.
Pubblicazione: (2024)
Domain-Specialized Object Detection via Model-Level Mixtures of Experts
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2026)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2026)
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
di: Liu, Yahui, et al.
Pubblicazione: (2025)
di: Liu, Yahui, et al.
Pubblicazione: (2025)
Merging Multi-Task Models via Weight-Ensembling Mixture of Experts
di: Tang, Anke, et al.
Pubblicazione: (2024)
di: Tang, Anke, et al.
Pubblicazione: (2024)
Video Relationship Detection Using Mixture of Experts
di: Shaabana, Ala, et al.
Pubblicazione: (2024)
di: Shaabana, Ala, et al.
Pubblicazione: (2024)
Robust Experts: the Effect of Adversarial Training on CNNs with Sparse Mixture-of-Experts Layers
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2025)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2025)
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging
di: Shen, Li, et al.
Pubblicazione: (2024)
di: Shen, Li, et al.
Pubblicazione: (2024)
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
di: Wang, Hongyu, et al.
Pubblicazione: (2025)
di: Wang, Hongyu, et al.
Pubblicazione: (2025)
Towards Adversarial Robustness of Model-Level Mixture-of-Experts Architectures for Semantic Segmentation
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2024)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2024)
LPT++: Efficient Training on Mixture of Long-tailed Experts
di: Dong, Bowen, et al.
Pubblicazione: (2024)
di: Dong, Bowen, et al.
Pubblicazione: (2024)
Mixture of Experts in Image Classification: What's the Sweet Spot?
di: Videau, Mathurin, et al.
Pubblicazione: (2024)
di: Videau, Mathurin, et al.
Pubblicazione: (2024)
Extracting Uncertainty Estimates from Mixtures of Experts for Semantic Segmentation
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2025)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2025)
BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma
di: Yang, Junlin, et al.
Pubblicazione: (2026)
di: Yang, Junlin, et al.
Pubblicazione: (2026)
Mixture of Group Experts for Learning Invariant Representations
di: Kang, Lei, et al.
Pubblicazione: (2025)
di: Kang, Lei, et al.
Pubblicazione: (2025)
Improving OOD Generalization of Pre-trained Encoders via Aligned Embedding-Space Ensembles
di: Peng, Shuman, et al.
Pubblicazione: (2024)
di: Peng, Shuman, et al.
Pubblicazione: (2024)
EC-DIT: Scaling Diffusion Transformers with Adaptive Expert-Choice Routing
di: Sun, Haotian, et al.
Pubblicazione: (2024)
di: Sun, Haotian, et al.
Pubblicazione: (2024)
Lightweight Metadata-Aware Mixture-of-Experts Masked Autoencoder for Earth Observation
di: Albughdadi, Mohanad
Pubblicazione: (2025)
di: Albughdadi, Mohanad
Pubblicazione: (2025)
From Sparse to Soft Mixtures of Experts
di: Puigcerver, Joan, et al.
Pubblicazione: (2023)
di: Puigcerver, Joan, et al.
Pubblicazione: (2023)
MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model Merging
di: Qiu, Zihuan, et al.
Pubblicazione: (2025)
di: Qiu, Zihuan, et al.
Pubblicazione: (2025)
Design and Behavior of Sparse Mixture-of-Experts Layers in CNN-based Semantic Segmentation
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2026)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2026)
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
di: Liu, Zhili, et al.
Pubblicazione: (2024)
di: Liu, Zhili, et al.
Pubblicazione: (2024)
Teacher-Guided Routing for Sparse Vision Mixture-of-Experts
di: Kada, Masahiro, et al.
Pubblicazione: (2026)
di: Kada, Masahiro, et al.
Pubblicazione: (2026)
Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of Experts
di: Liu, Yuejiang, et al.
Pubblicazione: (2024)
di: Liu, Yuejiang, et al.
Pubblicazione: (2024)
MoPD: Mixture-of-Prompts Distillation for Vision-Language Models
di: Chen, Yang, et al.
Pubblicazione: (2024)
di: Chen, Yang, et al.
Pubblicazione: (2024)
Mixture-of-Experts for Open Set Domain Adaptation: A Dual-Space Detection Approach
di: Du, Zhenbang, et al.
Pubblicazione: (2023)
di: Du, Zhenbang, et al.
Pubblicazione: (2023)
FLAVARS: A Multimodal Foundational Language and Vision Alignment Model for Remote Sensing
di: Corley, Isaac, et al.
Pubblicazione: (2025)
di: Corley, Isaac, et al.
Pubblicazione: (2025)
Parameter-Efficient Quantized Mixture-of-Experts Meets Vision-Language Instruction Tuning for Semiconductor Electron Micrograph Analysis
di: Srinivas, Sakhinana Sagar, et al.
Pubblicazione: (2024)
di: Srinivas, Sakhinana Sagar, et al.
Pubblicazione: (2024)
Sublinear Variational Optimization of Gaussian Mixture Models with Millions to Billions of Parameters
di: Salwig, Sebastian, et al.
Pubblicazione: (2025)
di: Salwig, Sebastian, et al.
Pubblicazione: (2025)
MoRE-Brain: Routed Mixture of Experts for Interpretable and Generalizable Cross-Subject fMRI Visual Decoding
di: Wei, Yuxiang, et al.
Pubblicazione: (2025)
di: Wei, Yuxiang, et al.
Pubblicazione: (2025)
FastMMoE: Accelerating Multimodal Large Language Models through Dynamic Expert Activation and Routing-Aware Token Pruning
di: Xia, Guoyang, et al.
Pubblicazione: (2025)
di: Xia, Guoyang, et al.
Pubblicazione: (2025)
IMPROVE: Iterative Model Pipeline Refinement and Optimization Leveraging LLM Experts
di: Xue, Eric, et al.
Pubblicazione: (2025)
di: Xue, Eric, et al.
Pubblicazione: (2025)
MoQE: Improve Quantization Model performance via Mixture of Quantization Experts
di: Zhang, Jinhao, et al.
Pubblicazione: (2025)
di: Zhang, Jinhao, et al.
Pubblicazione: (2025)
AMEND: A Mixture of Experts Framework for Long-tailed Trajectory Prediction
di: Mercurius, Ray Coden, et al.
Pubblicazione: (2024)
di: Mercurius, Ray Coden, et al.
Pubblicazione: (2024)
Adaptive Shared Experts with LoRA-Based Mixture of Experts for Multi-Task Learning
di: Yang, Minghao, et al.
Pubblicazione: (2025)
di: Yang, Minghao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
EMoE: Eigenbasis-Guided Routing for Mixture-of-Experts
di: Cheng, Anzhe, et al.
Pubblicazione: (2026) -
Stable Routing for Mixture-of-Experts in Class-Incremental Learning
di: Guo, Zirui, et al.
Pubblicazione: (2026) -
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
di: Lou, Meng, et al.
Pubblicazione: (2026) -
Expert Race: A Flexible Routing Strategy for Scaling Diffusion Transformer with Mixture of Experts
di: Yuan, Yike, et al.
Pubblicazione: (2025) -
Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language Models
di: Luo, Jun, et al.
Pubblicazione: (2024)