Guardado en:
| Autores principales: | Buginga, Gabriel, Silva, Edmundo de Souza e |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2405.15934 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Let the Experts Speak: Improving Survival Prediction & Calibration via Mixture-of-Experts Heads
por: Morrill, Todd, et al.
Publicado: (2025)
por: Morrill, Todd, et al.
Publicado: (2025)
Mixture-of-Clustered-Experts: Advancing Expert Specialization and Generalization in Instruction Tuning
por: Eo, Sugyeong, et al.
Publicado: (2025)
por: Eo, Sugyeong, et al.
Publicado: (2025)
Dual Mixture-of-Experts Framework for Discrete-Time Survival Analysis
por: Lee, Hyeonjun, et al.
Publicado: (2025)
por: Lee, Hyeonjun, et al.
Publicado: (2025)
Deep Clustering Survival Machines with Interpretable Expert Distributions
por: Hou, Bojian, et al.
Publicado: (2023)
por: Hou, Bojian, et al.
Publicado: (2023)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
por: Belenki, Lior, et al.
Publicado: (2025)
por: Belenki, Lior, et al.
Publicado: (2025)
Mixture of Online and Offline Experts for Non-stationary Time Series
por: Zhao, Zhilin, et al.
Publicado: (2022)
por: Zhao, Zhilin, et al.
Publicado: (2022)
Mixture of Experts in a Mixture of RL settings
por: Willi, Timon, et al.
Publicado: (2024)
por: Willi, Timon, et al.
Publicado: (2024)
Mixture of Experts Provably Detect and Learn the Latent Cluster Structure in Gradient-Based Learning
por: Kawata, Ryotaro, et al.
Publicado: (2025)
por: Kawata, Ryotaro, et al.
Publicado: (2025)
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
por: Park, Sejik
Publicado: (2024)
por: Park, Sejik
Publicado: (2024)
Similarity-Aware Mixture-of-Experts for Data-Efficient Continual Learning
por: Mclaughlin, Connor, et al.
Publicado: (2026)
por: Mclaughlin, Connor, et al.
Publicado: (2026)
Accurate Evaluation of Quickest Changepoint Detectors via Non-parametric Survival Analysis
por: Miyagawa, Taiki, et al.
Publicado: (2026)
por: Miyagawa, Taiki, et al.
Publicado: (2026)
Incremental Gaussian Mixture Clustering for Data Streams
por: Bhanderi, Aniket, et al.
Publicado: (2024)
por: Bhanderi, Aniket, et al.
Publicado: (2024)
Mixture of Autoencoder Experts Guidance using Unlabeled and Incomplete Data for Exploration in Reinforcement Learning
por: Malomgré, Elias, et al.
Publicado: (2025)
por: Malomgré, Elias, et al.
Publicado: (2025)
CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts
por: Yin, Xiangyang, et al.
Publicado: (2026)
por: Yin, Xiangyang, et al.
Publicado: (2026)
Path-Constrained Mixture-of-Experts
por: Gu, Zijin, et al.
Publicado: (2026)
por: Gu, Zijin, et al.
Publicado: (2026)
$μ$-Parametrization for Mixture of Experts
por: Małaśnicki, Jan, et al.
Publicado: (2025)
por: Małaśnicki, Jan, et al.
Publicado: (2025)
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
por: Nguyen, Dung V., et al.
Publicado: (2025)
por: Nguyen, Dung V., et al.
Publicado: (2025)
Expert-elicitation method for non-parametric joint priors using normalizing flows
por: Bockting, Florence, et al.
Publicado: (2024)
por: Bockting, Florence, et al.
Publicado: (2024)
Mixture of Raytraced Experts
por: Perin, Andrea, et al.
Publicado: (2025)
por: Perin, Andrea, et al.
Publicado: (2025)
Mixture of Lookup Experts
por: Jie, Shibo, et al.
Publicado: (2025)
por: Jie, Shibo, et al.
Publicado: (2025)
Mixture of Predefined Experts: Maximizing Data Usage on Vertical Federated Learning
por: Irureta, Jon, et al.
Publicado: (2026)
por: Irureta, Jon, et al.
Publicado: (2026)
Training of Neural Networks with Uncertain Data: A Mixture of Experts Approach
por: Luttner, Lucas
Publicado: (2023)
por: Luttner, Lucas
Publicado: (2023)
Speculating Experts Accelerates Inference for Mixture-of-Experts
por: Madan, Vivan, et al.
Publicado: (2026)
por: Madan, Vivan, et al.
Publicado: (2026)
Mixture of Experts Made Intrinsically Interpretable
por: Yang, Xingyi, et al.
Publicado: (2025)
por: Yang, Xingyi, et al.
Publicado: (2025)
Mixture of Experts (MoE): A Big Data Perspective
por: Gan, Wensheng, et al.
Publicado: (2025)
por: Gan, Wensheng, et al.
Publicado: (2025)
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
por: Liu, Zhili, et al.
Publicado: (2024)
por: Liu, Zhili, et al.
Publicado: (2024)
Mastery Guided Non-parametric Clustering to Scale-up Strategy Prediction
por: Shakya, Anup, et al.
Publicado: (2024)
por: Shakya, Anup, et al.
Publicado: (2024)
Robustness of Mixtures of Experts to Feature Noise
por: Sun, Dong, et al.
Publicado: (2026)
por: Sun, Dong, et al.
Publicado: (2026)
Temporally Extended Mixture-of-Experts Models
por: Shen, Zeyu, et al.
Publicado: (2026)
por: Shen, Zeyu, et al.
Publicado: (2026)
Generalizing GNNs with Tokenized Mixture of Experts
por: Guo, Xiaoguang, et al.
Publicado: (2026)
por: Guo, Xiaoguang, et al.
Publicado: (2026)
Mixture of Lookup Key-Value Experts
por: Wang, Zongcheng
Publicado: (2025)
por: Wang, Zongcheng
Publicado: (2025)
Hyperparameter Transfer with Mixture-of-Expert Layers
por: Jiang, Tianze, et al.
Publicado: (2026)
por: Jiang, Tianze, et al.
Publicado: (2026)
Mixture Compressor for Mixture-of-Experts LLMs Gains More
por: Huang, Wei, et al.
Publicado: (2024)
por: Huang, Wei, et al.
Publicado: (2024)
MC#: Mixture Compressor for Mixture-of-Experts Large Models
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Acquiring Diverse Skills using Curriculum Reinforcement Learning with Mixture of Experts
por: Celik, Onur, et al.
Publicado: (2024)
por: Celik, Onur, et al.
Publicado: (2024)
Mixture of Diverse Size Experts
por: Sun, Manxi, et al.
Publicado: (2024)
por: Sun, Manxi, et al.
Publicado: (2024)
Buffer Overflow in Mixture of Experts
por: Hayes, Jamie, et al.
Publicado: (2024)
por: Hayes, Jamie, et al.
Publicado: (2024)
Mixture of A Million Experts
por: He, Xu Owen
Publicado: (2024)
por: He, Xu Owen
Publicado: (2024)
Sparsity and Superposition in Mixture of Experts
por: Chaudhari, Marmik, et al.
Publicado: (2025)
por: Chaudhari, Marmik, et al.
Publicado: (2025)
Mixture of Concept Bottleneck Experts
por: De Santis, Francesco, et al.
Publicado: (2026)
por: De Santis, Francesco, et al.
Publicado: (2026)
Ejemplares similares
-
Let the Experts Speak: Improving Survival Prediction & Calibration via Mixture-of-Experts Heads
por: Morrill, Todd, et al.
Publicado: (2025) -
Mixture-of-Clustered-Experts: Advancing Expert Specialization and Generalization in Instruction Tuning
por: Eo, Sugyeong, et al.
Publicado: (2025) -
Dual Mixture-of-Experts Framework for Discrete-Time Survival Analysis
por: Lee, Hyeonjun, et al.
Publicado: (2025) -
Deep Clustering Survival Machines with Interpretable Expert Distributions
por: Hou, Bojian, et al.
Publicado: (2023) -
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
por: Belenki, Lior, et al.
Publicado: (2025)