Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency without Model Sweeps
Fuente:
arXiv
Saved in:
| Main Authors: | Hai, Do Tien, Mai, Trung Nguyen, Nguyen, TrungTin, Ho, Nhat, Nguyen, Binh T., Drovandi, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model Selection for Gaussian-gated Gaussian Mixture of Experts Using Dendrograms of Mixing Measures
by: Thai, Tuan, et al.
Published: (2025)
by: Thai, Tuan, et al.
Published: (2025)
Fast Model Selection and Stable Optimization for Softmax-Gated Multinomial-Logistic Mixture of Experts Models
by: Tran, TrungKhang, et al.
Published: (2026)
by: Tran, TrungKhang, et al.
Published: (2026)
A General Theory for Softmax Gating Multinomial Logistic Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
by: Tran, TrungKhang, et al.
Published: (2026)
by: Tran, TrungKhang, et al.
Published: (2026)
A Unified Framework for Variable Selection in Model-Based Clustering with Missing Not at Random
by: Ho, Binh H., et al.
Published: (2025)
by: Ho, Binh H., et al.
Published: (2025)
Towards Convergence Rates for Parameter Estimation in Gaussian-gated Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
Convergence Rates for Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
CompeteSMoE -- Effective Training of Sparse Mixture of Experts via Competition
by: Pham, Quang, et al.
Published: (2024)
by: Pham, Quang, et al.
Published: (2024)
On Least Square Estimation in Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Modifications of the BIC for order selection in finite mixture models
by: Nguyen, Hien Duy, et al.
Published: (2025)
by: Nguyen, Hien Duy, et al.
Published: (2025)
On Bayesian Softmax-Gated Mixture-of-Experts Models
by: Bariletto, Nicola, et al.
Published: (2026)
by: Bariletto, Nicola, et al.
Published: (2026)
Sigmoid Gating is More Sample Efficient than Softmax Gating in Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Non-asymptotic oracle inequalities for the Lasso in high-dimensional mixture of experts
by: Nguyen, TrungTin, et al.
Published: (2020)
by: Nguyen, TrungTin, et al.
Published: (2020)
Is Temperature Sample Efficient for Softmax Gaussian Mixture of Experts?
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Risk Bounds for Mixture Density Estimation on Compact Domains via the $h$-Lifted Kullback--Leibler Divergence
by: Chong, Mark Chiu, et al.
Published: (2024)
by: Chong, Mark Chiu, et al.
Published: (2024)
Statistical Perspective of Top-K Sparse Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Dendrogram of mixing measures: Hierarchical clustering and model selection for finite mixture models
by: Do, Dat, et al.
Published: (2024)
by: Do, Dat, et al.
Published: (2024)
Lightspeed Geometric Dataset Distance via Sliced Optimal Transport
by: Nguyen, Khai, et al.
Published: (2025)
by: Nguyen, Khai, et al.
Published: (2025)
On Parameter Estimation in Deviated Gaussian Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
High-dimensional Many-to-many-to-many Mediation Analysis
by: Nguyen, Tien Dat, et al.
Published: (2026)
by: Nguyen, Tien Dat, et al.
Published: (2026)
On the large-sample limits of some Bayesian model evaluation statistics
by: Nguyen, Hien Duy, et al.
Published: (2025)
by: Nguyen, Hien Duy, et al.
Published: (2025)
Approximation rates for finite mixtures of location-scale models and fast least-squares estimators
by: Nguyen, Hien Duy, et al.
Published: (2025)
by: Nguyen, Hien Duy, et al.
Published: (2025)
RepLoRA: Reparameterizing Low-Rank Adaptation via the Perspective of Mixture of Experts
by: Truong, Tuan, et al.
Published: (2025)
by: Truong, Tuan, et al.
Published: (2025)
Summarizing Bayesian Nonparametric Mixture Posterior -- Sliced Optimal Transport Metrics for Gaussian Mixtures
by: Nguyen, Khai, et al.
Published: (2024)
by: Nguyen, Khai, et al.
Published: (2024)
On Minimax Estimation of Parameters in Softmax-Contaminated Mixture of Experts
by: Yan, Fanqi, et al.
Published: (2025)
by: Yan, Fanqi, et al.
Published: (2025)
Sparse classification with positive-confidence data in high dimensions
by: Mai, The Tien, et al.
Published: (2025)
by: Mai, The Tien, et al.
Published: (2025)
ExGra-Med: Extended Context Graph Alignment for Medical Vision-Language Models
by: Nguyen, Duy M. H., et al.
Published: (2024)
by: Nguyen, Duy M. H., et al.
Published: (2024)
Bayesian Wasserstein Repulsive Gaussian Mixture Models
by: Huang, Weipeng, et al.
Published: (2025)
by: Huang, Weipeng, et al.
Published: (2025)
Granular Columns of Binary‐Size Mixtures Collapse on a Horizontal Plane
by: Thanh‐Trung Vo, et al.
Published: (2025)
by: Thanh‐Trung Vo, et al.
Published: (2025)
Quadratic Gating Mixture of Experts: Statistical Insights into Self-Attention
by: Akbarian, Pedram, et al.
Published: (2024)
by: Akbarian, Pedram, et al.
Published: (2024)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
by: Pham, Tuan Minh, et al.
Published: (2026)
by: Pham, Tuan Minh, et al.
Published: (2026)
Structure-Aware E(3)-Invariant Molecular Conformer Aggregation Networks
by: Nguyen, Duy M. H., et al.
Published: (2024)
by: Nguyen, Duy M. H., et al.
Published: (2024)
Uncovering Heterogeneity of Solar Flare Mechanism With Mixture Models
by: Do, Bach Viet, et al.
Published: (2024)
by: Do, Bach Viet, et al.
Published: (2024)
Consistent information criteria for regularized regression and loss-based learning problems
by: Zhang, Qingyuan, et al.
Published: (2024)
by: Zhang, Qingyuan, et al.
Published: (2024)
Exact Sampling of Gibbs Measures with Estimated Losses
by: Frazier, David T., et al.
Published: (2024)
by: Frazier, David T., et al.
Published: (2024)
Mixture-of-experts Wishart model for covariance matrices with an application to Cancer drug screening
by: Mai, The Tien, et al.
Published: (2026)
by: Mai, The Tien, et al.
Published: (2026)
Calibrated Generalized Bayesian Inference
by: Frazier, David T., et al.
Published: (2023)
by: Frazier, David T., et al.
Published: (2023)
Sigmoid Self-Attention has Lower Sample Complexity than Softmax Self-Attention: A Mixture-of-Experts Perspective
by: Yan, Fanqi, et al.
Published: (2025)
by: Yan, Fanqi, et al.
Published: (2025)
High-dimensional sparse classification using exponential weighting with empirical hinge loss
by: Mai, The Tien
Published: (2023)
by: Mai, The Tien
Published: (2023)
Similar Items
-
Model Selection for Gaussian-gated Gaussian Mixture of Experts Using Dendrograms of Mixing Measures
by: Thai, Tuan, et al.
Published: (2025) -
Fast Model Selection and Stable Optimization for Softmax-Gated Multinomial-Logistic Mixture of Experts Models
by: Tran, TrungKhang, et al.
Published: (2026) -
A General Theory for Softmax Gating Multinomial Logistic Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023) -
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
by: Tran, TrungKhang, et al.
Published: (2026) -
A Unified Framework for Variable Selection in Model-Based Clustering with Missing Not at Random
by: Ho, Binh H., et al.
Published: (2025)