Robust Heterogeneous Analog-Digital Computing for Mixture-of-Experts Models with Theoretical Generalization Guarantees
Fuente:
arXiv
Saved in:
| Main Authors: | Chowdhury, Mohammed Nowaz Rabbani, Tsai, Hsinyu, Burr, Geoffrey W., Maghraoui, Kaoutar El, Liu, Liu, Wang, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
A Provably Effective Method for Pruning Experts in Fine-tuned Sparse Mixture-of-Experts
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2024)
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2024)
On the Convergence Theory of Pipeline Gradient-based Analog In-memory Training
by: Wu, Zhaoxian, et al.
Published: (2024)
by: Wu, Zhaoxian, et al.
Published: (2024)
Analog In-Memory Computing with Uncertainty Quantification for Efficient Edge-based Medical Imaging Segmentation
by: Hamzaoui, Imane, et al.
Published: (2024)
by: Hamzaoui, Imane, et al.
Published: (2024)
Evaluating Fine-Tuned LLM Model For Medical Transcription With Small Low-Resource Languages Validated Dataset
by: Chowdhury, Mohammed Nowshad Ruhani, et al.
Published: (2026)
by: Chowdhury, Mohammed Nowshad Ruhani, et al.
Published: (2026)
Context-Aware Mixture-of-Experts Inference on CXL-Enabled GPU-NDP Systems
by: Fan, Zehao, et al.
Published: (2025)
by: Fan, Zehao, et al.
Published: (2025)
Using the IBM Analog In-Memory Hardware Acceleration Kit for Neural Network Training and Inference
by: Gallo, Manuel Le, et al.
Published: (2023)
by: Gallo, Manuel Le, et al.
Published: (2023)
Analog Foundation Models
by: Büchel, Julian, et al.
Published: (2025)
by: Büchel, Julian, et al.
Published: (2025)
SparseST: Exploiting Data Sparsity in Spatiotemporal Modeling and Prediction
by: Wu, Junfeng, et al.
Published: (2025)
by: Wu, Junfeng, et al.
Published: (2025)
Accelerating LLM Inference via Dynamic KV Cache Placement in Heterogeneous Memory System
by: Fang, Yunhua, et al.
Published: (2025)
by: Fang, Yunhua, et al.
Published: (2025)
Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning
by: Shemla, Yuval, et al.
Published: (2026)
by: Shemla, Yuval, et al.
Published: (2026)
EvidenceMoE: A Physics-Guided Mixture-of-Experts with Evidential Critics for Advancing Fluorescence Light Detection and Ranging in Scattering Media
by: Erbas, Ismail, et al.
Published: (2025)
by: Erbas, Ismail, et al.
Published: (2025)
CiMBA: Accelerating Genome Sequencing through On-Device Basecalling via Compute-in-Memory
by: Simon, William Andrew, et al.
Published: (2025)
by: Simon, William Andrew, et al.
Published: (2025)
Mixture of Heterogeneous Grouped Experts for Language Modeling
by: Ma, Zhicheng, et al.
Published: (2026)
by: Ma, Zhicheng, et al.
Published: (2026)
Paged Attention Meets FlexAttention: Unlocking Long-Context Efficiency in Deployed Inference
by: Joshi, Thomas, et al.
Published: (2025)
by: Joshi, Thomas, et al.
Published: (2025)
PHMForge: Evaluating LLM Agents on Industrial Prognostics through MCP-Native, Algorithm-Grounded Tools
by: Feng, Tianjun, et al.
Published: (2026)
by: Feng, Tianjun, et al.
Published: (2026)
The Ultimate Legal Deception Model – A Unified Theoretical Framework
by: Rabbani, Hassan
Published: (2025)
by: Rabbani, Hassan
Published: (2025)
Evaluating Temporal Semantic Caching and Workflow Optimization in Agentic Plan-Execute Pipelines
by: Merchant, Alimurtaza Mustafa, et al.
Published: (2026)
by: Merchant, Alimurtaza Mustafa, et al.
Published: (2026)
Robust Mixture Models for Algorithmic Fairness Under Latent Heterogeneity
by: Li, Siqi, et al.
Published: (2025)
by: Li, Siqi, et al.
Published: (2025)
Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts
by: Lyu, Boxuan, et al.
Published: (2026)
by: Lyu, Boxuan, et al.
Published: (2026)
MixtureKit: A General Framework for Composing, Training, and Visualizing Mixture-of-Experts Models
by: Chamma, Ahmad, et al.
Published: (2025)
by: Chamma, Ahmad, et al.
Published: (2025)
Rethinking Governance with Blockchain: An Actor-Network Theoretical Approach
by: HICOR, Zineb, et al.
Published: (2025)
by: HICOR, Zineb, et al.
Published: (2025)
Optimizing Distributed Deployment of Mixture-of-Experts Model Inference in Serverless Computing
by: Liu, Mengfan, et al.
Published: (2025)
by: Liu, Mengfan, et al.
Published: (2025)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
by: Wang, An, et al.
Published: (2024)
by: Wang, An, et al.
Published: (2024)
Sparse Mixture-of-Experts for Compositional Generalization: Empirical Evidence and Theoretical Foundations of Optimal Sparsity
by: Zhao, Jinze, et al.
Published: (2024)
by: Zhao, Jinze, et al.
Published: (2024)
HeterMoE: Efficient Training of Mixture-of-Experts Models on Heterogeneous GPUs
by: Wu, Yongji, et al.
Published: (2025)
by: Wu, Yongji, et al.
Published: (2025)
Mosaic: Data-Free Knowledge Distillation via Mixture-of-Experts for Heterogeneous Distributed Environments
by: Liu, Junming, et al.
Published: (2025)
by: Liu, Junming, et al.
Published: (2025)
ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems
by: Zhou, Wenyong, et al.
Published: (2026)
by: Zhou, Wenyong, et al.
Published: (2026)
Transcendental Regularization of Finite Mixtures:Theoretical Guarantees and Practical Limitations
by: Fokoué, Ernest
Published: (2026)
by: Fokoué, Ernest
Published: (2026)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
by: Huix, Tom, et al.
Published: (2024)
by: Huix, Tom, et al.
Published: (2024)
MC#: Mixture Compressor for Mixture-of-Experts Large Models
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Mixed Effects Mixture of Experts: Modeling Double Heterogeneous Trajectories
by: Yue, Xinkai, et al.
Published: (2026)
by: Yue, Xinkai, et al.
Published: (2026)
FNH-TTS: Mixture-of-Experts Duration Modeling for Robust Neural Speech Synthesis
by: Meng, Qingliang, et al.
Published: (2025)
by: Meng, Qingliang, et al.
Published: (2025)
Efficiently Editing Mixture-of-Experts Models with Compressed Experts
by: He, Yifei, et al.
Published: (2025)
by: He, Yifei, et al.
Published: (2025)
Rapid yet accurate Tile-circuit and device modeling for Analog In-Memory Computing
by: Luquin, J., et al.
Published: (2025)
by: Luquin, J., et al.
Published: (2025)
Compute SNR-Optimal Analog-to-Digital Converters for Analog In-Memory Computing
by: Kavishwar, Mihir, et al.
Published: (2025)
by: Kavishwar, Mihir, et al.
Published: (2025)
Towards Robust Learning to Optimize with Theoretical Guarantees
by: Song, Qingyu, et al.
Published: (2025)
by: Song, Qingyu, et al.
Published: (2025)
Robust Audiovisual Speech Recognition Models with Mixture-of-Experts
by: Wu, Yihan, et al.
Published: (2024)
by: Wu, Yihan, et al.
Published: (2024)
Mixture of Experts with Mixture of Precisions for Tuning Quality of Service
by: Imani, HamidReza, et al.
Published: (2024)
by: Imani, HamidReza, et al.
Published: (2024)
MIM-Reasoner: Learning with Theoretical Guarantees for Multiplex Influence Maximization
by: Do, Nguyen, et al.
Published: (2024)
by: Do, Nguyen, et al.
Published: (2024)
Similar Items
-
Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026) -
A Provably Effective Method for Pruning Experts in Fine-tuned Sparse Mixture-of-Experts
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2024) -
On the Convergence Theory of Pipeline Gradient-based Analog In-memory Training
by: Wu, Zhaoxian, et al.
Published: (2024) -
Analog In-Memory Computing with Uncertainty Quantification for Efficient Edge-based Medical Imaging Segmentation
by: Hamzaoui, Imane, et al.
Published: (2024) -
Evaluating Fine-Tuned LLM Model For Medical Transcription With Small Low-Resource Languages Validated Dataset
by: Chowdhury, Mohammed Nowshad Ruhani, et al.
Published: (2026)