TAMER: A Test-Time Adaptive MoE-Driven Framework for EHR Representation Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Yinghao, Zheng, Xiaochen, Allam, Ahmed, Krauthammer, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Simple Contrastive Representation Learning for Time Series Forecasting
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2023)
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2023)
Two-Stage Aggregation with Dynamic Local Attention for Irregular Time Series
von: Chen, Xingyu, et al.
Veröffentlicht: (2023)
von: Chen, Xingyu, et al.
Veröffentlicht: (2023)
Clustering of Disease Trajectories with Explainable Machine Learning: A Case Study on Postoperative Delirium Phenotypes
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2024)
PRISM: Mitigating EHR Data Sparsity via Learning from Missing Feature Calibrated Prototype Patient Representations
von: Zhu, Yinghao, et al.
Veröffentlicht: (2023)
von: Zhu, Yinghao, et al.
Veröffentlicht: (2023)
A Variational Perspective on Generative Protein Fitness Optimization
von: Bogensperger, Lea, et al.
Veröffentlicht: (2025)
von: Bogensperger, Lea, et al.
Veröffentlicht: (2025)
RAST-MoE-RL: A Regime-Aware Spatio-Temporal MoE Framework for Deep Reinforcement Learning in Ride-Hailing
von: Tang, Yuhan, et al.
Veröffentlicht: (2025)
von: Tang, Yuhan, et al.
Veröffentlicht: (2025)
Repurposing Protein Language Models for Latent Flow-Based Fitness Optimization
von: Arroyo, Amaru Caceres, et al.
Veröffentlicht: (2026)
von: Arroyo, Amaru Caceres, et al.
Veröffentlicht: (2026)
Geometric Asymmetry in MoE Specialization: Functional Decorrelation and Representational Overlap
von: Liu, Feilong
Veröffentlicht: (2026)
von: Liu, Feilong
Veröffentlicht: (2026)
DynaMo: Runtime Switchable Quantization for MoE with Cross-Dataset Adaptation
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
$μ$-MoE: Test-Time Pruning as Micro-Grained Mixture-of-Experts
von: Koike-Akino, Toshiaki, et al.
Veröffentlicht: (2025)
von: Koike-Akino, Toshiaki, et al.
Veröffentlicht: (2025)
Certain Head, Uncertain Tail: Expert-Sample for Test-Time Scaling in Fine-Grained MoE
von: Chen, Yuanteng, et al.
Veröffentlicht: (2026)
von: Chen, Yuanteng, et al.
Veröffentlicht: (2026)
MoE-PHDS: One MoE checkpoint for flexible runtime sparsity
von: Hannah, Lauren. A, et al.
Veröffentlicht: (2025)
von: Hannah, Lauren. A, et al.
Veröffentlicht: (2025)
AdapMoE: Adaptive Sensitivity-based Expert Gating and Management for Efficient MoE Inference
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
T-TAMER: Provably Taming Trade-offs in ML Serving
von: Yang, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Yang, Yuanyuan, et al.
Veröffentlicht: (2025)
Learning to Specialize: Joint Gating-Expert Training for Adaptive MoEs in Decentralized Settings
von: Farhat, Yehya, et al.
Veröffentlicht: (2023)
von: Farhat, Yehya, et al.
Veröffentlicht: (2023)
GW-MoE: Resolving Uncertainty in MoE Router with Global Workspace Theory
von: Wu, Haoze, et al.
Veröffentlicht: (2024)
von: Wu, Haoze, et al.
Veröffentlicht: (2024)
Delta Decompression for MoE-based LLMs Compression
von: Gu, Hao, et al.
Veröffentlicht: (2025)
von: Gu, Hao, et al.
Veröffentlicht: (2025)
Grouter: Decoupling Routing from Representation for Accelerated MoE Training
von: Xu, Yuqi, et al.
Veröffentlicht: (2026)
von: Xu, Yuqi, et al.
Veröffentlicht: (2026)
Expert Divergence Learning for MoE-based Language Models
von: Li, Jiaang, et al.
Veröffentlicht: (2026)
von: Li, Jiaang, et al.
Veröffentlicht: (2026)
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
von: Xue, Leyang, et al.
Veröffentlicht: (2024)
von: Xue, Leyang, et al.
Veröffentlicht: (2024)
MoE-Sieve: Routing-Guided LoRA for Efficient MoE Fine-Tuning
von: Manzoni, Andrea
Veröffentlicht: (2026)
von: Manzoni, Andrea
Veröffentlicht: (2026)
DDoS: A Graph Neural Network based Drug Synergy Prediction Algorithm
von: Schwarz, Kyriakos, et al.
Veröffentlicht: (2022)
von: Schwarz, Kyriakos, et al.
Veröffentlicht: (2022)
MoETTA: Test-Time Adaptation Under Mixed Distribution Shifts with MoE-LayerNorm
von: Fan, Xiao, et al.
Veröffentlicht: (2025)
von: Fan, Xiao, et al.
Veröffentlicht: (2025)
FloE: On-the-Fly MoE Inference on Memory-constrained GPU
von: Zhou, Yuxin, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxin, et al.
Veröffentlicht: (2025)
LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference
von: Chitty-Venkata, Krishna Teja, et al.
Veröffentlicht: (2025)
von: Chitty-Venkata, Krishna Teja, et al.
Veröffentlicht: (2025)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
MoE-GPS: Guidlines for Prediction Strategy for Dynamic Expert Duplication in MoE Load Balancing
von: Ma, Haiyue, et al.
Veröffentlicht: (2025)
von: Ma, Haiyue, et al.
Veröffentlicht: (2025)
Horseshoe Mixtures-of-Experts (HS-MoE)
von: Polson, Nick, et al.
Veröffentlicht: (2026)
von: Polson, Nick, et al.
Veröffentlicht: (2026)
DyMoE: Dynamic Expert Orchestration with Mixed-Precision Quantization for Efficient MoE Inference on Edge
von: Huang, Yuegui, et al.
Veröffentlicht: (2026)
von: Huang, Yuegui, et al.
Veröffentlicht: (2026)
MoBE: Mixture-of-Basis-Experts for Compressing MoE-based LLMs
von: Chen, Xiaodong, et al.
Veröffentlicht: (2025)
von: Chen, Xiaodong, et al.
Veröffentlicht: (2025)
MoESD: Unveil Speculative Decoding's Potential for Accelerating Sparse MoE
von: Huang, Zongle, et al.
Veröffentlicht: (2025)
von: Huang, Zongle, et al.
Veröffentlicht: (2025)
Moirai-MoE: Empowering Time Series Foundation Models with Sparse Mixture of Experts
von: Liu, Xu, et al.
Veröffentlicht: (2024)
von: Liu, Xu, et al.
Veröffentlicht: (2024)
ReLibra: Routing-Replay-Guided Load Balancing for MoE Training in Reinforcement Learning
von: Jin, Chao, et al.
Veröffentlicht: (2026)
von: Jin, Chao, et al.
Veröffentlicht: (2026)
MxMoE: Mixed-precision Quantization for MoE with Accuracy and Performance Co-Design
von: Duanmu, Haojie, et al.
Veröffentlicht: (2025)
von: Duanmu, Haojie, et al.
Veröffentlicht: (2025)
Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
von: Shi, Xiaoming, et al.
Veröffentlicht: (2024)
von: Shi, Xiaoming, et al.
Veröffentlicht: (2024)
MoORE: SVD-based Model MoE-ization for Conflict- and Oblivion-Resistant Multi-Task Adaptation
von: Yuan, Shen, et al.
Veröffentlicht: (2025)
von: Yuan, Shen, et al.
Veröffentlicht: (2025)
VA-MoE: Variables-Adaptive Mixture of Experts for Incremental Weather Forecasting
von: Chen, Hao, et al.
Veröffentlicht: (2024)
von: Chen, Hao, et al.
Veröffentlicht: (2024)
MoE Pathfinder: Trajectory-driven Expert Pruning
von: Yang, Xican, et al.
Veröffentlicht: (2025)
von: Yang, Xican, et al.
Veröffentlicht: (2025)
Llama 3 Meets MoE: Efficient Upcycling
von: Vavre, Aditya, et al.
Veröffentlicht: (2024)
von: Vavre, Aditya, et al.
Veröffentlicht: (2024)
MoE Lens -- An Expert Is All You Need
von: Chaudhari, Marmik, et al.
Veröffentlicht: (2026)
von: Chaudhari, Marmik, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Simple Contrastive Representation Learning for Time Series Forecasting
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2023) -
Two-Stage Aggregation with Dynamic Local Attention for Irregular Time Series
von: Chen, Xingyu, et al.
Veröffentlicht: (2023) -
Clustering of Disease Trajectories with Explainable Machine Learning: A Case Study on Postoperative Delirium Phenotypes
von: Zheng, Xiaochen, et al.
Veröffentlicht: (2024) -
PRISM: Mitigating EHR Data Sparsity via Learning from Missing Feature Calibrated Prototype Patient Representations
von: Zhu, Yinghao, et al.
Veröffentlicht: (2023) -
A Variational Perspective on Generative Protein Fitness Optimization
von: Bogensperger, Lea, et al.
Veröffentlicht: (2025)