Learning How Much to Think: Difficulty-Aware Dynamic MoEs for Graph Node Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Jiajun, Li, Yadong, Chen, Xuanze, Ma, Chen, Zhao, Chuang, Yu, Shanqing, Xuan, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mixture of Message Passing Experts with Routing Entropy Regularization for Node Classification
by: Chen, Xuanze, et al.
Published: (2025)
by: Chen, Xuanze, et al.
Published: (2025)
Mixture of Experts Meets Decoupled Message Passing: Towards General and Adaptive Node Classification
by: Chen, Xuanze, et al.
Published: (2024)
by: Chen, Xuanze, et al.
Published: (2024)
Rethinking Graph Transformer Architecture Design for Node Classification
by: Zhou, Jiajun, et al.
Published: (2024)
by: Zhou, Jiajun, et al.
Published: (2024)
CrossHGL: A Text-Free Foundation Model for Cross-Domain Heterogeneous Graph Learning
by: Chen, Xuanze, et al.
Published: (2026)
by: Chen, Xuanze, et al.
Published: (2026)
Clarify Confused Nodes via Separated Learning
by: Zhou, Jiajun, et al.
Published: (2023)
by: Zhou, Jiajun, et al.
Published: (2023)
Lateral Movement Detection via Time-aware Subgraph Classification on Authentication Logs
by: Zhou, Jiajun, et al.
Published: (2024)
by: Zhou, Jiajun, et al.
Published: (2024)
Traffic-MoE: A Sparse Foundation Model for Network Traffic Analysis
by: Zhou, Jiajun, et al.
Published: (2026)
by: Zhou, Jiajun, et al.
Published: (2026)
Facilitating Feature and Topology Lightweighting: An Ethereum Transaction Graph Compression Method for Malicious Account Detection
by: Zhou, Jiajun, et al.
Published: (2024)
by: Zhou, Jiajun, et al.
Published: (2024)
Network Anomaly Traffic Detection via Multi-view Feature Fusion
by: Hao, Song, et al.
Published: (2024)
by: Hao, Song, et al.
Published: (2024)
A Federated Parameter Aggregation Method for Node Classification Tasks with Different Graph Network Structures
by: Song, Hao, et al.
Published: (2024)
by: Song, Hao, et al.
Published: (2024)
Hypergraph-Based Dynamic Graph Node Classification
by: Ma, Xiaoxu, et al.
Published: (2024)
by: Ma, Xiaoxu, et al.
Published: (2024)
Correlation-Aware Graph Convolutional Networks for Multi-Label Node Classification
by: Bei, Yuanchen, et al.
Published: (2024)
by: Bei, Yuanchen, et al.
Published: (2024)
Dual-view Aware Smart Contract Vulnerability Detection for Ethereum
by: Yao, Jiacheng, et al.
Published: (2024)
by: Yao, Jiacheng, et al.
Published: (2024)
LoRALib: A Standardized Benchmark for Evaluating LoRA-MoE Methods
by: Wang, Shaoheng, et al.
Published: (2025)
by: Wang, Shaoheng, et al.
Published: (2025)
RAST-MoE-RL: A Regime-Aware Spatio-Temporal MoE Framework for Deep Reinforcement Learning in Ride-Hailing
by: Tang, Yuhan, et al.
Published: (2025)
by: Tang, Yuhan, et al.
Published: (2025)
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
by: Zhao, Xinyuan, et al.
Published: (2026)
by: Zhao, Xinyuan, et al.
Published: (2026)
MoE-GPS: Guidlines for Prediction Strategy for Dynamic Expert Duplication in MoE Load Balancing
by: Ma, Haiyue, et al.
Published: (2025)
by: Ma, Haiyue, et al.
Published: (2025)
Continual Pre-training of MoEs: How robust is your router?
by: Thérien, Benjamin, et al.
Published: (2025)
by: Thérien, Benjamin, et al.
Published: (2025)
Expert Selections In MoE Models Reveal (Almost) As Much As Text
by: Nuriyev, Amir, et al.
Published: (2026)
by: Nuriyev, Amir, et al.
Published: (2026)
MoEs Are Stronger than You Think: Hyper-Parallel Inference Scaling with RoE
by: Zibakhsh, Soheil, et al.
Published: (2025)
by: Zibakhsh, Soheil, et al.
Published: (2025)
Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts
by: Wu, Haoyuan, et al.
Published: (2025)
by: Wu, Haoyuan, et al.
Published: (2025)
MoE-Compression: How the Compression Error of Experts Affects the Inference Accuracy of MoE Model?
by: Ma, Songkai, et al.
Published: (2025)
by: Ma, Songkai, et al.
Published: (2025)
Staleness-Centric Optimizations for Parallel Diffusion MoE Inference
by: Luo, Jiajun, et al.
Published: (2024)
by: Luo, Jiajun, et al.
Published: (2024)
Adaptive Substructure-Aware Expert Model for Molecular Property Prediction
by: Jiang, Tianyi, et al.
Published: (2025)
by: Jiang, Tianyi, et al.
Published: (2025)
Pangu Ultra MoE: How to Train Your Big MoE on Ascend NPUs
by: Tang, Yehui, et al.
Published: (2025)
by: Tang, Yehui, et al.
Published: (2025)
Learning to Specialize: Joint Gating-Expert Training for Adaptive MoEs in Decentralized Settings
by: Farhat, Yehya, et al.
Published: (2023)
by: Farhat, Yehya, et al.
Published: (2023)
Sparse Crosscoders for diffing MoEs and Dense models
by: Chaudhari, Marmik, et al.
Published: (2026)
by: Chaudhari, Marmik, et al.
Published: (2026)
Hierarchical Local-Global Feature Learning for Few-shot Malicious Traffic Detection
by: Peng, Songtao, et al.
Published: (2025)
by: Peng, Songtao, et al.
Published: (2025)
Enhancing Ethereum Fraud Detection via Generative and Contrastive Self-supervision
by: Jin, Chenxiang, et al.
Published: (2024)
by: Jin, Chenxiang, et al.
Published: (2024)
Multi-view Correlation-aware Network Traffic Detection on Flow Hypergraph
by: Zhou, Jiajun, et al.
Published: (2025)
by: Zhou, Jiajun, et al.
Published: (2025)
Dirichlet-Prior Shaping: Guiding Expert Specialization in Upcycled MoEs
by: Mirvakhabova, Leyla, et al.
Published: (2025)
by: Mirvakhabova, Leyla, et al.
Published: (2025)
Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
MergeME: Model Merging Techniques for Homogeneous and Heterogeneous MoEs
by: Zhou, Yuhang, et al.
Published: (2025)
by: Zhou, Yuhang, et al.
Published: (2025)
UniMoE-Audio: Unified Speech and Music Generation with Dynamic-Capacity MoE
by: Liu, Zhenyu, et al.
Published: (2025)
by: Liu, Zhenyu, et al.
Published: (2025)
GRACE-MoE: Grouping and Replication with Locality-Aware Routing for Efficient Distributed MoE Inference
by: Han, Yu, et al.
Published: (2025)
by: Han, Yu, et al.
Published: (2025)
Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs
by: Ye, Charles, et al.
Published: (2026)
by: Ye, Charles, et al.
Published: (2026)
Does a Global Perspective Help Prune Sparse MoEs Elegantly?
by: Zhang, Zeliang, et al.
Published: (2026)
by: Zhang, Zeliang, et al.
Published: (2026)
Fast Graph Sharpness-Aware Minimization for Enhancing and Accelerating Few-Shot Node Classification
by: Luo, Yihong, et al.
Published: (2024)
by: Luo, Yihong, et al.
Published: (2024)
Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation
by: Chen, Minping, et al.
Published: (2026)
by: Chen, Minping, et al.
Published: (2026)
Think How to Think: Mitigating Overthinking with Autonomous Difficulty Cognition in Large Reasoning Models
by: Liu, Yongjiang, et al.
Published: (2025)
by: Liu, Yongjiang, et al.
Published: (2025)
Similar Items
-
Mixture of Message Passing Experts with Routing Entropy Regularization for Node Classification
by: Chen, Xuanze, et al.
Published: (2025) -
Mixture of Experts Meets Decoupled Message Passing: Towards General and Adaptive Node Classification
by: Chen, Xuanze, et al.
Published: (2024) -
Rethinking Graph Transformer Architecture Design for Node Classification
by: Zhou, Jiajun, et al.
Published: (2024) -
CrossHGL: A Text-Free Foundation Model for Cross-Domain Heterogeneous Graph Learning
by: Chen, Xuanze, et al.
Published: (2026) -
Clarify Confused Nodes via Separated Learning
by: Zhou, Jiajun, et al.
Published: (2023)