Mixture of Routers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jia-Chen, Xiong, Yu-Jie, Qiu, Xi-He, Xia, Chun-Ming, Dai, Fei, Zhou, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
Yuan 2.0-M32: Mixture of Experts with Attention Router
von: Wu, Shaohua, et al.
Veröffentlicht: (2024)
von: Wu, Shaohua, et al.
Veröffentlicht: (2024)
Wavelet Mixture of Experts for Time Series Forecasting
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
RouterDC: Query-Based Router by Dual Contrastive Learning for Assembling Large Language Models
von: Chen, Shuhao, et al.
Veröffentlicht: (2024)
von: Chen, Shuhao, et al.
Veröffentlicht: (2024)
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
CausalDiffTab: Mixed-Type Causal-Aware Diffusion for Tabular Data Generation
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
Multi-Sample Anti-Aliasing and Constrained Optimization for 3D Gaussian Splatting
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
von: Xie, Yanyue, et al.
Veröffentlicht: (2024)
von: Xie, Yanyue, et al.
Veröffentlicht: (2024)
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
von: Bao, Zhenghua, et al.
Veröffentlicht: (2026)
von: Bao, Zhenghua, et al.
Veröffentlicht: (2026)
Towards Fair and Comprehensive Evaluation of Routers in Collaborative LLM Systems
von: Wu, Wanxing, et al.
Veröffentlicht: (2026)
von: Wu, Wanxing, et al.
Veröffentlicht: (2026)
Gradient-Direction-Aware Density Control for 3D Gaussian Splatting
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
MemRouter: Memory-as-Embedding Routing for Long-Term Conversational Agents
von: Hu, Tianyu, et al.
Veröffentlicht: (2026)
von: Hu, Tianyu, et al.
Veröffentlicht: (2026)
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
Layerwise Recurrent Router for Mixture-of-Experts
von: Qiu, Zihan, et al.
Veröffentlicht: (2024)
von: Qiu, Zihan, et al.
Veröffentlicht: (2024)
RadialRouter: Structured Representation for Efficient and Robust Large Language Models Routing
von: Jin, Ruihan, et al.
Veröffentlicht: (2025)
von: Jin, Ruihan, et al.
Veröffentlicht: (2025)
More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection
von: Sun, Runze, et al.
Veröffentlicht: (2026)
von: Sun, Runze, et al.
Veröffentlicht: (2026)
Mixture of Reasonings: Teach Large Language Models to Reason with Adaptive Strategies
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
LoRA$^2$ : Multi-Scale Low-Rank Approximations for Fine-Tuning Large Language Models
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2024)
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2024)
An Entailment Tree Generation Approach for Multimodal Multi-Hop Question Answering with Mixture-of-Experts and Iterative Feedback Mechanism
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
Performance Characterization of Expert Router for Scalable LLM Inference
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
Route-and-Reason: Scaling Large Language Model Reasoning with Reinforced Model Router
von: Shao, Chenyang, et al.
Veröffentlicht: (2025)
von: Shao, Chenyang, et al.
Veröffentlicht: (2025)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
MoSEs: Uncertainty-Aware AI-Generated Text Detection via Mixture of Stylistics Experts with Conditional Thresholds
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
Toward Super Agent System with Hybrid AI Routers
von: Yao, Yuhang, et al.
Veröffentlicht: (2025)
von: Yao, Yuhang, et al.
Veröffentlicht: (2025)
Meta-RTL: Reinforcement-Based Meta-Transfer Learning for Low-Resource Commonsense Reasoning
von: Fu, Yu, et al.
Veröffentlicht: (2024)
von: Fu, Yu, et al.
Veröffentlicht: (2024)
Understand Then Memory: A Cognitive Gist-Driven RAG Framework with Global Semantic Diffusion
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2026)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2026)
When to Reason: Semantic Router for vLLM
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
MoA: Heterogeneous Mixture of Adapters for Parameter-Efficient Fine-Tuning of Large Language Models
von: Cao, Jie, et al.
Veröffentlicht: (2025)
von: Cao, Jie, et al.
Veröffentlicht: (2025)
Mixture-of-Depths Attention
von: Zhu, Lianghui, et al.
Veröffentlicht: (2026)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2026)
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
von: Chen, Kedi, et al.
Veröffentlicht: (2024)
von: Chen, Kedi, et al.
Veröffentlicht: (2024)
Mixture-of-Minds: Multi-Agent Reinforcement Learning for Table Understanding
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
GRAPHMOE: Amplifying Cognitive Depth of Mixture-of-Experts Network via Introducing Self-Rethinking Mechanism
von: Lv, Bo, et al.
Veröffentlicht: (2025)
von: Lv, Bo, et al.
Veröffentlicht: (2025)
Olympus: A Universal Task Router for Computer Vision Tasks
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers
von: Ma, Wenhan, et al.
Veröffentlicht: (2025)
von: Ma, Wenhan, et al.
Veröffentlicht: (2025)
Mobile-Agent-V: A Video-Guided Approach for Effortless and Efficient Operational Knowledge Injection in Mobile Automation
von: Wang, Junyang, et al.
Veröffentlicht: (2025)
von: Wang, Junyang, et al.
Veröffentlicht: (2025)
VL-RouterBench: A Benchmark for Vision-Language Model Routing
von: Huang, Zhehao, et al.
Veröffentlicht: (2025)
von: Huang, Zhehao, et al.
Veröffentlicht: (2025)
Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource
von: Li, Houyi, et al.
Veröffentlicht: (2025)
von: Li, Houyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025) -
Yuan 2.0-M32: Mixture of Experts with Attention Router
von: Wu, Shaohua, et al.
Veröffentlicht: (2024) -
Wavelet Mixture of Experts for Time Series Forecasting
von: Zhou, Zheng, et al.
Veröffentlicht: (2025) -
RouterDC: Query-Based Router by Dual Contrastive Learning for Assembling Large Language Models
von: Chen, Shuhao, et al.
Veröffentlicht: (2024) -
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)