HierRouter: Coordinated Routing of Specialized Large Language Models via Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Nikunj, Guo, Bill, Kannan, Rajgopal, Prasanna, Viktor K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do LLM-derived graph priors improve multi-agent coordination?
by: Gupta, Nikunj, et al.
Published: (2026)
by: Gupta, Nikunj, et al.
Published: (2026)
SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2026)
by: Gupta, Nikunj, et al.
Published: (2026)
Deep Meta Coordination Graphs for Multi-agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2025)
by: Gupta, Nikunj, et al.
Published: (2025)
Action-Graph Policies: Learning Action Co-dependencies in Multi-Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2026)
by: Gupta, Nikunj, et al.
Published: (2026)
TIGER-MARL: Enhancing Multi-Agent Reinforcement Learning with Temporal Information through Graph-based Embeddings and Representations
by: Gupta, Nikunj, et al.
Published: (2025)
by: Gupta, Nikunj, et al.
Published: (2025)
LYNX: Learning Dynamic Exits for Confidence-Controlled Reasoning
by: Akgül, Ömer Faruk, et al.
Published: (2025)
by: Akgül, Ömer Faruk, et al.
Published: (2025)
Adversarial Training in Low-Label Regimes with Margin-Based Interpolation
by: Ye, Tian, et al.
Published: (2024)
by: Ye, Tian, et al.
Published: (2024)
Attention, Distillation, and Tabularization: Towards Practical Neural Network-Based Prefetching
by: Zhang, Pengmiao, et al.
Published: (2023)
by: Zhang, Pengmiao, et al.
Published: (2023)
FACTUAL: A Novel Framework for Contrastive Learning Based Robust SAR Image Classification
by: Wang, Xu, et al.
Published: (2024)
by: Wang, Xu, et al.
Published: (2024)
Studying the Effects of Self-Attention on SAR Automatic Target Recognition
by: Fein-Ashley, Jacob, et al.
Published: (2024)
by: Fein-Ashley, Jacob, et al.
Published: (2024)
PaCKD: Pattern-Clustered Knowledge Distillation for Compressing Memory Access Prediction Models
by: Gupta, Neelesh, et al.
Published: (2024)
by: Gupta, Neelesh, et al.
Published: (2024)
A Persistent-State Dataflow Accelerator for Memory-Bound Linear Attention Decode on FPGA
by: Gupta, Neelesh, et al.
Published: (2026)
by: Gupta, Neelesh, et al.
Published: (2026)
Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
by: Zhang, Haozhen, et al.
Published: (2025)
by: Zhang, Haozhen, et al.
Published: (2025)
SPECTRE: An FFT-Based Efficient Drop-In Replacement to Self-Attention for Long Contexts
by: Fein-Ashley, Jacob, et al.
Published: (2025)
by: Fein-Ashley, Jacob, et al.
Published: (2025)
Conformal Prediction for Federated Graph Neural Networks with Missing Neighbor Information
by: Akgül, Ömer Faruk, et al.
Published: (2024)
by: Akgül, Ömer Faruk, et al.
Published: (2024)
Contextual Feedback Loops: Amplifying Deep Reasoning with Iterative Top-Down Feedback
by: Fein-Ashley, Jacob, et al.
Published: (2024)
by: Fein-Ashley, Jacob, et al.
Published: (2024)
RouterDC: Query-Based Router by Dual Contrastive Learning for Assembling Large Language Models
by: Chen, Shuhao, et al.
Published: (2024)
by: Chen, Shuhao, et al.
Published: (2024)
LLM Router: Rethinking Routing with Prefill Activations
by: Varshney, Tanay, et al.
Published: (2026)
by: Varshney, Tanay, et al.
Published: (2026)
VL-RouterBench: A Benchmark for Vision-Language Model Routing
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
Rethinking RL for LLM Reasoning: It's Sparse Policy Selection, Not Capability Learning
by: Akgül, Ömer Faruk, et al.
Published: (2026)
by: Akgül, Ömer Faruk, et al.
Published: (2026)
TypeBandit: Type-Level Context Allocation and Reweighting for Effective Attribute Completion in Heterogeneous Graph Neural Networks
by: Wang, Ta-Yang, et al.
Published: (2026)
by: Wang, Ta-Yang, et al.
Published: (2026)
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
by: Zhao, Siyan, et al.
Published: (2025)
by: Zhao, Siyan, et al.
Published: (2025)
Mixture of Thoughts: Learning to Aggregate What Experts Think, Not Just What They Say
by: Fein-Ashley, Jacob, et al.
Published: (2025)
by: Fein-Ashley, Jacob, et al.
Published: (2025)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
by: Heuillet, Maxime, et al.
Published: (2025)
by: Heuillet, Maxime, et al.
Published: (2025)
DFA-RAG: Conversational Semantic Router for Large Language Model with Definite Finite Automaton
by: Sun, Yiyou, et al.
Published: (2024)
by: Sun, Yiyou, et al.
Published: (2024)
Benchmarking Deep Learning Classifiers for SAR Automatic Target Recognition
by: Fein-Ashley, Jacob, et al.
Published: (2023)
by: Fein-Ashley, Jacob, et al.
Published: (2023)
A Single Graph Convolution Is All You Need: Efficient Grayscale Image Classification
by: Fein-Ashley, Jacob, et al.
Published: (2024)
by: Fein-Ashley, Jacob, et al.
Published: (2024)
Mixture of Scope Experts at Test: Generalizing Deeper Graph Neural Networks with Shallow Variants
by: Deng, Gangda, et al.
Published: (2024)
by: Deng, Gangda, et al.
Published: (2024)
Towards Ideal Temporal Graph Neural Networks: Evaluations and Conclusions after 10,000 GPU Hours
by: Yang, Yuxin, et al.
Published: (2024)
by: Yang, Yuxin, et al.
Published: (2024)
PAHD: Perception-Action based Human Decision Making using Explainable Graph Neural Networks on SAR Images
by: Wijeratne, Sasindu, et al.
Published: (2024)
by: Wijeratne, Sasindu, et al.
Published: (2024)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
by: Xu, Xin, et al.
Published: (2026)
by: Xu, Xin, et al.
Published: (2026)
xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning
by: Qian, Cheng, et al.
Published: (2025)
by: Qian, Cheng, et al.
Published: (2025)
Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers
by: Ma, Wenhan, et al.
Published: (2025)
by: Ma, Wenhan, et al.
Published: (2025)
Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization
by: Tang, Haochun, et al.
Published: (2026)
by: Tang, Haochun, et al.
Published: (2026)
Routoo: Learning to Route to Large Language Models Effectively
by: Mohammadshahi, Alireza, et al.
Published: (2024)
by: Mohammadshahi, Alireza, et al.
Published: (2024)
Large Language Model Confidence Estimation via Black-Box Access
by: Pedapati, Tejaswini, et al.
Published: (2024)
by: Pedapati, Tejaswini, et al.
Published: (2024)
Thermometer: Towards Universal Calibration for Large Language Models
by: Shen, Maohao, et al.
Published: (2024)
by: Shen, Maohao, et al.
Published: (2024)
EmbedLLM: Learning Compact Representations of Large Language Models
by: Zhuang, Richard, et al.
Published: (2024)
by: Zhuang, Richard, et al.
Published: (2024)
RouteNator: A Router-Based Multi-Modal Architecture for Generating Synthetic Training Data for Function Calling LLMs
by: Belavadi, Vibha, et al.
Published: (2025)
by: Belavadi, Vibha, et al.
Published: (2025)
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
by: Lv, Ang, et al.
Published: (2025)
by: Lv, Ang, et al.
Published: (2025)
Similar Items
-
Do LLM-derived graph priors improve multi-agent coordination?
by: Gupta, Nikunj, et al.
Published: (2026) -
SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2026) -
Deep Meta Coordination Graphs for Multi-agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2025) -
Action-Graph Policies: Learning Action Co-dependencies in Multi-Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2026) -
TIGER-MARL: Enhancing Multi-Agent Reinforcement Learning with Temporal Information through Graph-based Embeddings and Representations
by: Gupta, Nikunj, et al.
Published: (2025)