RouteLLM: Learning to Route LLMs with Preference Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ong, Isaac, Almahairi, Amjad, Wu, Vincent, Chiang, Wei-Lin, Wu, Tianhao, Gonzalez, Joseph E., Kadous, M Waleed, Stoica, Ion |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
von: Li, Tianle, et al.
Veröffentlicht: (2024)
von: Li, Tianle, et al.
Veröffentlicht: (2024)
Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
von: Chiang, Wei-Lin, et al.
Veröffentlicht: (2024)
von: Chiang, Wei-Lin, et al.
Veröffentlicht: (2024)
OR-Bench: An Over-Refusal Benchmark for Large Language Models
von: Cui, Justin, et al.
Veröffentlicht: (2024)
von: Cui, Justin, et al.
Veröffentlicht: (2024)
RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment
von: Luo, Yingfeng, et al.
Veröffentlicht: (2026)
von: Luo, Yingfeng, et al.
Veröffentlicht: (2026)
No One Fits All: From Fixed Prompting to Learned Routing in Multilingual LLMs
von: Wu, Wei-Chi, et al.
Veröffentlicht: (2026)
von: Wu, Wei-Chi, et al.
Veröffentlicht: (2026)
Arch-Router: Aligning LLM Routing with Human Preferences
von: Tran, Co, et al.
Veröffentlicht: (2025)
von: Tran, Co, et al.
Veröffentlicht: (2025)
EvoRoute: Experience-Driven Self-Routing LLM Agent Systems
von: Zhang, Guibin, et al.
Veröffentlicht: (2026)
von: Zhang, Guibin, et al.
Veröffentlicht: (2026)
TagRouter: Learning Route to LLMs through Tags for Open-Domain Text Generation Tasks
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing
von: Li, Rongjun, et al.
Veröffentlicht: (2026)
von: Li, Rongjun, et al.
Veröffentlicht: (2026)
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
von: Zheng, Lianmin, et al.
Veröffentlicht: (2023)
von: Zheng, Lianmin, et al.
Veröffentlicht: (2023)
Dr.LLM: Dynamic Layer Routing in LLMs
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
Prompt-to-Leaderboard
von: Frick, Evan, et al.
Veröffentlicht: (2025)
von: Frick, Evan, et al.
Veröffentlicht: (2025)
Route to Reason: Adaptive Routing for LLM and Reasoning Strategy Selection
von: Pan, Zhihong, et al.
Veröffentlicht: (2025)
von: Pan, Zhihong, et al.
Veröffentlicht: (2025)
Informed Routing in LLMs: Smarter Token-Level Computation for Faster Inference
von: Han, Chao, et al.
Veröffentlicht: (2025)
von: Han, Chao, et al.
Veröffentlicht: (2025)
Learning to Route LLMs with Confidence Tokens
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute
von: Ding, Dujian, et al.
Veröffentlicht: (2025)
von: Ding, Dujian, et al.
Veröffentlicht: (2025)
LLMAP: LLM-Assisted Multi-Objective Route Planning with User Preferences
von: Yuan, Liangqi, et al.
Veröffentlicht: (2025)
von: Yuan, Liangqi, et al.
Veröffentlicht: (2025)
Break the Sequential Dependency of LLM Inference Using Lookahead Decoding
von: Fu, Yichao, et al.
Veröffentlicht: (2024)
von: Fu, Yichao, et al.
Veröffentlicht: (2024)
Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback
von: Miranda, Lester James V., et al.
Veröffentlicht: (2024)
von: Miranda, Lester James V., et al.
Veröffentlicht: (2024)
Specifications: The missing link to making the development of LLM systems an engineering discipline
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
Post-Training Sparse Attention with Double Sparsity
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
How to Evaluate Reward Models for RLHF
von: Frick, Evan, et al.
Veröffentlicht: (2024)
von: Frick, Evan, et al.
Veröffentlicht: (2024)
DiSRouter: Distributed Self-Routing for LLM Selections
von: Zheng, Hang, et al.
Veröffentlicht: (2025)
von: Zheng, Hang, et al.
Veröffentlicht: (2025)
RouteProfile: Graph-Based Profiling for Cold-Start LLM Routing
von: Xu, Jingjun, et al.
Veröffentlicht: (2026)
von: Xu, Jingjun, et al.
Veröffentlicht: (2026)
A Unified Approach to Routing and Cascading for LLMs
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
Sleep-time Compute: Beyond Inference Scaling at Test-time
von: Lin, Kevin, et al.
Veröffentlicht: (2025)
von: Lin, Kevin, et al.
Veröffentlicht: (2025)
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
von: Shen, Jingyan, et al.
Veröffentlicht: (2025)
von: Shen, Jingyan, et al.
Veröffentlicht: (2025)
Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection
von: Chen, Yiwen, et al.
Veröffentlicht: (2026)
von: Chen, Yiwen, et al.
Veröffentlicht: (2026)
HAPS: Hierarchical LLM Routing with Joint Architecture and Parameter Search
von: Tian, Zihang, et al.
Veröffentlicht: (2026)
von: Tian, Zihang, et al.
Veröffentlicht: (2026)
OmniRouter: Budget and Performance Controllable Multi-LLM Routing
von: Mei, Kai, et al.
Veröffentlicht: (2025)
von: Mei, Kai, et al.
Veröffentlicht: (2025)
Causal LLM Routing: End-to-End Regret Minimization from Observational Data
von: Tsiourvas, Asterios, et al.
Veröffentlicht: (2025)
von: Tsiourvas, Asterios, et al.
Veröffentlicht: (2025)
Self-Routing RAG: Binding Selective Retrieval with Knowledge Verbalization
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
Learning to Route Languages for Multilingual Policy Optimization
von: Guo, Geyang, et al.
Veröffentlicht: (2026)
von: Guo, Geyang, et al.
Veröffentlicht: (2026)
LLMRank: Understanding LLM Strengths for Model Routing
von: Agrawal, Shubham, et al.
Veröffentlicht: (2025)
von: Agrawal, Shubham, et al.
Veröffentlicht: (2025)
Harnessing the Power of Multiple Minds: Lessons Learned from LLM Routing
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
RouteGoT: Node-Adaptive Routing for Cost-Efficient Graph of Thoughts Reasoning
von: Liu, Yuhang, et al.
Veröffentlicht: (2026)
von: Liu, Yuhang, et al.
Veröffentlicht: (2026)
RAGRouter: Learning to Route Queries to Multiple Retrieval-Augmented Language Models
von: Zhang, Jiarui, et al.
Veröffentlicht: (2025)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
von: Li, Tianle, et al.
Veröffentlicht: (2024) -
Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
von: Chiang, Wei-Lin, et al.
Veröffentlicht: (2024) -
OR-Bench: An Over-Refusal Benchmark for Large Language Models
von: Cui, Justin, et al.
Veröffentlicht: (2024) -
RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment
von: Luo, Yingfeng, et al.
Veröffentlicht: (2026) -
No One Fits All: From Fixed Prompting to Learned Routing in Multilingual LLMs
von: Wu, Wei-Chi, et al.
Veröffentlicht: (2026)