Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics
Fuente:
arXiv
Guardado en:
| Autores principales: | Piskala, Deepak Babu, Raajaa, Vijay, Mishra, Sachin, Bozza, Bruno |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
por: Piskala, Deepak Babu
Publicado: (2026)
por: Piskala, Deepak Babu
Publicado: (2026)
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
por: Piskala, Deepak Babu
Publicado: (2026)
por: Piskala, Deepak Babu
Publicado: (2026)
LLMAP: LLM-Assisted Multi-Objective Route Planning with User Preferences
por: Yuan, Liangqi, et al.
Publicado: (2025)
por: Yuan, Liangqi, et al.
Publicado: (2025)
Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR
por: Ingle, Yash, et al.
Publicado: (2026)
por: Ingle, Yash, et al.
Publicado: (2026)
RouteLLM: Learning to Route LLMs with Preference Data
por: Ong, Isaac, et al.
Publicado: (2024)
por: Ong, Isaac, et al.
Publicado: (2024)
Hardware Aware Ensemble Selection for Balancing Predictive Accuracy and Cost
por: Maier, Jannis, et al.
Publicado: (2024)
por: Maier, Jannis, et al.
Publicado: (2024)
Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
por: Bhandarkar, Abhay, et al.
Publicado: (2025)
por: Bhandarkar, Abhay, et al.
Publicado: (2025)
Probabilistic Trust Intervals for Out of Distribution Detection
por: Singh, Gagandeep, et al.
Publicado: (2021)
por: Singh, Gagandeep, et al.
Publicado: (2021)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
por: Xu, Zaiyan, et al.
Publicado: (2025)
por: Xu, Zaiyan, et al.
Publicado: (2025)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
por: Ding, Dujian, et al.
Publicado: (2024)
por: Ding, Dujian, et al.
Publicado: (2024)
Trust by Design: Skill Profiles for Transparent, Cost-Aware LLM Routing
por: Okamoto, Mika, et al.
Publicado: (2026)
por: Okamoto, Mika, et al.
Publicado: (2026)
Breaking Model Lock-in: Cost-Efficient Zero-Shot LLM Routing via a Universal Latent Space
por: Yan, Cheng, et al.
Publicado: (2026)
por: Yan, Cheng, et al.
Publicado: (2026)
Optimal Signal Decomposition-based Multi-Stage Learning for Battery Health Estimation
por: Pamshetti, Vijay Babu, et al.
Publicado: (2025)
por: Pamshetti, Vijay Babu, et al.
Publicado: (2025)
Dr.LLM: Dynamic Layer Routing in LLMs
por: Heakl, Ahmed, et al.
Publicado: (2025)
por: Heakl, Ahmed, et al.
Publicado: (2025)
Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge
por: Zhang, Wenbo, et al.
Publicado: (2026)
por: Zhang, Wenbo, et al.
Publicado: (2026)
Preference Guided Iterated Pareto Referent Optimisation for Accessible Route Planning
por: Speziali, Paolo, et al.
Publicado: (2026)
por: Speziali, Paolo, et al.
Publicado: (2026)
LENSLLM: Unveiling Fine-Tuning Dynamics for LLM Selection
por: Zeng, Xinyue, et al.
Publicado: (2025)
por: Zeng, Xinyue, et al.
Publicado: (2025)
VAGPO: Vision-augmented Asymmetric Group Preference Optimization for Graph Routing Problems
por: Liu, Shiyan, et al.
Publicado: (2025)
por: Liu, Shiyan, et al.
Publicado: (2025)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
por: Fu, Qichen, et al.
Publicado: (2024)
por: Fu, Qichen, et al.
Publicado: (2024)
Less is More: Improving LLM Alignment via Preference Data Selection
por: Deng, Xun, et al.
Publicado: (2025)
por: Deng, Xun, et al.
Publicado: (2025)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
por: Metcalf, Katherine, et al.
Publicado: (2024)
por: Metcalf, Katherine, et al.
Publicado: (2024)
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs
por: Nguyen, Quang H., et al.
Publicado: (2024)
por: Nguyen, Quang H., et al.
Publicado: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
por: Niu, Tianyi, et al.
Publicado: (2026)
por: Niu, Tianyi, et al.
Publicado: (2026)
A Bi-Objective Approach to Last-Mile Delivery Routing Considering Driver Preferences
por: Mesa, Juan Pablo, et al.
Publicado: (2024)
por: Mesa, Juan Pablo, et al.
Publicado: (2024)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
por: Sun, Shengjie, et al.
Publicado: (2025)
por: Sun, Shengjie, et al.
Publicado: (2025)
GraNNite: Enabling High-Performance Execution of Graph Neural Networks on Resource-Constrained Neural Processing Units
por: Das, Arghadip, et al.
Publicado: (2025)
por: Das, Arghadip, et al.
Publicado: (2025)
SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning
por: Xu, Zhongling, et al.
Publicado: (2026)
por: Xu, Zhongling, et al.
Publicado: (2026)
Adversarial Preference Learning for Robust LLM Alignment
por: Wang, Yuanfu, et al.
Publicado: (2025)
por: Wang, Yuanfu, et al.
Publicado: (2025)
Machine Learning Framework for Early Power, Performance, and Area Estimation of RTL
por: Chattopadhyay, Anindita, et al.
Publicado: (2025)
por: Chattopadhyay, Anindita, et al.
Publicado: (2025)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
por: Zhang, Yi-Kai, et al.
Publicado: (2025)
por: Zhang, Yi-Kai, et al.
Publicado: (2025)
LLM Data Selection and Utilization via Dynamic Bi-level Optimization
por: Yu, Yang, et al.
Publicado: (2025)
por: Yu, Yang, et al.
Publicado: (2025)
PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment
por: Verma, Richa, et al.
Publicado: (2026)
por: Verma, Richa, et al.
Publicado: (2026)
A Cost-Effective LLM-based Approach to Identify Wildlife Trafficking in Online Marketplaces
por: Barbosa, Juliana, et al.
Publicado: (2025)
por: Barbosa, Juliana, et al.
Publicado: (2025)
MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness
por: Hathidara, Ashutosh, et al.
Publicado: (2026)
por: Hathidara, Ashutosh, et al.
Publicado: (2026)
Aligning LLM Agents by Learning Latent Preference from User Edits
por: Gao, Ge, et al.
Publicado: (2024)
por: Gao, Ge, et al.
Publicado: (2024)
Automatic Demonstration Selection for LLM-based Tabular Data Classification
por: Han, Shuchu, et al.
Publicado: (2025)
por: Han, Shuchu, et al.
Publicado: (2025)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
por: Gao, Chen-Xiao, et al.
Publicado: (2024)
por: Gao, Chen-Xiao, et al.
Publicado: (2024)
On the Performance of Imputation Techniques for Missing Values on Healthcare Datasets
por: Joel, Luke Oluwaseye, et al.
Publicado: (2024)
por: Joel, Luke Oluwaseye, et al.
Publicado: (2024)
3D Optimization for AI Inference Scaling: Balancing Accuracy, Cost, and Latency
por: Jung, Minseok, et al.
Publicado: (2025)
por: Jung, Minseok, et al.
Publicado: (2025)
Ejemplares similares
-
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
por: Piskala, Deepak Babu
Publicado: (2026) -
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
por: Piskala, Deepak Babu
Publicado: (2026) -
LLMAP: LLM-Assisted Multi-Objective Route Planning with User Preferences
por: Yuan, Liangqi, et al.
Publicado: (2025) -
Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR
por: Ingle, Yash, et al.
Publicado: (2026) -
RouteLLM: Learning to Route LLMs with Preference Data
por: Ong, Isaac, et al.
Publicado: (2024)