Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Piskala, Deepak Babu, Raajaa, Vijay, Mishra, Sachin, Bozza, Bruno |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
LLMAP: LLM-Assisted Multi-Objective Route Planning with User Preferences
von: Yuan, Liangqi, et al.
Veröffentlicht: (2025)
von: Yuan, Liangqi, et al.
Veröffentlicht: (2025)
Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR
von: Ingle, Yash, et al.
Veröffentlicht: (2026)
von: Ingle, Yash, et al.
Veröffentlicht: (2026)
RouteLLM: Learning to Route LLMs with Preference Data
von: Ong, Isaac, et al.
Veröffentlicht: (2024)
von: Ong, Isaac, et al.
Veröffentlicht: (2024)
Hardware Aware Ensemble Selection for Balancing Predictive Accuracy and Cost
von: Maier, Jannis, et al.
Veröffentlicht: (2024)
von: Maier, Jannis, et al.
Veröffentlicht: (2024)
Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
von: Bhandarkar, Abhay, et al.
Veröffentlicht: (2025)
von: Bhandarkar, Abhay, et al.
Veröffentlicht: (2025)
Probabilistic Trust Intervals for Out of Distribution Detection
von: Singh, Gagandeep, et al.
Veröffentlicht: (2021)
von: Singh, Gagandeep, et al.
Veröffentlicht: (2021)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
Trust by Design: Skill Profiles for Transparent, Cost-Aware LLM Routing
von: Okamoto, Mika, et al.
Veröffentlicht: (2026)
von: Okamoto, Mika, et al.
Veröffentlicht: (2026)
Breaking Model Lock-in: Cost-Efficient Zero-Shot LLM Routing via a Universal Latent Space
von: Yan, Cheng, et al.
Veröffentlicht: (2026)
von: Yan, Cheng, et al.
Veröffentlicht: (2026)
Optimal Signal Decomposition-based Multi-Stage Learning for Battery Health Estimation
von: Pamshetti, Vijay Babu, et al.
Veröffentlicht: (2025)
von: Pamshetti, Vijay Babu, et al.
Veröffentlicht: (2025)
Dr.LLM: Dynamic Layer Routing in LLMs
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge
von: Zhang, Wenbo, et al.
Veröffentlicht: (2026)
von: Zhang, Wenbo, et al.
Veröffentlicht: (2026)
Preference Guided Iterated Pareto Referent Optimisation for Accessible Route Planning
von: Speziali, Paolo, et al.
Veröffentlicht: (2026)
von: Speziali, Paolo, et al.
Veröffentlicht: (2026)
LENSLLM: Unveiling Fine-Tuning Dynamics for LLM Selection
von: Zeng, Xinyue, et al.
Veröffentlicht: (2025)
von: Zeng, Xinyue, et al.
Veröffentlicht: (2025)
VAGPO: Vision-augmented Asymmetric Group Preference Optimization for Graph Routing Problems
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
von: Fu, Qichen, et al.
Veröffentlicht: (2024)
von: Fu, Qichen, et al.
Veröffentlicht: (2024)
Less is More: Improving LLM Alignment via Preference Data Selection
von: Deng, Xun, et al.
Veröffentlicht: (2025)
von: Deng, Xun, et al.
Veröffentlicht: (2025)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
von: Xia, Yuchen, et al.
Veröffentlicht: (2024)
von: Xia, Yuchen, et al.
Veröffentlicht: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
A Bi-Objective Approach to Last-Mile Delivery Routing Considering Driver Preferences
von: Mesa, Juan Pablo, et al.
Veröffentlicht: (2024)
von: Mesa, Juan Pablo, et al.
Veröffentlicht: (2024)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
GraNNite: Enabling High-Performance Execution of Graph Neural Networks on Resource-Constrained Neural Processing Units
von: Das, Arghadip, et al.
Veröffentlicht: (2025)
von: Das, Arghadip, et al.
Veröffentlicht: (2025)
SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning
von: Xu, Zhongling, et al.
Veröffentlicht: (2026)
von: Xu, Zhongling, et al.
Veröffentlicht: (2026)
Adversarial Preference Learning for Robust LLM Alignment
von: Wang, Yuanfu, et al.
Veröffentlicht: (2025)
von: Wang, Yuanfu, et al.
Veröffentlicht: (2025)
Machine Learning Framework for Early Power, Performance, and Area Estimation of RTL
von: Chattopadhyay, Anindita, et al.
Veröffentlicht: (2025)
von: Chattopadhyay, Anindita, et al.
Veröffentlicht: (2025)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
LLM Data Selection and Utilization via Dynamic Bi-level Optimization
von: Yu, Yang, et al.
Veröffentlicht: (2025)
von: Yu, Yang, et al.
Veröffentlicht: (2025)
PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment
von: Verma, Richa, et al.
Veröffentlicht: (2026)
von: Verma, Richa, et al.
Veröffentlicht: (2026)
A Cost-Effective LLM-based Approach to Identify Wildlife Trafficking in Online Marketplaces
von: Barbosa, Juliana, et al.
Veröffentlicht: (2025)
von: Barbosa, Juliana, et al.
Veröffentlicht: (2025)
MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness
von: Hathidara, Ashutosh, et al.
Veröffentlicht: (2026)
von: Hathidara, Ashutosh, et al.
Veröffentlicht: (2026)
Aligning LLM Agents by Learning Latent Preference from User Edits
von: Gao, Ge, et al.
Veröffentlicht: (2024)
von: Gao, Ge, et al.
Veröffentlicht: (2024)
Automatic Demonstration Selection for LLM-based Tabular Data Classification
von: Han, Shuchu, et al.
Veröffentlicht: (2025)
von: Han, Shuchu, et al.
Veröffentlicht: (2025)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2024)
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2024)
On the Performance of Imputation Techniques for Missing Values on Healthcare Datasets
von: Joel, Luke Oluwaseye, et al.
Veröffentlicht: (2024)
von: Joel, Luke Oluwaseye, et al.
Veröffentlicht: (2024)
3D Optimization for AI Inference Scaling: Balancing Accuracy, Cost, and Latency
von: Jung, Minseok, et al.
Veröffentlicht: (2025)
von: Jung, Minseok, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
von: Piskala, Deepak Babu
Veröffentlicht: (2026) -
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
von: Piskala, Deepak Babu
Veröffentlicht: (2026) -
LLMAP: LLM-Assisted Multi-Objective Route Planning with User Preferences
von: Yuan, Liangqi, et al.
Veröffentlicht: (2025) -
Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR
von: Ingle, Yash, et al.
Veröffentlicht: (2026) -
RouteLLM: Learning to Route LLMs with Preference Data
von: Ong, Isaac, et al.
Veröffentlicht: (2024)