IPR: Intelligent Prompt Routing with User-Controlled Quality-Cost Trade-offs
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Aosong, Srinivasan, Balasubramaniam, Zhou, Yun, Xu, Zhichao, Zhou, Kang, Guan, Sheng, Chen, Yueyan, Wu, Xian, Kulkarni, Ninad, Zhang, Yi, Shen, Zhengyuan, Bespalov, Dmitriy, Mishra, Soumya Smruti, Teng, Yifei, Wang, Darren Yow-Bang, Ding, Haibo, Cheong, Lin Lee |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SLOT: Structuring the Output of Large Language Models
by: Wang, Darren Yow-Bang, et al.
Published: (2025)
by: Wang, Darren Yow-Bang, et al.
Published: (2025)
Diffusion Language Model Inference with Monte Carlo Tree Search
by: Huang, Zheng, et al.
Published: (2025)
by: Huang, Zheng, et al.
Published: (2025)
TaeBench: Improving Quality of Toxic Adversarial Examples
by: Zhu, Xuan, et al.
Published: (2024)
by: Zhu, Xuan, et al.
Published: (2024)
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation
by: Schwartz, Daniel, et al.
Published: (2025)
by: Schwartz, Daniel, et al.
Published: (2025)
BayesFlow: A Probability Inference Framework for Meta-Agent Assisted Workflow Generation
by: Yuan, Bo, et al.
Published: (2026)
by: Yuan, Bo, et al.
Published: (2026)
A Systematic Survey of Automatic Prompt Optimization Techniques
by: Ramnath, Kiran, et al.
Published: (2025)
by: Ramnath, Kiran, et al.
Published: (2025)
Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Task
by: Ghashami, Mina, et al.
Published: (2024)
by: Ghashami, Mina, et al.
Published: (2024)
CSPLADE: Learned Sparse Retrieval with Causal Language Models
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
An Empirical Study of Automating Agent Evaluation
by: Zhou, Kang, et al.
Published: (2026)
by: Zhou, Kang, et al.
Published: (2026)
Reinforcement Learning for Self-Improving Agent with Skill Library
by: Wang, Jiongxiao, et al.
Published: (2025)
by: Wang, Jiongxiao, et al.
Published: (2025)
Towards Building a Robust Toxicity Predictor
by: Bespalov, Dmitriy, et al.
Published: (2024)
by: Bespalov, Dmitriy, et al.
Published: (2024)
TurboFuzzLLM: Turbocharging Mutation-based Fuzzing for Effectively Jailbreaking Large Language Models in Practice
by: Goel, Aman, et al.
Published: (2025)
by: Goel, Aman, et al.
Published: (2025)
LaRS: Latent Reasoning Skills for Chain-of-Thought Reasoning
by: Xu, Zifan, et al.
Published: (2023)
by: Xu, Zifan, et al.
Published: (2023)
Optimal Routing in the Presence of Hooks: Three Case Studies
by: Chitra, Tarun, et al.
Published: (2025)
by: Chitra, Tarun, et al.
Published: (2025)
Stochastic Multipath Routing for High-Throughput Entanglement Distribution in Quantum Repeater Networks
by: Mishra, Ankit, et al.
Published: (2026)
by: Mishra, Ankit, et al.
Published: (2026)
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
PromptPrism: A Linguistically-Inspired Taxonomy for Prompts
by: Jeoung, Sullam, et al.
Published: (2025)
by: Jeoung, Sullam, et al.
Published: (2025)
IPR-1: Interactive Physical Reasoner
by: Zhang, Mingyu, et al.
Published: (2025)
by: Zhang, Mingyu, et al.
Published: (2025)
Less is More for Improving Automatic Evaluation of Factual Consistency
by: Wang, Tong, et al.
Published: (2024)
by: Wang, Tong, et al.
Published: (2024)
Sequential Decision Making in Stochastic Games with Incomplete Preferences over Temporal Objectives
by: Kulkarni, Abhishek Ninad, et al.
Published: (2025)
by: Kulkarni, Abhishek Ninad, et al.
Published: (2025)
IPR Management Smart Contracts
by: Giorgino, Luca
Published: (2026)
by: Giorgino, Luca
Published: (2026)
Enhancing Trust and Empathy in Marketing: Strategic AI and Human Influencer Selection for Optimized Content Persuasion
by: Yueyan Zhang, et al.
Published: (2024)
by: Yueyan Zhang, et al.
Published: (2024)
Learning to Ideate for Machine Learning Engineering Agents
by: Zhang, Yunxiang, et al.
Published: (2026)
by: Zhang, Yunxiang, et al.
Published: (2026)
Route Experts by Sequence, not by Token
by: Wen, Tiansheng, et al.
Published: (2025)
by: Wen, Tiansheng, et al.
Published: (2025)
The Mechanism Behind the Therapeutic Role of Alpha‐Tocopherol in Mitigating Hypobaric Hypoxia–Induced Eye Defect in Drosophila melanogaster
by: Seekha Naik, et al.
Published: (2025)
by: Seekha Naik, et al.
Published: (2025)
Mixed-Type Tabular Data Synthesis with Score-based Diffusion in Latent Space
by: Zhang, Hengrui, et al.
Published: (2023)
by: Zhang, Hengrui, et al.
Published: (2023)
OpenTab: Advancing Large Language Models as Open-domain Table Reasoners
by: Kong, Kezhi, et al.
Published: (2024)
by: Kong, Kezhi, et al.
Published: (2024)
Dementia Strategies of England and Glasgow Declaration: What Requires to be Done?
by: Smruti Bulsari
Published: (2024)
by: Smruti Bulsari
Published: (2024)
An adaptive multimesh rational approximation scheme for the spectral fractional Laplacian
by: Bespalov, Alex, et al.
Published: (2025)
by: Bespalov, Alex, et al.
Published: (2025)
Convergence analysis of the adaptive stochastic collocation finite element method
by: Bespalov, Alex, et al.
Published: (2024)
by: Bespalov, Alex, et al.
Published: (2024)
Impact of Community Proactive Health Management Application on Electronic Health Literacy and Self‐Management of Hypertensive Patients
by: Yueyan Jiang
Published: (2024)
by: Yueyan Jiang
Published: (2024)
Improved Object-Based Style Transfer with Single Deep Network
by: Kulkarni, Harshmohan, et al.
Published: (2024)
by: Kulkarni, Harshmohan, et al.
Published: (2024)
Role Of IPR In Digital Marketing And E-Commerce
by: Dr. Profe. Babasaheb savant, et al.
Published: (2025)
by: Dr. Profe. Babasaheb savant, et al.
Published: (2025)
Routing End User Queries to Enterprise Databases
by: Sudarshan, Saikrishna, et al.
Published: (2026)
by: Sudarshan, Saikrishna, et al.
Published: (2026)
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
by: Naik, Ninad
Published: (2024)
by: Naik, Ninad
Published: (2024)
On extended 1-perfect bitrades
by: Bespalov, Evgeny A., et al.
Published: (2020)
by: Bespalov, Evgeny A., et al.
Published: (2020)
Cloud Computing Review: A Decade of Research
by: Swain, Smruti Rekha
Published: (2026)
by: Swain, Smruti Rekha
Published: (2026)
QuantAgent: Price-Driven Multi-Agent LLMs for High-Frequency Trading
by: Xiong, Fei, et al.
Published: (2025)
by: Xiong, Fei, et al.
Published: (2025)
On the Last-Iterate Convergence of Shuffling Gradient Methods
by: Liu, Zijian, et al.
Published: (2024)
by: Liu, Zijian, et al.
Published: (2024)
Nonconvex Stochastic Optimization under Heavy-Tailed Noises: Optimal Convergence without Gradient Clipping
by: Liu, Zijian, et al.
Published: (2024)
by: Liu, Zijian, et al.
Published: (2024)
Similar Items
-
SLOT: Structuring the Output of Large Language Models
by: Wang, Darren Yow-Bang, et al.
Published: (2025) -
Diffusion Language Model Inference with Monte Carlo Tree Search
by: Huang, Zheng, et al.
Published: (2025) -
TaeBench: Improving Quality of Toxic Adversarial Examples
by: Zhu, Xuan, et al.
Published: (2024) -
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation
by: Schwartz, Daniel, et al.
Published: (2025) -
BayesFlow: A Probability Inference Framework for Meta-Agent Assisted Workflow Generation
by: Yuan, Bo, et al.
Published: (2026)