Prompt Smart, Pay Less: Cost-Aware APO for Real-World Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Choudhari, Jayesh, Singh, Piyush Kumar, McIlwraith, Douglas, Nair, Snehal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Hidden Costs of Domain Fine-Tuning: Pii-Bearing Data Degrades Safety and Increases Leakage
by: Choudhari, Jayesh, et al.
Published: (2026)
by: Choudhari, Jayesh, et al.
Published: (2026)
MOSAIC: Modular Opinion Summarization using Aspect Identification and Clustering
by: Singh, Piyush Kumar, et al.
Published: (2026)
by: Singh, Piyush Kumar, et al.
Published: (2026)
The Biosecurity Blind Spot: Systematic Dual-use Detection in Open Science Infrastructure
by: Sharma, Vasudha, et al.
Published: (2026)
by: Sharma, Vasudha, et al.
Published: (2026)
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025)
by: Jain, Naman, et al.
Published: (2025)
APO: Alpha-Divergence Preference Optimization
by: Zixian, Wang
Published: (2025)
by: Zixian, Wang
Published: (2025)
Multilingual-To-Multimodal (M2M): Unlocking New Languages with Monolingual Text
by: Pasi, Piyush Singh
Published: (2026)
by: Pasi, Piyush Singh
Published: (2026)
Paying Alignment Tax with Contrastive Learning
by: Korkmaz, Buse Sibel, et al.
Published: (2025)
by: Korkmaz, Buse Sibel, et al.
Published: (2025)
Weak Supervision for Real World Graphs
by: Nair, Pratheeksha, et al.
Published: (2025)
by: Nair, Pratheeksha, et al.
Published: (2025)
Mask-Enhanced Autoregressive Prediction: Pay Less Attention to Learn More
by: Zhuang, Xialie, et al.
Published: (2025)
by: Zhuang, Xialie, et al.
Published: (2025)
Pay Less Attention to Function Words for Free Robustness of Vision-Language Models
by: Tian, Qiwei, et al.
Published: (2025)
by: Tian, Qiwei, et al.
Published: (2025)
Pay for Hints, Not Answers: LLM Shepherding for Cost-Efficient Inference
by: Dong, Ziming, et al.
Published: (2026)
by: Dong, Ziming, et al.
Published: (2026)
Medial femorotibial and femoropatellar articular cartilage delamination following intra‐articular therapy with triamcinolone and gentamicin in five yearlings
by: C. P. Beggan, et al.
Published: (2025)
by: C. P. Beggan, et al.
Published: (2025)
MDPs with a State Sensing Cost
by: Kapoor, Vansh, et al.
Published: (2025)
by: Kapoor, Vansh, et al.
Published: (2025)
PromptWise: Online Learning for Cost-Aware Prompt Assignment in Generative Models
by: Hu, Xiaoyan, et al.
Published: (2025)
by: Hu, Xiaoyan, et al.
Published: (2025)
Paying Less Generalization Tax: A Cross-Domain Generalization Study of RL Training for LLM Agents
by: Liu, Zhihan, et al.
Published: (2026)
by: Liu, Zhihan, et al.
Published: (2026)
CAPO: Cost-Aware Prompt Optimization
by: Zehle, Tom, et al.
Published: (2025)
by: Zehle, Tom, et al.
Published: (2025)
Risk-Aware Linear Bandits: Theory and Applications in Smart Order Routing
by: Ji, Jingwei, et al.
Published: (2022)
by: Ji, Jingwei, et al.
Published: (2022)
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization
by: Hong, Minjie, et al.
Published: (2025)
by: Hong, Minjie, et al.
Published: (2025)
MO-CAPO: Multi-Objective Cost-Aware Prompt Optimization
by: Büssing, Jan, et al.
Published: (2026)
by: Büssing, Jan, et al.
Published: (2026)
Cost-Aware Learning
by: Mohri, Clara, et al.
Published: (2026)
by: Mohri, Clara, et al.
Published: (2026)
PromptWizard: Task-Aware Prompt Optimization Framework
by: Agarwal, Eshaan, et al.
Published: (2024)
by: Agarwal, Eshaan, et al.
Published: (2024)
Debiasing Text Safety Classifiers through a Fairness-Aware Ensemble
by: Sturman, Olivia, et al.
Published: (2024)
by: Sturman, Olivia, et al.
Published: (2024)
Prompting Forgetting: Unlearning in GANs via Textual Guidance
by: Nagasubramaniam, Piyush, et al.
Published: (2025)
by: Nagasubramaniam, Piyush, et al.
Published: (2025)
Frontiers of Deep Learning: From Novel Application to Real-World Deployment
by: Xie, Rui
Published: (2024)
by: Xie, Rui
Published: (2024)
Pay Less Attention to Deceptive Artifacts: Robust Detection of Compressed Deepfakes on Online Social Networks
by: Li, Manyi, et al.
Published: (2025)
by: Li, Manyi, et al.
Published: (2025)
Continuous Approximations for Improving Quantization Aware Training of LLMs
by: Li, He, et al.
Published: (2024)
by: Li, He, et al.
Published: (2024)
Towards Scalable & Efficient Interaction-Aware Planning in Autonomous Vehicles using Knowledge Distillation
by: Gupta, Piyush, et al.
Published: (2024)
by: Gupta, Piyush, et al.
Published: (2024)
Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data
by: de Campos, Andresa Rodrigues, et al.
Published: (2026)
by: de Campos, Andresa Rodrigues, et al.
Published: (2026)
What is in a Price? Estimating Willingness-to-Pay with Bayesian Hierarchical Models
by: Pillai, Srijesh, et al.
Published: (2025)
by: Pillai, Srijesh, et al.
Published: (2025)
Warmer for Less: A Cost-Efficient Strategy for Cold-Start Recommendations at Pinterest
by: Ebrahimi, Saeed, et al.
Published: (2025)
by: Ebrahimi, Saeed, et al.
Published: (2025)
InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting
by: Sabbaghi, Mahdi, et al.
Published: (2026)
by: Sabbaghi, Mahdi, et al.
Published: (2026)
Integrating Asynchronous AdaBoost into Federated Learning: Five Real World Applications
by: Oghlukyan, Arthur, et al.
Published: (2025)
by: Oghlukyan, Arthur, et al.
Published: (2025)
Information Filtering Networks: Theoretical Foundations, Generative Methodologies, and Real-World Applications
by: Aste, Tomaso
Published: (2025)
by: Aste, Tomaso
Published: (2025)
Capacity Provisioning Motivated Online Non-Convex Optimization Problem with Memory and Switching Cost
by: Vaze, Rahul, et al.
Published: (2024)
by: Vaze, Rahul, et al.
Published: (2024)
Solve Smart, Not Often: Policy Learning for Costly MILP Re-solving
by: Ai, Rui, et al.
Published: (2025)
by: Ai, Rui, et al.
Published: (2025)
When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding
by: Yang, Xu, et al.
Published: (2026)
by: Yang, Xu, et al.
Published: (2026)
Watch Your Step: A Cost-Sensitive Framework for Accelerometer-Based Fall Detection in Real-World Streaming Scenarios
by: Aderinola, Timilehin B., et al.
Published: (2025)
by: Aderinola, Timilehin B., et al.
Published: (2025)
Data Efficient Subset Training with Differential Privacy
by: Gandhi, Ninad Jayesh, et al.
Published: (2025)
by: Gandhi, Ninad Jayesh, et al.
Published: (2025)
Signal Fidelity Index-Aware Calibration for Dementia Predictions Across Heterogeneous Real-World Data
by: Cheng, Jingya, et al.
Published: (2025)
by: Cheng, Jingya, et al.
Published: (2025)
SEAFormer: A Spatial Proximity and Edge-Aware Transformer for Real-World Vehicle Routing Problems
by: Basharzad, Saeed Nasehi, et al.
Published: (2026)
by: Basharzad, Saeed Nasehi, et al.
Published: (2026)
Similar Items
-
The Hidden Costs of Domain Fine-Tuning: Pii-Bearing Data Degrades Safety and Increases Leakage
by: Choudhari, Jayesh, et al.
Published: (2026) -
MOSAIC: Modular Opinion Summarization using Aspect Identification and Clustering
by: Singh, Piyush Kumar, et al.
Published: (2026) -
The Biosecurity Blind Spot: Systematic Dual-use Detection in Open Science Infrastructure
by: Sharma, Vasudha, et al.
Published: (2026) -
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025) -
APO: Alpha-Divergence Preference Optimization
by: Zixian, Wang
Published: (2025)