CATP-LLM: Empowering Large Language Models for Cost-Aware Tool Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Duo, Wang, Jinghe, Meng, Yuan, Zhang, Yanning, Sun, Le, Wang, Zhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use
von: Liu, Hanbing, et al.
Veröffentlicht: (2026)
von: Liu, Hanbing, et al.
Veröffentlicht: (2026)
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
von: Chen, Lei, et al.
Veröffentlicht: (2026)
von: Chen, Lei, et al.
Veröffentlicht: (2026)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
MoE$^2$: Optimizing Collaborative Inference for Edge Large Language Models
von: Jin, Lyudong, et al.
Veröffentlicht: (2025)
von: Jin, Lyudong, et al.
Veröffentlicht: (2025)
WirelessLLM: Empowering Large Language Models Towards Wireless Intelligence
von: Shao, Jiawei, et al.
Veröffentlicht: (2024)
von: Shao, Jiawei, et al.
Veröffentlicht: (2024)
A Cost-Benefit Analysis of On-Premise Large Language Model Deployment: Breaking Even with Commercial LLM Services
von: Pan, Guanzhong, et al.
Veröffentlicht: (2025)
von: Pan, Guanzhong, et al.
Veröffentlicht: (2025)
LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddings
von: Wang, Duo, et al.
Veröffentlicht: (2024)
von: Wang, Duo, et al.
Veröffentlicht: (2024)
EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation
von: Zhang, Shu-Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Shu-Hao, et al.
Veröffentlicht: (2026)
BudgetThinker: Empowering Budget-aware LLM Reasoning with Control Tokens
von: Wen, Hao, et al.
Veröffentlicht: (2025)
von: Wen, Hao, et al.
Veröffentlicht: (2025)
Context-Aware Probabilistic Modeling with LLM for Multimodal Time Series Forecasting
von: Yao, Yueyang, et al.
Veröffentlicht: (2025)
von: Yao, Yueyang, et al.
Veröffentlicht: (2025)
ClimateLLM: Efficient Weather Forecasting via Frequency-Aware Large Language Models
von: Li, Shixuan, et al.
Veröffentlicht: (2025)
von: Li, Shixuan, et al.
Veröffentlicht: (2025)
GPrune-LLM: Generalization-Aware Structured Pruning for Large Language Models
von: Liu, Xiaoyun, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoyun, et al.
Veröffentlicht: (2026)
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model
von: Jiang, Haitao, et al.
Veröffentlicht: (2023)
von: Jiang, Haitao, et al.
Veröffentlicht: (2023)
PAT: Pruning-Aware Tuning for Large Language Models
von: Liu, Yijiang, et al.
Veröffentlicht: (2024)
von: Liu, Yijiang, et al.
Veröffentlicht: (2024)
EfficientLLM: Efficiency in Large Language Models
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2025)
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2025)
Large Language Models as Tool Makers
von: Cai, Tianle, et al.
Veröffentlicht: (2023)
von: Cai, Tianle, et al.
Veröffentlicht: (2023)
Empowering GNNs via Edge-Aware Weisfeiler-Leman Algorithm
von: Liu, Meng, et al.
Veröffentlicht: (2022)
von: Liu, Meng, et al.
Veröffentlicht: (2022)
Infeasibility Aware Large Language Models for Combinatorial Optimization
von: Wang, Yakun, et al.
Veröffentlicht: (2026)
von: Wang, Yakun, et al.
Veröffentlicht: (2026)
PiSSA: Principal Singular Values and Singular Vectors Adaptation of Large Language Models
von: Meng, Fanxu, et al.
Veröffentlicht: (2024)
von: Meng, Fanxu, et al.
Veröffentlicht: (2024)
Aligning Diffusion Language Models via Unpaired Preference Optimization
von: Jindal, Vaibhav, et al.
Veröffentlicht: (2025)
von: Jindal, Vaibhav, et al.
Veröffentlicht: (2025)
Demonstration of DB-GPT: Next Generation Data Interaction System Empowered by Large Language Models
von: Xue, Siqiao, et al.
Veröffentlicht: (2024)
von: Xue, Siqiao, et al.
Veröffentlicht: (2024)
PickLLM: Context-Aware RL-Assisted Large Language Model Routing
von: Sikeridis, Dimitrios, et al.
Veröffentlicht: (2024)
von: Sikeridis, Dimitrios, et al.
Veröffentlicht: (2024)
Time-LLM: Time Series Forecasting by Reprogramming Large Language Models
von: Jin, Ming, et al.
Veröffentlicht: (2023)
von: Jin, Ming, et al.
Veröffentlicht: (2023)
TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models
von: Mu, Lin, et al.
Veröffentlicht: (2026)
von: Mu, Lin, et al.
Veröffentlicht: (2026)
Empowering Autonomous Driving with Large Language Models: A Safety Perspective
von: Wang, Yixuan, et al.
Veröffentlicht: (2023)
von: Wang, Yixuan, et al.
Veröffentlicht: (2023)
MetaTool: Facilitating Large Language Models to Master Tools with Meta-task Augmentation
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
Empowering Time Series Analysis with Foundation Models: A Comprehensive Survey
von: Ye, Jiexia, et al.
Veröffentlicht: (2024)
von: Ye, Jiexia, et al.
Veröffentlicht: (2024)
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
LLM4GNAS: A Large Language Model Based Toolkit for Graph Neural Architecture Search
von: Gao, Yang, et al.
Veröffentlicht: (2025)
von: Gao, Yang, et al.
Veröffentlicht: (2025)
Beyond Correctness: Confidence-Aware Reward Modeling for Enhancing Large Language Model Reasoning
von: He, Qianxi, et al.
Veröffentlicht: (2025)
von: He, Qianxi, et al.
Veröffentlicht: (2025)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
Saliency-Aware Regularized Quantization Calibration for Large Language Models
von: Zhao, Yanlong, et al.
Veröffentlicht: (2026)
von: Zhao, Yanlong, et al.
Veröffentlicht: (2026)
T-LLM: Teaching Large Language Models to Forecast Time Series via Temporal Distillation
von: Guo, Suhan, et al.
Veröffentlicht: (2026)
von: Guo, Suhan, et al.
Veröffentlicht: (2026)
Sparsity-Aware Low-Rank Representation for Efficient Fine-Tuning of Large Language Models
von: Zhang, Longteng, et al.
Veröffentlicht: (2026)
von: Zhang, Longteng, et al.
Veröffentlicht: (2026)
PT$^2$-LLM: Post-Training Ternarization for Large Language Models
von: Yan, Xianglong, et al.
Veröffentlicht: (2025)
von: Yan, Xianglong, et al.
Veröffentlicht: (2025)
CataLM: Empowering Catalyst Design Through Large Language Models
von: Wang, Ludi, et al.
Veröffentlicht: (2024)
von: Wang, Ludi, et al.
Veröffentlicht: (2024)
ICAD-LLM: One-for-All Anomaly Detection via In-Context Learning with Large Language Models
von: Wu, Zhongyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhongyuan, et al.
Veröffentlicht: (2025)
OR-Toolformer: Modeling and Solving Operations Research Problems with Tool Augmented Large Language Models
von: Zhang, Jianzhang, et al.
Veröffentlicht: (2025)
von: Zhang, Jianzhang, et al.
Veröffentlicht: (2025)
Provable Benefits of In-Tool Learning for Large Language Models
von: Houliston, Sam, et al.
Veröffentlicht: (2025)
von: Houliston, Sam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use
von: Liu, Hanbing, et al.
Veröffentlicht: (2026) -
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
von: Chen, Lei, et al.
Veröffentlicht: (2026) -
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
von: Luo, Haipeng, et al.
Veröffentlicht: (2025) -
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models
von: Xu, Hefei, et al.
Veröffentlicht: (2026) -
MoE$^2$: Optimizing Collaborative Inference for Edge Large Language Models
von: Jin, Lyudong, et al.
Veröffentlicht: (2025)