ToolCaching: Towards Efficient Caching for LLM Tool-calling
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhai, Yi, Shen, Dian, Luo, Junzhou, Yang, Bin |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
par: Sakizli, Furkan
Publié: (2026)
par: Sakizli, Furkan
Publié: (2026)
AgentPulse: A Continuous Multi-Signal Framework for Evaluating AI Agents in Deployment
par: Gao, Yuxuan, et autres
Publié: (2026)
par: Gao, Yuxuan, et autres
Publié: (2026)
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems
par: Fakih, Mohamad, et autres
Publié: (2024)
par: Fakih, Mohamad, et autres
Publié: (2024)
Inference-Time Intervention in Large Language Models for Reliable Requirement Verification
par: Darm, Paul, et autres
Publié: (2025)
par: Darm, Paul, et autres
Publié: (2025)
NormCode Canvas: Making LLM Agentic Workflows Development Sustainable via Case-Based Reasoning
par: Guan, Xin, et autres
Publié: (2026)
par: Guan, Xin, et autres
Publié: (2026)
Codebase-Memory: Tree-Sitter-Based Knowledge Graphs for LLM Code Exploration via MCP
par: Vogel, Martin, et autres
Publié: (2026)
par: Vogel, Martin, et autres
Publié: (2026)
Toolsuite for Implementing Multiagent Systems Based on Communication Protocols
par: Chopra, Amit K., et autres
Publié: (2025)
par: Chopra, Amit K., et autres
Publié: (2025)
Semantic Consensus: Process-Aware Conflict Detection and Resolution for Enterprise Multi-Agent LLM Systems
par: Acharya, Vivek
Publié: (2026)
par: Acharya, Vivek
Publié: (2026)
ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents
par: Rafique, Mofasshara, et autres
Publié: (2026)
par: Rafique, Mofasshara, et autres
Publié: (2026)
Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets
par: Sakizli, Furkan
Publié: (2026)
par: Sakizli, Furkan
Publié: (2026)
A Framework for Testing and Adapting REST APIs as LLM Tools
par: Bandlamudi, Jayachandu, et autres
Publié: (2025)
par: Bandlamudi, Jayachandu, et autres
Publié: (2025)
Neural Theorem Proving for Verification Conditions: A Real-World Benchmark
par: Xu, Qiyuan, et autres
Publié: (2026)
par: Xu, Qiyuan, et autres
Publié: (2026)
GRETEL: A Goal-driven Retrieval and Execution-based Trial Framework for LLM Tool Selection Enhancing
par: Wu, Zongze, et autres
Publié: (2025)
par: Wu, Zongze, et autres
Publié: (2025)
CoTran: An LLM-based Code Translator using Reinforcement Learning with Feedback from Compiler and Symbolic Execution
par: Jana, Prithwish, et autres
Publié: (2023)
par: Jana, Prithwish, et autres
Publié: (2023)
Engineering Systems for Data Analysis Using Interactive Structured Inductive Programming
par: Surana, Shraddha, et autres
Publié: (2025)
par: Surana, Shraddha, et autres
Publié: (2025)
Mind the GAP: Text Safety Does Not Transfer to Tool-Call Safety in LLM Agents
par: Cartagena, Arnold, et autres
Publié: (2026)
par: Cartagena, Arnold, et autres
Publié: (2026)
AI-PROPELLER: Warehouse-Scale Interprocedural Code Layout Optimization with AlphaEvolve
par: Ananda, Chaitanya Mamatha, et autres
Publié: (2026)
par: Ananda, Chaitanya Mamatha, et autres
Publié: (2026)
Bridging MDE and AI: A Systematic Review of Domain-Specific Languages and Model-Driven Practices in AI Software Systems Engineering
par: Raedler, Simon, et autres
Publié: (2023)
par: Raedler, Simon, et autres
Publié: (2023)
Speculative Automated Refactoring of Imperative Deep Learning Programs to Graph Execution
par: Khatchadourian, Raffi, et autres
Publié: (2025)
par: Khatchadourian, Raffi, et autres
Publié: (2025)
Context: Proactive Goal-Directed Intelligence via Composable Sandboxed Programs, Declarative Wiring, and Structured Interaction
par: Magarshak, Gregory
Publié: (2026)
par: Magarshak, Gregory
Publié: (2026)
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
par: Iannelli, Michael, et autres
Publié: (2024)
par: Iannelli, Michael, et autres
Publié: (2024)
ATLAS: A Layered Constraint-Guided Framework for Structured Artifact Generation in LLM-Assisted MDE
par: Ma, Tong, et autres
Publié: (2025)
par: Ma, Tong, et autres
Publié: (2025)
Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents
par: Tuan, Thanh Luong, et autres
Publié: (2026)
par: Tuan, Thanh Luong, et autres
Publié: (2026)
An advanced AI driven database system
par: Tedeschi, M., et autres
Publié: (2025)
par: Tedeschi, M., et autres
Publié: (2025)
Proof Automation with Large Language Models
par: Lu, Minghai, et autres
Publié: (2024)
par: Lu, Minghai, et autres
Publié: (2024)
Securing the Agent: Vendor-Neutral, Multitenant Enterprise Retrieval and Tool Use
par: Arceo, Francisco Javier, et autres
Publié: (2026)
par: Arceo, Francisco Javier, et autres
Publié: (2026)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
par: Xia, Bowei, et autres
Publié: (2026)
par: Xia, Bowei, et autres
Publié: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
par: Baldonado, Juan Manuel, et autres
Publié: (2025)
par: Baldonado, Juan Manuel, et autres
Publié: (2025)
UI-CUBE: Enterprise-Grade Computer Use Agent Benchmarking Beyond Task Accuracy to Operational Reliability
par: Cristescu, Horia, et autres
Publié: (2025)
par: Cristescu, Horia, et autres
Publié: (2025)
Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains
par: Calboreanu, Elias
Publié: (2026)
par: Calboreanu, Elias
Publié: (2026)
A Machine Learning Approach Towards SKILL Code Autocompletion
par: Dehaerne, Enrique, et autres
Publié: (2023)
par: Dehaerne, Enrique, et autres
Publié: (2023)
When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges
par: Darshan, Parth, et autres
Publié: (2026)
par: Darshan, Parth, et autres
Publié: (2026)
Multicalibration for LLM-based Code Generation
par: Campos, Viola, et autres
Publié: (2025)
par: Campos, Viola, et autres
Publié: (2025)
Evaluating Temporal Semantic Caching and Workflow Optimization in Agentic Plan-Execute Pipelines
par: Merchant, Alimurtaza Mustafa, et autres
Publié: (2026)
par: Merchant, Alimurtaza Mustafa, et autres
Publié: (2026)
StepCache: Step-Level Reuse with Lightweight Verification and Selective Patching for LLM Serving
par: Nouri, Azam
Publié: (2026)
par: Nouri, Azam
Publié: (2026)
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
par: Zhang, Jiahao, et autres
Publié: (2026)
par: Zhang, Jiahao, et autres
Publié: (2026)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
par: Agrawal, Lakshya A, et autres
Publié: (2025)
par: Agrawal, Lakshya A, et autres
Publié: (2025)
Efficacy of Various Large Language Models in Generating Smart Contracts
par: Chatterjee, Siddhartha, et autres
Publié: (2024)
par: Chatterjee, Siddhartha, et autres
Publié: (2024)
Evaluating the Limitations of Local LLMs in Solving Complex Programming Challenges
par: Matotek, Kadin, et autres
Publié: (2025)
par: Matotek, Kadin, et autres
Publié: (2025)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
par: Sun, Zhensu, et autres
Publié: (2024)
par: Sun, Zhensu, et autres
Publié: (2024)
Documents similaires
-
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
par: Sakizli, Furkan
Publié: (2026) -
AgentPulse: A Continuous Multi-Signal Framework for Evaluating AI Agents in Deployment
par: Gao, Yuxuan, et autres
Publié: (2026) -
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems
par: Fakih, Mohamad, et autres
Publié: (2024) -
Inference-Time Intervention in Large Language Models for Reliable Requirement Verification
par: Darm, Paul, et autres
Publié: (2025) -
NormCode Canvas: Making LLM Agentic Workflows Development Sustainable via Case-Based Reasoning
par: Guan, Xin, et autres
Publié: (2026)