Don't Break the Cache: An Evaluation of Prompt Caching for Long-Horizon Agentic Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lumer, Elias, Nizar, Faheem, Jangiti, Akshaya, Frank, Kevin, Gulati, Anmol, Phadate, Mandar, Subbiah, Vamse Kumar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agent-as-a-Graph: Knowledge Graph-Based Tool and Agent Retrieval for LLM Multi-Agent Systems
von: Nizar, Faheem, et al.
Veröffentlicht: (2025)
von: Nizar, Faheem, et al.
Veröffentlicht: (2025)
Tool-to-Agent Retrieval: Bridging Tools and Agents for Scalable LLM Multi-Agent Systems
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
Jackal: A Real-World Execution-Based Benchmark Evaluating Large Language Models on Text-to-JQL Tasks
von: Frank, Kevin, et al.
Veröffentlicht: (2025)
von: Frank, Kevin, et al.
Veröffentlicht: (2025)
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
von: Gulati, Anmol, et al.
Veröffentlicht: (2026)
von: Gulati, Anmol, et al.
Veröffentlicht: (2026)
Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL
von: Murali, Vishnu, et al.
Veröffentlicht: (2026)
von: Murali, Vishnu, et al.
Veröffentlicht: (2026)
Is Grep All You Need? How Agent Harnesses Reshape Agentic Search
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
MemTool: Optimizing Short-Term Memory Management for Dynamic Tool Calling in LLM Agent Multi-Turn Conversations
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents
von: Onweller, Hailey, et al.
Veröffentlicht: (2026)
von: Onweller, Hailey, et al.
Veröffentlicht: (2026)
Graph RAG-Tool Fusion
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
Toolshed: Scale Tool-Equipped Agents with Advanced RAG-Tool Fusion and Tool Knowledge Bases
von: Lumer, Elias, et al.
Veröffentlicht: (2024)
von: Lumer, Elias, et al.
Veröffentlicht: (2024)
Rethinking Retrieval: From Traditional Retrieval Augmented Generation to Agentic and Non-Vector Reasoning Systems in the Financial Domain for Large Language Models
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks
von: Chan, Brian J, et al.
Veröffentlicht: (2024)
von: Chan, Brian J, et al.
Veröffentlicht: (2024)
Comparison of Text-Based and Image-Based Retrieval in Multimodal Retrieval Augmented Generation Large Language Model Systems
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
von: Lumer, Elias, et al.
Veröffentlicht: (2025)
SideQuest: Model-Driven KV Cache Management for Long-Horizon Agentic Reasoning
von: Kariyappa, Sanjay, et al.
Veröffentlicht: (2026)
von: Kariyappa, Sanjay, et al.
Veröffentlicht: (2026)
Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs
von: Boroujeni, Sayed Pedram Haeri, et al.
Veröffentlicht: (2026)
von: Boroujeni, Sayed Pedram Haeri, et al.
Veröffentlicht: (2026)
Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL
von: Wu, Ian, et al.
Veröffentlicht: (2026)
von: Wu, Ian, et al.
Veröffentlicht: (2026)
Beyond Rows to Reasoning: Agentic Retrieval for Multimodal Spreadsheet Understanding and Editing
von: Gulati, Anmol, et al.
Veröffentlicht: (2026)
von: Gulati, Anmol, et al.
Veröffentlicht: (2026)
vCache: Verified Semantic Prompt Caching
von: Schroeder, Luis Gaspar, et al.
Veröffentlicht: (2025)
von: Schroeder, Luis Gaspar, et al.
Veröffentlicht: (2025)
Don't be so Stief! Learning KV Cache low-rank approximation over the Stiefel manifold
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break
von: Wang, Xinyu Jessica, et al.
Veröffentlicht: (2026)
von: Wang, Xinyu Jessica, et al.
Veröffentlicht: (2026)
AgenticCache: Cache-Driven Asynchronous Planning for Embodied AI Agents
von: Kim, Hojoon, et al.
Veröffentlicht: (2026)
von: Kim, Hojoon, et al.
Veröffentlicht: (2026)
CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs
von: Fahey, Ryan
Veröffentlicht: (2026)
von: Fahey, Ryan
Veröffentlicht: (2026)
Evaluating Temporal Semantic Caching and Workflow Optimization in Agentic Plan-Execute Pipelines
von: Merchant, Alimurtaza Mustafa, et al.
Veröffentlicht: (2026)
von: Merchant, Alimurtaza Mustafa, et al.
Veröffentlicht: (2026)
Agentic Aggregation for Parallel Scaling of Long-Horizon Agentic Tasks
von: Lee, Yoonsang, et al.
Veröffentlicht: (2026)
von: Lee, Yoonsang, et al.
Veröffentlicht: (2026)
Probing the Prompt KV Cache: Where It Becomes Dispensable
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
Financial self‐efficacy of consumers: A review and research agenda
von: Anmol Gulati, et al.
Veröffentlicht: (2024)
von: Anmol Gulati, et al.
Veröffentlicht: (2024)
Efficient Long-Horizon GUI Agents via Training-Free KV Cache Compression
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
(Don't) Mind the Gap
von: Frank, Natalie Priebe, et al.
Veröffentlicht: (2025)
von: Frank, Natalie Priebe, et al.
Veröffentlicht: (2025)
Coded Caching with Shared Caches and Private Caches
von: Peter, Elizabath, et al.
Veröffentlicht: (2022)
von: Peter, Elizabath, et al.
Veröffentlicht: (2022)
Leyline: KV Cache Directives for Agentic Inference
von: Ma, Bole, et al.
Veröffentlicht: (2026)
von: Ma, Bole, et al.
Veröffentlicht: (2026)
VNF-Cache: An In-Network Key-Value Store Cache Based on Network Function Virtualization
von: Farias, Bruno E., et al.
Veröffentlicht: (2025)
von: Farias, Bruno E., et al.
Veröffentlicht: (2025)
Systematic Evaluation of Randomized Cache Designs against Cache Occupancy
von: Chakraborty, Anirban, et al.
Veröffentlicht: (2023)
von: Chakraborty, Anirban, et al.
Veröffentlicht: (2023)
End-to-End Long Document Summarization using Gradient Caching
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
Prompt Injection Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Semantic Caching
von: Gosmar, Diego, et al.
Veröffentlicht: (2026)
von: Gosmar, Diego, et al.
Veröffentlicht: (2026)
You Don't Need Prompt Engineering Anymore: The Prompting Inversion
von: Khan, Imran
Veröffentlicht: (2025)
von: Khan, Imran
Veröffentlicht: (2025)
Cache Your Prompt When It's Green: Carbon-Aware Caching for Large Language Model Serving
von: Tian, Yuyang, et al.
Veröffentlicht: (2025)
von: Tian, Yuyang, et al.
Veröffentlicht: (2025)
Addendum: Systematic Evaluation of Randomized Cache Designs against Cache Occupancy
von: Chakraborty, Anirban, et al.
Veröffentlicht: (2025)
von: Chakraborty, Anirban, et al.
Veröffentlicht: (2025)
IE as Cache: Information Extraction Enhanced Agentic Reasoning
von: Lv, Hang, et al.
Veröffentlicht: (2026)
von: Lv, Hang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Agent-as-a-Graph: Knowledge Graph-Based Tool and Agent Retrieval for LLM Multi-Agent Systems
von: Nizar, Faheem, et al.
Veröffentlicht: (2025) -
Tool-to-Agent Retrieval: Bridging Tools and Agents for Scalable LLM Multi-Agent Systems
von: Lumer, Elias, et al.
Veröffentlicht: (2025) -
Jackal: A Real-World Execution-Based Benchmark Evaluating Large Language Models on Text-to-JQL Tasks
von: Frank, Kevin, et al.
Veröffentlicht: (2025) -
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
von: Sen, Sahil, et al.
Veröffentlicht: (2026) -
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
von: Gulati, Anmol, et al.
Veröffentlicht: (2026)