Lei, J., & Ilager, S. (2026). ARKV: Adaptive and Resource-Efficient KV Cache Management under Limited Memory Budget for Long-Context Inference in LLMs.
Chicago Style (17th ed.) CitationLei, Jianlong, and Shashikant Ilager. ARKV: Adaptive and Resource-Efficient KV Cache Management Under Limited Memory Budget for Long-Context Inference in LLMs. 2026.
MLA (9th ed.) CitationLei, Jianlong, and Shashikant Ilager. ARKV: Adaptive and Resource-Efficient KV Cache Management Under Limited Memory Budget for Long-Context Inference in LLMs. 2026.
Warning: These citations may not always be 100% accurate.