Tummalapalli, P., Arayakandy, S., Pal, R., & Kundan, K. (2026). LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load.
Chicago Style (17th ed.) CitationTummalapalli, Pranay, Sahil Arayakandy, Ritam Pal, and Kautuk Kundan. LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load. 2026.
MLA (9th ed.) CitationTummalapalli, Pranay, et al. LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load. 2026.
Warning: These citations may not always be 100% accurate.