Evaluating Serverless Machine Learning Performance on Google Cloud Run
Fuente:
arXiv
Saved in:
| Main Authors: | Khatiwada, Prerana, Dhakal, Pranjal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dirigent: Lightweight Serverless Orchestration
by: Cvetković, Lazar, et al.
Published: (2024)
by: Cvetković, Lazar, et al.
Published: (2024)
Imaginary Machines: A Serverless Model for Cloud Applications
by: Wawrzoniak, Michael, et al.
Published: (2024)
by: Wawrzoniak, Michael, et al.
Published: (2024)
Taming Serverless Cold Starts Through OS Co-Design
by: Holmes, Ben, et al.
Published: (2025)
by: Holmes, Ben, et al.
Published: (2025)
Nexus: Transparent I/O Offloading for High-Density Serverless Computing
by: Park, JooYoung, et al.
Published: (2026)
by: Park, JooYoung, et al.
Published: (2026)
TrEnv-X: Transparently Share Serverless Execution Environments Across Different Functions and Nodes
by: Huang, Jialiang, et al.
Published: (2025)
by: Huang, Jialiang, et al.
Published: (2025)
Object as a Service: Simplifying Cloud-Native Development through Serverless Object Abstraction
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
Why iCloud Fails: The Category Mistake of Cloud Synchronization
by: Borrill, Paul
Published: (2026)
by: Borrill, Paul
Published: (2026)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
by: Lin, Changyuan, et al.
Published: (2025)
by: Lin, Changyuan, et al.
Published: (2025)
Funky: Cloud-Native FPGA Virtualization and Orchestration
by: Koshiba, Atsushi, et al.
Published: (2025)
by: Koshiba, Atsushi, et al.
Published: (2025)
Unlocking True Elasticity for the Cloud-Native Era with Dandelion
by: Kuchler, Tom, et al.
Published: (2025)
by: Kuchler, Tom, et al.
Published: (2025)
EdgeWeaver: Accelerating IoT Application Development Across Edge-Cloud Continuum
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2026)
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2026)
"Range as a Key" is the Key! Fast and Compact Cloud Block Store Index with RASK
by: Zhao, Haoru, et al.
Published: (2026)
by: Zhao, Haoru, et al.
Published: (2026)
FastMig: Leveraging FastFreeze to Establish Robust Service Liquidity in Cloud 2.0
by: Manatura, Sorawit, et al.
Published: (2024)
by: Manatura, Sorawit, et al.
Published: (2024)
TIDAL: Recovering Temporal Phase for Cloud Block Storage Placement from LLM-Derived Semantics
by: Tan, Difan, et al.
Published: (2026)
by: Tan, Difan, et al.
Published: (2026)
Nanvix: A Multikernel OS Design for High-Density Serverless Deployments
by: Segarra, Carlos, et al.
Published: (2026)
by: Segarra, Carlos, et al.
Published: (2026)
Optimizing SSD Caches for Cloud Block Storage Systems Using Machine Learning Approaches
by: Cheng, Chiyu, et al.
Published: (2024)
by: Cheng, Chiyu, et al.
Published: (2024)
Performance Isolation and Semantic Determinism in Efficient GPU Spatial Sharing
by: Yang, Zhenyuan, et al.
Published: (2026)
by: Yang, Zhenyuan, et al.
Published: (2026)
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
Optimizing CPU Cache Utilization in Cloud VMs with Accurate Cache Abstraction
by: Tofigh, Mani, et al.
Published: (2025)
by: Tofigh, Mani, et al.
Published: (2025)
SteelDB: Diagnosing Kernel-Space Bottlenecks in Cloud OLTP Databases
by: Kondo, Mitsumasa
Published: (2026)
by: Kondo, Mitsumasa
Published: (2026)
CPU-Limits kill Performance: Time to rethink Resource Control
by: Shetty, Chirag, et al.
Published: (2025)
by: Shetty, Chirag, et al.
Published: (2025)
THEMIS: Time, Heterogeneity, and Energy Minded Scheduling for Fair Multi-Tenant Use in FPGAs
by: Karabulut, Emre, et al.
Published: (2024)
by: Karabulut, Emre, et al.
Published: (2024)
Mewz: Lightweight Execution Environment for WebAssembly with High Isolation and Portability using Unikernels
by: Ueda, Soichiro, et al.
Published: (2024)
by: Ueda, Soichiro, et al.
Published: (2024)
Optimizing Task Scheduling in Heterogeneous Computing Environments: A Comparative Analysis of CPU, GPU, and ASIC Platforms Using E2C Simulator
by: Mohammadjafari, Ali, et al.
Published: (2024)
by: Mohammadjafari, Ali, et al.
Published: (2024)
Telepathic Datacenters: Fast RPCs using Shared CXL Memory
by: Mahar, Suyash, et al.
Published: (2024)
by: Mahar, Suyash, et al.
Published: (2024)
Agent Centric Operating System -- a Comprehensive Review and Outlook for Operating System
by: Jia, Shian, et al.
Published: (2024)
by: Jia, Shian, et al.
Published: (2024)
PhoenixOS: Concurrent OS-level GPU Checkpoint and Restore with Validated Speculation
by: Wei, Xingda, et al.
Published: (2024)
by: Wei, Xingda, et al.
Published: (2024)
BLITZSCALE: Fast and Live Large Model Autoscaling with O(1) Host Caching
by: Zhang, Dingyan, et al.
Published: (2024)
by: Zhang, Dingyan, et al.
Published: (2024)
Mercury: QoS-Aware Tiered Memory System
by: Lu, Jiaheng, et al.
Published: (2024)
by: Lu, Jiaheng, et al.
Published: (2024)
Zero-consistency root emulation for unprivileged container image build
by: Priedhorsky, Reid, et al.
Published: (2024)
by: Priedhorsky, Reid, et al.
Published: (2024)
Enabling performance portability of data-parallel OpenMP applications on asymmetric multicore processors
by: Saez, Juan Carlos, et al.
Published: (2024)
by: Saez, Juan Carlos, et al.
Published: (2024)
Skip TLB flushes for reused pages within mmap's
by: Schimmelpfennig, Frederic, et al.
Published: (2024)
by: Schimmelpfennig, Frederic, et al.
Published: (2024)
FALCON: Pinpointing and Mitigating Stragglers for Large-Scale Hybrid-Parallel Training
by: Wu, Tianyuan, et al.
Published: (2024)
by: Wu, Tianyuan, et al.
Published: (2024)
Mitigating context switching in densely packed Linux clusters with Latency-Aware Group Scheduling
by: Isstaif, Al Amjad Tawfiq, et al.
Published: (2025)
by: Isstaif, Al Amjad Tawfiq, et al.
Published: (2025)
DPC: A Distributed Page Cache over CXL
by: Bergman, Shai, et al.
Published: (2026)
by: Bergman, Shai, et al.
Published: (2026)
EdgeFlow: Fast Cold Starts for LLMs on Mobile Devices
by: Yan, Yongsheng, et al.
Published: (2026)
by: Yan, Yongsheng, et al.
Published: (2026)
A Periodic Space of Distributed Computing: Vision & Framework
by: Salehi, Mohsen Amini, et al.
Published: (2026)
by: Salehi, Mohsen Amini, et al.
Published: (2026)
Fix: externalizing network I/O in serverless computing
by: Deng, Yuhan, et al.
Published: (2025)
by: Deng, Yuhan, et al.
Published: (2025)
RAGDoll: Efficient Offloading-based Online RAG System on a Single GPU
by: Yu, Weiping, et al.
Published: (2025)
by: Yu, Weiping, et al.
Published: (2025)
Equilibria: Fair Multi-Tenant CXL Memory Tiering At Scale
by: Zhao, Kaiyang, et al.
Published: (2026)
by: Zhao, Kaiyang, et al.
Published: (2026)
Similar Items
-
Dirigent: Lightweight Serverless Orchestration
by: Cvetković, Lazar, et al.
Published: (2024) -
Imaginary Machines: A Serverless Model for Cloud Applications
by: Wawrzoniak, Michael, et al.
Published: (2024) -
Taming Serverless Cold Starts Through OS Co-Design
by: Holmes, Ben, et al.
Published: (2025) -
Nexus: Transparent I/O Offloading for High-Density Serverless Computing
by: Park, JooYoung, et al.
Published: (2026) -
TrEnv-X: Transparently Share Serverless Execution Environments Across Different Functions and Nodes
by: Huang, Jialiang, et al.
Published: (2025)