Serverless Cold Starts and Where to Find Them
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Joosen, Artjom, Hassan, Ahmed, Asenov, Martin, Singh, Rajkarn, Darlow, Luke, Wang, Jianfeng, Deng, Qiwen, Barker, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Taming Serverless Cold Starts Through OS Co-Design
von: Holmes, Ben, et al.
Veröffentlicht: (2025)
von: Holmes, Ben, et al.
Veröffentlicht: (2025)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
Shattering the Ephemeral Storage Cost Barrier for Data-Intensive Serverless Workflows
von: Ustiugov, Dmitrii, et al.
Veröffentlicht: (2023)
von: Ustiugov, Dmitrii, et al.
Veröffentlicht: (2023)
Nanvix: A Multikernel OS Design for High-Density Serverless Deployments
von: Segarra, Carlos, et al.
Veröffentlicht: (2026)
von: Segarra, Carlos, et al.
Veröffentlicht: (2026)
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
Chameleon: Adaptive Caching and Scheduling for Many-Adapter LLM Inference Environments
von: Iliakopoulou, Nikoleta, et al.
Veröffentlicht: (2024)
von: Iliakopoulou, Nikoleta, et al.
Veröffentlicht: (2024)
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
von: Sun, Bowen, et al.
Veröffentlicht: (2026)
von: Sun, Bowen, et al.
Veröffentlicht: (2026)
Reexamining Paradigms of End-to-End Data Movement
von: Fang, Chin, et al.
Veröffentlicht: (2025)
von: Fang, Chin, et al.
Veröffentlicht: (2025)
Optimal Configuration of API Resources in Cloud Native Computing
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2024)
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2024)
Mitigating GIL Bottlenecks in Edge AI Systems
von: Mandal, Mridankan, et al.
Veröffentlicht: (2026)
von: Mandal, Mridankan, et al.
Veröffentlicht: (2026)
RAID Organizations for Improved Reliability and Performance: A Not Entirely Unbiased Tutorial (1st revision)
von: Thomasian, Alexander
Veröffentlicht: (2024)
von: Thomasian, Alexander
Veröffentlicht: (2024)
SwitchFS: Asynchronous Metadata Updates for Distributed Filesystems with In-Network Coordination
von: Xu, Jingwei, et al.
Veröffentlicht: (2024)
von: Xu, Jingwei, et al.
Veröffentlicht: (2024)
CPU-Limits kill Performance: Time to rethink Resource Control
von: Shetty, Chirag, et al.
Veröffentlicht: (2025)
von: Shetty, Chirag, et al.
Veröffentlicht: (2025)
Optimizing CPU Cache Utilization in Cloud VMs with Accurate Cache Abstraction
von: Tofigh, Mani, et al.
Veröffentlicht: (2025)
von: Tofigh, Mani, et al.
Veröffentlicht: (2025)
GPUVM: GPU-driven Unified Virtual Memory
von: Nazaraliyev, Nurlan, et al.
Veröffentlicht: (2024)
von: Nazaraliyev, Nurlan, et al.
Veröffentlicht: (2024)
GridPilot: Real-Time Grid-Responsive Control for AI Supercomputers
von: Constantinescu, Denisa-Andreea, et al.
Veröffentlicht: (2026)
von: Constantinescu, Denisa-Andreea, et al.
Veröffentlicht: (2026)
eScope: A Fine-Grained Power Prediction Mechanism for Mobile Applications
von: Mukherjee, Dipayan, et al.
Veröffentlicht: (2024)
von: Mukherjee, Dipayan, et al.
Veröffentlicht: (2024)
FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines
von: He, Jiaao, et al.
Veröffentlicht: (2024)
von: He, Jiaao, et al.
Veröffentlicht: (2024)
Comprehensive Plugin-Based Monitoring of Nexflow Workflow Executions
von: Kharma, Sami, et al.
Veröffentlicht: (2026)
von: Kharma, Sami, et al.
Veröffentlicht: (2026)
EdgeFlow: Fast Cold Starts for LLMs on Mobile Devices
von: Yan, Yongsheng, et al.
Veröffentlicht: (2026)
von: Yan, Yongsheng, et al.
Veröffentlicht: (2026)
DPDPU: Data Processing with DPUs
von: Hu, Jiasheng, et al.
Veröffentlicht: (2024)
von: Hu, Jiasheng, et al.
Veröffentlicht: (2024)
Flex-MIG: Enabling Distributed Execution on MIG
von: Kim, Myeongsu, et al.
Veröffentlicht: (2025)
von: Kim, Myeongsu, et al.
Veröffentlicht: (2025)
Dirigent: Lightweight Serverless Orchestration
von: Cvetković, Lazar, et al.
Veröffentlicht: (2024)
von: Cvetković, Lazar, et al.
Veröffentlicht: (2024)
Flexible Swapping for the Cloud
von: Pandurov, Milan, et al.
Veröffentlicht: (2024)
von: Pandurov, Milan, et al.
Veröffentlicht: (2024)
Microsecond-scale Dynamic Validation of Idempotency for GPU Kernels
von: Han, Mingcong, et al.
Veröffentlicht: (2024)
von: Han, Mingcong, et al.
Veröffentlicht: (2024)
Hiku: Pull-Based Scheduling for Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Random Adaptive Cache Placement Policy
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
The Hitchhiker's Guide to Programming and Optimizing Cache Coherent Heterogeneous Systems: CXL, NVLink-C2C, and AMD Infinity Fabric
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
Intent-driven scheduling of backup jobs
von: Dutta, Souvik, et al.
Veröffentlicht: (2024)
von: Dutta, Souvik, et al.
Veröffentlicht: (2024)
Evaluating Serverless Machine Learning Performance on Google Cloud Run
von: Khatiwada, Prerana, et al.
Veröffentlicht: (2024)
von: Khatiwada, Prerana, et al.
Veröffentlicht: (2024)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
von: Qi, S., et al.
Veröffentlicht: (2024)
von: Qi, S., et al.
Veröffentlicht: (2024)
Nexus: Transparent I/O Offloading for High-Density Serverless Computing
von: Park, JooYoung, et al.
Veröffentlicht: (2026)
von: Park, JooYoung, et al.
Veröffentlicht: (2026)
Evaluating Emerging AI/ML Accelerators: IPU, RDU, and NVIDIA/AMD GPUs
von: Peng, Hongwu, et al.
Veröffentlicht: (2023)
von: Peng, Hongwu, et al.
Veröffentlicht: (2023)
Cold-Start Anti-Patterns and Refactorings in Serverless Systems: An Empirical Study
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
A Case for CATS: A Conductor-driven Asymmetric Transport Scheme for Semantic Prioritization
von: Rizvi, Syed Muhammad Aqdas
Veröffentlicht: (2026)
von: Rizvi, Syed Muhammad Aqdas
Veröffentlicht: (2026)
Ähnliche Einträge
-
Taming Serverless Cold Starts Through OS Co-Design
von: Holmes, Ben, et al.
Veröffentlicht: (2025) -
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
von: Lin, Changyuan, et al.
Veröffentlicht: (2025) -
Shattering the Ephemeral Storage Cost Barrier for Data-Intensive Serverless Workflows
von: Ustiugov, Dmitrii, et al.
Veröffentlicht: (2023) -
Nanvix: A Multikernel OS Design for High-Density Serverless Deployments
von: Segarra, Carlos, et al.
Veröffentlicht: (2026) -
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)