Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Tariq, Syed Salauddin Mohammad, Zein, Ali Al, Vaidya, Soumya Sripad, Khanolkar, Arati, Song, Zheng, Roy, Probir |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2024)
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2024)
Cold-Start Anti-Patterns and Refactorings in Serverless Systems: An Empirical Study
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2025)
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2025)
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
par: Nguyen, Chanh, et autres
Publié: (2025)
par: Nguyen, Chanh, et autres
Publié: (2025)
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
par: Akbari, Saman, et autres
Publié: (2025)
par: Akbari, Saman, et autres
Publié: (2025)
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
par: Sun, Bowen, et autres
Publié: (2026)
par: Sun, Bowen, et autres
Publié: (2026)
Serverless Cold Starts and Where to Find Them
par: Joosen, Artjom, et autres
Publié: (2024)
par: Joosen, Artjom, et autres
Publié: (2024)
Hiku: Pull-Based Scheduling for Serverless Computing
par: Akbari, Saman, et autres
Publié: (2025)
par: Akbari, Saman, et autres
Publié: (2025)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
par: Besozzi, Valerio, et autres
Publié: (2025)
par: Besozzi, Valerio, et autres
Publié: (2025)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
par: Qi, S., et autres
Publié: (2024)
par: Qi, S., et autres
Publié: (2024)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
par: Lin, Changyuan, et autres
Publié: (2025)
par: Lin, Changyuan, et autres
Publié: (2025)
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
par: Taylor, Maya, et autres
Publié: (2026)
par: Taylor, Maya, et autres
Publié: (2026)
Extrae.jl: Julia bindings for the Extrae HPC Profiler
par: Sanchez-Ramirez, Sergio, et autres
Publié: (2025)
par: Sanchez-Ramirez, Sergio, et autres
Publié: (2025)
Temporal Load Imbalance on Ondes3D Seismic Simulator for Different Multicore Architectures
par: Solórzano, Ana Luisa Veroneze, et autres
Publié: (2024)
par: Solórzano, Ana Luisa Veroneze, et autres
Publié: (2024)
AARC: Automated Affinity-aware Resource Configuration for Serverless Workflows
par: Jin, Lingxiao, et autres
Publié: (2025)
par: Jin, Lingxiao, et autres
Publié: (2025)
Profiling and optimization of multi-card GPU machine learning jobs
par: Lawenda, Marcin, et autres
Publié: (2025)
par: Lawenda, Marcin, et autres
Publié: (2025)
CUTHERMO: Understanding GPU Memory Inefficiencies with Heat Map Profiling
par: Zhao, Yanbo, et autres
Publié: (2025)
par: Zhao, Yanbo, et autres
Publié: (2025)
Reducing Tail Latencies Through Environment- and Neighbour-aware Thread Management
par: Jeffery, Andrew, et autres
Publié: (2024)
par: Jeffery, Andrew, et autres
Publié: (2024)
TaxBreak: Unmasking the Hidden Costs of LLM Inference Through Overhead Decomposition
par: Vellaisamy, Prabhu, et autres
Publié: (2026)
par: Vellaisamy, Prabhu, et autres
Publié: (2026)
Data-Driven Analysis to Understand GPU Hardware Resource Usage of Optimizations
par: Islam, Tanzima Z., et autres
Publié: (2024)
par: Islam, Tanzima Z., et autres
Publié: (2024)
HydraServe: Minimizing Cold Start Latency for Serverless LLM Serving in Public Clouds
par: Lou, Chiheng, et autres
Publié: (2025)
par: Lou, Chiheng, et autres
Publié: (2025)
Cold Start Latency in Serverless Computing: A Systematic Review, Taxonomy, and Future Directions
par: Golec, Muhammed, et autres
Publié: (2023)
par: Golec, Muhammed, et autres
Publié: (2023)
Characterizing WebGPU Dispatch Overhead for LLM Inference Across Four GPU Vendors, Three Backends, and Three Browsers
par: Maczan, Jędrzej
Publié: (2026)
par: Maczan, Jędrzej
Publié: (2026)
Orthrus: Accelerating Multi-BFT Consensus through Concurrent Partial Ordering of Transactions (Extended Version)
par: Lyu, Hanzheng, et autres
Publié: (2024)
par: Lyu, Hanzheng, et autres
Publié: (2024)
Taming Serverless Cold Starts Through OS Co-Design
par: Holmes, Ben, et autres
Publié: (2025)
par: Holmes, Ben, et autres
Publié: (2025)
EdgeProfiler: A Fast Profiling Framework for Lightweight LLMs on Edge Using Analytical Model
par: Pinnock, Alyssa, et autres
Publié: (2025)
par: Pinnock, Alyssa, et autres
Publié: (2025)
MPI Implementation Profiling for Better Application Performance
par: Shipley, Riley, et autres
Publié: (2024)
par: Shipley, Riley, et autres
Publié: (2024)
Compiler Support for Speculation in Decoupled Access/Execute Architectures
par: Szafarczyk, Robert, et autres
Publié: (2025)
par: Szafarczyk, Robert, et autres
Publié: (2025)
Extracting Practical, Actionable Energy Insights from Supercomputer Telemetry and Logs
par: Cornelius, Melanie, et autres
Publié: (2025)
par: Cornelius, Melanie, et autres
Publié: (2025)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
par: Zhuang, Chen, et autres
Publié: (2024)
par: Zhuang, Chen, et autres
Publié: (2024)
Optimal Parallel Scheduling under Concave Speedup Functions
par: Li, Chengzhang, et autres
Publié: (2025)
par: Li, Chengzhang, et autres
Publié: (2025)
Efficient GPU-Centered Singular Value Decomposition Using the Divide-and-Conquer Method
par: Liu, Shifang, et autres
Publié: (2025)
par: Liu, Shifang, et autres
Publié: (2025)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
par: Mao, Ying, et autres
Publié: (2020)
par: Mao, Ying, et autres
Publié: (2020)
Staging Blocked Evaluation over Structured Sparse Matrices
par: Das, Pratyush, et autres
Publié: (2024)
par: Das, Pratyush, et autres
Publié: (2024)
Cloud Performance Decomposition for Long-Term Performance Engineering: A Case Study
par: Debnath, Shimul, et autres
Publié: (2026)
par: Debnath, Shimul, et autres
Publié: (2026)
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
par: Ng, Nathan, et autres
Publié: (2026)
par: Ng, Nathan, et autres
Publié: (2026)
Preliminary report: Initial evaluation of StdPar implementations on AMD GPUs for HPC
par: Lin, Wei-Chen, et autres
Publié: (2024)
par: Lin, Wei-Chen, et autres
Publié: (2024)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
par: Sun, Tingyang, et autres
Publié: (2026)
par: Sun, Tingyang, et autres
Publié: (2026)
Dissecting the software-based measurement of CPU energy consumption: a comparative analysis
par: Raffin, Guillaume, et autres
Publié: (2024)
par: Raffin, Guillaume, et autres
Publié: (2024)
Bridding OT and PaaS in Edge-to-Cloud Continuum
par: Barrios, Carlos J, et autres
Publié: (2025)
par: Barrios, Carlos J, et autres
Publié: (2025)
RAPID-LLM: Resilience-Aware Performance analysis of Infrastructure for Distributed LLM Training and Inference
par: Karfakis, George, et autres
Publié: (2025)
par: Karfakis, George, et autres
Publié: (2025)
Documents similaires
-
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2024) -
Cold-Start Anti-Patterns and Refactorings in Serverless Systems: An Empirical Study
par: Tariq, Syed Salauddin Mohammad, et autres
Publié: (2025) -
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
par: Nguyen, Chanh, et autres
Publié: (2025) -
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
par: Akbari, Saman, et autres
Publié: (2025) -
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
par: Sun, Bowen, et autres
Publié: (2026)