Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
Fuente:
arXiv
Guardado en:
| Autores principales: | Sun, Bowen, Antonopoulos, Christos D., Smirni, Evgenia, Ren, Bin, Bellas, Nikolaos, Lalis, Spyros |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
por: Akbari, Saman, et al.
Publicado: (2025)
por: Akbari, Saman, et al.
Publicado: (2025)
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
por: Nguyen, Chanh, et al.
Publicado: (2025)
por: Nguyen, Chanh, et al.
Publicado: (2025)
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2025)
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2025)
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2024)
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2024)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
por: Qi, S., et al.
Publicado: (2024)
por: Qi, S., et al.
Publicado: (2024)
The Energy Cost of Execution-Idle in GPU Clusters
por: Lei, Yiran, et al.
Publicado: (2026)
por: Lei, Yiran, et al.
Publicado: (2026)
Hiku: Pull-Based Scheduling for Serverless Computing
por: Akbari, Saman, et al.
Publicado: (2025)
por: Akbari, Saman, et al.
Publicado: (2025)
Serverless Cold Starts and Where to Find Them
por: Joosen, Artjom, et al.
Publicado: (2024)
por: Joosen, Artjom, et al.
Publicado: (2024)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
por: Besozzi, Valerio, et al.
Publicado: (2025)
por: Besozzi, Valerio, et al.
Publicado: (2025)
Cold Start Latency in Serverless Computing: A Systematic Review, Taxonomy, and Future Directions
por: Golec, Muhammed, et al.
Publicado: (2023)
por: Golec, Muhammed, et al.
Publicado: (2023)
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
por: Taylor, Maya, et al.
Publicado: (2026)
por: Taylor, Maya, et al.
Publicado: (2026)
AARC: Automated Affinity-aware Resource Configuration for Serverless Workflows
por: Jin, Lingxiao, et al.
Publicado: (2025)
por: Jin, Lingxiao, et al.
Publicado: (2025)
Cold-Start Anti-Patterns and Refactorings in Serverless Systems: An Empirical Study
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2025)
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2025)
CGSim: A Simulation Framework for Large Scale Distributed Computing Environment
por: Vatsavai, Sairam Sri, et al.
Publicado: (2025)
por: Vatsavai, Sairam Sri, et al.
Publicado: (2025)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
por: Lin, Changyuan, et al.
Publicado: (2025)
por: Lin, Changyuan, et al.
Publicado: (2025)
Fast and Scalable Mixed Precision Euclidean Distance Calculations Using GPU Tensor Cores
por: Curless, Brian, et al.
Publicado: (2025)
por: Curless, Brian, et al.
Publicado: (2025)
Energy-Aware Computing in the Year 2026
por: Tchakoute, Roblex Nana, et al.
Publicado: (2026)
por: Tchakoute, Roblex Nana, et al.
Publicado: (2026)
Optimal Configuration of API Resources in Cloud Native Computing
por: Truyen, Eddy, et al.
Publicado: (2025)
por: Truyen, Eddy, et al.
Publicado: (2025)
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
por: Kulkarni, Apurv Deepak, et al.
Publicado: (2025)
por: Kulkarni, Apurv Deepak, et al.
Publicado: (2025)
Automated Calibration of Parallel and Distributed Computing Simulators: A Case Study
por: McDonald, Jesse, et al.
Publicado: (2024)
por: McDonald, Jesse, et al.
Publicado: (2024)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
por: Ramesh, Risshab Srinivas
Publicado: (2024)
por: Ramesh, Risshab Srinivas
Publicado: (2024)
Optimizations on Graph-Level for Domain Specific Computations in Julia and Application to QED
por: Reinhard, Anton, et al.
Publicado: (2025)
por: Reinhard, Anton, et al.
Publicado: (2025)
Modeling the Effect of Data Redundancy on Speedup in MLFMA Near-Field Computation
por: Sadeghi, Morteza
Publicado: (2025)
por: Sadeghi, Morteza
Publicado: (2025)
ADELIA: Automatic Differentiation for Efficient Laplace Inference Approximations
por: Boudaoud, Afif, et al.
Publicado: (2026)
por: Boudaoud, Afif, et al.
Publicado: (2026)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
por: Mao, Ying, et al.
Publicado: (2020)
por: Mao, Ying, et al.
Publicado: (2020)
Cost-Performance Evaluation of General Compute Instances: AWS, Azure, GCP, and OCI
por: Tharwani, Jay, et al.
Publicado: (2024)
por: Tharwani, Jay, et al.
Publicado: (2024)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
por: Lin, Mao, et al.
Publicado: (2026)
por: Lin, Mao, et al.
Publicado: (2026)
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
por: Andersson, Måns I., et al.
Publicado: (2025)
por: Andersson, Måns I., et al.
Publicado: (2025)
DREAMS: Decentralized Resource Allocation and Service Management across the Compute Continuum Using Service Affinity
por: Dinh-Tuan, Hai, et al.
Publicado: (2025)
por: Dinh-Tuan, Hai, et al.
Publicado: (2025)
The SAP Cloud Infrastructure Dataset: A Reality Check of Scheduling and Placement of VMs in Cloud Computing
por: Uhlig, Arno, et al.
Publicado: (2025)
por: Uhlig, Arno, et al.
Publicado: (2025)
Modeling the Impact of Fiber Latency on Compute-Communication Overlap in Geo-Distributed Multi-Datacenter AI Training
por: Papavasileiou, Ioannis, et al.
Publicado: (2026)
por: Papavasileiou, Ioannis, et al.
Publicado: (2026)
mLR: Scalable Laminography Reconstruction based on Memoization
por: Ma, Bin, et al.
Publicado: (2025)
por: Ma, Bin, et al.
Publicado: (2025)
A Multi-Port Concurrent Communication Model for handling Compute Intensive Tasks on Distributed Satellite System Constellations
por: Veeravalli, Bharadwaj
Publicado: (2026)
por: Veeravalli, Bharadwaj
Publicado: (2026)
HydraServe: Minimizing Cold Start Latency for Serverless LLM Serving in Public Clouds
por: Lou, Chiheng, et al.
Publicado: (2025)
por: Lou, Chiheng, et al.
Publicado: (2025)
Opt4GPTQ: Co-Optimizing Memory and Computation for 4-bit GPTQ Quantized LLM Inference on Heterogeneous Platforms
por: Zhang, Yaozheng, et al.
Publicado: (2025)
por: Zhang, Yaozheng, et al.
Publicado: (2025)
Synthesizing Proxy Applications for MPI Programs
por: Luo, Jiyu, et al.
Publicado: (2023)
por: Luo, Jiyu, et al.
Publicado: (2023)
HeteGen: Heterogeneous Parallel Inference for Large Language Models on Resource-Constrained Devices
por: Zhao, Xuanlei, et al.
Publicado: (2024)
por: Zhao, Xuanlei, et al.
Publicado: (2024)
MP-SL: Multihop Parallel Split Learning
por: Tirana, Joana, et al.
Publicado: (2024)
por: Tirana, Joana, et al.
Publicado: (2024)
Learning-Augmented Performance Model for Tensor Product Factorization in High-Order FEM
por: Ren, Xuanzhengbo, et al.
Publicado: (2026)
por: Ren, Xuanzhengbo, et al.
Publicado: (2026)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
por: Sun, Tingyang, et al.
Publicado: (2026)
por: Sun, Tingyang, et al.
Publicado: (2026)
Ejemplares similares
-
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
por: Akbari, Saman, et al.
Publicado: (2025) -
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
por: Nguyen, Chanh, et al.
Publicado: (2025) -
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2025) -
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
por: Tariq, Syed Salauddin Mohammad, et al.
Publicado: (2024) -
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
por: Qi, S., et al.
Publicado: (2024)