CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, S., Moore, H., Hogade, N., Milojicic, D., Bash, C., Pasricha, S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Framework for SLO, Carbon, and Wastewater-Aware Sustainable FaaS Cloud Platform Management
von: Qi, Sirui, et al.
Veröffentlicht: (2024)
von: Qi, Sirui, et al.
Veröffentlicht: (2024)
Sustainable Carbon-Aware and Water-Efficient LLM Scheduling in Geo-Distributed Cloud Datacenters
von: Moore, Hayden, et al.
Veröffentlicht: (2025)
von: Moore, Hayden, et al.
Veröffentlicht: (2025)
MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters
von: Moore, H., et al.
Veröffentlicht: (2026)
von: Moore, H., et al.
Veröffentlicht: (2026)
Sustainable Graph Analytics Workload Scheduling with Evolutionary Reinforcement Learning in Edge-Cloud Systems
von: Ramicetty, P., et al.
Veröffentlicht: (2026)
von: Ramicetty, P., et al.
Veröffentlicht: (2026)
Hiku: Pull-Based Scheduling for Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
The SAP Cloud Infrastructure Dataset: A Reality Check of Scheduling and Placement of VMs in Cloud Computing
von: Uhlig, Arno, et al.
Veröffentlicht: (2025)
von: Uhlig, Arno, et al.
Veröffentlicht: (2025)
Universal Workers: A Vision for Eliminating Cold Starts in Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
von: Sun, Bowen, et al.
Veröffentlicht: (2026)
von: Sun, Bowen, et al.
Veröffentlicht: (2026)
An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
von: Punniyamoorthy, Vinoth, et al.
Veröffentlicht: (2025)
von: Punniyamoorthy, Vinoth, et al.
Veröffentlicht: (2025)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
Game-Theoretic Deep Reinforcement Learning to Minimize Carbon Emissions and Energy Costs for AI Inference Workloads in Geo-Distributed Data Centers
von: Hogade, Ninad, et al.
Veröffentlicht: (2024)
von: Hogade, Ninad, et al.
Veröffentlicht: (2024)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
von: Tariq, Syed Salauddin Mohammad, et al.
Veröffentlicht: (2025)
Optimal Configuration of API Resources in Cloud Native Computing
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
von: Muhammad, Said, et al.
Veröffentlicht: (2025)
von: Muhammad, Said, et al.
Veröffentlicht: (2025)
Energy-Aware Computing in the Year 2026
von: Tchakoute, Roblex Nana, et al.
Veröffentlicht: (2026)
von: Tchakoute, Roblex Nana, et al.
Veröffentlicht: (2026)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
von: Mao, Ying, et al.
Veröffentlicht: (2020)
von: Mao, Ying, et al.
Veröffentlicht: (2020)
EcoLife: Carbon-Aware Serverless Function Scheduling for Sustainable Computing
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
von: Hu, Kan, et al.
Veröffentlicht: (2024)
von: Hu, Kan, et al.
Veröffentlicht: (2024)
Optimal Parallel Scheduling under Concave Speedup Functions
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
Asymptotically Optimal Scheduling of Multiple Parallelizable Job Classes
von: Berg, Benjamin, et al.
Veröffentlicht: (2024)
von: Berg, Benjamin, et al.
Veröffentlicht: (2024)
CGSim: A Simulation Framework for Large Scale Distributed Computing Environment
von: Vatsavai, Sairam Sri, et al.
Veröffentlicht: (2025)
von: Vatsavai, Sairam Sri, et al.
Veröffentlicht: (2025)
Cloud Resource Allocation with Convex Optimization
von: Boghani, Shayan, et al.
Veröffentlicht: (2025)
von: Boghani, Shayan, et al.
Veröffentlicht: (2025)
Usability Evaluation of Cloud for HPC Applications
von: Sochat, Vanessa, et al.
Veröffentlicht: (2025)
von: Sochat, Vanessa, et al.
Veröffentlicht: (2025)
Operational Strategies for Non-Disruptive Scheduling Transitions in Production HPC Systems
von: MacLachlan, Glen, et al.
Veröffentlicht: (2026)
von: MacLachlan, Glen, et al.
Veröffentlicht: (2026)
Unleashing the Power of Preemptive Priority-based Scheduling for Real-Time GPU Tasks
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
Bridding OT and PaaS in Edge-to-Cloud Continuum
von: Barrios, Carlos J, et al.
Veröffentlicht: (2025)
von: Barrios, Carlos J, et al.
Veröffentlicht: (2025)
Kubernetes in Action: Exploring the Performance of Kubernetes Distributions in the Cloud
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
Flowshop Machine Scheduling: Markov Modeling, Optimal Schedules and Heuristics
von: Ghanem, Samah A. M.
Veröffentlicht: (2025)
von: Ghanem, Samah A. M.
Veröffentlicht: (2025)
Sampling in Cloud Benchmarking: A Critical Review and Methodological Guidelines
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
The Sunk Carbon Fallacy: Rethinking Carbon Footprint Metrics for Effective Carbon-Aware Scheduling
von: Bashir, Noman, et al.
Veröffentlicht: (2024)
von: Bashir, Noman, et al.
Veröffentlicht: (2024)
Evaluating HPC-Style CPU Performance and Cost in Virtualized Cloud Infrastructures
von: Tharwani, Jay, et al.
Veröffentlicht: (2025)
von: Tharwani, Jay, et al.
Veröffentlicht: (2025)
Is Intelligence the Right Direction in New OS Scheduling for Multiple Resources in Cloud Environments?
von: Dou, Xinglei, et al.
Veröffentlicht: (2025)
von: Dou, Xinglei, et al.
Veröffentlicht: (2025)
SLO-Aware Scheduling for Large Language Model Inferences
von: Huang, Jinqi, et al.
Veröffentlicht: (2025)
von: Huang, Jinqi, et al.
Veröffentlicht: (2025)
Cloud Performance Decomposition for Long-Term Performance Engineering: A Case Study
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
AARC: Automated Affinity-aware Resource Configuration for Serverless Workflows
von: Jin, Lingxiao, et al.
Veröffentlicht: (2025)
von: Jin, Lingxiao, et al.
Veröffentlicht: (2025)
Advanced Scheduling Strategies for Distributed Quantum Computing Jobs
von: Ni, Gongyu, et al.
Veröffentlicht: (2026)
von: Ni, Gongyu, et al.
Veröffentlicht: (2026)
Efficient Fault Localization in a Cloud Stack Using End-to-End Application Service Topology
von: Mathews, Dhanya R, et al.
Veröffentlicht: (2025)
von: Mathews, Dhanya R, et al.
Veröffentlicht: (2025)
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
von: Taylor, Maya, et al.
Veröffentlicht: (2026)
von: Taylor, Maya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Framework for SLO, Carbon, and Wastewater-Aware Sustainable FaaS Cloud Platform Management
von: Qi, Sirui, et al.
Veröffentlicht: (2024) -
Sustainable Carbon-Aware and Water-Efficient LLM Scheduling in Geo-Distributed Cloud Datacenters
von: Moore, Hayden, et al.
Veröffentlicht: (2025) -
MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters
von: Moore, H., et al.
Veröffentlicht: (2026) -
Sustainable Graph Analytics Workload Scheduling with Evolutionary Reinforcement Learning in Edge-Cloud Systems
von: Ramicetty, P., et al.
Veröffentlicht: (2026) -
Hiku: Pull-Based Scheduling for Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)