SpotKube: Cost-Optimal Microservices Deployment with Cluster Autoscaling and Spot Pricing
Fuente:
arXiv
Saved in:
| Main Authors: | Edirisinghe, Dasith, Rajapakse, Kavinda, Abeysinghe, Pasindu, Rathnayake, Sunimal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
by: Kim, Taeyoon, et al.
Published: (2026)
by: Kim, Taeyoon, et al.
Published: (2026)
Self-adaptive, Requirements-driven Autoscaling of Microservices
by: Nunes, João Paulo Karol Santos, et al.
Published: (2024)
by: Nunes, João Paulo Karol Santos, et al.
Published: (2024)
ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
by: Bai, Haoyu, et al.
Published: (2026)
by: Bai, Haoyu, et al.
Published: (2026)
SpotVista: Availability-Aware Recommendation System for Reliable and Cost-Efficient Multi-Node Spot Instances
by: Kim, Taeyoon, et al.
Published: (2026)
by: Kim, Taeyoon, et al.
Published: (2026)
DeepVM: Integrating Spot and On-Demand VMs for Cost-Efficient Deep Learning Clusters in the Cloud
by: Kim, Yoochan, et al.
Published: (2024)
by: Kim, Yoochan, et al.
Published: (2024)
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
by: Hu, Kan, et al.
Published: (2024)
by: Hu, Kan, et al.
Published: (2024)
HGraphScale: Hierarchical Graph Learning for Autoscaling Microservice Applications in Container-based Cloud Computing
by: Fang, Zhengxin, et al.
Published: (2025)
by: Fang, Zhengxin, et al.
Published: (2025)
Carbon-Aware Microservice Deployment for Optimal User Experience on a Budget
by: Kreutz, Kevin, et al.
Published: (2025)
by: Kreutz, Kevin, et al.
Published: (2025)
KubeDSM: A Kubernetes-based Dynamic Scheduling and Migration Framework for Cloud-Assisted Edge Clusters
by: Pashaeehir, Amirhossein, et al.
Published: (2025)
by: Pashaeehir, Amirhossein, et al.
Published: (2025)
GFS: A Preemption-aware Scheduling Framework for GPU Clusters with Predictive Spot Instance Management
by: Duan, Jiaang, et al.
Published: (2025)
by: Duan, Jiaang, et al.
Published: (2025)
Ding-Dong Ditch: Peeking Into Spot Instance Availability
by: Kim, Kyumin, et al.
Published: (2026)
by: Kim, Kyumin, et al.
Published: (2026)
An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
by: Punniyamoorthy, Vinoth, et al.
Published: (2025)
by: Punniyamoorthy, Vinoth, et al.
Published: (2025)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
by: Ponce, Francisco, et al.
Published: (2026)
by: Ponce, Francisco, et al.
Published: (2026)
Proactive and Reactive Autoscaling Techniques for Edge Computing
by: Gupta, Suhrid, et al.
Published: (2025)
by: Gupta, Suhrid, et al.
Published: (2025)
AI-Driven Multi-Region Provisioning for Cloud Services Using Spot Fleets
by: Fabra, Javier, et al.
Published: (2026)
by: Fabra, Javier, et al.
Published: (2026)
KubeIntellect: A Modular LLM-Orchestrated Agent Framework for End-to-End Kubernetes Management
by: Ardebili, Mohsen Seyedkazemi, et al.
Published: (2025)
by: Ardebili, Mohsen Seyedkazemi, et al.
Published: (2025)
ENOVA: Autoscaling towards Cost-effective and Stable Serverless LLM Serving
by: Huang, Tao, et al.
Published: (2024)
by: Huang, Tao, et al.
Published: (2024)
Mitigating Interference of Microservices with a Scoring Mechanism in Large-scale Clusters
by: Yang, Dingyu, et al.
Published: (2024)
by: Yang, Dingyu, et al.
Published: (2024)
Visualizing Cloud-native Applications with KubeDiagrams
by: Merle, Philippe, et al.
Published: (2025)
by: Merle, Philippe, et al.
Published: (2025)
HuntMS: A Framework for Microservice Geo-Distribution for Carbon and Cost Reduction
by: Christofidi, Georgia, et al.
Published: (2026)
by: Christofidi, Georgia, et al.
Published: (2026)
Multi-Objective Optimization of Consumer Group Autoscaling in Message Broker Systems
by: Landau, Diogo, et al.
Published: (2024)
by: Landau, Diogo, et al.
Published: (2024)
TokenScale: Timely and Accurate Autoscaling for Disaggregated LLM Serving with Token Velocity
by: Lai, Ruiqi, et al.
Published: (2025)
by: Lai, Ruiqi, et al.
Published: (2025)
C-Koordinator: Interference-aware Management for Large-scale and Co-located Microservice Clusters
by: Song, Shengye, et al.
Published: (2025)
by: Song, Shengye, et al.
Published: (2025)
Daedalus: Self-Adaptive Horizontal Autoscaling for Resource Efficiency of Distributed Stream Processing Systems
by: Pfister, Benjamin J. J., et al.
Published: (2024)
by: Pfister, Benjamin J. J., et al.
Published: (2024)
Complexity at Scale: A Quantitative Analysis of an Alibaba Microservice Deployment
by: Winchester, Giles, et al.
Published: (2025)
by: Winchester, Giles, et al.
Published: (2025)
SAIR: Cost-Efficient Multi-Stage ML Pipeline Autoscaling via In-Context Reinforcement Learning
by: Su, Jianchang, et al.
Published: (2026)
by: Su, Jianchang, et al.
Published: (2026)
Shifting the Sweet Spot: High-Performance Matrix-Free Method for High-Order Elasticity
by: Chang, Dali, et al.
Published: (2026)
by: Chang, Dali, et al.
Published: (2026)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026)
by: Li, Zhifei, et al.
Published: (2026)
LA-IMR: Latency-Aware, Predictive In-Memory Routing and Proactive Autoscaling for Tail-Latency-Sensitive Cloud Robotics
by: Seo, Eunil, et al.
Published: (2025)
by: Seo, Eunil, et al.
Published: (2025)
Metric Criticality Identification for Cloud Microservices
by: Singal, Akanksha, et al.
Published: (2025)
by: Singal, Akanksha, et al.
Published: (2025)
NotNets: Accelerating Microservices by Bypassing the Network
by: Alvaro, Peter, et al.
Published: (2024)
by: Alvaro, Peter, et al.
Published: (2024)
Signalling Health for Improved Kubernetes Microservice Availability
by: Roberts, Jacob, et al.
Published: (2025)
by: Roberts, Jacob, et al.
Published: (2025)
Finding Nemo-Nemo: CFT DAG-based Consensus in the WAN
by: Kerur, Rithwik, et al.
Published: (2026)
by: Kerur, Rithwik, et al.
Published: (2026)
Contention-Aware Microservice Deployment in Collaborative Mobile Edge Networks
by: Ge, Xinlei, et al.
Published: (2024)
by: Ge, Xinlei, et al.
Published: (2024)
Hierarchical Autoscaling for Large Language Model Serving with Chiron
by: Patke, Archit, et al.
Published: (2025)
by: Patke, Archit, et al.
Published: (2025)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
by: Qi, S., et al.
Published: (2024)
by: Qi, S., et al.
Published: (2024)
Energy-aware Distributed Microservice Request Placement at the Edge
by: Toczé, Klervie, et al.
Published: (2024)
by: Toczé, Klervie, et al.
Published: (2024)
Energy Metrics for Edge Microservice Request Placement Strategies
by: Toczé, Klervie, et al.
Published: (2025)
by: Toczé, Klervie, et al.
Published: (2025)
The High Cost of Keeping Warm: Characterizing Overhead in Serverless Autoscaling Policies
by: Kondrashov, Leonid, et al.
Published: (2025)
by: Kondrashov, Leonid, et al.
Published: (2025)
Eva: Cost-Efficient Cloud-Based Cluster Scheduling
by: Chang, Tzu-Tao, et al.
Published: (2025)
by: Chang, Tzu-Tao, et al.
Published: (2025)
Similar Items
-
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
by: Kim, Taeyoon, et al.
Published: (2026) -
Self-adaptive, Requirements-driven Autoscaling of Microservices
by: Nunes, João Paulo Karol Santos, et al.
Published: (2024) -
ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
by: Bai, Haoyu, et al.
Published: (2026) -
SpotVista: Availability-Aware Recommendation System for Reliable and Cost-Efficient Multi-Node Spot Instances
by: Kim, Taeyoon, et al.
Published: (2026) -
DeepVM: Integrating Spot and On-Demand VMs for Cost-Efficient Deep Learning Clusters in the Cloud
by: Kim, Yoochan, et al.
Published: (2024)