Exploiting Spot Instances for Time-Critical Cloud Workloads Using Optimal Randomized Strategies
Fuente:
arXiv
Saved in:
| Main Authors: | Bhuyan, Neelkamal, Bhatia, Randeep, Kodialam, Murali, Lakshman, TV |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Opportunistic Scheduling for Optimal Spot Instance Savings in the Cloud
by: Bhuyan, Neelkamal, et al.
Published: (2026)
by: Bhuyan, Neelkamal, et al.
Published: (2026)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026)
by: Li, Zhifei, et al.
Published: (2026)
Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
Directives for Function Offloading in 5G Networks Based on a Performance Characteristics Analysis
by: Dettinger, Falk, et al.
Published: (2025)
by: Dettinger, Falk, et al.
Published: (2025)
Swing: Short-cutting Rings for Higher Bandwidth Allreduce
by: De Sensi, Daniele, et al.
Published: (2024)
by: De Sensi, Daniele, et al.
Published: (2024)
LLAMP: Assessing Network Latency Tolerance of HPC Applications with Linear Programming
by: Shen, Siyuan, et al.
Published: (2024)
by: Shen, Siyuan, et al.
Published: (2024)
Exploring GPU-to-GPU Communication: Insights into Supercomputer Interconnects
by: De Sensi, Daniele, et al.
Published: (2024)
by: De Sensi, Daniele, et al.
Published: (2024)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
by: Talluri, Sacheendra, et al.
Published: (2025)
by: Talluri, Sacheendra, et al.
Published: (2025)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
by: Erben, Alexander, et al.
Published: (2023)
by: Erben, Alexander, et al.
Published: (2023)
Moonshot: Optimizing Chain-Based Rotating Leader BFT via Optimistic Proposals
by: Doidge, Isaac, et al.
Published: (2024)
by: Doidge, Isaac, et al.
Published: (2024)
Cognitive Infrastructure: A Unified DCIM Framework for AI Data Centers
by: Sunkara, Krishna Chaitanya
Published: (2026)
by: Sunkara, Krishna Chaitanya
Published: (2026)
Evaluating Emerging AI/ML Accelerators: IPU, RDU, and NVIDIA/AMD GPUs
by: Peng, Hongwu, et al.
Published: (2023)
by: Peng, Hongwu, et al.
Published: (2023)
Tetris: An SLA-aware Application Placement Strategy in the Edge-Cloud Continuum
by: Almeida, Lucas, et al.
Published: (2025)
by: Almeida, Lucas, et al.
Published: (2025)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
by: Xie, Tianfang
Published: (2026)
by: Xie, Tianfang
Published: (2026)
Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters
by: Da, Wei, et al.
Published: (2025)
by: Da, Wei, et al.
Published: (2025)
Bine Trees: Enhancing Collective Operations by Optimizing Communication Locality
by: De Sensi, Daniele, et al.
Published: (2025)
by: De Sensi, Daniele, et al.
Published: (2025)
FCDP: Fully Cached Data Parallel for Communication-Avoiding Large-Scale Training
by: Park, Gyeongseo, et al.
Published: (2026)
by: Park, Gyeongseo, et al.
Published: (2026)
eScope: A Fine-Grained Power Prediction Mechanism for Mobile Applications
by: Mukherjee, Dipayan, et al.
Published: (2024)
by: Mukherjee, Dipayan, et al.
Published: (2024)
FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines
by: He, Jiaao, et al.
Published: (2024)
by: He, Jiaao, et al.
Published: (2024)
Comprehensive Plugin-Based Monitoring of Nexflow Workflow Executions
by: Kharma, Sami, et al.
Published: (2026)
by: Kharma, Sami, et al.
Published: (2026)
Nezha: Deployable and High-Performance Consensus Using Synchronized Clocks
by: Geng, Jinkun, et al.
Published: (2022)
by: Geng, Jinkun, et al.
Published: (2022)
Security Analysis of Bitcoin's V2 Transport Protocol: Exploiting Design Implications for Sustained Eclipse and Downgrade Attacks
by: Ndolo, Charmaine, et al.
Published: (2026)
by: Ndolo, Charmaine, et al.
Published: (2026)
A Survey on Heterogeneous Computing Using SmartNICs and Emerging Data Processing Units
by: Tibbetts, Nathan, et al.
Published: (2025)
by: Tibbetts, Nathan, et al.
Published: (2025)
Konnektor: Connection Protocol for Ensuring Peer Uniqueness in Decentralized P2P Networks
by: Ozkan, Onur
Published: (2024)
by: Ozkan, Onur
Published: (2024)
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)
by: Leung, Derek, et al.
Published: (2025)
Distributed Recoverable Sketches (Extended Version)
by: Cohen, Diana, et al.
Published: (2025)
by: Cohen, Diana, et al.
Published: (2025)
Intersections of Web3 and AI -- View in 2024
by: Hyland-Wood, David, et al.
Published: (2024)
by: Hyland-Wood, David, et al.
Published: (2024)
Artifact Evaluation for Distributed Systems: Current Practices and Beyond
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
NotebookOS: A Replicated Notebook Platform for Interactive Training with On-Demand GPUs
by: Carver, Benjamin, et al.
Published: (2025)
by: Carver, Benjamin, et al.
Published: (2025)
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
Generic Multicast (Extended Version)
by: Bolina, José Augusto, et al.
Published: (2024)
by: Bolina, José Augusto, et al.
Published: (2024)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
by: Yuan, Renzhong, et al.
Published: (2026)
by: Yuan, Renzhong, et al.
Published: (2026)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
by: Guan, Bo
Published: (2026)
by: Guan, Bo
Published: (2026)
OPTIMUMP2P: Fast and Reliable Gossiping in P2P Networks
by: Nicolaou, Nicolas, et al.
Published: (2025)
by: Nicolaou, Nicolas, et al.
Published: (2025)
Rotary GPU: Exploring Local Execution Paths for Large Mixture-of-Experts Models Under Limited GPU Memory
by: Jo, Myeong Jun
Published: (2026)
by: Jo, Myeong Jun
Published: (2026)
Chameleon: Adaptive Caching and Scheduling for Many-Adapter LLM Inference Environments
by: Iliakopoulou, Nikoleta, et al.
Published: (2024)
by: Iliakopoulou, Nikoleta, et al.
Published: (2024)
T3: Transparent Tracking & Triggering for Fine-grained Overlap of Compute & Collectives
by: Pati, Suchita, et al.
Published: (2024)
by: Pati, Suchita, et al.
Published: (2024)
NimbusGuard: A Novel Framework for Proactive Kubernetes Autoscaling Using Deep Q-Networks
by: Wanigasooriya, Chamath, et al.
Published: (2026)
by: Wanigasooriya, Chamath, et al.
Published: (2026)
dpBento: Benchmarking DPUs for Data Processing
by: Hu, Jiasheng, et al.
Published: (2025)
by: Hu, Jiasheng, et al.
Published: (2025)
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026)
by: Bai, Tianyu, et al.
Published: (2026)
Similar Items
-
Opportunistic Scheduling for Optimal Spot Instance Savings in the Cloud
by: Bhuyan, Neelkamal, et al.
Published: (2026) -
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026) -
Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads
by: Stokely, Murray, et al.
Published: (2025) -
Directives for Function Offloading in 5G Networks Based on a Performance Characteristics Analysis
by: Dettinger, Falk, et al.
Published: (2025) -
Swing: Short-cutting Rings for Higher Bandwidth Allreduce
by: De Sensi, Daniele, et al.
Published: (2024)