Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads
Fuente:
arXiv
Saved in:
| Main Authors: | Stokely, Murray, Nadgir, Neel, Peele, Jack, Kostakis, Orestis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
by: Xie, Tianfang
Published: (2026)
by: Xie, Tianfang
Published: (2026)
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026)
by: Bai, Tianyu, et al.
Published: (2026)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
by: Talluri, Sacheendra, et al.
Published: (2025)
by: Talluri, Sacheendra, et al.
Published: (2025)
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)
by: Leung, Derek, et al.
Published: (2025)
FCDP: Fully Cached Data Parallel for Communication-Avoiding Large-Scale Training
by: Park, Gyeongseo, et al.
Published: (2026)
by: Park, Gyeongseo, et al.
Published: (2026)
Commitment Against Front Running Attacks
by: Canidio, Andrea, et al.
Published: (2023)
by: Canidio, Andrea, et al.
Published: (2023)
Exploiting Spot Instances for Time-Critical Cloud Workloads Using Optimal Randomized Strategies
by: Bhuyan, Neelkamal, et al.
Published: (2026)
by: Bhuyan, Neelkamal, et al.
Published: (2026)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026)
by: Li, Zhifei, et al.
Published: (2026)
Distributed Recoverable Sketches (Extended Version)
by: Cohen, Diana, et al.
Published: (2025)
by: Cohen, Diana, et al.
Published: (2025)
Intersections of Web3 and AI -- View in 2024
by: Hyland-Wood, David, et al.
Published: (2024)
by: Hyland-Wood, David, et al.
Published: (2024)
Artifact Evaluation for Distributed Systems: Current Practices and Beyond
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
NotebookOS: A Replicated Notebook Platform for Interactive Training with On-Demand GPUs
by: Carver, Benjamin, et al.
Published: (2025)
by: Carver, Benjamin, et al.
Published: (2025)
Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters
by: Da, Wei, et al.
Published: (2025)
by: Da, Wei, et al.
Published: (2025)
Generic Multicast (Extended Version)
by: Bolina, José Augusto, et al.
Published: (2024)
by: Bolina, José Augusto, et al.
Published: (2024)
dpBento: Benchmarking DPUs for Data Processing
by: Hu, Jiasheng, et al.
Published: (2025)
by: Hu, Jiasheng, et al.
Published: (2025)
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
by: Pathak, Prashant Kumar, et al.
Published: (2026)
by: Pathak, Prashant Kumar, et al.
Published: (2026)
OPTIMUMP2P: Fast and Reliable Gossiping in P2P Networks
by: Nicolaou, Nicolas, et al.
Published: (2025)
by: Nicolaou, Nicolas, et al.
Published: (2025)
Studying the Effect of Schedule Preemption on Dynamic Task Graph Scheduling
by: Khodabandehlou, Mohammadali, et al.
Published: (2026)
by: Khodabandehlou, Mohammadali, et al.
Published: (2026)
GPUnion: Autonomous GPU Sharing on Campus
by: Li, Yufang, et al.
Published: (2025)
by: Li, Yufang, et al.
Published: (2025)
NimbusGuard: A Novel Framework for Proactive Kubernetes Autoscaling Using Deep Q-Networks
by: Wanigasooriya, Chamath, et al.
Published: (2026)
by: Wanigasooriya, Chamath, et al.
Published: (2026)
Social Dynamics of DAOs: Power, Onboarding, and Inclusivity
by: Kozlova, Victoria, et al.
Published: (2025)
by: Kozlova, Victoria, et al.
Published: (2025)
Knowledge Graphs-Driven Intelligence for Distributed Decision Systems
by: Napoli, Rosario, et al.
Published: (2026)
by: Napoli, Rosario, et al.
Published: (2026)
Advocate -- Trustworthy Evidence in Cloud Systems
by: Werner, Sebastian, et al.
Published: (2024)
by: Werner, Sebastian, et al.
Published: (2024)
Opportunistic Scheduling for Optimal Spot Instance Savings in the Cloud
by: Bhuyan, Neelkamal, et al.
Published: (2026)
by: Bhuyan, Neelkamal, et al.
Published: (2026)
Cognitive Infrastructure: A Unified DCIM Framework for AI Data Centers
by: Sunkara, Krishna Chaitanya
Published: (2026)
by: Sunkara, Krishna Chaitanya
Published: (2026)
EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM Inference
by: Sidik, Bronislav, et al.
Published: (2026)
by: Sidik, Bronislav, et al.
Published: (2026)
Alea-BFT: Practical Asynchronous Byzantine Fault Tolerance
by: Antunes, Diogo S., et al.
Published: (2024)
by: Antunes, Diogo S., et al.
Published: (2024)
Laminar: A Probe-First Scheduling Paradigm with Deterministic Runtime Survival
by: Chu, Zhengyan
Published: (2026)
by: Chu, Zhengyan
Published: (2026)
Trident: Adaptive Scheduling for Heterogeneous Multimodal Data Pipelines
by: Pan, Ding, et al.
Published: (2026)
by: Pan, Ding, et al.
Published: (2026)
Directives for Function Offloading in 5G Networks Based on a Performance Characteristics Analysis
by: Dettinger, Falk, et al.
Published: (2025)
by: Dettinger, Falk, et al.
Published: (2025)
Pioplat: A Scalable, Low-Cost Framework for Latency Reduction in Ethereum Blockchain
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
by: Murimi, Almond Kiruthu
Published: (2025)
by: Murimi, Almond Kiruthu
Published: (2025)
A Comprehensive Survey on Orbital Edge Computing: Systems, Applications, and Algorithms
by: Wu, Changhao, et al.
Published: (2023)
by: Wu, Changhao, et al.
Published: (2023)
A Preliminary Model of Coordination-free Consistency
by: Li, Shulu, et al.
Published: (2025)
by: Li, Shulu, et al.
Published: (2025)
Scalable Genomic Context Analysis with GCsnap2 on HPC Clusters
by: Krummenacher, Reto, et al.
Published: (2025)
by: Krummenacher, Reto, et al.
Published: (2025)
Parallel/Distributed Tabu Search for Scheduling Microprocessor Tasks in Hybrid Flowshop
by: Janiak, Adam, et al.
Published: (2025)
by: Janiak, Adam, et al.
Published: (2025)
A trustless society? A political look at the blockchain vision
by: Rehak, Rainer
Published: (2024)
by: Rehak, Rainer
Published: (2024)
HashKitty: Distributed Password Analysis
by: Antunes, Pedro, et al.
Published: (2025)
by: Antunes, Pedro, et al.
Published: (2025)
Simulating Cloud Environments of Connected Vehicles for Anomaly Detection
by: Weiß, M., et al.
Published: (2024)
by: Weiß, M., et al.
Published: (2024)
Similar Items
-
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
by: Stokely, Murray, et al.
Published: (2025) -
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
by: Xie, Tianfang
Published: (2026) -
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026) -
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
by: Talluri, Sacheendra, et al.
Published: (2025) -
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)