Asymptotically Optimal Scheduling of Multiple Parallelizable Job Classes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Berg, Benjamin, Moseley, Benjamin, Wang, Weina, Harchol-Balter, Mor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How to Rent GPUs on a Budget
von: Li, Zhouzi, et al.
Veröffentlicht: (2024)
von: Li, Zhouzi, et al.
Veröffentlicht: (2024)
Mean field optimal Core Allocation across Malleable jobs
von: Li, Zhouzi, et al.
Veröffentlicht: (2026)
von: Li, Zhouzi, et al.
Veröffentlicht: (2026)
BOA Constrictor: Squeezing Performance out of GPUs in the Cloud via Budget-Optimal Allocation
von: Li, Zhouzi, et al.
Veröffentlicht: (2026)
von: Li, Zhouzi, et al.
Veröffentlicht: (2026)
SHIRO: Near-Optimal Communication Strategies for Distributed Sparse Matrix Multiplication
von: Zhuang, Chen, et al.
Veröffentlicht: (2025)
von: Zhuang, Chen, et al.
Veröffentlicht: (2025)
Optimal Parallel Scheduling under Concave Speedup Functions
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
Flowshop Machine Scheduling: Markov Modeling, Optimal Schedules and Heuristics
von: Ghanem, Samah A. M.
Veröffentlicht: (2025)
von: Ghanem, Samah A. M.
Veröffentlicht: (2025)
Chopin: An Open Source R-language Tool to Support Spatial Analysis on Parallelizable Infrastructure
von: Song, Insang, et al.
Veröffentlicht: (2024)
von: Song, Insang, et al.
Veröffentlicht: (2024)
Advanced Scheduling Strategies for Distributed Quantum Computing Jobs
von: Ni, Gongyu, et al.
Veröffentlicht: (2026)
von: Ni, Gongyu, et al.
Veröffentlicht: (2026)
LLload: Simplifying Real-Time Job Monitoring for HPC Users
von: Byun, Chansup, et al.
Veröffentlicht: (2024)
von: Byun, Chansup, et al.
Veröffentlicht: (2024)
Unleashing the Power of Preemptive Priority-based Scheduling for Real-Time GPU Tasks
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
Hiku: Pull-Based Scheduling for Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
"Two-Stagification": Job Dispatching in Large-Scale Clusters via a Two-Stage Architecture
von: Yildiz, Mert, et al.
Veröffentlicht: (2025)
von: Yildiz, Mert, et al.
Veröffentlicht: (2025)
Taming Cold Starts: Proactive Serverless Scheduling with Model Predictive Control
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
Operational Strategies for Non-Disruptive Scheduling Transitions in Production HPC Systems
von: MacLachlan, Glen, et al.
Veröffentlicht: (2026)
von: MacLachlan, Glen, et al.
Veröffentlicht: (2026)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
von: Qi, S., et al.
Veröffentlicht: (2024)
von: Qi, S., et al.
Veröffentlicht: (2024)
The SAP Cloud Infrastructure Dataset: A Reality Check of Scheduling and Placement of VMs in Cloud Computing
von: Uhlig, Arno, et al.
Veröffentlicht: (2025)
von: Uhlig, Arno, et al.
Veröffentlicht: (2025)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
von: Muhammad, Said, et al.
Veröffentlicht: (2025)
von: Muhammad, Said, et al.
Veröffentlicht: (2025)
Is Sparse Matrix Reordering Effective for Sparse Matrix-Vector Multiplication?
von: Asudeh, Omid, et al.
Veröffentlicht: (2025)
von: Asudeh, Omid, et al.
Veröffentlicht: (2025)
Optimal Configuration of API Resources in Cloud Native Computing
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
von: Truyen, Eddy, et al.
Veröffentlicht: (2025)
Is Intelligence the Right Direction in New OS Scheduling for Multiple Resources in Cloud Environments?
von: Dou, Xinglei, et al.
Veröffentlicht: (2025)
von: Dou, Xinglei, et al.
Veröffentlicht: (2025)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
Alya towards Exascale: Optimal OpenACC Performance of the Navier-Stokes Finite Element Assembly on GPUs
von: Owen, Herbert, et al.
Veröffentlicht: (2024)
von: Owen, Herbert, et al.
Veröffentlicht: (2024)
Performance and scaling of the LFRic weather and climate model on different generations of HPE Cray EX supercomputers
von: Bull, J. Mark, et al.
Veröffentlicht: (2024)
von: Bull, J. Mark, et al.
Veröffentlicht: (2024)
PolyTOPS: Reconfigurable and Flexible Polyhedral Scheduler
von: Consolaro, Gianpietro, et al.
Veröffentlicht: (2024)
von: Consolaro, Gianpietro, et al.
Veröffentlicht: (2024)
Toward Smart Scheduling in Tapis
von: Stubbs, Joe, et al.
Veröffentlicht: (2024)
von: Stubbs, Joe, et al.
Veröffentlicht: (2024)
A Pilot Study on Tunable Precision Emulation via Automatic BLAS Offloading
von: Liu, Hang, et al.
Veröffentlicht: (2025)
von: Liu, Hang, et al.
Veröffentlicht: (2025)
BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
Cloud Performance Decomposition for Long-Term Performance Engineering: A Case Study
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
mLR: Scalable Laminography Reconstruction based on Memoization
von: Ma, Bin, et al.
Veröffentlicht: (2025)
von: Ma, Bin, et al.
Veröffentlicht: (2025)
Recorder: Comprehensive Parallel I/O Tracing and Analysis
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
von: Lin, Mao, et al.
Veröffentlicht: (2026)
von: Lin, Mao, et al.
Veröffentlicht: (2026)
GPU Cluster Scheduling for Network-Sensitive Deep Learning
von: Sharma, Aakash, et al.
Veröffentlicht: (2024)
von: Sharma, Aakash, et al.
Veröffentlicht: (2024)
Scheduling Languages: A Past, Present, and Future Taxonomy
von: Hall, Mary, et al.
Veröffentlicht: (2024)
von: Hall, Mary, et al.
Veröffentlicht: (2024)
Energy-Optimized Scheduling for AIoT Workloads Using TOPSIS
von: Pradeep, Preethika, et al.
Veröffentlicht: (2025)
von: Pradeep, Preethika, et al.
Veröffentlicht: (2025)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
von: Ather, Hammad, et al.
Veröffentlicht: (2024)
von: Ather, Hammad, et al.
Veröffentlicht: (2024)
Opt4GPTQ: Co-Optimizing Memory and Computation for 4-bit GPTQ Quantized LLM Inference on Heterogeneous Platforms
von: Zhang, Yaozheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yaozheng, et al.
Veröffentlicht: (2025)
Matryoshka: Optimization of Dynamic Diverse Quantum Chemistry Systems via Elastic Parallelism Transformation
von: Wang, Tuowei, et al.
Veröffentlicht: (2024)
von: Wang, Tuowei, et al.
Veröffentlicht: (2024)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
von: Wang, Yuxin, et al.
Veröffentlicht: (2023)
von: Wang, Yuxin, et al.
Veröffentlicht: (2023)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
von: Lin, Changyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
How to Rent GPUs on a Budget
von: Li, Zhouzi, et al.
Veröffentlicht: (2024) -
Mean field optimal Core Allocation across Malleable jobs
von: Li, Zhouzi, et al.
Veröffentlicht: (2026) -
BOA Constrictor: Squeezing Performance out of GPUs in the Cloud via Budget-Optimal Allocation
von: Li, Zhouzi, et al.
Veröffentlicht: (2026) -
SHIRO: Near-Optimal Communication Strategies for Distributed Sparse Matrix Multiplication
von: Zhuang, Chen, et al.
Veröffentlicht: (2025) -
Optimal Parallel Scheduling under Concave Speedup Functions
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)