Eva: Cost-Efficient Cloud-Based Cluster Scheduling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Tzu-Tao, Venkataraman, Shivaram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PAL: A Variability-Aware Policy for Scheduling ML Workloads in GPU Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2024)
von: Jain, Rutwik, et al.
Veröffentlicht: (2024)
LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models
von: Chang, Tzu-Tao, et al.
Veröffentlicht: (2025)
von: Chang, Tzu-Tao, et al.
Veröffentlicht: (2025)
Minos: Systematically Classifying Performance and Power Characteristics of GPU Workloads on HPC Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2026)
von: Jain, Rutwik, et al.
Veröffentlicht: (2026)
SYMPHONY: Improving Memory Management for LLM Inference Workloads
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
Efficient Probabilistic Workflow Scheduling for IaaS Clouds
von: Russo, Gabriele Russo, et al.
Veröffentlicht: (2024)
von: Russo, Gabriele Russo, et al.
Veröffentlicht: (2024)
DeepVM: Integrating Spot and On-Demand VMs for Cost-Efficient Deep Learning Clusters in the Cloud
von: Kim, Yoochan, et al.
Veröffentlicht: (2024)
von: Kim, Yoochan, et al.
Veröffentlicht: (2024)
CarbonFlex: Enabling Carbon-aware Provisioning and Scheduling for Cloud Clusters
von: Hanafy, Walid A., et al.
Veröffentlicht: (2025)
von: Hanafy, Walid A., et al.
Veröffentlicht: (2025)
Tesserae: Scalable Placement Policies for Deep Learning Workloads
von: Bian, Song, et al.
Veröffentlicht: (2025)
von: Bian, Song, et al.
Veröffentlicht: (2025)
A Deep Reinforcement Learning Approach for Cost Optimized Workflow Scheduling in Cloud Computing Environments
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
Armada: Memory-Efficient Distributed Training of Large-Scale Graph Neural Networks
von: Waleffe, Roger, et al.
Veröffentlicht: (2025)
von: Waleffe, Roger, et al.
Veröffentlicht: (2025)
KubeDSM: A Kubernetes-based Dynamic Scheduling and Migration Framework for Cloud-Assisted Edge Clusters
von: Pashaeehir, Amirhossein, et al.
Veröffentlicht: (2025)
von: Pashaeehir, Amirhossein, et al.
Veröffentlicht: (2025)
Harpagon: Minimizing DNN Serving Cost via Efficient Dispatching, Scheduling and Splitting
von: Zhao, Zhixin, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixin, et al.
Veröffentlicht: (2024)
Bandwidth-Aware and Cost-Efficient Pipeline Parallel Scheduling in Geo-Distributed LLM Training
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
An Elastic Job Scheduler for HPC Applications on the Cloud
von: Bhosale, Aditya, et al.
Veröffentlicht: (2025)
von: Bhosale, Aditya, et al.
Veröffentlicht: (2025)
In Serverless, OS Scheduler Choice Costs Money: A Hybrid Scheduling Approach for Cheaper FaaS
von: Zhao, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhao, Yuxuan, et al.
Veröffentlicht: (2024)
FedCostAware: Enabling Cost-Aware Federated Learning on the Cloud
von: Sinha, Aditya, et al.
Veröffentlicht: (2025)
von: Sinha, Aditya, et al.
Veröffentlicht: (2025)
Efficient Data Labeling and Optimal Device Scheduling in HWNs Using Clustered Federated Semi-Supervised Learning
von: Hamood, Moqbel, et al.
Veröffentlicht: (2024)
von: Hamood, Moqbel, et al.
Veröffentlicht: (2024)
Exploring Distributed Vector Databases Performance on HPC Platforms: A Study with Qdrant
von: Ockerman, Seth, et al.
Veröffentlicht: (2025)
von: Ockerman, Seth, et al.
Veröffentlicht: (2025)
Minimizing Energy in Reliability and Deadline-Ensured Workflow Scheduling in Cloud
von: Sarkar, Suvarthi, et al.
Veröffentlicht: (2025)
von: Sarkar, Suvarthi, et al.
Veröffentlicht: (2025)
Megha: Decentralized Global Fair Scheduling for Federated Clusters
von: Thiyyakat, Meghana, et al.
Veröffentlicht: (2021)
von: Thiyyakat, Meghana, et al.
Veröffentlicht: (2021)
An Online Fragmentation-Aware Scheduler for Managing GPU-Sharing Workloads on Multi-Instance GPUs
von: Ting, Hsu-Tzu, et al.
Veröffentlicht: (2025)
von: Ting, Hsu-Tzu, et al.
Veröffentlicht: (2025)
COUNTER: Cluster GCN based Energy Efficient Resource Management for Sustainable Cloud Computing Environments
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
von: Kim, Taeyoon, et al.
Veröffentlicht: (2026)
von: Kim, Taeyoon, et al.
Veröffentlicht: (2026)
Hestia: Hyperthread-Level Scheduling for Cloud Microservices with Interference-Aware Attention
von: Yang, Dingyu, et al.
Veröffentlicht: (2026)
von: Yang, Dingyu, et al.
Veröffentlicht: (2026)
Exact, Efficient, and Reliable Multi-Objective and Multi-Constrained IoT Workflow Scheduling in Edge-Hub-Cloud Cyber-Physical Systems
von: Kouloumpris, Andreas, et al.
Veröffentlicht: (2026)
von: Kouloumpris, Andreas, et al.
Veröffentlicht: (2026)
Power-Aware Scheduling for Multi-Center HPC Electricity Cost Optimization
von: Hossain, Abrar, et al.
Veröffentlicht: (2025)
von: Hossain, Abrar, et al.
Veröffentlicht: (2025)
Rubick: Exploiting Job Reconfigurability for Deep Learning Cluster Scheduling
von: Zhang, Xinyi, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2024)
Energy Efficient Scheduling for Serverless Systems
von: Tsenos, Michail, et al.
Veröffentlicht: (2024)
von: Tsenos, Michail, et al.
Veröffentlicht: (2024)
A Poly-Log Approximation for Transaction Scheduling in Fog-Cloud Computing and Beyond
von: Adhikari, Ramesh, et al.
Veröffentlicht: (2025)
von: Adhikari, Ramesh, et al.
Veröffentlicht: (2025)
Scientific Workflow Scheduling in Cloud Considering Cold Start and Variable Pricing Model
von: Sarkar, Suvarthi, et al.
Veröffentlicht: (2025)
von: Sarkar, Suvarthi, et al.
Veröffentlicht: (2025)
Adaptive Heuristics for Scheduling DNN Inferencing on Edge and Cloud for Personalized UAV Fleets
von: Raj, Suman, et al.
Veröffentlicht: (2024)
von: Raj, Suman, et al.
Veröffentlicht: (2024)
A Survey on Scheduling Techniques in the Edge Cloud: Issues, Challenges and Future Directions
von: Asghar, Hassan, et al.
Veröffentlicht: (2022)
von: Asghar, Hassan, et al.
Veröffentlicht: (2022)
A Taxonomy of Schedulers -- Operating Systems, Clusters and Big Data Frameworks
von: Sliwko, Leszek
Veröffentlicht: (2025)
von: Sliwko, Leszek
Veröffentlicht: (2025)
GraphSnapShot: Caching Local Structure for Fast Graph Learning
von: Liu, Dong, et al.
Veröffentlicht: (2024)
von: Liu, Dong, et al.
Veröffentlicht: (2024)
A Decentralized Microservice Scheduling Approach Using Service Mesh in Cloud-Edge Systems
von: Wen, Yangyang, et al.
Veröffentlicht: (2025)
von: Wen, Yangyang, et al.
Veröffentlicht: (2025)
KCES: A Workflow Containerization Scheduling Scheme Under Cloud-Edge Collaboration Framework
von: Shan, Chenggang, et al.
Veröffentlicht: (2024)
von: Shan, Chenggang, et al.
Veröffentlicht: (2024)
Sustainable Graph Analytics Workload Scheduling with Evolutionary Reinforcement Learning in Edge-Cloud Systems
von: Ramicetty, P., et al.
Veröffentlicht: (2026)
von: Ramicetty, P., et al.
Veröffentlicht: (2026)
HyperDrive: Scheduling Serverless Functions in the Edge-Cloud-Space 3D Continuum
von: Pusztai, Thomas, et al.
Veröffentlicht: (2024)
von: Pusztai, Thomas, et al.
Veröffentlicht: (2024)
Towards Energy Efficient Co-Scheduling in HPC
von: Zheng, Zhong, et al.
Veröffentlicht: (2026)
von: Zheng, Zhong, et al.
Veröffentlicht: (2026)
PGT-I: Scaling Spatiotemporal GNNs with Memory-Efficient Distributed Training
von: Ockerman, Seth, et al.
Veröffentlicht: (2025)
von: Ockerman, Seth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PAL: A Variability-Aware Policy for Scheduling ML Workloads in GPU Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2024) -
LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models
von: Chang, Tzu-Tao, et al.
Veröffentlicht: (2025) -
Minos: Systematically Classifying Performance and Power Characteristics of GPU Workloads on HPC Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2026) -
SYMPHONY: Improving Memory Management for LLM Inference Workloads
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024) -
Efficient Probabilistic Workflow Scheduling for IaaS Clouds
von: Russo, Gabriele Russo, et al.
Veröffentlicht: (2024)