Duration-Informed Workload Scheduler
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Loreti, Daniela, Leone, Davide, Borghesi, Andrea |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Workload Schedulers -- Genesis, Algorithms and Differences
par: Sliwko, Leszek, et autres
Publié: (2025)
par: Sliwko, Leszek, et autres
Publié: (2025)
Topology-aware Preemptive Scheduling for Co-located LLM Workloads
par: Zhang, Ping, et autres
Publié: (2024)
par: Zhang, Ping, et autres
Publié: (2024)
Capacity Planning and Scheduling for Jobs with Uncertainty in Resource Usage and Duration
par: Patra, Sunandita, et autres
Publié: (2025)
par: Patra, Sunandita, et autres
Publié: (2025)
HiveMind: OS-Inspired Scheduling for Concurrent LLM Agent Workloads
par: Agyemang, Justice Owusu, et autres
Publié: (2026)
par: Agyemang, Justice Owusu, et autres
Publié: (2026)
Resource Allocation and Workload Scheduling for Large-Scale Distributed Deep Learning: A Survey
par: Liang, Feng, et autres
Publié: (2024)
par: Liang, Feng, et autres
Publié: (2024)
Tesserae: Scalable Placement Policies for Deep Learning Workloads
par: Bian, Song, et autres
Publié: (2025)
par: Bian, Song, et autres
Publié: (2025)
An Advanced Reinforcement Learning Framework for Online Scheduling of Deferrable Workloads in Cloud Computing
par: Dong, Hang, et autres
Publié: (2024)
par: Dong, Hang, et autres
Publié: (2024)
Hybrid Learning and Optimization-Based Dynamic Scheduling for DL Workloads on Heterogeneous GPU Clusters
par: Dongare, Shruti, et autres
Publié: (2025)
par: Dongare, Shruti, et autres
Publié: (2025)
Compass: A Decentralized Scheduler for Latency-Sensitive ML Workflows
par: Yang, Yuting, et autres
Publié: (2024)
par: Yang, Yuting, et autres
Publié: (2024)
Hybrid Heterogeneous Clusters Can Lower the Energy Consumption of LLM Inference Workloads
par: Wilkins, Grant, et autres
Publié: (2024)
par: Wilkins, Grant, et autres
Publié: (2024)
Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption
par: Yildiz, Mert, et autres
Publié: (2026)
par: Yildiz, Mert, et autres
Publié: (2026)
Towards Carbon-Aware Container Orchestration: Predicting Workload Energy Consumption with Federated Learning
par: Saad, Zainab, et autres
Publié: (2025)
par: Saad, Zainab, et autres
Publié: (2025)
Quantifying Energy and Cost Benefits of Hybrid Edge Cloud: Analysis of Traditional and Agentic Workloads
par: Alamouti, Siavash
Publié: (2025)
par: Alamouti, Siavash
Publié: (2025)
Mixture-of-Schedulers: An Adaptive Scheduling Agent as a Learned Router for Expert Policies
par: Wang, Xinbo, et autres
Publié: (2025)
par: Wang, Xinbo, et autres
Publié: (2025)
ReCoVer: Resilient LLM Pre-Training System via Fault-Tolerant Collective and Versatile Workload
par: Liu, Ziyue, et autres
Publié: (2026)
par: Liu, Ziyue, et autres
Publié: (2026)
Enabling Secure and Ephemeral AI Workloads in Data Mesh Environments
par: Patel, Chinkit, et autres
Publié: (2025)
par: Patel, Chinkit, et autres
Publié: (2025)
Power- and Fragmentation-aware Online Scheduling for GPU Datacenters
par: Lettich, Francesco, et autres
Publié: (2024)
par: Lettich, Francesco, et autres
Publié: (2024)
Dynamic Scheduling Strategies for Resource Optimization in Computing Environments
par: Wang, Xiaoye
Publié: (2024)
par: Wang, Xiaoye
Publié: (2024)
Equinox: Holistic Fair Scheduling in Serving Large Language Models
par: Wei, Zhixiang, et autres
Publié: (2025)
par: Wei, Zhixiang, et autres
Publié: (2025)
Block: Balancing Load in LLM Serving with Context, Knowledge and Predictive Scheduling
par: Da, Wei, et autres
Publié: (2025)
par: Da, Wei, et autres
Publié: (2025)
TAPAS: Thermal- and Power-Aware Scheduling for LLM Inference in Cloud Platforms
par: Stojkovic, Jovan, et autres
Publié: (2025)
par: Stojkovic, Jovan, et autres
Publié: (2025)
Evaluating the Efficacy of LLM-Based Reasoning for Multiobjective HPC Job Scheduling
par: Jadhav, Prachi, et autres
Publié: (2025)
par: Jadhav, Prachi, et autres
Publié: (2025)
TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference
par: Papaioannou, Konstantinos, et autres
Publié: (2026)
par: Papaioannou, Konstantinos, et autres
Publié: (2026)
WORKSWORLD: A Domain for Integrated Numeric Planning and Scheduling of Distributed Pipelined Workflows
par: Paul, Taylor, et autres
Publié: (2026)
par: Paul, Taylor, et autres
Publié: (2026)
Reconstruction-Based Adaptive Scheduling Using AI Inferences in Safety-Critical Systems
par: Alshaer, Samer, et autres
Publié: (2025)
par: Alshaer, Samer, et autres
Publié: (2025)
Reinforcement Learning-driven Data-intensive Workflow Scheduling for Volunteer Edge-Cloud
par: Mounesan, Motahare, et autres
Publié: (2024)
par: Mounesan, Motahare, et autres
Publié: (2024)
Reducing Fragmentation and Starvation in GPU Clusters through Dynamic Multi-Objective Scheduling
par: Mamirov, Akhmadillo
Publié: (2025)
par: Mamirov, Akhmadillo
Publié: (2025)
Efficient MoE Inference with Fine-Grained Scheduling of Disaggregated Expert Parallelism
par: Pan, Xinglin, et autres
Publié: (2025)
par: Pan, Xinglin, et autres
Publié: (2025)
SparOA: Sparse and Operator-aware Hybrid Scheduling for Edge DNN Inference
par: Zhang, Ziyang, et autres
Publié: (2025)
par: Zhang, Ziyang, et autres
Publié: (2025)
CoRaiS: Lightweight Real-Time Scheduler for Multi-Edge Cooperative Computing
par: Hu, Yujiao, et autres
Publié: (2024)
par: Hu, Yujiao, et autres
Publié: (2024)
Eventually-Consistent Federated Scheduling for Data Center Workloads
par: Thiyyakat, Meghana, et autres
Publié: (2023)
par: Thiyyakat, Meghana, et autres
Publié: (2023)
A Scheduling Framework for Efficient MoE Inference on Edge GPU-NDP Systems
par: Wu, Qi, et autres
Publié: (2026)
par: Wu, Qi, et autres
Publié: (2026)
TS-EoH: An Edge Server Task Scheduling Algorithm Based on Evolution of Heuristic
par: Yatong, Wang, et autres
Publié: (2024)
par: Yatong, Wang, et autres
Publié: (2024)
DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling
par: Pathak, Nilavra, et autres
Publié: (2026)
par: Pathak, Nilavra, et autres
Publié: (2026)
iScheduler: Reinforcement Learning-Driven Continual Optimization for Large-Scale Resource Investment Problems
par: Hu, Yi-Xiang, et autres
Publié: (2026)
par: Hu, Yi-Xiang, et autres
Publié: (2026)
Deep Reinforcement Learning for Job Scheduling and Resource Management in Cloud Computing: An Algorithm-Level Review
par: Gu, Yan, et autres
Publié: (2025)
par: Gu, Yan, et autres
Publié: (2025)
LLMSched: Uncertainty-Aware Workload Scheduling for Compound LLM Applications
par: Zhu, Botao, et autres
Publié: (2025)
par: Zhu, Botao, et autres
Publié: (2025)
FlowPrefill: Decoupling Preemption from Prefill Scheduling Granularity to Mitigate Head-of-Line Blocking in LLM Serving
par: Hsieh, Chia-chi, et autres
Publié: (2026)
par: Hsieh, Chia-chi, et autres
Publié: (2026)
Learning to Schedule: A Supervised Learning Framework for Network-Aware Scheduling of Data-Intensive Workloads
par: Timilsina, Sankalpa, et autres
Publié: (2025)
par: Timilsina, Sankalpa, et autres
Publié: (2025)
Decentralized Distributed Proximal Policy Optimization (DD-PPO) for High Performance Computing Scheduling on Multi-User Systems
par: Sgambati, Matthew, et autres
Publié: (2025)
par: Sgambati, Matthew, et autres
Publié: (2025)
Documents similaires
-
Workload Schedulers -- Genesis, Algorithms and Differences
par: Sliwko, Leszek, et autres
Publié: (2025) -
Topology-aware Preemptive Scheduling for Co-located LLM Workloads
par: Zhang, Ping, et autres
Publié: (2024) -
Capacity Planning and Scheduling for Jobs with Uncertainty in Resource Usage and Duration
par: Patra, Sunandita, et autres
Publié: (2025) -
HiveMind: OS-Inspired Scheduling for Concurrent LLM Agent Workloads
par: Agyemang, Justice Owusu, et autres
Publié: (2026) -
Resource Allocation and Workload Scheduling for Large-Scale Distributed Deep Learning: A Survey
par: Liang, Feng, et autres
Publié: (2024)