Trustworthy Scheduling for Big Data Applications

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Tomaras, Dimitrios, Kalogeraki, Vana, Gunopulos, Dimitrios
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911401659858944
author Tomaras, Dimitrios
Kalogeraki, Vana
Gunopulos, Dimitrios
author_facet Tomaras, Dimitrios
Kalogeraki, Vana
Gunopulos, Dimitrios
contents Recent advances in modern containerized execution environments have resulted in substantial benefits in terms of elasticity and more efficient utilization of computing resources. Although existing schedulers strive to optimize performance metrics like task execution times and resource utilization, they provide limited transparency into their decision-making processes or the specific actions developers must take to meet Service Level Objectives (SLOs). In this work, we propose X-Sched, a middleware that uses explainability techniques to generate actionable guidance on resource configurations that makes task execution in containerized environments feasible, under resource and time constraints. X-Sched addresses this gap by integrating counterfactual explanations with advanced machine learning models, such as Random Forests, to efficiently identify optimal configurations. This approach not only ensures that tasks are executed in line with performance goals but also gives users clear, actionable insights into the rationale behind scheduling decisions. Our experimental results validated with data from real-world execution environments, illustrate the efficiency, benefits and practicality of our approach.
format Preprint
id arxiv_https___arxiv_org_abs_2601_18983
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Trustworthy Scheduling for Big Data Applications
Tomaras, Dimitrios
Kalogeraki, Vana
Gunopulos, Dimitrios
Distributed, Parallel, and Cluster Computing
Recent advances in modern containerized execution environments have resulted in substantial benefits in terms of elasticity and more efficient utilization of computing resources. Although existing schedulers strive to optimize performance metrics like task execution times and resource utilization, they provide limited transparency into their decision-making processes or the specific actions developers must take to meet Service Level Objectives (SLOs). In this work, we propose X-Sched, a middleware that uses explainability techniques to generate actionable guidance on resource configurations that makes task execution in containerized environments feasible, under resource and time constraints. X-Sched addresses this gap by integrating counterfactual explanations with advanced machine learning models, such as Random Forests, to efficiently identify optimal configurations. This approach not only ensures that tasks are executed in line with performance goals but also gives users clear, actionable insights into the rationale behind scheduling decisions. Our experimental results validated with data from real-world execution environments, illustrate the efficiency, benefits and practicality of our approach.
title Trustworthy Scheduling for Big Data Applications
topic Distributed, Parallel, and Cluster Computing
url https://arxiv.org/abs/2601.18983