SGPRS: Seamless GPU Partitioning Real-Time Scheduler for Periodic Deep Learning Workloads
Fuente:
arXiv
Saved in:
| Main Authors: | Babaei, Amir Fakhim, Chantem, Thidapat |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DARIS: An Oversubscribed Spatio-Temporal Scheduler for Real-Time DNN Inference on GPUs
by: Babaei, Amir Fakhim, et al.
Published: (2025)
by: Babaei, Amir Fakhim, et al.
Published: (2025)
Cost-effective Deep Learning Infrastructure with NVIDIA GPU
by: Ghimire, Aatiz, et al.
Published: (2025)
by: Ghimire, Aatiz, et al.
Published: (2025)
Modern Middlewares for Automated Vehicles: A Tutorial
by: Klüner, David Philipp, et al.
Published: (2024)
by: Klüner, David Philipp, et al.
Published: (2024)
Evaluating Asynchronous Semantics in Trace-Discovered Resilience Models: A Case Study on the OpenTelemetry Demo
by: Krasnovsky, Anatoly A.
Published: (2025)
by: Krasnovsky, Anatoly A.
Published: (2025)
Emergence-as-Code for Self-Governing Reliable Systems
by: Krasnovsky, Anatoly A.
Published: (2026)
by: Krasnovsky, Anatoly A.
Published: (2026)
Hydra: Brokering Cloud and HPC Resources to Support the Execution of Heterogeneous Workloads at Scale
by: Alsaadi, Aymen, et al.
Published: (2024)
by: Alsaadi, Aymen, et al.
Published: (2024)
STaleX: A Spatiotemporal-Aware Adaptive Auto-scaling Framework for Microservices
by: Dashtbani, Majid, et al.
Published: (2025)
by: Dashtbani, Majid, et al.
Published: (2025)
Metronome: Differentiated Delay Scheduling for Serverless Functions
by: Chen, Zhuangbin, et al.
Published: (2025)
by: Chen, Zhuangbin, et al.
Published: (2025)
Fixed-Priority and EDF Schedules for ROS2 Graphs on Uniprocessor
by: Bell, Oren, et al.
Published: (2025)
by: Bell, Oren, et al.
Published: (2025)
Investigating Matrix Repartitioning to Address the Over- and Undersubscription Challenge for a GPU-based CFD Solver
by: Olenik, Gregor, et al.
Published: (2025)
by: Olenik, Gregor, et al.
Published: (2025)
Energy-Optimized Scheduling for AIoT Workloads Using TOPSIS
by: Pradeep, Preethika, et al.
Published: (2025)
by: Pradeep, Preethika, et al.
Published: (2025)
ATOM: Asynchronous Training of Massive Models for Deep Learning in a Decentralized Environment
by: Wu, Xiaofeng, et al.
Published: (2024)
by: Wu, Xiaofeng, et al.
Published: (2024)
A Blockchain-Oriented Software Engineering Architecture for Carbon Credit Certification Systems
by: Vaccargiu, Matteo, et al.
Published: (2026)
by: Vaccargiu, Matteo, et al.
Published: (2026)
Software Engineering for Collective Cyber-Physical Ecosystems
by: Casadei, Roberto, et al.
Published: (2024)
by: Casadei, Roberto, et al.
Published: (2024)
Macroprogramming: Concepts, State of the Art, and Opportunities of Macroscopic Behaviour Modelling
by: Casadei, Roberto
Published: (2022)
by: Casadei, Roberto
Published: (2022)
TorchGWAS : GPU-accelerated GWAS for thousands of quantitative phenotypes
by: Zhao, Xingzhong, et al.
Published: (2026)
by: Zhao, Xingzhong, et al.
Published: (2026)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
by: Nichols, Daniel, et al.
Published: (2025)
by: Nichols, Daniel, et al.
Published: (2025)
UPC Sentinel: An Accurate Approach for Detecting Upgradeability Proxy Contracts in Ethereum
by: Ebrahimi, Amir M., et al.
Published: (2024)
by: Ebrahimi, Amir M., et al.
Published: (2024)
A Large-Scale Exploratory Study on the Proxy Pattern in Ethereum
by: Ebrahimi, Amir M., et al.
Published: (2025)
by: Ebrahimi, Amir M., et al.
Published: (2025)
Real-Time GPU-Accelerated Monte Carlo Evaluation of Safety-Critical AEB Systems Under Uncertainty
by: Karjol, Akshay, et al.
Published: (2026)
by: Karjol, Akshay, et al.
Published: (2026)
A Survey of Real-Time Support, Analysis, and Advancements in ROS 2
by: Casini, Daniel, et al.
Published: (2025)
by: Casini, Daniel, et al.
Published: (2025)
$μ$OpTime: Statically Reducing the Execution Time of Microbenchmark Suites Using Stability Metrics
by: Japke, Nils, et al.
Published: (2025)
by: Japke, Nils, et al.
Published: (2025)
Object as a Service: Simplifying Cloud-Native Development through Serverless Object Abstraction
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
CodeAD: Synthesize Code of Rules for Log-based Anomaly Detection with LLMs
by: Huang, Junjie, et al.
Published: (2025)
by: Huang, Junjie, et al.
Published: (2025)
Interferences within a certifiable design methodology for high-performance multi-core platforms
by: Khelassi, Mohamed Amine, et al.
Published: (2026)
by: Khelassi, Mohamed Amine, et al.
Published: (2026)
Who is in Charge here? Understanding How Runtime Configuration Affects Software along with Variables&Constants
by: Luo, Chaopeng, et al.
Published: (2025)
by: Luo, Chaopeng, et al.
Published: (2025)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
by: Jain, Kunal, et al.
Published: (2024)
by: Jain, Kunal, et al.
Published: (2024)
Balancing Fairness and Performance in Multi-User Spark Workloads with Dynamic Scheduling (extended version)
by: Kažemaks, Dāvis, et al.
Published: (2025)
by: Kažemaks, Dāvis, et al.
Published: (2025)
Lockbox -- A Zero Trust Architecture for Secure Processing of Sensitive Cloud Workloads
by: Thotempudi, Vamshi Krishna, et al.
Published: (2026)
by: Thotempudi, Vamshi Krishna, et al.
Published: (2026)
Supercharging Federated Learning with Flower and NVIDIA FLARE
by: Roth, Holger R., et al.
Published: (2024)
by: Roth, Holger R., et al.
Published: (2024)
Learning Recovery Strategies for Dynamic Self-healing in Reactive Systems
by: Sanabria, Mateo, et al.
Published: (2024)
by: Sanabria, Mateo, et al.
Published: (2024)
Model Discovery and Graph Simulation: A Lightweight Gateway to Chaos Engineering
by: Krasnovsky, Anatoly A.
Published: (2025)
by: Krasnovsky, Anatoly A.
Published: (2025)
Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads
by: Agyemang, Justice Owusu, et al.
Published: (2026)
by: Agyemang, Justice Owusu, et al.
Published: (2026)
Blink: CPU-Free LLM Inference by Delegating the Serving Stack to GPU and SmartNIC
by: Siavashi, Mohammad, et al.
Published: (2026)
by: Siavashi, Mohammad, et al.
Published: (2026)
Communication-Avoiding SpGEMM via Trident Partitioning on Hierarchical GPU Interconnects
by: Bellavita, Julian, et al.
Published: (2026)
by: Bellavita, Julian, et al.
Published: (2026)
MARCO: Multi-Agent Code Optimization with Real-Time Knowledge Integration for High-Performance Computing
by: Rahman, Asif, et al.
Published: (2025)
by: Rahman, Asif, et al.
Published: (2025)
FIRST: Federated Inference Resource Scheduling Toolkit for Scientific AI Model Access
by: Tanikanti, Aditya, et al.
Published: (2025)
by: Tanikanti, Aditya, et al.
Published: (2025)
SeBS-Flow: Benchmarking Serverless Cloud Function Workflows
by: Schmid, Larissa, et al.
Published: (2024)
by: Schmid, Larissa, et al.
Published: (2024)
CloudHeatMap: Heatmap-Based Monitoring for Large-Scale Cloud Systems
by: Sohana, Sarah, et al.
Published: (2024)
by: Sohana, Sarah, et al.
Published: (2024)
Adaptable TeaStore
by: Bliudze, Simon, et al.
Published: (2024)
by: Bliudze, Simon, et al.
Published: (2024)
Similar Items
-
DARIS: An Oversubscribed Spatio-Temporal Scheduler for Real-Time DNN Inference on GPUs
by: Babaei, Amir Fakhim, et al.
Published: (2025) -
Cost-effective Deep Learning Infrastructure with NVIDIA GPU
by: Ghimire, Aatiz, et al.
Published: (2025) -
Modern Middlewares for Automated Vehicles: A Tutorial
by: Klüner, David Philipp, et al.
Published: (2024) -
Evaluating Asynchronous Semantics in Trace-Discovered Resilience Models: A Case Study on the OpenTelemetry Demo
by: Krasnovsky, Anatoly A.
Published: (2025) -
Emergence-as-Code for Self-Governing Reliable Systems
by: Krasnovsky, Anatoly A.
Published: (2026)