MQFQ-Sticky: Fair Queueing For Serverless GPU Functions
Fuente:
arXiv
Saved in:
| Main Authors: | Fuerst, Alexander, Anil, Siddharth, Dixit, Vishakha, Purushottam, Kulkarni, Sharma, Prateek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FaasMeter: Energy-First Serverless Computing
by: Rehman, Abdul, et al.
Published: (2024)
by: Rehman, Abdul, et al.
Published: (2024)
A Deep Recurrent-Reinforcement Learning Method for Intelligent AutoScaling of Serverless Functions
by: Agarwal, Siddharth, et al.
Published: (2023)
by: Agarwal, Siddharth, et al.
Published: (2023)
Input-Based Ensemble-Learning Method for Dynamic Memory Configuration of Serverless Computing Functions
by: Agarwal, Siddharth, et al.
Published: (2024)
by: Agarwal, Siddharth, et al.
Published: (2024)
On-demand Cold Start Frequency Reduction with Off-Policy Reinforcement Learning in Serverless Computing
by: Agarwal, Siddharth, et al.
Published: (2023)
by: Agarwal, Siddharth, et al.
Published: (2023)
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
by: Zhao, Chen, et al.
Published: (2026)
by: Zhao, Chen, et al.
Published: (2026)
AARC: Automated Affinity-aware Resource Configuration for Serverless Workflows
by: Jin, Lingxiao, et al.
Published: (2025)
by: Jin, Lingxiao, et al.
Published: (2025)
Comprehensive Deadlock Prevention for GPU Collective Communication
by: Pan, Lichen, et al.
Published: (2023)
by: Pan, Lichen, et al.
Published: (2023)
emucxl: an emulation framework for CXL-based disaggregated memory applications
by: Gond, Raja, et al.
Published: (2024)
by: Gond, Raja, et al.
Published: (2024)
cuNRTO: GPU-Accelerated Nonlinear Robust Trajectory Optimization
by: Wang, Jiawei, et al.
Published: (2026)
by: Wang, Jiawei, et al.
Published: (2026)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
by: Jain, Kunal, et al.
Published: (2024)
by: Jain, Kunal, et al.
Published: (2024)
Balancing Fairness and Performance in Multi-User Spark Workloads with Dynamic Scheduling (extended version)
by: Kažemaks, Dāvis, et al.
Published: (2025)
by: Kažemaks, Dāvis, et al.
Published: (2025)
Litmus: Fair Pricing for Serverless Computing
by: Pei, Qi, et al.
Published: (2024)
by: Pei, Qi, et al.
Published: (2024)
Towards Fast Setup and High Throughput of GPU Serverless Computing
by: Zhao, Han, et al.
Published: (2024)
by: Zhao, Han, et al.
Published: (2024)
Dirigent: Lightweight Serverless Orchestration
by: Cvetković, Lazar, et al.
Published: (2024)
by: Cvetković, Lazar, et al.
Published: (2024)
The Jevons Paradox In Cloud Computing: A Thermodynamics Perspective
by: Sharma, Prateek
Published: (2024)
by: Sharma, Prateek
Published: (2024)
SGPRS: Seamless GPU Partitioning Real-Time Scheduler for Periodic Deep Learning Workloads
by: Babaei, Amir Fakhim, et al.
Published: (2024)
by: Babaei, Amir Fakhim, et al.
Published: (2024)
TrEnv-X: Transparently Share Serverless Execution Environments Across Different Functions and Nodes
by: Huang, Jialiang, et al.
Published: (2025)
by: Huang, Jialiang, et al.
Published: (2025)
FaaSTube: Optimizing GPU-oriented Data Transfer for Serverless Computing
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
Saarthi: An End-to-End Intelligent Platform for Optimising Distributed Serverless Workloads
by: Agarwal, Siddharth, et al.
Published: (2025)
by: Agarwal, Siddharth, et al.
Published: (2025)
HAS-GPU: Efficient Hybrid Auto-scaling with Fine-grained GPU Allocation for SLO-aware Serverless Inferences
by: Gu, Jianfeng, et al.
Published: (2025)
by: Gu, Jianfeng, et al.
Published: (2025)
Torpor: GPU-Enabled Serverless Computing for Low-Latency, Resource-Efficient Inference
by: Yu, Minchen, et al.
Published: (2023)
by: Yu, Minchen, et al.
Published: (2023)
Robust Set Partitioning Strategy for Malicious Information Detection in Large-Scale Internet of Things
by: Suo, Yuhan, et al.
Published: (2025)
by: Suo, Yuhan, et al.
Published: (2025)
Deployment of Containerized Simulations in an API-Driven Distributed Infrastructure
by: Kraus, Tim, et al.
Published: (2025)
by: Kraus, Tim, et al.
Published: (2025)
Distribution and Management of Datacenter Load Decoupling
by: Lin, Liuzixuan, et al.
Published: (2025)
by: Lin, Liuzixuan, et al.
Published: (2025)
Stannic: Systolic STochAstic ONliNe SchedulIng AcCelerator
by: Ross, Adam H., et al.
Published: (2025)
by: Ross, Adam H., et al.
Published: (2025)
H-MBR: Hypervisor-level Memory Bandwidth Reservation for Mixed Criticality Systems
by: Oliveira, Afonso, et al.
Published: (2025)
by: Oliveira, Afonso, et al.
Published: (2025)
GPUArmor: A Hardware-Software Co-design for Efficient and Scalable Memory Safety on GPUs
by: Ziad, Mohamed Tarek Ibn, et al.
Published: (2025)
by: Ziad, Mohamed Tarek Ibn, et al.
Published: (2025)
Towards Sustainable Computing: Exploring Energy Consumption Efficiency of Alternative Configurations and Workloads in an Open Source Messaging System
by: Voreakou, Maria, et al.
Published: (2025)
by: Voreakou, Maria, et al.
Published: (2025)
20 Years in Life of a Smart Building: A retrospective
by: Skrivankova, Karolina, et al.
Published: (2025)
by: Skrivankova, Karolina, et al.
Published: (2025)
Beyond the Bermuda Triangle of Contention: IOMMU Interference in Mixed Criticality Systems
by: Costa, Diogo, et al.
Published: (2025)
by: Costa, Diogo, et al.
Published: (2025)
Investigating Timing-Based Information Leakage in Data Flow-Driven Real-Time Systems
by: Babar, Mohammad Fakhruddin, et al.
Published: (2025)
by: Babar, Mohammad Fakhruddin, et al.
Published: (2025)
Optimizing Microgrid Composition for Sustainable Data Centers
by: Irion, Julius, et al.
Published: (2025)
by: Irion, Julius, et al.
Published: (2025)
DeepCEE: Efficient Cross-Region Model Distributed Training System under Heterogeneous GPUs and Networks
by: Wang, Jinquan, et al.
Published: (2025)
by: Wang, Jinquan, et al.
Published: (2025)
Optimizing Sensor Node Localization for Achieving Sustainable Smart Agriculture System Connectivity
by: Naeem, Mohamed
Published: (2025)
by: Naeem, Mohamed
Published: (2025)
Designing Dense Satellite Clusters for Distributed Space-based Datacenters
by: Pénot, Jules, et al.
Published: (2026)
by: Pénot, Jules, et al.
Published: (2026)
Load Balancing Using Sparse Communication
by: Mendelson, Gal, et al.
Published: (2022)
by: Mendelson, Gal, et al.
Published: (2022)
Vessim: A Testbed for Carbon-Aware Applications and Systems
by: Wiesner, Philipp, et al.
Published: (2023)
by: Wiesner, Philipp, et al.
Published: (2023)
Iterative Thresholding and Projection Algorithms and Model-Based Deep Neural Networks for Sparse LQR Control Design
by: Cho, Myung
Published: (2022)
by: Cho, Myung
Published: (2022)
Thinking fast and slow -- a cognitive inspired framework for decision intelligence for power systems
by: Mathur, Apoorv
Published: (2026)
by: Mathur, Apoorv
Published: (2026)
SPARe: Stacked Parallelism with Adaptive Reordering for Fault-Tolerant LLM Pretraining Systems with 100k+ GPUs
by: Lee, Jin, et al.
Published: (2026)
by: Lee, Jin, et al.
Published: (2026)
Similar Items
-
FaasMeter: Energy-First Serverless Computing
by: Rehman, Abdul, et al.
Published: (2024) -
A Deep Recurrent-Reinforcement Learning Method for Intelligent AutoScaling of Serverless Functions
by: Agarwal, Siddharth, et al.
Published: (2023) -
Input-Based Ensemble-Learning Method for Dynamic Memory Configuration of Serverless Computing Functions
by: Agarwal, Siddharth, et al.
Published: (2024) -
On-demand Cold Start Frequency Reduction with Off-Policy Reinforcement Learning in Serverless Computing
by: Agarwal, Siddharth, et al.
Published: (2023) -
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
by: Zhao, Chen, et al.
Published: (2026)