KUBEDIRECT: Unleashing the Full Power of the Cluster Manager for Serverless Computing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Sheng, Zhang, Zhiquan, Liu, Xuanzhe, Jin, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HydraServe: Minimizing Cold Start Latency for Serverless LLM Serving in Public Clouds
von: Lou, Chiheng, et al.
Veröffentlicht: (2025)
von: Lou, Chiheng, et al.
Veröffentlicht: (2025)
Litmus: Fair Pricing for Serverless Computing
von: Pei, Qi, et al.
Veröffentlicht: (2024)
von: Pei, Qi, et al.
Veröffentlicht: (2024)
Caching Aided Multi-Tenant Serverless Computing
von: Qiao, Chu, et al.
Veröffentlicht: (2024)
von: Qiao, Chu, et al.
Veröffentlicht: (2024)
Making Serverless Computing Extensible: A Case Study of Serverless Data Analytics
von: Yu, Minchen, et al.
Veröffentlicht: (2025)
von: Yu, Minchen, et al.
Veröffentlicht: (2025)
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
von: Chang, Zihan, et al.
Veröffentlicht: (2024)
von: Chang, Zihan, et al.
Veröffentlicht: (2024)
Multi-Event Triggers for Serverless Computing
von: Carl, Natalie, et al.
Veröffentlicht: (2025)
von: Carl, Natalie, et al.
Veröffentlicht: (2025)
Serverless Computing: Architecture, Concepts, and Applications
von: Ghorbian, Mohsen, et al.
Veröffentlicht: (2025)
von: Ghorbian, Mohsen, et al.
Veröffentlicht: (2025)
Lagom: Unleashing the Power of Communication and Computation Overlapping for Distributed LLM Training
von: Xu, Guanbin, et al.
Veröffentlicht: (2026)
von: Xu, Guanbin, et al.
Veröffentlicht: (2026)
ClusterLess: Deadline-Aware Serverless Workflow Orchestration on Federated Edge Clusters
von: Farahani, Reza, et al.
Veröffentlicht: (2026)
von: Farahani, Reza, et al.
Veröffentlicht: (2026)
FaasMeter: Energy-First Serverless Computing
von: Rehman, Abdul, et al.
Veröffentlicht: (2024)
von: Rehman, Abdul, et al.
Veröffentlicht: (2024)
Software Resource Disaggregation for HPC with Serverless Computing
von: Copik, Marcin, et al.
Veröffentlicht: (2024)
von: Copik, Marcin, et al.
Veröffentlicht: (2024)
Jiagu: Optimizing Serverless Computing Resource Utilization with Harmonized Efficiency and Practicability
von: Liu, Qingyuan, et al.
Veröffentlicht: (2024)
von: Liu, Qingyuan, et al.
Veröffentlicht: (2024)
SeaLLM: Service-Aware and Latency-Optimized Resource Sharing for Large Language Model Inference
von: Zhao, Yihao, et al.
Veröffentlicht: (2025)
von: Zhao, Yihao, et al.
Veröffentlicht: (2025)
Towards Fast Setup and High Throughput of GPU Serverless Computing
von: Zhao, Han, et al.
Veröffentlicht: (2024)
von: Zhao, Han, et al.
Veröffentlicht: (2024)
Towards Energy-Efficient Serverless Computing with Hardware Isolation
von: Carl, Natalie, et al.
Veröffentlicht: (2025)
von: Carl, Natalie, et al.
Veröffentlicht: (2025)
GreenWhisk: Emission-Aware Computing for Serverless Platform
von: Serenari, Jayden, et al.
Veröffentlicht: (2024)
von: Serenari, Jayden, et al.
Veröffentlicht: (2024)
Unleashing the Power of Tree-of-Thoughts for Edge-Enabled AIGC Service Provisioning
von: Liu, Zhang, et al.
Veröffentlicht: (2026)
von: Liu, Zhang, et al.
Veröffentlicht: (2026)
Cicada: A Pipeline-Efficient Approach to Serverless Inference with Decoupled Management
von: Wu, Z., et al.
Veröffentlicht: (2025)
von: Wu, Z., et al.
Veröffentlicht: (2025)
FaaSTube: Optimizing GPU-oriented Data Transfer for Serverless Computing
von: Wu, Hao, et al.
Veröffentlicht: (2024)
von: Wu, Hao, et al.
Veröffentlicht: (2024)
DistServe: Disaggregating Prefill and Decoding for Goodput-optimized Large Language Model Serving
von: Zhong, Yinmin, et al.
Veröffentlicht: (2024)
von: Zhong, Yinmin, et al.
Veröffentlicht: (2024)
ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block Decoding
von: Jin, ChiHeng, et al.
Veröffentlicht: (2026)
von: Jin, ChiHeng, et al.
Veröffentlicht: (2026)
Towards Seamless Serverless Computing Across an Edge-Cloud Continuum
von: Simion, Emilian, et al.
Veröffentlicht: (2024)
von: Simion, Emilian, et al.
Veröffentlicht: (2024)
HotSwap: Enabling Live Dependency Sharing in Serverless Computing
von: Li, Rui, et al.
Veröffentlicht: (2024)
von: Li, Rui, et al.
Veröffentlicht: (2024)
Leveraging Core and Uncore Frequency Scaling for Power-Efficient Serverless Workflows
von: Tzenetopoulos, Achilleas, et al.
Veröffentlicht: (2024)
von: Tzenetopoulos, Achilleas, et al.
Veröffentlicht: (2024)
WarmServe: Enabling One-for-Many GPU Prewarming for Multi-LLM Serving
von: Lou, Chiheng, et al.
Veröffentlicht: (2025)
von: Lou, Chiheng, et al.
Veröffentlicht: (2025)
Scale: Deep Reinforcement Learning for Container Scheduling in Serverless Edge Computing
von: Chen, Chen, et al.
Veröffentlicht: (2026)
von: Chen, Chen, et al.
Veröffentlicht: (2026)
Tutorial: Object as a Service (OaaS) Serverless Cloud Computing Paradigm
von: Lertpongrujikorn, Pawissanutt, et al.
Veröffentlicht: (2024)
von: Lertpongrujikorn, Pawissanutt, et al.
Veröffentlicht: (2024)
Silent Failures in Stateless Systems: Rethinking Anomaly Detection for Serverless Computing
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Chanh, et al.
Veröffentlicht: (2025)
Trabant: A Serverless Architecture for Multi-Tenant Orbital Edge Computing
von: Pfandzelter, Tobias, et al.
Veröffentlicht: (2025)
von: Pfandzelter, Tobias, et al.
Veröffentlicht: (2025)
EcoLife: Carbon-Aware Serverless Function Scheduling for Sustainable Computing
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
Hiku: Pull-Based Scheduling for Serverless Computing
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
von: Akbari, Saman, et al.
Veröffentlicht: (2025)
Code once, Run Green: Automated Green Code Translation in Serverless Computing
von: Werner, Sebastian, et al.
Veröffentlicht: (2025)
von: Werner, Sebastian, et al.
Veröffentlicht: (2025)
Cosmos: A Cost Model for Serverless Workflows in the 3D Compute Continuum
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
Gaia: Hybrid Hardware Acceleration for Serverless AI in the 3D Compute Continuum
von: Reisecker, Maximilian, et al.
Veröffentlicht: (2025)
von: Reisecker, Maximilian, et al.
Veröffentlicht: (2025)
Torpor: GPU-Enabled Serverless Computing for Low-Latency, Resource-Efficient Inference
von: Yu, Minchen, et al.
Veröffentlicht: (2023)
von: Yu, Minchen, et al.
Veröffentlicht: (2023)
Energy Efficiency Support for Software Defined Networks: a Serverless Computing Approach
von: Banaie, Fatemeh, et al.
Veröffentlicht: (2024)
von: Banaie, Fatemeh, et al.
Veröffentlicht: (2024)
Unleashing Collaborative Computing for Adaptive Video Streaming with Multi-objective Optimization in Satellite Terrestrial Networks
von: Shen, Zhishu, et al.
Veröffentlicht: (2024)
von: Shen, Zhishu, et al.
Veröffentlicht: (2024)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
von: Qi, S., et al.
Veröffentlicht: (2024)
von: Qi, S., et al.
Veröffentlicht: (2024)
Cold Start Latency in Serverless Computing: A Systematic Review, Taxonomy, and Future Directions
von: Golec, Muhammed, et al.
Veröffentlicht: (2023)
von: Golec, Muhammed, et al.
Veröffentlicht: (2023)
Databelt: A Continuous Data Path for Serverless Workflows in the 3D Compute Continuum
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HydraServe: Minimizing Cold Start Latency for Serverless LLM Serving in Public Clouds
von: Lou, Chiheng, et al.
Veröffentlicht: (2025) -
Litmus: Fair Pricing for Serverless Computing
von: Pei, Qi, et al.
Veröffentlicht: (2024) -
Caching Aided Multi-Tenant Serverless Computing
von: Qiao, Chu, et al.
Veröffentlicht: (2024) -
Making Serverless Computing Extensible: A Case Study of Serverless Data Analytics
von: Yu, Minchen, et al.
Veröffentlicht: (2025) -
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
von: Chang, Zihan, et al.
Veröffentlicht: (2024)