FaaSTube: Optimizing GPU-oriented Data Transfer for Serverless Computing
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Hao, Deng, Junxiao, Yu, Minchen, Yu, Yue, Liu, Yaochen, Fan, Hao, Wu, Song, Wang, Wei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Making Serverless Computing Extensible: A Case Study of Serverless Data Analytics
di: Yu, Minchen, et al.
Pubblicazione: (2025)
di: Yu, Minchen, et al.
Pubblicazione: (2025)
Torpor: GPU-Enabled Serverless Computing for Low-Latency, Resource-Efficient Inference
di: Yu, Minchen, et al.
Pubblicazione: (2023)
di: Yu, Minchen, et al.
Pubblicazione: (2023)
FaasMeter: Energy-First Serverless Computing
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
Joint$λ$: Orchestrating Serverless Workflows on Jointcloud FaaS Systems
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
ProFaaStinate: Delaying Serverless Function Calls to Optimize Platform Performance
di: Schirmer, Trever, et al.
Pubblicazione: (2023)
di: Schirmer, Trever, et al.
Pubblicazione: (2023)
FaaS Is Not Enough: Serverless Handling of Burst-Parallel Jobs
di: Barcelona-Pons, Daniel, et al.
Pubblicazione: (2024)
di: Barcelona-Pons, Daniel, et al.
Pubblicazione: (2024)
gFaaS: Enabling Generic Functions in Serverless Computing
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
FaaSKeeper: Learning from Building Serverless Services with ZooKeeper as an Example
di: Copik, Marcin, et al.
Pubblicazione: (2022)
di: Copik, Marcin, et al.
Pubblicazione: (2022)
FaaSMT: Lightweight Serverless Framework for Intrusion Detection Using Merkle Tree and Task Inlining
di: Li, Chuang, et al.
Pubblicazione: (2025)
di: Li, Chuang, et al.
Pubblicazione: (2025)
Towards Fast Setup and High Throughput of GPU Serverless Computing
di: Zhao, Han, et al.
Pubblicazione: (2024)
di: Zhao, Han, et al.
Pubblicazione: (2024)
In Serverless, OS Scheduler Choice Costs Money: A Hybrid Scheduling Approach for Cheaper FaaS
di: Zhao, Yuxuan, et al.
Pubblicazione: (2024)
di: Zhao, Yuxuan, et al.
Pubblicazione: (2024)
λScale: Enabling Fast Scaling for Serverless Large Language Model Inference
di: Yu, Minchen, et al.
Pubblicazione: (2025)
di: Yu, Minchen, et al.
Pubblicazione: (2025)
Optimizing Distributed Deployment of Mixture-of-Experts Model Inference in Serverless Computing
di: Liu, Mengfan, et al.
Pubblicazione: (2025)
di: Liu, Mengfan, et al.
Pubblicazione: (2025)
RLHFless: Serverless Computing for Efficient RLHF
di: Wei, Rui, et al.
Pubblicazione: (2026)
di: Wei, Rui, et al.
Pubblicazione: (2026)
GeoFaaS: An Edge-to-Cloud FaaS Platform
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2024)
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2024)
It Takes Two to Tango: Serverless Workflow Serving via Bilaterally Engaged Resource Adaptation
di: Wu, Jing, et al.
Pubblicazione: (2025)
di: Wu, Jing, et al.
Pubblicazione: (2025)
GreenFaaS: Maximizing Energy Efficiency of HPC Workloads with FaaS
di: Kamatar, Alok, et al.
Pubblicazione: (2024)
di: Kamatar, Alok, et al.
Pubblicazione: (2024)
Jiagu: Optimizing Serverless Computing Resource Utilization with Harmonized Efficiency and Practicability
di: Liu, Qingyuan, et al.
Pubblicazione: (2024)
di: Liu, Qingyuan, et al.
Pubblicazione: (2024)
Optimizing FaaS Platforms for MCP-enabled Agentic Workflows
di: Kulkarni, Varad, et al.
Pubblicazione: (2026)
di: Kulkarni, Varad, et al.
Pubblicazione: (2026)
Application-Centric Benchmarking of Distributed FaaS Platforms using BeFaaS
di: Grambow, Martin, et al.
Pubblicazione: (2023)
di: Grambow, Martin, et al.
Pubblicazione: (2023)
ServerlessLoRA: Minimizing Latency and Cost in Serverless Inference for LoRA-Based LLMs
di: Sui, Yifan, et al.
Pubblicazione: (2025)
di: Sui, Yifan, et al.
Pubblicazione: (2025)
FaaSMoE: A Serverless Framework for Multi-Tenant Mixture-of-Experts Serving
di: Wang, Minghe, et al.
Pubblicazione: (2026)
di: Wang, Minghe, et al.
Pubblicazione: (2026)
EdgeFaaS: A Function-based Framework for Edge Computing
di: Jin, Runyu, et al.
Pubblicazione: (2022)
di: Jin, Runyu, et al.
Pubblicazione: (2022)
DeepServe: Serverless Large Language Model Serving at Scale
di: Hu, Junhao, et al.
Pubblicazione: (2025)
di: Hu, Junhao, et al.
Pubblicazione: (2025)
HAS-GPU: Efficient Hybrid Auto-scaling with Fine-grained GPU Allocation for SLO-aware Serverless Inferences
di: Gu, Jianfeng, et al.
Pubblicazione: (2025)
di: Gu, Jianfeng, et al.
Pubblicazione: (2025)
Multi-Event Triggers for Serverless Computing
di: Carl, Natalie, et al.
Pubblicazione: (2025)
di: Carl, Natalie, et al.
Pubblicazione: (2025)
Litmus: Fair Pricing for Serverless Computing
di: Pei, Qi, et al.
Pubblicazione: (2024)
di: Pei, Qi, et al.
Pubblicazione: (2024)
Serverless Computing: Architecture, Concepts, and Applications
di: Ghorbian, Mohsen, et al.
Pubblicazione: (2025)
di: Ghorbian, Mohsen, et al.
Pubblicazione: (2025)
Warp-STAR: High-performance, Differentiable GPU-Accelerated Static Timing Analysis through Warp-oriented Parallel Orchestration
di: Huang, En-Ming, et al.
Pubblicazione: (2026)
di: Huang, En-Ming, et al.
Pubblicazione: (2026)
Dilu: Enabling GPU Resourcing-on-Demand for Serverless DL Serving via Introspective Elasticity
di: Lv, Cunchi, et al.
Pubblicazione: (2025)
di: Lv, Cunchi, et al.
Pubblicazione: (2025)
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
di: Chang, Zihan, et al.
Pubblicazione: (2024)
di: Chang, Zihan, et al.
Pubblicazione: (2024)
Barrier-Augmented Lagrangian for GPU-based Elastodynamic Contact
di: Guo, Dewen, et al.
Pubblicazione: (2024)
di: Guo, Dewen, et al.
Pubblicazione: (2024)
KUBEDIRECT: Unleashing the Full Power of the Cluster Manager for Serverless Computing
di: Qi, Sheng, et al.
Pubblicazione: (2026)
di: Qi, Sheng, et al.
Pubblicazione: (2026)
Caching Aided Multi-Tenant Serverless Computing
di: Qiao, Chu, et al.
Pubblicazione: (2024)
di: Qiao, Chu, et al.
Pubblicazione: (2024)
Software Resource Disaggregation for HPC with Serverless Computing
di: Copik, Marcin, et al.
Pubblicazione: (2024)
di: Copik, Marcin, et al.
Pubblicazione: (2024)
Cicada: A Pipeline-Efficient Approach to Serverless Inference with Decoupled Management
di: Wu, Z., et al.
Pubblicazione: (2025)
di: Wu, Z., et al.
Pubblicazione: (2025)
Databelt: A Continuous Data Path for Serverless Workflows in the 3D Compute Continuum
di: Marcelino, Cynthia, et al.
Pubblicazione: (2025)
di: Marcelino, Cynthia, et al.
Pubblicazione: (2025)
Konflux: Optimized Function Fusion for Serverless Applications
di: Kowallik, Niklas, et al.
Pubblicazione: (2026)
di: Kowallik, Niklas, et al.
Pubblicazione: (2026)
Junctiond: Extending FaaS Runtimes with Kernel-Bypass
di: Saurez, Enrique, et al.
Pubblicazione: (2024)
di: Saurez, Enrique, et al.
Pubblicazione: (2024)
Towards a Testbed for Scalable FaaS Platforms
di: Schirmer, Trever, et al.
Pubblicazione: (2025)
di: Schirmer, Trever, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Making Serverless Computing Extensible: A Case Study of Serverless Data Analytics
di: Yu, Minchen, et al.
Pubblicazione: (2025) -
Torpor: GPU-Enabled Serverless Computing for Low-Latency, Resource-Efficient Inference
di: Yu, Minchen, et al.
Pubblicazione: (2023) -
FaasMeter: Energy-First Serverless Computing
di: Rehman, Abdul, et al.
Pubblicazione: (2024) -
Joint$λ$: Orchestrating Serverless Workflows on Jointcloud FaaS Systems
di: Li, Rui, et al.
Pubblicazione: (2025) -
ProFaaStinate: Delaying Serverless Function Calls to Optimize Platform Performance
di: Schirmer, Trever, et al.
Pubblicazione: (2023)