Shabari: Delayed Decision-Making for Faster and Efficient Serverless Functions
Fuente:
arXiv
Salvato in:
| Autori principali: | Sinha, Prasoon, Kaffes, Kostis, Yadwadkar, Neeraja J. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cortex: Workflow-Aware Resource Pooling and Scheduling for Agentic Serving
di: Pagonas, Nikos, et al.
Pubblicazione: (2025)
di: Pagonas, Nikos, et al.
Pubblicazione: (2025)
Scalable and Cost-Efficient ML Inference: Parallel Batch Processing with Serverless Functions
di: Barrak, Amine, et al.
Pubblicazione: (2025)
di: Barrak, Amine, et al.
Pubblicazione: (2025)
Apodotiko: Enabling Efficient Serverless Federated Learning in Heterogeneous Environments
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
Toward Systems Foundations for Agentic Exploration
di: Xu, Jiakai, et al.
Pubblicazione: (2025)
di: Xu, Jiakai, et al.
Pubblicazione: (2025)
ServerlessLLM: Low-Latency Serverless Inference for Large Language Models
di: Fu, Yao, et al.
Pubblicazione: (2024)
di: Fu, Yao, et al.
Pubblicazione: (2024)
Speculative Actions: A Lossless Framework for Faster Agentic Systems
di: Ye, Naimeng, et al.
Pubblicazione: (2025)
di: Ye, Naimeng, et al.
Pubblicazione: (2025)
Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud
di: Ghosh, Himel
Pubblicazione: (2024)
di: Ghosh, Himel
Pubblicazione: (2024)
ServerlessLoRA: Minimizing Latency and Cost in Serverless Inference for LoRA-Based LLMs
di: Sui, Yifan, et al.
Pubblicazione: (2025)
di: Sui, Yifan, et al.
Pubblicazione: (2025)
High-Performance Serverless Computing: A Systematic Literature Review on Serverless for HPC, AI, and Big Data
di: Besozzi, Valerio, et al.
Pubblicazione: (2026)
di: Besozzi, Valerio, et al.
Pubblicazione: (2026)
LIFL: A Lightweight, Event-driven Serverless Platform for Federated Learning
di: Qi, Shixiong, et al.
Pubblicazione: (2024)
di: Qi, Shixiong, et al.
Pubblicazione: (2024)
Optimizing Distributed Deployment of Mixture-of-Experts Model Inference in Serverless Computing
di: Liu, Mengfan, et al.
Pubblicazione: (2025)
di: Liu, Mengfan, et al.
Pubblicazione: (2025)
Pursuing Overall Welfare in Federated Learning through Sequential Decision Making
di: Hahn, Seok-Ju, et al.
Pubblicazione: (2024)
di: Hahn, Seok-Ju, et al.
Pubblicazione: (2024)
ProFaaStinate: Delaying Serverless Function Calls to Optimize Platform Performance
di: Schirmer, Trever, et al.
Pubblicazione: (2023)
di: Schirmer, Trever, et al.
Pubblicazione: (2023)
Reinforcement Learning-Based Dynamic Management of Structured Parallel Farm Skeletons on Serverless Platforms
di: Li, Lanpei, et al.
Pubblicazione: (2026)
di: Li, Lanpei, et al.
Pubblicazione: (2026)
FaaSMoE: A Serverless Framework for Multi-Tenant Mixture-of-Experts Serving
di: Wang, Minghe, et al.
Pubblicazione: (2026)
di: Wang, Minghe, et al.
Pubblicazione: (2026)
Making Serverless Computing Extensible: A Case Study of Serverless Data Analytics
di: Yu, Minchen, et al.
Pubblicazione: (2025)
di: Yu, Minchen, et al.
Pubblicazione: (2025)
Metronome: Differentiated Delay Scheduling for Serverless Functions
di: Chen, Zhuangbin, et al.
Pubblicazione: (2025)
di: Chen, Zhuangbin, et al.
Pubblicazione: (2025)
Towards Seamless Hierarchical Federated Learning under Intermittent Client Participation: A Stagewise Decision-Making Methodology
di: Wu, Minghong, et al.
Pubblicazione: (2025)
di: Wu, Minghong, et al.
Pubblicazione: (2025)
FedFetch: Faster Federated Learning with Adaptive Downstream Prefetching
di: Yan, Qifan, et al.
Pubblicazione: (2025)
di: Yan, Qifan, et al.
Pubblicazione: (2025)
MoEless: Efficient MoE LLM Serving via Serverless Computing
di: Yu, Hanfei, et al.
Pubblicazione: (2026)
di: Yu, Hanfei, et al.
Pubblicazione: (2026)
Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection
di: Gupta, Vima, et al.
Pubblicazione: (2024)
di: Gupta, Vima, et al.
Pubblicazione: (2024)
Converge Faster, Talk Less: Hessian-Informed Federated Zeroth-Order Optimization
di: Li, Zhe, et al.
Pubblicazione: (2025)
di: Li, Zhe, et al.
Pubblicazione: (2025)
Hybrid Decentralized Optimization: Leveraging Both First- and Zeroth-Order Optimizers for Faster Convergence
di: Ansaripour, Matin, et al.
Pubblicazione: (2022)
di: Ansaripour, Matin, et al.
Pubblicazione: (2022)
Faster Distributed Inference-Only Recommender Systems via Bounded Lag Synchronous Collectives
di: Dichev, Kiril, et al.
Pubblicazione: (2025)
di: Dichev, Kiril, et al.
Pubblicazione: (2025)
VineLM: Trie-Based Fine-Grained Control for Agentic Workflows
di: Pagonas, Nikos, et al.
Pubblicazione: (2026)
di: Pagonas, Nikos, et al.
Pubblicazione: (2026)
Raptor: Distributed Scheduling for Serverless Functions
di: Exton, Kevin, et al.
Pubblicazione: (2024)
di: Exton, Kevin, et al.
Pubblicazione: (2024)
Affinity-aware Serverless Function Scheduling
di: De Palma, Giuseppe, et al.
Pubblicazione: (2024)
di: De Palma, Giuseppe, et al.
Pubblicazione: (2024)
Energy Efficient Scheduling for Serverless Systems
di: Tsenos, Michail, et al.
Pubblicazione: (2024)
di: Tsenos, Michail, et al.
Pubblicazione: (2024)
Orchestrating the Execution of Serverless Functions in Hybrid Clouds
di: Peri, Aristotelis, et al.
Pubblicazione: (2024)
di: Peri, Aristotelis, et al.
Pubblicazione: (2024)
On the Complexity of Reachability Properties in Serverless Function Scheduling
di: De Palma, Giuseppe, et al.
Pubblicazione: (2024)
di: De Palma, Giuseppe, et al.
Pubblicazione: (2024)
Konflux: Optimized Function Fusion for Serverless Applications
di: Kowallik, Niklas, et al.
Pubblicazione: (2026)
di: Kowallik, Niklas, et al.
Pubblicazione: (2026)
Asynchronous Federated Stochastic Optimization for Heterogeneous Objectives Under Arbitrary Delays
di: Iakovidou, Charikleia, et al.
Pubblicazione: (2024)
di: Iakovidou, Charikleia, et al.
Pubblicazione: (2024)
Zenix: Efficient Execution of Bulky Serverless Applications
di: Guo, Zhiyuan, et al.
Pubblicazione: (2022)
di: Guo, Zhiyuan, et al.
Pubblicazione: (2022)
Dependency-aware Resource Allocation for Serverless Functions at the Edge
di: Baresi, Luciano, et al.
Pubblicazione: (2023)
di: Baresi, Luciano, et al.
Pubblicazione: (2023)
Distributed Stochastic Gradient Descent with Staleness: A Stochastic Delay Differential Equation Based Framework
di: Yu, Siyuan, et al.
Pubblicazione: (2024)
di: Yu, Siyuan, et al.
Pubblicazione: (2024)
FSD-Inference: Fully Serverless Distributed Inference with Scalable Cloud Communication
di: Oakley, Joe, et al.
Pubblicazione: (2024)
di: Oakley, Joe, et al.
Pubblicazione: (2024)
Towards Energy-Efficient Serverless Computing with Hardware Isolation
di: Carl, Natalie, et al.
Pubblicazione: (2025)
di: Carl, Natalie, et al.
Pubblicazione: (2025)
Towards Resource-Efficient Serverless LLM Inference with SLINFER
di: Xu, Chuhao, et al.
Pubblicazione: (2025)
di: Xu, Chuhao, et al.
Pubblicazione: (2025)
Making MoE-based LLM Inference Resilient with Tarragon
di: Zhang, Songyu, et al.
Pubblicazione: (2026)
di: Zhang, Songyu, et al.
Pubblicazione: (2026)
Training Heterogeneous Client Models using Knowledge Distillation in Serverless Federated Learning
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
di: Chadha, Mohak, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cortex: Workflow-Aware Resource Pooling and Scheduling for Agentic Serving
di: Pagonas, Nikos, et al.
Pubblicazione: (2025) -
Scalable and Cost-Efficient ML Inference: Parallel Batch Processing with Serverless Functions
di: Barrak, Amine, et al.
Pubblicazione: (2025) -
Apodotiko: Enabling Efficient Serverless Federated Learning in Heterogeneous Environments
di: Chadha, Mohak, et al.
Pubblicazione: (2024) -
Toward Systems Foundations for Agentic Exploration
di: Xu, Jiakai, et al.
Pubblicazione: (2025) -
ServerlessLLM: Low-Latency Serverless Inference for Large Language Models
di: Fu, Yao, et al.
Pubblicazione: (2024)