DRPC: Distributed Reinforcement Learning Approach for Scalable Resource Provisioning in Container-based Clusters
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bai, Haoyu, Xu, Minxian, Ye, Kejiang, Buyya, Rajkumar, Xu, Chengzhong |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
par: Xu, Minxian, et autres
Publié: (2025)
par: Xu, Minxian, et autres
Publié: (2025)
TD3-Sched: Learning to Orchestrate Container-based Cloud-Edge Resources via Distributed Reinforcement Learning
par: Song, Shengye, et autres
Publié: (2025)
par: Song, Shengye, et autres
Publié: (2025)
DOPD: A Dynamic PD-Disaggregation Architecture for Maximizing Goodput in LLM Inference Serving
par: Liao, Junhan, et autres
Publié: (2025)
par: Liao, Junhan, et autres
Publié: (2025)
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
par: Tang, Lujie, et autres
Publié: (2024)
par: Tang, Lujie, et autres
Publié: (2024)
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
par: Hu, Kan, et autres
Publié: (2024)
par: Hu, Kan, et autres
Publié: (2024)
BrownoutServe: SLO-Aware Inference Serving under Bursty Workloads for MoE-based LLMs
par: Hu, Jianmin, et autres
Publié: (2025)
par: Hu, Jianmin, et autres
Publié: (2025)
ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
par: Bai, Haoyu, et autres
Publié: (2026)
par: Bai, Haoyu, et autres
Publié: (2026)
An Interference-aware Approach for Co-located Container Orchestration with Novel Metric
par: Li, Xiang, et autres
Publié: (2024)
par: Li, Xiang, et autres
Publié: (2024)
MSARS: A Meta-Learning and Reinforcement Learning Framework for SLO Resource Allocation and Adaptive Scaling for Microservices
par: Hu, Kan, et autres
Publié: (2024)
par: Hu, Kan, et autres
Publié: (2024)
UELLM: A Unified and Efficient Approach for LLM Inference Serving
par: He, Yiyuan, et autres
Publié: (2024)
par: He, Yiyuan, et autres
Publié: (2024)
CloudNativeSim: a toolkit for modeling and simulation of cloud-native applications
par: Wu, Jingfeng, et autres
Publié: (2024)
par: Wu, Jingfeng, et autres
Publié: (2024)
SealOS+: A Sealos-based Approach for Adaptive Resource Optimization Under Dynamic Workloads for Securities Trading System
par: Jia, Haojie, et autres
Publié: (2025)
par: Jia, Haojie, et autres
Publié: (2025)
Cloud Native System for LLM Inference Serving
par: Xu, Minxian, et autres
Publié: (2025)
par: Xu, Minxian, et autres
Publié: (2025)
Unlock the Potential of Fine-grained LLM Serving via Dynamic Module Scaling
par: Wu, Jingfeng, et autres
Publié: (2025)
par: Wu, Jingfeng, et autres
Publié: (2025)
C-Koordinator: Interference-aware Management for Large-scale and Co-located Microservice Clusters
par: Song, Shengye, et autres
Publié: (2025)
par: Song, Shengye, et autres
Publié: (2025)
TempoScale: A Cloud Workloads Prediction Approach Integrating Short-Term and Long-Term Information
par: Wen, Linfeng, et autres
Publié: (2024)
par: Wen, Linfeng, et autres
Publié: (2024)
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda
par: Xu, Minxian, et autres
Publié: (2026)
par: Xu, Minxian, et autres
Publié: (2026)
Multi-Layer Scheduling for MoE-Based LLM Reasoning
par: Sun, Yifan, et autres
Publié: (2026)
par: Sun, Yifan, et autres
Publié: (2026)
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
par: Bai, Xu, et autres
Publié: (2025)
par: Bai, Xu, et autres
Publié: (2025)
ReinFog: A Deep Reinforcement Learning Empowered Framework for Resource Management in Edge and Cloud Computing Environments
par: Wang, Zhiyu, et autres
Publié: (2024)
par: Wang, Zhiyu, et autres
Publié: (2024)
EnFed: An Energy-aware Federated Learning in Resource Constrained Environments for Human Activity Recognition
par: Mukherjee, Anwesha, et autres
Publié: (2024)
par: Mukherjee, Anwesha, et autres
Publié: (2024)
A Deep Reinforcement Learning Approach for Cost Optimized Workflow Scheduling in Cloud Computing Environments
par: Jayanetti, Amanda, et autres
Publié: (2024)
par: Jayanetti, Amanda, et autres
Publié: (2024)
FlexPipe: Adapting Dynamic LLM Serving Through Inflight Pipeline Refactoring in Fragmented Serverless Clusters
par: Lin, Yanying, et autres
Publié: (2025)
par: Lin, Yanying, et autres
Publié: (2025)
HGraphScale: Hierarchical Graph Learning for Autoscaling Microservice Applications in Container-based Cloud Computing
par: Fang, Zhengxin, et autres
Publié: (2025)
par: Fang, Zhengxin, et autres
Publié: (2025)
StatuScale: Status-aware and Elastic Scaling Strategy for Microservice Applications
par: Wen, Linfeng, et autres
Publié: (2024)
par: Wen, Linfeng, et autres
Publié: (2024)
Reinforcement Learning based Workflow Scheduling in Cloud and Edge Computing Environments: A Taxonomy, Review and Future Directions
par: Jayanetti, Amanda, et autres
Publié: (2024)
par: Jayanetti, Amanda, et autres
Publié: (2024)
Deep Reinforcement Learning-based Methods for Resource Scheduling in Cloud Computing: A Review and Future Directions
par: Zhou, Guangyao, et autres
Publié: (2021)
par: Zhou, Guangyao, et autres
Publié: (2021)
A Joint Time and Energy-Efficient Federated Learning-based Computation Offloading Method for Mobile Edge Computing
par: Mukherjee, Anwesha, et autres
Publié: (2024)
par: Mukherjee, Anwesha, et autres
Publié: (2024)
TrustMesh: A Blockchain-Enabled Trusted Distributed Computing Framework for Open Heterogeneous IoT Environments
par: Rangwala, Murtaza, et autres
Publié: (2024)
par: Rangwala, Murtaza, et autres
Publié: (2024)
Generative Federated Learning for Smart Prediction and Recommendation Applications
par: Mukherjee, Anwesha, et autres
Publié: (2025)
par: Mukherjee, Anwesha, et autres
Publié: (2025)
BanaServe: Unified KV Cache and Dynamic Module Migration for Balancing Disaggregated LLM Serving in AI Infrastructure
par: He, Yiyuan, et autres
Publié: (2025)
par: He, Yiyuan, et autres
Publié: (2025)
BucketServe: Bucket-Based Dynamic Batching for Smart and Efficient LLM Inference Serving
par: Zheng, Wanyi, et autres
Publié: (2025)
par: Zheng, Wanyi, et autres
Publié: (2025)
A System Aware Resource Allocation for Distributed Workflows in Quantum Computing Environments
par: Sawaika, Abhishek, et autres
Publié: (2026)
par: Sawaika, Abhishek, et autres
Publié: (2026)
DRLQ: A Deep Reinforcement Learning-based Task Placement for Quantum Cloud Computing
par: Nguyen, Hoa T., et autres
Publié: (2024)
par: Nguyen, Hoa T., et autres
Publié: (2024)
TF-DDRL: A Transformer-enhanced Distributed DRL Technique for Scheduling IoT Applications in Edge and Cloud Computing Environments
par: Wang, Zhiyu, et autres
Publié: (2024)
par: Wang, Zhiyu, et autres
Publié: (2024)
A Knowledge Distillation-empowered Adaptive Federated Reinforcement Learning Framework for Multi-Domain IoT Applications Scheduling
par: Wang, Zhiyu, et autres
Publié: (2025)
par: Wang, Zhiyu, et autres
Publié: (2025)
Deep Reinforcement Learning (DRL)-based Methods for Serverless Stream Processing Engines: A Vision, Architectural Elements, and Future Directions
par: Read, Maria R., et autres
Publié: (2024)
par: Read, Maria R., et autres
Publié: (2024)
Saarthi: An End-to-End Intelligent Platform for Optimising Distributed Serverless Workloads
par: Agarwal, Siddharth, et autres
Publié: (2025)
par: Agarwal, Siddharth, et autres
Publié: (2025)
Placement of Microservices-based IoT Applications in Fog Computing: A Taxonomy and Future Directions
par: Pallewatta, Samodha, et autres
Publié: (2022)
par: Pallewatta, Samodha, et autres
Publié: (2022)
A Decentralized Root Cause Localization Approach for Edge Computing Environments
par: Fernando, Duneesha, et autres
Publié: (2025)
par: Fernando, Duneesha, et autres
Publié: (2025)
Documents similaires
-
Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
par: Xu, Minxian, et autres
Publié: (2025) -
TD3-Sched: Learning to Orchestrate Container-based Cloud-Edge Resources via Distributed Reinforcement Learning
par: Song, Shengye, et autres
Publié: (2025) -
DOPD: A Dynamic PD-Disaggregation Architecture for Maximizing Goodput in LLM Inference Serving
par: Liao, Junhan, et autres
Publié: (2025) -
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
par: Tang, Lujie, et autres
Publié: (2024) -
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
par: Hu, Kan, et autres
Publié: (2024)