SealOS+: A Sealos-based Approach for Adaptive Resource Optimization Under Dynamic Workloads for Securities Trading System
Fuente:
arXiv
Guardado en:
| Autores principales: | Jia, Haojie, Li, Zhenhao, Li, Gen, Xu, Minxian, Ye, Kejiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
por: Tang, Lujie, et al.
Publicado: (2024)
por: Tang, Lujie, et al.
Publicado: (2024)
TempoScale: A Cloud Workloads Prediction Approach Integrating Short-Term and Long-Term Information
por: Wen, Linfeng, et al.
Publicado: (2024)
por: Wen, Linfeng, et al.
Publicado: (2024)
BrownoutServe: SLO-Aware Inference Serving under Bursty Workloads for MoE-based LLMs
por: Hu, Jianmin, et al.
Publicado: (2025)
por: Hu, Jianmin, et al.
Publicado: (2025)
MSARS: A Meta-Learning and Reinforcement Learning Framework for SLO Resource Allocation and Adaptive Scaling for Microservices
por: Hu, Kan, et al.
Publicado: (2024)
por: Hu, Kan, et al.
Publicado: (2024)
An Interference-aware Approach for Co-located Container Orchestration with Novel Metric
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
DRPC: Distributed Reinforcement Learning Approach for Scalable Resource Provisioning in Container-based Clusters
por: Bai, Haoyu, et al.
Publicado: (2024)
por: Bai, Haoyu, et al.
Publicado: (2024)
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
por: Hu, Kan, et al.
Publicado: (2024)
por: Hu, Kan, et al.
Publicado: (2024)
TD3-Sched: Learning to Orchestrate Container-based Cloud-Edge Resources via Distributed Reinforcement Learning
por: Song, Shengye, et al.
Publicado: (2025)
por: Song, Shengye, et al.
Publicado: (2025)
UELLM: A Unified and Efficient Approach for LLM Inference Serving
por: He, Yiyuan, et al.
Publicado: (2024)
por: He, Yiyuan, et al.
Publicado: (2024)
BucketServe: Bucket-Based Dynamic Batching for Smart and Efficient LLM Inference Serving
por: Zheng, Wanyi, et al.
Publicado: (2025)
por: Zheng, Wanyi, et al.
Publicado: (2025)
Cloud Native System for LLM Inference Serving
por: Xu, Minxian, et al.
Publicado: (2025)
por: Xu, Minxian, et al.
Publicado: (2025)
Unlock the Potential of Fine-grained LLM Serving via Dynamic Module Scaling
por: Wu, Jingfeng, et al.
Publicado: (2025)
por: Wu, Jingfeng, et al.
Publicado: (2025)
CloudNativeSim: a toolkit for modeling and simulation of cloud-native applications
por: Wu, Jingfeng, et al.
Publicado: (2024)
por: Wu, Jingfeng, et al.
Publicado: (2024)
Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
por: Xu, Minxian, et al.
Publicado: (2025)
por: Xu, Minxian, et al.
Publicado: (2025)
DOPD: A Dynamic PD-Disaggregation Architecture for Maximizing Goodput in LLM Inference Serving
por: Liao, Junhan, et al.
Publicado: (2025)
por: Liao, Junhan, et al.
Publicado: (2025)
Resource Optimization with MPI Process Malleability for Dynamic Workloads in HPC Clusters
por: Iserte, Sergio, et al.
Publicado: (2025)
por: Iserte, Sergio, et al.
Publicado: (2025)
BanaServe: Unified KV Cache and Dynamic Module Migration for Balancing Disaggregated LLM Serving in AI Infrastructure
por: He, Yiyuan, et al.
Publicado: (2025)
por: He, Yiyuan, et al.
Publicado: (2025)
C-Koordinator: Interference-aware Management for Large-scale and Co-located Microservice Clusters
por: Song, Shengye, et al.
Publicado: (2025)
por: Song, Shengye, et al.
Publicado: (2025)
StatuScale: Status-aware and Elastic Scaling Strategy for Microservice Applications
por: Wen, Linfeng, et al.
Publicado: (2024)
por: Wen, Linfeng, et al.
Publicado: (2024)
ARC-V: Vertical Resource Adaptivity for HPC Workloads in Containerized Environments
por: Medeiros, Daniel, et al.
Publicado: (2025)
por: Medeiros, Daniel, et al.
Publicado: (2025)
Crossword: Adaptive Consensus for Dynamic Data-Heavy Workloads
por: Hu, Guanzhou, et al.
Publicado: (2025)
por: Hu, Guanzhou, et al.
Publicado: (2025)
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
FlexPipe: Adapting Dynamic LLM Serving Through Inflight Pipeline Refactoring in Fragmented Serverless Clusters
por: Lin, Yanying, et al.
Publicado: (2025)
por: Lin, Yanying, et al.
Publicado: (2025)
ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
por: Bai, Haoyu, et al.
Publicado: (2026)
por: Bai, Haoyu, et al.
Publicado: (2026)
Dynamic Client Clustering, Bandwidth Allocation, and Workload Optimization for Semi-synchronous Federated Learning
por: Yu, Liangkun, et al.
Publicado: (2024)
por: Yu, Liangkun, et al.
Publicado: (2024)
PRISM: Dynamic Primitive-Based Forecasting for Large-Scale GPU Cluster Workloads
por: Wu, Xin, et al.
Publicado: (2026)
por: Wu, Xin, et al.
Publicado: (2026)
A Review of Tools and Techniques for Optimization of Workload Mapping and Scheduling in Heterogeneous HPC System
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
Workload Buoyancy: Keeping Apps Afloat by Identifying Shared Resource Bottlenecks
por: Larsson, Oliver, et al.
Publicado: (2026)
por: Larsson, Oliver, et al.
Publicado: (2026)
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda
por: Xu, Minxian, et al.
Publicado: (2026)
por: Xu, Minxian, et al.
Publicado: (2026)
GENSERVE: Efficient Co-Serving of Heterogeneous Diffusion Model Workloads
por: Ye, Fanjiang, et al.
Publicado: (2026)
por: Ye, Fanjiang, et al.
Publicado: (2026)
BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems
por: Wang, Yuxin, et al.
Publicado: (2024)
por: Wang, Yuxin, et al.
Publicado: (2024)
Data Management System Analysis for Distributed Computing Workloads
por: Hsu, Kuan-Chieh, et al.
Publicado: (2025)
por: Hsu, Kuan-Chieh, et al.
Publicado: (2025)
Orchestrating Mixed-Criticality Cloud Workloads in Reconfigurable Manufacturing Systems
por: Barletta, Marco, et al.
Publicado: (2024)
por: Barletta, Marco, et al.
Publicado: (2024)
TierBase: A Workload-Driven Cost-Optimized Key-Value Store
por: Shen, Zhitao, et al.
Publicado: (2025)
por: Shen, Zhitao, et al.
Publicado: (2025)
DynaShard: Secure and Adaptive Blockchain Sharding Protocol with Hybrid Consensus and Dynamic Shard Management
por: Liu, Ao, et al.
Publicado: (2024)
por: Liu, Ao, et al.
Publicado: (2024)
Profiling and Modeling of Power Characteristics of Leadership-Scale HPC System Workloads
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
Characterizing Production GPU Workloads using System-wide Telemetry Data
por: Cankur, Onur, et al.
Publicado: (2025)
por: Cankur, Onur, et al.
Publicado: (2025)
Shift Parallelism: Low-Latency, High-Throughput LLM Inference for Dynamic Workloads
por: Hidayetoglu, Mert, et al.
Publicado: (2025)
por: Hidayetoglu, Mert, et al.
Publicado: (2025)
Autopoiesis: A Self-Evolving System Paradigm for LLM Serving Under Runtime Dynamics
por: Jiang, Youhe, et al.
Publicado: (2026)
por: Jiang, Youhe, et al.
Publicado: (2026)
Edge AI: A Taxonomy, Systematic Review and Future Directions
por: Gill, Sukhpal Singh, et al.
Publicado: (2024)
por: Gill, Sukhpal Singh, et al.
Publicado: (2024)
Ejemplares similares
-
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
por: Tang, Lujie, et al.
Publicado: (2024) -
TempoScale: A Cloud Workloads Prediction Approach Integrating Short-Term and Long-Term Information
por: Wen, Linfeng, et al.
Publicado: (2024) -
BrownoutServe: SLO-Aware Inference Serving under Bursty Workloads for MoE-based LLMs
por: Hu, Jianmin, et al.
Publicado: (2025) -
MSARS: A Meta-Learning and Reinforcement Learning Framework for SLO Resource Allocation and Adaptive Scaling for Microservices
por: Hu, Kan, et al.
Publicado: (2024) -
An Interference-aware Approach for Co-located Container Orchestration with Novel Metric
por: Li, Xiang, et al.
Publicado: (2024)