Mutiny! How does Kubernetes fail, and what can we do about it?
Fuente:
arXiv
Salvato in:
| Autori principali: | Barletta, Marco, Cinque, Marcello, Di Martino, Catello, Kalbarczyk, Zbigniew T., Iyer, Ravishankar K. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Orchestrating Mixed-Criticality Cloud Workloads in Reconfigurable Manufacturing Systems
di: Barletta, Marco, et al.
Pubblicazione: (2024)
di: Barletta, Marco, et al.
Pubblicazione: (2024)
INDIGO: Page Migration for Hardware Memory Disaggregation Across a Network
di: Patke, Archit, et al.
Pubblicazione: (2025)
di: Patke, Archit, et al.
Pubblicazione: (2025)
Hierarchical Autoscaling for Large Language Model Serving with Chiron
di: Patke, Archit, et al.
Pubblicazione: (2025)
di: Patke, Archit, et al.
Pubblicazione: (2025)
Story of Two GPUs: Characterizing the Resilience of Hopper H100 and Ampere A100 GPUs
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
Queue management for slo-oriented large language model serving
di: Patke, Archit, et al.
Pubblicazione: (2024)
di: Patke, Archit, et al.
Pubblicazione: (2024)
Efficient Interactive LLM Serving with Proxy Model-based Sequence Length Prediction
di: Qiu, Haoran, et al.
Pubblicazione: (2024)
di: Qiu, Haoran, et al.
Pubblicazione: (2024)
Enhancing Kubernetes Resilience through Anomaly Detection and Prediction
di: Anemogiannis, V., et al.
Pubblicazione: (2025)
di: Anemogiannis, V., et al.
Pubblicazione: (2025)
Kubernetes in Action: Exploring the Performance of Kubernetes Distributions in the Cloud
di: Aqasizade, Hossein, et al.
Pubblicazione: (2024)
di: Aqasizade, Hossein, et al.
Pubblicazione: (2024)
Adaptive Resource Allocation for Workflow Containerization on Kubernetes
di: Shan, Chenggang, et al.
Pubblicazione: (2023)
di: Shan, Chenggang, et al.
Pubblicazione: (2023)
Signalling Health for Improved Kubernetes Microservice Availability
di: Roberts, Jacob, et al.
Pubblicazione: (2025)
di: Roberts, Jacob, et al.
Pubblicazione: (2025)
A Contention-Free Model for Converged Kubernetes on HPC
di: Sochat, Vanessa, et al.
Pubblicazione: (2024)
di: Sochat, Vanessa, et al.
Pubblicazione: (2024)
Running Cloud-native Workloads on HPC with High-Performance Kubernetes
di: Chazapis, Antony, et al.
Pubblicazione: (2024)
di: Chazapis, Antony, et al.
Pubblicazione: (2024)
The National Research Platform: Stretched, Multi-Tenant, Scientific Kubernetes Cluster
di: Weitzel, Derek, et al.
Pubblicazione: (2025)
di: Weitzel, Derek, et al.
Pubblicazione: (2025)
QONNECT: A QoS-Aware Orchestration System for Distributed Kubernetes Clusters
di: Aslan, Haci Ismail, et al.
Pubblicazione: (2025)
di: Aslan, Haci Ismail, et al.
Pubblicazione: (2025)
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
A User-centric Kubernetes-based Architecture for Green Cloud Computing
di: Zanotto, Matteo, et al.
Pubblicazione: (2025)
di: Zanotto, Matteo, et al.
Pubblicazione: (2025)
Resilience Evaluation of Kubernetes in Cloud-Edge Environments via Failure Injection
di: Chen, Zihao, et al.
Pubblicazione: (2025)
di: Chen, Zihao, et al.
Pubblicazione: (2025)
Container late-binding in unprivileged dHTC pilot systems on Kubernetes resources
di: Sfiligoi, Igor, et al.
Pubblicazione: (2025)
di: Sfiligoi, Igor, et al.
Pubblicazione: (2025)
A Kubernetes custom scheduler based on reinforcement learning for compute-intensive pods
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
di: Zhang, Guilin, et al.
Pubblicazione: (2025)
di: Zhang, Guilin, et al.
Pubblicazione: (2025)
Kubernetes in the Cloud vs. Bare Metal: A Comparative Study of Network Costs
di: Redoli, Rodrigo Mompo, et al.
Pubblicazione: (2025)
di: Redoli, Rodrigo Mompo, et al.
Pubblicazione: (2025)
Priority Matters: Optimising Kubernetes Clusters Usage with Constraint-Based Pod Packing
di: Christensen, Henrik Daniel, et al.
Pubblicazione: (2025)
di: Christensen, Henrik Daniel, et al.
Pubblicazione: (2025)
Comparative Analysis of Lightweight Kubernetes Distributions for Edge Computing: Performance and Resource Efficiency
di: Yakubov, Diyaz, et al.
Pubblicazione: (2025)
di: Yakubov, Diyaz, et al.
Pubblicazione: (2025)
KIS-S: A GPU-Aware Kubernetes Inference Simulator with RL-Based Auto-Scaling
di: Zhang, Guilin, et al.
Pubblicazione: (2025)
di: Zhang, Guilin, et al.
Pubblicazione: (2025)
KubeIntellect: A Modular LLM-Orchestrated Agent Framework for End-to-End Kubernetes Management
di: Ardebili, Mohsen Seyedkazemi, et al.
Pubblicazione: (2025)
di: Ardebili, Mohsen Seyedkazemi, et al.
Pubblicazione: (2025)
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
di: Kim, Taeyoon, et al.
Pubblicazione: (2026)
di: Kim, Taeyoon, et al.
Pubblicazione: (2026)
NL-CPS: Reinforcement Learning-Based Kubernetes Control Plane Placement in Multi-Region Clusters
di: Alam, Sajid, et al.
Pubblicazione: (2026)
di: Alam, Sajid, et al.
Pubblicazione: (2026)
Implementation of New Security Features in CMSWEB Kubernetes Cluster at CERN
di: Ali, Aamir, et al.
Pubblicazione: (2024)
di: Ali, Aamir, et al.
Pubblicazione: (2024)
KubeDSM: A Kubernetes-based Dynamic Scheduling and Migration Framework for Cloud-Assisted Edge Clusters
di: Pashaeehir, Amirhossein, et al.
Pubblicazione: (2025)
di: Pashaeehir, Amirhossein, et al.
Pubblicazione: (2025)
Asynchronous Checkpoint for Eventually Consistent Databases
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2025)
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2025)
Kubernetes Deployment Options for On-Prem Clusters
di: Bryant, Lincoln, et al.
Pubblicazione: (2024)
di: Bryant, Lincoln, et al.
Pubblicazione: (2024)
Quantum-Enhanced Distributed Sensor Fusion: Lower Bounds on Aggregation from Projection Noise to Heisenberg-Limited Byzantine-Tolerant Networks
di: Iyer, Vasanth, et al.
Pubblicazione: (2026)
di: Iyer, Vasanth, et al.
Pubblicazione: (2026)
Transforming Lock-free Linked Lists into Distributed Lock-free Linked Lists
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2025)
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2025)
Distributing Context-Aware Shared Memory Data Structures: A Case Study on Singly-Linked Lists
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2024)
di: Ravishankar, Raaghav, et al.
Pubblicazione: (2024)
PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
RISC-V for HPC: Where we are and where we need to go
di: Brown, Nick
Pubblicazione: (2024)
di: Brown, Nick
Pubblicazione: (2024)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
di: Mao, Ying, et al.
Pubblicazione: (2020)
di: Mao, Ying, et al.
Pubblicazione: (2020)
Container-level Energy Observability in Kubernetes Clusters
di: Pijnacker, Bjorn, et al.
Pubblicazione: (2025)
di: Pijnacker, Bjorn, et al.
Pubblicazione: (2025)
An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
di: Punniyamoorthy, Vinoth, et al.
Pubblicazione: (2025)
di: Punniyamoorthy, Vinoth, et al.
Pubblicazione: (2025)
AntBatchInfer: Elastic Batch Inference in the Kubernetes Cluster
di: Li, Siyuan, et al.
Pubblicazione: (2024)
di: Li, Siyuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Orchestrating Mixed-Criticality Cloud Workloads in Reconfigurable Manufacturing Systems
di: Barletta, Marco, et al.
Pubblicazione: (2024) -
INDIGO: Page Migration for Hardware Memory Disaggregation Across a Network
di: Patke, Archit, et al.
Pubblicazione: (2025) -
Hierarchical Autoscaling for Large Language Model Serving with Chiron
di: Patke, Archit, et al.
Pubblicazione: (2025) -
Story of Two GPUs: Characterizing the Resilience of Hopper H100 and Ampere A100 GPUs
di: Cui, Shengkun, et al.
Pubblicazione: (2025) -
Queue management for slo-oriented large language model serving
di: Patke, Archit, et al.
Pubblicazione: (2024)