Quantifying Autoscaler Vulnerabilities: An Empirical Study of Resource Misallocation Induced by Cloud Infrastructure Faults
Fuente:
arXiv
Guardado en:
| Autor principal: | Park, Gijun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
por: Zhang, Guilin, et al.
Publicado: (2025)
por: Zhang, Guilin, et al.
Publicado: (2025)
ADAPT: A Self-Calibrating Proactive Autoscaler for Container Orchestration
por: Baghel, Himanshu Singh
Publicado: (2026)
por: Baghel, Himanshu Singh
Publicado: (2026)
A Dynamic Approach to Load Balancing in Cloud Infrastructure: Enhancing Energy Efficiency and Resource Utilization
por: Sakib, Shadman, et al.
Publicado: (2025)
por: Sakib, Shadman, et al.
Publicado: (2025)
Repurposing of the Run 2 CMS High Level Trigger Infrastructure as a Cloud Resource for Offline Computing
por: Mascheroni, Marco, et al.
Publicado: (2024)
por: Mascheroni, Marco, et al.
Publicado: (2024)
An Empirical Study of Production Incidents in Generative AI Cloud Services
por: Yan, Haoran, et al.
Publicado: (2025)
por: Yan, Haoran, et al.
Publicado: (2025)
Optimization Opportunities for Cloud-Based Data Pipeline Infrastructures
por: Jablonski, Johannes, et al.
Publicado: (2026)
por: Jablonski, Johannes, et al.
Publicado: (2026)
LaissezCloud: Continuous Resource Renegotiation for the Public Cloud
por: Harith, Tejas, et al.
Publicado: (2026)
por: Harith, Tejas, et al.
Publicado: (2026)
Host-Side Telemetry for Performance Diagnosis in Cloud and HPC GPU Infrastructure
por: Darzi, Erfan, et al.
Publicado: (2025)
por: Darzi, Erfan, et al.
Publicado: (2025)
SuperBench: Improving Cloud AI Infrastructure Reliability with Proactive Validation
por: Xiong, Yifan, et al.
Publicado: (2024)
por: Xiong, Yifan, et al.
Publicado: (2024)
Formal and Empirical Study of Metadata-Based Profiling for Resource Management in the Computing Continuum
por: Morichetta, Andrea, et al.
Publicado: (2025)
por: Morichetta, Andrea, et al.
Publicado: (2025)
Domain-Adversarial Transfer Learning for Fault Root Cause Identification in Cloud Computing Systems
por: Fang, Bruce, et al.
Publicado: (2025)
por: Fang, Bruce, et al.
Publicado: (2025)
A Self-Healing and Fault-Tolerant Cloud-based Digital Twin Processing Management Model
por: Saxena, Deepika, et al.
Publicado: (2025)
por: Saxena, Deepika, et al.
Publicado: (2025)
Agora: Bridging the GPU Cloud Resource-Price Disconnect
por: McDougall, Ian, et al.
Publicado: (2025)
por: McDougall, Ian, et al.
Publicado: (2025)
Leveraging Public Cloud Infrastructure for Real-time Connected Vehicle Speed Advisory at a Signalized Corridor
por: Deng, Hsien-Wen, et al.
Publicado: (2024)
por: Deng, Hsien-Wen, et al.
Publicado: (2024)
BSODiag: A Global Diagnosis Framework for Batch Servers Outage in Large-scale Cloud Infrastructure Systems
por: Duan, Tao, et al.
Publicado: (2025)
por: Duan, Tao, et al.
Publicado: (2025)
The SAP Cloud Infrastructure Dataset: A Reality Check of Scheduling and Placement of VMs in Cloud Computing
por: Uhlig, Arno, et al.
Publicado: (2025)
por: Uhlig, Arno, et al.
Publicado: (2025)
Pico-Cloud: Cloud Infrastructure for Tiny Edge Devices
por: Guri, Mordechai
Publicado: (2025)
por: Guri, Mordechai
Publicado: (2025)
Distributed Resource Selection for Self-Organising Cloud-Edge Systems
por: Renau, Quentin, et al.
Publicado: (2025)
por: Renau, Quentin, et al.
Publicado: (2025)
EES-CND: Collaborative Neural Decision-Making for Drift-Aware Fault-Tolerant Edge-Cloud Service Placement
por: Herabad, Mohammadsadeq Garshasbi, et al.
Publicado: (2026)
por: Herabad, Mohammadsadeq Garshasbi, et al.
Publicado: (2026)
M$^2$-MFP: A Multi-Scale and Multi-Level Memory Failure Prediction Framework for Reliable Cloud Infrastructure
por: Xie, Hongyi, et al.
Publicado: (2025)
por: Xie, Hongyi, et al.
Publicado: (2025)
Evaluating HPC-Style CPU Performance and Cost in Virtualized Cloud Infrastructures
por: Tharwani, Jay, et al.
Publicado: (2025)
por: Tharwani, Jay, et al.
Publicado: (2025)
HPCAdvisor: A Tool for Assisting Users in Selecting HPC Resources in the Cloud
por: Netto, Marco A. S.
Publicado: (2024)
por: Netto, Marco A. S.
Publicado: (2024)
Collaborative Multi-Agent Reinforcement Learning Approach for Elastic Cloud Resource Scaling
por: Fang, Bruce, et al.
Publicado: (2025)
por: Fang, Bruce, et al.
Publicado: (2025)
ICPS: Real-Time Resource Configuration for Cloud Serverless Functions Considering Affinity
por: Chen, Long, et al.
Publicado: (2025)
por: Chen, Long, et al.
Publicado: (2025)
An Empirical Study on Governance in Bitcoin's Consensus Evolution
por: Notland, Jakob Svennevik, et al.
Publicado: (2023)
por: Notland, Jakob Svennevik, et al.
Publicado: (2023)
Permuting Transactions in Ethereum Blocks: An Empirical Study
por: Droll, Jan
Publicado: (2025)
por: Droll, Jan
Publicado: (2025)
Cloud Resource Allocation with Convex Optimization
por: Boghani, Shayan, et al.
Publicado: (2025)
por: Boghani, Shayan, et al.
Publicado: (2025)
Literature Study on Operational Data Analytics Frameworks in Large-scale Computing Infrastructures
por: Suman, Shekhar, et al.
Publicado: (2026)
por: Suman, Shekhar, et al.
Publicado: (2026)
H-EYE: Holistic Resource Modeling and Management for Diversely Scaled Edge-Cloud Systems
por: Dagli, Ismet, et al.
Publicado: (2024)
por: Dagli, Ismet, et al.
Publicado: (2024)
Flora: Efficient Cloud Resource Selection for Big Data Processing via Job Classification
por: Will, Jonathan, et al.
Publicado: (2025)
por: Will, Jonathan, et al.
Publicado: (2025)
Modeling Anomaly Detection in Cloud Services: Analysis of the Properties that Impact Latency and Resource Consumption
por: Grabher, Gabriel Job Antunes, et al.
Publicado: (2025)
por: Grabher, Gabriel Job Antunes, et al.
Publicado: (2025)
HybridFlow: Resource-Adaptive Subtask Routing for Efficient Edge-Cloud LLM Inference
por: Dong, Jiangwen, et al.
Publicado: (2025)
por: Dong, Jiangwen, et al.
Publicado: (2025)
ARRC: Explainable, Workflow-Integrated Recommender for Sustainable Resource Optimization Across the Edge-Cloud Continuum
por: Jahnke, Brian-Frederik, et al.
Publicado: (2025)
por: Jahnke, Brian-Frederik, et al.
Publicado: (2025)
Resource Slicing through Intelligent Orchestration of Energy-aware IoT services in Edge-Cloud Continuum
por: Shahid, Hafiz Faheem, et al.
Publicado: (2024)
por: Shahid, Hafiz Faheem, et al.
Publicado: (2024)
COUNTER: Cluster GCN based Energy Efficient Resource Management for Sustainable Cloud Computing Environments
por: Wang, Han, et al.
Publicado: (2025)
por: Wang, Han, et al.
Publicado: (2025)
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
por: Tang, Lujie, et al.
Publicado: (2024)
por: Tang, Lujie, et al.
Publicado: (2024)
Building Castles in the Cloud: Architecting Resilient and Scalable Infrastructure
por: Gundla, Naresh Kumar
Publicado: (2024)
por: Gundla, Naresh Kumar
Publicado: (2024)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
por: Lu, Zhengxian, et al.
Publicado: (2024)
por: Lu, Zhengxian, et al.
Publicado: (2024)
An Empirical Study of Cross-Language Interoperability in Replicated Data Systems
por: Mondal, Provakar, et al.
Publicado: (2025)
por: Mondal, Provakar, et al.
Publicado: (2025)
Deep Reinforcement Learning-based Methods for Resource Scheduling in Cloud Computing: A Review and Future Directions
por: Zhou, Guangyao, et al.
Publicado: (2021)
por: Zhou, Guangyao, et al.
Publicado: (2021)
Ejemplares similares
-
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
por: Zhang, Guilin, et al.
Publicado: (2025) -
ADAPT: A Self-Calibrating Proactive Autoscaler for Container Orchestration
por: Baghel, Himanshu Singh
Publicado: (2026) -
A Dynamic Approach to Load Balancing in Cloud Infrastructure: Enhancing Energy Efficiency and Resource Utilization
por: Sakib, Shadman, et al.
Publicado: (2025) -
Repurposing of the Run 2 CMS High Level Trigger Infrastructure as a Cloud Resource for Offline Computing
por: Mascheroni, Marco, et al.
Publicado: (2024) -
An Empirical Study of Production Incidents in Generative AI Cloud Services
por: Yan, Haoran, et al.
Publicado: (2025)