A Holistic Framework for Automated Configuration Recommendation for Cloud Service Monitoring
Fuente:
arXiv
Salvato in:
| Autori principali: | Bastos, Anson, Venneti, Shreeya, Parayil, Anjaly, Choure, Ayush, Bansal, Chetan, Wang, Rujia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Attention Enhanced Entity Recommendation for Intelligent Monitoring in Cloud Systems
di: Hussain, Fiza, et al.
Pubblicazione: (2025)
di: Hussain, Fiza, et al.
Pubblicazione: (2025)
Towards Cloud Efficiency with Large-scale Workload Characterization
di: Parayil, Anjaly, et al.
Pubblicazione: (2024)
di: Parayil, Anjaly, et al.
Pubblicazione: (2024)
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025)
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025)
Sutradhara: An Intelligent Orchestrator-Engine Co-design for Tool-based Agentic Inference
di: Biswas, Anish, et al.
Pubblicazione: (2026)
di: Biswas, Anish, et al.
Pubblicazione: (2026)
Serving Heterogeneous LoRA Adapters in Distributed LLM Inference Systems
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025)
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025)
Dependency Aware Incident Linking in Large Cloud Systems
di: Ghosh, Supriyo, et al.
Pubblicazione: (2024)
di: Ghosh, Supriyo, et al.
Pubblicazione: (2024)
An Empirical Study of Production Incidents in Generative AI Cloud Services
di: Yan, Haoran, et al.
Pubblicazione: (2025)
di: Yan, Haoran, et al.
Pubblicazione: (2025)
Ensuring Fair LLM Serving Amid Diverse Applications
di: Khan, Redwan Ibne Seraj, et al.
Pubblicazione: (2024)
di: Khan, Redwan Ibne Seraj, et al.
Pubblicazione: (2024)
Workload Intelligence: Punching Holes Through the Cloud Abstraction
di: Huang, Lexiang, et al.
Pubblicazione: (2024)
di: Huang, Lexiang, et al.
Pubblicazione: (2024)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
di: Jain, Kunal, et al.
Pubblicazione: (2024)
di: Jain, Kunal, et al.
Pubblicazione: (2024)
Intelligent Monitoring Framework for Cloud Services: A Data-Driven Approach
di: Srinivas, Pooja, et al.
Pubblicazione: (2024)
di: Srinivas, Pooja, et al.
Pubblicazione: (2024)
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
di: Chen, Yinfang, et al.
Pubblicazione: (2025)
di: Chen, Yinfang, et al.
Pubblicazione: (2025)
MAS-H2: A Hierarchical Multi-Agent System for Holistic Cloud-Native Autoscaling
di: Hamzeh, Hamed, et al.
Pubblicazione: (2026)
di: Hamzeh, Hamed, et al.
Pubblicazione: (2026)
Large Language Model Aided QoS Prediction for Service Recommendation
di: Liu, Huiying, et al.
Pubblicazione: (2024)
di: Liu, Huiying, et al.
Pubblicazione: (2024)
MAIZX: A Carbon-Aware Framework for Optimizing Cloud Computing Emissions
di: Ruilova, Federico, et al.
Pubblicazione: (2025)
di: Ruilova, Federico, et al.
Pubblicazione: (2025)
Federated Learning Framework for Scalable AI in Heterogeneous HPC and Cloud Environments
di: Ghimire, Sangam, et al.
Pubblicazione: (2025)
di: Ghimire, Sangam, et al.
Pubblicazione: (2025)
A Robust Power Model Training Framework for Cloud Native Runtime Energy Metric Exporter
di: Choochotkaew, Sunyanan, et al.
Pubblicazione: (2024)
di: Choochotkaew, Sunyanan, et al.
Pubblicazione: (2024)
Holistic Evaluation Metrics: Use Case Sensitive Evaluation Metrics for Federated Learning
di: Li, Yanli, et al.
Pubblicazione: (2024)
di: Li, Yanli, et al.
Pubblicazione: (2024)
A Distributed Framework for Causal Modeling of Performance Variability in GPU Traces
di: Lahiry, Ankur, et al.
Pubblicazione: (2025)
di: Lahiry, Ankur, et al.
Pubblicazione: (2025)
Stress Monitoring in Healthcare: An Ensemble Machine Learning Framework Using Wearable Sensor Data
di: Sinhal, Arpana, et al.
Pubblicazione: (2025)
di: Sinhal, Arpana, et al.
Pubblicazione: (2025)
Simplified Swarm Learning Framework for Robust and Scalable Diagnostic Services in Cancer Histopathology
di: Wu, Yanjie, et al.
Pubblicazione: (2025)
di: Wu, Yanjie, et al.
Pubblicazione: (2025)
SERFLOW: A Cross-Service Cost Optimization Framework for SLO-Aware Dynamic ML Inference
di: Zhang, Zongshun, et al.
Pubblicazione: (2025)
di: Zhang, Zongshun, et al.
Pubblicazione: (2025)
DNN-Powered MLOps Pipeline Optimization for Large Language Models: A Framework for Automated Deployment and Resource Management
di: Krishnamoorthy, Mahesh Vaijainthymala, et al.
Pubblicazione: (2025)
di: Krishnamoorthy, Mahesh Vaijainthymala, et al.
Pubblicazione: (2025)
TurboGR: An Accelerated Training System for Large-Scale Generative Recommendation
di: Chai, Huichao, et al.
Pubblicazione: (2026)
di: Chai, Huichao, et al.
Pubblicazione: (2026)
Pipette: Automatic Fine-grained Large Language Model Training Configurator for Real-World Clusters
di: Yim, Jinkyu, et al.
Pubblicazione: (2024)
di: Yim, Jinkyu, et al.
Pubblicazione: (2024)
Governing Cloud Data Pipelines with Agentic AI
di: Kirubakaran, Aswathnarayan Muthukrishnan, et al.
Pubblicazione: (2025)
di: Kirubakaran, Aswathnarayan Muthukrishnan, et al.
Pubblicazione: (2025)
Improvements & Evaluations on the MLCommons CloudMask Benchmark
di: Chennamsetti, Varshitha, et al.
Pubblicazione: (2024)
di: Chennamsetti, Varshitha, et al.
Pubblicazione: (2024)
Local-Cloud Inference Offloading for LLMs in Multi-Modal, Multi-Task, Multi-Dialogue Settings
di: Yuan, Liangqi, et al.
Pubblicazione: (2025)
di: Yuan, Liangqi, et al.
Pubblicazione: (2025)
Federated Automated Feature Engineering
di: Overman, Tom, et al.
Pubblicazione: (2024)
di: Overman, Tom, et al.
Pubblicazione: (2024)
AI-Driven Health Monitoring of Distributed Computing Architecture: Insights from XGBoost and SHAP
di: Sun, Xiaoxuan, et al.
Pubblicazione: (2024)
di: Sun, Xiaoxuan, et al.
Pubblicazione: (2024)
Synthetic Time Series for Anomaly Detection in Cloud Microservices
di: Allam, Mohamed, et al.
Pubblicazione: (2024)
di: Allam, Mohamed, et al.
Pubblicazione: (2024)
FL-GUARD: A Holistic Framework for Run-Time Detection and Recovery of Negative Federated Learning
di: Lin, Hong, et al.
Pubblicazione: (2024)
di: Lin, Hong, et al.
Pubblicazione: (2024)
Agglomerative Federated Learning: Empowering Larger Model Training via End-Edge-Cloud Collaboration
di: Wu, Zhiyuan, et al.
Pubblicazione: (2023)
di: Wu, Zhiyuan, et al.
Pubblicazione: (2023)
DiSCo: Device-Server Collaborative LLM-Based Text Streaming Services
di: Sun, Ting, et al.
Pubblicazione: (2025)
di: Sun, Ting, et al.
Pubblicazione: (2025)
Computing in the Era of Large Generative Models: From Cloud-Native to AI-Native
di: Lu, Yao, et al.
Pubblicazione: (2024)
di: Lu, Yao, et al.
Pubblicazione: (2024)
Beyond Model Scale Limits: End-Edge-Cloud Federated Learning with Self-Rectified Knowledge Agglomeration
di: Wu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wu, Zhiyuan, et al.
Pubblicazione: (2025)
Two-dimensional Sparse Parallelism for Large Scale Deep Learning Recommendation Model Training
di: Zhang, Xin, et al.
Pubblicazione: (2025)
di: Zhang, Xin, et al.
Pubblicazione: (2025)
Reducing Energy Bloat in Large Model Training
di: Chung, Jae-Won, et al.
Pubblicazione: (2023)
di: Chung, Jae-Won, et al.
Pubblicazione: (2023)
PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
AIConfigurator: Lightning-Fast Configuration Optimization for Multi-Framework LLM Serving
di: Xu, Tianhao, et al.
Pubblicazione: (2026)
di: Xu, Tianhao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Attention Enhanced Entity Recommendation for Intelligent Monitoring in Cloud Systems
di: Hussain, Fiza, et al.
Pubblicazione: (2025) -
Towards Cloud Efficiency with Large-scale Workload Characterization
di: Parayil, Anjaly, et al.
Pubblicazione: (2024) -
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025) -
Sutradhara: An Intelligent Orchestrator-Engine Co-design for Tool-based Agentic Inference
di: Biswas, Anish, et al.
Pubblicazione: (2026) -
Serving Heterogeneous LoRA Adapters in Distributed LLM Inference Systems
di: Jaiswal, Shashwat, et al.
Pubblicazione: (2025)