Intelligent Cloud Orchestration: A Hybrid Predictive and Heuristic Framework for Cost Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Nagoriya, Heet, Rohit, Komal |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Next-Generation Event-Driven Architectures: Performance, Scalability, and Intelligent Orchestration Across Messaging Frameworks
por: Arafat, Jahidul, et al.
Publicado: (2025)
por: Arafat, Jahidul, et al.
Publicado: (2025)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
por: Zhang, Guilin, et al.
Publicado: (2025)
por: Zhang, Guilin, et al.
Publicado: (2025)
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
por: Lysenstøen, Christian
Publicado: (2026)
por: Lysenstøen, Christian
Publicado: (2026)
EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM Inference
por: Sidik, Bronislav, et al.
Publicado: (2026)
por: Sidik, Bronislav, et al.
Publicado: (2026)
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
por: Will, Jonathan, et al.
Publicado: (2025)
por: Will, Jonathan, et al.
Publicado: (2025)
Cognitive Infrastructure: A Unified DCIM Framework for AI Data Centers
por: Sunkara, Krishna Chaitanya
Publicado: (2026)
por: Sunkara, Krishna Chaitanya
Publicado: (2026)
HiDVFS: A Hierarchical Multi-Agent DVFS Scheduler for OpenMP DAG Workloads
por: Pivezhandi, Mohammad, et al.
Publicado: (2026)
por: Pivezhandi, Mohammad, et al.
Publicado: (2026)
Feature-Aware Task-to-Core Allocation in Embedded Multi-core Platforms via Statistical Learning
por: Pivezhandi, Mohammad, et al.
Publicado: (2025)
por: Pivezhandi, Mohammad, et al.
Publicado: (2025)
ACME: Adaptive Customization of Large Models via Distributed Systems
por: Dai, Ziming, et al.
Publicado: (2025)
por: Dai, Ziming, et al.
Publicado: (2025)
An Empirical Study of the Impact of Federated Learning on Machine Learning Model Accuracy
por: Yang, Haotian, et al.
Publicado: (2025)
por: Yang, Haotian, et al.
Publicado: (2025)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
por: Polyakov, Igor, et al.
Publicado: (2025)
por: Polyakov, Igor, et al.
Publicado: (2025)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
por: Yuan, Renzhong, et al.
Publicado: (2026)
por: Yuan, Renzhong, et al.
Publicado: (2026)
Laminar: A Probe-First Scheduling Paradigm with Deterministic Runtime Survival
por: Chu, Zhengyan
Publicado: (2026)
por: Chu, Zhengyan
Publicado: (2026)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
por: Xie, Tianfang
Publicado: (2026)
por: Xie, Tianfang
Publicado: (2026)
Parameter-Efficient and Personalized Federated Training of Generative Models at the Edge
por: Khan, Kabir, et al.
Publicado: (2025)
por: Khan, Kabir, et al.
Publicado: (2025)
CarbonEdge: Carbon-Aware Deep Learning Inference Framework for Sustainable Edge Computing
por: Zhang, Guilin, et al.
Publicado: (2026)
por: Zhang, Guilin, et al.
Publicado: (2026)
Readout-Side Bypass for Residual Hybrid Quantum-Classical Models
por: Zhang, Guilin, et al.
Publicado: (2025)
por: Zhang, Guilin, et al.
Publicado: (2025)
ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving
por: Li, Xiangchen, et al.
Publicado: (2026)
por: Li, Xiangchen, et al.
Publicado: (2026)
SRFed: Mitigating Poisoning Attacks in Privacy-Preserving Federated Learning with Heterogeneous Data
por: Lu, Yiwen
Publicado: (2026)
por: Lu, Yiwen
Publicado: (2026)
SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis Detection
por: Ospitia, Santiago, et al.
Publicado: (2026)
por: Ospitia, Santiago, et al.
Publicado: (2026)
Cross-Platform Fused MoE Dispatch in Triton: Portable Expert Routing Without CUDA
por: Mitra, Subhadip
Publicado: (2026)
por: Mitra, Subhadip
Publicado: (2026)
Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
por: Woisetschläger, Herbert, et al.
Publicado: (2023)
por: Woisetschläger, Herbert, et al.
Publicado: (2023)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
por: Rosendal, Daan, et al.
Publicado: (2026)
por: Rosendal, Daan, et al.
Publicado: (2026)
Serverless GPU Architecture for Enterprise HR Analytics: A Production-Scale BDaaS Implementation
por: Zhang, Guilin, et al.
Publicado: (2025)
por: Zhang, Guilin, et al.
Publicado: (2025)
Uncovering Bugs in Formal Explainers: A Case Study with PyXAI
por: Huang, Xuanxiang, et al.
Publicado: (2025)
por: Huang, Xuanxiang, et al.
Publicado: (2025)
Knowledge Graphs-Driven Intelligence for Distributed Decision Systems
por: Napoli, Rosario, et al.
Publicado: (2026)
por: Napoli, Rosario, et al.
Publicado: (2026)
WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
por: Li, Xiangchen, et al.
Publicado: (2026)
por: Li, Xiangchen, et al.
Publicado: (2026)
Shipwright: Proving liveness of distributed systems with Byzantine participants
por: Leung, Derek, et al.
Publicado: (2025)
por: Leung, Derek, et al.
Publicado: (2025)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
por: Erben, Alexander, et al.
Publicado: (2023)
por: Erben, Alexander, et al.
Publicado: (2023)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
por: Guan, Bo
Publicado: (2026)
por: Guan, Bo
Publicado: (2026)
LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability
por: Lucas, Tom, et al.
Publicado: (2026)
por: Lucas, Tom, et al.
Publicado: (2026)
LAMMPS-KOKKOS: Performance Portable Molecular Dynamics Across Exascale Architectures
por: Johansson, Anders, et al.
Publicado: (2025)
por: Johansson, Anders, et al.
Publicado: (2025)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
por: Talluri, Sacheendra, et al.
Publicado: (2025)
por: Talluri, Sacheendra, et al.
Publicado: (2025)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
por: Motta, Steven, et al.
Publicado: (2026)
por: Motta, Steven, et al.
Publicado: (2026)
Network Structures as an Attack Surface: Topology-Based Privacy Leakage in Federated Learning
por: Rangwala, Murtaza, et al.
Publicado: (2025)
por: Rangwala, Murtaza, et al.
Publicado: (2025)
Vectorized Adaptive Histograms for Sparse Oblique Forests
por: Lubonja, Ariel, et al.
Publicado: (2026)
por: Lubonja, Ariel, et al.
Publicado: (2026)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
por: Putra, Jody Almaida
Publicado: (2026)
por: Putra, Jody Almaida
Publicado: (2026)
A Selective Homomorphic Encryption Approach for Faster Privacy-Preserving Federated Learning
por: Korkmaz, Abdulkadir, et al.
Publicado: (2025)
por: Korkmaz, Abdulkadir, et al.
Publicado: (2025)
Data Race Satisfiability on Array Elements
por: Shim, Junhyung, et al.
Publicado: (2025)
por: Shim, Junhyung, et al.
Publicado: (2025)
Privacy-Aware Split Inference with Speculative Decoding for Large Language Models over Wide-Area Networks
por: Cunningham, Michael
Publicado: (2026)
por: Cunningham, Michael
Publicado: (2026)
Ejemplares similares
-
Next-Generation Event-Driven Architectures: Performance, Scalability, and Intelligent Orchestration Across Messaging Frameworks
por: Arafat, Jahidul, et al.
Publicado: (2025) -
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
por: Zhang, Guilin, et al.
Publicado: (2025) -
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
por: Lysenstøen, Christian
Publicado: (2026) -
EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM Inference
por: Sidik, Bronislav, et al.
Publicado: (2026) -
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
por: Will, Jonathan, et al.
Publicado: (2025)