Machine learning-based cloud resource allocation algorithms: a comprehensive comparative review
Fuente:
arXiv
Saved in:
| Main Authors: | Bodra, Deep, Khairnar, Sushil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparative Performance Analysis of Modern NoSQL Data Technologies: Redis, Aerospike, and Dragonfly
by: Bodra, Deep, et al.
Published: (2025)
by: Bodra, Deep, et al.
Published: (2025)
Adaptive multi-criteria-based load balancing technique for resource allocation in fog-cloud environments
by: Gad-Elrab, Ahmed A. A., et al.
Published: (2024)
by: Gad-Elrab, Ahmed A. A., et al.
Published: (2024)
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026)
by: Mehta, Deep Pankajbhai
Published: (2026)
The intelligent prediction and assessment of financial information risk in the cloud computing model
by: Wang, Yufu, et al.
Published: (2024)
by: Wang, Yufu, et al.
Published: (2024)
Regression prediction algorithm for energy consumption regression in cloud computing based on horned lizard algorithm optimised convolutional neural network-bidirectional gated recurrent unit
by: Li, Feiyang, et al.
Published: (2024)
by: Li, Feiyang, et al.
Published: (2024)
A privacy-preserving, distributed and cooperative FCM-based learning approach for cancer research
by: Salmeron, Jose L., et al.
Published: (2024)
by: Salmeron, Jose L., et al.
Published: (2024)
Dynamic Resource Allocation for Virtual Machine Migration Optimization using Machine Learning
by: Gong, Yulu, et al.
Published: (2024)
by: Gong, Yulu, et al.
Published: (2024)
Training Through Failure: Effects of Data Consistency in Parallel Machine Learning Training
by: Cao, Ray, et al.
Published: (2024)
by: Cao, Ray, et al.
Published: (2024)
Practical offloading for fine-tuning LLM on commodity GPU via learned sparse projectors
by: Chen, Siyuan, et al.
Published: (2024)
by: Chen, Siyuan, et al.
Published: (2024)
Block size estimation for data partitioning in HPC applications using machine learning techniques
by: Cantini, Riccardo, et al.
Published: (2022)
by: Cantini, Riccardo, et al.
Published: (2022)
High-Dimensional Data Processing: Benchmarking Machine Learning and Deep Learning Architectures in Local and Distributed Environments
by: Rodriguez, Julian, et al.
Published: (2025)
by: Rodriguez, Julian, et al.
Published: (2025)
A Nonlinear Hash-based Optimization Method for SpMV on GPUs
by: Yan, Chen, et al.
Published: (2025)
by: Yan, Chen, et al.
Published: (2025)
BanditWare: A Contextual Bandit-based Framework for Hardware Prediction
by: Coleman, Tainã, et al.
Published: (2025)
by: Coleman, Tainã, et al.
Published: (2025)
Adaptive AI-based Decentralized Resource Management in the Cloud-Edge Continuum
by: Li, Lanpei, et al.
Published: (2025)
by: Li, Lanpei, et al.
Published: (2025)
A Blockchain and Artificial Intelligence based System for Halal Food Traceability
by: Alourani, Abdulla, et al.
Published: (2024)
by: Alourani, Abdulla, et al.
Published: (2024)
Accelerating Large Language Model Training with Hybrid GPU-based Compression
by: Xu, Lang, et al.
Published: (2024)
by: Xu, Lang, et al.
Published: (2024)
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems
by: Zhou, Wenqing, et al.
Published: (2025)
by: Zhou, Wenqing, et al.
Published: (2025)
Speeding up Local Optimization in Vehicle Routing with Tensor-based GPU Acceleration
by: Lei, Zhenyu, et al.
Published: (2025)
by: Lei, Zhenyu, et al.
Published: (2025)
A Survey on Large Language Model Acceleration based on KV Cache Management
by: Li, Haoyang, et al.
Published: (2024)
by: Li, Haoyang, et al.
Published: (2024)
Blockchain-aided wireless federated learning: Resource allocation and client scheduling
by: Li, Jun, et al.
Published: (2024)
by: Li, Jun, et al.
Published: (2024)
Venus: An Efficient Edge Memory-and-Retrieval System for VLM-based Online Video Understanding
by: Ye, Shengyuan, et al.
Published: (2025)
by: Ye, Shengyuan, et al.
Published: (2025)
GWLZ: A Group-wise Learning-based Lossy Compression Framework for Scientific Data
by: Jia, Wenqi, et al.
Published: (2024)
by: Jia, Wenqi, et al.
Published: (2024)
Distributed Inference on Mobile Edge and Cloud: A Data-Cartography based Clustering Approach
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
ProMoE: Fast MoE-based LLM Serving using Proactive Caching
by: Song, Xiaoniu, et al.
Published: (2024)
by: Song, Xiaoniu, et al.
Published: (2024)
D$^{2}$MoE: Dual Routing and Dynamic Scheduling for Efficient On-Device MoE-based LLM Serving
by: Wang, Haodong, et al.
Published: (2025)
by: Wang, Haodong, et al.
Published: (2025)
Atmosphere: Context and situational-aware collaborative IoT architecture for edge-fog-cloud computing
by: Ortiz, Guadalupe, et al.
Published: (2024)
by: Ortiz, Guadalupe, et al.
Published: (2024)
Tutoring LLM into a Better CUDA Optimizer
by: Brabec, Matyáš, et al.
Published: (2025)
by: Brabec, Matyáš, et al.
Published: (2025)
SMART: When is it Actually Worth Expanding a Speculative Tree?
by: Wang, Lifu, et al.
Published: (2026)
by: Wang, Lifu, et al.
Published: (2026)
MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service
by: Yu, Timothy Tin Long, et al.
Published: (2026)
by: Yu, Timothy Tin Long, et al.
Published: (2026)
An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU
by: Yang, Ruijia, et al.
Published: (2026)
by: Yang, Ruijia, et al.
Published: (2026)
FLAS: a combination of proactive and reactive auto-scaling architecture for distributed services
by: Rampérez, Víctor, et al.
Published: (2025)
by: Rampérez, Víctor, et al.
Published: (2025)
Mixture-of-Schedulers: An Adaptive Scheduling Agent as a Learned Router for Expert Policies
by: Wang, Xinbo, et al.
Published: (2025)
by: Wang, Xinbo, et al.
Published: (2025)
Isambard-AI: a leadership class supercomputer optimised specifically for Artificial Intelligence
by: McIntosh-Smith, Simon, et al.
Published: (2024)
by: McIntosh-Smith, Simon, et al.
Published: (2024)
KVCache Cache in the Wild: Characterizing and Optimizing KVCache Cache at a Large Cloud Provider
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research
by: Heredia, Ignacio, et al.
Published: (2025)
by: Heredia, Ignacio, et al.
Published: (2025)
zIA: a GenAI-powered local auntie assists tourists in Italy
by: Cassani, Alexio, et al.
Published: (2024)
by: Cassani, Alexio, et al.
Published: (2024)
Design a Win-Win Strategy That Is Fair to Both Service Providers and Tasks When Rejection Is Not an Option
by: Trabelsi, Yohai, et al.
Published: (2024)
by: Trabelsi, Yohai, et al.
Published: (2024)
DIAP: A Decentralized Agent Identity Protocol with Zero-Knowledge Proofs and a Hybrid P2P Stack
by: Liu, Yuanjie, et al.
Published: (2025)
by: Liu, Yuanjie, et al.
Published: (2025)
Accelerating a Triton Fused Kernel for W4A16 Quantized Inference with SplitK work decomposition
by: Hoque, Adnan, et al.
Published: (2024)
by: Hoque, Adnan, et al.
Published: (2024)
An Upload-Efficient Scheme for Transferring Knowledge From a Server-Side Pre-trained Generator to Clients in Heterogeneous Federated Learning
by: Zhang, Jianqing, et al.
Published: (2024)
by: Zhang, Jianqing, et al.
Published: (2024)
Similar Items
-
Comparative Performance Analysis of Modern NoSQL Data Technologies: Redis, Aerospike, and Dragonfly
by: Bodra, Deep, et al.
Published: (2025) -
Adaptive multi-criteria-based load balancing technique for resource allocation in fog-cloud environments
by: Gad-Elrab, Ahmed A. A., et al.
Published: (2024) -
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026) -
The intelligent prediction and assessment of financial information risk in the cloud computing model
by: Wang, Yufu, et al.
Published: (2024) -
Regression prediction algorithm for energy consumption regression in cloud computing based on horned lizard algorithm optimised convolutional neural network-bidirectional gated recurrent unit
by: Li, Feiyang, et al.
Published: (2024)