ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Haoyu, Islam, Muhammed Tawfiqul, Xu, Minxian, Buyya, Rajkumar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Proactive and Reactive Autoscaling Techniques for Edge Computing
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025)
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025)
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
von: Bai, Xu, et al.
Veröffentlicht: (2025)
von: Bai, Xu, et al.
Veröffentlicht: (2025)
Adaptive Management of Microservices in Dynamic Computing Environments: A Taxonomy and Future Directions
von: Chen, Ming, et al.
Veröffentlicht: (2026)
von: Chen, Ming, et al.
Veröffentlicht: (2026)
iDynamics: A Configurable Emulation Framework for Evaluating Microservice Scheduling Policies under Controllable Cloud-Edge Dynamics
von: Chen, Ming, et al.
Veröffentlicht: (2025)
von: Chen, Ming, et al.
Veröffentlicht: (2025)
A Hybrid Reactive-Proactive Auto-scaling Algorithm for SLA-Constrained Edge Computing
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025)
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025)
TraDE: Network and Traffic-aware Adaptive Scheduling for Microservices Under Dynamics
von: Chen, Ming, et al.
Veröffentlicht: (2024)
von: Chen, Ming, et al.
Veröffentlicht: (2024)
HGraphScale: Hierarchical Graph Learning for Autoscaling Microservice Applications in Container-based Cloud Computing
von: Fang, Zhengxin, et al.
Veröffentlicht: (2025)
von: Fang, Zhengxin, et al.
Veröffentlicht: (2025)
DRPC: Distributed Reinforcement Learning Approach for Scalable Resource Provisioning in Container-based Clusters
von: Bai, Haoyu, et al.
Veröffentlicht: (2024)
von: Bai, Haoyu, et al.
Veröffentlicht: (2024)
LSRAM: A Lightweight Autoscaling and SLO Resource Allocation Framework for Microservices Based on Gradient Descent
von: Hu, Kan, et al.
Veröffentlicht: (2024)
von: Hu, Kan, et al.
Veröffentlicht: (2024)
Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
von: Xu, Minxian, et al.
Veröffentlicht: (2025)
von: Xu, Minxian, et al.
Veröffentlicht: (2025)
Multi-Layer Scheduling for MoE-Based LLM Reasoning
von: Sun, Yifan, et al.
Veröffentlicht: (2026)
von: Sun, Yifan, et al.
Veröffentlicht: (2026)
Placement of Microservices-based IoT Applications in Fog Computing: A Taxonomy and Future Directions
von: Pallewatta, Samodha, et al.
Veröffentlicht: (2022)
von: Pallewatta, Samodha, et al.
Veröffentlicht: (2022)
Self-adaptive, Requirements-driven Autoscaling of Microservices
von: Nunes, João Paulo Karol Santos, et al.
Veröffentlicht: (2024)
von: Nunes, João Paulo Karol Santos, et al.
Veröffentlicht: (2024)
DOPD: A Dynamic PD-Disaggregation Architecture for Maximizing Goodput in LLM Inference Serving
von: Liao, Junhan, et al.
Veröffentlicht: (2025)
von: Liao, Junhan, et al.
Veröffentlicht: (2025)
PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving
von: Bai, Xu, et al.
Veröffentlicht: (2026)
von: Bai, Xu, et al.
Veröffentlicht: (2026)
LLM-Driven Intent-Based Privacy-Aware Orchestration Across the Cloud-Edge Continuum
von: Su, Zijie, et al.
Veröffentlicht: (2026)
von: Su, Zijie, et al.
Veröffentlicht: (2026)
A Joint Time and Energy-Efficient Federated Learning-based Computation Offloading Method for Mobile Edge Computing
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
EnFed: An Energy-aware Federated Learning in Resource Constrained Environments for Human Activity Recognition
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
Generative Federated Learning for Smart Prediction and Recommendation Applications
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2025)
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2025)
TrustMesh: A Blockchain-Enabled Trusted Distributed Computing Framework for Open Heterogeneous IoT Environments
von: Rangwala, Murtaza, et al.
Veröffentlicht: (2024)
von: Rangwala, Murtaza, et al.
Veröffentlicht: (2024)
A Deep Reinforcement Learning Approach for Cost Optimized Workflow Scheduling in Cloud Computing Environments
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
SpotKube: Cost-Optimal Microservices Deployment with Cluster Autoscaling and Spot Pricing
von: Edirisinghe, Dasith, et al.
Veröffentlicht: (2024)
von: Edirisinghe, Dasith, et al.
Veröffentlicht: (2024)
MSARS: A Meta-Learning and Reinforcement Learning Framework for SLO Resource Allocation and Adaptive Scaling for Microservices
von: Hu, Kan, et al.
Veröffentlicht: (2024)
von: Hu, Kan, et al.
Veröffentlicht: (2024)
Microservices-based Software Systems Reengineering: State-of-the-Art and Future Directions
von: Mohottige, Thakshila Imiya, et al.
Veröffentlicht: (2024)
von: Mohottige, Thakshila Imiya, et al.
Veröffentlicht: (2024)
A Risk-Aware UAV-Edge Service Framework for Wildfire Monitoring and Emergency Response
von: Huang, Yulun, et al.
Veröffentlicht: (2026)
von: Huang, Yulun, et al.
Veröffentlicht: (2026)
Reinforcement Learning based Workflow Scheduling in Cloud and Edge Computing Environments: A Taxonomy, Review and Future Directions
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
von: Jayanetti, Amanda, et al.
Veröffentlicht: (2024)
TF-DDRL: A Transformer-enhanced Distributed DRL Technique for Scheduling IoT Applications in Edge and Cloud Computing Environments
von: Wang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyu, et al.
Veröffentlicht: (2024)
ReinFog: A Deep Reinforcement Learning Empowered Framework for Resource Management in Edge and Cloud Computing Environments
von: Wang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyu, et al.
Veröffentlicht: (2024)
StatuScale: Status-aware and Elastic Scaling Strategy for Microservice Applications
von: Wen, Linfeng, et al.
Veröffentlicht: (2024)
von: Wen, Linfeng, et al.
Veröffentlicht: (2024)
A Cascaded Graph Neural Network for Joint Root Cause Localization and Analysis in Edge Computing Environments
von: Fernando, Duneesha, et al.
Veröffentlicht: (2026)
von: Fernando, Duneesha, et al.
Veröffentlicht: (2026)
A Decentralized Root Cause Localization Approach for Edge Computing Environments
von: Fernando, Duneesha, et al.
Veröffentlicht: (2025)
von: Fernando, Duneesha, et al.
Veröffentlicht: (2025)
iAnomaly: A Toolkit for Generating Performance Anomaly Datasets in Edge-Cloud Integrated Computing Environments
von: Fernando, Duneesha, et al.
Veröffentlicht: (2024)
von: Fernando, Duneesha, et al.
Veröffentlicht: (2024)
Saarthi: An End-to-End Intelligent Platform for Optimising Distributed Serverless Workloads
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
C-Koordinator: Interference-aware Management for Large-scale and Co-located Microservice Clusters
von: Song, Shengye, et al.
Veröffentlicht: (2025)
von: Song, Shengye, et al.
Veröffentlicht: (2025)
Performance and Security Aware Distributed Service Placement in Fog Computing
von: Goudarzi, Mohammad, et al.
Veröffentlicht: (2026)
von: Goudarzi, Mohammad, et al.
Veröffentlicht: (2026)
A Knowledge Distillation-empowered Adaptive Federated Reinforcement Learning Framework for Multi-Domain IoT Applications Scheduling
von: Wang, Zhiyu, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyu, et al.
Veröffentlicht: (2025)
BrownoutServe: SLO-Aware Inference Serving under Bursty Workloads for MoE-based LLMs
von: Hu, Jianmin, et al.
Veröffentlicht: (2025)
von: Hu, Jianmin, et al.
Veröffentlicht: (2025)
Multi-Objective Optimization of Consumer Group Autoscaling in Message Broker Systems
von: Landau, Diogo, et al.
Veröffentlicht: (2024)
von: Landau, Diogo, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning-based Methods for Resource Scheduling in Cloud Computing: A Review and Future Directions
von: Zhou, Guangyao, et al.
Veröffentlicht: (2021)
von: Zhou, Guangyao, et al.
Veröffentlicht: (2021)
Deep Reinforcement Learning (DRL)-based Methods for Serverless Stream Processing Engines: A Vision, Architectural Elements, and Future Directions
von: Read, Maria R., et al.
Veröffentlicht: (2024)
von: Read, Maria R., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Proactive and Reactive Autoscaling Techniques for Edge Computing
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025) -
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
von: Bai, Xu, et al.
Veröffentlicht: (2025) -
Adaptive Management of Microservices in Dynamic Computing Environments: A Taxonomy and Future Directions
von: Chen, Ming, et al.
Veröffentlicht: (2026) -
iDynamics: A Configurable Emulation Framework for Evaluating Microservice Scheduling Policies under Controllable Cloud-Edge Dynamics
von: Chen, Ming, et al.
Veröffentlicht: (2025) -
A Hybrid Reactive-Proactive Auto-scaling Algorithm for SLA-Constrained Edge Computing
von: Gupta, Suhrid, et al.
Veröffentlicht: (2025)