KubeIntellect: A Modular LLM-Orchestrated Agent Framework for End-to-End Kubernetes Management
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ardebili, Mohsen Seyedkazemi, Bartolini, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KubeDSM: A Kubernetes-based Dynamic Scheduling and Migration Framework for Cloud-Assisted Edge Clusters
von: Pashaeehir, Amirhossein, et al.
Veröffentlicht: (2025)
von: Pashaeehir, Amirhossein, et al.
Veröffentlicht: (2025)
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
von: Kim, Taeyoon, et al.
Veröffentlicht: (2026)
von: Kim, Taeyoon, et al.
Veröffentlicht: (2026)
QONNECT: A QoS-Aware Orchestration System for Distributed Kubernetes Clusters
von: Aslan, Haci Ismail, et al.
Veröffentlicht: (2025)
von: Aslan, Haci Ismail, et al.
Veröffentlicht: (2025)
Environmentally-Conscious Cloud Orchestration Considering Geo-Distributed Data Centers
von: Attenni, Giulio, et al.
Veröffentlicht: (2025)
von: Attenni, Giulio, et al.
Veröffentlicht: (2025)
DECICE: AI-Driven Scheduling and Digital Twin Integration for the Cloud-HPC-Edge Compute Continuum
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2026)
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2026)
GenKubeSec: LLM-Based Kubernetes Misconfiguration Detection, Localization, Reasoning, and Remediation
von: Malul, Ehud, et al.
Veröffentlicht: (2024)
von: Malul, Ehud, et al.
Veröffentlicht: (2024)
Beyond End-to-End: Dynamic Chain Optimization for Private LLM Adaptation on the Edge
von: Wu, Yebo, et al.
Veröffentlicht: (2026)
von: Wu, Yebo, et al.
Veröffentlicht: (2026)
ScaleLLM: A Resource-Frugal LLM Serving Framework by Optimizing End-to-End Efficiency
von: Yao, Yuhang, et al.
Veröffentlicht: (2024)
von: Yao, Yuhang, et al.
Veröffentlicht: (2024)
A Proposed End-To-End Principle for Data Commons
von: Grossman, Robert L.
Veröffentlicht: (2025)
von: Grossman, Robert L.
Veröffentlicht: (2025)
SpotKube: Cost-Optimal Microservices Deployment with Cluster Autoscaling and Spot Pricing
von: Edirisinghe, Dasith, et al.
Veröffentlicht: (2024)
von: Edirisinghe, Dasith, et al.
Veröffentlicht: (2024)
End-to-End and Phase-Level Performance Optimization for Hyperledger Fabric
von: Sollu, Pavan, et al.
Veröffentlicht: (2026)
von: Sollu, Pavan, et al.
Veröffentlicht: (2026)
Saarthi: An End-to-End Intelligent Platform for Optimising Distributed Serverless Workloads
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
A Survey of End-to-End Modeling for Distributed DNN Training: Workloads, Simulators, and TCO
von: Svedas, Jonas, et al.
Veröffentlicht: (2025)
von: Svedas, Jonas, et al.
Veröffentlicht: (2025)
Visualizing Cloud-native Applications with KubeDiagrams
von: Merle, Philippe, et al.
Veröffentlicht: (2025)
von: Merle, Philippe, et al.
Veröffentlicht: (2025)
Cost-Effective Edge Data Distribution with End-To-End Delay Guarantees in Edge Computing
von: Shankar, Ravi, et al.
Veröffentlicht: (2025)
von: Shankar, Ravi, et al.
Veröffentlicht: (2025)
Kubernetes in Action: Exploring the Performance of Kubernetes Distributions in the Cloud
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
Exploring Influence Factors on LLM Suitability for No-Code Development of End User IoT Applications
von: Wang, Minghe, et al.
Veröffentlicht: (2025)
von: Wang, Minghe, et al.
Veröffentlicht: (2025)
nncase: An End-to-End Compiler for Efficient LLM Deployment on Heterogeneous Storage Architectures
von: Guo, Hui, et al.
Veröffentlicht: (2025)
von: Guo, Hui, et al.
Veröffentlicht: (2025)
A Contention-Free Model for Converged Kubernetes on HPC
von: Sochat, Vanessa, et al.
Veröffentlicht: (2024)
von: Sochat, Vanessa, et al.
Veröffentlicht: (2024)
Signalling Health for Improved Kubernetes Microservice Availability
von: Roberts, Jacob, et al.
Veröffentlicht: (2025)
von: Roberts, Jacob, et al.
Veröffentlicht: (2025)
Adaptive Resource Allocation for Workflow Containerization on Kubernetes
von: Shan, Chenggang, et al.
Veröffentlicht: (2023)
von: Shan, Chenggang, et al.
Veröffentlicht: (2023)
Odyssey: An End-to-End System for Pareto-Optimal Serverless Query Processing
von: Jesalpura, Shyam, et al.
Veröffentlicht: (2025)
von: Jesalpura, Shyam, et al.
Veröffentlicht: (2025)
Enhancing Kubernetes Resilience through Anomaly Detection and Prediction
von: Anemogiannis, V., et al.
Veröffentlicht: (2025)
von: Anemogiannis, V., et al.
Veröffentlicht: (2025)
Improving the End-to-End Efficiency of Offline Inference for Multi-LLM Applications Based on Sampling and Simulation
von: Fang, Jingzhi, et al.
Veröffentlicht: (2025)
von: Fang, Jingzhi, et al.
Veröffentlicht: (2025)
A User-centric Kubernetes-based Architecture for Green Cloud Computing
von: Zanotto, Matteo, et al.
Veröffentlicht: (2025)
von: Zanotto, Matteo, et al.
Veröffentlicht: (2025)
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
von: Mao, Ying, et al.
Veröffentlicht: (2020)
von: Mao, Ying, et al.
Veröffentlicht: (2020)
A Unified Ontology for Scalable Knowledge Graph-Driven Operational Data Analytics in High-Performance Computing Systems
von: Khan, Junaid Ahmed, et al.
Veröffentlicht: (2025)
von: Khan, Junaid Ahmed, et al.
Veröffentlicht: (2025)
LuWu: An End-to-End In-Network Out-of-Core Optimizer for 100B-Scale Model-in-Network Data-Parallel Training on Distributed GPUs
von: Sun, Mo, et al.
Veröffentlicht: (2024)
von: Sun, Mo, et al.
Veröffentlicht: (2024)
Running Cloud-native Workloads on HPC with High-Performance Kubernetes
von: Chazapis, Antony, et al.
Veröffentlicht: (2024)
von: Chazapis, Antony, et al.
Veröffentlicht: (2024)
Kubernetes in the Cloud vs. Bare Metal: A Comparative Study of Network Costs
von: Redoli, Rodrigo Mompo, et al.
Veröffentlicht: (2025)
von: Redoli, Rodrigo Mompo, et al.
Veröffentlicht: (2025)
A Kubernetes custom scheduler based on reinforcement learning for compute-intensive pods
von: Zhou, Hanlin, et al.
Veröffentlicht: (2026)
von: Zhou, Hanlin, et al.
Veröffentlicht: (2026)
The National Research Platform: Stretched, Multi-Tenant, Scientific Kubernetes Cluster
von: Weitzel, Derek, et al.
Veröffentlicht: (2025)
von: Weitzel, Derek, et al.
Veröffentlicht: (2025)
Resilience Evaluation of Kubernetes in Cloud-Edge Environments via Failure Injection
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
From Data Center IoT Telemetry to Data Analytics Chatbots -- Virtual Knowledge Graph is All You Need
von: Khan, Junaid Ahmed, et al.
Veröffentlicht: (2025)
von: Khan, Junaid Ahmed, et al.
Veröffentlicht: (2025)
Efficient Fault Localization in a Cloud Stack Using End-to-End Application Service Topology
von: Mathews, Dhanya R, et al.
Veröffentlicht: (2025)
von: Mathews, Dhanya R, et al.
Veröffentlicht: (2025)
KIS-S: A GPU-Aware Kubernetes Inference Simulator with RL-Based Auto-Scaling
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
eLLM: Elastic Memory Management Framework for Efficient LLM Serving
von: Xu, Jiale, et al.
Veröffentlicht: (2025)
von: Xu, Jiale, et al.
Veröffentlicht: (2025)
OServe: Accelerating LLM Serving via Spatial-Temporal Workload Orchestration
von: Jiang, Youhe, et al.
Veröffentlicht: (2026)
von: Jiang, Youhe, et al.
Veröffentlicht: (2026)
An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
von: Punniyamoorthy, Vinoth, et al.
Veröffentlicht: (2025)
von: Punniyamoorthy, Vinoth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KubeDSM: A Kubernetes-based Dynamic Scheduling and Migration Framework for Cloud-Assisted Edge Clusters
von: Pashaeehir, Amirhossein, et al.
Veröffentlicht: (2025) -
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
von: Kim, Taeyoon, et al.
Veröffentlicht: (2026) -
QONNECT: A QoS-Aware Orchestration System for Distributed Kubernetes Clusters
von: Aslan, Haci Ismail, et al.
Veröffentlicht: (2025) -
Environmentally-Conscious Cloud Orchestration Considering Geo-Distributed Data Centers
von: Attenni, Giulio, et al.
Veröffentlicht: (2025) -
DECICE: AI-Driven Scheduling and Digital Twin Integration for the Cloud-HPC-Edge Compute Continuum
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2026)