Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
Fuente:
arXiv
Saved in:
| Main Author: | Mukkolakkal, Muhamed Ramees Cheriya |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
1.5 Million Messages Per Second on 3 Machines: Benchmarking and Latency Optimization of Apache Pulsar at Enterprise Scale
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
by: Will, Jonathan, et al.
Published: (2025)
by: Will, Jonathan, et al.
Published: (2025)
Deadline-Aware Joint Task Scheduling and Offloading in Mobile Edge Computing Systems
by: Nguyen, Ngoc Hung, et al.
Published: (2025)
by: Nguyen, Ngoc Hung, et al.
Published: (2025)
GPUnion: Autonomous GPU Sharing on Campus
by: Li, Yufang, et al.
Published: (2025)
by: Li, Yufang, et al.
Published: (2025)
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
by: Pathak, Prashant Kumar, et al.
Published: (2026)
by: Pathak, Prashant Kumar, et al.
Published: (2026)
Knowledge Graphs-Driven Intelligence for Distributed Decision Systems
by: Napoli, Rosario, et al.
Published: (2026)
by: Napoli, Rosario, et al.
Published: (2026)
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
by: Woisetschläger, Herbert, et al.
Published: (2023)
by: Woisetschläger, Herbert, et al.
Published: (2023)
Studying the Effect of Schedule Preemption on Dynamic Task Graph Scheduling
by: Khodabandehlou, Mohammadali, et al.
Published: (2026)
by: Khodabandehlou, Mohammadali, et al.
Published: (2026)
Serverless GPU Architecture for Enterprise HR Analytics: A Production-Scale BDaaS Implementation
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
by: Woisetschläger, Herbert, et al.
Published: (2023)
by: Woisetschläger, Herbert, et al.
Published: (2023)
Operational Memory Architecture for Kubernetes:Preserving Causal Context Across the Evidence Horizon
by: Khan, Shamsher
Published: (2026)
by: Khan, Shamsher
Published: (2026)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
by: Murimi, Almond Kiruthu
Published: (2025)
by: Murimi, Almond Kiruthu
Published: (2025)
Cognitive Infrastructure: A Unified DCIM Framework for AI Data Centers
by: Sunkara, Krishna Chaitanya
Published: (2026)
by: Sunkara, Krishna Chaitanya
Published: (2026)
Mobile Traffic Prediction at the Edge Through Distributed and Deep Transfer Learning
by: Petrella, Alfredo, et al.
Published: (2023)
by: Petrella, Alfredo, et al.
Published: (2023)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)
by: Leung, Derek, et al.
Published: (2025)
Analysis of Design Patterns and Benchmark Practices in Apache Kafka Event-Streaming Systems
by: Mohammad, Muzeeb
Published: (2025)
by: Mohammad, Muzeeb
Published: (2025)
Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems
by: Sharma, Aasish Kumar, et al.
Published: (2026)
by: Sharma, Aasish Kumar, et al.
Published: (2026)
Nezha: Deployable and High-Performance Consensus Using Synchronized Clocks
by: Geng, Jinkun, et al.
Published: (2022)
by: Geng, Jinkun, et al.
Published: (2022)
Vertical Federated Image Segmentation
by: Mandal, Paul K., et al.
Published: (2024)
by: Mandal, Paul K., et al.
Published: (2024)
Horizontal Federated Computer Vision
by: Mandal, Paul K., et al.
Published: (2023)
by: Mandal, Paul K., et al.
Published: (2023)
StepCache: Step-Level Reuse with Lightweight Verification and Selective Patching for LLM Serving
by: Nouri, Azam
Published: (2026)
by: Nouri, Azam
Published: (2026)
Spark-LLM-Eval: A Distributed Framework for Statistically Rigorous Large Language Model Evaluation
by: Mitra, Subhadip
Published: (2026)
by: Mitra, Subhadip
Published: (2026)
SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis Detection
by: Ospitia, Santiago, et al.
Published: (2026)
by: Ospitia, Santiago, et al.
Published: (2026)
Intersections of Web3 and AI -- View in 2024
by: Hyland-Wood, David, et al.
Published: (2024)
by: Hyland-Wood, David, et al.
Published: (2024)
Accelerating Causal Algorithms for Industrial-scale Data: A Distributed Computing Approach with Ray Framework
by: Verma, Vishal, et al.
Published: (2024)
by: Verma, Vishal, et al.
Published: (2024)
dpBento: Benchmarking DPUs for Data Processing
by: Hu, Jiasheng, et al.
Published: (2025)
by: Hu, Jiasheng, et al.
Published: (2025)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026)
by: Li, Zhifei, et al.
Published: (2026)
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026)
by: Bai, Tianyu, et al.
Published: (2026)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM Inference
by: Sidik, Bronislav, et al.
Published: (2026)
by: Sidik, Bronislav, et al.
Published: (2026)
ACME: Adaptive Customization of Large Models via Distributed Systems
by: Dai, Ziming, et al.
Published: (2025)
by: Dai, Ziming, et al.
Published: (2025)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
by: Li, Zonghang, et al.
Published: (2024)
by: Li, Zonghang, et al.
Published: (2024)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
by: Nguyen, Thinh, et al.
Published: (2025)
by: Nguyen, Thinh, et al.
Published: (2025)
An Empirical Study of the Impact of Federated Learning on Machine Learning Model Accuracy
by: Yang, Haotian, et al.
Published: (2025)
by: Yang, Haotian, et al.
Published: (2025)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
by: Erben, Alexander, et al.
Published: (2023)
by: Erben, Alexander, et al.
Published: (2023)
Autonomous Edge-Deployed AI Agents for Electric Vehicle Charging Infrastructure Management
by: Cherifi, Mohammed
Published: (2026)
by: Cherifi, Mohammed
Published: (2026)
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
by: Iannelli, Michael, et al.
Published: (2024)
by: Iannelli, Michael, et al.
Published: (2024)
Efficient Construction of Large Search Spaces for Auto-Tuning
by: Willemsen, Floris-Jan, et al.
Published: (2025)
by: Willemsen, Floris-Jan, et al.
Published: (2025)
DDS: DPU-optimized Disaggregated Storage [Extended Report]
by: Zhang, Qizhen, et al.
Published: (2024)
by: Zhang, Qizhen, et al.
Published: (2024)
Similar Items
-
1.5 Million Messages Per Second on 3 Machines: Benchmarking and Latency Optimization of Apache Pulsar at Enterprise Scale
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026) -
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
by: Will, Jonathan, et al.
Published: (2025) -
Deadline-Aware Joint Task Scheduling and Offloading in Mobile Edge Computing Systems
by: Nguyen, Ngoc Hung, et al.
Published: (2025) -
GPUnion: Autonomous GPU Sharing on Campus
by: Li, Yufang, et al.
Published: (2025) -
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
by: Pathak, Prashant Kumar, et al.
Published: (2026)