Building AI Agents for Autonomous Clouds: Challenges and Design Principles
Fuente:
arXiv
Saved in:
| Main Authors: | Shetty, Manish, Chen, Yinfang, Somashekar, Gagan, Ma, Minghua, Simmhan, Yogesh, Zhang, Xuchao, Mace, Jonathan, Vandevoorde, Dax, Las-Casas, Pedro, Gupta, Shachee Mishra, Nath, Suman, Bansal, Chetan, Rajmohan, Saravan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
by: Chen, Yinfang, et al.
Published: (2025)
by: Chen, Yinfang, et al.
Published: (2025)
Continuous Benchmark Generation for Evaluating Enterprise-scale LLM Agents
by: Saxena, Divyanshu, et al.
Published: (2025)
by: Saxena, Divyanshu, et al.
Published: (2025)
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4
by: Zhang, Xuchao, et al.
Published: (2024)
by: Zhang, Xuchao, et al.
Published: (2024)
Exploring LLM-based Agents for Root Cause Analysis
by: Roy, Devjeet, et al.
Published: (2024)
by: Roy, Devjeet, et al.
Published: (2024)
AMPO: Active Multi-Preference Optimization for Self-play Preference Selection
by: Gupta, Taneesh, et al.
Published: (2025)
by: Gupta, Taneesh, et al.
Published: (2025)
REFA: Reference Free Alignment for multi-preference optimization
by: Gupta, Taneesh, et al.
Published: (2024)
by: Gupta, Taneesh, et al.
Published: (2024)
Intent-based System Design and Operation
by: Anand, Vaastav, et al.
Published: (2025)
by: Anand, Vaastav, et al.
Published: (2025)
AutoAdapt: An Automated Domain Adaptation Framework for LLMs
by: Sinha, Sidharth, et al.
Published: (2026)
by: Sinha, Sidharth, et al.
Published: (2026)
Multi-Preference Optimization: Generalizing DPO via Set-Level Contrasts
by: Gupta, Taneesh, et al.
Published: (2024)
by: Gupta, Taneesh, et al.
Published: (2024)
X-lifecycle Learning for Cloud Incident Management using LLMs
by: Goel, Drishti, et al.
Published: (2024)
by: Goel, Drishti, et al.
Published: (2024)
Generative Caching for Structurally Similar Prompts and Responses
by: Chakraborty, Sarthak, et al.
Published: (2025)
by: Chakraborty, Sarthak, et al.
Published: (2025)
An Empirical Study of Production Incidents in Generative AI Cloud Services
by: Yan, Haoran, et al.
Published: (2025)
by: Yan, Haoran, et al.
Published: (2025)
Intelligent Monitoring Framework for Cloud Services: A Data-Driven Approach
by: Srinivas, Pooja, et al.
Published: (2024)
by: Srinivas, Pooja, et al.
Published: (2024)
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
by: Jaiswal, Shashwat, et al.
Published: (2025)
by: Jaiswal, Shashwat, et al.
Published: (2025)
CARMO: Dynamic Criteria Generation for Context-Aware Reward Modelling
by: Gupta, Taneesh, et al.
Published: (2024)
by: Gupta, Taneesh, et al.
Published: (2024)
eARCO: Efficient Automated Root Cause Analysis with Prompt Optimization
by: Goel, Drishti, et al.
Published: (2025)
by: Goel, Drishti, et al.
Published: (2025)
Dependency Aware Incident Linking in Large Cloud Systems
by: Ghosh, Supriyo, et al.
Published: (2024)
by: Ghosh, Supriyo, et al.
Published: (2024)
Attention Enhanced Entity Recommendation for Intelligent Monitoring in Cloud Systems
by: Hussain, Fiza, et al.
Published: (2025)
by: Hussain, Fiza, et al.
Published: (2025)
Adaptive Heuristics for Scheduling DNN Inferencing on Edge and Cloud for Personalized UAV Fleets
by: Raj, Suman, et al.
Published: (2024)
by: Raj, Suman, et al.
Published: (2024)
Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity
by: Xia, Menglin, et al.
Published: (2026)
by: Xia, Menglin, et al.
Published: (2026)
EvidenT: An Evidence-Preserving Framework for Iterative System-Level Package Repair
by: Zhao, Chenyu, et al.
Published: (2026)
by: Zhao, Chenyu, et al.
Published: (2026)
Can Language Models Go Beyond Coding? Assessing the Capability of Language Models to Build Real-World Systems
by: Zhao, Chenyu, et al.
Published: (2025)
by: Zhao, Chenyu, et al.
Published: (2025)
Synergistic Weak-Strong Collaboration by Aligning Preferences
by: Jiao, Yizhu, et al.
Published: (2025)
by: Jiao, Yizhu, et al.
Published: (2025)
Large Language Models can Deliver Accurate and Interpretable Time Series Anomaly Detection
by: Liu, Jun, et al.
Published: (2024)
by: Liu, Jun, et al.
Published: (2024)
WebXSkill: Skill Learning for Autonomous Web Agents
by: Wang, Zhaoyang, et al.
Published: (2026)
by: Wang, Zhaoyang, et al.
Published: (2026)
Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning
by: Li, Shangzhe, et al.
Published: (2026)
by: Li, Shangzhe, et al.
Published: (2026)
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents
by: Zhao, Chenyu, et al.
Published: (2026)
by: Zhao, Chenyu, et al.
Published: (2026)
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
by: Xia, Menglin, et al.
Published: (2023)
by: Xia, Menglin, et al.
Published: (2023)
AeroDaaS: A Programmable Drones-as-a-Service Platform for Intelligent Aerial Systems
by: Astu, Kautuk, et al.
Published: (2026)
by: Astu, Kautuk, et al.
Published: (2026)
AerialDB: A Federated Peer-to-Peer Spatio-temporal Edge Datastore for Drone Fleets
by: Jaiswal, Shashwat, et al.
Published: (2025)
by: Jaiswal, Shashwat, et al.
Published: (2025)
AeroDaaS: Towards an Application Programming Framework for Drones-as-a-Service
by: Raj, Suman, et al.
Published: (2025)
by: Raj, Suman, et al.
Published: (2025)
A Benchmark for Language Models in Real-World System Building
by: Jin, Weilin, et al.
Published: (2026)
by: Jin, Weilin, et al.
Published: (2026)
SynthAgent: Adapting Web Agents with Synthetic Supervision
by: Wang, Zhaoyang, et al.
Published: (2025)
by: Wang, Zhaoyang, et al.
Published: (2025)
LEGOMem: Modular Procedural Memory for Multi-agent LLM Systems for Workflow Automation
by: Han, Dongge, et al.
Published: (2025)
by: Han, Dongge, et al.
Published: (2025)
Understanding the Performance and Power of LLM Inferencing on Edge Accelerators
by: Arya, Mayank, et al.
Published: (2025)
by: Arya, Mayank, et al.
Published: (2025)
Towards AI Agents for Course Instruction in Higher Education: Early Experiences from the Field
by: Simmhan, Yogesh, et al.
Published: (2025)
by: Simmhan, Yogesh, et al.
Published: (2025)
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2026)
by: Naman, Pranjal, et al.
Published: (2026)
Optimizing Federated Learning using Remote Embeddings for Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2025)
by: Naman, Pranjal, et al.
Published: (2025)
Per-Phase Fidelity Attribution for Quantum Compilers using HBR Decomposition
by: Pati, Chandrachud, et al.
Published: (2026)
by: Pati, Chandrachud, et al.
Published: (2026)
OptimES: Optimizing Federated Learning Using Remote Embeddings for Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2025)
by: Naman, Pranjal, et al.
Published: (2025)
Similar Items
-
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
by: Chen, Yinfang, et al.
Published: (2025) -
Continuous Benchmark Generation for Evaluating Enterprise-scale LLM Agents
by: Saxena, Divyanshu, et al.
Published: (2025) -
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4
by: Zhang, Xuchao, et al.
Published: (2024) -
Exploring LLM-based Agents for Root Cause Analysis
by: Roy, Devjeet, et al.
Published: (2024) -
AMPO: Active Multi-Preference Optimization for Self-play Preference Selection
by: Gupta, Taneesh, et al.
Published: (2025)