PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Cui, Shengkun, Krishna, Rahul, Jha, Saurabh, Iyer, Ravishankar K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simplifying Root Cause Analysis in Kubernetes with StateGraph and LLM
by: Xiang, Yong, et al.
Published: (2025)
by: Xiang, Yong, et al.
Published: (2025)
Causal AI-based Root Cause Identification: Research to Practice at Scale
by: Jha, Saurabh, et al.
Published: (2025)
by: Jha, Saurabh, et al.
Published: (2025)
A Survey on Failure Analysis and Fault Injection in AI Systems
by: Yu, Guangba, et al.
Published: (2024)
by: Yu, Guangba, et al.
Published: (2024)
Root Cause Analysis for Microservice Systems via Cascaded Conditional Learning with Hypergraphs
by: Xie, Shuaiyu, et al.
Published: (2025)
by: Xie, Shuaiyu, et al.
Published: (2025)
LLOR: Automated Repair of OpenMP Programs
by: Bora, Utpal, et al.
Published: (2024)
by: Bora, Utpal, et al.
Published: (2024)
Adaptive Protein Design Protocols and Middleware
by: Alsaadi, Aymen, et al.
Published: (2025)
by: Alsaadi, Aymen, et al.
Published: (2025)
Hierarchical Autoscaling for Large Language Model Serving with Chiron
by: Patke, Archit, et al.
Published: (2025)
by: Patke, Archit, et al.
Published: (2025)
Container-level Energy Observability in Kubernetes Clusters
by: Pijnacker, Bjorn, et al.
Published: (2025)
by: Pijnacker, Bjorn, et al.
Published: (2025)
SemanticForge: Repository-Level Code Generation through Semantic Knowledge Graphs and Constraint Satisfaction
by: Zhang, Wuyang, et al.
Published: (2025)
by: Zhang, Wuyang, et al.
Published: (2025)
Intelligent Load Balancing in Cloud Computer Systems
by: Sliwko, Leszek
Published: (2025)
by: Sliwko, Leszek
Published: (2025)
On the Inference (In-)Security of Vertical Federated Learning: Efficient Auditing against Inference Tampering Attack
by: Huang, Chung-ju, et al.
Published: (2025)
by: Huang, Chung-ju, et al.
Published: (2025)
SysMoBench: Evaluating AI on Formally Modeling Complex Real-World Systems
by: Cheng, Qian, et al.
Published: (2025)
by: Cheng, Qian, et al.
Published: (2025)
VibeCodeHPC: An Agent-Based Iterative Prompting Auto-Tuner for HPC Code Generation Using LLMs
by: Hayashi, Shun-ichiro, et al.
Published: (2025)
by: Hayashi, Shun-ichiro, et al.
Published: (2025)
LLMs as Packagers of HPC Software
by: Melone, Caetano, et al.
Published: (2025)
by: Melone, Caetano, et al.
Published: (2025)
FIRST: Federated Inference Resource Scheduling Toolkit for Scientific AI Model Access
by: Tanikanti, Aditya, et al.
Published: (2025)
by: Tanikanti, Aditya, et al.
Published: (2025)
HPCAgentTester: A Multi-Agent LLM Approach for Enhanced HPC Unit Test Generation
by: Karanjai, Rabimba, et al.
Published: (2025)
by: Karanjai, Rabimba, et al.
Published: (2025)
From Edge to HPC: Investigating Cross-Facility Data Streaming Architectures
by: George, Anjus, et al.
Published: (2025)
by: George, Anjus, et al.
Published: (2025)
Domain Adaptation-based Edge Computing for Cross-Conditions Fault Diagnosis
by: Wang, Yanzhi, et al.
Published: (2024)
by: Wang, Yanzhi, et al.
Published: (2024)
Bridging the Gap: Empowering Small Models in Reliable OpenACC-based Parallelization via GEPA-Optimized Prompting
by: Jhaveri, Samyak, et al.
Published: (2026)
by: Jhaveri, Samyak, et al.
Published: (2026)
MMO: Meta Multi-Objectivization for Software Configuration Tuning
by: Chen, Pengzhou, et al.
Published: (2021)
by: Chen, Pengzhou, et al.
Published: (2021)
Performance-Aligned LLMs for Generating Fast Code
by: Nichols, Daniel, et al.
Published: (2024)
by: Nichols, Daniel, et al.
Published: (2024)
Atmosphere: Context and situational-aware collaborative IoT architecture for edge-fog-cloud computing
by: Ortiz, Guadalupe, et al.
Published: (2024)
by: Ortiz, Guadalupe, et al.
Published: (2024)
Ten Years of Teaching Empirical Software Engineering in the context of Energy-efficient Software
by: Malavolta, Ivano, et al.
Published: (2024)
by: Malavolta, Ivano, et al.
Published: (2024)
Why Does the LLM Stop Computing: An Empirical Study of User-Reported Failures in Open-Source LLMs
by: Yu, Guangba, et al.
Published: (2026)
by: Yu, Guangba, et al.
Published: (2026)
From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications
by: Thareja, Komal, et al.
Published: (2026)
by: Thareja, Komal, et al.
Published: (2026)
Building AI Agents for Autonomous Clouds: Challenges and Design Principles
by: Shetty, Manish, et al.
Published: (2024)
by: Shetty, Manish, et al.
Published: (2024)
(POSTER) From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications
by: Thareja, Komal, et al.
Published: (2026)
by: Thareja, Komal, et al.
Published: (2026)
Adapting Multi-objectivized Software Configuration Tuning
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads
by: Agyemang, Justice Owusu, et al.
Published: (2026)
by: Agyemang, Justice Owusu, et al.
Published: (2026)
Hydra: Brokering Cloud and HPC Resources to Support the Execution of Heterogeneous Workloads at Scale
by: Alsaadi, Aymen, et al.
Published: (2024)
by: Alsaadi, Aymen, et al.
Published: (2024)
Story of Two GPUs: Characterizing the Resilience of Hopper H100 and Ampere A100 GPUs
by: Cui, Shengkun, et al.
Published: (2025)
by: Cui, Shengkun, et al.
Published: (2025)
Testing learning-enabled cyber-physical systems with Large-Language Models: A Formal Approach
by: Zheng, Xi, et al.
Published: (2023)
by: Zheng, Xi, et al.
Published: (2023)
An Analysis of HPC and Edge Architectures in the Cloud
by: Santillan, Steven, et al.
Published: (2025)
by: Santillan, Steven, et al.
Published: (2025)
On the correlation between Architectural Smells and Static Analysis Warnings
by: Esposito, Matteo, et al.
Published: (2024)
by: Esposito, Matteo, et al.
Published: (2024)
Cilium and VDM -- Towards Formal Analysis of Cilium Policies
by: Kulik, Tomas, et al.
Published: (2024)
by: Kulik, Tomas, et al.
Published: (2024)
Complexity at Scale: A Quantitative Analysis of an Alibaba Microservice Deployment
by: Winchester, Giles, et al.
Published: (2025)
by: Winchester, Giles, et al.
Published: (2025)
AI for Distributed Systems Design: Scalable Cloud Optimization Through Repeated LLMs Sampling And Simulators
by: Tagliabue, Jacopo
Published: (2025)
by: Tagliabue, Jacopo
Published: (2025)
Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions
by: Teranishi, Keita, et al.
Published: (2025)
by: Teranishi, Keita, et al.
Published: (2025)
Open Challenges in the Formal Verification of Autonomous Driving
by: Burgio, Paolo, et al.
Published: (2024)
by: Burgio, Paolo, et al.
Published: (2024)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
by: Liu, Shu, et al.
Published: (2026)
by: Liu, Shu, et al.
Published: (2026)
Similar Items
-
Simplifying Root Cause Analysis in Kubernetes with StateGraph and LLM
by: Xiang, Yong, et al.
Published: (2025) -
Causal AI-based Root Cause Identification: Research to Practice at Scale
by: Jha, Saurabh, et al.
Published: (2025) -
A Survey on Failure Analysis and Fault Injection in AI Systems
by: Yu, Guangba, et al.
Published: (2024) -
Root Cause Analysis for Microservice Systems via Cascaded Conditional Learning with Hypergraphs
by: Xie, Shuaiyu, et al.
Published: (2025) -
LLOR: Automated Repair of OpenMP Programs
by: Bora, Utpal, et al.
Published: (2024)