RuntimeSlicer: Towards Generalizable Unified Runtime State Representation for Failure Management
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Lingzhe, Jia, Tong, Hong, Weijie, Wang, Mingyu, Duan, Chiming, He, Minghua, Wang, Rongqian, Peng, Xi, Wang, Meiling, Zhang, Gong, Chen, Renhai, Li, Ying |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Failure Management for Multi-Agent Systems with Reasoning Trace Representation
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Adaptive Root Cause Localization for Microservice Systems with Multi-Agent Recursion-of-Thought
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
AgentFM: Role-Aware Failure Management for Distributed Databases with LLM-Driven Multi-Agents
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
E2E-REME: Towards End-to-End Microservices Auto-Remediation via Experience-Simulation Reinforcement Fine-Tuning
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Agentic Memory Enhanced Recursive Reasoning for Root Cause Localization in Microservices
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Hypothesize-Then-Verify: Speculative Root Cause Analysis for Microservices with Pathwise Parallelism
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
United We Stand: Towards End-to-End Log-based Fault Diagnosis via Interactive Multi-Task Learning
by: He, Minghua, et al.
Published: (2025)
by: He, Minghua, et al.
Published: (2025)
Walk the Talk: Is Your Log-based Software Reliability Maintenance System Really Reliable?
by: He, Minghua, et al.
Published: (2025)
by: He, Minghua, et al.
Published: (2025)
MicroRemed: Benchmarking LLMs in Microservices Remediation
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
ThinkFL: Self-Refining Failure Localization for Microservice Systems via Reinforcement Fine-Tuning
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
LogDB: Multivariate Log-based Failure Diagnosis for Distributed Databases (Extended from MultiLog)
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
Towards Robust LLM Post-Training: Automatic Failure Management for Reinforcement Fine-Tuning
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents
by: Zhang, Xi, et al.
Published: (2026)
by: Zhang, Xi, et al.
Published: (2026)
LogAction: Consistent Cross-system Anomaly Detection through Logs via Active Domain Adaptation
by: Duan, Chiming, et al.
Published: (2025)
by: Duan, Chiming, et al.
Published: (2025)
Runtime Analysis of the SMS-EMOA for Many-Objective Optimization
by: Zheng, Weijie, et al.
Published: (2023)
by: Zheng, Weijie, et al.
Published: (2023)
ACE Runtime - A ZKP-Native Blockchain Runtime with Sub-Second Cryptographic Finality
by: Wang, Jian Sheng
Published: (2026)
by: Wang, Jian Sheng
Published: (2026)
Failure Prediction at Runtime for Generative Robot Policies
by: Römer, Ralf, et al.
Published: (2025)
by: Römer, Ralf, et al.
Published: (2025)
Towards Agentic Runtime Healing
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
Slicer Networks
by: Zhang, Hang, et al.
Published: (2024)
by: Zhang, Hang, et al.
Published: (2024)
A Survey of AIOps for Failure Management in the Era of Large Language Models
by: Zhang, Lingzhe, et al.
Published: (2024)
by: Zhang, Lingzhe, et al.
Published: (2024)
Clove: Object-Level CXL Memory Management in Managed Runtimes
by: Son, Sam, et al.
Published: (2026)
by: Son, Sam, et al.
Published: (2026)
Runtime Analysis for the NSGA-II: Proving, Quantifying, and Explaining the Inefficiency For Many Objectives
by: Zheng, Weijie, et al.
Published: (2022)
by: Zheng, Weijie, et al.
Published: (2022)
Research on WebAssembly Runtimes: A Survey
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
A Readiness-Driven Runtime for Pipeline-Parallel Training under Runtime Variability
by: Liu, Ruitao, et al.
Published: (2026)
by: Liu, Ruitao, et al.
Published: (2026)
Reducing Events to Augment Log-based Anomaly Detection Models: An Empirical Study
by: Zhang, Lingzhe, et al.
Published: (2024)
by: Zhang, Lingzhe, et al.
Published: (2024)
Autonomous Action Runtime Management(AARM):A System Specification for Securing AI-Driven Actions at Runtime
by: Errico, Herman
Published: (2026)
by: Errico, Herman
Published: (2026)
Sealing the Audit-Runtime Gap for LLM Skills
by: Shen, Tingda, et al.
Published: (2026)
by: Shen, Tingda, et al.
Published: (2026)
AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills
by: Zhuang, Haomin, et al.
Published: (2026)
by: Zhuang, Haomin, et al.
Published: (2026)
ScaleDL: Towards Scalable and Efficient Runtime Prediction for Distributed Deep Learning Workloads
by: Wang, Xiaokai, et al.
Published: (2025)
by: Wang, Xiaokai, et al.
Published: (2025)
Runtime Consultants
by: Fisman, Dana, et al.
Published: (2025)
by: Fisman, Dana, et al.
Published: (2025)
Fourier Analysis Meets Runtime Analysis: Precise Runtimes on Plateaus
by: Doerr, Benjamin, et al.
Published: (2023)
by: Doerr, Benjamin, et al.
Published: (2023)
GNMR: Runtime Stability Control for Low-Precision Large Language Model Training
by: Kong, Boao, et al.
Published: (2026)
by: Kong, Boao, et al.
Published: (2026)
Runtime Backdoor Detection for Federated Learning via Representational Dissimilarity Analysis
by: Zhang, Xiyue, et al.
Published: (2025)
by: Zhang, Xiyue, et al.
Published: (2025)
RvLLM: LLM Runtime Verification with Domain Knowledge
by: Zhang, Yedi, et al.
Published: (2025)
by: Zhang, Yedi, et al.
Published: (2025)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
by: Agia, Christopher, et al.
Published: (2024)
by: Agia, Christopher, et al.
Published: (2024)
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
by: Park, Seongheon, et al.
Published: (2026)
by: Park, Seongheon, et al.
Published: (2026)
Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents
by: Meng, M.
Published: (2026)
by: Meng, M.
Published: (2026)
LLMs as Firmware Experts: A Runtime-Grown Tree-of-Agents Framework
by: Zhang, Xiangrui, et al.
Published: (2025)
by: Zhang, Xiangrui, et al.
Published: (2025)
Similar Items
-
Efficient Failure Management for Multi-Agent Systems with Reasoning Trace Representation
by: Zhang, Lingzhe, et al.
Published: (2026) -
Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought
by: Zhang, Lingzhe, et al.
Published: (2026) -
Adaptive Root Cause Localization for Microservice Systems with Multi-Agent Recursion-of-Thought
by: Zhang, Lingzhe, et al.
Published: (2025) -
AgentFM: Role-Aware Failure Management for Distributed Databases with LLM-Driven Multi-Agents
by: Zhang, Lingzhe, et al.
Published: (2025) -
E2E-REME: Towards End-to-End Microservices Auto-Remediation via Experience-Simulation Reinforcement Fine-Tuning
by: Zhang, Lingzhe, et al.
Published: (2026)