Saved in:
| Main Authors: | Nguyen, Hai-Duong, Tran, Xuan-The |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.17998 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Engineering AI Agents for Clinical Workflows: A Case Study in Architecture,MLOps, and Governance
by: Lopes, Cláudio Lúcio do Val, et al.
Published: (2026)
by: Lopes, Cláudio Lúcio do Val, et al.
Published: (2026)
A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents
by: Srinivasan, Vasundra
Published: (2026)
by: Srinivasan, Vasundra
Published: (2026)
Carbon-Aware Governance Gates: An Architecture for Sustainable GenAI Development
by: Abbasi, Mateen A., et al.
Published: (2026)
by: Abbasi, Mateen A., et al.
Published: (2026)
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026)
by: Garralda-Barrio, Mariano
Published: (2026)
AgentGuard: Runtime Verification of AI Agents
by: Koohestani, Roham
Published: (2025)
by: Koohestani, Roham
Published: (2025)
Leveraging AI for Enhanced Software Effort Estimation: A Comprehensive Study and Framework Proposal
by: Tran, Nhi, et al.
Published: (2024)
by: Tran, Nhi, et al.
Published: (2024)
CaveAgent: Transforming LLMs into Stateful Runtime Operators
by: Ran, Maohao, et al.
Published: (2026)
by: Ran, Maohao, et al.
Published: (2026)
Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory
by: Li, Ruiyin, et al.
Published: (2026)
by: Li, Ruiyin, et al.
Published: (2026)
MAAD: Automate Software Architecture Design through Knowledge-Driven Multi-Agent Collaboration
by: Li, Ruiyin, et al.
Published: (2025)
by: Li, Ruiyin, et al.
Published: (2025)
Debugging and Runtime Analysis of Neural Networks with VLMs (A Case Study)
by: Hu, Boyue Caroline, et al.
Published: (2025)
by: Hu, Boyue Caroline, et al.
Published: (2025)
Loosely-Structured Software: Engineering Context, Structure, and Evolution Entropy in Runtime-Rewired Multi-Agent Systems
by: Zhang, Weihao, et al.
Published: (2026)
by: Zhang, Weihao, et al.
Published: (2026)
REDO: Execution-Free Runtime Error Detection for COding Agents
by: Li, Shou, et al.
Published: (2024)
by: Li, Shou, et al.
Published: (2024)
ProbGuard: Probabilistic Runtime Monitoring for LLM Agent Safety
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents
by: Zhong, Hailin, et al.
Published: (2026)
by: Zhong, Hailin, et al.
Published: (2026)
Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes
by: Metere, Alfredo
Published: (2026)
by: Metere, Alfredo
Published: (2026)
AgentHub: A Registry for Discoverable, Verifiable, and Reproducible AI Agents
by: Pautsch, Erik, et al.
Published: (2025)
by: Pautsch, Erik, et al.
Published: (2025)
SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces
by: Xu, Duling, et al.
Published: (2026)
by: Xu, Duling, et al.
Published: (2026)
VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
MORTAR: A Model-based Runtime Action Repair Framework for AI-enabled Cyber-Physical Systems
by: Wang, Renzhi, et al.
Published: (2024)
by: Wang, Renzhi, et al.
Published: (2024)
Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study
by: Chand, Sivajeet, et al.
Published: (2026)
by: Chand, Sivajeet, et al.
Published: (2026)
Adaptive Confidence Gating in Multi-Agent Collaboration for Efficient and Optimized Code Generation
by: Zhang, Haoji, et al.
Published: (2026)
by: Zhang, Haoji, et al.
Published: (2026)
Instruction-Driven Game Engine: A Poker Case Study
by: Wu, Hongqiu, et al.
Published: (2024)
by: Wu, Hongqiu, et al.
Published: (2024)
RuntimeSlicer: Towards Generalizable Unified Runtime State Representation for Failure Management
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Automating a Complete Software Test Process Using LLMs: An Automotive Case Study
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
TDD Governance for Multi-Agent Code Generation via Prompt Engineering
by: Hasanli, Tarlan, et al.
Published: (2026)
by: Hasanli, Tarlan, et al.
Published: (2026)
SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion
by: Ma, George, et al.
Published: (2025)
by: Ma, George, et al.
Published: (2025)
Governance by Construction for Generalist Agents
by: Shlomov, Segev, et al.
Published: (2026)
by: Shlomov, Segev, et al.
Published: (2026)
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
by: Wei, Jinbiao, et al.
Published: (2026)
by: Wei, Jinbiao, et al.
Published: (2026)
AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
by: Luo, Weilin, et al.
Published: (2025)
by: Luo, Weilin, et al.
Published: (2025)
Every Software as an Agent: Blueprint and Case Study
by: Xu, Mengwei
Published: (2025)
by: Xu, Mengwei
Published: (2025)
PyVeritas: On Verifying Python via LLM-Based Transpilation and Bounded Model Checking for C
by: Orvalho, Pedro, et al.
Published: (2025)
by: Orvalho, Pedro, et al.
Published: (2025)
MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering
by: Guo, Chuanzhe, et al.
Published: (2026)
by: Guo, Chuanzhe, et al.
Published: (2026)
Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
by: Yu, Simon, et al.
Published: (2026)
by: Yu, Simon, et al.
Published: (2026)
Practical Limits of Autonomous Test Repair: A Multi-Agent Case Study with LLM-Driven Discovery and Self-Correction
by: Lee, Hyukjoo
Published: (2026)
by: Lee, Hyukjoo
Published: (2026)
RepoReviewer: A Local-First Multi-Agent Architecture for Repository-Level Code Review
by: Zhang, Peng
Published: (2026)
by: Zhang, Peng
Published: (2026)
MemGovern: Enhancing Code Agents through Learning from Governed Human Experiences
by: Wang, Qihao, et al.
Published: (2026)
by: Wang, Qihao, et al.
Published: (2026)
Agent Behavioral Contracts: Formal Specification and Runtime Enforcement for Reliable Autonomous AI Agents
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
by: Li, Yikun, et al.
Published: (2026)
by: Li, Yikun, et al.
Published: (2026)
Runtime-Structured Task Decomposition for Agentic Coding Systems
by: Asthana, Shubhi, et al.
Published: (2026)
by: Asthana, Shubhi, et al.
Published: (2026)
Similar Items
-
Engineering AI Agents for Clinical Workflows: A Case Study in Architecture,MLOps, and Governance
by: Lopes, Cláudio Lúcio do Val, et al.
Published: (2026) -
A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents
by: Srinivasan, Vasundra
Published: (2026) -
Carbon-Aware Governance Gates: An Architecture for Sustainable GenAI Development
by: Abbasi, Mateen A., et al.
Published: (2026) -
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026) -
AgentGuard: Runtime Verification of AI Agents
by: Koohestani, Roham
Published: (2025)