Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Simon, Chong, Derek, Nandi, Ananjan, Soylu, Dilara, Sun, Jiuding, Manning, Christopher D, Shi, Weiyan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SWE-Shepherd: Advancing PRMs for Reinforcing Code Agents
by: Dihan, Mahir Labib, et al.
Published: (2026)
by: Dihan, Mahir Labib, et al.
Published: (2026)
REDO: Execution-Free Runtime Error Detection for COding Agents
by: Li, Shou, et al.
Published: (2024)
by: Li, Shou, et al.
Published: (2024)
Runtime Execution Traces Guided Automated Program Repair with Multi-Agent Debate
by: Wu, Jiaqing, et al.
Published: (2026)
by: Wu, Jiaqing, et al.
Published: (2026)
REDriver: Runtime Enforcement for Autonomous Vehicles
by: Sun, Yang, et al.
Published: (2024)
by: Sun, Yang, et al.
Published: (2024)
AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents
by: Zhong, Hailin, et al.
Published: (2026)
by: Zhong, Hailin, et al.
Published: (2026)
ProbGuard: Probabilistic Runtime Monitoring for LLM Agent Safety
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026)
by: Garralda-Barrio, Mariano
Published: (2026)
Logging Like Humans for LLMs: Rethinking Logging via Execution and Runtime Feedback
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
Simple Fault Localization using Execution Traces
by: Prenner, Julian Aron, et al.
Published: (2025)
by: Prenner, Julian Aron, et al.
Published: (2025)
Agent Behavioral Contracts: Formal Specification and Runtime Enforcement for Reliable Autonomous AI Agents
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation
by: He, Minghua, et al.
Published: (2025)
by: He, Minghua, et al.
Published: (2025)
AgentGuard: Runtime Verification of AI Agents
by: Koohestani, Roham
Published: (2025)
by: Koohestani, Roham
Published: (2025)
Code Digital Twin: Empowering LLMs with Tacit Knowledge for Complex Software Development
by: Peng, Xin, et al.
Published: (2025)
by: Peng, Xin, et al.
Published: (2025)
Event-B Agent: Towards LLM Agent for Formal Model Synthesis and Repair
by: Wang, Hongshu, et al.
Published: (2026)
by: Wang, Hongshu, et al.
Published: (2026)
Formalisms for Robotic Mission Specification and Execution: A Comparative Analysis
by: Filippone, Gianluca, et al.
Published: (2026)
by: Filippone, Gianluca, et al.
Published: (2026)
MCP-SandboxScan: WASM-based Secure Execution and Runtime Analysis for MCP Tools
by: Tan, Zhuoran, et al.
Published: (2026)
by: Tan, Zhuoran, et al.
Published: (2026)
Demystifying Errors in LLM Reasoning Traces: An Empirical Study of Code Execution Simulation
by: Abdollahi, Mohammad, et al.
Published: (2025)
by: Abdollahi, Mohammad, et al.
Published: (2025)
Synthesizing Efficient and Permissive Programmatic Runtime Shields for Neural Policies
by: Shi, Jieke, et al.
Published: (2024)
by: Shi, Jieke, et al.
Published: (2024)
Requirements Development and Formalization for Reliable Code Generation: A Multi-Agent Vision
by: Lu, Xu, et al.
Published: (2025)
by: Lu, Xu, et al.
Published: (2025)
CaveAgent: Transforming LLMs into Stateful Runtime Operators
by: Ran, Maohao, et al.
Published: (2026)
by: Ran, Maohao, et al.
Published: (2026)
Empowering Autonomous Debugging Agents with Efficient Dynamic Analysis
by: Xiang, Jiahong, et al.
Published: (2026)
by: Xiang, Jiahong, et al.
Published: (2026)
KAIJU: An Executive Kernel for Intent-Gated Execution of LLM Agents
by: Guerin, Cormac, et al.
Published: (2026)
by: Guerin, Cormac, et al.
Published: (2026)
Assessing the Capability of Android Dynamic Analysis Tools to Combat Anti-Runtime Analysis Techniques
by: Suo, Dewen, et al.
Published: (2025)
by: Suo, Dewen, et al.
Published: (2025)
Runtime Enforcement for Operationalizing Ethics in Autonomous Systems
by: De Sanctis, Martina, et al.
Published: (2026)
by: De Sanctis, Martina, et al.
Published: (2026)
The Runtime Dimension of Ethics in Self-Adaptive Systems
by: Autili, Marco, et al.
Published: (2026)
by: Autili, Marco, et al.
Published: (2026)
Research on WebAssembly Runtimes: A Survey
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
Formal Analysis of the Contract Automata Runtime Environment with Uppaal: Modelling, Verification and Testing
by: Basile, Davide
Published: (2025)
by: Basile, Davide
Published: (2025)
Towards Effective Detection of Ponzi schemes on Ethereum with Contract Runtime Behavior Graph
by: Liang, Ruichao, et al.
Published: (2024)
by: Liang, Ruichao, et al.
Published: (2024)
LLMs as Firmware Experts: A Runtime-Grown Tree-of-Agents Framework
by: Zhang, Xiangrui, et al.
Published: (2025)
by: Zhang, Xiangrui, et al.
Published: (2025)
SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces
by: Xu, Duling, et al.
Published: (2026)
by: Xu, Duling, et al.
Published: (2026)
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs
by: Haque, Mirazul, et al.
Published: (2025)
by: Haque, Mirazul, et al.
Published: (2025)
CASET: Complexity Analysis using Simple Execution Traces for CS* submissions
by: Mehta, Aaryen, et al.
Published: (2024)
by: Mehta, Aaryen, et al.
Published: (2024)
ArgRE: Formal Argumentation for Conflict Resolution in Multi-Agent Requirements Negotiation
by: Cheng, Haowei, et al.
Published: (2026)
by: Cheng, Haowei, et al.
Published: (2026)
Trace: Securing Smart Contract Repository Against Access Control Vulnerability
by: Chen, Chong, et al.
Published: (2025)
by: Chen, Chong, et al.
Published: (2025)
Constraint-Guided Multi-Agent Decompilation for Executable Binary Recovery
by: Zhang, Yifan, et al.
Published: (2026)
by: Zhang, Yifan, et al.
Published: (2026)
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
by: Ashrafi, Nazmus, et al.
Published: (2025)
by: Ashrafi, Nazmus, et al.
Published: (2025)
XAI for Coding Agent Failures: Transforming Raw Execution Traces into Actionable Insights
by: Joshi, Arun
Published: (2026)
by: Joshi, Arun
Published: (2026)
Compiling Code LLMs into Lightweight Executables
by: Shi, Jieke, et al.
Published: (2026)
by: Shi, Jieke, et al.
Published: (2026)
K-CIRCT: A Layered, Composable, and Executable Formal Semantics for CIRCT Hardware IRs
by: Zhao, Jianhong, et al.
Published: (2024)
by: Zhao, Jianhong, et al.
Published: (2024)
A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents
by: Srinivasan, Vasundra
Published: (2026)
by: Srinivasan, Vasundra
Published: (2026)
Similar Items
-
SWE-Shepherd: Advancing PRMs for Reinforcing Code Agents
by: Dihan, Mahir Labib, et al.
Published: (2026) -
REDO: Execution-Free Runtime Error Detection for COding Agents
by: Li, Shou, et al.
Published: (2024) -
Runtime Execution Traces Guided Automated Program Repair with Multi-Agent Debate
by: Wu, Jiaqing, et al.
Published: (2026) -
REDriver: Runtime Enforcement for Autonomous Vehicles
by: Sun, Yang, et al.
Published: (2024) -
AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents
by: Zhong, Hailin, et al.
Published: (2026)