AgentGuard: Runtime Verification of AI Agents
Fuente:
arXiv
Saved in:
| Main Author: | Koohestani, Roham |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking IDE Customization for Enhanced HAX: A Hyperdimensional Perspective
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
Benchmarking AI Models in Software Engineering: A Review, Search Tool, and Unified Approach for Elevating Benchmark Quality
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
by: Deljouyi, Amirhossein, et al.
Published: (2024)
by: Deljouyi, Amirhossein, et al.
Published: (2024)
TriCEGAR: A Trace-Driven Abstraction Mechanism for Agentic AI
by: Koohestani, Roham, et al.
Published: (2026)
by: Koohestani, Roham, et al.
Published: (2026)
AST-PAC: AST-guided Membership Inference for Code
by: Koohestani, Roham, et al.
Published: (2026)
by: Koohestani, Roham, et al.
Published: (2026)
ProbGuard: Probabilistic Runtime Monitoring for LLM Agent Safety
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Code4MeV2: a Research-oriented Code-completion Platform
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
Does In-IDE Calibration of Large Language Models work at Scale?
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents
by: Zhong, Hailin, et al.
Published: (2026)
by: Zhong, Hailin, et al.
Published: (2026)
HyperSeq: A Hyper-Adaptive Representation for Predictive Sequencing of States
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
CaveAgent: Transforming LLMs into Stateful Runtime Operators
by: Ran, Maohao, et al.
Published: (2026)
by: Ran, Maohao, et al.
Published: (2026)
Agent Behavioral Contracts: Formal Specification and Runtime Enforcement for Reliable Autonomous AI Agents
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
REDO: Execution-Free Runtime Error Detection for COding Agents
by: Li, Shou, et al.
Published: (2024)
by: Li, Shou, et al.
Published: (2024)
AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime
by: Su, Jianhao, et al.
Published: (2026)
by: Su, Jianhao, et al.
Published: (2026)
SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces
by: Xu, Duling, et al.
Published: (2026)
by: Xu, Duling, et al.
Published: (2026)
A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents
by: Srinivasan, Vasundra
Published: (2026)
by: Srinivasan, Vasundra
Published: (2026)
BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks
by: Tu, Xinming, et al.
Published: (2026)
by: Tu, Xinming, et al.
Published: (2026)
UnitTenX: Generating Tests for Legacy Packages with AI Agents Powered by Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2025)
by: Charalambous, Yiannis, et al.
Published: (2025)
AdaptiveGuard: Towards Adaptive Runtime Safety for LLM-Powered Software
by: Yang, Rui, et al.
Published: (2025)
by: Yang, Rui, et al.
Published: (2025)
AgentGuard: A Multi-Agent Framework for Robust Package Confusion Detection via Hybrid Search and Metadata-Content Fusion
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
Loosely-Structured Software: Engineering Context, Structure, and Evolution Entropy in Runtime-Rewired Multi-Agent Systems
by: Zhang, Weihao, et al.
Published: (2026)
by: Zhang, Weihao, et al.
Published: (2026)
The AI Agent Index
by: Casper, Stephen, et al.
Published: (2025)
by: Casper, Stephen, et al.
Published: (2025)
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026)
by: Garralda-Barrio, Mariano
Published: (2026)
Solver-Aided Verification of Policy Compliance in Tool-Augmented LLM Agents
by: Winston, Cailin, et al.
Published: (2026)
by: Winston, Cailin, et al.
Published: (2026)
An Empirical Study of Agent Developer Practices in AI Agent Frameworks
by: Wang, Yanlin, et al.
Published: (2025)
by: Wang, Yanlin, et al.
Published: (2025)
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
by: Nguyen, Hai-Duong, et al.
Published: (2026)
by: Nguyen, Hai-Duong, et al.
Published: (2026)
ROSMonitoring 2.0: Extending ROS Runtime Verification to Services and Ordered Topics
by: Saadat, Maryam Ghaffari, et al.
Published: (2024)
by: Saadat, Maryam Ghaffari, et al.
Published: (2024)
AgentSLA : Towards a Service Level Agreement for AI Agents
by: Jouneaux, Gwendal, et al.
Published: (2025)
by: Jouneaux, Gwendal, et al.
Published: (2025)
AgentHub: A Registry for Discoverable, Verifiable, and Reproducible AI Agents
by: Pautsch, Erik, et al.
Published: (2025)
by: Pautsch, Erik, et al.
Published: (2025)
Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification
by: He, Zehai, et al.
Published: (2026)
by: He, Zehai, et al.
Published: (2026)
AgentMesh: A Cooperative Multi-Agent Generative AI Framework for Software Development Automation
by: Khanzadeh, Sourena
Published: (2025)
by: Khanzadeh, Sourena
Published: (2025)
Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
by: Yu, Simon, et al.
Published: (2026)
by: Yu, Simon, et al.
Published: (2026)
RuntimeSlicer: Towards Generalizable Unified Runtime State Representation for Failure Management
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Unified Software Engineering Agent as AI Software Engineer
by: Applis, Leonhard, et al.
Published: (2025)
by: Applis, Leonhard, et al.
Published: (2025)
MASAI: Modular Architecture for Software-engineering AI Agents
by: Arora, Daman, et al.
Published: (2024)
by: Arora, Daman, et al.
Published: (2024)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Skilled AI Agents for Embedded and IoT Systems Development
by: Li, Yiming, et al.
Published: (2026)
by: Li, Yiming, et al.
Published: (2026)
Agentsway -- Software Development Methodology for AI Agents-based Teams
by: Bandara, Eranga, et al.
Published: (2025)
by: Bandara, Eranga, et al.
Published: (2025)
PBT-Bench: Benchmarking AI Agents on Property-Based Testing
by: Jing, Lucas, et al.
Published: (2026)
by: Jing, Lucas, et al.
Published: (2026)
Formal Architecture Descriptors as Navigation Primitives for AI Coding Agents
by: Jin, Ruoqi
Published: (2026)
by: Jin, Ruoqi
Published: (2026)
Similar Items
-
Rethinking IDE Customization for Enhanced HAX: A Hyperdimensional Perspective
by: Koohestani, Roham, et al.
Published: (2025) -
Benchmarking AI Models in Software Engineering: A Review, Search Tool, and Unified Approach for Elevating Benchmark Quality
by: Koohestani, Roham, et al.
Published: (2025) -
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
by: Deljouyi, Amirhossein, et al.
Published: (2024) -
TriCEGAR: A Trace-Driven Abstraction Mechanism for Agentic AI
by: Koohestani, Roham, et al.
Published: (2026) -
AST-PAC: AST-guided Membership Inference for Code
by: Koohestani, Roham, et al.
Published: (2026)