Governance by Construction for Generalist Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Shlomov, Segev, Shoham, Iftach, Oved, Alon, Levy, Ido, Marreed, Sami, Ship, Harold, Akrabi, Offer, Zeltyn, Sergey, Yaeli, Avi, Mashkif, Nir |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Enterprise-Ready Computer Using Generalist Agent
by: Marreed, Sami, et al.
Published: (2025)
by: Marreed, Sami, et al.
Published: (2025)
SNAP: Semantic Stories for Next Activity Prediction
by: Oved, Alon, et al.
Published: (2024)
by: Oved, Alon, et al.
Published: (2024)
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
by: Shlomov, Segev, et al.
Published: (2025)
by: Shlomov, Segev, et al.
Published: (2025)
ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
by: Levy, Ido, et al.
Published: (2024)
by: Levy, Ido, et al.
Published: (2024)
IDA: Breaking Barriers in No-code UI Automation Through Large Language Models and Human-Centric Design
by: Shlomov, Segev, et al.
Published: (2024)
by: Shlomov, Segev, et al.
Published: (2024)
TabAgent: A Framework for Replacing Agentic Generative Components with Tabular-Textual Classifiers
by: Levy, Ido, et al.
Published: (2026)
by: Levy, Ido, et al.
Published: (2026)
AgentFixer: From Failure Detection to Fix Recommendations in LLM Agentic Systems
by: Mulian, Hadar, et al.
Published: (2026)
by: Mulian, Hadar, et al.
Published: (2026)
Textual Planning with Explicit Latent Transitions
by: Shlomi, Eliezer, et al.
Published: (2026)
by: Shlomi, Eliezer, et al.
Published: (2026)
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
by: Shlomov, Segev, et al.
Published: (2024)
by: Shlomov, Segev, et al.
Published: (2024)
An Agent-Based Framework for the Automatic Validation of Mathematical Optimization Models
by: Zadorojniy, Alexander, et al.
Published: (2025)
by: Zadorojniy, Alexander, et al.
Published: (2025)
HyperAgent: Generalist Software Engineering Agents to Solve Coding Tasks at Scale
by: Phan, Huy Nhat, et al.
Published: (2024)
by: Phan, Huy Nhat, et al.
Published: (2024)
Correct-by-Construction Design of Contextual Robotic Missions Using Contracts
by: Mallozzi, Piergiuseppe, et al.
Published: (2023)
by: Mallozzi, Piergiuseppe, et al.
Published: (2023)
OpenHands: An Open Platform for AI Software Developers as Generalist Agents
by: Wang, Xingyao, et al.
Published: (2024)
by: Wang, Xingyao, et al.
Published: (2024)
SetupBench: Assessing Software Engineering Agents' Ability to Bootstrap Development Environments
by: Arora, Avi, et al.
Published: (2025)
by: Arora, Avi, et al.
Published: (2025)
Experimenting with Multi-Agent Software Development: Towards a Unified Platform
by: Sami, Malik Abdul, et al.
Published: (2024)
by: Sami, Malik Abdul, et al.
Published: (2024)
MemGovern: Enhancing Code Agents through Learning from Governed Human Experiences
by: Wang, Qihao, et al.
Published: (2026)
by: Wang, Qihao, et al.
Published: (2026)
From Generalist to Specialist: Exploring CWE-Specific Vulnerability Detection
by: Atiiq, Syafiq Al, et al.
Published: (2024)
by: Atiiq, Syafiq Al, et al.
Published: (2024)
Towards Automated Governance: A DSL for Human-Agent Collaboration in Software Projects
by: Ait, Adem, et al.
Published: (2025)
by: Ait, Adem, et al.
Published: (2025)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
by: Huang, Huihui, et al.
Published: (2026)
by: Huang, Huihui, et al.
Published: (2026)
An Exploratory Study of the Relationship between SATD and Other Software Development Activities
by: Esfandiari, Shima, et al.
Published: (2024)
by: Esfandiari, Shima, et al.
Published: (2024)
A Construction-Phase Digital Twin Framework for Quality Assurance and Decision Support in Civil Infrastructure Projects
by: Islam, Md Asiful, et al.
Published: (2026)
by: Islam, Md Asiful, et al.
Published: (2026)
Automata Models for Effective Bug Pattern Description
by: Yaacov, Tom, et al.
Published: (2025)
by: Yaacov, Tom, et al.
Published: (2025)
TDD Governance for Multi-Agent Code Generation via Prompt Engineering
by: Hasanli, Tarlan, et al.
Published: (2026)
by: Hasanli, Tarlan, et al.
Published: (2026)
CodePori: Large-Scale System for Autonomous Software Development Using Multi-Agent Technology
by: Rasheed, Zeeshan, et al.
Published: (2024)
by: Rasheed, Zeeshan, et al.
Published: (2024)
Autonomous Legacy Web Application Upgrades Using a Multi-Agent System
by: Ala-Salmi, Valtteri, et al.
Published: (2025)
by: Ala-Salmi, Valtteri, et al.
Published: (2025)
From Specification to Service: Accelerating API-First Development Using Multi-Agent Systems
by: Chauhan, Saurabh, et al.
Published: (2025)
by: Chauhan, Saurabh, et al.
Published: (2025)
Distributed Approach to Haskell Based Applications Refactoring with LLMs Based Multi-Agent Systems
by: Siddeeq, Shahbaz, et al.
Published: (2025)
by: Siddeeq, Shahbaz, et al.
Published: (2025)
CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval
by: Liu, Ye, et al.
Published: (2024)
by: Liu, Ye, et al.
Published: (2024)
From Craft to Constitution: A Governance-First Paradigm for Principled Agent Engineering
by: Xu, Qiang, et al.
Published: (2025)
by: Xu, Qiang, et al.
Published: (2025)
Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents
by: Liu, Ting
Published: (2026)
by: Liu, Ting
Published: (2026)
AI-Assisted Requirements Engineering: An Empirical Evaluation Relative to Expert Judgment
by: Levy, Oz, et al.
Published: (2026)
by: Levy, Oz, et al.
Published: (2026)
Work in Progress: AI-Powered Engineering-Bridging Theory and Practice
by: Levy, Oz, et al.
Published: (2025)
by: Levy, Oz, et al.
Published: (2025)
The Ramanujan Library -- Automated Discovery on the Hypergraph of Integer Relations
by: Beit-Halachmi, Itay, et al.
Published: (2024)
by: Beit-Halachmi, Itay, et al.
Published: (2024)
MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering
by: Guo, Chuanzhe, et al.
Published: (2026)
by: Guo, Chuanzhe, et al.
Published: (2026)
Engineering AI Agents for Clinical Workflows: A Case Study in Architecture,MLOps, and Governance
by: Lopes, Cláudio Lúcio do Val, et al.
Published: (2026)
by: Lopes, Cláudio Lúcio do Val, et al.
Published: (2026)
DataGovBench: Benchmarking LLM Agents for Real-World Data Governance Workflows
by: Liu, Zhou, et al.
Published: (2025)
by: Liu, Zhou, et al.
Published: (2025)
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026)
by: Garralda-Barrio, Mariano
Published: (2026)
Demystifying the DAO Governance Process
by: Ma, Junjie, et al.
Published: (2024)
by: Ma, Junjie, et al.
Published: (2024)
Quality Evaluation of COBOL to Java Code Transformation
by: Froimovich, Shmulik, et al.
Published: (2025)
by: Froimovich, Shmulik, et al.
Published: (2025)
Automated Game Testing With Online Search Agent and Model Construction, a Study
by: Samira Shirzadehhajimahmood, et al.
Published: (2025)
by: Samira Shirzadehhajimahmood, et al.
Published: (2025)
Similar Items
-
Towards Enterprise-Ready Computer Using Generalist Agent
by: Marreed, Sami, et al.
Published: (2025) -
SNAP: Semantic Stories for Next Activity Prediction
by: Oved, Alon, et al.
Published: (2024) -
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
by: Shlomov, Segev, et al.
Published: (2025) -
ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
by: Levy, Ido, et al.
Published: (2024) -
IDA: Breaking Barriers in No-code UI Automation Through Large Language Models and Human-Centric Design
by: Shlomov, Segev, et al.
Published: (2024)