Beyond Task Success: An Evidence-Synthesis Framework for Evaluating, Governing, and Orchestrating Agentic AI
Fuente:
arXiv
Saved in:
| Main Authors: | Koch, Christopher, Wellbrock, Joshua Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agile V: A Compliance-Ready Framework for AI-Augmented Engineering -- From Concept to Audit-Ready Delivery
by: Koch, Christopher, et al.
Published: (2026)
by: Koch, Christopher, et al.
Published: (2026)
Runtime Advocates: A Persona-Driven Framework for Requirements@Runtime Decision Support
by: Hernandez, Demetrius, et al.
Published: (2025)
by: Hernandez, Demetrius, et al.
Published: (2025)
Agent-Oriented Visual Programming for the Web of Things
by: Burattini, Samuele, et al.
Published: (2025)
by: Burattini, Samuele, et al.
Published: (2025)
SAGE: Tool-Augmented LLM Task Solving Strategies in Scalable Multi-Agent Environments
by: Strehlow, Robert K., et al.
Published: (2026)
by: Strehlow, Robert K., et al.
Published: (2026)
From Governance Norms to Enforceable Controls: A Layered Translation Method for Runtime Guardrails in Agentic AI
by: Koch, Christopher
Published: (2026)
by: Koch, Christopher
Published: (2026)
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
by: Akshathala, Sreemaee, et al.
Published: (2025)
by: Akshathala, Sreemaee, et al.
Published: (2025)
AIAP: A No-Code Workflow Builder for Non-Experts with Natural Language and Multi-Agent Collaboration
by: An, Hyunjn, et al.
Published: (2025)
by: An, Hyunjn, et al.
Published: (2025)
Feynman: Knowledge-Infused Diagramming Agent for Scalable Visual Designs
by: Wen, Zixin, et al.
Published: (2026)
by: Wen, Zixin, et al.
Published: (2026)
Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development
by: Koch, Christopher
Published: (2026)
by: Koch, Christopher
Published: (2026)
Agentic Lybic: Multi-Agent Execution System with Tiered Reasoning and Orchestration
by: Guo, Liangxuan, et al.
Published: (2025)
by: Guo, Liangxuan, et al.
Published: (2025)
Beyond Banning AI: A First Look at GenAI Governance in Open Source Software Communities
by: Yang, Wenhao, et al.
Published: (2026)
by: Yang, Wenhao, et al.
Published: (2026)
Murakkab: Resource-Efficient Agentic Workflow Orchestration in Cloud Platforms
by: Chaudhry, Gohar Irfan, et al.
Published: (2025)
by: Chaudhry, Gohar Irfan, et al.
Published: (2025)
Multi-Stakeholder Alignment in LLM-Powered Collaborative AI Systems: A Multi-Agent Framework for Intelligent Tutoring
by: Uchoa, Alexandre P, et al.
Published: (2025)
by: Uchoa, Alexandre P, et al.
Published: (2025)
A Visualization Framework for Exploring Multi-Agent-Based Simulations Case Study of an Electric Vehicle Home Charging Ecosystem
by: Christensen, Kristoffer, et al.
Published: (2025)
by: Christensen, Kristoffer, et al.
Published: (2025)
V-SHiNE: A Virtual Smart Home Framework for Explainability Evaluation
by: Sadeghi, Mersedeh, et al.
Published: (2026)
by: Sadeghi, Mersedeh, et al.
Published: (2026)
Toward Adaptive Categories: Dimensional Governance for Agentic AI
by: Engin, Zeynep, et al.
Published: (2025)
by: Engin, Zeynep, et al.
Published: (2025)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
by: Lwin, Naing Oo, et al.
Published: (2026)
by: Lwin, Naing Oo, et al.
Published: (2026)
Human-Artificial Interaction in the Age of Agentic AI: A System-Theoretical Approach
by: Borghoff, Uwe M., et al.
Published: (2025)
by: Borghoff, Uwe M., et al.
Published: (2025)
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026)
by: Zhang, Shuyu, et al.
Published: (2026)
AutoClimDS: Climate Data Science Agentic AI -- A Knowledge Graph is All You Need
by: Jaber, Ahmed, et al.
Published: (2025)
by: Jaber, Ahmed, et al.
Published: (2025)
AgenticAD: A Specialized Multiagent System Framework for Holistic Alzheimer Disease Management
by: Bazgir, Adib, et al.
Published: (2025)
by: Bazgir, Adib, et al.
Published: (2025)
EmBARDiment: an Embodied AI Agent for Productivity in XR
by: Bovo, Riccardo, et al.
Published: (2024)
by: Bovo, Riccardo, et al.
Published: (2024)
Agentic AI: The Era of Semantic Decoding
by: Peyrard, Maxime, et al.
Published: (2024)
by: Peyrard, Maxime, et al.
Published: (2024)
Beyond Dark Patterns: A Concept-Based Framework for Ethical Software Design
by: Caragay, Evan, et al.
Published: (2023)
by: Caragay, Evan, et al.
Published: (2023)
Crowdsourcing: A Framework for Usability Evaluation
by: Nasir, Muhammad
Published: (2024)
by: Nasir, Muhammad
Published: (2024)
A2H: Agent-to-Human Protocol for AI Agent
by: Liang, Zhiyuan, et al.
Published: (2025)
by: Liang, Zhiyuan, et al.
Published: (2025)
On the Utility of External Agent Intention Predictor for Human-AI Coordination
by: Wang, Chenxu, et al.
Published: (2024)
by: Wang, Chenxu, et al.
Published: (2024)
Evaluating Citizen Satisfaction with Saudi Arabia's E-Government Services: A Standards-Based, Theory-Informed Approach
by: Alannsary, Mohammed O.
Published: (2025)
by: Alannsary, Mohammed O.
Published: (2025)
FlowEval: Reference-based Evaluation of Generated User Interfaces
by: Wu, Jason, et al.
Published: (2026)
by: Wu, Jason, et al.
Published: (2026)
A Unified, Cross-Platform Framework for Automatic GUI and Plugin Generation in Structural Bioinformatics and Beyond
by: Guo, Sikao, et al.
Published: (2026)
by: Guo, Sikao, et al.
Published: (2026)
Robust, Observable, and Evolvable Agentic Systems Engineering: A Principled Framework Validated via the Fairy GUI Agent
by: Sun, Jiazheng, et al.
Published: (2025)
by: Sun, Jiazheng, et al.
Published: (2025)
Verified Synthesis of Optimal Safety Controllers for Human-Robot Collaboration
by: Gleirscher, Mario, et al.
Published: (2021)
by: Gleirscher, Mario, et al.
Published: (2021)
Sherlock: Reliable and Efficient Agentic Workflow Execution
by: Ro, Yeonju, et al.
Published: (2025)
by: Ro, Yeonju, et al.
Published: (2025)
CandorMD: An AI-Assisted Audio Simulation and Feedback System for Training Clinicians for Medical Error Disclosure
by: Lin, Inna Wanyin, et al.
Published: (2026)
by: Lin, Inna Wanyin, et al.
Published: (2026)
Human-in-the-Loop Testing of AI Agents for Air Traffic Control with a Regulated Assessment Framework
by: Carvell, Ben, et al.
Published: (2026)
by: Carvell, Ben, et al.
Published: (2026)
Decoupled Intelligence: A Multi-Agent LLM Framework for Controllable Traffic Scenario Generation in SUMO
by: Li, Shuyang, et al.
Published: (2026)
by: Li, Shuyang, et al.
Published: (2026)
Qualixar OS: A Universal Operating System for AI Agent Orchestration
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care
by: Cho, Andrew, et al.
Published: (2025)
by: Cho, Andrew, et al.
Published: (2025)
SmartEx: A Framework for Generating User-Centric Explanations in Smart Environments
by: Sadeghi, Mersedeh, et al.
Published: (2024)
by: Sadeghi, Mersedeh, et al.
Published: (2024)
Teaming in the AI Era: AI-Augmented Frameworks for Forming, Simulating, and Optimizing Human Teams
by: Almutairi, Mohammed
Published: (2025)
by: Almutairi, Mohammed
Published: (2025)
Similar Items
-
Agile V: A Compliance-Ready Framework for AI-Augmented Engineering -- From Concept to Audit-Ready Delivery
by: Koch, Christopher, et al.
Published: (2026) -
Runtime Advocates: A Persona-Driven Framework for Requirements@Runtime Decision Support
by: Hernandez, Demetrius, et al.
Published: (2025) -
Agent-Oriented Visual Programming for the Web of Things
by: Burattini, Samuele, et al.
Published: (2025) -
SAGE: Tool-Augmented LLM Task Solving Strategies in Scalable Multi-Agent Environments
by: Strehlow, Robert K., et al.
Published: (2026) -
From Governance Norms to Enforceable Controls: A Layered Translation Method for Runtime Guardrails in Agentic AI
by: Koch, Christopher
Published: (2026)