Quantifying the Expectation-Realisation Gap for Agentic AI Systems
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Lobentanzer, Sebastian |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Agentic Automation Canvas: a structured framework for agentic AI project design
par: Lobentanzer, Sebastian
Publié: (2026)
par: Lobentanzer, Sebastian
Publié: (2026)
From Goals to Aspects, Revisited: An NFR Pattern Language for Agentic AI Systems
par: Yu, Yijun
Publié: (2026)
par: Yu, Yijun
Publié: (2026)
Describing Agentic AI Systems with C4: Lessons from Industry Projects
par: Rausch, Andreas, et autres
Publié: (2026)
par: Rausch, Andreas, et autres
Publié: (2026)
Expectations vs Reality -- A Secondary Study on AI Adoption in Software Testing
par: Karhu, Katja, et autres
Publié: (2025)
par: Karhu, Katja, et autres
Publié: (2025)
Agentic AI Software Engineers: Programming with Trust
par: Roychoudhury, Abhik, et autres
Publié: (2025)
par: Roychoudhury, Abhik, et autres
Publié: (2025)
Agentic Business Process Management Systems
par: Dumas, Marlon, et autres
Publié: (2026)
par: Dumas, Marlon, et autres
Publié: (2026)
Project Prometheus: Bridging the Intent Gap in Agentic Program Repair via Reverse-Engineered Executable Specifications
par: Wang, Yongchao, et autres
Publié: (2026)
par: Wang, Yongchao, et autres
Publié: (2026)
Pragmos: A Process Agentic Modeling System
par: Hernández-Ávalos, Pedro-Aarón, et autres
Publié: (2026)
par: Hernández-Ávalos, Pedro-Aarón, et autres
Publié: (2026)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
par: Li, Jingyue, et autres
Publié: (2026)
par: Li, Jingyue, et autres
Publié: (2026)
Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
par: Salim, Mohamad, et autres
Publié: (2026)
par: Salim, Mohamad, et autres
Publié: (2026)
TriCEGAR: A Trace-Driven Abstraction Mechanism for Agentic AI
par: Koohestani, Roham, et autres
Publié: (2026)
par: Koohestani, Roham, et autres
Publié: (2026)
Code for Machines, Not Just Humans: Quantifying AI-Friendliness with Code Health Metrics
par: Borg, Markus, et autres
Publié: (2026)
par: Borg, Markus, et autres
Publié: (2026)
Runtime-Structured Task Decomposition for Agentic Coding Systems
par: Asthana, Shubhi, et autres
Publié: (2026)
par: Asthana, Shubhi, et autres
Publié: (2026)
From Code-Centric to Intent-Centric Software Engineering: A Reflexive Thematic Analysis of Generative AI, Agentic Systems, and Engineering Accountability
par: De La Cruz, Elyson
Publié: (2026)
par: De La Cruz, Elyson
Publié: (2026)
Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review
par: Kamalı, Hüseyin Özgür, et autres
Publié: (2026)
par: Kamalı, Hüseyin Özgür, et autres
Publié: (2026)
Agentic AI in 6G Software Businesses: A Layered Maturity Model
par: Zohaib, Muhammad, et autres
Publié: (2025)
par: Zohaib, Muhammad, et autres
Publié: (2025)
LLM-Based Agentic Systems for Software Engineering: Challenges and Opportunities
par: Tang, Yongjian, et autres
Publié: (2026)
par: Tang, Yongjian, et autres
Publié: (2026)
PEFA-AI: Advancing Open-source LLMs for RTL generation using Progressive Error Feedback Agentic-AI
par: Narayanan, Athma, et autres
Publié: (2025)
par: Narayanan, Athma, et autres
Publié: (2025)
REGAL: A Registry-Driven Architecture for Deterministic Grounding of Agentic AI in Enterprise Telemetry
par: Agrawal, Yuvraj
Publié: (2026)
par: Agrawal, Yuvraj
Publié: (2026)
Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems
par: Ouyang, Yipeng, et autres
Publié: (2026)
par: Ouyang, Yipeng, et autres
Publié: (2026)
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
par: Akshathala, Sreemaee, et autres
Publié: (2025)
par: Akshathala, Sreemaee, et autres
Publié: (2025)
Precision in Practice: Knowledge Guided Code Summarizing Grounded in Industrial Expectations
par: Li, Jintai, et autres
Publié: (2026)
par: Li, Jintai, et autres
Publié: (2026)
ReusStdFlow: A Standardized Reusability Framework for Dynamic Workflow Construction in Agentic AI
par: Zhang, Gaoyang, et autres
Publié: (2026)
par: Zhang, Gaoyang, et autres
Publié: (2026)
A Dual-Helix Governance Approach Towards Reliable Agentic AI for WebGIS Development
par: Boyuan, et autres
Publié: (2026)
par: Boyuan, et autres
Publié: (2026)
Digital Twin and Agentic AI for Wild Fire Disaster Management: Intelligent Virtual Situation Room
par: Morsali, Mohammad, et autres
Publié: (2026)
par: Morsali, Mohammad, et autres
Publié: (2026)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
par: Casserini, Matteo, et autres
Publié: (2026)
par: Casserini, Matteo, et autres
Publié: (2026)
Quantifying Uncertainty in Machine Learning-Based Pervasive Systems: Application to Human Activity Recognition
par: Balditsyn, Vladimir, et autres
Publié: (2025)
par: Balditsyn, Vladimir, et autres
Publié: (2025)
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
par: Naqvi, Saba, et autres
Publié: (2026)
par: Naqvi, Saba, et autres
Publié: (2026)
From Expectation to Habit: Why Do Software Practitioners Adopt Fairness Toolkits?
par: Voria, Gianmario, et autres
Publié: (2024)
par: Voria, Gianmario, et autres
Publié: (2024)
Toward a Science of Intent: Closure Gaps and Delegation Envelopes for Open-World AI Agents
par: Armesto, Maximiliano, et autres
Publié: (2026)
par: Armesto, Maximiliano, et autres
Publié: (2026)
ToolMisuseBench: An Offline Deterministic Benchmark for Tool Misuse and Recovery in Agentic Systems
par: Sigdel, Akshey, et autres
Publié: (2026)
par: Sigdel, Akshey, et autres
Publié: (2026)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
par: Bhat, Vishvesh, et autres
Publié: (2025)
par: Bhat, Vishvesh, et autres
Publié: (2025)
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
par: Ehsani, Ramtin, et autres
Publié: (2026)
par: Ehsani, Ramtin, et autres
Publié: (2026)
Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI
par: Sapkota, Ranjan, et autres
Publié: (2025)
par: Sapkota, Ranjan, et autres
Publié: (2025)
Bridging the Gap between Real-world and Synthetic Images for Testing Autonomous Driving Systems
par: Amini, Mohammad Hossein, et autres
Publié: (2024)
par: Amini, Mohammad Hossein, et autres
Publié: (2024)
Engineering AI Judge Systems
par: Lin, Jiahuei, et autres
Publié: (2024)
par: Lin, Jiahuei, et autres
Publié: (2024)
Agentic Harness for Real-World Compilers
par: Zheng, Yingwei, et autres
Publié: (2026)
par: Zheng, Yingwei, et autres
Publié: (2026)
DeepCode: Open Agentic Coding
par: Li, Zongwei, et autres
Publié: (2025)
par: Li, Zongwei, et autres
Publié: (2025)
ROSBag MCP Server: Analyzing Robot Data with LLMs for Agentic Embodied AI Applications
par: Fu, Lei, et autres
Publié: (2025)
par: Fu, Lei, et autres
Publié: (2025)
AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems
par: Badagi, Chitra, et autres
Publié: (2026)
par: Badagi, Chitra, et autres
Publié: (2026)
Documents similaires
-
The Agentic Automation Canvas: a structured framework for agentic AI project design
par: Lobentanzer, Sebastian
Publié: (2026) -
From Goals to Aspects, Revisited: An NFR Pattern Language for Agentic AI Systems
par: Yu, Yijun
Publié: (2026) -
Describing Agentic AI Systems with C4: Lessons from Industry Projects
par: Rausch, Andreas, et autres
Publié: (2026) -
Expectations vs Reality -- A Secondary Study on AI Adoption in Software Testing
par: Karhu, Katja, et autres
Publié: (2025) -
Agentic AI Software Engineers: Programming with Trust
par: Roychoudhury, Abhik, et autres
Publié: (2025)