AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Badagi, Chitra, Singh, Divye, Sen, Animesh, Shirsath, Adinath |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
by: Naqvi, Saba, et al.
Published: (2026)
by: Naqvi, Saba, et al.
Published: (2026)
Automated Population-Level Audit Assurance via AI-Based Document Intelligence
by: Vasudevan, Santosh, et al.
Published: (2026)
by: Vasudevan, Santosh, et al.
Published: (2026)
AI-Driven Tools in Modern Software Quality Assurance: An Assessment of Benefits, Challenges, and Future Directions
by: Pysmennyi, Ihor, et al.
Published: (2025)
by: Pysmennyi, Ihor, et al.
Published: (2025)
Generative AI for Testing of Autonomous Driving Systems: A Survey
by: Song, Qunying, et al.
Published: (2025)
by: Song, Qunying, et al.
Published: (2025)
Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents
by: Liu, Ting
Published: (2026)
by: Liu, Ting
Published: (2026)
A Framework for the Adoption and Integration of Generative AI in Midsize Organizations and Enterprises (FAIGMOE)
by: Weinberg, Abraham Itzhak
Published: (2025)
by: Weinberg, Abraham Itzhak
Published: (2025)
REGAL: A Registry-Driven Architecture for Deterministic Grounding of Agentic AI in Enterprise Telemetry
by: Agrawal, Yuvraj
Published: (2026)
by: Agrawal, Yuvraj
Published: (2026)
Ethics Testing: Proactive Identification of Generative AI System Harms
by: Tan, Shin Hwei, et al.
Published: (2026)
by: Tan, Shin Hwei, et al.
Published: (2026)
Impact and Implications of Generative AI for Enterprise Architects in Agile Environments: A Systematic Literature Review
by: Kooy, Stefan Julian, et al.
Published: (2025)
by: Kooy, Stefan Julian, et al.
Published: (2025)
Using AI/ML to Find and Remediate Enterprise Secrets in Code & Document Sharing Platforms
by: Kerr, Gregor, et al.
Published: (2024)
by: Kerr, Gregor, et al.
Published: (2024)
Lessons Learned from the Use of Generative AI in Engineering and Quality Assurance of a WEB System for Healthcare
by: Travassos, Guilherme H., et al.
Published: (2025)
by: Travassos, Guilherme H., et al.
Published: (2025)
Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution
by: Pan, Kai, et al.
Published: (2026)
by: Pan, Kai, et al.
Published: (2026)
Disrupting Test Development with AI Assistants
by: Joshi, Vijay, et al.
Published: (2024)
by: Joshi, Vijay, et al.
Published: (2024)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
by: Baqar, Mohammad, et al.
Published: (2024)
by: Baqar, Mohammad, et al.
Published: (2024)
Generative AI to Generate Test Data Generators
by: Baudry, Benoit, et al.
Published: (2024)
by: Baudry, Benoit, et al.
Published: (2024)
A Defect Classification Framework for AI-Based Software Systems (AI-ODC)
by: Alannsary, Mohammed O.
Published: (2025)
by: Alannsary, Mohammed O.
Published: (2025)
AI-Assisted Unit Test Writing and Test-Driven Code Refactoring: A Case Study
by: Smolic, Ema, et al.
Published: (2026)
by: Smolic, Ema, et al.
Published: (2026)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
by: Gao, Xuanqi, et al.
Published: (2025)
by: Gao, Xuanqi, et al.
Published: (2025)
Where AI Assurance Might Go Wrong: Initial lessons from engineering of critical systems
by: Bloomfield, Robin, et al.
Published: (2025)
by: Bloomfield, Robin, et al.
Published: (2025)
Engineering AI Judge Systems
by: Lin, Jiahuei, et al.
Published: (2024)
by: Lin, Jiahuei, et al.
Published: (2024)
Leveraging LLMs for User Stories in AI Systems: UStAI Dataset
by: Yamani, Asma, et al.
Published: (2025)
by: Yamani, Asma, et al.
Published: (2025)
Leveraging AI for Enhanced Software Effort Estimation: A Comprehensive Study and Framework Proposal
by: Tran, Nhi, et al.
Published: (2024)
by: Tran, Nhi, et al.
Published: (2024)
PBT-Bench: Benchmarking AI Agents on Property-Based Testing
by: Jing, Lucas, et al.
Published: (2026)
by: Jing, Lucas, et al.
Published: (2026)
Breaking Barriers in Software Testing: The Power of AI-Driven Automation
by: Naqvi, Saba, et al.
Published: (2025)
by: Naqvi, Saba, et al.
Published: (2025)
Causal Reasoning in Software Quality Assurance: A Systematic Review
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
World of Workflows: A Benchmark for Bringing World Models to Enterprise Systems
by: Gupta, Lakshya, et al.
Published: (2026)
by: Gupta, Lakshya, et al.
Published: (2026)
Expectations vs Reality -- A Secondary Study on AI Adoption in Software Testing
by: Karhu, Katja, et al.
Published: (2025)
by: Karhu, Katja, et al.
Published: (2025)
A Multi-agent AI System for Deep Learning Model Migration from TensorFlow to JAX
by: Nikolov, Stoyan, et al.
Published: (2026)
by: Nikolov, Stoyan, et al.
Published: (2026)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
by: Cheng, Audrey, et al.
Published: (2025)
by: Cheng, Audrey, et al.
Published: (2025)
Greening AI-enabled Systems with Software Engineering: A Research Agenda for Environmentally Sustainable AI Practices
by: Cruz, Luís, et al.
Published: (2025)
by: Cruz, Luís, et al.
Published: (2025)
GBQA: A Game Benchmark for Evaluating LLMs as Quality Assurance Engineers
by: Jiang, Shufan, et al.
Published: (2026)
by: Jiang, Shufan, et al.
Published: (2026)
Unit Test Generation using Generative AI : A Comparative Performance Analysis of Autogeneration Tools
by: Bhatia, Shreya, et al.
Published: (2023)
by: Bhatia, Shreya, et al.
Published: (2023)
CoDefeater: Using LLMs To Find Defeaters in Assurance Cases
by: Gohar, Usman, et al.
Published: (2024)
by: Gohar, Usman, et al.
Published: (2024)
On the Impact of Black-box Deployment Strategies for Edge AI on Latency and Model Performance
by: Singh, Jaskirat, et al.
Published: (2024)
by: Singh, Jaskirat, et al.
Published: (2024)
FinRobot: Generative Business Process AI Agents for Enterprise Resource Planning in Finance
by: Yang, Hongyang, et al.
Published: (2025)
by: Yang, Hongyang, et al.
Published: (2025)
Skilled AI Agents for Embedded and IoT Systems Development
by: Li, Yiming, et al.
Published: (2026)
by: Li, Yiming, et al.
Published: (2026)
Quantifying the Expectation-Realisation Gap for Agentic AI Systems
by: Lobentanzer, Sebastian
Published: (2026)
by: Lobentanzer, Sebastian
Published: (2026)
WALL: A Web Application for Automated Quality Assurance using Large Language Models
by: Abtahi, Seyed Moein, et al.
Published: (2025)
by: Abtahi, Seyed Moein, et al.
Published: (2025)
On Unified Prompt Tuning for Request Quality Assurance in Public Code Review
by: Chen, Xinyu, et al.
Published: (2024)
by: Chen, Xinyu, et al.
Published: (2024)
MultiAIGCD: A Comprehensive dataset for AI Generated Code Detection Covering Multiple Languages, Models,Prompts, and Scenarios
by: Demirok, Basak, et al.
Published: (2025)
by: Demirok, Basak, et al.
Published: (2025)
Similar Items
-
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
by: Naqvi, Saba, et al.
Published: (2026) -
Automated Population-Level Audit Assurance via AI-Based Document Intelligence
by: Vasudevan, Santosh, et al.
Published: (2026) -
AI-Driven Tools in Modern Software Quality Assurance: An Assessment of Benefits, Challenges, and Future Directions
by: Pysmennyi, Ihor, et al.
Published: (2025) -
Generative AI for Testing of Autonomous Driving Systems: A Survey
by: Song, Qunying, et al.
Published: (2025) -
Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents
by: Liu, Ting
Published: (2026)