Quantifying Automation Risk in High-Automation AI Systems: A Bayesian Framework for Failure Propagation and Optimal Oversight
Fuente:
arXiv
Saved in:
| Main Authors: | Srivastava, Vishal, Sah, Tanmay |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI and Human Oversight: A Risk-Based Framework for Alignment
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
Deceptive Automated Interpretability: Language Models Coordinating to Fool Oversight Systems
by: Lermen, Simon, et al.
Published: (2025)
by: Lermen, Simon, et al.
Published: (2025)
Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs
by: Agrawal, Tanmay
Published: (2025)
by: Agrawal, Tanmay
Published: (2025)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
by: Gaube, Susanne, et al.
Published: (2026)
by: Gaube, Susanne, et al.
Published: (2026)
Automating the Refinement of Reinforcement Learning Specifications
by: Ambadkar, Tanmay, et al.
Published: (2025)
by: Ambadkar, Tanmay, et al.
Published: (2025)
Fundamental Limits of Black-Box Safety Evaluation: Information-Theoretic and Computational Barriers from Latent Context Conditioning
by: Srivastava, Vishal
Published: (2026)
by: Srivastava, Vishal
Published: (2026)
Between Innovation and Oversight: A Cross-Regional Study of AI Risk Management Frameworks in the EU, U.S., UK, and China
by: Al-Maamari, Amir
Published: (2025)
by: Al-Maamari, Amir
Published: (2025)
Automating AI Failure Tracking: Semantic Association of Reports in AI Incident Database
by: Russo, Diego, et al.
Published: (2025)
by: Russo, Diego, et al.
Published: (2025)
Bridging Temporal and Textual Modalities: A Multimodal Framework for Automated Cloud Failure Root Cause Analysis
by: Park, Gijun
Published: (2026)
by: Park, Gijun
Published: (2026)
On the Prospects of Incorporating Large Language Models (LLMs) in Automated Planning and Scheduling (APS)
by: Pallagani, Vishal, et al.
Published: (2024)
by: Pallagani, Vishal, et al.
Published: (2024)
Automated Auditing of Hospital Discharge Summaries for Care Transitions
by: Dasula, Akshat, et al.
Published: (2026)
by: Dasula, Akshat, et al.
Published: (2026)
RIFT: A RubrIc Failure Mode Taxonomy and Automated Diagnostics
by: Qi, Zhengyang, et al.
Published: (2026)
by: Qi, Zhengyang, et al.
Published: (2026)
Native Explainability for Bayesian Confidence Propagation Neural Networks: A Framework for Trusted Brain-Like AI
by: Makridis, Georgios, et al.
Published: (2026)
by: Makridis, Georgios, et al.
Published: (2026)
Hierarchical Pedagogical Oversight: A Multi-Agent Adversarial Framework for Reliable AI Tutoring
by: Sadhu, Saisab, et al.
Published: (2025)
by: Sadhu, Saisab, et al.
Published: (2025)
Listening Alone, Understanding Together: Collaborative Context Recovery for Privacy-Aware AI
by: Srivastava, Tanmay, et al.
Published: (2026)
by: Srivastava, Tanmay, et al.
Published: (2026)
The Case for Developing a Foundation Model for Planning-like Tasks from Scratch
by: Srivastava, Biplav, et al.
Published: (2024)
by: Srivastava, Biplav, et al.
Published: (2024)
Enhancing Safety Standards in Automated Systems Using Dynamic Bayesian Networks
by: Talluri, Kranthi Kumar, et al.
Published: (2025)
by: Talluri, Kranthi Kumar, et al.
Published: (2025)
Automated Kernel Discovery Towards Understanding High-dimensional Bayesian Optimization
by: Yun, Taeyoung, et al.
Published: (2026)
by: Yun, Taeyoung, et al.
Published: (2026)
SynthAI: A Multi Agent Generative AI Framework for Automated Modular HLS Design Generation
by: Sheikholeslam, Seyed Arash, et al.
Published: (2024)
by: Sheikholeslam, Seyed Arash, et al.
Published: (2024)
The Verifier Tax: Horizon Dependent Safety Success Tradeoffs in Tool Using LLM Agents
by: Sah, Tanmay, et al.
Published: (2026)
by: Sah, Tanmay, et al.
Published: (2026)
An Empirical Study on Failures in Automated Issue Solving
by: Liu, Simiao, et al.
Published: (2025)
by: Liu, Simiao, et al.
Published: (2025)
Transforming NLU with Babylon: A Case Study in Development of Real-time, Edge-Efficient, Multi-Intent Translation System for Automated Drive-Thru Ordering
by: Varzaneh, Mostafa, et al.
Published: (2024)
by: Varzaneh, Mostafa, et al.
Published: (2024)
Human-AI Complementarity: A Goal for Amplified Oversight
by: Jain, Rishub, et al.
Published: (2025)
by: Jain, Rishub, et al.
Published: (2025)
Automated Computation of Therapies Using Failure Mode and Effects Analysis in the Medical Domain
by: Luttermann, Malte, et al.
Published: (2024)
by: Luttermann, Malte, et al.
Published: (2024)
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
by: Ziftci, Celal, et al.
Published: (2026)
by: Ziftci, Celal, et al.
Published: (2026)
Oversight Structures for Agentic AI in Public-Sector Organizations
by: Schmitz, Chris, et al.
Published: (2025)
by: Schmitz, Chris, et al.
Published: (2025)
Abduct, Act, Predict: Scaffolding Causal Inference for Automated Failure Attribution in Multi-Agent Systems
by: West, Alva, et al.
Published: (2025)
by: West, Alva, et al.
Published: (2025)
GamED.AI: A Hierarchical Multi-Agent Framework for Automated Educational Game Generation
by: Agarwal, Shiven, et al.
Published: (2026)
by: Agarwal, Shiven, et al.
Published: (2026)
Automated Design of Agentic Systems
by: Hu, Shengran, et al.
Published: (2024)
by: Hu, Shengran, et al.
Published: (2024)
HonestCyberEval: An AI Cyber Risk Benchmark for Automated Software Exploitation
by: Ristea, Dan, et al.
Published: (2024)
by: Ristea, Dan, et al.
Published: (2024)
Healthcare AI for Automation or Allocation? A Transaction Cost Economics Framework
by: Ercole, Ari
Published: (2026)
by: Ercole, Ari
Published: (2026)
A Benchmark for Scalable Oversight Protocols
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2025)
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2025)
STAF: Leveraging LLMs for Automated Attack Tree-Based Security Test Generation
by: Khule, Tanmay, et al.
Published: (2025)
by: Khule, Tanmay, et al.
Published: (2025)
The Vibe-Automation of Automation: A Proactive Education Framework for Computer Science in the Age of Generative AI
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
AI Agent-Driven Framework for Automated Product Knowledge Graph Construction in E-Commerce
by: Peshevski, Dimitar, et al.
Published: (2025)
by: Peshevski, Dimitar, et al.
Published: (2025)
Calibrating Conservatism for Scalable Oversight
by: Overman, William, et al.
Published: (2026)
by: Overman, William, et al.
Published: (2026)
A Novel Framework for Automated Warehouse Layout Generation
by: Shahroudnejad, Atefeh, et al.
Published: (2024)
by: Shahroudnejad, Atefeh, et al.
Published: (2024)
Explainable Bayesian Optimization
by: Chakraborty, Tanmay, et al.
Published: (2024)
by: Chakraborty, Tanmay, et al.
Published: (2024)
Tiered Agentic Oversight: A Hierarchical Multi-Agent System for Healthcare Safety
by: Kim, Yubin, et al.
Published: (2025)
by: Kim, Yubin, et al.
Published: (2025)
A DeepSeek-Powered AI System for Automated Chest Radiograph Interpretation in Clinical Practice
by: Bai, Yaowei, et al.
Published: (2025)
by: Bai, Yaowei, et al.
Published: (2025)
Similar Items
-
AI and Human Oversight: A Risk-Based Framework for Alignment
by: Kandikatla, Laxmiraju, et al.
Published: (2025) -
Deceptive Automated Interpretability: Language Models Coordinating to Fool Oversight Systems
by: Lermen, Simon, et al.
Published: (2025) -
Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs
by: Agrawal, Tanmay
Published: (2025) -
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
by: Gaube, Susanne, et al.
Published: (2026) -
Automating the Refinement of Reinforcement Learning Specifications
by: Ambadkar, Tanmay, et al.
Published: (2025)