STAMP/STPA Informed Characterization of Factors Leading to Loss of Control in AI Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Barrett, Steve, Bruvere, Anna, Fillingham, Sean P., Rhodes, Catherine, Vergani, Stefano |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Lessons from External Review of DeepMind's Scheming Inability Safety Case
di: Barrett, Stephen, et al.
Pubblicazione: (2026)
di: Barrett, Stephen, et al.
Pubblicazione: (2026)
Systematic Hazard Analysis for Frontier AI using STPA
di: Mylius, Simon
Pubblicazione: (2025)
di: Mylius, Simon
Pubblicazione: (2025)
Introducing a Novel Systems Thinking approach inspired by STPA: Road Safety Intervention design case study
di: Badaoui, Halima El, et al.
Pubblicazione: (2026)
di: Badaoui, Halima El, et al.
Pubblicazione: (2026)
To Build or Not to Build? Factors that Lead to Non-Development or Abandonment of AI Systems
di: Chappidi, Shreya, et al.
Pubblicazione: (2026)
di: Chappidi, Shreya, et al.
Pubblicazione: (2026)
STAMP: Multi-pattern Attention-aware Multiple Instance Learning for STAS Diagnosis in Multi-center Histopathology Images
di: Pan, Liangrui, et al.
Pubblicazione: (2025)
di: Pan, Liangrui, et al.
Pubblicazione: (2025)
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT
di: Qi, Yi, et al.
Pubblicazione: (2023)
di: Qi, Yi, et al.
Pubblicazione: (2023)
A Methodology for Quantitative AI Risk Modeling
di: Murray, Malcolm, et al.
Pubblicazione: (2025)
di: Murray, Malcolm, et al.
Pubblicazione: (2025)
The Role of Risk Modeling in Advanced AI Risk Management
di: Touzet, Chloé, et al.
Pubblicazione: (2025)
di: Touzet, Chloé, et al.
Pubblicazione: (2025)
Who Controls the Conversation? User Perspectives On Generative AI (LLM) System Prompts
di: Neumann, Anna, et al.
Pubblicazione: (2026)
di: Neumann, Anna, et al.
Pubblicazione: (2026)
AI Loss of Control Incident Management: Response & Resilience
di: Gruetzemacher, Ross
Pubblicazione: (2026)
di: Gruetzemacher, Ross
Pubblicazione: (2026)
Climate Implications of Diffusion-based Generative Visual AI Systems and their Mass Adoption
di: Utz, Vanessa, et al.
Pubblicazione: (2025)
di: Utz, Vanessa, et al.
Pubblicazione: (2025)
A Survey of Physics-Informed AI for Complex Urban Systems
di: Xu, En, et al.
Pubblicazione: (2025)
di: Xu, En, et al.
Pubblicazione: (2025)
Efficiency Will Not Lead to Sustainable Reasoning AI
di: Wiesner, Philipp, et al.
Pubblicazione: (2025)
di: Wiesner, Philipp, et al.
Pubblicazione: (2025)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
di: Brundage, Miles, et al.
Pubblicazione: (2026)
di: Brundage, Miles, et al.
Pubblicazione: (2026)
The Narrative Continuity Test: A Conceptual Framework for Evaluating Identity Persistence in AI Systems
di: Natangelo, Stefano
Pubblicazione: (2025)
di: Natangelo, Stefano
Pubblicazione: (2025)
Assessing confidence in frontier AI safety cases
di: Barrett, Stephen, et al.
Pubblicazione: (2025)
di: Barrett, Stephen, et al.
Pubblicazione: (2025)
Funding AI for Good: A Call for Meaningful Engagement
di: Lin, Hongjin, et al.
Pubblicazione: (2025)
di: Lin, Hongjin, et al.
Pubblicazione: (2025)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
di: Magesh, Varun, et al.
Pubblicazione: (2024)
di: Magesh, Varun, et al.
Pubblicazione: (2024)
AI-Powered Legal Intelligence System Architecture: A Comprehensive Framework for Automated Legal Consultation and Analysis
di: Kalaycioglu, Sean, et al.
Pubblicazione: (2025)
di: Kalaycioglu, Sean, et al.
Pubblicazione: (2025)
A Practical Guide for Supporting Formative Assessment and Feedback Using Generative AI
di: Prompiengchai, Sapolnach, et al.
Pubblicazione: (2025)
di: Prompiengchai, Sapolnach, et al.
Pubblicazione: (2025)
Giving AI Personalities Leads to More Human-Like Reasoning
di: Nighojkar, Animesh, et al.
Pubblicazione: (2025)
di: Nighojkar, Animesh, et al.
Pubblicazione: (2025)
Enhancing Equitable Access to AI in Housing and Homelessness System of Care through Federated Learning
di: Taib, Musa, et al.
Pubblicazione: (2024)
di: Taib, Musa, et al.
Pubblicazione: (2024)
The Trust Calibration Maturity Model for Characterizing and Communicating Trustworthiness of AI Systems
di: Steinmetz, Scott T, et al.
Pubblicazione: (2025)
di: Steinmetz, Scott T, et al.
Pubblicazione: (2025)
Concerning the Responsible Use of AI in the US Criminal Justice System
di: Moore, Cristopher, et al.
Pubblicazione: (2025)
di: Moore, Cristopher, et al.
Pubblicazione: (2025)
The Loss of Control Playbook: Degrees, Dynamics, and Preparedness
di: Stix, Charlotte, et al.
Pubblicazione: (2025)
di: Stix, Charlotte, et al.
Pubblicazione: (2025)
PICA: A Data-driven Synthesis of Peer Instruction and Continuous Assessment
di: Geinitz, Steve
Pubblicazione: (2024)
di: Geinitz, Steve
Pubblicazione: (2024)
Detection and Characterization of Coordinated Online Behavior: A Survey
di: Mannocci, Lorenzo, et al.
Pubblicazione: (2024)
di: Mannocci, Lorenzo, et al.
Pubblicazione: (2024)
Toward Quantitative Modeling of Cybersecurity Risks Due to AI Misuse
di: Barrett, Steve, et al.
Pubblicazione: (2025)
di: Barrett, Steve, et al.
Pubblicazione: (2025)
Generative AI as a Geopolitical Factor in Industry 5.0: Sovereignty, Access, and Control
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2025)
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2025)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
di: McGregor, Sean, et al.
Pubblicazione: (2024)
di: McGregor, Sean, et al.
Pubblicazione: (2024)
AI Governance Control Stack for Operational Stability: Achieving Hardened Governance in AI Systems
di: Morgan, Horatio
Pubblicazione: (2026)
di: Morgan, Horatio
Pubblicazione: (2026)
Military AI Needs Technically-Informed Regulation to Safeguard AI Research and its Applications
di: Simmons-Edler, Riley, et al.
Pubblicazione: (2025)
di: Simmons-Edler, Riley, et al.
Pubblicazione: (2025)
AI and Supercomputing are Powering the Next Wave of Breakthrough Science - But at What Cost?
di: Bianchini, Stefano, et al.
Pubblicazione: (2025)
di: Bianchini, Stefano, et al.
Pubblicazione: (2025)
Informing AI Risk Assessment with News Media: Analyzing National and Political Variation in the Coverage of AI Risks
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
The Decision Path to Control AI Risks Completely: Fundamental Control Mechanisms for AI Governance
di: Tao, Yong
Pubblicazione: (2025)
di: Tao, Yong
Pubblicazione: (2025)
Evaluating AI-Generated Images of Cultural Artifacts with Community-Informed Rubrics
di: Johnson, Nari, et al.
Pubblicazione: (2026)
di: Johnson, Nari, et al.
Pubblicazione: (2026)
Characterizing AI Fact-Checkers and Their Contributions on Community Notes
di: Gong, Yilin, et al.
Pubblicazione: (2026)
di: Gong, Yilin, et al.
Pubblicazione: (2026)
Agentic AI and the Cyber Arms Race
di: Oesch, Sean, et al.
Pubblicazione: (2025)
di: Oesch, Sean, et al.
Pubblicazione: (2025)
A.I. In All The Wrong Places
di: Böhlen, Marc, et al.
Pubblicazione: (2024)
di: Böhlen, Marc, et al.
Pubblicazione: (2024)
Critical Transit Infrastructure in Smart Cities and Urban Air Quality: A Multi-City Seasonal Comparison of Ridership and PM2.5
di: Elliott, Sean, et al.
Pubblicazione: (2026)
di: Elliott, Sean, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Lessons from External Review of DeepMind's Scheming Inability Safety Case
di: Barrett, Stephen, et al.
Pubblicazione: (2026) -
Systematic Hazard Analysis for Frontier AI using STPA
di: Mylius, Simon
Pubblicazione: (2025) -
Introducing a Novel Systems Thinking approach inspired by STPA: Road Safety Intervention design case study
di: Badaoui, Halima El, et al.
Pubblicazione: (2026) -
To Build or Not to Build? Factors that Lead to Non-Development or Abandonment of AI Systems
di: Chappidi, Shreya, et al.
Pubblicazione: (2026) -
STAMP: Multi-pattern Attention-aware Multiple Instance Learning for STAS Diagnosis in Multi-center Histopathology Images
di: Pan, Liangrui, et al.
Pubblicazione: (2025)