AI Safety as Control of Irreversibility: A Systems Framework for Decision-Energy and Sovereignty Boundaries
Fuente:
arXiv
Salvato in:
| Autori principali: | Shu, Wesley, Wei, Peng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control
di: Wei, Peng, et al.
Pubblicazione: (2026)
di: Wei, Peng, et al.
Pubblicazione: (2026)
A Public Theory of Distillation Resistance via Constraint-Coupled Reasoning Architectures
di: Wei, Peng, et al.
Pubblicazione: (2026)
di: Wei, Peng, et al.
Pubblicazione: (2026)
Generative AI as a Geopolitical Factor in Industry 5.0: Sovereignty, Access, and Control
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2025)
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2025)
A Conceptual Framework for AI-based Decision Systems in Critical Infrastructures
di: Leyli-abadi, Milad, et al.
Pubblicazione: (2025)
di: Leyli-abadi, Milad, et al.
Pubblicazione: (2025)
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
di: Dobbe, Roel
Pubblicazione: (2025)
di: Dobbe, Roel
Pubblicazione: (2025)
Defining AI Models and AI Systems: A Framework to Resolve the Boundary Problem
di: Sun, Yuanyuan, et al.
Pubblicazione: (2026)
di: Sun, Yuanyuan, et al.
Pubblicazione: (2026)
The Missing Red Line: How Commercial Pressure Erodes AI Safety Boundaries
di: Petrova, Nora, et al.
Pubblicazione: (2026)
di: Petrova, Nora, et al.
Pubblicazione: (2026)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
di: Lin, Zhiqiang, et al.
Pubblicazione: (2025)
di: Lin, Zhiqiang, et al.
Pubblicazione: (2025)
Institutional AI Sovereignty Through Gateway Architecture: Implementation Report from Fontys ICT
di: Huijts, Ruud, et al.
Pubblicazione: (2025)
di: Huijts, Ruud, et al.
Pubblicazione: (2025)
Safety Cases: How to Justify the Safety of Advanced AI Systems
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
Emerging Practices in Frontier AI Safety Frameworks
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
Human-AI Safety: A Descendant of Generative AI and Control Systems Safety
di: Bajcsy, Andrea, et al.
Pubblicazione: (2024)
di: Bajcsy, Andrea, et al.
Pubblicazione: (2024)
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
di: Staufer, Leon, et al.
Pubblicazione: (2026)
di: Staufer, Leon, et al.
Pubblicazione: (2026)
Catalyzing Informed Residential Energy Retrofit Decisions via Domain-Specific LLM
di: Shu, Lei, et al.
Pubblicazione: (2026)
di: Shu, Lei, et al.
Pubblicazione: (2026)
Diagnosing Hallucination Risk in AI Surgical Decision-Support: A Sequential Framework for Sequential Validation
di: Chen, Dong, et al.
Pubblicazione: (2025)
di: Chen, Dong, et al.
Pubblicazione: (2025)
Safety Cases: A Scalable Approach to Frontier AI Safety
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
The Controllability Trap: A Governance Framework for Military AI Agents
di: Sahoo, Subramanyam
Pubblicazione: (2026)
di: Sahoo, Subramanyam
Pubblicazione: (2026)
Case-based Reasoning Augmented Large Language Model Framework for Decision Making in Realistic Safety-Critical Driving Scenarios
di: Gan, Wenbin, et al.
Pubblicazione: (2025)
di: Gan, Wenbin, et al.
Pubblicazione: (2025)
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems
di: Ide, Shun, et al.
Pubblicazione: (2024)
di: Ide, Shun, et al.
Pubblicazione: (2024)
International Agreements on AI Safety: Review and Recommendations for a Conditional AI Safety Treaty
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
Agentic Microphysics: A Manifesto for Generative AI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
Toward an African Agenda for AI Safety
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
Concrete Problems in AI Safety, Revisited
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
An Online Hierarchical Energy Management System for Energy Communities, Complying with the Current Technical Legislation Framework
di: Capillo, Antonino, et al.
Pubblicazione: (2024)
di: Capillo, Antonino, et al.
Pubblicazione: (2024)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
di: Xia, Boming, et al.
Pubblicazione: (2024)
di: Xia, Boming, et al.
Pubblicazione: (2024)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
di: Vaccaro, Michelle, et al.
Pubblicazione: (2026)
di: Vaccaro, Michelle, et al.
Pubblicazione: (2026)
Evaluation Framework for AI Systems in "the Wild"
di: Jabbour, Sarah, et al.
Pubblicazione: (2025)
di: Jabbour, Sarah, et al.
Pubblicazione: (2025)
Assistive AI for Augmenting Human Decision-making
di: Gyöngyössy, Natabara Máté, et al.
Pubblicazione: (2024)
di: Gyöngyössy, Natabara Máté, et al.
Pubblicazione: (2024)
AI Governance Control Stack for Operational Stability: Achieving Hardened Governance in AI Systems
di: Morgan, Horatio
Pubblicazione: (2026)
di: Morgan, Horatio
Pubblicazione: (2026)
AI Safety: Necessary, but insufficient and possibly problematic
di: P, Deepak
Pubblicazione: (2024)
di: P, Deepak
Pubblicazione: (2024)
Agentic LLM Framework for Adaptive Decision Discourse
di: Dolant, Antoine, et al.
Pubblicazione: (2025)
di: Dolant, Antoine, et al.
Pubblicazione: (2025)
The Singapore Consensus on Global AI Safety Research Priorities
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
Annotating the Chain-of-Thought: A Behavior-Labeled Dataset for AI Safety
di: Menke, Antonio-Gabriel Chacón, et al.
Pubblicazione: (2025)
di: Menke, Antonio-Gabriel Chacón, et al.
Pubblicazione: (2025)
The Ghost in the Grammar: Methodological Anthropomorphism in AI Safety Evaluations
di: Costa, Mariana Lins
Pubblicazione: (2026)
di: Costa, Mariana Lins
Pubblicazione: (2026)
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
di: Li, Miles Q., et al.
Pubblicazione: (2026)
di: Li, Miles Q., et al.
Pubblicazione: (2026)
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
Upstream and Downstream AI Safety: Both on the Same River?
di: McDermid, John, et al.
Pubblicazione: (2024)
di: McDermid, John, et al.
Pubblicazione: (2024)
Combining Cost-Constrained Runtime Monitors for AI Safety
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
Probabilistic Analysis of Copyright Disputes and Generative AI Safety
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
Building Effective Safety Guardrails in AI Education Tools
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control
di: Wei, Peng, et al.
Pubblicazione: (2026) -
A Public Theory of Distillation Resistance via Constraint-Coupled Reasoning Architectures
di: Wei, Peng, et al.
Pubblicazione: (2026) -
Generative AI as a Geopolitical Factor in Industry 5.0: Sovereignty, Access, and Control
di: Wasi, Azmine Toushik, et al.
Pubblicazione: (2025) -
A Conceptual Framework for AI-based Decision Systems in Critical Infrastructures
di: Leyli-abadi, Milad, et al.
Pubblicazione: (2025) -
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
di: Dobbe, Roel
Pubblicazione: (2025)