AI Safety Frameworks Should Include Procedures for Model Access Decisions
Fuente:
arXiv
Salvato in:
| Autori principali: | Kembery, Edward, Reed, Tom |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Responsible Governing AI Proliferation
di: Kembery, Edward
Pubblicazione: (2024)
di: Kembery, Edward
Pubblicazione: (2024)
Position Paper: Model Access should be a Key Concern in AI Governance
di: Kembery, Edward, et al.
Pubblicazione: (2024)
di: Kembery, Edward, et al.
Pubblicazione: (2024)
How Should AI Safety Benchmarks Benchmark Safety?
di: Yu, Cheng, et al.
Pubblicazione: (2026)
di: Yu, Cheng, et al.
Pubblicazione: (2026)
AI Safety Should Prioritize the Future of Work
di: Hazra, Sanchaita, et al.
Pubblicazione: (2025)
di: Hazra, Sanchaita, et al.
Pubblicazione: (2025)
AI Safety as Control of Irreversibility: A Systems Framework for Decision-Energy and Sovereignty Boundaries
di: Shu, Wesley, et al.
Pubblicazione: (2026)
di: Shu, Wesley, et al.
Pubblicazione: (2026)
AI Companies Should Report Pre- and Post-Mitigation Safety Evaluations
di: Bowen, Dillon, et al.
Pubblicazione: (2025)
di: Bowen, Dillon, et al.
Pubblicazione: (2025)
A Grading Rubric for AI Safety Frameworks
di: Alaga, Jide, et al.
Pubblicazione: (2024)
di: Alaga, Jide, et al.
Pubblicazione: (2024)
Evaluating AI Providers' Frontier Safety Frameworks
di: Stelling, Lily, et al.
Pubblicazione: (2025)
di: Stelling, Lily, et al.
Pubblicazione: (2025)
Verification methods for international AI agreements
di: Wasil, Akash R., et al.
Pubblicazione: (2024)
di: Wasil, Akash R., et al.
Pubblicazione: (2024)
Who Should Run Advanced AI Evaluations -- AISIs?
di: Stein, Merlin, et al.
Pubblicazione: (2024)
di: Stein, Merlin, et al.
Pubblicazione: (2024)
Why AI Is WEIRD and Should Not Be This Way: Towards AI For Everyone, With Everyone, By Everyone
di: Mihalcea, Rada, et al.
Pubblicazione: (2024)
di: Mihalcea, Rada, et al.
Pubblicazione: (2024)
What do model reports say about their ChemBio benchmark evaluations? Comparing recent releases to the STREAM framework
di: Reed, Tom, et al.
Pubblicazione: (2025)
di: Reed, Tom, et al.
Pubblicazione: (2025)
Students Know AI Should Not Replace Thinking, but How Do They Regulate It? The TACO Framework for Human-AI Cognitive Partnership
di: Chan, Cecilia Ka Yuk
Pubblicazione: (2026)
di: Chan, Cecilia Ka Yuk
Pubblicazione: (2026)
Institutional AI: A Governance Framework for Distributional AGI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
When Should AI Read the Room? Public Perceptions of Social Intelligence in AI Agents
di: Mathur, Leena, et al.
Pubblicazione: (2026)
di: Mathur, Leena, et al.
Pubblicazione: (2026)
Emerging Practices in Frontier AI Safety Frameworks
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
Questionnaire Responses Do not Capture the Safety of AI Agents
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2026)
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2026)
AI Safety in Generative AI Large Language Models: A Survey
di: Chua, Jaymari, et al.
Pubblicazione: (2024)
di: Chua, Jaymari, et al.
Pubblicazione: (2024)
STREAM (ChemBio): A Standard for Transparently Reporting Evaluations in AI Model Reports
di: McCaslin, Tegan, et al.
Pubblicazione: (2025)
di: McCaslin, Tegan, et al.
Pubblicazione: (2025)
AI Safety for Everyone
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
di: Fort, Kristina
Pubblicazione: (2024)
di: Fort, Kristina
Pubblicazione: (2024)
A Framework for Exploring the Consequences of AI-Mediated Enterprise Knowledge Access and Identifying Risks to Workers
di: Gausen, Anna, et al.
Pubblicazione: (2023)
di: Gausen, Anna, et al.
Pubblicazione: (2023)
Case-based Reasoning Augmented Large Language Model Framework for Decision Making in Realistic Safety-Critical Driving Scenarios
di: Gan, Wenbin, et al.
Pubblicazione: (2025)
di: Gan, Wenbin, et al.
Pubblicazione: (2025)
Safety First: Psychological Safety as the Key to AI Transformation
di: Reich, Aaron, et al.
Pubblicazione: (2026)
di: Reich, Aaron, et al.
Pubblicazione: (2026)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
di: Charnock, Jacob, et al.
Pubblicazione: (2026)
di: Charnock, Jacob, et al.
Pubblicazione: (2026)
Access to Personal Data and the Right to Good Governance during Asylum Procedures after the CJEU's YS. and M. and S. judgment
di: Brouwer, Evelien, et al.
Pubblicazione: (2025)
di: Brouwer, Evelien, et al.
Pubblicazione: (2025)
AI Safety, Alignment, and Ethics (AI SAE)
di: Waldner, Dylan
Pubblicazione: (2025)
di: Waldner, Dylan
Pubblicazione: (2025)
Safety cases for frontier AI
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
When Should Algorithms Resign? A Proposal for AI Governance
di: Bhatt, Umang, et al.
Pubblicazione: (2024)
di: Bhatt, Umang, et al.
Pubblicazione: (2024)
AI Agents Should be Regulated Based on the Extent of Their Autonomous Operations
di: Osogami, Takayuki
Pubblicazione: (2025)
di: Osogami, Takayuki
Pubblicazione: (2025)
Agentic AI Systems Should Be Designed as Marginal Token Allocators
di: Zhu, Siqi
Pubblicazione: (2026)
di: Zhu, Siqi
Pubblicazione: (2026)
International AI Safety Report 2026
di: Bengio, Yoshua, et al.
Pubblicazione: (2026)
di: Bengio, Yoshua, et al.
Pubblicazione: (2026)
Social Theory Should Be a Structural Prior for Agentic AI: A Formal Framework for Multi-Agent Social Systems
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2026)
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2026)
World Models Should Prioritize the Unification of Physical and Social Dynamics
di: Zhang, Xiaoyuan, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoyuan, et al.
Pubblicazione: (2025)
(When) Should We Delegate AI Governance to AIs? Some Lessons from Administrative Law
di: Caputo, Nicholas
Pubblicazione: (2025)
di: Caputo, Nicholas
Pubblicazione: (2025)
Governing dual-use technologies: Case studies of international security agreements and lessons for AI governance
di: Wasil, Akash R., et al.
Pubblicazione: (2024)
di: Wasil, Akash R., et al.
Pubblicazione: (2024)
A Conceptual Framework for AI-based Decision Systems in Critical Infrastructures
di: Leyli-abadi, Milad, et al.
Pubblicazione: (2025)
di: Leyli-abadi, Milad, et al.
Pubblicazione: (2025)
Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI
di: Ji, Eugene Yu, et al.
Pubblicazione: (2026)
di: Ji, Eugene Yu, et al.
Pubblicazione: (2026)
The BIG Argument for AI Safety Cases
di: Habli, Ibrahim, et al.
Pubblicazione: (2025)
di: Habli, Ibrahim, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Responsible Governing AI Proliferation
di: Kembery, Edward
Pubblicazione: (2024) -
Position Paper: Model Access should be a Key Concern in AI Governance
di: Kembery, Edward, et al.
Pubblicazione: (2024) -
How Should AI Safety Benchmarks Benchmark Safety?
di: Yu, Cheng, et al.
Pubblicazione: (2026) -
AI Safety Should Prioritize the Future of Work
di: Hazra, Sanchaita, et al.
Pubblicazione: (2025) -
AI Safety as Control of Irreversibility: A Systems Framework for Decision-Energy and Sovereignty Boundaries
di: Shu, Wesley, et al.
Pubblicazione: (2026)